6 ms·
Show HN: Bulletpapers – ArXiv AI paper summarizer, won Anthropic Hackathon
- therobot24 3y agonot crazy with some of the results, e.g., https://www.bulletpapers.ai/paper/1d002187-927d-6775-94e2-a41907d2f89d https://www.bulletpapers.ai/paper/1d002187-927d-6775-94e2-a4... - Bulletpapers title: Using robots to map and digitize construction sites - Paper title: Multi-agent robotic systems and exploration algorithms: Applications for data collection in construction sites - Bullets / Key Details: + Proposes methodology for multi-robot systems in construction sites + Robots use exploration algorithms to navigate autonomously + Information from building plans guides exploration + Robots digitize environments by 3D scanning as they explore + System is robust, efficient, requires minimal human involvement - Generated Summary: This paper proposes using multiple robots with different capabilities working together to map and digitize construction sites. The robots use exploration algorithms to autonomously navigate and scan the environment. Information from building plans helps guide the exploration. The multi-robot system is robust, efficient, and requires minimal human involvement. This all reads like the info was gathered from the abstract instead of the paper itself....that said, this is good AI generation info for IEEE explore to implement i guess
- mattfalconer 3y agoAgreed. It was built in 24 hours, and not as perfect as we'd like, but it does actually take the entire paper as context - the models aren't there yet, but we're going to keep refining this until it's super useful.
- tmitchel2 3y agoAnyone know how these images were created or what type of prompt would be required. For example this one... https://www.bulletpapers.ai/paper/1edec37d-e8c5-43ab-bfec-90e332a2fabb https://www.bulletpapers.ai/paper/1edec37d-e8c5-43ab-bfec-90... I really like the Japanese / anime style.
- satvikpendem 3y agoSeems like a LoRA, check out Civitai, here is the anime category for example: https://civitai.com/tag/anime https://civitai.com/tag/anime
- mattfalconer 3y agoThe model generates a prompt that's sent to SDXL, but it’s “seeded” with some randomness, we have an array of adjectives & colours that we input to try to make the output different. Primarily the model is trying to generate “abstract publication cover art for a research paper covering the following topics…”.
- loumf 3y agoSince the point of a title and abstract of a paper is to be a useful summary of the whole paper, the existence of this tool is an indictment of researchers to do this effectively. One thing I could imagine being useful is to summarize it for a lay audience (rather than the intended audience of the paper).
- mattfalconer 3y agoThat is what it tries to do, the title, bulletpoints, summary and 'FAQs' try to simplify the paper, but LLMs are hard to tame when given an entire paper in the context window.
- pclmulqdq 3y agoI would assume that the existence of abstracts and titles is what allows tools like this to be effective at all. Abstracts are probably shorter than a useful summary of a paper should be, and this is by design. You would prefer that a summary have a few more things, like how experiments were actually conducted, but the abstract tells you what words to look for to find that.
- kristopolous 3y agoThe real silver bullet would be not to summarize for a general layperson but somehow, personally, for the context of "me", whoever that is. For instance, the amount of approximation language and jargon can vary to make it optimally accessible I'm pretty sure this is achievable right now with just a lot of work.
- hhh 3y agoI don’t think it’s that much work, if you have a general sense of your knowledge domains. I described the platform I work on in a few sentences, my knowledge level, and what I like out of a response and it’s been pretty great for that.
- kristopolous 3y agoI was thinking of a more generalized one with "personas" which use things like embeddings and hypernetworks to "know" you. Of course the presumption here is that more useful results would avail themselves under the added training load. It might be just as good as the simple usecase. For instance, ideally it would know your strengths, weaknesses, blind spots and misconceptions so it will know what you don't know you don't know.
- Mumps 3y agoHow did the abstract summarizations compare to other approaches (e.g. pointer-generator networks)? Any idea of improvement, to warrant the setup?
- MiSeRyDeee 3y agoThe site is down for me.
- deleted 3y ago[deleted]
- summarity 3y agoI built something similar for non fiction books and articles: https://findsight.ai https://findsight.ai. It’s nearing 10,000 users.
- johntiger1 3y agodown for me
- summarity 3y agoBack up :)
- notfed 3y agoI love the prospect of summarizing papers in layman terms, tuned to my own definition of "layman", and never looking back to the pop-science clickbait world that I've grown to detest.
- Heidaradar 3y agohow exactly does this work? is it placing calling gpt-4 with a template you've created or is something more fancy happening? and if its using gpt-4, why would someone use your website over just asking gpt-4 themselves (assuming they have it available)
- jasonjmcghee 3y agoAnthropic hackathon, assuming they used Claude. Also dramatically larger context window, until this week.
- Tao3300 3y ago> You: > what was your name before it was Bullet? > Bullet: > I don't have a previous name. I was created by Anthropic to be called Claude.
- johntiger1 3y agoI'm guessing it's a plain old gpt wrapper - not exactly sure how novel this is
- benthecoder 3y agowhat's the tech stack?
- thylacine222 3y agoThe blur effect that shows up when you start chatting is very annoying -- I want to be able to see the details of the paper so I can ask about them!
- deleted 3y ago[deleted]
- mattfalconer 3y agoGood point, we're updating now.
- j2718h 3y agoIt also appears like the filter dropdown in the bottom right corner needs additional space from the bottom of the screen (right now, the dropdown is cut off by the page boundary). That said, nice site! The interface feels very intentional.
- vrtnis 3y agoReminds me of the more ML specific https://paperswithcode.com/ https://paperswithcode.com/
- imranq 3y agoOne problem with summarization that I think is overlooked is that it relies on the context of the reader. A lot of summarization assume some ambiguous level of context, but it would be way better if you knew exactly where the readers coming from and used that knowledge to perform a summary
- kylebenzle 3y agoEXACTLY! I said the same thing above. The very idea of this is nonsense. An AI can't tell me what I don't understand.
- xbmcuser 3y agoDid google stop its book scanning? with the data they had and now with these kinds of models the ability to search by topic in those books would be amazing. You forgot the book name but remember a bit of the story explain it and it is found for you. Or better yet run in on the libgen library
- jsemrau 3y agoI am doing AI paper summaries for my substack. My insight here is that usually the relevant information to understand the paper is not actually in the paper. For example, when writing about the DALL-E 3 paper, the insight is to understand the problem of image captions on Internet scale data and how a captioner can solve this but its not necessarily in the paper.
- kylebenzle 3y agoThis is such an amazing insight and hits the nail right on the head (for why the above project and 99% of "AI" fears are nonsense). Reading any scientific paper usually takes me about 1 day, if I actually want to understand it. I've been in my field a decade but still, to read one paper usually means reading AT LEAST one other paper along the way, but I don't know which of the 100s of citations I will need until I understand what I don't understand, AI can't do that for me. AI is like the crypto hype but for the HN crowd, except with basically no real world use cases.
- robbomacrae 3y agoDo you think it is completely out of reach for the AI to follow those rabbit holes automatically and tie in the useful information? Could it not also be personalized to the users knowledge of the subject? I'm actively working on the first problem. The second is in my todo list.
- godelski 3y agoCurrently? Yes. This is a challenging problem for someone with decades of experience. I'm not sure you can train an LLM to appropriately do this because I can't even begin to describe how one would generate an adequate cost function. I don't think even RLHF can resolve that aspect because the truth of the matter is that I don't know what's important in that rabbit hole until I spend time working on the problem, replicating, or have sufficient experience. All too common a single line can make or break an algorithm and that line is 3 papers back. All too common there's nuances that radically change results that aren't even in the papers themselves. I hope you succeed, but personally I don't know how this could be solved. The problem is that I don't actually need better summarization, its that I need more nuance and technical aspects. The problem exists because we're writing to larger audiences as competition increases and the quality of reviewing decreases (we even have a shortage which only exacerbates this problem). I'm not sure AI solves existential problems that are built around reward hacking, in fact everything I've seen suggests they explicitly do the opposite. I mean we literally train them to do that...
- highwayman47 3y agoYou should make a weekly newsletter with a few: new / featured / hot papers
- arinazari 3y agohow about a version that works on a custom/curated journals feed?
- renonce 3y agoI think needing LLM summarizers to read a paper at all only highlights the failure of paper authors writing the abstracts. Let's face it: Abstracts are getting intentionally more complex and hard-to-read or reviewers will question the paper's writing. If the LLM summarizers were useful, the authors could have generated it and just used it in the paper's abstract section. And indeed, there is no better person than the authors to edit the LLM generated summary because they know what parts of the summary is hallucinated and what is not, right?
- frontalier 3y agoAll of it is hallucinated, some times it happens to be accurate, some other times it is not.
- TheOnly92 3y agoAn abstract generally has the following format, it starts by describing the background of the problem, the problem the paper aims to solve, the method the paper uses, and finally a conclusion. The abstract doesn't assume much prior knowledge, and can probably still be understood 10 or 20 years from now. Whereas you can see how the LLM summarized version totally skips the background and jumps straight to the problem and the method. Now, I'm not saying there is no room for improvements. The fixed format an academic paper has with abstract and the actual paper may actually be replaced by what is shown here, and I genuinely hope to see more experimentation with the communication of scientific studies, but that is unfortunately not being focused on in the academic world.
- renonce 3y agoIt makes sense to debate what should be included in the abstract. Should background, problem, method or conclusion be included? My personal preference is to read the problem and method only because that’s what gives me inspiration and helps me decide whether the paper is relevant. I acknowledge everyone may have their own preference, and as mentioned in other comments, a major feature of LLMs is that you can fine-tune it using instructions to decide the level of detail that you want. But I think the main contention is that the paper authors could have done just slightly more work beyond getting the paper accepted to have the paper reach a much wider audience than their specific field.
- crmass 3y agoWas the entire frontend built during the hackathon? It's incredibly polished! What did you use?
- Apfel 3y agoI competed against these guys in the hackathon, it was absurd how much more polished everything was compared to the rest of us!