5 ms·
Show HN: I made an Ollama summarizer for Firefox
Source: https://github.com/tcsenpai/spacellama https://github.com/tcsenpai/spacellama
- RicoElectrico 2y agoI've found that for the most part the articles that I want summarized are those which only fit the largest context models such as Claude. Because otherwise I can skim-read the article possibly in reader mode for legibility. Is llama 2 a good fit considering its small context window?
- tcsenpai 2y agoPersonally I use llama3.1:8b or mistral-nemo:latest which have a decent contex window (even if it is less than the commercial ones usually). I am working on a token calculator / division of the content method too but is very early
- garyfirestorm 2y agowhy not llama3.2:3B? it has fairly large context window too
- reissbaker 2y agoI assume because the 8B model is smarter than the 3B model; it outperforms it on almost every benchmark: https://huggingface.co/meta-llama/Llama-3.2-3B https://huggingface.co/meta-llama/Llama-3.2-3B If you have the compute, might as well use the better model :) The 3.2 series wasn't the kind of leap that 3.0 -> 3.1 was in terms of intelligence; it was just: 1. Meta releasing multimodal vision models for the first time (11B and 90B), and 2. Meta releasing much smaller models than the 3.1 series (1B and 3B).
- reissbaker 2y agoI don't think this is intended for Llama 2? The Llama 3.1 and 3.2 series have very long context windows (128k tokens).
- tempodox 2y agoWhat about using a Modelfile for ollama that tweaks the context window size? I seem to remember parameters for that in the ollama GitHub docs.
- tcsenpai 2y agoI applied (for now) a pre-filled table with a 4096 default limit. Users can also specify an upper or lower limit from the UI directly now. Added chunk and recursive summarization too.
- htrp 2y agodo multi stage summarization?
- tcsenpai 2y agoHi! This was a good suggestion! I implemented it in v 1.1 which is already out :)
- donclark 2y agoIf we can get this as the default for all the newly posted HN articles please and thank you?
- totallymike 2y agoI sincerely hope this never happens
- ukuina 2y agoThis is why I built https://hackyournews.com https://hackyournews.com It summarizes via Puter (free).
- iJohnDoe 2y agoSo cool! Thanks. Bookmarked.
- deleted 2y ago[deleted]
- chx 2y agoHelp me understand why people are using these. I presume you want information of some value to you otherwise you wouldn't bother reading an article. Then you feed it to a probabilistic algorithm and so you can not have any idea what the output has to do with the input. Like https://i.imgur.com/n6hFwVv.png https://i.imgur.com/n6hFwVv.png you can somewhat decipher what this slop wants to be but what if the summary leaves out or invents or inverts some crucial piece of info?
- andrewmcwatters 2y agoPeople write too much. Get to the point.
- chx 2y agoany point? regardless of what's written? does that work for you?
- garyfirestorm 2y agosometimes you don't have time to read an entirety of a large article. You want a quick summary, some people are poor at summarizing things in their head as they go and can get lost in dense text. Extensions like these really help me with headers, structure that I want to follow, quick overview and gives me an idea if I want to deep dive further.
- drdaeman 2y agoSometimes it's not even an article, but a video. And sometimes all you care is just a single tiny fact from that video. Although I don't think this particular summarizer works for videos. And I don't think Ollama API supports audio ingestion for transcription. There are some summarizers that work with YouTube specifically (using automatic subtitles).
- tcsenpai 2y agoSpeaking of, I made also a youtube summarizer at https://github.com/tcsenpai/youlama https://github.com/tcsenpai/youlama
- asdev 2y agoI built a chrome version of this for summarizing HN comments: https://github.com/built-by-as/FastDigest https://github.com/built-by-as/FastDigest
- deleted 2y ago[deleted]
- oneshtein 2y agoI use PageAssist with Ollama for two months, but I never called "Summarise" option in menu. :-/
- tcsenpai 2y agoTIL, I am experimenting with PageAssist right now
- tcsenpai 2y agoUpdate: v 1.1 is out! - # Changelog ## [1.1] - 2024-03-19 ### Added - New `model_tokens.json` file containing token limits for various Ollama models. - Dynamic token limit updating based on selected model in options. - Automatic loading of model-specific token limits from `model_tokens.json`. - Chunking and recursive summary for long pages - Better handling of markdown returns ### Changed - Updated `manifest.json` to include `model_tokens.json` as a web accessible resource. - Modified `options.js` to handle dynamic token limit updates: - Added `loadModelTokens()` function to fetch model token data. - Added `updateTokenLimit()` function to update token limit based on selected model. - Updated `restoreOptions()` function to incorporate dynamic token limit updating. - Added event listener for model selection changes. ### Improved - User experience in options page with automatic token limit updates. - Flexibility in handling different models and their respective token limits. ### Fixed - Potential issues with incorrect token limits for different models.