11 ms·
Show HN: Spegel, a Terminal Browser That Uses LLMs to Rewrite Webpages
- bubblyworld 1y agoClassic that the first example is for parsing the goddamn recipe from the goddamn recipe site. Instant thumbs up from me haha, looks like a neat little project.
- andrepd 1y agoWhich it apparently does by completely changing the recipe in random places including ingredients and amounts thereof. It is _indeed_ a very good microcosm of what LLMs are, just not in the way these comments think.
- throwawayoldie 1y agoThe output was then posted to the Internet for everyone to see, without the minimal amount of proofreading that would be necessary to catch that, which gives us a good microcosm of how LLMs are used. On a more pleasant topic the original recipe sounds delicious, I may give it a try when the weather cools off a little.
- simedw 1y agoIt was actually a bit worse than that the LLM never got the full recipe due to some truncation logic I had added. So it regurgitated the recipe from training, and apparently, it couldn't do both that and convert units at the same time with the lite model (it worked for just flash). I should have caught that, and there are probably other bugs too waiting to be found. That said, it's still a great recipe.
- andrepd 1y ago[flagged]
- deleted 1y ago[deleted]
- 0x696C6961 1y agoWhat is the point?
- plonq 1y agoI’m someone else but for me the point is a serious bug resulted _incorrect data_, making it impossible to trust the output.
- bubblyworld 1y agoAssuming you are responding in good faith - the author politely acknowledged the bug (despite the snark in the comment they responded to), explained what happened and fixed it. I'm not sure what more I could expect here? Bugs are inevitable, I think it's how they are handled that drives trust for me.
- andrepd 1y agoThe point is LLMs are fundamentally unreliable algorithms for generating plausible text, and as such entirely unsuitable for this task. "But the recipe is probably delicious anyway" is beside the point, when it completely corrupted the meaning of the original. Which is annoying when it's a recipe but potentially very damaging when it's something else. Techies seem to pretend this doesn't happen, and the general public who doesn't understand will trust the aforementioned techies. So what we see is these tools being used en masse and uncritically for purposes to which they are unsuited. I don't think this is good.
- bubblyworld 1y agoWhat do you mean? The recipes in the screenshot look more or less the same, the formatting has just changed in the Spiegel one (which is what was asked for, so no surprises there). Edit: just saw the author's comment, I think I'm looking at the fixed page
- IncreasePosts 1y agoThere are extensions that do that for you, in a deterministic way and not relying on LLMs. For example, Recipe Filter for chrome. It just shows a pop up over the page when it loads if it detects a recipe
- bubblyworld 1y agoThanks, I already use that plugin, actually, I just found the problem amusingly familiar. Recipe sites are the original AI slop =P
- lpribis 1y agoAnother great example of LLM hype train re-inventing something that already existed [1] (and was actually thought out) but making it worse and non-deterministic in the worst ways possible. https://schema.org/Recipe https://schema.org/Recipe
- deleted 1y ago[deleted]
- deleted 1y ago[deleted]
- soap- 1y agoAnd that would be great, if anyone used it. LLMs are specifically good at a task like this because they can extract content from any webpage, regardless of it supports whatever standard that no one implements
- komali2 1y agoThat's a cool schema, but the LLM solution is necessary because recipe website makers will never use the schema because they want you to have to read through garbage, with some misguided belief that this helps their SEO or something. Or maybe they get more money if you scroll through more ads?
- bubblyworld 1y agoI'm genuinely a bit confused by the recipe blog business model. Like there's got to be one, right? People don't usually spew the same story about their grandma hundreds of times on a real blog. Just hitting keywords for search? Many of them don't even have ads so I feel like that can't be it. Maybe referrals?
- Revisional_Sin 1y agoSEO. Longer articles get ranked higher.
- cyrillite 1y agoI have been thinking of a project extremely similar to this for a totally different purpose. It’s lovely to see something like this. Thank you for sharing it, inspiring
- amelius 1y agoCurious about that different purpose ...
- crest 1y agoA cool hack, but also impressive to come up with a CLI "browser" that's even more expensive to run than Chromium.
- leroman 1y agoCool idea! but kind of wasteful.. I just feel wrong if I waste energy.. At least you could first turn it into markdown with a library that preserves semantic web structures (I authored this- https://github.com/romansky/dom-to-semantic-markdown https://github.com/romansky/dom-to-semantic-markdown) saving many tokens = much less energy used..
- otabdeveloper4 1y agoThis is exactly the sort of thing that should be running on a local LLM. Using a big cloud provider for this is madness.
- remram 1y agoNot to be confused with Kubernetes' Spiegel: https://spegel.dev/ https://spegel.dev/ https://github.com/spegel-org/spegel https://github.com/spegel-org/spegel
- herval 1y agoWe’re back to the BBS days, 30 years later!
- robbles 1y agoI'm curious whether anyone has run into hallucinations with this kind of use of an LLM. They are pretty great at converting data between formats, but I always worry there's a small chance it changes the actual data in the output in some small but misleading way.
- throwawayoldie 1y agoGuess you didn't see the old version of the screenshots on the page, which showed things like 1.5 pounds of lamb being converted into 1.5 kg, that is, more than doubling it.
- barrenko 1y agoI need this, but for the new forum formats such as Discourse or Discuss or whatever it's called. An eyesore and a brainsore.
- Jotalea 1y agoInsanely resource expensive, but still a very interesting "why not?" idea. I think a fitting use case would be adapting newer websites for them to work on older hardware. That is, assuming the new technologies used are not vital to the functionality of the website (ex. Spotify, YouTube, WhatsApp) and can be adapted to older technologies (ex. Google Search, from all the styles that it has, to a simple input and a button). In theory this could be used for ad blocking; though more expensive and less efficient, but the idea is there. So, it is a very curious idea, but we still have to find an appropriate use case.
- hambes 1y agoI've thought about getting a web browser to work on the terminal for a while now. This is an idea that hadn't occured to me yet and I'm intrigued. But I feel it doesn't solve the main issue of terminal-based web browsing. Displaying HTML in the terminal is often kind of ugly and css-based fanciness does not work at all, but that can usually just be ignored. The main problem is javascript and dynamic content, which this approach just ignores. So no real step forward for cli web browsing, imo.
- ghm2180 1y agoThis is great! Another useful amendment to this that would make me use it add a chrome browser tool to allow access to pages that need authn and then scrape them for you. My #1 usecase is fetching wikis on my hard drive and letting a local coding agent use it for creating plans.
- gvison 1y agoGreat project, much less memory than opening a web page in a browser.
- js2 1y agoI was unfamiliar with Textual which seems more interesting than Slowly Braised Lamb Ragu: https://textual.textualize.io/ https://textual.textualize.io/
- qsort 1y agoThis is actually very cool. Not really replacing a browser, but it could enable an alternative way of browsing the web with a combination of deterministic search and prompts. It would probably work even better as a command line tool. A natural next step could be doing things with multiple "tabs" at once, e.g: tab 1 contains news outlet A's coverage of a story, tab 2 has outlet B's coverage, tab 3 has Wikipedia; summarize and provide references. I guess the problem at that point is whether the underlying model can support this type of workflow, which doesn't really seem to be the case even with SOTA models.
- hliyan 1y agoFor me, a natural next step would be to turn this into a service -- rather than doing it in the browser, this acts as a proxy, strips away all the crud and serves your browser clean text. No need to install a new browser, just point the browser to the URL via the service. But if we do it, we have to admit something hilarious: we will soon be using AI to convert text provided by the website creator into elaborate web experiences, which end users will strip away before consuming it in a form very close to what the creator wrote down in the first place (this is already happening with beautifully worded emails that start with "I hope this email finds you well").
- npmipg 1y agoworking on this as we speak!
- simedw 1y agoThank you. I was thinking of showing multiple tabs/views at the same time, but only from the same source. Maybe we could have one tab with the original content optimised for cli viewing, and another tab just doing fact checking (can ground it with google search or brave). Would be a fun experiment.
- phatskat 1y ago> I was thinking of showing multiple tabs/views at the same time, but only from the same source. I think the primary reason I use multiple tabs but _especially_ multiple splits is to show content from various sources. Obviously this is different that a terminal context, as I usually have figma or api docs in one split and the dev server on the other. Still, being able to have textual content from multiple sources visible or quickly accessible would probably be helpful for a number of users
- ohadron 1y agoThis is a terrific idea and could also have a lot of value with regards to accessibility.
- taco_emoji 1y agoThe problem, as always, is that LLMs are not deterministic. Accessibility needs to be reliable and predictable above all else.
- pepperonipboy 1y agoCould work great with emacs' eww!
- thephotonsphere 1y agoalso with lynx because it can browse from stdin
- sammy0910 1y agoI built a project that basically does this for emacs https://github.com/sstraust/simpleweb https://github.com/sstraust/simpleweb
- clbrmbr 1y agoSuggestion: add a -p option: spegel -p "extract only the product reviews" > REVIEWS.md
- sammy0910 1y agoI built something that did this a bit ago https://github.com/sstraust/simpleweb https://github.com/sstraust/simpleweb
- deleted 1y ago[deleted]
- sammy0910 1y agosomething I found challenging when I was building was -- how do you make the speed fast enough so that it still creates a smooth browsing experience? I'm curious how you tackled that problem
- simedw 1y agoThat's a cool project. I think most of it comes down to Flash-Lite being really fast, and the fact that I'm only outputting markdown, which is fairly easy and streams well.
- 4b11b4 1y agohttps://github.com/sstraust/simpleweb/blob/79294b461b2e67a243a6db750eeadd2fff296fde/simpleweb/src/simpleweb/compute_simple_page.clj#L15 https://github.com/sstraust/simpleweb/blob/79294b461b2e67a24... Not the answer to your question but here's the prompt
- busssard 1y agowhat does it do about javascript?
- anonu 1y agoDon't you need javascript to make most webpages useful?
- inetknght 1y agoGood sir, no. The web has existed for long before javascript was around. The web was useful for long before javascript was around. I literally hate javascript -- not the language itself but the way it is used. It has enabled some pretty cool things, yes. But javascript is not required to make useful webpages.
- pmxi 1y agoI think you misunderstood him. Yes, it’s possible to CREATE a useful webpage without JavaScript, but many EXISTING webpages rely on JavaScript to be functional.
- jazzyjackson 1y agoIf Amazon.com can work with JavaScript disabled, any site could be rewritten to do without. But I think to even get to the content on a lot of SPAs this would need to be running a headless browser to render the page, before extracting the static content unfortunately
- IncreasePosts 1y agoNo - an experiment: try disabling javascript in your browser settings, and then whenever you see a webpage that isn't working, enable javascript for that domain. You'd be surprised how fast 90% of the web feels with JS disabled.
- nicklo 1y agoHave you considered making an MCP for this? Would be great for use in vibe-coding
- ktpsns 1y agoReminds me of https://www.brow.sh/ https://www.brow.sh/ which is not AI related at all but just a very powerful terminal browser which in fact supports JS, even videos.
- deleted 1y ago[deleted]
- cheevly 1y agoVery cool! My retired AI agent transformed live webpage content, here's an old video clip of transforming HN to My Little Pony (with some annoying sounds): https://www.youtube.com/watch?v=1_j6cYeByOU https://www.youtube.com/watch?v=1_j6cYeByOU. Skip to ~37 seconds for the outcome. I made an open-source standalone Chrome extension as well, it should probably still work for anyone curious: https://github.com/joshgriffith/ChromeGPT https://github.com/joshgriffith/ChromeGPT
- ghaering 1y ago[dead]
- Klaster_1 1y agoNow that's a user agent!
- CaptainFever 1y agoFinally, web browsers work for the user, not the website owners!
- adrianpike 1y agoSuper neat - I did something similar on a lark to enable useful "web browsing" over 1200 baud packet - I have Starlink back at my camp but might be a few miles away, so as long as I can get line of sight I can Google up stuff, albeit slow. Worked well but I never really productionalized it beyond some weekend tinkering.
- eniac111 1y agoCool! It would be even better if it was able to create simple web pages for vintage browsers.
- stronglikedan 1y agoThat would violate the do-one-thing-and-do-it-well principle for no apparent benefit. There are plenty of tools to convert markdown to basic HTML already.
- treyd 1y agoI wonder if you could use a less sophisticated model (maybe even something based on LSTMs) to walk over the DOM and extract just the chunks that should be emitted and collected into the browsable data structure, but doing it all locally. I feel like it'd be straightforward to generate training data for this, using an LLM-based toolchain like what the author wrote to be used directly.
- askonomm 1y agoUnfortunately in the modern web simply walking the DOM doesn't cut it if the website's content loads in with JS. You could only walk the DOM once the JS has loaded, and all the requests it makes have finished, and at that point you're already using a whole browser renderer anyway.
- kccqzy 1y agoYeah but this project doesn't use JS anyway.
- deepdarkforest 1y agoThe main problem with these approaches is that most sites now are useless without JS or having access to the accessibility tree. Projects like browser-use or other DOM based approaches at least see the DOM(and screenshots). I wonder if you could turn this into a chrome extension that at least filters and parses the DOM
- jadbox 1y agoI actually made a CLI tool recently that uses Puppeteer to render the page including JS, summarizes key info and actions, and enables simple form filling all from a CLI menu. I built it for my own use-cases (checking and paying power bills from CLI), but I'd love to get feedback on the core concept: https://github.com/jadbox/solomonagent https://github.com/jadbox/solomonagent
- andoando 1y agoDude I love this. I've been thinking of doing this exactly this, but for as a screen reader for accessibility reasons.
- jadbox 1y agoThanks, it's alpha at the moment- next feature is complex forms and bug fixing broken actions (downloading). Do give it a spin! Welcome to contribute or drop feedback on the repo :)
- willsmith72 1y agoTrue for stuff requiring interaction, but to help their LCP/SEO lots of sites these days render plain html first. It's not "usable" but for viewing it's pretty good
- fzaninotto 1y agoCongrats! Now you need an entire datacenter to visualize a web page.
- juujian 1y agoCouldn't this time reasonably well on a local machine is you have some kind of neutral processing chip and enough ram? Conversion to MD shouldn't require a huge model.
- busssard 1y agoonly if you use an API and not a dedicated distill/tune for html to MD conversion. But the question of Javascript remains
- stared 1y agoAny chance it would work for pages like Facebook or LinkedIn? I would love to have a distraction-free way of searching information there. Obviously, against wishes of these social networks, which want us to be addicted... I mean, engaged.
- simedw 1y agoWe’ll probably have to add some custom code to log in, get an auth token, and then browse with it. Not sure if LinkedIn would like that, but I certainly would.
- aydyn 1y agoDoes anyone really get addicted to linkedin? Its so sanitized and clinical. Nobody acts real on there or even pretends to.
- encom 1y agoThe worst[1] part about losing my job last month was having to take LinkedIn seriously, and the best[2] part about now having found a new job is logging off LinkedIn, for a very long time hopefully. The self-aggrandising, pretentious, occasionally virtue signalling, performance-posting make me want to throw up. It takes a considerable amount of effort on my part to not make sarcastic shitposts, but in the interest of self preservation, I restrain myself. My header picture, however, is my extremely messy desk, full of electronics, tools, test equipment, drawings, computers and coffee cups. Because that's just how I work when I'm in the zone, and it serves as a quiet counterpoint to the polished self-promotion people do. And I didn't even get the new job through LinkedIn, though it did yield one interview. [1] Not the actual worst. [2] Not the actual best.
- b0a04gl 1y agothis is another layer of abstraction on top of an already broken system. you're running html through an llm to get markdown that gets rendered in a terminal browser. that's like... three format conversions just to read text. the original web had simple html that was readable in any terminal browser already. now they arent designed as documents anymore but rather designed as applications that happen to deliver some content as a side effect
- amelius 1y agoI take it you never use "Reader mode" in your browser?
- MangoToupe 1y agoThat's the world we live in. You can either not have access to content or you must accept abstractions to remove all the bad decisions browser vendors have forced on us the last 30 years to support ad-browsing.
- _joel 1y ago> this is another layer of abstraction on top of an already broken system pretty much like all modern computing then, hey.
- nashashmi 1y agoThink of it as a secretary that is transforming and formatting information. You may desire for the original medium to be something like what you want but you don’t get that so you can get a cheap dumber secretary instead.
- worldsayshi 1y agoIf the web site is a SPA that is hydrated using an API it would be conceivable that the LLM can build a reusable interface around the API while taking inspiration from the original page. That interface can then be stored in some cache. I'm not saying it's necessarily a good idea but perhaps a bad/fun idea that can inspire good ideas?
- deleted 1y ago[deleted]
- 098799 1y agoYou could also use headless selenium under the hood and pipe to the model the entire Dom of the document after the JavaScript was loaded. Of course it would make it much slower but also would amend the main worry people have which is many websites will flat out not show anything in the initial GET request.
- busssard 1y agocan you flesh this out a tiny bit? because for indy-crawlers the javascript rendering is the main problem.
- 098799 1y agoHere's a sketch: https://chatgpt.com/share/68640b97-9a48-8007-a27c-fdf85ff41200 https://chatgpt.com/share/68640b97-9a48-8007-a27c-fdf85ff412... -- selenium drives your actual browser under the hood.
- web3aj 1y agoVery cool. I’ve been interested in browsing the web directly from my terminal; this feels accessible.
- insane_dreamer 1y agoInteresting, but why round-trip through an LLM just to convert HTML to Markdown?
- markstos 1y agoBecause the modern web isn't reliably HTML, it's "web apps" with heavy use of JavaScript and API calls. To first display the HTML that you see in your browser, you need a user agent that runs JavaScript and makes all the backend calls that Chrome would make to put together some HTML. Some websites may still return some static upfront that could be usefully understood without JavaScript processing, but a lot don't. That's not to say you need an LLM, there are projects like Puppeteer that are like headless browsers that can return the rendered HTML, which can then be sent through an HTML to Markdown filter. That would be less computationally intensive.
- insane_dreamer 1y ago> That's not to say you need an LLM, ... then be sent through an HTML to Markdown filter. That would be less computationally intensive. which was exactly my point
- crent 1y agoBecause this isn't just converting HTML to markdown. I'd recommend taking another look at the website and particularly read the recipe example as it demonstrates the goal of the project pretty well.
- nashashmi 1y agoYou should call this software a lens and filter instead of a mirror. It takes the essential information and transforms it into another medium.
- amelius 1y agoCan it strip ads?
- tossandthrow 1y agoIt can inject its own!
- amelius 1y agoYou have a point as it uses Gemini under the hood. However, the moment Google introduces ads in the model users will run away. So Google really has no opportunity here to inject ads. And wouldn't it be ironic if Gemini was used to strip ads from webpages?
- tossandthrow 1y agoThe field of "seo for Ai", ie, seeking to have your company featured in LLMs, is already established. In the rare cases where the model would jam on its own, this will likely already happen.
- mromanuk 1y agoI definitely like the LLM in the middle, it’s a nice way to circumvent the SEO machine and how Google has optimized writing in recent years. Removing all the cruft from a recipe is a brilliant case for an LLM. And I suspect more of this is coming: LLMs to filter. I mean, it would be nice to just read the recipe from HTML, but SEO has turned everything into an arms race.
- hirako2000 1y agoDo you also like what it costs you to browse the web via an LLM potentially swallowing millions of tokens per minutes ?
- prophesi 1y agoThis seems like a suitable job for a small language model. Bit biased since I just read this paper[0] [0] https://research.nvidia.com/labs/lpr/slm-agents/ https://research.nvidia.com/labs/lpr/slm-agents/
- yellow_lead 1y agoLLM adds cruft, LLM removes cruft, never a miscommunication
- visarga 1y agoI foreseen this a couple years ago. We already have web search tools in LLMs, and they are amazing when they chain multiple searches. But Spegel is a completely different take. I think the ad blocker of the future will be a local LLM, small and efficient. Want to sort your timeline chronologically? Or want a different UI? Want some things removed, and others promoted? Hide low quality comments in a thread? All are possible with LLM in the middle, in either agent or proxy mode. I bet this will be unpleasant for advertisers.
- tines 1y ago> Removing all the cruft from a recipe is a brilliant case for an LLM Is it though, when the LLM might mutate the recipe unpredictably? I can't believe people trust probabilistic software for cases that cannot tolerate error.
- kelsey98765431 1y agoPeople here are not realizing that html is just the start. If you can render a webpage into a view, you can render any input the model accepts. PDF to this view. Zip file of images to this view. Giant json file into this view. Whatever. The view is the product here, not the html input.
- nartho 1y agoI think the project itself is really cool, that said I really don't like the trend of having LLMs regurgitate content back to us. That said, this kinda makes me think of Browsh, who took the opposite approach and tries to render the HTML in the terminal (without LLMs as far as I know) https://github.com/browsh-org/browsh https://github.com/browsh-org/browsh https://www.youtube.com/watch?v=HZq86XfBoRo https://www.youtube.com/watch?v=HZq86XfBoRo
- hirako2000 1y agoThat would also keep your wallet or GPU rag coller
- deleted 1y ago[deleted]
- hyperific 1y agoWhy not use pandoc to convert html to markdown and have the LLM condense from there?
- __MatrixMan__ 1y agoIt would be cool of it were smart enough to figure out whether it was necessary to rewrite the page on every visit. There's a large chunk of the web where one of us could visit once, rewrite to markdown, and then serve the cleaned up version to each other without requiring a distinct rebuild on each visit.
- pmxi 1y agoThe author says this is for “personalized views using your own prompts.” Though, I suppose it’s still useful to cache the outputs for the default prompt.
- __MatrixMan__ 1y agoOr to cache the output for whatever prompt your peers think is most appropriate for that particular site.
- myfonj 1y agoEach user have distinct needs, and has a distinct prior knowledge about the topic, so even the "raw" super clean source form will probably be eventually adjusted differently for most users. But yes, having some global shared redundant P2P cache (of the "raw" data), like IPFS (?) could possibly help and save some processing power and help with availability and data preservation.
- __MatrixMan__ 1y agoI imagine it sort of like a microscope. For any chunk of data that people bothered to annotate with prompts re: how it should be rendered you'd end up with two or three "lenses" that you could toggle between. Or, if the existing lenses don't do the trick, you could publish your own and, if your immediate peers find them useful, maybe your transitive peers will end up knowing about them as well.
- simedw 1y agoIf the goal is to have a more consistent layout on each visit, I think we could save the last page's markdown and send it to the model as a one-shot example...
- WD-42 1y agoDoes anyone know why LLMs love emojis so much?
- userbinator 1y agoLikely because it was trained on such material... which is just as authentic and vapid.
- coder543 1y agoJust a typo note: the flow diagram in the article says "Gemini 2.5 Pro Lite", but there is no such thing.
- simedw 1y agoYou are right, it's Gemini 2.5 Flash Lite
- neocodesoftware 1y agoDoes it fail cloudflare captcha?
- ospider 1y agoI think it will, it uses requests, and cloudflare blocks traffic from non-browser, e.g. python http clients. It would be better to use something like curl_cffi.
- willm 1y agoWhy not just use ncurses?
- deleted 1y ago[deleted]
- Bluestein 1y agoGosh. Lovely project and cool, and - likewise - a bit scary: This is where the "bubble" seals itself "from the inside" and custom (or cloud, biased) LLMs sear the "bubble" in.- The ultimate rose (or red, or blue or black ...) coloured glasses.-
- wayeq 1y ago... what?
- rrnechmech 1y agoI think OP means that this "filtering" the on the fly conversion does can amplify the content bubbles we live in
- mossTechnician 1y agoChanges Spegel made to the linked recipe's ingredients: Pounds of lamb become kilograms (more than doubling the quantity of meat), a medium onion turns large, one celery stalk becomes two, six cloves of garlic turn into four, tomato paste vanishes, we lose nearly half a cup of wine, beef stock gets an extra ¾ cup, rosemary is replaced with oregano.
- achierius 1y agoDid you actually observe this, or is just meant to be illustrative of what could happen?
- mossTechnician 1y agoThis is what actually happened in the linked article. The recipe is around the text that says > Sometimes you don't want to read through someone's life story just to get to a recipe... That said, this is a great recipe I compared the list of ingredients to the screenshot, did a couple unit conversions, and these are the discrepancies I saw.
- orliesaurus 1y agooh damn...
- jugglinmike 1y agoGreat catch. I was getting ready to mention the theoretical risk of asking an LLM be your arbiter of truth; it didn't even occur to me to check the chosen example for correctness. In a way, this blog post is a useful illustration not just of the hazards of LLMs, but also of our collective tendency to eschew verity for novelty.
- throwawayoldie 1y ago> the theoretical risk of asking an LLM be your arbiter of truth "Theoretical"? I think you misspelled "ubiquitous".
- 1y ago
- jannniii 1y agoGopher is back!
- jannniii 1y ago[dead]
- IncreasePosts 1y agoI did something similar, but with a chrome extension. Basically, for every web page, I feed the HTML to a local LLM (well, on a server in my basement). I ask it to consider if the content is likely clickbait or can be summarized without losing too many interesting details, and if so, it adds a little floating icon to the top of the page that I can click on to see the summary instead. My next plan is to rewrite hyperlinks to provide a summary of the page on hover, or possibly to rewrite the hyperlinks to be more indicative of the content at the end of it(no more complaining about the titles of HN posts...). But, my machine isn't too beefy and I'm not sure how well that will work, or how to prioritize links on the page.
- benrutter 1y agoWelcome to 2025 where it's more reasonable to filter all content through an LLM than to expect web developers to make use of the semantic web that's existed for more than a decade. . . Serioisly though, looks like a novel fix for the problem that most terminal browsers face. Namely that terminals are text based, but the web, whilst it contains text, is often subdivided up in a way that only really makes sense graphically. I wonder if a similar type of thing might work for screen readers or other accessibility features
- cout 1y agoThis is a neat idea! I wonder if it could be adapted to render as gopher pages.
- Buttons840 1y agoA step towards the future of ad-blocking maybe? Just rewrite every page?
- userbinator 1y agoMany people were doing that at the turn of the century(!) with filtering proxies, more deterministically and with far less computing power. Some still do today.
- conradkay 1y agoSomething tells me we'll see more ad-inserting
- Modified3019 1y ago>Companies burning energy with llms to dynamically hide ads and bullshit on every pageload >Individuals burning energy using personal llm internet condoms to strips ads and bullshit from every pageload Eventually there will be a project where volunteers use llms to harvest the real internet and “launder” both the copyright and content into some kind of pre-processed distributed shadow internet where things are actual useable, while being just as wrong as the real internet. What a future.
- revskill 1y agoUse uv instead of pip
- tartoran 1y agoLoving the text only browsing. Is this as fast as in the preview?
- eevmanu 1y agogreat POC looks very similar to a chrome extension i use for a similar goal: reader view - https://chromewebstore.google.com/detail/ecabifbgmdmgdllomnfinbmaellmclnh https://chromewebstore.google.com/detail/ecabifbgmdmgdllomnf...
- deadbabe 1y agoI would like to see a version of this where an LLM just takes the highlights of various social media content from your feed and just gives you the stuff worth watching. This also means excluding crap you had no interest in and was simply inserted into your feed. Fight algorithms with algorithms. Eliminate doom scrolling.