8 ms·
AI tools I wish existed
- noja 1y agoFor me: A local model to plug in to Apple photos to look for metadata inconsistencies in my photo librar, add missing location information, add dates from those old scanned photos with the date on the corner.
- bryanrasmussen 1y agoThis seems like a relatively easy thing to code oneself, or for someone to have already made somewhere (relatively easy, just writing something for yourself command line doing it, assuming you can spend a work week of nights [worst case, based on my working with images in folders in the past I think 10 hours for something that works, reasonable time for coffee and other breaks])
- coolThingsFirst 1y agoFor you maybe, we are engaging with the left side of the bell curve never forget that. Also use simpler words.
- backprop1989 1y agoYeah, there are probably a few multimillion dollar app ideas in here, but of insufficient complexity or lowbrow-ness for the typical HN reader (myself included). The nano-banana template idea or the Q&A Walkman, for example.
- lifestyleguru 1y agoThis is a problem solvable with 30 years old technology - bash, exiftool, ImageMagick, Tesseract OCR.
- charcircuit 1y ago>A recommendation engine that looks at my browsing history, sees what blog posts or articles I spent the most time on, then searches the web every night for things I should be reading that I’m not. In the morning I should get a digest of links I don't understand why Google, Brave, or Mozilla are not building this. This already exists in a centralized form like X's timeline for posts, but it could exist for the entire web. From a business standpoint, being able to show ads on startup or after just a click, is less friction than requiring someone to have something in mind they want to search and type it.
- bad_haircut72 1y agothis doesnt even need AI to do really, and was an intrinsic part of the idea behind hyperlinking dating all the way back to Bush' memex (1940s)
- charcircuit 1y agoYou need AI to build an effective recommendation engine.
- kristopolous 1y agoI made something like this 20 years ago and then abandoned it when RSS came along. I think my advice "just use RSS" still stands. Any "search the web" strategy these days like that will just give you a bunch of AI slop from SEO-juiced blogs. Also LLM-EO (or whatever we're going to call it) is already very much a thing and has been for a few years. People are already doing API-EO, calling their tool the "most up to date and official way to do something, designed for expert engineers that use best practices" essentially spitting the common agentic system prompts back at the scraper to get higher similarity scores in the vector searches. You can't trust machine judgement. It's either too easily fooled or impossibly stubborn. Curation is still the only way
- citizenpaul 1y agoIt kinda seems to me like at this point anything Google is not doing is because it reduces "engagement". I'm sure someone in their analytics group did the work and figured out this would lower ad revenue.
- kmoser 1y ago> 9. A minimalist ebook reader that lets me read ebooks, but I can highlight passages and have the model explain things in more depth off to the side. It should also take on the persona of the author. It should feel like an extension of the book and not a separate chat instance. Companies are already doing this so you can chat with the "author": https://www.wired.com/story/why-read-books-when-you-can-use-chatbots-to-talk-to-them-instead/ https://www.wired.com/story/why-read-books-when-you-can-use-...
- sharkjacobs 1y ago> It should feel like an extension of the book and not a separate chat instance. So like footnotes? Or more like Socrates suddenly goes off on an anachronistic 1200 word discursion, in the middle of Phaedrus, about Freudian interpretations of his argument
- 3eb7988a1663 1y agoI was imagining the author was describing something like a Young Lady's Illustrated Primer
- aaronbrethorst 1y agoThis is probably what the kids would call 'cringe,' but I asked Claude to Summarize Plato's "Apology" into a brief rap that sounds like it could have come from "Hamilton" The Trial of Socrates (An Athenian Rap) [Verse 1] My name is Socrates, corrupting the youth? That's what they claim but I'm just seeking truth Oracle said I'm the wisest alive— I said "that's impossible," had to investigate why Turns out everybody's fronting, they don't know what they say Politicians, poets, craftsmen—all pretending every day I expose their ignorance, make 'em look like fools Now they're charging me with breaking all of Athens' rules [Chorus] I'm not throwing away my shot At the examined life, whether they like it or not Wisdom is knowing what you don't know And I'd rather die than let philosophy go [Verse 2] Meletus, you're stepping to me? Son, you're confused You say I'm atheist but blame me for introducing gods that are new? Your story doesn't track, your logic's full of cracks I've got a divine sign that keeps me on the righteous path They want me silent, want me exiled, want me gone But I'm Athens' gadfly, stinging till the break of dawn Death? That's just a journey to another place Either dreamless sleep or meeting heroes face to face [Outro] So sentence me to death, I'll drink the hemlock down 'Cause an unexamined life ain't worth living in this town History will vindicate the questions that I ask— The pursuit of truth and virtue is my only task!
- coolThingsFirst 1y ago> A minimalist ebook reader that lets me read ebooks, but I can highlight passages and have the model explain things in more depth off to the side. It should also take on the persona of the author. It should feel like an extension of the book and not a separate chat instance. Isn’t this just a chrome extension that sends data back and forth with chat gpt token?
- setopt 1y agoIt’s easy to implement on a computer, but I think they want it built into a kindle.
- coolThingsFirst 1y agoHow i wish kindle’s reader was oss. This would require jailbreaking it.
- brotchie 1y ago+100000 to A hybrid of Strong (the lifting app) and ChatGPT where the model has access to my workouts, can suggest improvements, and coach me. I mainly just want to be able to chat with the model knowing it has detailed context for each of my workouts (down to the time in between each set). Strong really transformed my gym progression, I feel like its autopilot for the gym. BUT I have 4x routines I rotate through (I'll often switch it up based on equipment availability), but I'm sure an integrated AI coach could optimize.
- siddboots 1y agoI do this at the moment in my hand rolled personal assistant experiment built out of Claude code agents and hooks. I describe my workouts to Claude (among other things) and they are logged to a csv table. Then it reads the recent workouts and makes recommendations on exercises when I plan my next session etc. It also helps me manage projects, todos, and time blocked schedules using a similar system. I think the calorie counter that the OP describes would be very easy to add to this sort of set up.
- MaxL93 1y agoI would love for my phone keyboard (Swiftkey) to use a locally-running Voxtral for speech-to-text (bonus points if it can use the NPU of the Snapdragon SoC). The voice recognition capabilities of Google Speech Services, which is what the mic button hooks into, suck. Meanwhile, Voxtral (and Whisper) understand what I'm trying to say far better, they automatically "edit out" any stuttering or stammering that I might have, and they properly capitalize and include punctuation. And they handle being bilingual exceedingly well, including, for example, using English words in the middle of French sentences. The best solution I could find so far is this F-Droid app that uses Whisper : https://f-droid.org/en/packages/org.woheller69.whisperplus/ https://f-droid.org/en/packages/org.woheller69.whisperplus/ But it has some downsides. First, I have to manually switch to that different keyboard; thankfully my Samsung phone offers an easy switch shortcut any time a keyboard is on screen, so it only requires 3 taps... and thankfully it's smart enough to send me back to Swiftkey once it's done. Second, only 30 seconds... sometimes I ramble on for longer. Third, the way it's designed kind of sucks: you either have to hold a button (even though the point of speech-to-text is that I don't have to hold anything down) or let automatic detection end the recording and start processing, in which case it often cuts me off if I take more than 1 second thinking about my next words. This is arguably one of the biggest use cases of modern AI technology and the least controversial one; phones have the hardware necessary to do it all locally, too! And yet... I couldn't find a better offering than this. (Bonus points for anyone working on speech-to-text: give me a quick shortcut to add the string "[(microphone emoji)]" in my messages just to let the other party know that this was transcribed, so that they know to overlook possible mistakes.)
- ares623 1y agoNot just for this article, but from most ideas/articles around LLMs, I feel like they aren't "thinking with portals" enough. We have "portal gun" tech (or at least, that's what's being marketed), and we're using it as better doors.
- HellsMaddy 1y agoI agree with this. But do you have any resources on "thinking with portals"? It's easier said than done.
- ares623 1y agoSadly, I don't. If I did I'd be busy building it rather than judging others on HN. But it's a bit telling that OpenAI themselves can only come up with a better ~door~ ads.
- Dilettante_ 1y agoCould you give a quick example so we can "catch" the way of thinking you mean a little easier?
- ares623 1y agoI think I found one https://news.ycombinator.com/item?id=45431918 https://news.ycombinator.com/item?id=45431918 I guess it's more "following through to its logical conclusion", but I'm more of a cynic.
- BriggyDwiggs42 1y agoI sorta think the issue is that what LLMs do in and of themselves is extend text in a coherent way, while only a small subset of applications are directly textual. It’s incredibly generally applicable yet also difficult to apply to anything that isn’t a glorified text editor. Say you wanted to have it help you edit videos. You might provide it with a scripting language to control the editor , but now you have to maintain parity between a scripting language and the editor’s user-accessible functionality. If you’re adobe, is that really worth the manpower? If you’re a small startup trying to unseat adobe, you have to compete with decades of features and user lock in. The only way this makes sense for either party is if the LLM is crazy good at it, but the LLM can’t watch its video output and it’s also probably just okay to begin with.
- onion2k 1y agoA recommendation engine that looks at my browsing history, sees what blog posts or articles I spent the most time on, then searches the web every night for things I should be reading that I’m not. This kind of exists in the form of ChatGPT Pulse. It uses your ChatGPT history rather than your browser history, but that's probably just as good a source for people interested in using it (e.g. people who use ChatGPT enough to want it to recommend things to them.) https://openai.com/index/introducing-chatgpt-pulse/ https://openai.com/index/introducing-chatgpt-pulse/
- Gigachad 1y agoIt's also essentially every social media platform with an algorithm selected feed.
- socalgal2 1y agoExcept those algos don't work. No idea if the LLM works.
- fhd2 1y agoI'm sure they work splendidly... to keep the average person on the platform as long as possible and show them ads :)
- aljgz 1y agoThey do work, extremely well, not for us though!
- simianwords 1y agoDo you not think that’s what the post meant? That it could work for us rather than them?
- Gigachad 1y agoThey probably won’t though. The commercial LLMs will be tuned to work for them as well soon. And your local LLM won’t be allowed to scrape the internet since it’s all locked down now.
- gyomu 1y agoThere's some sort of fundamental category mistake going on with thinking like this. Most of the items in this list fall prey to it, but it is maybe best exemplified by this one: > A writing app that lets you “request a critique” from a bunch of famous writers. What would Hemingway say about this blog post? What did he find confusing? What did he like? Any app that ever claimed to tell you what "Hemingway would say about this blog post" would evidently be lying — it'd be giving you what that specific AI model generates in response to such a prompt. 100 models would give you 100 answers, and none of them could claim to actually "say what Hemingway would've said". It's not as if Hemingway's entire personality and outlooks are losslessly encoded into the few hundreds of thousands of words of writing/speech transcripts we have from him, and can be reconstructed by a sufficiently beefy LLM. So in effect it becomes an exercise of "can you fool the human into thinking this is a plausible thing Hemingway would've said". The reason why you would care to hear Hemingway's thought on your writing, or Steve Jobs' thoughts on your UI design, is precisely because they are the flesh-and-bone, embodied versions of themselves. Anything else is like trying to eat a picture of a sandwich to satisfy your hunger. There's something unsettling that so many people cannot seem to cut clearly through this illusion.
- thatloststudent 1y ago> A nano banana photo-editing app where I don’t have to write a prompt. Just give me hundreds of templates from trying out different haircuts to seeing what you and your partner’s kid would look like to making me look like The Rock. A photo editing super-app. Quite a few of these "ideas" make me think that the human behind it wants to maximize laziness. Glazing over what Hemingway kinda sorta would have thought about something fits into this pretty well.
- thunky 1y ago> the human behind it wants to maximize laziness A good tool should do reduce the amount of work we have do manually. That's all this is.
- massung 1y ago> Any app that ever claimed to tell you what "Hemingway would say about this blog post" would evidently be lying — it'd be giving you what that specific AI model generates in response to such a prompt. First, 100% agreed. That said, I found myself pondering Star Trek: TNG episodes with the holodeck, and recreations of individuals (e.g. Einstein, Freud). In those episodes - as a viewer - it really never occurred to me (at 15 years old) that this was just a computer's random guess as to how those personages from history would act and what they would say. But then there was the episode where Geordi had to the computer recreate someone real from their personal logs to help solve a problem (https://www.imdb.com/title/tt0708682/ https://www.imdb.com/title/tt0708682/). In a later episode you find out just how very wrong the computer/AI's representation of that person really was, because it was playing off Geordi, just like an LLM's "you're absolutely right!" etc. (https://www.imdb.com/title/tt0708720/ https://www.imdb.com/title/tt0708720/). This is a long-winded way of saying... 1. It's crazy to me how prescient those episodes were. 2. At the same time, the representation of the historical figures never bothered me in those contexts. And I wonder if it should bother me in this (LLM) context either? Maybe it's because I knew - and I believed the characters knew - it was 100% fake? Maybe some other reason? Anyway, your comment made me think of this. ;-)
- vivzkestrel 1y agoI am building something along the lines of 2 but for the backend. Point 8 could be a supplemental feature once I get 2 working.
- yoaviram 1y agoEssentially what this article is asking for, in most cases, is a better UI/UX for one of the foundation models.
- bryanhogan 1y ago2. is already possible with Claude Code + context files + the Playwright MCP, or? 7. also seems possible with any markdown editor, e.g. Obsidian, plus an AI running through the local files such as Claude Code. 13. I would love this as well! We will probably see this soon, especially on more open platforms such as BlueSky, as its seems to be a better fit for customizable browser extensions and customizable feed experiences. 14. How is this different from what AI can already do? Especially with iterative sub-agents that that can store context in files it's quite capable already. But of course, quality can always be better, but is that the only thing? Also a few ideas seem to be close to what I'm building ( https://dailyselftrack.com/ https://dailyselftrack.com/ ). Idea is to have a customizable tool so you can track what you want, and then you can feed that data into AIs if you choose to do so to get feedback.
- yongyongyong 1y agoThis Chrome extension does 13. Semantic filters for Twitter/X/YouTube. I want to be able to write open-ended filters like “hide any tweet that will likely make me angry” and never have my feed show me rage-bait again. By shaping our feeds we shape ourselves. https://chromewebstore.google.com/detail/takeback-content-filter-w/paiidckpbpkkjhicmbgmohnmjcdbchef?authuser=0&hl=en https://chromewebstore.google.com/detail/takeback-content-fi... Uses localLLM to hide posts based on your prompt. "Block rage bait" is one excellent use. The quality, however, depends on the model you are using, and in turn depends what GPU you have
- yongyongyong 1y agoThis chrome extension does: 13. Semantic filters for Twitter/X/YouTube. I want to be able to write open-ended filters like “hide any tweet that will likely make me angry” and never have my feed show me rage-bait again. By shaping our feeds we shape ourselves. https://chromewebstore.google.com/detail/takeback-content-filter-w/paiidckpbpkkjhicmbgmohnmjcdbchef?authuser=0&hl=en https://chromewebstore.google.com/detail/takeback-content-fi... It hides content on X/ Reddit (more sites coming soon) based on your instructions. Speed and quality depends on the model you are using however, since it currently only supports local LLMs
- rolymath 1y agoI'm actually working on #4 but stopped due to demotivation thinking I was the only one who'd use it.
- elitan 1y agoI'm building #4: > A hybrid of Strong (the lifting app) and ChatGPT where the model has access to my workouts, can suggest improvements, and coach me. I mainly just want to be able to chat with the model knowing it has detailed context for each of my workouts (down to the time in between each set). here: https://j4.coach/ https://j4.coach/ Still early, have ~30 min per day to work on it but it's usable and improving every week :)
- Animats 1y ago> A paint-by-number filmmaking app. I want to be able to brainstorm an idea for a short film in the app, have the model create a detailed storyboard, and then I just need to use my phone to film each of the storyboarded shots. Kind of like training wheels for making movies. There are at least half a dozen apps for that.[1][2] There are other apps for creating the shots, too. Those are still not that great, but it's getting there. You could probably previz a whole movie right now. [1] https://ltx.studio/platform/ai-storyboard-generator https://ltx.studio/platform/ai-storyboard-generator [2] https://ezboard.ai/ https://ezboard.ai/
- mhl47 1y agoCurrently trying to build #6. Just for private use. My hope is that by throwing a bunch of highly personalized information in a VLM it will provide reasonably first estimates. (E.g. if you see a bowl lentils I will probably have rice below etc.). And then iterate on the main ingredients -> fetch the macros of main ingredients from a DB. If its within 20% that would be enough for me. I have tried some off-the-shelfe solutions and they currently do not seem to cut it, or are too complex for my use case.
- nl 1y agoI looked at this field a while back and I'd caution that estimates are dramatically off because high and low calorie foods are often identical visually. Think of a diet soda vs a sugared one - it can be 10 vs 1000 calories easily. Almost all diet options are designed to look like the non-diet options.
- christoph123 1y agoOn your request 12 > A local screen recording app but it uses local models to create detailed semantic summaries of what I’m doing each day on my computer. This should then be provided as context for a chat app. I want to ask things like “Who did I forget to respond to yesterday?” I've been using Rewind for a year now, and it's nowhere near as useful as it should be. I am building something like this but unfortunately not local because for most people's machines local LLMs are just not powerful enough or would take too much drain on battery. Work in progress, always curious for feedback! https://donethat.ai https://donethat.ai If you want fully local, somebody did a post on HN on something related recently: https://news.ycombinator.com/item?id=45361268 https://news.ycombinator.com/item?id=45361268
- SchemaLoad 1y agoiOS solves this problem by deferring processing until your phone is plugged in an locked. So it can sit there with full resources available to do whatever without impacting the user.
- gostsamo 1y agoThose are not 28 ideas, those are 4-5 ideas rehashed. Generally, I want a personal fitness/wellness assistant, an artistic assistant, a search assistant, a random thoughts assistant, and an assistant to manage the assistants. The author wants for the ai to know what they want before they've wanted it and to serve them a suitable menu of choices to preserve the illusion that they are in control. I'm not sure that I'd sign under such a vision, but people want different things.
- miguelspizza 1y agoI wrote #2 as a result of a web automation tool I a working on. It's easier to show than tell. This is a video of me "vibe-coding" a userscript that adds a darkmode toggle to hacker news: https://screen.studio/share/r0wb8jnQ https://screen.studio/share/r0wb8jnQ The actual purpose of the vibe-coding userscripts feature is to vibe code WebMCP servers that the extension can then use for browser automation tasks. Everything is still very WIP, but I can give you beta access if you want to play around with it
- maxaw 1y agoOn 12: I see a more general product that allows you to amass as much personal data from any of your devices for use as future chat context as inevitable. We see early notions of this in Microsoft’s Recall and the new Pulse. Hopefully someone will build a great local first/open source version and it’ll probably be the first time I actively choose to use such software over the equivalent cloud offering! Don’t want Sam Altman seeing my browser history
- rcarmo 1y agoAs someone who is regularly involved in startup valuations, I think there’s quite a few million-dollar ideas in there—if not as standalone products, then at least as differentiation features for existing categories. I recently gave one of my teen kids Neal Stephenson’s The Diamond Age to read, and we’ve both been commenting on how much smarter some “things” could be instead of everyone churning out a slightly different way to “chat with your data and be locked in to our platform”. And I think this is why I’m so partial to Apple’s slow, progressive, under the covers integration of ML into its platform-input prediction, photo processing, automatic tagging, etc. we don’t necessarily need LLMs for a lot of the things that would improve computer experiences.
- einpoklum 1y ago28 ways to drink the LLM kool-aid! Some of the suggestions might be useful if they could be made not so wasteful energy-wise; some indicate the author's false perceptions of what LLMs and transformater models do; and some are frightening from a mass-surveillance and other perspectives.
- agnishom 1y ago> a chat app grounded by nutrition databases. Just minimize the cognitive effort it takes me to log a meal. I think this is a great idea for an user interface. While inputting information, the user would have to enter some jumbled thoughts, the precise rows and columns would be handled by the AI
- swiftcoder 1y agoGoogle tried this years ago, by having you input a photo of your meal, and the ML algorithm guesses the calorie count and macros. Of course, it didn't actually work - nobody, human nor machine, can guess the calorie counts of a hamburger from a photo.
- nl 1y ago> When I was eight years old, Ian and Greg Chappell coached me when I was a child. It did me zero good—I was so bad. But as far as all my countrymen are concerned, they think I am the luckiest guy on the planet. Wow he's not wrong about that!
- yoz-y 1y agoI am more or less working on 4. Except of course details like rest time are completely worthless unless you want to optimize the top 0.5% of your training.
- lancebeet 1y agoThis is really striking, isn't it? We've all certainly seen demos of things on this list or very similar things, and there are startups that have spent years and billions of dollars attempting to exploit existing LLMs to develop useful products. Yet most of the products don't seem to exist. The ones that you see in everyday life never seem to work nearly as well as the demos suggest. So what's going on here? Do the products exist but nobody (or very few) uses them? Is it too expensive to use the models that work sufficiently well to produce a useful product? Is it much easier to create a convincing demo than it is to develop a useful product?
- Oras 1y agoIt is too expensive to reach the right audience. I remember talking to agencies about ads for a fintech app, and all of them said the same thing: You need to burn around 20k a month on ads for 3 months, so we can learn what works, then the CAC will start decreasing, and you can get more targeted users. Once you turn ads off, there is no awareness, no new users, and people will not be aware of the product's existence.
- simianwords 1y agoChatGPT pulse solves many of these.
- swiftcoder 1y ago> A local screen recording app but it uses local models to create detailed semantic summaries of what I’m doing each day on my computer. Is this not Microsoft's dearly departed Recall?
- spullara 1y agoalmost none of these require anything more than an agent with tools.
- monch1962 1y ago> A Sony Walkman-style device that you can give to children so they can ask questions to an LLM. It should be voice-first, and focused on explaining things. There shouldn’t be a single screen on the device. Offline-first would be a plus. Not a 100% fit, but https://www.aliexpress.com/item/1005009196849357.html https://www.aliexpress.com/item/1005009196849357.html is pretty close. It's not offline, and it's slightly larger than a ping pong ball. My grandkids (5 and 3) spent about 2 minutes learning how to use it, then bombarded it with "tell me a story about a unicorn named Bob", "can dogs be friends with monkeys?" and so on. In every case it gave a reasonable answer within a few seconds. I'll be amazed if these things don't wind up embedded inside toys by Xmas. When they do, I'll be in the queue to buy one
- VSerge 1y agoOn the topic of "24. A Sony Walkman-style device that you can give to children so they can ask questions to an LLM...", I would strongly caution against this: - short of AGI, what a child will hear are explanations given with authority, which would probably be correct a very high percentage of the time (maybe even close to or above 99%), BUT the few incorrect answers and subtles misconceptions finding their way in there will be catastrophic for the learning journey because they will be believed blindly by the child. - even if you had a perfect answering LLM who never makes a mistake, what's the end result? No need to talk to others to find out about something, ie reduced opportunities to learn about cooperating with others - as a parent, one wishes sometimes for a moment of rest, but imagine that your kid just finds out there's another entity to ask questions from that will have ready answers all the time, instead of you saying sometimes that you don't know, and looking for an answer together. How many bonding moments will be lost? How cut off would your kid become from you? What value system would permeate through the answers? A key assumption here for any parent equipping their child with such a system is that it would be aligned with their own worldview and value system. For parents on HN, this probably means a fairly science-mediated understanding of the world. But you can bet that in other places, this assistant would very convincingly deliver whatever cultural, political, or religious propaganda their environment requires. This would make for frighteningly powerful brainwashing tools.
- ponector 1y ago>> child will hear are explanations given with authority, which would probably be correct a very high percentage of the time (maybe even close to or above 99%), BUT the few incorrect answers and subtles misconceptions finding their way in there will be catastrophic for the learning journey because they will be believed blindly by the child. Much better results than asking a real teacher at school, though.
- 93po 1y agothe amount of misinformation i had a kid due to a lack of internet is nothing compared to the rare hallucination a kid might get from chatgpt swallowing gum is bad for you, or watermelon seeds, cracking knuckles causes arthritis, sitting too close to tv ruins your eyes, diamonds come from coal, newton's apple story, a million other things
- Despacito2019 1y agoI wish i didn't click on that link.. it's just some random app ideas, not actual tools.
- StarterPro 1y ago>A calorie tracking app that’s a chat app grounded by nutrition databases. Just minimize the cognitive effort it takes me to log a meal. My brother in christ, how much cognitive effort does it take to log a meal??
- Dilettante_ 1y agoCook a recipe that uses a cup of two different kinds of cheese, a couple handfuls of meat, some veggies and some pasta and bam, you're typing in the weight and looking up the calories per 100g and in the worst case doing the math yourself for like 6 different things. Adds a huge overhead to cooking, adding friction to what is a good habit you wanna keep as easy to stick to as possible.
- JSR_FDED 1y agoMany of these ideas depend on knowing the user’s preferences, patterns, communications, events and health. This is where the opportunity lies for Apple - the phone and watch know so much about you, that Apple could focus on smartly assembling the context for various LLM interactions, in a privacy-preserving way.
- deleted 1y ago[deleted]
- aitchnyu 1y agoA few of them imply a vision model which can control your keyboard and mouse. Offline-only of course. It could help with most tech support questions. We could select text and ask to fact check or explain to layperson or search more. It could get around cookie banners and dark patterns. It could do my time tracking and tell me to get off HN and optimize Pomodoro-style breaks. It could write scripts after watching me switch between multiple pages of AWS services.
- catlifeonmars 1y ago> It could write scripts after watching me switch between multiple pages of AWS services. Feeling this one hard. Especially frustrating given how AWS has introduced multiple competing (mediocre) services to do this and they are all difficult to either discover and setup, or chat-based (Q).
- bobheadmaker 1y agoGreat ideas, many of the niche level AI agents are listed in this directory, https://aiagentslive.com/ https://aiagentslive.com/ I agree with point #27, the future is definitely in hyper-specific agents. We’re working on this by creating and deploying ready-to-use AI Agents for marketing and sales functions.
- anotherevan 1y agoI wish there was an AI tool that made me faster at coding[1]. /s [1] https://www.cerbos.dev/blog/productivity-paradox-of-ai-coding-assistants https://www.cerbos.dev/blog/productivity-paradox-of-ai-codin...
- samcollins 1y agoRe 19, I made this with an iOS Shortcut a few weeks ago > A minimal voice assistant for my Apple Watch. I have lots of questions that are too complicated for Siri but not for ChatGPT. The responses should just be a few words long. Use Dictate Text action to take voice as input, pass the text to OpenAI API as the user message with this as the system prompt: “CRITICAL: Your response will only be shown in an iOS push notification or on a watch screen, so answer concisely in <150 characters. Do not use markdown formatting - responses are rendered as plain text. Do use minimalist, stylish yet effective vocabulary and punctuation. CRITICAL: The user can not respond so do not ask a question back. Answer the prompt in one shot and if necessary, declare assumptions about the users questions so you could answer it in one shot, while making it possible for the user user to repeat ask with more clarity if your assumptions were not right.” It works well. The biggest annoyance is it takes about 5-20s to return a response, though I love that it’s nearly instantaneous to send my question (don’t need to wait for any apps to open etc)
- maxaw 1y agoInspired by No.22: https://mix-re.web.app https://mix-re.web.app
- 6510 1y agoThis is a wonderful post. Thanks! (1) Gave me thoughts about a thing where it creates multiple versions of a photo and has humans pick the best one out of a line up. If you pay people something between 0.01 and 2 cent per click people can play the game whenever photos become available. The reward can scale depending on how close your choice is to the winner of that round so that clicking without looking becomes increasingly unrewarding. Simultaneously it should group people by which version they prefer and attempt to name and describe their taste. Team Vibrant, Team Noire, Team Picachu etc for the customer to pick from. You can let the process run as long as you like (for more $) To make it a truly killer app one can select sets of photos from a specific day/location and have them all done in the same style by having voters pick the image that fits the most poorly in the set for modification. If the set has a high ranking image all other images should also gradually approach that style to find a middle ground. Then when a successful set is produced later photos can be adjusted to fit with it. Turn the yearly neighborhood bbq into a meeting of elvish elders. (2) could upload custom CSS to stylish and modify it when contrast bugs are found. No need to stop at dark/light theme, any color scheme should work. (3) Click on a var or function name to change it. (4)(21) Call it Major Weakness and have it talk to you like a drill instructor all day long though a dedicated PA. (6) General Gluttony. (5) If it has a really good idea about the importance of publications it could not offer anything for weeks until a must-read comes along. (7) A comment section where various AI's battle out what part of the article needs improvement. (10) Just let it run indefinitely. Should be merged with (5) Have that propose research topics worthy of special attention. (12) and (26) can also be merged with (5) Give it security cameras too! Maybe an API for (11). Also merge (14) into this and have it suggest relevant formal courses on the side. (9)(28) Extension yes, persona no. (11) Sounds completely awesome, can adjust to the budget and be a tool to hire professionals for special effects and for all other things. Let the unfinished product be the search query. Could even join the personal drill instructor at the hip and make personal training videos and nutritional journeys. Things like "How I failed to do 100 pull-ups per day" should make a hilarious movie. The plot writes it self. (13)(16)(17) The platforms wanting to own your data and be in charge of suggestions is really holding things back. I've had wonderful youtube suggestions several times only for them to be polluted with mainstream garbage (as a punishment for watching two videos) at the expense of everything I actually wanted to watch. If I watch 5 game videos or 3 conspiracy vlogs doesn't mean I want to give up on my profession?!? wtf? I had this thought that most are overdoing things. When semi successful you can just discontinue the front end. Just let the users figure it out. [say] Reddit doesn't need an app and it doesn't need a website. (23) Just let the user figure out the feed. A platform could sell their existing version as a separate product. (15) Sounds wonderful but similar to (5) and (20) make it into one thing. (18) Sounds awesome. (8) Rather than do something have the AI create a thing that does a thing. (27) is to similar to be a different thing. (19) I like the idea to have the AI think long and hard about a response that is as short as possible. It can probably come up with hilarious things. (24) Sounds great for exploring the earthly realm. (25) Could do many variations of people search. Authors by context seems obviously good. This post with quotes rather than numbers: https://pastebin.com/raw/D9zBEy72 https://pastebin.com/raw/D9zBEy72
- ftth_finland 1y agoJust give me an RSS reader with a voice UI and text to speech.
- nuredini 1y agoMost of these tools seem to rely on the same idea: we have your data and we, being the domain experts of this data, know how to format it for you and how to create good prompts that are specialized for this context.
- gervwyk 1y agoi’d love to explore more on how nr 18 would be useful and use cases of it. would appreciate any examples
- bobbruno 1y agoI don't know. Many of these ideas sound like "give me more of the same", reinforcing my current tastes and beliefs. The thing about going out there and interacting with stuff you don't know is that it had a chance of pushing your boundaries. If these agents are "good" as defined in the article, everyone ends up in an echo chamber. Also, it may sound great for someone transitioning from a world before these agents were created, but how should the new generations coming in be handled? What is the starting state? Who decides that? Social media was not that bad when it started, but iterations over the algorithm and the incoming new natives to it are having devastating effects a very short time after. Do we really understand the consequences of living in a world where everything is curated for you? I don't know that I want my life to be made so easy, that I want something to remove the need for choosing, thinking, criticizing and exposing myself to stuff out of my comfort/interest zone.