4 ms·
I have a mini PC with an n100 CPU connected to a small 7" monitor sitting on my desk, under the regular PC. I have llama 3b (q4) generating endless stories in
by behohippy 2y ago
I have a mini PC with an n100 CPU connected to a small 7" monitor sitting on my desk, under the regular PC. I have llama 3b (q4) generating endless stories in different genres and styles. It's fun to glance over at it and read whatever it's in the middle of making. I gave llama.cpp one CPU core and it generates slow enough to just read at a normal pace, and the CPU fans don't go nuts. Totally not productive or really useful but I like it.
- bithavoc 2y agothis is so cool, any chance you post a video?
- behohippy 2y agoJust this pic: https://imgur.com/ip8GWIh https://imgur.com/ip8GWIh
- Dansvidania 2y agothis sounds pretty cool, do you have any video/media of it?
- Uehreka 2y agoDo you find that it actually generates varied and diverse stories? Or does it just fall into the same 3 grooves? Last week I tried to get an LLM (one of the recent Llama models running through Groq, it was 70B I believe) to produce randomly generated prompts in a variety of styles and it kept producing cyberpunk scifi stuff. When I told it to stop doing cyberpunk scifi stuff it went completely to wild west.
- o11c 2y agoYou should not ever expect an LLM to actually do what you want without handholding, and randomness in particular is one of the places it fails badly. This is probably fundamental. That said, this is also not helped by the fact that all of the default interfaces lack many essential features, so you have to build the interface yourself. Neither "clear the context on every attempt" nor "reuse the context repeatedly" will give good results, but having one context producing just one-line summaries, then fresh contexts expanding each one will do slightly less badly. (If you actually want the LLM to do something useful, there are many more things that need to be added beyond this)
- dotancohen 2y agoSounds to me like you might want to reduce the Top P - that will prevent the really unlikely next tokens from ever being selected, while still providing nice randomness in the remaining next tokens so you continue to get diverse stories.
- janalsncm 2y agoGenerate a list of 5000 possible topics you’d like it to talk about. Randomly pick one and inject that into your prompt.
- coder543 2y agoSomeone mentioned generating millions of (very short) stories with an LLM a few weeks ago: https://news.ycombinator.com/item?id=42577644 https://news.ycombinator.com/item?id=42577644 They linked to an interactive explorer that nicely shows the diversity of the dataset, and the HF repo links to the GitHub repo that has the code that generated the stories: https://github.com/lennart-finke/simple_stories_generate https://github.com/lennart-finke/simple_stories_generate So, it seems there are ways to get varied stories.
- fi-le 2y agoI was wondering where the traffic came from, thanks for mentioning it!
- keeganpoppen 2y agooh wow that is actually such a brilliant little use case-- really cuts to the core of the real "magic" of ai: that it can just keep running continuously. it never gets tired, and never gets tired of thinking.
- ipython 2y agoThat's neat. I just tried something similar: FORTUNE=$(fortune) && echo $FORTUNE && echo "Convert the following output of the Unix `fortune` command into a small screenplay in the style of Shakespeare: \n\n $FORTUNE" | ollama run phi4
- watermelon0 2y agoDoesn't `fortune` inside double quotes execute the command in bash? You should use single quotes instead of backticks.
- droideqa 2y agoThat's awesome!