Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
BoredomIsFun
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
61.
▲
by
BoredomIsFun
1mo ago
/r/localllama, /r/sillytavernai for leads and then personal vibe check.
62.
▲
by
BoredomIsFun
1mo ago
> Are there any creative-writing LORAs published for open-weight models? Oddly enough standalone LoRA adapters are very popular in image generation world and utterly unpopular in LLM world - there it is customary just to distribute full
63.
▲
by
BoredomIsFun
1mo ago
Just buy 2x5060ti's and run local Qwen 3.8. Not Claude of course, but a good backup anyway.
64.
▲
by
BoredomIsFun
1mo ago
Stiff, repetive style is often a result of very small, less than 0.5 sampling temperature, very small top-k, very high min-p etc. Most of "normies" (wrt to /r/localllama and /r/sillytavernai) never tweak the s
65.
▲
by
BoredomIsFun
1mo ago
There are many different ways to make an LLM sound more human-like (the author has actually explicitly mentioned ChatGPT sounds more natural). For their public announcements etc. they might as well used specially trained small 24-32B creati
66.
▲
by
BoredomIsFun
1mo ago
> No it's not useful, no you don't know what you're doing, and no it is not 'proper'. But you're learning. It is sold by PG as something special though. > Maybe just accept that not everyone starts from f
67.
▲
by
BoredomIsFun
1mo ago
> That's what I would call exploring. Then it is not "building llm from scratch" in my book. Just mindlees following instructions. Could be educational yes, but only trivially useful, if you have no bloody idea what you ar
68.
▲
by
BoredomIsFun
1mo ago
We can continue this pointless conversation, in the tone "who you are to tell what is fun to others and whst is not". You'd be impervious to any argument stating that dealing with far beyound someone understanding and requiri
69.
▲
by
BoredomIsFun
1mo ago
> Life is about fun, not extraction, not value, not avoiding waste of time but literally enjoy what you do, nothing is more important (IMO). This is, pardon, demagoguery. There is always "future fun" and "present fun"
70.
▲
by
BoredomIsFun
1mo ago
How would you filter out garbage from your training data, for example? If you are trying to use someone elses corpus, would it be "from the scratch" then?
71.
▲
by
BoredomIsFun
1mo ago
That'd would be a terrible advice if there weren't a plenty of other things "you are not equipped for", but far less daunting both theoretically and practically. Such as, say, convolutional neural networks, or some older
72.
▲
by
BoredomIsFun
1mo ago
I just voiced my opinion. I just think buiding an LLM from the scratch for 17 y.o. is pointless exercise, advising a teenager to do so is borderline irresponsible, and frankly PG is simply virtue signalling here, as LLMs are still trendy,
73.
▲
by
BoredomIsFun
1mo ago
> Why would you tell people that the correct order Because I can?. JK. Because that was my experience, of someone who is 2.5 older than 17? > For some (many?) people, a 'proper' understanding develops _after_ the exploration
74.
▲
by
BoredomIsFun
1mo ago
I do not think it is a proper thing to do for 17 y.o., unless they are exceptionally mathematically gifted, as proper understanding of how LLMs are trained requires a good grasp of calculus, understanding modern OS and SDE tools for proper
75.
▲
by
BoredomIsFun
1mo ago
Are you having some rage fit? Using dashes to construct new words, especially in ironic, humorous context is widely used prasctice in English (which you don't seem to know well either).
76.
▲
by
BoredomIsFun
1mo ago
Are you an LLM? Because, I, human just used them.
77.
▲
by
BoredomIsFun
1mo ago
> I've never heard of the word "hardboiled" referring to anything but eggs. I am glad you've lerned something new.
78.
▲
by
BoredomIsFun
1mo ago
> Straight out of another LLM ??? Not a native speaker, turns out short word is hardboiled. > Please name some books in this style, since it seems to be on the tip of your tongue. Googled for you: https://www.google.com
79.
▲
by
BoredomIsFun
1mo ago
> it's with using an AI to write up the article for human consumption about it. Nothing wrong with it, as soon as it is not obviously AI-generated looking and has good, high quality content.
80.
▲
by
BoredomIsFun
1mo ago
[flagged]
81.
▲
by
BoredomIsFun
1mo ago
Llama 3.2 3B was surprisingly good for general-purpose text manipulation tasks. I'd argue it might be better than many modern tiny models for that.
82.
▲
by
BoredomIsFun
2mo ago
> make AI have experience like we do. We'll probably have to go analog for that. The only known systems that certainly can experience are mammals (with apparently analog brains).
83.
▲
by
BoredomIsFun
2mo ago
> If they AI-generated it, it's because they didn't care about the quality anyway. No? There are far far more AI generated images you see daily, than the awful stuff from 2024 that you immediately discern as AI-generated. You c
84.
▲
by
BoredomIsFun
2mo ago
Just use Ideogram. Locally.
85.
▲
by
BoredomIsFun
2mo ago
These kinds of images are disingenuous. Modern image generating models have way higher quality of output, and in most cases you wouldn't even know it is generated.
86.
▲
by
BoredomIsFun
2mo ago
> I imagine those latter two can be engineered, no? Ultimately yes, but not in modern AI systems.
87.
▲
by
BoredomIsFun
2mo ago
> GTK 5 will drop the X11 back-end entirely, so at some point in the coming decade, GTK applications will gradually stop working. Hopefully someone will come up with Wayland emulation layer for x11 then.
88.
▲
by
BoredomIsFun
2mo ago
I agree. They have serious issues with coherence.
89.
▲
by
BoredomIsFun
2mo ago
> I suppose for dialogue generation in games? I use it to write short sci-fi stories. Life is not only about being an SDE. > Why on Earth would you ever want to write code with a model that is supposedly "jailbroken"? I need
90.
▲
by
BoredomIsFun
2mo ago
yes. you can even parallelize two cards and get 1.7 times the speed.
More ›