Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
storus
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
storus
2mo ago
I hope they somewhat fixed the hallucination and forgetting plagued V4 previews and that it wasn't just benchmaxxed but the numbers hold in reality. Then it would be my choice for 2x DGX Spark or 2x RTX Pro 6000.
32.
▲
by
storus
2mo ago
What really helped me was to run any .md/prompts through an LLM to find contradictions, duplicates or ambiguities, repeatedly. That led to agents much better following instructions.
33.
▲
by
storus
2mo ago
There was a presentation at CS 25 Transformers United about some phenomena like chain-of-thought emerging only after training LLMs with at least 1T tokens and only in LLMs of certain size. https://youtu.be/tVtOevLrt5U?t=923
34.
▲
by
storus
2mo ago
The speech he had at congress didn't sound very flattering towards the beginning of his today's statement: https://www.youtube.com/watch?v=_i91NSOyxHM He didn't mention outright banning open source LLMs, just
35.
▲
by
storus
2mo ago
If your brain has a local metabolic failure leading to the inability to replete neurotransmitters, how taking a walk is a solution to your problem? You need to improve energetics first. Or what if some virus/auto-antibody is occupying
36.
▲
by
storus
2mo ago
This feels like Nvidia's strategy doesn't really care about the few companies they invested to pushing frontier AI, but about scaling up the amount of GPUs/datacenters needed; open models forcing more GPU purchases would driv
37.
▲
by
storus
2mo ago
Aren't depressions coming with side-effects like inability to do anything at all? The block is somewhere else to overcome before even thinking about exercise and the brain might have a good reason to block it (e.g. some local energetic
38.
▲
by
storus
2mo ago
What would be the current best method to fine-tune it for my own specific agentic tasks? LoRA + DPO? GRPO? Something else?
39.
▲
by
storus
2mo ago
Getting 404 on the OP's link. Does it mean it got banned or self-censored in the meantime?
40.
▲
by
storus
2mo ago
It's not really Markov chain as you need full P(x_t|x_{t-1}, x_{t-2}... x_1) instead of just P(x_t|x_{t-1}).
41.
▲
by
storus
2mo ago
That's one way to pull the ladder if regulatory capture fails...
42.
▲
by
storus
2mo ago
Dunno, now the workflow is like agent makes code changes, ruff complains, agent fixes complaints at the cost of code bloat, agent makes a PR, another agent reviews the PR, agent makes changes, PR is approved and merged. Nobody reads PRs any
43.
▲
by
storus
2mo ago
No, they are using long optical cables attached to drones. The battlefield looks like a giant spider web afterwards.
44.
▲
by
storus
2mo ago
One could argue that the interesting parts of simulation require higher precision. When you are in stable conditions, that's where your intuition might be sufficient already; once you hit those ugly parts then you start requiring as
45.
▲
by
storus
2mo ago
FP64 is not that precise; proper simulations usually need much higher precision. Even ancient Intel could do 80-bit FP.
46.
▲
by
storus
3mo ago
Borland was way ahead of its time, nowadays there is no non-AI way to build a whole UI app in on afternoon like with Delphi/C++ Builder. Pity their greed prevented them from dominating via simplicity.
47.
▲
by
storus
3mo ago
I doubt they did any distillation as Hinton defined it (requiring logit access). They most likely ran a bunch of prompts/conversations and captured the results. Those conversations already missed thinking tokens, replaced by some confu
48.
▲
by
storus
3mo ago
Old Reddit is the only way to read text there for me so if that's login-walled or gone I am done with them as well.
49.
▲
by
storus
3mo ago
Isn't it right next to the largest agglomeration in Poland? Also close to two (maybe three) other countries connected via highways?
50.
▲
by
storus
3mo ago
DeepSeek V4 hallucinates like crazy and often forgets explicitly mentioned parts of the context. I guess compressing tokens and cherry-picking attention comes at a cost.
51.
▲
by
storus
3mo ago
I would rather see them releasing 3.7-27B, 3.7-122B or their 3.8 versions. Qwen/QwQ were always about the best available local inference at home.
52.
▲
by
storus
3mo ago
Africa is going to be full of old Versaces, Balenciagas, Guccis and Valentinos.
53.
▲
by
storus
3mo ago
Windows Phone aesthetics was repulsive to most people at that time; we finally got TrueColor 4k screens and all MS could do was to use 10 colors everywhere and start the flat fad that destroyed UX on most systems. What a waste.
54.
▲
by
storus
3mo ago
Capacitive screens were out of possibility for them as Apple bought 2 year production in advance, a trick Tim deployed repeatedly in many areas. MeeGo had a chance but US funds didn't want to allow a state where an EU company would rul
55.
▲
by
storus
3mo ago
Well, not exactly, for example if I search for LEGO I get the original but not the 10x cheaper compatible knockoffs.
56.
▲
by
storus
3mo ago
They are orthogonal; preference optimization like RLHF can be done on the base model which can later be quantized, or it could be done on a new LoRA that is then converted to QLoRA.
57.
▲
by
storus
3mo ago
Can you do the inverse as well? Like Amazon but only the knockoffs?
58.
▲
by
storus
3mo ago
One could have anticipated XBox getting slowly destroyed by appointing a young clueless outsider as its boss.
59.
▲
by
storus
3mo ago
First they tried to approve software patents during an agriculture and fisheries council session, now they are bending procedural rules to hack it in before summer vacations. Some weird form of democracy™.
60.
▲
by
storus
3mo ago
Math is static, CS is dynamic. In math you describe static idealized "worlds", in CS you look at any discrete dynamic process in detail via algorithms. Many folks doing math can't understand algorithms, and many coders can&#x
More ›