Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
helloplanets
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
31.
▲
by
helloplanets
2mo ago
With the new Ultracode modes in Claude Code and Codex it's been taken to the next level. I mean, multi agent systems have been available for a long time, but the newest models seem to be much more RL'd for them. Both GPT-5.6 and F
32.
▲
by
helloplanets
2mo ago
People don't realize the exponential diffictlty curve with juggling. The highest amount of balls ever juggled is 11.
33.
▲
Video Game Critic
(videogamecritic.com)
3 points
by
helloplanets
3mo ago
|
0 comments
34.
▲
by
helloplanets
3mo ago
The actual part on fine-tuning seems very short in the article. Did I miss a page where they have examples of fine-tuning it for different niche use cases? Optimizing models to be fine-tuned is an amazing direction, but just makes me wonder
35.
▲
Weirdest Commits in Linux Kernel Git History
(destroyallsoftware.com)
2 points
by
helloplanets
3mo ago
|
0 comments
36.
▲
by
helloplanets
3mo ago
Shouldn't the valuation be in Bs instead of Ms?
37.
▲
by
helloplanets
3mo ago
Yes, it is a 10x markup on the API prices. Depending on whether you factor in cooling costs, data center staff, etc. Or GPU costs and the electricity the GPUs are using only. Either way, inference is very much where the money is made, train
38.
▲
by
helloplanets
3mo ago
True, it'd be a whole other situation if the tokens limits were cumulative. I guess it would all come down to whether their Claude Code subscription plans are turning in a profit or not. At least for the segment of 20$ subscribers who
39.
▲
by
helloplanets
3mo ago
The subscription based plans are heavily subsidized, but the direct API inference pricing (which larger companies need to pay) is profitable. Using a full Claude Max 20x plan to 100% of weekly usage would easily cost you 2k through the API.
40.
▲
by
helloplanets
3mo ago
I just did this on one .claude directory and >20% of the answers there included some variation of "real", "actual", "exact", "honest", "genuine", "valid", "true". ~15%
41.
▲
by
helloplanets
3mo ago
It's kind of offputting how much Anthropic models these days keep repeating "real", "genuine" and "honest". They've RL'd that way over the top.
42.
▲
California Institute for Machine Consciousness – Research Program Whitepaper [pdf]
(cimc.ai)
1 points
by
helloplanets
3mo ago
|
0 comments
43.
▲
by
helloplanets
3mo ago
Definitely not just a classifier layered on top, although there is one of those as well. Pretty sure it's different post-training / finetune run and the model weights are different between them. Which is bound to have effects on t
44.
▲
Can you run every line of code in Super Mario Bros.? [video]
(youtube.com)
1 points
by
helloplanets
3mo ago
|
0 comments
45.
▲
by
helloplanets
3mo ago
But this guy's been all over the news? He's not some made up fantasy person with absolutely no real world footprint. Even if his website is sloppy. https://en.wikipedia.org/wiki/Peter_Young_(activist)
46.
▲
kinopio.club
(kinopio.club)
3 points
by
helloplanets
3mo ago
|
0 comments
47.
▲
by
helloplanets
3mo ago
Not at all. This looks just like someone trying to make a quick buck, hyping their product up with bad benchmarks.
48.
▲
by
helloplanets
3mo ago
> Both conditions used GitHub Copilot (Claude Sonnet 4.5 or Haiku 4.5, depending on study) running in VS Code within isolated Docker containers. The only difference was Mouse tool availability. ( https://hic-ai.com/papers&
49.
▲
Moby Dick Workout (2022)
(hogbaysoftware.com)
109 points
by
helloplanets
3mo ago
|
34 comments
50.
▲
by
helloplanets
3mo ago
> These are my personal beliefs, not those of Nym. Why are you posting this on your company's site, littered with ads for the company's product? Post it on a personal blog, or just say that these indeed are the company's b
51.
▲
by
helloplanets
3mo ago
Doesn't make sense to fixate on LLMs and not the actual Transformer/attention foundation. The Transformer/attention architecture is the breakthrough, not LLMs. Especially the RLHF chat paradigm is 100% a byproduct. Which is e
52.
▲
by
helloplanets
3mo ago
Tangential, but I'm pretty sad about EU having absolutely nothing in the actual SotA LLM market. Especially given the recent events of US completely restricting the actual SotA models. Has this been just pure lack of funding and infra?
53.
▲
by
helloplanets
3mo ago
If that ain't getting steganographically tagged...
54.
▲
by
helloplanets
3mo ago
Dario's been openly talking how worried he is about China and labs getting synthetic training data off their models, for years. Most recently in relation to "Mythos level" capabilities. Not really distillation, just syntheti
55.
▲
by
helloplanets
3mo ago
The issue is that using Claude Code is an easy compromise for most to make, when you get to use the models 10x cheaper than through API pricing with a custom harness. The cheap tokens are the product.
56.
▲
Universities are studying how they lost the public's trust
(theatlantic.com)
6 points
by
helloplanets
3mo ago
|
7 comments
57.
▲
by
helloplanets
3mo ago
Slide number 55 is a beauty.
58.
▲
by
helloplanets
3mo ago
Which great writers are you thinking of here? True outsider art is very rare afaik.
59.
▲
by
helloplanets
3mo ago
I actually thought about that while writing the original comment as well. For Emma, Forever Ago is one of my all time favorite albums, good example of raw emotion with no need for any bells or whistles. The big thing there is, that he alrea
60.
▲
by
helloplanets
3mo ago
I don't believe this is how great music usually comes about, not even Techno. It's missing the other essential piece. Being influenced by and completely immersed in a niche of other brilliant people. (The most extreme example of t
More ›