Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
thorum
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
1.
▲
by
thorum
13d ago
Sure, but the point is that the labs use more powerful internal models for research work, not public models. Public models tend to lag the internal frontier by a decent margin, and are constrained in other ways by monitoring. It’s just not
2.
▲
by
thorum
13d ago
> The researchers asked Anthropic’s Claude Opus 4.8, running on open-source software called OpenClaw Meanwhile, Navier–Stokes was solved by an internal model significantly more capable than Astra (and therefore more capable than Mythos&#
3.
▲
by
thorum
13d ago
Why does it have to be about that? The author says they agree with the letter in the first paragraph.
4.
▲
by
thorum
16d ago
FWIW, I flagged and downvoted both of your comments. The article is a very interesting overview of the space and recent developments, and having to scroll past multiple paragraphs about your low attention span and reading difficulties to ge
5.
▲
by
thorum
18d ago
It reminds me of the Cognitive Dark Forest hypotheses recently shared here: > “You are creating your cool streaming platform in your bedroom. Nobody is stopping you, but if you succeed, if you get the signal out, if you are being noticed
6.
▲
by
thorum
24d ago
There’s nothing surprising or confusing here. Outside of tech communities like HN, I see anti-AI rhetoric everywhere. A very large number of humans see AI as a purely evil and parasitic thing, no matter how it’s used. It’s an especially co
7.
▲
by
thorum
1mo ago
Prior discussion from when the service stopped accepting new customers in July: https://news.ycombinator.com/item?id=48803886
8.
▲
by
thorum
1mo ago
When I have some important problem I don’t know how to solve, I of course think about it all the time. Every now and again I’ll think of something that seems extra useful and write it down. Eventually I’m looking at a whole page of good ide
9.
▲
by
thorum
1mo ago
Unfortunately once you accept this premise, the next logical conclusion is that the physical world and analog systems are just as vulnerable to exploitation by embodied AI.
10.
▲
by
thorum
2mo ago
Free users have had access to reasoning for a while. o4-mini and the initial GPT-5 launch both included reasoning modes for free users. They took away the button a few months ago and are now putting it back.
11.
▲
by
thorum
3mo ago
Interesting that all four models converge on such similar designs, for such short prompts.
12.
▲
by
thorum
3mo ago
Interesting read! Creating tests is highlighted as something Claude did well, but it strikes me that all the weaker rejected solutions could have been avoided if it were really good at designing intelligent tests for itself. For example,
13.
▲
by
thorum
3mo ago
The “correct”, elegant way for AI to interact with existing software would take decades and billions of dollars to build. Someone would have to do the hard work of building new APIs, solving decades of accessibility issues, etc. Or you can
14.
▲
by
thorum
3mo ago
The team with the most star power and hype tends to attract the best young talent. If the next big breakthrough in AI comes from Anthropic, good chance it comes from some genius you’ve never heard of who decided to work there because of [fa
15.
▲
by
thorum
3mo ago
I wish them all the best and hope they succeed, but can’t help but suspect they’ve fallen into deep LLM psychosis. Even if you assume they can build this thing and it works as described and then get past all the regulatory hurdles, the scal
16.
▲
by
thorum
3mo ago
I actually think “explore Claude’s understanding of colors” is an interesting concept. A lot of fascinating cultural information gets compressed into LLMs.
17.
▲
by
thorum
4mo ago
Unfortunately for the people mad about this, I predict the only thing they will accomplish by pressuring the rsync maintainers, is to discourage everyone else from responsibly disclosing their use of AI. You’re just going to make people dis
18.
▲
by
thorum
4mo ago
Somewhat useless article. To summarize, we have anecdotes suggesting they may work but no one has figured out how to prove or disprove it in a study, and the author has some doubts. Meanwhile supplements can be dangerous if you take too muc
19.
▲
by
thorum
4mo ago
I was really unimpressed by the free Codex (for nodejs/react dev). I think it must be using a less powerful model or they’re limiting it in some other way.
20.
▲
by
thorum
5mo ago
You’re probably right in a literal technical sense, but a very large number of people (maybe most?) would choose “no” if properly informed and asked for consent, and lots of people are morally opposed even in principle to downloading a larg
21.
▲
The physics slop that YouTube wants me to make [video]
(youtube.com)
2 points
by
thorum
5mo ago
|
0 comments
22.
▲
by
thorum
5mo ago
The models are primitive right now, but we’re clearly heading toward “AI as sound synthesis, human as artist” - much like how producers currently use a DAW to assemble premade loops and sounds from Splice, but with the producer now able to
23.
▲
by
thorum
5mo ago
Isn’t this a permissions issue? Your “opt out” is using a GitHub access token that doesn’t allow it to happen.
24.
▲
by
thorum
6mo ago
I have the opposite experience: random HN/Reddit comments saying “this sucks” or “whoa this is a huge improvement” are the only benchmark that means anything. Standard benchmarks are all gamed and don’t capture the complexity of the re
25.
▲
by
thorum
6mo ago
Stars have been useless as signals for project quality for a while. They’re mostly bought, at this point. I regularly see obviously vibe-coded nonsense projects on GitHub’s Trending page with 10,000 stars. I don’t believe 10,000 people have
26.
▲
by
thorum
6mo ago
Good day for Kling.
27.
▲
by
thorum
7mo ago
Ape thinking is a cognitive practice where a human deliberately solves problems with their own mind. Practitioners of ape thinking will typically author thoughts by thinking them with their own brain, using neurons and synapses. The term wa
28.
▲
by
thorum
7mo ago
Their design approach wasn’t particularly unusual, so I’m not sure what that sentence means. I do miss the days when technical reports were clear and concise. This one has some interesting information, but it’s buried under a mountain of em
29.
▲
by
thorum
8mo ago
AI for help figuring things out and Timeshift for when you accidentally break something. One reboot and it’s fixed.
30.
▲
by
thorum
8mo ago
> but the number of problems requiring deep creative solutions feels like it is diminishing rapidly. If anything, we have more intractable problems needing deep creative solutions than ever before. People are dying as I write this. We’ve
More ›