Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ddp26
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
ddp26
3mo ago
Yeah, a great developer I know showed me how he could use it to get a safe dev container for Claude Code, in a way that wasn't doable with Docker.
32.
▲
Porting the Moebius 0.2B image model to run in Claude Code on web
(simonwillison.net)
2 points
by
ddp26
3mo ago
|
0 comments
33.
▲
The Wealth of the Richest People in AI
(futuresearch.ai)
4 points
by
ddp26
3mo ago
|
0 comments
34.
▲
by
ddp26
3mo ago
Fair point.
35.
▲
by
ddp26
3mo ago
Is this the trend? There have been various points where one of Anthropic or OpenAI was substantially ahead. Sure, many times they're close, but now doesn't seem like one of them.
36.
▲
by
ddp26
3mo ago
Based on my conjecture that Anthropic is ahead on AI research, and that OpenAI doesn't know how to make Fable-class models.
37.
▲
by
ddp26
3mo ago
I'm going to pre-register my prediction that GPT-5.6 Sol is significantly behind Claude Fable 5, as evaluated by general consensus once time has passed for people to get familiar with both.
38.
▲
by
ddp26
4mo ago
If it's helpful, I'm still holding at July 9 as my median date that Fable gets re-released to Americans, the news of the last 24 hours didn't update the model meaningfully.
39.
▲
by
ddp26
4mo ago
People forget that Meta already did this years ago, before prediction markets became the next big consumer trend for them to chase. The app was called Forecast, and launched in June 2020. (Around the same time that Kalshi and Polymarket lau
40.
▲
World-Modeling the US vs. Anthropic on Claude Fable
(lesswrong.com)
9 points
by
ddp26
4mo ago
|
1 comments
41.
▲
by
ddp26
4mo ago
You mean chatgpt style AI won't help them with those skills? If a human parent or teacher can help with skills like reading, an AI system can too, once it's trained and designed to do so. (How good are humans at teaching reading a
42.
▲
by
ddp26
4mo ago
Yes, I have, but comments are still useful. Don't you think this is an overcorrection?
43.
▲
Conscripting engineers to make training data won't push AI
(futuresearch.ai)
1 points
by
ddp26
4mo ago
|
0 comments
44.
▲
How the US vs. Anthropic Standoff on Claude Fable Will End
(futuresearch.ai)
2 points
by
ddp26
4mo ago
|
1 comments
45.
▲
Claude can miss the motives of politicians
(futuresearch.ai)
10 points
by
ddp26
4mo ago
|
0 comments
46.
▲
Measuring one way AIs lack self-awareness
(futuresearch.ai)
1 points
by
ddp26
4mo ago
|
0 comments
47.
▲
by
ddp26
4mo ago
It is refreshing but perhaps actually not warranted this time? I mostly study web research, and Opus 4.7 was a regression on BrowseComp compared to Opus 4.6, which has been born out by my usage. Opus 4.8 is now much better than either 4.7 o
48.
▲
by
ddp26
4mo ago
What's a definition of AGI you would use, for either time, tasks, value, or job descriptions?
49.
▲
by
ddp26
4mo ago
I linked elsewhere in a comment, Metaculus has AGI forecasts. You can also now use AI forecasters like FutureSearch [1] (disclaimer: I work there), which are competitive with the best humans / teams of humans. And since you aren't
50.
▲
by
ddp26
4mo ago
Thank you! Tok me a few hours, without Claude Code I don't think I would have even attempted this.
51.
▲
by
ddp26
4mo ago
It's been a big problem for a while. The big Metaculus question about AGI has depends on the game "Montezuma's revenge" (!), and there have been many debates about this going back to at least 2020: https://www
52.
▲
by
ddp26
4mo ago
Author here, I agree, I'd be happy if admins want to change the title of this submission to the title of the piece.
53.
▲
by
ddp26
4mo ago
Author here, I drew on this from AI 2027. Yes, a very-expensive AGI, e.g. $1 million / day to simulate a smart human, would be a huge deal. But it would have meaningfully different effects than a cheap one. Here's one definition A
54.
▲
by
ddp26
4mo ago
I see a lot of comments like this is the blocking of prediction markets about politics, war, etc. It's important to remember that ~80% of activity Polymarket and ~90% of Kalshi, by volume, are sports. These are effectively sports betti
55.
▲
Some rare examples of AIs being underconfident
(futuresearch.ai)
6 points
by
ddp26
4mo ago
|
0 comments
56.
▲
by
ddp26
4mo ago
Snake oil is a bit strong, no? I would agree that the burden of proof is on multi-agent systems to show they are outperforming single-agent systems. On my own evals I have seen this, though the improvement may not have been worth the extra
57.
▲
by
ddp26
4mo ago
I like this, though it does leave me feeling more nervous when I really don't know how I'd solve the problem, still requires trust.
58.
▲
History doesn't repeat itself as often as LLMs think
(futuresearch.ai)
1 points
by
ddp26
5mo ago
|
0 comments
59.
▲
by
ddp26
5mo ago
It would have to be an incredibly tiny tax, no?
60.
▲
by
ddp26
5mo ago
What was your use case?
More ›