Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
rbuccigrossi
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
rbuccigrossi
2mo ago
No, I believe you are misunderstanding the quote. Each “question” in ARC-AGI-3 is a game that has hidden rules that you can understand if you look at the game board long enough. This quote means that Opus 5 is looking at the game board, fig
2.
▲
by
rbuccigrossi
3mo ago
The others are an IQ test (TrackingAI), a test of Ph.D. questions across multiple domains (Humanity's Last Exam), and graphical pattern matching (ARC-AGI-2). What's interesting is that while they are rather different in nature (ye
3.
▲
by
rbuccigrossi
3mo ago
The text is my transcript of the video formatted by Claude and with sources added at the end. I do use deep research (across Gemini, ChatGPT, and Claude) to gather background and ideas, and Claude for editing. I started machine learning and
4.
▲
by
rbuccigrossi
3mo ago
Author here. In short there are 4 different metrics: - METR's time horizon - TrackingAI's offline cognitive test - Humanity's Last Exam, and - ARC-AGI-2 that have lasted longer than 2 years (though ARC-AGI-2 is now saturated)
5.
▲
Is AI Progress Real? Four Independent Metrics Show It
(skepticcto.substack.com)
6 points
by
rbuccigrossi
3mo ago
|
6 comments
6.
▲
by
rbuccigrossi
4mo ago
In short, running a $3,299 GMKtek EVO-X2 (Ryzen AI Max+395 with 198 GB) 24/7 with the Gemma 4 26B-A4B model, being as generous as possible, only saves you $1,279.07/year in inference costs. (120 t/s for output tokens at $0.34
7.
▲
Local AI Hardware: Break Even in 2.6 Years?
(skepticcto.com)
3 points
by
rbuccigrossi
4mo ago
|
1 comments
8.
▲
by
rbuccigrossi
4mo ago
You're first :D
9.
▲
Show HN: Decoding the Language Machine – AI video series and CC repo
(github.com)
2 points
by
rbuccigrossi
4mo ago
|
2 comments
10.
▲
by
rbuccigrossi
1y ago
We work in the arena of automated AI workflows where consistency of success is vital. When you threaten an LLM you are drawing the LLM into the texts where threats occur (flame wars, parody, etc.). So intuitively you would expect it to work
11.
▲
Threatening AI Does Not Make It More Useful. Why Sergey Brin Is Wrong
(tcg.com)
5 points
by
rbuccigrossi
1y ago
|
3 comments
12.
▲
by
rbuccigrossi
1y ago
Treating an LLM with respect is not about pretending it has feelings; it’s about understanding that every word in your prompt is a signal that shifts the probabilistic landscape from which the model draws its answer. It’s about probability,
13.
▲
“End of Support” Does Not Mean “End of Life” for Open Source Projects
(tcg.com)
2 points
by
rbuccigrossi
6y ago
|
0 comments