Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
akitaonrails
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
akitaonrails
25d ago
Every major open source project today ships with a CODE_OF_CONDUCT.md in the repository. The idea was sold that this existed to “protect minorities” and “create safe spaces.” In practice, what I watched happen over the last decade was somet
2.
▲
by
akitaonrails
1mo ago
For many years I had a fixed idea in my head, and I always thought it wasn’t practical: converting NES games to run on the Master System.
3.
▲
by
akitaonrails
1mo ago
TL;DR: ai-memory 2.0 is out, with the open OKF format, local embeddings turned on by default, and real support for several agents and a whole team working on the same project in parallel.
4.
▲
by
akitaonrails
2mo ago
Last week I wrote about the Discord censorship and Brazil’s Digital ECA law, the Digital ECA (“Estatuto Digital da Criança e do Adolescente”, the digital version of Brazil’s Child and Adolescent Statute): the ANPD (Brazil’s data protection
5.
▲
by
akitaonrails
2mo ago
Claude now ships with a stamp on it. Since August 2, 2026, Anthropic has been invisibly marking the text of its newest models, with the older ones migrating over the following months. No off switch. The backlash came fast: a wave of subscri
6.
▲
by
akitaonrails
2mo ago
I’ve run four models: Qwen 3.8 Max, GLM 5.3, Gemini 3.7 Flash, and a 27B Qwen 3.8 running locally on my RTX 5090. One of them closed in on the leading group. Another pulled off the biggest jump this test has ever recorded. A third got caugh
7.
▲
by
akitaonrails
2mo ago
I spent 1 week and hundreds of dollars to run a real-world programming benchmark (real project back to back) through 30 LLMs to assert which is best and why.
8.
▲
by
akitaonrails
3mo ago
Yes, it is.
9.
▲
by
akitaonrails
3mo ago
ai-memory run keeps the same programming session while switching among Claude Code, Codex, and other harnesses, with searchable workstreams and integration with ai-jail and ai-usagebar.
10.
▲
by
akitaonrails
3mo ago
This post has two parts. First, the news: Microsoft’s Majorana 2 announcement and why the physicists in the field remain with one foot (both, actually) behind. Then, the educational part: a decent explanation of Shor’s algorithm, because a
11.
▲
by
akitaonrails
3mo ago
Everything below is on my GitHub, with binaries, releases and install instructions. In this post: clock-tui (with the google-calendar-tui and ghpending widgets), github-visualize, frank_geary, frank_scanlation, frank_lyrics, frank_type, dis
12.
▲
by
akitaonrails
4mo ago
This week a curious controversy showed up in the open LLM world: Rio 3.5 Open 397B, presented by the City of Rio / IplanRIO in the Rio 3 Open announcement as an open model. The accusation: in practice, the published checkpoint, meaning
13.
▲
by
akitaonrails
4mo ago
Another round of my coding benchmark. This time it’s three open source entries: Kimi K2.7 Code, GLM 5.2, and the MiniMax M3 that got open weights but that I can’t run at home no matter how hard I squeeze. Before the numbers, the usual conte
14.
▲
by
akitaonrails
4mo ago
Anthropic shipped Claude Fable 5 this week, and before we get to my benchmark numbers (spoiler: a technical tie with Opus), the whole soap opera deserves a retelling, because the context matters more than the model. And it’s a good one, the
15.
▲
by
akitaonrails
4mo ago
Did you know many manga artists compose key moments as two-page spreads? If you read on typical fansub sites, you are probably missing many of them because most of those pages only show one long, vertical, single-page scroll. Prettify Manga
16.
▲
by
akitaonrails
4mo ago
TL;DR: Opus 4.8 doesn’t feel much different from Opus 4.7, not in daily use and not in the benchmark. 95/100 against 97/100, inside the noise. It’s the fastest Opus I’ve measured, but the day-to-day experience is the same. Grok 4.
17.
▲
by
akitaonrails
4mo ago
I have a strong opinion about this, so let me just plant the flag: no open source project is ready to be published without three things. In order of importance: Installation surface. The new user has to be able to install and try the tool w
18.
▲
by
akitaonrails
4mo ago
I’m a subscriber to the MANGA Plus by Shueisha app on Android’s Google Play Store. For those who don’t know, this is Shueisha’s official channel (the Japanese publisher behind Shonen Jump) for reading their manga legally, with chapters rele
19.
▲
by
akitaonrails
4mo ago
Most people just trust Google and leave everything there. Anyone who follows this blog knows I don’t trust anyone: if it’s in the cloud, it’s not mine. That’s why I keep a NAS at home to back up everything that matters. I don’t trust Google
20.
▲
by
akitaonrails
5mo ago
Every time I mention coding agent harnesses, somebody shows up asking: “but why don’t you use Pi?” or “why not Oh-My-Pi?” I hate that question when it comes like that. Not because I have anything against Pi. I don’t. The problem is that alm
21.
▲
by
akitaonrails
5mo ago
In last week’s post, “Wrapping Up My AI Marathon: Success or Failure?”, I closed the first phase of my marathon. More than 600 hours of heavy coding-agent use, dozens of projects, hundreds of thousands of lines of code and documentation, an
22.
▲
by
akitaonrails
5mo ago
Anyone following the blog already knows I use a whole bunch of different LLM vendors. The main ones are Claude and GPT, via Claude Code and Codex. I use each one’s harness because I’m locked into their subscription plans (Pro, Plus, Max), w
23.
▲
by
akitaonrails
5mo ago
Five days ago I wrote a long post on coding-agent memory where I recommended agentmemory as the answer. After a week running it in personal production, I’m walking it back. This post explains what went wrong and the open source project I st
24.
▲
by
akitaonrails
6mo ago
TL;DR: No. Across all three rounds of experiments I ran, mixing “strong frontier planner + cheap executor” loses to just using Opus 4.7 alone in a mature harness. Solo Opus 4.7 in opencode delivers Tier A (97/100) in 18 minutes for ~$4
25.
▲
by
akitaonrails
6mo ago
This is possibly one of the most relatable benchmarks for real programmers, to compare most of the more popular open source and commercial LLMs available. You will be surprised by those findings and insights.
26.
▲
by
akitaonrails
6mo ago
Clean Code was originally written for human programmers. But how does it change when we now write for AI agents?
27.
▲
by
akitaonrails
6mo ago
I spent the whole Sunday on this, and it was one of the most productive Sundays I’ve had in a while. The mission was specific: close the loop on the simcade racing games that are hardest to emulate, the ones I’ve wanted to run properly for
28.
▲
by
akitaonrails
6mo ago
TL;DR: Yes, the title is clickbait. The answer is no, it’s not worth it. Keep using Claude Code with Opus 4.6 or 4.7. Details below. A few weeks ago I wrote a detailed LLM coding benchmark comparing 33 open source and commercial models on t
29.
▲
by
akitaonrails
6mo ago
I bought a Lenovo Thinkpad T14 Gen 6 and installed Omarchy on it. It’s not my main machine, and it’s not supposed to be. It’s a companion: a notebook I can open on the 3D printing office desk, SSH into the desktop, fire up Claude Code, acce
30.
▲
by
akitaonrails
6mo ago
I updated my report and ranking on the LLMs benchmarks to add the newest Claude Opus 4.7 and Qwen 3.6. Check it out (all the benchmark code is open sourced in my GitHub)
More ›