Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
aroman
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
aroman
6d ago
The reverse captcha really made me feel something in my bones. Like for a moment I was a second-class citizen of the web. I wonder if this is how it "feels" to be an LLM attempting to use the web...
2.
▲
by
aroman
7d ago
I want to try Planetscale... but we're addicted to (and totally dependent on) Neon's branching model. They really got us hooked on that!
3.
▲
by
aroman
8d ago
Finally, I can delete `sync-agent-docs.sh`, which recursively symlinked AGENTS.md to GEMINI.md and CLAUDE.md...
4.
▲
by
aroman
8d ago
nix + LLMs is the most fun I've had with personal computing since I burned an Ubuntu live CD in middle school
5.
▲
by
aroman
8d ago
Shall we also ban projects that are not authored directly in bytecode? Surely it's not really "made by hand" if you used a higher level interpreted language. But seriously: software engineering is all about leveraging abstra
6.
▲
by
aroman
13d ago
> It also seems like frontier models reached some limit, whether this is capex related, business model related or something else. Nobody knows but it's happening to all frontier labs it seems. What’s your evidence of this? On the co
7.
▲
by
aroman
13d ago
And again the boy will cry out “the wolf is here!” But this time the wolf really will be here, and no one will believe him.
8.
▲
by
aroman
28d ago
I read it as: we acted as quickly as we could once it became possible to do so , i.e., the change in control was completed. not that their special agreement carved out some maximum notice term.
9.
▲
Bill Gates says world has "no plan" on AI in new essay
(gatesnotes.com)
17 points
by
aroman
1mo ago
|
3 comments
10.
▲
by
aroman
1mo ago
I think you need to spend more time with Sol. If you think there is nothing even close to as good as Fable - my guess is you haven’t spent as much time getting as familiar with working with those models as you have with Claude’s. Codex is m
11.
▲
by
aroman
1mo ago
The LLM is reasoning about estimates from its training data... which is to say, from human engineering timescales. I suspect the labs could improve the models such that they are estimating these sorts of things but they don't prioritiz
12.
▲
by
aroman
2mo ago
I tried Fedora Silverblue and I found its notion/implementation of immutability fairly frustrating. It’s actually the reason I went to nixOS. It seemed to me that it was all the friction of immutability without any of the benefits of r
13.
▲
by
aroman
2mo ago
What alternative have you moved on to?
14.
▲
by
aroman
2mo ago
Revenue, data, and stamina.
15.
▲
by
aroman
2mo ago
I was an early and passionate adopter and paying customer of Cursor (since 2023!), but it’s probably been 6 months since I opened it. These days I “write” code with claude code and codex, and read/review it on GitHub. If I need to read
16.
▲
by
aroman
2mo ago
I simply do not believe the switching costs are high enough that they could eliminate those plans. The Chinese models will eat their lunch.
17.
▲
by
aroman
2mo ago
Right, it's about how much time the human spent - the time spent by the machine itself is irrelevant. As you rightly point out: that is why we measure programmer experience by wall clock time, not CPU time :)
18.
▲
by
aroman
2mo ago
No, it wouldn’t. The hours in question are human experience, not that of the agent.
19.
▲
by
aroman
2mo ago
Sun Tzu said: if you scrape your enemy, call it training; if your enemy scrapes you, call it an attack.
20.
▲
by
aroman
2mo ago
You have hundreds of hours with a model that was barely even released hundreds of hours ago? The perception of capability varies greatly between task. For my needs for example sol xhigh consistently outperforms fable xhigh.
21.
▲
by
aroman
3mo ago
No. I used to use Cursor, but now my workflow is that I use an inhouse CLI tool I wrote called "bud" that wraps/seeds the harnesses per-worktree, and boots a full copy of the game so each worktree can work independently. If g
22.
▲
by
aroman
3mo ago
Indeed, much of the scariness is how fearlessly and confidently it writes them with little regard to their actual usefulness or value. When I find it adding a lot of tests, I often say something like: "audit each test carefully, and co
23.
▲
by
aroman
3mo ago
Like before AI, the scrutiny varies with the sensitivity of the area being edited. Simple UI change? I do an AI review, but otherwise neither read nor write the code. The models are good enough they write better UI code than me, 9 out of 10
24.
▲
by
aroman
3mo ago
In terms of ability to ship? Easily tenfold. We literally ship 10 times more than before AI. This does not, however, translate into a tenfold increase in actual business success, of course :)
25.
▲
by
aroman
3mo ago
Claude Code fan here... Codex is very good. Sometimes better. The killer feature is price. After 6+ months of exclusive Claude Code usage, I was begrudgingly forced to try Codex once Anthropic rejiggered their limits such that I kept maxing
26.
▲
by
aroman
3mo ago
I've been doing this for ages - you just spin up harness B as a subprocess/tool call from harness A. For example, I had a "/codex-review" claude skill for ages that did exactly that. Technically you're right it
27.
▲
by
aroman
3mo ago
This makes me think they really are quite capacity constrained at the moment. I had assumed they were primarily limiting it to entice people to upgrade, but I feel like these limits are so low and so temporary (especially over July 4th week
28.
▲
by
aroman
3mo ago
I'm not sure I follow - 614 GB/sec is pretty squarely in dGPU territory (~5070 level). External GPUs can definitely exceed that on the very high end, but it seems pretty competitive, no?
29.
▲
by
aroman
3mo ago
For sure, on paper - I'm curious, do you actually notice that difference in your day-to-day? I struggle to think of times in my usage of my computer where I think "this feels slow", but maybe I'm blind to it.
30.
▲
by
aroman
3mo ago
You're arguing against a point I did not make. I observed that Cloudflare prioritizes expanding to new products over making improvements to existing ones. I did not claim they do not improve their products. There are numerous example
More ›