Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
extr
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
extr
12d ago
It's a great model and you're right it does feel quite natural at times while Fable 5.1 still has a claude-ish shape to it. Unfortunately I just find that it's not reliable enough as a daily driver and ends up performing spec
2.
▲
by
extr
12d ago
I used to do this but recently I switched to having Fable 5.1 spawn forks of itself rather than Opus subagents. Yes it's more expensive but you don't pay for reads that already happened pre-fork, and you end up doing less rework s
3.
▲
by
extr
12d ago
- It doesn't write great code. - Occasionally has strange tics around asking for permission for obvious next-steps, implied actions, etc. - It's very expensive, both in terms of tokens and % usage on subscription plans. - Relatedl
4.
▲
by
extr
12d ago
I just tell everyone to use Fable 5.1 for everything at this point. Astra is unfortunately a dud, I'm sure they will try to fix a bunch of it with GPT-6.1 but OAI has had this issue for awhile now where every other generation has some
5.
▲
by
extr
22d ago
I use NextDNS at home. It's great to have lots of control around DNS while also running a locked-down router (Eero). I personally use it in combination with a raspberry pi to run a version of this https://cazander.ca/20
6.
▲
by
extr
1mo ago
Sounds like you have a lot of axes to grind outside of simply "the latest models regressed on delivering concise prose".
7.
▲
by
extr
1mo ago
It's "literally unbearable" when the AI that completes software engineering tasks at 100x speed and quality from 2 years ago uses too much jargon?
8.
▲
by
extr
1mo ago
I'm sorry but the whining over LLM output styles is embarrassing. Do Claude and GPT models always respond in exactly the way my most articulate coworker would? No. The overused jargon is absolutely annoying. But these things aren'
9.
▲
by
extr
1mo ago
Unfortunately Sol does not compare to Fable at all.
10.
▲
by
extr
2mo ago
keep in mind fable = mythos which as been "done" since february. so the gap is not 2 months, it's more like - techniques probably started "working" in late 2025, now are trickling down to 2nd tier labs 9 months late
11.
▲
by
extr
2mo ago
It's because Fable is just synthetic RL tasks + scale. The secret has been out for awhile now.
12.
▲
by
extr
2mo ago
Manus did a lot of harness work to make up for gaps in Opus 4.5 tier models. I tried it for a time - they had a great deep research/PDF generation pipeline, parallelization, etc. The bitter lesson has now come for them: the latest mode
13.
▲
by
extr
2mo ago
yeah it's true, you do have to guide them. i find that the key is you have to know what's possible. you have to have the instinct for "this really shouldn't be so difficult". my junior SWE coworkers have the same tr
14.
▲
by
extr
2mo ago
$80 is definitely low now that I look at my numbers. but not OOMs low, it's closer to like $200 on heavy days. i don't know how you're doing $3k/day, that's wild. i'm pretty aggressive about compaction and sess
15.
▲
by
extr
2mo ago
Yes lol. Of all things people are getting on me for it's the number of LoC x Years In Business of this startup. I don't fucking know, I didn't start the company and I wasn't here for several of those industrious years. L
16.
▲
by
extr
2mo ago
> "unguided" means "I typed a prompt into claude code and waited yolo" Yes, this is literally what that means.
17.
▲
by
extr
2mo ago
This is a great point and I agree. My own productivity varies based on what part of the codebase I'm working on. If it's "been in there before" and I know the right questions to ask, I can one-shot a good design/imp
18.
▲
by
extr
2mo ago
It's a fair point, it's not truly unlimited and I do wonder how that would change my workflow. I can definitely imagine if I was inside Anthropic or OAI with unlimited "fast" tokens, you would be more tempted to hand ove
19.
▲
by
extr
2mo ago
No, actually. The point is to build a profitable business.
20.
▲
by
extr
2mo ago
> unguided LLM usage Why aren't you guiding your LLM usage? Is that what I said - to spam it and not guide anything? Or to have a careful workflow where you agree on design and maximize your human judgement/leverage? > any s
21.
▲
by
extr
2mo ago
How is this not true? Taking a Senior SWE @ ~$200K, even just the base salary cost / 2080 working hours is $100/hr. Fully loaded employer cost + accounting for non-coding time gets you to upper 100s easily. Even for a junior makin
22.
▲
by
extr
2mo ago
Yes 100%. This morning I casually prompted Codex to drive the browser to complete extensive performance testing in-situ that would have literally been weeks of work before. Probably in reality it just wouldn't have been done, and perfo
23.
▲
by
extr
2mo ago
Have you worked at many startups?
24.
▲
by
extr
2mo ago
This was more true a few months ago but Fable has improved the situation considerably. Also just remember - minimalist code looks and feels great but customers do not read your code. I have caught myself many times providing "correctio
25.
▲
by
extr
2mo ago
Disagree. I operate this way inside a multi-million line legacy codebase.
26.
▲
by
extr
2mo ago
Keep the decision-making and execution separate. Use the high IQ models to chat about the design and make them drive subagents to do the actual work. "Chat" style threads are actually quite cheap. Where it gets expensive is having
27.
▲
by
extr
2mo ago
Performance is better than ever. It's never been more practical to set up wildly complex synthetic test environments and measure perf wins. Plus the models will find every possible algorithmic/design improvement. It actually gives
28.
▲
by
extr
2mo ago
I would be really curious to hear from devs at Databricks what the experience of development is like internally. I work at a small startup with essentially unlimited AI spend budget - the entire point is that I should be turning to it at ev
29.
▲
by
extr
2mo ago
I don't really care if it was made by AI or smeared onto the keyboard by a monkey. It was effective in it's job and communicated clearly.
30.
▲
by
extr
2mo ago
Well made website IMO. Gets to the point with plenty of easy-to-understand evidence and calls to action.
More ›