Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
efromvt
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
efromvt
3mo ago
Out of curiosity, how often are the resource limits the bottlenecks? What do harnesses do to help here - limit parallelism? More efficient tools?
32.
▲
by
efromvt
3mo ago
Second comment, having read in more depth (really love the auto-layout detail!) - the spec doesn't seem to naturally support layering (which is useful in some multi-axis automatic cases) - any plans for composability?
33.
▲
by
efromvt
3mo ago
Op1M5 was scarring the first (dozen) time around! Thank you minelayers
34.
▲
by
efromvt
3mo ago
I wonder if all those different heavy shenanigans are just to get the hollander's gauss/hunchback autoconnon in the right torso by default. (original Mechcommander is still the best of the series, even if mission 5 broke my soul f
35.
▲
by
efromvt
3mo ago
Semantic types as the extra formatting factor is super useful because they are concise encodings of a lot of formatting boilerplate.[1] Do you envision the flint type registry being shared/extensible? Why not have that be a data proper
36.
▲
by
efromvt
3mo ago
Very much in agreement with this but especially on the UX side I find that 'what is intuitive to me' is not universal because you come with a bunch of contextual priors you can't clear. I wish there was more tooling focused o
37.
▲
by
efromvt
3mo ago
Exactly, so it has a success rate of 0 and infinite cost/completion on your relevant benchmark. If the benchmark doesn't map to what you need it to, then yeah, it's not a useful input.
38.
▲
by
efromvt
3mo ago
If you don't have a backend, then it's all telemetry, right? And backend logs don't capture a lot of the UX side of things - how a call got triggered, from where, etc (which yeah you can start to instrument, but then that
39.
▲
by
efromvt
3mo ago
Isn't the benchmark working exactly how it should in that case?
40.
▲
by
efromvt
3mo ago
It feels like you could argue that since you control nature/nurture it's very possible to create a model aligned to an arbitrary spec - there is no theoretical reason it's not possible given N runs, and you only need to tak
41.
▲
by
efromvt
3mo ago
I'm not sure the security/safety stuff is entirely in their control at this point, though you can argue that they are indirectly responsible through encouraging regulation via their positions on safety/risk.
42.
▲
by
efromvt
3mo ago
I find it interesting that when drawing this parallel you mention that some devs 'get it' and 'build a great culture'; I think this is exactly where the analogy breaks down. Good managers get great results from people (a
43.
▲
by
efromvt
3mo ago
I think it often useful to push the conversation down "we built a system for humans that dealt with this, what from that is or is not applicable for agents in the same context"? Humans randomizing resume review for screening is pr
44.
▲
by
efromvt
3mo ago
This starts to sound more like ‘social engineering a human assistant’, so there’s a degree of required specialization that does meaningfully increase costs.
45.
▲
by
efromvt
4mo ago
I think the theory is that it’s not purely static - you need to keep training and tuning the model (even just for general knowledge upkeep with current architecture) and so the infra/data is a contributory moat. Exfiltrating weights wo
46.
▲
by
efromvt
4mo ago
million times this - getattr on every dataclass is a wild choice
47.
▲
by
efromvt
4mo ago
The SELECT machinery is the product with databases! SQL often the shortest description of the processing logic, and the database has an efficient local execution engine that can prune/reduce data read based on the plan. Very hard to
48.
▲
by
efromvt
4mo ago
Why would an agent supersede a service for a well-defined workflow contract that does not require an agentic loop? I assume both will need to exist.
49.
▲
by
efromvt
4mo ago
I’d be interested in the benchmarking if you ever write it up! People do seem to assume LLM as a judge/panel improves outcomes (and arguably it does in cases like code review?) but I suspect it is very situational and the priors from h
50.
▲
by
efromvt
4mo ago
Been working on optimizing CLIs for cheap agent use and figuring out how to build integrated agentic features that aren’t a full chat interface. Agent UX optimization is kind of fun! Much more testable than human UX, though it’ll be interes
51.
▲
by
efromvt
4mo ago
Fantastic to have PyO3/Maturin guides too - the rust/python/typescript turducken I’ve always wanted.
52.
▲
by
efromvt
4mo ago
repetition of "belt-and-suspenders" kills me with opus, especially because it always means the model is suppressing something I would want to be an actual failure
53.
▲
by
efromvt
4mo ago
I think the perception is that it is not 'only marginally better'; whether or not you specifically agree that perceived quality gap lets them differentiate on price. I'd further say that there are probably enough rational a
54.
▲
by
efromvt
4mo ago
I think you can sympathize with the safety motives while still thinking this was a dumb implementation to degrade silently? I actually have faith in them getting the guardrail triggers pretty good, but consensus seems like they’re not yet t
55.
▲
by
efromvt
4mo ago
The openrouter provider flakiness with deepseek was infuriating, but I’m happy in hindsight because direct deepseek has been very pleasant. Shocked by how low spend is.
56.
▲
by
efromvt
4mo ago
I do slightly prefer 5.5 for complex work but Claude quota usage has gotten infinitely better since the dark days a few months back - has gone from being infuriating to something I pretty much don’t have to worry about with it as a daily dr
57.
▲
by
efromvt
4mo ago
I'd be very curious about the bottleneck breakdown in most current software dev - I suspect inference is far from the bottleneck in most things I do, though driving it to 0 would still be nice . I do agree that if it was 0 we'd p
58.
▲
by
efromvt
4mo ago
Deepseek cost/performance is incredible. That said, I still feel like for agentic coding we haven't plateaued (I slightly prefer GPT 5.5 to Claude for complex stuff, to be honest), and so the extra price is absolutely worth it to
59.
▲
by
efromvt
4mo ago
There's always a bulk insert, but I wouldn't say every engine has always had a reasonable way to bulk load truly large data... parquet really helped with interop but before that when your best option was a CSV and bcp life was n
60.
▲
by
efromvt
4mo ago
A pretty common request is to lift the FROM up before the select, like the below. I'm pretty fine with status quo since my mind is usually "hmm what do I need to get" first, then I figure out how to get it, but some engines (
More ›