Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
tibbar
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
61.
▲
by
tibbar
7mo ago
I have no skin in the game here, but this seems a bit "sharp-edged", do you have something against the guy? He just seems deep into his influencer/retired hobbyist arc to me...
62.
▲
by
tibbar
7mo ago
I think the big frustration I've had in learning modern ML is that the entire owl is just so complicated. A poor explainer reads like "black box is black boxing the other black box", completely undecipherable. A mediocre-to
63.
▲
by
tibbar
8mo ago
Yeah, I think you're right that LLMs are overused. In most cases where a deterministic system is feasible and desirable, it's also much faster and cheaper than using an LLM, too..
64.
▲
by
tibbar
8mo ago
To be clear, it's absolutely impossible for OpenAI and the others to stop. The valuation and honestly the global markets depend on them staying leveraged to the hilt. So they're not going to stop. However, the point is that the mo
65.
▲
by
tibbar
8mo ago
Determinism in agents is a complex topic because there are several different layers of abstraction, each of which may introduce its own non-determinism. But yeah, it is going to be difficult to induce determinism in a commercial coding agen
66.
▲
by
tibbar
8mo ago
Respectfully, was this comment AI generated? It has all the signs. And scaffolding does matter a lot, but mostly because the models just got a lot better and the corresponding scaffolding for long running tasks hasn't really caught up
67.
▲
by
tibbar
8mo ago
Right, but if OpenAI wanted to stop doing research and just monetize its current models, all indications are that it would be profitable. If not, various adjustments to pricing/ads/ etc could get it there. However, it has no reaso
68.
▲
by
tibbar
8mo ago
It was a modest update to a UX ... certainly nothing world-changing. (It's also had success with some backend performance refactors, but this particular change was all frontend.) The note was basically just a transcription of what I wa
69.
▲
by
tibbar
8mo ago
Today I got a feature request from another team in a call. I typed into our slack channel as a note. Someone typed @cursor and moments later the feature was implemented (correctly) and ready to merge. The tools are good! The main bottleneck
70.
▲
by
tibbar
8mo ago
The metaphor for the original post was more like "You're already wearing a raincoat and umbrella, and you're forecasting a flood warning?" So, the flood warning (project revenue) may be completely incorrect, but it'
71.
▲
by
tibbar
8mo ago
For me, math was a way to study structure. I find this sort of thing tremendously beautiful on its own, but as it happens "finding the structure in things" turns out to be quite lucrative in the professional world as well, and I o
72.
▲
by
tibbar
8mo ago
There's a good chance that in the long run LLMs can become good at this, but this would require them e.g. being plugged into the meetings and so on that led to a particular feature request. To be a good software engineer, you need all
73.
▲
by
tibbar
8mo ago
the problem with LLM code review is that it's good at checking local consistency and minor bugs, but it generally can't tell you if you are solving the wrong problem or if your approach is a bad one for non-technical reasons. This
74.
▲
by
tibbar
8mo ago
I can't speak to your specific set up, but it sounds like you're halfway there if you can access the previous traces? All anyone can ask for is "show me the traces that led up to this point"; the "why did you do thi
75.
▲
by
tibbar
8mo ago
For a typical coding agent, there are intermediate tool call outputs and LLM commentary produced while it works on a task and passed to the LLM as context for follow up requests. (Hence the term agent: it is an LLM call in a loop.) You can
76.
▲
by
tibbar
8mo ago
I think this statement is on the same level as "a human cannot explain why they gave the answer they gave because they cannot actually introspect the chemical reactions in their brain." That is true, but a human often has an inter
77.
▲
by
tibbar
8mo ago
if the agent can review its reasoning traces, which i think is often true in this era of 1M token context, then it may be able to provide a meaningful answer to the question.
78.
▲
by
tibbar
8mo ago
Haha, I guess we see this in reverse: I see the specific framework as the implementation detail (I use, and enjoy, temporal!) I think it's the part about automatically launching a metric ton of agents that is (to use the term again) mi
79.
▲
by
tibbar
8mo ago
Oh, like, I wouldn't actually use this specific implementation. I used to work at a shop with thousands of lines of Oracle triggers that you had to edit inline in the web browser with no version control and I shudder to think of retu
80.
▲
by
tibbar
8mo ago
I can only speak for my own mind ;) but the most advanced thing I'd seen prior in this regard was Google Sheets' =AI function, which is pretty convenient (if awkward) when you want to map values to LLM output. What I specifically
81.
▲
by
tibbar
8mo ago
Is it? pgvector is primarily about indexing, I think? This feels more similar =AI in Google Sheets.
82.
▲
by
tibbar
8mo ago
This is mind-bending. I can't imagine that it performs well enough to be particularly fit for production just yet but..... wow.
83.
▲
by
tibbar
8mo ago
this is the real reason why people are switching to claude code.
84.
▲
by
tibbar
8mo ago
Here’s a post from 2021 about the migration! [0] I guess 2021 is a long time ago now. How did that happen… [0] https://github.blog/engineering/infrastructure/partitioning-...
85.
▲
by
tibbar
8mo ago
It's interesting to take the counterfactual, what it looks like when large projects are run poorly. The answer often looks like: * Poorly defined goals / definition of success * Overly-complex plans, slowly executed against * A fo
86.
▲
by
tibbar
8mo ago
Github used to publish some pretty interesting postmortems. Maybe they still do. IIRC that they were struggling with scaling their SQL db and were starting to hit the limits. It's a tough position to be in because you have to either to
87.
▲
by
tibbar
8mo ago
I think most large platforms eventually split the tools out because you indeed can get MUCH better CI/CD, ticket management, documentation, etc from dedicated platforms for each. However when you're just starting out the cognitive
88.
▲
by
tibbar
8mo ago
Lots of dedicated CI/CD out there that works well. CircleCI has worked for me
89.
▲
by
tibbar
8mo ago
You would kind of expect with the pressure of supporting OpenAI and GitHub etc. that Azure would have been whipped into shape by now.
90.
▲
by
tibbar
8mo ago
GitHub isn't VC funded at the moment, though. It's owned by Microsoft. Not that this necessarily changes your point.
More ›