Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sdesol
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
sdesol
3mo ago
This is a hard problem, but one worth solving, I think, since it means less tokens and better AI reasoning. I believe LLMs are good enough that, if given the right context, it can very much solve almost all tasks. If this works, it means we
32.
▲
by
sdesol
3mo ago
> So I want to say there's still a lot of value in context engineering though it seems to diminish with each model release I can't see how it would diminish unless you are literally working on public domain stuff. Unless stuffi
33.
▲
by
sdesol
3mo ago
The user flow I am trying to get adopted for sessions is to turn them into notes and lessons when you have finished and it should be part of the code review process. By propery categorizing lessons and notes, it should make it easy to scrub
34.
▲
by
sdesol
3mo ago
> You don't need another layer I do think we need another layer, but it should be a routing layer. I am finalizing my pi-brains extension for Pi ( https://github.com/earendil-works/pi ) which does this: https:&#
35.
▲
by
sdesol
3mo ago
> For AI agents, it means well-designed CLI tools that help the agent orient itself in a task and pull exactly the 'context-for-the-job' it needs right then. This is exactly what I am trying to solve and I have what I call smar
36.
▲
by
sdesol
3mo ago
I talked about this before but China would be in much better position if LLMs turn into a commodidty. Where they can dominate is in hardware, as fast and cheap inference is probably going to be the moat.
37.
▲
by
sdesol
4mo ago
> It’s demoralized a lot of people Isn't this a good thing? Doesn't it mean your company doesn't want employees to treat LLMs like a slot machines?
38.
▲
by
sdesol
4mo ago
Was there at least performance gains to be measured?
39.
▲
by
sdesol
4mo ago
The big thing is, the western world has moved so much of the manufacturing to China and think a lot of people will not forgive Samsung and others, so I can see China owning a good portion of the supply chain.
40.
▲
by
sdesol
4mo ago
> Microsoft adding Deepseek support I believe this hasn't been confirmed yet but I think it speaks to a bigger problem for the AI companies which is, if you give capable developers a good reasoning LLM, they can make it work like it
41.
▲
by
sdesol
4mo ago
I think the issues is, it is going against a very well established pattern. I have a tool that wraps ripgrep so that search results always includes context and from time to time, the agent will use ripgrep by itself and when I ask why, it w
42.
▲
by
sdesol
4mo ago
> Great! So next time the human will prompt the agent to watch out for and avoid this bug. I actually created a system for something like this. The basic idea is, once you have identified what the issue was and fixed it, you can create l
43.
▲
by
sdesol
4mo ago
What behavioural things are you noticing that Claude does not pickup from the CLAUDE.md. I am working on a `pi-brains` extension for the Pi agent. It is designed to inject rules into write and edits tool calls for matching files. I am curi
44.
▲
by
sdesol
4mo ago
> Filling context with what you think the model needs adds nothing and possibly just inflates context which is harmful. The solution that I've developed is, let the agent figure things out efficiently, without inflating the context.
45.
▲
by
sdesol
4mo ago
The AI agreed > There isn’t a magic prompt I could have written before knowing the solution. A retrospective prompt can reproduce the 200 lines only because it embeds the invariants you discovered. The useful AI task would have been syst
46.
▲
by
sdesol
4mo ago
The code in http://github.com/gitsense/gsc-cli contains over 300 go files and I would say 90% were one shots. I can even prove that they are AI generated since I attribute the AI in each file, as shown in the example b
47.
▲
by
sdesol
4mo ago
> playing a slot machine hoping I still believe the slot machine analogy holds to some extent, but I can honestly say my winning percentage is at least 90% for one shot generated code now. I think if you know it's limitations (inlcl
48.
▲
by
sdesol
4mo ago
> and no ways to measure whether AI is working better. What I do with my product is I explicity tell you to ask your agent. I have real world examples and real world repositories that you can try with: https://gitsense.com
49.
▲
by
sdesol
4mo ago
I am working on solving the AI Code Provenance problem and I believe my repos may be the first that provides AI code provenance. See the following example: https://github.com/gitsense/gsc-cli/blob/main/in
50.
▲
by
sdesol
4mo ago
> GLM writing This is honestly what I care bout the most now, which is how well they can write. I think we have reached a point now, if you know how to program, you can provide enough information for the models to pretty much do what you
51.
▲
by
sdesol
4mo ago
> So it's a scary time to work in tech even if I think the trend will ultimately reverse. I honestly can't see things going back to what it was 5 years ago. We will probably not have the future that Anthropic hopes for, but I t
52.
▲
by
sdesol
4mo ago
> A well thought out pros and cons story wins over binary yes/no answers at pro and anti ai companies alike. The issue with this is, you need to know how to really program to be able to articulate the pros and cons, which a new grad
53.
▲
by
sdesol
4mo ago
> playing trendy tech lottery. I don't know about that, and I am 100% biased so take what I say with a grain of salt. My position is very much this: you may not trust coding agents to make code changes, but if you're not willin
54.
▲
by
sdesol
4mo ago
> replacing deterministic systems in their support flows The issue is, they don't want to provide "better" support but "cheaper" support. Imagine a trained agent that understands the big picture. Now imagine a
55.
▲
by
sdesol
4mo ago
> first for the OSS project and again for a commercial product. Is there a way to reach out to you as I would like to hear what you have to say about what I am working on. You can update your HN profile to include contact information or
56.
▲
by
sdesol
4mo ago
> What is that worth? :-) This is one of those double edge sword situations. It is on the front page and it stays because it will trigger a lot of people and he has to spend a lot of effort explaining himself. What is that worth? His ex
57.
▲
by
sdesol
4mo ago
I created this and I would say glm-4.7 accounts for 80% of the code in https://github.com/gitsense/gsc-cli If you look at a file like: https://github.com/gitsense/gsc-cli/blob/main/i
58.
▲
by
sdesol
4mo ago
Fixed, thanks!
59.
▲
by
sdesol
4mo ago
This sounds like it is more aligned with what I have created which is "We need to capture your conversations with AI". If you look at https://github.com/gitsense/gsc-cli/blob/main/internal/
60.
▲
by
sdesol
4mo ago
> at a Sonnet 4.6-level model MiMo v2.5.0-Pro is honestly the first Chinese model that I've tried where I really though why should I use Claude Sonnet when I can get the same results for a fraction of the cost. There was always some
More ›