Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
wek
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
An Empirical Study of Harness Design for Coding Agents
(arxiv.org)
29 points
by
wek
8d ago
|
1 comments
2.
▲
by
wek
9d ago
I agree with this. Why get locked in at the harness layer. I want my workspace to be independent of my coding agents and models so I can use multiple and switch as I wish
3.
▲
by
wek
9d ago
This is cool! We have a similar approach in Nimbalyst... agent makes extensions that are human editable and agent editable ... includes Jupyter, markdown, csv, excalidraw, code, pdfs, browser. Its so nice not to have to app switch all the
4.
▲
by
wek
11d ago
Thank you for this. This seems like a promising approach to agent permissioning. What are the performance implications?
5.
▲
Multiplayer for Claude Code and Codex
(nimbalyst.com)
1 points
by
wek
24d ago
|
0 comments
6.
▲
by
wek
1mo ago
From their abstract: "Our findings reveal that despite continuous development activity and growing codebase complexity of the agent harnesses, there is no statistically significant improvement in SWE-bench benchmark score (i.e., resolv
7.
▲
Agent Harness Evolution Shapes Coding Agent Quality
(arxiv.org)
2 points
by
wek
1mo ago
|
1 comments
8.
▲
AI-to-AI Code Reviews of GitHub Pull Requests
(arxiv.org)
1 points
by
wek
1mo ago
|
0 comments
9.
▲
by
wek
1mo ago
Exactly. Expertise is frequently required to steer, correct, rework, know what is good and what isn't. This happens 500 times a day on both the micro level and the architecture level.
10.
▲
by
wek
2mo ago
Use Nimbalyst (I work on it). Multi-sessions, visual editing, open-source
11.
▲
Study: Coding agents rarely retrieve open-source contribution rules
(arxiv.org)
2 points
by
wek
2mo ago
|
0 comments
12.
▲
by
wek
2mo ago
I have Max subscription to both Claude Code and Codex. I use both fulltime throughout the day. I have them check each other's work, and route tasks they are best at to the right one. The checking each other's work is so helpful. N
13.
▲
Position: Coding Benchmarks Are Misaligned with Agentic Software Engineering
(arxiv.org)
2 points
by
wek
3mo ago
|
0 comments
14.
▲
Configuring Agentic AI Coding Tools: An Exploratory Study
(arxiv.org)
3 points
by
wek
4mo ago
|
0 comments
15.
▲
Building the harness around our coding agents. Eight failure modes and pillars
(nimbalyst.com)
3 points
by
wek
4mo ago
|
0 comments
16.
▲
Constraint Decay: The Fragility of LLM Agents in Back End Code Generation
(arxiv.org)
287 points
by
wek
4mo ago
|
197 comments
17.
▲
by
wek
4mo ago
The article is correct to emphasize the importance of definition and that this can be a bottleneck. But it is incorrect to show the layered documentation and development lines taking just as long as they did pre-coding-agents. Our team is c
18.
▲
Show HN: Nimbalyst open source Obsidian, Codex app, and Linear for coding agents
(github.com)
7 points
by
wek
5mo ago
|
1 comments
19.
▲
by
wek
5mo ago
Well said. It is so much better for when the human can see what the LLM has changed and iterate on it with the LLM, making their own changes. Markdown is better for that.
20.
▲
by
wek
5mo ago
I kind of felt the same way reading the article! It felt so unusual to encounter someone who is both smart and humble and willing to admit they were learning. And I was happy to encounter it and sad that I was so surprised by it.
21.
▲
by
wek
5mo ago
Congratulations! I think that a lot of the value will be in the judgement of the maintainer about the marginal next feature (and saying no to all those other features) ... if software is a stream then the value is in what gets into the stre
22.
▲
by
wek
5mo ago
What an excellent article by a smart, humble, still-learning person! Favorite quote:" There are a whole bunch of reasons I’m not scared that my career as a software engineer is over now that computers can write their own code, partly b
23.
▲
by
wek
5mo ago
And as of today Nimbalyst is open source!
24.
▲
by
wek
5mo ago
This is on our wish list for a custom extension. Many users have asked for it. If someone in the community wants to take a crack at it....
25.
▲
by
wek
5mo ago
Thanks! A good way to learn the platform is by building an extension.
26.
▲
Agyn: A Multi-Agent System for Team-Based Autonomous Software Engineering
(arxiv.org)
2 points
by
wek
5mo ago
|
0 comments
27.
▲
The Code Agent Orchestra: what makes multi-agent coding work
(addyosmani.com)
2 points
by
wek
5mo ago
|
1 comments
28.
▲
by
wek
5mo ago
I use Nano Banana all the time and this seems like a step up
29.
▲
by
wek
5mo ago
I agree. Why would they not keep the $20 plan as a gateway drug?
30.
▲
by
wek
5mo ago
What are the implications of this for Cursor being model agnostic?
More ›