Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
oren1531
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
oren1531
7mo ago
Try npx agent-triage demo — runs on sample data locally. Would love to hear what you find when you point it at your own traces.
2.
▲
Show HN: Agent-triage – diagnosis of agent failures from production traces
(github.com)
4 points
by
oren1531
7mo ago
|
2 comments
3.
▲
by
oren1531
7mo ago
I built agent-triage - CLI that automates diagnosing AI agent failures in production. I was spending way too much time staring at logs and web dashboards trying to figure out why my multi-agent setups kept failing. You just point it at your
4.
▲
by
oren1531
7mo ago
Grok's value was never really about model quality - it was the only model with real-time access to what's actually being said on X. And it's less filtered than the others, which matters for certain topics where ChatGPT/C
5.
▲
by
oren1531
7mo ago
Good point on the green/red dashboard. The opportunity cost angle is worth adding though. A failed run isn't just the wasted tokens and retry cost - it's also the task that didn't get done and the engineering required to