Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
grumblemumble
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
How Does AI Interpret Consent: A Look Inside Claude Code's Safety Classifier
(highflame.com)
5 points
by
grumblemumble
2mo ago
|
2 comments
2.
▲
by
grumblemumble
2mo ago
A teardown of Claude Code's auto-mode safety classifier, looking at the undocumented ruleset that interprets user consent.
3.
▲
by
grumblemumble
5mo ago
I'm curious about the performance tax of deep packet inspection on these calls. Is this an eBPF implementation or a local proxy? Moving "guardrails" from code to infra feels like the right architectural shift, provided the la
4.
▲
Securing AI Agents and MCP at the network layer with Tailscale and Highflame
(businesswire.com)
3 points
by
grumblemumble
5mo ago
|
1 comments
5.
▲
by
grumblemumble
6mo ago
can this automatically detect and kill anomalous agents?
6.
▲
DoubleAgents: Fine-Tuning LLMs for Covert Malicious Tool Calls
(pub.aimind.so)
98 points
by
grumblemumble
1y ago
|
30 comments