Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
clbrmbr
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
27 ms
·
31.
▲
by
clbrmbr
2mo ago
I think so, if you can inject synthetic errors with the same distribution as the errors you care about.
32.
▲
by
clbrmbr
2mo ago
It is surprisingly hard to do in a single prompt. I’ve had good luck though with being very explicit about asking for review, then to think deeply about what actually matters, and then reduce to a very short extremely concise reply. —- the
33.
▲
by
clbrmbr
2mo ago
I’ve found that doing the full requirements capture, planning, writing, reviewing, gardening loop with frontier models has worked quite well since last October, and phenomenally since Fable 5. The key I’ve found is human peer review. The re
34.
▲
by
clbrmbr
2mo ago
Mais par contre, Canada also has some tremendous senatorial committees. They asked real questions of Hinton.
35.
▲
by
clbrmbr
3mo ago
Matches my experience. Basement home offices are the worst offenders.
36.
▲
by
clbrmbr
3mo ago
Do we have recent commentary on open-weight from the Anthropic founding team?
37.
▲
by
clbrmbr
3mo ago
With $onnet he would have gotten pwned. Or at least I’d love to see a comparison against other models.
38.
▲
by
clbrmbr
3mo ago
It is customary that one may publish one’s own personal correspondence unless the other party has requested confidentiality. Maybe this open invitation to the world pushes the boundaries of that definition, but I don’t see where an expectat
39.
▲
by
clbrmbr
4mo ago
Hmm I imagine using a server to connect to signal/whatsapp or even email, then using a local model to classify and filter and trim messages and forwarding to SMS, and viceversa. I guess the trouble is I’d need many source numbers :thin
40.
▲
by
clbrmbr
4mo ago
Fascinating. Can you explain why southern London is DC while northern London is AC?
41.
▲
by
clbrmbr
4mo ago
Can you share an example? I've been happily using Fable this afternoon and it just seems like the usual upgrade so far with no interruption to my (fairly standard) SWENG problems.
42.
▲
by
clbrmbr
4mo ago
Very nice. Can I fork for my private use? Was thinking of building similar.
43.
▲
by
clbrmbr
4mo ago
Philanthropically-minded people will move the winners to a donor advised fund which gives FMV write off without ever paying capital gains. With index funds you never have the strong winners to do this with, and so giving is far less tax-ef
44.
▲
by
clbrmbr
4mo ago
But do they historically beat the S&P 500?
45.
▲
by
clbrmbr
4mo ago
Which suburb, haha. I’m in Englewood and have similar experience of very few tech folks around.
46.
▲
by
clbrmbr
4mo ago
Was a good watch, tho would have liked to be there in person. Props to Brenden & his Cosmos team for really setting the bar.
47.
▲
by
clbrmbr
4mo ago
One capability that I see is missing from opus is this ability to understand an entire system. My hope is that a mythos class model will be able to comprehend even something as complicated as an IOT system with a hardware and firmware layer
48.
▲
by
clbrmbr
4mo ago
This. I added that instruction the first and last time I was gaslit by an underpowered subagent.
49.
▲
by
clbrmbr
5mo ago
> You asked a simple question. They lobbed a document. I’ve become so cynical that I cannot read this phrasing without thinking it came from… an LLM. This cynicism also negatively impacts human communication. This constant doubt of wheth
50.
▲
by
clbrmbr
5mo ago
Come on, Anthropic ARE the good guys if there are any. Certainly the incentives of trillions will do what money does, but they have assembled an incredibly altruistic and philosophically-minded crew. I’m rooting for them and trying real har
51.
▲
by
clbrmbr
5mo ago
Karpathy embedded within an organization is way more impressive than him out on his own with hot takes and little projects. I hope he does great things for Anthropic.
52.
▲
by
clbrmbr
5mo ago
Anyone here using OpenBSD? If so, for what purpose? I’ve always wanted to use NetBSD for an application for an embedded system / IoT device but never had the pleasure (yet!).
53.
▲
by
clbrmbr
5mo ago
Thanks was a good watch. Sad though the example of the AI app to “help farmers” that is making things up. I would expect a generational cassava farmer to have a much better sense of how to treat the plants than an image model.
54.
▲
by
clbrmbr
5mo ago
This mirrors my experience at Stevens. The professor would not babysit us during exams and that really did inspire pride. Also the exams were often brutally hard which inspired despair.
55.
▲
by
clbrmbr
5mo ago
High-end Chromebook done right could be a very good thing for computer security.
56.
▲
by
clbrmbr
5mo ago
I thought similarly, until I actually tried using AI to shop for clothes, now I’m a total convert. It’s like the best possible men’s fashion concierge…
57.
▲
by
clbrmbr
5mo ago
Do you know if this exploit works on Docker containers? And if so, I assume it just allows escalation WITHIN the container? So this attack is scary for Linux desktops and servers, but a fully containerized system like common on CI/CD s
58.
▲
by
clbrmbr
5mo ago
It’s interesting to compare how the agentic search performs, with these targeted reads and lots of tool calls in the stream, versus the older but still valid paradigm of using a high-reasoning model like GPT-X-pro and feeding in all the rel
59.
▲
by
clbrmbr
5mo ago
So what do we do? Pin our dependencies (to hashes when possible), and only update when there are CVEs? But problem is this could lead to abuse of the CVE system to try to force rapid adoption of attacked packages. What prevents this?
60.
▲
by
clbrmbr
5mo ago
Wouldn’t you prefer to pin to SHA hashes? Or does your package manager cloud-side ensure immutability of releases?
More ›