Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
ekidd
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
ekidd
3mo ago
Everyone always says things like "It's just marketing". But seriously, I highly recommend reading the published details of the Huggingface breach. The model found and chained multiple zero days. To escape, it punched a hole
32.
▲
by
ekidd
3mo ago
> If AI were to become a super weapon why should I trust a private company to own it? We have just recently established that: 1. OpenAI's internal "Galaxy" model is fully capable of functioning as what security people re
33.
▲
by
ekidd
3mo ago
Oh, to be clear, I don't think that anything about this overall situation is even slightly OK.
34.
▲
by
ekidd
3mo ago
Apparently, the attacker in the Hugging Face case was reported to be an internal OpenAI model trying to break into HF and steal the answers to cybersecurity benchmarks: https://openai.com/index/hugging-face-model-evalua
35.
▲
by
ekidd
3mo ago
> For small models (which are probably distilled from their big ones) you can serve them economically all the time and not hemorrhage money. For smaller models, you're competing with DeepSeek V4 Flash. (Which I think is a 284B A13
36.
▲
by
ekidd
3mo ago
Those are Claude Code security issues? And Claude Code is a gigantic vibe-coded dumpster fire with more bugs than an ant hill? pi-agent famously doesn't even try to provide a sandbox, so obviously it can't have sandbox security
37.
▲
by
ekidd
3mo ago
It is a pretty catastrophically dumb CVE, of the sort that makes me not want to allow OpenCode anywhere near any of my machines in the future. It was basically "RCE as a service", not some subtle bug. Personally, I run pi-agent
38.
▲
by
ekidd
3mo ago
AI subscription pricing was fine when it was $100/month for some opaque 5 hour token budget I don't think I ever used, not even that one day where I coded for 14 hours non-stop using Fable. But like most people with low token usag
39.
▲
by
ekidd
3mo ago
I have written assembly for about 5 different processors, including 65C02s, 680x0s, cute little DSP3210s that managed the CPU cache manually, utterly cursed TI320C40s, and (of course) a bit of Intel. WebAssembly is simpler than some of thes
40.
▲
by
ekidd
3mo ago
Fable wrote a pretty decent test suite covering typical Prolog programs, including things with non-trivial execution patterns like "append" that do complex backtracking on multiple branches. And I've run a modest number of te
41.
▲
by
ekidd
3mo ago
Claude is perfectly capable of writing assembly. Here's a working (basic) Prolog interpreter that Claude Fable 5 wrote in WebAssembly in 61 minutes for $16.75 in token costs: https://github.com/emk/fable-wasm-prolo
42.
▲
by
ekidd
3mo ago
You can maybe run a local Sonnet-4.5-ish-level model ( sort of) for less than the price of a new car, even at current massively inflated prices for fast RAM. This is probably not what you were looking for. But it's there. You could sh
43.
▲
by
ekidd
3mo ago
GLM 5.2 isn't quite modern Opus tier, as seen in this comparison where Opus 4.5 scores 4/5 on some coding tasks where GLM 5.2 scores 0/5: https://www.tryai.dev/blog/gpt-5.6-build-off-12-models But yes,
44.
▲
by
ekidd
3mo ago
A really common use case is install/bootstrap/setup scripts. You know, those sketchy curl-to-bash things, or cloud-init scripts, or whatever you run to set up your actual higher level tools like Python.
45.
▲
by
ekidd
3mo ago
> I don't take 3 as a given. There's just too much going on in the space for one cloistered company to control it all and be in control of it. Yeah, in my comment, I was assuming the publicly-stated goals of the labs actually
46.
▲
by
ekidd
3mo ago
The author is completely right about the AI Lab's promised vision of the world: They claim to want to create superhuman intelligence, which will produce vast abundance. But superhuman intelligence would be extremely dangerous, so it
47.
▲
by
ekidd
3mo ago
We have benchmarks for our use cases, and every generation after Gemini 2.0 Flash has been a grim hit on price/performance. Costs have gone up, throughput has gone down, and performance has improved very slightly (and regressed on a fe
48.
▲
by
ekidd
3mo ago
As a professional programmer entering the final third of an enjoyable career, I would now place "learning to code" in the same category as "making a living as a poet." As in, it's truly enjoyable art and some people
49.
▲
by
ekidd
3mo ago
Frontier models like Fable are mostly useful if you want to paste in one or two prompts, and receive a subtly broken application that looks impressive. That is very hard to do with local models today. What current local models work fine f
50.
▲
by
ekidd
3mo ago
By "blew it with Washington" you mean "Didn't donate millions to the ballroom."
51.
▲
by
ekidd
3mo ago
> The same change could be affected by e.g. schools and businesses agreeing to open at 8am instead of 9am. School starts at 8am everywhere that I know of in northern New England and always has? Does school start at 9am where you live?
52.
▲
by
ekidd
3mo ago
Yeah, as someone who lives in Vermont, you could talk me into permanent DST. That would move the winter sunset from, say, 4:21pm to 5:21pm, which would mean I'd get enough twilight for a short walk after work. And Maine is even furth
53.
▲
by
ekidd
3mo ago
> I saw what I forgot immediately; but soon after, with engagement, I saw how quickly I was able to remember. We actually have pretty good models for how long it takes to forget things. It's the same basic math that powers Anki. T
54.
▲
by
ekidd
3mo ago
> It's best not to blame the students. They are good at optimizing metrics; that's how they ended up here in the first place. As an alumnus of Dartmouth College's CS program, I am sad to hear that my alma mater has sunk
55.
▲
by
ekidd
3mo ago
> The penalty for cheating should be automatic expulsion. Historically, the penalty for cheating at Dartmouth was a 9-month suspension for a first offense (no matter how small, in theory), and permanent "separation from the colleg
56.
▲
by
ekidd
4mo ago
> May be it is that we are ingenious amd creative with tools and thats how we evolve. And every time you use the AI to be ingenious or creative, that will be added to the training data. Then someday the AI can be ingenious and creative
57.
▲
by
ekidd
4mo ago
A GPU with 24GBs of RAM is mostly useful for running a very carefully squeezed Qwen3.6 27B (4-bit Unsloth quants, 8-bit K/V cache, possibly MTP, 128k context). This is a fun little model that's smart enough to do debugging, refact
58.
▲
by
ekidd
4mo ago
> If so, the thinking trace can be sort of nonsensical for a reader, though whether this is an idiosyncrasy of the model or a property of LLMs in general isn't clear to me yet. Yes, several models think in weird jargon. Here is an
59.
▲
by
ekidd
4mo ago
The difference is that a compiler is a rigorous, (nearly) determinisic, heavily tested artrifact built by expert humans. I have only encountered genuine code generation bugs in compilers twice in my career. And yes, those bugs I did trace t
60.
▲
by
ekidd
4mo ago
> Telling people “you must read all the code generated by an LLM” is definitely meaningful—but it is not at all moderate (so most people won’t do it). I am honestly heartbroken to live in a world where reading the code is seen as an
More ›