Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
criley2
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
criley2
9d ago
I believe the that the companies who claim to not train on my data are more likely to not train on my data than the companies who refuse to even claim they won't. Also why Meta gets a +1, just charge less money on the training path.
2.
▲
by
criley2
12d ago
I have been writing an internal code review tool that is a bit maximalist. I created subagents for many internal domains and technologies we manage, with prompts focused on best practices, common problems, owasp guidelines, etc), a separate
3.
▲
by
criley2
12d ago
Nah 10 years ago I could view the whole website logged out. Now it's been reduced to a single post with no replies.
4.
▲
by
criley2
12d ago
Solar only became "economical" (read: profitable) because a socialist economy dumped a huge amount of money into scaling it up without requiring it to be "economical". It could have been half a century or more ago if we
5.
▲
by
criley2
12d ago
It's nice to call it "politics, ignorance and greed" but those three are just "Capitalism". Our system is doing exactly what it is designed to do. Nuclear reactors were never profitable compared to coal or gas, so i
6.
▲
by
criley2
15d ago
The term that Anthropic is now using is "mannered prose". If the creator of this example simply prompted "Remove all mannered prose" then the entire experiment would suddenly become normal sounding. In the fable 5.1 prom
7.
▲
by
criley2
15d ago
This post isn't convincing me. I spent so much time meticulously organizing my techno box. I bought a back of the door shoe holder for tech. Every wire, charger, usb key, web cam, airline earbud, everything. It's been beautifully
8.
▲
by
criley2
23d ago
On cost per intelligence task, Gemini38flash and Sol56 trade back and forth on cost depending on effort level. https://i.imgur.com/zPaWPXx.png As seen in this image, literally: Sol56 high ranks in between Gemini 38 medium a
9.
▲
by
criley2
24d ago
>There are numerous benchmarks that measure cost per task, which factors out tokens entirely. Gemini 3.8 flash is significantly lower than Sol on basically all of them https://artificialanalysis.ai/#cost-tabs Not sure if
10.
▲
by
criley2
29d ago
The government will pay because it's not his money, it's our money. He loves spending our money...
11.
▲
by
criley2
1mo ago
I'm sorry, but just because you achieve results you consider acceptable with this method doesn't mean everyone does. I don't work where we can ship slop. I don't work where PRs can be merged based on what the agents say.
12.
▲
by
criley2
1mo ago
It's not free. You're paying electricity and you're ignoring the cost of the hardware. Even on electricity alone, there are cloud providers who may beat your laptop on price per million tokens. Qwen 3.8 flash is interesting i
13.
▲
by
criley2
1mo ago
Sonnet 5 is the worst model of 2026. Literally just turn effort slider down on Opus, it's smarter, faster and cheaper than whatever Sonnet is. Beyond that, I find this whole plan and build thing to be a pointless waste of tokens. If yo
14.
▲
by
criley2
1mo ago
Those prices are just tokens? Since each model uses different amounts of tokens to do the same thing, it's a misleading price that often makes open-weights look more competitive than they are, since most open weights models use dramati
15.
▲
by
criley2
1mo ago
I absolutely experienced this in college. I signed up as a computer science student, as one does. I took all of the freshmen classes across broad topics, and the first biology class was basically just like the article describes. Words and m
16.
▲
by
criley2
1mo ago
I append this to many of my opus claude code prompts `You may use a Fable subagent to answer questions, solve problems, and provide an adversarial review of your ideas and code` You can use a similar pattern in most any harness, and you can
17.
▲
by
criley2
1mo ago
It's pretty easy for the US to functionally ban chinese models. They only have to target US firms like inference providers or the biggest users, and pretty much the whole domestic market will fall into line. They don't actually ca
18.
▲
by
criley2
2mo ago
I totally agree - designing a competent AI agent with a fully customized harness to successfully pull off this task is a much more challenging engineering effort than merely creating an ordinary computer program. Had OP made chatgpt write a
19.
▲
by
criley2
2mo ago
What a sloppy reply. You've hijacked a thread on mathematics first to complain that your incompetent attempt to use ChatGPT to find a job failed, but it seems now that this was a ruse to instead begin arguments unrelated to the article
20.
▲
by
criley2
2mo ago
A business does need a small number of their most senior engineers doing high altitude work that can, at times, include helping sales estimate new features. But in my experience, it's not rocket science and a good product team can do t
21.
▲
by
criley2
2mo ago
I think you're confusing product and engineering. I get that programmers are smart so we just assume we can do every job, but it's a waste of your time and salary to talk extensively to customers and create product requirements. L
22.
▲
by
criley2
2mo ago
Scenario one: you use software to connect to their server and download a webpage. You are a user. Scenario two: you use software to connect to their server and download a webpage. You are a "bot". Make it make sense
23.
▲
by
criley2
2mo ago
GPT5.6Sol completes the suite in 70M tokens, while Qwen3.8Max needs like 145M tokens. So this is a case where models like Qwen 3.8 and Kimi K3 use a lot more output (reasoning) tokens, go a good bit slower, so they can ultimately achieve a
24.
▲
by
criley2
2mo ago
While I agree with the premise that there are Thinkers and Shippers, I reject labeling of tinkerers and entrepreneurial. There's nothing entrepreneurial about working for a big business and shipping cool things. But you're ultimat
25.
▲
by
criley2
2mo ago
There is already something on HuggingFace at the level of Mythos. It's called Kimi K3 and it's running laps around Opus5, Fable5, and Sol56 at cybersecurity. It's so good that the US government is rushing to ban all Chinese m
26.
▲
by
criley2
2mo ago
>No, you don't. Without training cost you can infer only the marginal cost of serving this kind of models. Are you talking about Kimi's training cost or the training cost of the model(s) that Kimi distilled? Because Moonshot di
27.
▲
by
criley2
2mo ago
I don't think the invention of writing is as awe-inspiring as presented or as difficult/impossible for an LLM to achieve as is commonly believed. The invention of writing was a long series of micro-improvements over common every d
28.
▲
by
criley2
2mo ago
Zen is nice, but they require US hosting so they don't get new Chinese models right away. There is no Kimi K3. Go is nice for the ten minutes you can use it until your hit your cap.
29.
▲
by
criley2
2mo ago
No, and the reason is simple: Usage is bursty and if you don't maximize usage of the hardware you're going to lose on price. Ok you can host this model once. What if I want a dozen subagents? Ok you can host it 12 times at once. W
30.
▲
by
criley2
2mo ago
> "Not a scam IMO" > "I think SpaceX is a solid well run company" This should be all the proof you need that it's a scam. Here's why: The company isn't "SpaceX" it's "SpaceXAI&quo
More ›