Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
swingboy
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
61.
▲
by
swingboy
3mo ago
“Anthropic wants to MAKE A DEAL!”
62.
▲
by
swingboy
3mo ago
These guesstimate release dates seem…soon. Usually the Polymarket markets are pretty accurate and really accurate when an insider at one of the companies puts a bet.
63.
▲
by
swingboy
3mo ago
How is this enforceable?
64.
▲
by
swingboy
3mo ago
And it’s a paper from Alibaba researchers, the company/lab that Anthropic called out by name.
65.
▲
by
swingboy
3mo ago
I’ve always assumed any LLM output that was some type of rating or score was bullshit. Unless the LLM writes a Python script to calculate the score (and even then…) then the score it outputs is just the next most likely token, taking into a
66.
▲
by
swingboy
3mo ago
Did he not listen to music made by others when he was isolated in the cabin?
67.
▲
by
swingboy
3mo ago
All members of the Epstein class.
68.
▲
by
swingboy
3mo ago
Nice try, Barbra.
69.
▲
by
swingboy
3mo ago
How would export controls apply if OpenAI or Anthropic released a model as open weights? Not that they would, but asking out of curiosity.
70.
▲
by
swingboy
3mo ago
Because the American people weren’t outraged enough to push them out.
71.
▲
by
swingboy
3mo ago
Here’s to hoping that Alibaba (and other Chinese labs) have collected some really good distilled data.
72.
▲
by
swingboy
3mo ago
My work involves asking LLMs about both Tianenmen Square and what’s going on in Gaza, so I can’t use Chinese or American models!
73.
▲
by
swingboy
4mo ago
I get a 500 when clicking “Explore the Models”
74.
▲
by
swingboy
4mo ago
Anthropic’s best practices still include the use of XML: https://platform.claude.com/docs/en/build-with-claude/prompt...
75.
▲
by
swingboy
4mo ago
*Advice only applies to neighborhoods without an HOA.
76.
▲
by
swingboy
4mo ago
Does “the best machine for AI use” apply here considering these models are still server-side?
77.
▲
by
swingboy
4mo ago
GPT5.5 xhigh seems to benchmark about on par with Mythos for cybersecurity.
78.
▲
by
swingboy
4mo ago
The Trump administration would never do anything to manipulate the markets. /s
79.
▲
by
swingboy
4mo ago
I realize these models are locked up pretty tight and terabytes in size, but in a future like that, I don’t see them not being leaked via an insider. The weights have to be loaded into VRAM at some point.
80.
▲
by
swingboy
4mo ago
Same model that costs $12 in tokens to finally add “overflow-x: hidden;” to an element, by the way. https://news.ycombinator.com/item?id=48498573
81.
▲
by
swingboy
4mo ago
overflow is CSS 101
82.
▲
by
swingboy
4mo ago
Immediately I thought “isn’t this just an overflow issue?” Amazing how far these models still have to go and also how many people don’t know basic CSS.
83.
▲
by
swingboy
4mo ago
Interesting that the `brew-rs` experiment has concluded and didn't find much of a performance increase. I suppose that is expected though with a lot of the bottleneck being network IO?
84.
▲
by
swingboy
4mo ago
What file format(s) are giant LLM models distributed in? I’m surprised they don’t get leaked by employees.
85.
▲
by
swingboy
4mo ago
[flagged]
86.
▲
by
swingboy
4mo ago
What is revising in this context? For Americans.
87.
▲
by
swingboy
4mo ago
Yes, but harnesses don’t automatically include the README.md in the system prompt like they do AGENTS.md.
88.
▲
by
swingboy
4mo ago
Apologies for the naivety, but, why is SpaceX valued so high? Starlink? Are rockets really a lucrative business? Don’t get me wrong, being able to send objects up into orbit is cool, but is it $1.8T cool?
89.
▲
by
swingboy
4mo ago
Nice! Looks like it’s topping the two coding ones. I noticed it is absent from the Social Intelligence board though?
90.
▲
by
swingboy
4mo ago
Looking forward to the results. Thanks for your work.
More ›