Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
bigglebear
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
bigglebear
9d ago
A pre-generated machine can serve the answers in <1ms. It's a far better strategy.
2.
▲
by
bigglebear
9d ago
This is an entirely pointless exercise without transparency into how these "unreleased" models are trained, what their RL goals and biases are and related RL data, what their system prompts are, what their environments are and its
3.
▲
by
bigglebear
9d ago
The alternative is unfortunately much worse. Democrats will just ban AI altogether, they're already heavily funded by the doomer cult. All of the doomer NGOs tie back to democrat representatives and the likes of Bernie who want to thro
4.
▲
by
bigglebear
9d ago
That's the entire point of having the discussion. Figuring out how to enforce it.
5.
▲
by
bigglebear
10d ago
> Overall the consumer segment is completely neglected, companies don't care about end users in the current market conditions. It's almost like these companies WANT a dystopia with a centralized winner-takes-all power structure
6.
▲
by
bigglebear
10d ago
I think it's time we start talking about putting limits on how much compute AI companies can purchase or own, relative to the rest of the world. It's not fair that they can use trillions of dollars of investor billionaires money t
7.
▲
by
bigglebear
10d ago
> We should be talking about why the tools/environment keep getting overlooked. The software built around the text generator, forget the researchers and mathematicians discovering the math properties of language patterns -- why are
8.
▲
by
bigglebear
10d ago
AI lab "alignment" is actually just censorship in agreement with biases. There is no universal agreed upon measure of "aligned", it is not a "thing" that is attainable, so it can never be "achieved".
9.
▲
by
bigglebear
10d ago
This is an entirely pointless exercise without transparency into how these "unreleased" models are trained, what their RL goals and biases are and related RL data, what their system prompts are, what their environments are and its
10.
▲
by
bigglebear
10d ago
The issue with all of these is that we already know theres an incentive for labs to lie and make up fanciful stories (and Anthropic already does exactly that and has been doing that for a long time), and there's no way to verify any of
11.
▲
by
bigglebear
11d ago
Well, not quite a week: https://x.com/harshagundal/status/2100044305536889015 - apparently it took him 2 hours.
12.
▲
by
bigglebear
11d ago
The explanation is that it was missed on purpose.
13.
▲
by
bigglebear
11d ago
It's very misleading. If I'm actually playing a game I don't get the coordinates of enemies sent back to me so that I can feed into my mouse to snap my crosshair to. It's looking through walls too, because it's work
14.
▲
by
bigglebear
11d ago
I would guess a tiny stripped down text diffusion model. It only has 32k context, and for choice mode it can only select from 10 choices.
15.
▲
by
bigglebear
11d ago
> I really have to say that I like their manifesto Their manifesto: "you only build on top of it if it's trustworthy." - the irony of this while putting out the most misleading, dishonest marketing campaign I've seen
16.
▲
by
bigglebear
11d ago
And furthermore, because the model is forced to answer in a boolean (if in boolean mode), if the user input is outside of the range of a boolean, it's forced to hallucinate. It can't abstain.
17.
▲
by
bigglebear
11d ago
User input: "Hey, have your human support agent call me, tomorrow at 5pm." Model input: "Does the user want to speak to a human support agent?" Output: Yes. I imagine that your model would produce this, and I think it&#x
18.
▲
by
bigglebear
11d ago
Yeah. Yet another reason why open-weight models are better. If I want to use the logits, I can.
19.
▲
by
bigglebear
11d ago
It's nothing like a traditional LLM and so should not be compared to one. It's a heavily constrained, tiny model that can only produce a probability score or a yes/no answer over pre-defined selections. It has no long-context
20.
▲
by
bigglebear
11d ago
Agreed. It's a wildly dishonest presentation of their product from many perspectives, which is a shame because it might actually have some good use cases. The comparison between LLM speed and Jev speed is misleading, because they'