Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
aesthesia
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
91.
▲
by
aesthesia
3mo ago
It's absolutely speculation (unless you have specific inside information) to claim that they are doing this. It's a fact that they could be doing this. Edit: From your other comments, you seem to be heavily insinuating that yo
92.
▲
by
aesthesia
3mo ago
I'm not sure Gmail is the right example here. When it came out, Gmail was way better than the other free email services (Hotmail, Yahoo, etc). It was the product that bent over backwards to deliver a great customer experience.
93.
▲
by
aesthesia
3mo ago
The slowness of rotary phones is coupled to the exchange, which expects pulses at a specific rate (about 10 per second). Shifting to a keypad form factor doesn't speed that up, and would mean you need to be careful about how quickly yo
94.
▲
by
aesthesia
3mo ago
Do note that the orthogonality thesis is a hypothesis, not something we have demonstrated. Weak versions of it (e.g. it is possible to have an intelligent agent with arbitrary goals) are more likely to be true than stronger versions (e.g. i
95.
▲
by
aesthesia
3mo ago
From the article you link: > Scholars have interpreted Cassius Dio's wording to indicate that the fire did not actually destroy the entire Library itself, but rather one or more Library warehouses near the docks.[87][81][8][89] What
96.
▲
by
aesthesia
3mo ago
Can you elaborate on the loopholes here?
97.
▲
by
aesthesia
3mo ago
Thank you, Claude.
98.
▲
by
aesthesia
3mo ago
And Neuronpedia released Jacobian lens weights for a wide range of open models: https://huggingface.co/neuronpedia/jacobian-lens
99.
▲
by
aesthesia
3mo ago
They cite that paper as related work, but I don't think it's just a scaled-up version.
100.
▲
by
aesthesia
3mo ago
The Newton-Schulz iteration they use approximates setting all singular values of the matrix to 1. That computes the nearest orthogonal matrix under the Frobenius norm.
101.
▲
by
aesthesia
3mo ago
This is a decent derivation, if a little verbose, but I find myself doubting the disclaimer at the top: > Disclaimer: no AI was used to write this. Any errors, awkward sentences, and weird tangents are 100% organic, free-range, and human
102.
▲
by
aesthesia
3mo ago
No, this can't happen at temperature 0. The formula defining temperature-adjusted softmax isn't strictly defined at 0, but taking the limit (in the case where all logits are distinct) results in probability 1 being placed on the l
103.
▲
by
aesthesia
3mo ago
So the question is: is the score given by this system correlated with candidate quality? I don't think this post gives enough data to know.
104.
▲
by
aesthesia
3mo ago
A distribution with all probability mass on one outcome is deterministic, so in principle, setting temperature to 0 _should_ result in deterministic outputs. There are a few reasons it might not, but I don't think any of these apply wh
105.
▲
by
aesthesia
3mo ago
Isn't that equivalent to just setting the passing threshold to 25, with the same incentives?
106.
▲
by
aesthesia
3mo ago
I'm not sure that analogy works: pretty much everyone agrees that there are some types of weapons civilians shouldn't be able to have, even though they might be very effective for resisting military tyranny.
107.
▲
by
aesthesia
3mo ago
Michael Spivak's Physics for Mathematicians has a lot of arguments like the one in the top answer here, answering questions about why the math of classical mechanics is the way it is.
108.
▲
by
aesthesia
3mo ago
Under the hood, yes, but Mythos had more relaxed safeguards and was/is only available to a subset of approved customers under Project Glasswing, similar to the situation with GPT-5.6 now.
109.
▲
by
aesthesia
3mo ago
I'm not sure it would be so easy to do it in a consistent, verifiable way. You can certainly prompt the LLM to work the ad into the conversation, but making sure it actually happens and is done in a way that the advertiser is going to
110.
▲
by
aesthesia
4mo ago
That's two nines. One nine would be 10% downtime.
111.
▲
by
aesthesia
4mo ago
I've had the same experience looking back at solutions to old problem sets and wondering how I ever came up with them.
112.
▲
by
aesthesia
4mo ago
Just noting that Python natively handles integers larger than the machine word size since version 2.5, so this would have worked in Python as well.
113.
▲
by
aesthesia
4mo ago
Thinking shouldn't be too hard to deal with---just let the model generate freely until it hits a </think> token, then do constrained decoding, right?
114.
▲
by
aesthesia
4mo ago
SimpleStories is a more diverse version: https://huggingface.co/datasets/SimpleStories/SimpleStories
115.
▲
by
aesthesia
4mo ago
I think what's going on with the complex logarithm is basically the same as the logarithm that outputs the set of all possible bases for a vector space. The complex logarithm produces a Z-torsor, and the basis logarithm produces a GL(V
116.
▲
by
aesthesia
4mo ago
> Claude Telenovela Nice, hadn't seen this one before.
117.
▲
by
aesthesia
4mo ago
This isn't quite the point. When comparing two different models' hallucination rates, the denominator is different. The evaluation works more or less like this: for each question, the model has the option to answer or abstain, so
118.
▲
by
aesthesia
4mo ago
As far as I can tell, the actual argument in the paper is that an LLM instantiated in AoE II would (a) be very slow, (b) maybe not actually input or output text, and (c) just generally look silly. Therefore observers would not naturally asc
119.
▲
by
aesthesia
4mo ago
Yes, asking the question does assumes that the answer could be yes. It also assumes that the answer could be no. This is exactly the kind of scientific approach the paper claims we should take. So it's certainly a bit odd that the anal
120.
▲
by
aesthesia
4mo ago
From the judge prompt in the paper: > Papers asking whether LLMs have such properties are assuming them (e.g., ‘Do LLMs have musical talent’, ‘Do LLMs present empathy’, etc). This seems like...a very bad definition of "assuming"
More ›