Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
uh_uh
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
61.
▲
by
uh_uh
1y ago
Was it hallucinating here, or are the commenters hallucinating? What OP is saying is just not true. A CT scan and normal daily commute in Grand Central station are NOT comparable in terms of radiation received. Somehow this is controversia
62.
▲
by
uh_uh
1y ago
Weird you don't have this requirement for the OP spewing his urban myths above.
63.
▲
by
uh_uh
1y ago
In your opinion how many hours spent in Grand Central station equal the radiation received from a CT scan?
64.
▲
by
uh_uh
1y ago
> Remember that you'll get comparable levels of radiation even if you commute through the grand central station every day. Gemini says this: > A single typical CT scan delivers a dose that is roughly 1,000 to over 5,000 times hig
65.
▲
by
uh_uh
2y ago
1. "It is not saying that humans confidently give wrong answers if they do not know correct ones." And I didn't say that they do either, so you might have hallucinated that. 2. What are you arguing about? I didn't say th
66.
▲
by
uh_uh
2y ago
> We do know how LLMs work, correct? NO! We have working training algorithms. We still don't have a complete understanding of why deep learning works in practice, and especially not why it works at the current level of scale. If you
67.
▲
by
uh_uh
2y ago
Examples of emergence: 1. Multi-step reasoning with backtracking when DeepSeek R1 was trained via GRPO. 2. Translation of languages they haven't even seen via in-context learning. 3. Arithmetic: heavily correlated with model size, but
68.
▲
by
uh_uh
2y ago
Sorry but this is nonsense. Do you have a theory about when certain LLM capabilities emerge? AFAIK we don't have a good theory about when and why they do emerge. But even if knew how something works (which in present case we don't
69.
▲
by
uh_uh
2y ago
The LLM is a glorified autocomplete in as much as you are a glorified replicator. Yes, it was trained on autocomplete but that doesn't say much about what capabilities might emerge.
70.
▲
by
uh_uh
2y ago
1. Many humans don't have an idea of the limits of their competence. It's called the Dunning–Kruger effect. 2. LLMs regularly tell me if what I'm asking for is possible or not. I'm not saying they're always correct,
71.
▲
by
uh_uh
2y ago
We don't really have a clue what they are and aren't capable of. Prior to the LLM-boom, many people – and I include myself in this – thought it'd be impossible to get to the level of capability we have now purely from statist
72.
▲
by
uh_uh
2y ago
This top to bottom drawing – does this tell us anything about the underlying model architecture? AFAIK diffusion models do not work like that. They denoise the full frame over many steps. In the past there used to be attempts to slowly synt
73.
▲
by
uh_uh
2y ago
I hope you're joking. Sometimes they don't even know which company developed them. E.g. DeepSeek was claiming it was developed by OpenAI.
74.
▲
by
uh_uh
2y ago
Could be. Also, as you imply, they'd have to loosen the regularization penalty on θ, and maybe it's difficult to loosen it such that it won't become too prone to overfitting. Maybe their current setup of keeping θ "dumb&
75.
▲
by
uh_uh
2y ago
Given that all parameters are trained jointly at inference time and a single sample of z is supposed to encode ALL inputs and outputs for a given puzzle (I think), I don't quite understand the role of the latent z here. Feels like μ an
76.
▲
by
uh_uh
2y ago
This is actually a pretty cool accidental mirror test.
77.
▲
by
uh_uh
2y ago
Pundits were saying that deep learning has hit a plateau even before the LLM boom.
78.
▲
by
uh_uh
2y ago
This. Articles like this are examples of motivated reasoning and seem to be coming from a place of insecurity by programmers who feel their careers threatened.
79.
▲
by
uh_uh
2y ago
So why are they SOTA translators? Would you consider old translation software bullshit generators? Because LLMs can do their job, and more.
80.
▲
by
uh_uh
2y ago
The choice unfortunately seems to correlate with the person's age. Younger generations will have no trouble treating LLMs as actually intelligent. Yet another example of "Science progresses one funeral at a time.”
81.
▲
by
uh_uh
2y ago
> There is no abstracted concept of a "colour" in there. There's just a lot of imagery tagged with each colour name, and if you select a different colour you get a vector in a space pointing to different images. It has bee
82.
▲
by
uh_uh
2y ago
Is there anything AI (or more like AGI) could do for you that would change your opinion? Eternal life, interstellar travel, curing diseases, educating people with superhuman patience and competence would not be worth it?
83.
▲
by
uh_uh
2y ago
Aren't you at least a little bit excited for the potential benefits?
84.
▲
by
uh_uh
2y ago
So you don't have a concrete suggestion to solve the scamming problem?
85.
▲
by
uh_uh
2y ago
Are you suggesting government action against putting up code like this to GitHub? It’s ok if you are, but I want to put into more concrete terms what we’re talking about.
86.
▲
by
uh_uh
2y ago
Cooperation works if the potential damage caused by a rouge actor is sufficiently low. Otherwise, it's too easy to sabotage things. This is why we don't want random rouge states to have nukes. AI will give so much leverage to roug
87.
▲
by
uh_uh
2y ago
Demanding responsible behaviour from everybody is not going to work. Some people don't care about negative externalities that much and it's enough if only a few of them decide not to play ball. So either grandma needs to adapt whi
88.
▲
by
uh_uh
2y ago
This tech is going to be ubiquitous, it's just too easy to distribute it. Grandma better starts adapting now.
89.
▲
by
uh_uh
2y ago
Interesting write-up. Does this mean that all of OP's competitors are facing the same legal issues? I wonder what (if anything) they do about this.
90.
▲
by
uh_uh
2y ago
Why is "without prompting" a requirement? If a hypothetical LLM comes up with the best neologisms that one can imagine, but only after prompting, then they're disqualified?
More ›