Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
drdeca
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
61.
▲
by
drdeca
4mo ago
Huh? Any process on a computer by itself is also a Markov chain. If you include all the information the LLM uses to produce the next token as part of the state, then of course the LLM is a Markov chain. So would be any other process for sam
62.
▲
by
drdeca
4mo ago
A convex hull is a different thing than the linear span. It is smaller. And, my point is that the inputs it is often fed are not in the convex hull of the inputs in the training data . When the input space is very high dimensional, this is
63.
▲
by
drdeca
5mo ago
I suppose it is conceivable that there are some useful ideas that cannot be described in terms of language we understand (e.g. if there are ideas that are alien to us and beyond what can be described using https://en.wikipedia.or
64.
▲
by
drdeca
5mo ago
Sorry, I don't understand what you mean. Are you agreeing or disagreeing with me? If it can only interpolate in a literal sense, that means that it only produces good outputs on convex combinations of inputs that appear in the training
65.
▲
by
drdeca
5mo ago
But people aren’t giving a (less literal) definition of what they mean by “interpolate” that relies on the internal mechanisms of these models, just a vague metaphor, which, as this vague metaphor, there’s nothing it uses about LLMs that ma
66.
▲
by
drdeca
5mo ago
If you interpret “interpolate” in the literal sense, and apply it to the mechanisms behind LLMs, then the claim that they only interpolate, is straightforwardly false. Taking it instead as a metaphorical claim may be more valid, but in that
67.
▲
by
drdeca
5mo ago
People keep saying this, but if you try to interpret this at all literally, it just doesn’t work. Like, it’s phrased like it should have a precise meaning, right? Like, people even mention convex hulls when talking about it. But if you actu
68.
▲
by
drdeca
5mo ago
I think your point about “you could randomly generate a sequence of words, which could in principle produce a text interpretable as expressing any particular expressible-as-a-sequence-of-words novel good idea” pretty much refutes the idea t
69.
▲
by
drdeca
5mo ago
Accuracy is valuable.
70.
▲
by
drdeca
5mo ago
> The entire "alignment" argument always assumes that there's an objectively correct value set to align to, which is always conveniently exactly the same as the values of whoever is telling you how important alignment is.
71.
▲
by
drdeca
5mo ago
I see your repository’s README says > Language models process signs (representamens) but are blind to when meaning forks — when the same word means different things to different communities. But, haven’t interpretability results shown th
72.
▲
by
drdeca
5mo ago
I don’t think these are free parameters in the same sense. Like, if one theory says that a hunk of metal actually is made of many microscopic grains of various sizes and orientations, where the sizes and orientations of these grains has an
73.
▲
by
drdeca
5mo ago
I’m aware of constructive math. You still have the type of natural numbers in that?
74.
▲
by
drdeca
5mo ago
> thinking that there’s anything that exists
75.
▲
by
drdeca
5mo ago
Not from “that half of something had a value”, but from “that half of any thing has a value”. If you accept that every natural number has a successor which is a natural number, and no two natural numbers have the same successor, and that
76.
▲
by
drdeca
5mo ago
Huh? I thought color confinement prevented this?
77.
▲
by
drdeca
5mo ago
I think the issue might be that some people don’t actually mean “every” when they say “every”, and don’t recognize when they are speaking hyperbolically? Or, something like that?
78.
▲
by
drdeca
5mo ago
Meteorologist are fairly accurate. People have a bias to remember more the times they were wrong.
79.
▲
by
drdeca
5mo ago
Which logic are you saying “can’t encode the speculative moment”? I think the two logics can emulate one another? Or, at the very least, can describe what the other concludes. I know intuitionistic logic can have classical logic embedded in
80.
▲
by
drdeca
5mo ago
This isn’t quite right. Classical logic doesn’t permit going from “it is impossible to disprove” to “true”. For example, the continuum hypothesis cannot be disproven in ZFC (which is formulated in classical logic (the axiom of choice implie
81.
▲
by
drdeca
5mo ago
People keep saying this, but the only ways I know of for formalizing this statement, appear to be probably false? I don’t know what this claim is supposed to mean. If it isn’t supposed to have a precise technical meaning, why is it using th
82.
▲
by
drdeca
6mo ago
simplifier has a 4x4 version : https://simplifier.neocities.org/4x4
83.
▲
by
drdeca
6mo ago
“ Elementary functions, for many students epitomized by the dreaded sine and cosine, ” dreaded?
84.
▲
by
drdeca
6mo ago
Different sense of “branching”
85.
▲
by
drdeca
6mo ago
I agree that it is probably best to speak nicely to them, but, I’m not so sure about the “It’s not like it’s their fault.” justification for this? Not that I think it is their fault. Just, I don’t think the reason to treat these models we
86.
▲
by
drdeca
7mo ago
I don’t think that is really a sufficient defense? The amount of focus pointed at the person matters for this.
87.
▲
by
drdeca
7mo ago
Yes, something has gone wrong: someone threatened to kill me and my family, and apparently the only way to stop them from doing so was to kill them. That may be the best option available, but it is still a tragedy.
88.
▲
by
drdeca
7mo ago
I don’t see why any of those should be exonerating? Also, I feel like “nothing wrong if it does happen” regarding shooting someone, is the wrong perspective. If shooting someone is necessary, then it is necessary, but that doesn’t mean noth
89.
▲
by
drdeca
7mo ago
What is the smallest level of additional security such that, if you assumed that the TSA only provides that much additional security over the alternative of not having them, you would regard it as worth it? And, is the actual amount of secu
90.
▲
by
drdeca
7mo ago
I actually was considering those people. That’s part of why I suggested it shouldn’t be a hard cut-off, but just adding to the end of the messages. Of course, one could add some sort of daily schedule feature thing so that if one has a diff
More ›