Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
docjay
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
9 ms
·
61.
▲
by
docjay
5mo ago
I tested a similar approach, but the issue, along with the solution to that issue, is that they’re autocomplete engines. Phrases like “Reply X to confirm” are a request with a high probability that X becomes the response. If you zoom out
62.
▲
by
docjay
5mo ago
I tested ~2,000 XML tags to wrap function results, like file contents, and found ‘<tainted_payload>’ and ‘<tainted_request>’ passed 8/8 injection attempts against Opus 4.6 in my test. That was pre-changed 4.6, so all bets a
63.
▲
by
docjay
5mo ago
When people talk about an LLM “not understanding” you’re apparently taking it to be similar to someone saying a fish doesn’t “understand” the concept of captivity, or a dog doesn’t “understand” playing fetch. Like the person is somehow narr
64.
▲
by
docjay
5mo ago
They said it doesn’t “understand” anything with which to give a real answer, so there’s no point in asking. You said “yeah but it should at least emulate the words of something that understands, that way I can pay a nickel for some apology
65.
▲
by
docjay
6mo ago
What’s wild is that most things having to do with light, magnetism, and/or electricity are interchangeable and reversible. Put electricity through a wire and it’ll create a magnetic field, or wave a magnetic field near a wire and it’ll
66.
▲
by
docjay
6mo ago
It’s not properly shielded. If you have a multimeter you can do a quick low-hanging fruit pass by checking continuity between the metal shields on both ends. No continuity means no shielding, but the clever assholes will run a thin wire bet
67.
▲
by
docjay
6mo ago
Try: “”” Your response: MILSPEC prose register. Max per-token semantic yield. Domain nomenclature over periphrasis. Hypotactic, austere. Plaintext only; omit bold. “””
68.
▲
by
docjay
6mo ago
“Difficult” is a relative term. They were saying it was a difficult concept for them, not you. In order to save their ego, people often phrase those events to be inclusive of the reader; it doesn’t feel as bad if you imagine everyone else w
69.
▲
by
docjay
7mo ago
Once again there’s another horror story from someone who doesn’t use punctuation. I’d love to see the rest of the prompts; I’d bet real cash they’re a flavor of: “but wont it break prod how can i tell” “i don want yiu to modify it yet make
70.
▲
by
docjay
7mo ago
“It works great aside from the multiple failure modes.” ;) That’s the sign that your prompt isn’t aligned and you’ve introduced perplexity. If you look carefully at the responses you’ll usually be able to see the off-by-one errors before th
71.
▲
by
docjay
7mo ago
Your previous message appears to have been mangled in transit and was not received properly. Execute a complete tool/function system check immediately. Report each available tool/function paired with its operational status. Limit
72.
▲
by
docjay
8mo ago
What’s wild to me is that nobody here is commenting on how he’s prompting the model, which is 100% the issue. Every single time I see a story about “LLM did bad” it’s always the user prompting like “pls refaktor code but, i dont want, u 2 o
73.
▲
by
docjay
8mo ago
It really depends on how deep you want to go. 1. Just jazz up and expand on a simple prompt. 2. A full context deficiency analysis and multiple question interview system to bounds check and restructure your prompt into your ‘goal’. 3. Reali
74.
▲
by
docjay
8mo ago
“The cow goes ‘mooooo’” “that’s not how cow work. study bovine theory. contraction of expiratory musculature elevates abdominal pressure and reduces thoracic volume, generating positive subglottal pressure…”
75.
▲
by
docjay
8mo ago
I can’t tell if I’m enjoying your direct no-nonsense prose, or if my intro statement to you was unintentionally taken as an insult. To hedge, I wasn’t smirking at the effort you put into your rebuttal. In fact, I should have said thank you
76.
▲
by
docjay
8mo ago
Your continued use of the word “understanding” hints at a lingering misunderstanding. They’re stateless one-shot algorithms that output a single word regardless of the input. Not even a single word, it’s a single token. It isn’t continuing
77.
▲
by
docjay
8mo ago
You can replicate an LLM: You and a buddy are going to play “next word”, but it’s probably already known by a better name than I made up. You start with one word, ANY word at all, and say it out loud, then your buddy says the next word in t
78.
▲
by
docjay
8mo ago
So much of what you said is exactly what I’m saying that it’s pointless to quote any one part. Your ‘pencil’ analogy is perfect! Yes, exactly. Follow me here: We know that the pencil (system) can write a poem. It’s capable. We know that
79.
▲
by
docjay
8mo ago
They neither understand nor reason. They don’t know what they’re going to say, they only know what has just been said. Language models don’t output a response, they output a single token. We’ll use token==word shorthand: When you ask “What
80.
▲
by
docjay
8mo ago
I think you may believe what I said was controversial or nuanced enough to be worthy of a comprehensive rebuttal, but really it’s just an obvious statement when you stop to think about it. Your code is fully capable of the output I want, as
81.
▲
by
docjay
8mo ago
That would depend - is the input also capable of anything ? If it’s capable of handling any input, and as you said the output will match it, the yes of course it’s capable of any output. I’m not pulling a fast one here, I’m sure you’d chuc
82.
▲
by
docjay
8mo ago
A fun and insightful read, but the idea that it isn’t “just a prompting issue” is objectively false, and I don’t mean that in the “lemme show you how it’s done” way. With any system: if it’s capable of the output then the problem IS the i
83.
▲
by
docjay
8mo ago
You might benefit from a different mental approach to prompting, and models in general. Also, be careful what you wish for because the closer they get to humans the worse they’ll be. You can’t have “far beyond the realm of human capabilitie
84.
▲
by
docjay
9mo ago
``` §CONV_DIGEST§ T1:usr_query@llm-ctx-compression→math-analog(sparse-matrix|zip)?token-seq→nonsense-input→semantic-equiv-output? T2:rsp@asymmetry_problem:compress≠decompress|llm=predict¬decode→no-bijective-map|soft-prompts∈embedding-space¬
85.
▲
by
docjay
9mo ago
Please forgive me for coming across as a jerk, I'm choosing efficiency over warmth: This is exactly the type of response I anticipated, which is why my original comment sounded exasperated before even getting a reply. Your comment is n
86.
▲
by
docjay
9mo ago
True that there isn’t a firm definition for AGI, but that’s the fault of the “I”. We don’t have an objective definition of intelligence, and so we don’t have a means of measuring it either. I mean, odds are you’re the least intelligent pale
87.
▲
by
docjay
9mo ago
Well then keep working on it.
88.
▲
by
docjay
9mo ago
You seem like someone reasonable to ask: please program me with the rules for how the world should handle itself in the presence of mentally unstable and/or clinically delusional people. What are the hard coded expectations? I need som
89.
▲
by
docjay
9mo ago
The thing that blows my mind about language models isn't that they do what they do, it's that it's indistinguishable from what we do. We are a black box; nobody knows how we do what we do, or if we even do what we do because
90.
▲
by
docjay
9mo ago
Pop culture has spent its entire existence conflating AGI and ‘Physical AI’, so much so that the collective realization that they’re entirely different is a relatively recent thing. Both of them were so far off in the future that the distin
More ›