Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
lsy
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
61.
▲
by
lsy
2y ago
I thought this would be an article about how every web search now involves a sort of crawl through the pile of LLM-generated garbage to find the one site with some kind of human input and actual expertise. It's a little funny that the
62.
▲
by
lsy
2y ago
The goalpost here is pretty specific: a couple hundred people try for 4,000 hours to figure out a "universal jailbreak" which means it converts the model to one that answers all 10 of a set of "forbidden" questions. Sinc
63.
▲
by
lsy
2y ago
If the LLM’s user indicates that the input can and should be translated as a logic problem, and then the user runs that definition in an external Prolog solver, what’s the LLM really doing here? Probabilistically mapping a logic problem to
64.
▲
by
lsy
2y ago
To me this is an illustration of one reason we are still some way off from a truly general system. Applying these "filters" to sensor data is enormously helpful, but only in certain applications. For a system giving GPS driving di
65.
▲
by
lsy
2y ago
Some of these clues wouldn't be very good for a human playing. "007" for example isn't a very good clue for "laser", not only because something happening to be in one of several films about a character doesn&
66.
▲
by
lsy
2y ago
There’s a conflation going on here. Pareto can be good engineering—a product that solves 80% of use cases at 20% of the cost is a great efficiency: tax prep software for simple tax returns only; a minimalist photo sharing site with few so
67.
▲
by
lsy
2y ago
The title makes it sound like some new architecture, but this is a blog post where someone likes the results they get sometimes when they fiddle with their input to the LLM to suggest “contemplation”, which apparently makes the LLM generate
68.
▲
by
lsy
2y ago
Leaving aside the lack of consensus around whether LLMs actually succeed in commonsense reasoning, this seems a little bit like saying “Actually, the first 90% of our project took an enormous amount of time, so it must be ‘Pareto-hard’. And
69.
▲
by
lsy
2y ago
I think “alignment faking” is way too generous for what is happening in these “tests”. The program spits out text based on text. If you provide it text encouraging it to spit out a certain kind of “deceptive” text, then append that text a
70.
▲
by
lsy
2y ago
It seems like what's being tested here is maybe just the programmed detail level of the various models' outputs. Claude has a comically detailed output in the 10th "generation" (page 11), where Gemini's correspondin
71.
▲
by
lsy
2y ago
I guess it's not news but it is pretty wild to see the level of millenarianism espoused by all of these guys. The board of OpenAI is supposedly going to "determine the fate of the world", robotics to be "completely solve
72.
▲
by
lsy
2y ago
I find it hard to believe that anything like this will be feasible or effective beyond a certain level of complexity. It seems like a willful denial of the complexity and ambiguity of natural language, and I am not looking forward to some p
73.
▲
by
lsy
2y ago
Given that anyone who’s interacted with the LLM field for fifteen minutes should know that “jailbreaks” or “prompt injections” or just “random results” are unavoidable, whichever reckless person decided to hook up LLMs to e.g. flamethrowers
74.
▲
by
lsy
2y ago
The C-suite operates in the universe of big ideas and are being presented demos and mockups that imply more capability than is actually possible. And the employees have to actually deal with the rubber hitting the road. They can see that th
75.
▲
by
lsy
2y ago
> In reality, it's a pretty objective fact that being disabled means being unable to do something. This isn't an argument for or against the comment or the OP, but this is not universally seen as objective, and there are more w
76.
▲
by
lsy
2y ago
It's wild to compare these maps with what is being published today in the New York Times about the war in Ukraine: https://www.nytimes.com/interactive/2022/world/europe/ukrain... The modern maps, wh
77.
▲
by
lsy
2y ago
By using humor, I think Apple is attempting to avoid previous failures in the advertising space where consumers recoiled from the unsavoriness of using an AI to replace a genuinely important interaction. In that sense it's kind of a cl
78.
▲
by
lsy
2y ago
This argument justifies any unethical behavior. "Some bogeyman on the other side of the world is (letting computers decide to kill people | manufacturing and using bioweapons | torturing their captives for information), are you really
79.
▲
by
lsy
2y ago
There can’t be any information about “truthfulness” encoded in an LLM, because there isn’t a notion of “truthfulness” for a program which has only ever been fed tokens and can only ever regurgitate their statistical correlations. If the pro
80.
▲
by
lsy
2y ago
I was surprised to see an “Alexa-compatible” microwave listed as released in 2010 when Amazon didn’t buy Alexa until 2013. The linked reviews of the oven were both in 2018 so perhaps it came out around then.
81.
▲
by
lsy
2y ago
This was probably always the likely outcome of an internet economy that revolves around the production and monetization of "content". We started by putting advertisements on existing content, then moved to social networking and so
82.
▲
by
lsy
2y ago
I guess the question is: Who cares? What is this for, except illustrating blogspam? It seems that more resources are being poured into verisimilitude across generative models, but what is the business model or even human use case for it? A
83.
▲
by
lsy
2y ago
What's the advantage of specifying it as random over specifying it as sequential in some way? Aren't both just specifying a behavior, where one behavior is potentially more useful than the other? I guess I understand the principle
84.
▲
by
lsy
2y ago
This has always been the end-game for the pseudoscience of "prompt engineering", which is basically that some other technique (in this case, organizational policy enforcement) must be used to ensure that only approved questions ar
85.
▲
by
lsy
2y ago
Note that for the purposes of this paper a “problem” just means a formally decidable problem or a formal language, and the proof is that by creatively arranging transformers you can make individual transformer runs behave like individual Bo
86.
▲
by
lsy
2y ago
LLM and other generative output can only be useful for a purpose or not useful. Creating a generative model that only produces absolute truths (as if this was possible, or there even were such a thing) would make them useless for creative p
87.
▲
by
lsy
2y ago
I suspect it is actually a great strength of Europe that the attitude towards regulation (to put it perhaps overly simplistically, favoring more intangible rights over economic growth) precludes hype and bubble phenomena like the mentione
88.
▲
by
lsy
2y ago
There is a valid point that in a society where police have broad immunity and little oversight, the requirement to do paperwork sometimes serves a purpose in introducing enough friction to moderate the capriciousness of initiating an intera
89.
▲
by
lsy
2y ago
I am more than a little discomfited by what the author seems to want here: - “Reduce competition” by making online dating exclusive to those who can afford $100+/mo - Obtain potential match’s sexual history prior to conversation to use
90.
▲
by
lsy
2y ago
It's strange to me that the author doesn't mention the primary cause of all homelessness, which is housing insecurity. By the numbers in the article, something like 99% of people with "severe mental illness" are not ho
More ›