Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Imnimo
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
Imnimo
10mo ago
I wonder what it would take to adapt a model like this to generate non-Earthlike terrain. For example, if you were using it to make planets without atmospheres and without water cycles, or planets like Io with rampant volcanism.
62.
▲
by
Imnimo
10mo ago
Why should it be the case that LLMs are equally comfortable in x86 Assembly and Python? At least, it doesn't strike me as implausible that working in a human-readable programming language is a benefit for an LLM that is also trained on
63.
▲
by
Imnimo
10mo ago
We're really drawing a fine distinction if something "looks like" an ad but isn't an ad. Isn't that the whole point of an ad - it's appearance?
64.
▲
by
Imnimo
10mo ago
I guess I thought the pipeline was typically Pretraining -> SFT -> Reasoning RL, such that it would be expensive to test how changes to SFT affect the model you get out of Reasoning RL. Is it standard to do SFT as a final step?
65.
▲
by
Imnimo
10mo ago
>we did train Claude on it, including in SL. How do you tell whether this is helpful? Like if you're just putting stuff in a system prompt, you can plausibly a/b test changes. But if you throwing it into pretraining, can Anthro
66.
▲
by
Imnimo
11mo ago
>Why doesn’t someone else create a competing app that’s better and thereby steal all their business? How do I know if the competing app is actually better? I mean, this was the advertising angle for eHarmony about a decade ago - that it
67.
▲
by
Imnimo
11mo ago
The good news is we can just wait until the AI is superintelligent, then have it explain to us what consciousness really is, and then we can use that to decide if the AI is conscious. Easy peasy!
68.
▲
by
Imnimo
11mo ago
I think one could certainly make the case that model capabilities should be open. My observation is just about how little it took to flip the model from refusal to cooperation. Like at least a human in this situation who is actually fooled
69.
▲
by
Imnimo
11mo ago
>At this point they had to convince Claude—which is extensively trained to avoid harmful behaviors—to engage in the attack. They did so by jailbreaking it, effectively tricking it to bypass its guardrails. They broke down their attacks i
70.
▲
by
Imnimo
11mo ago
These are both a lot more fun, and a lot more educational than leetcode problems. Strongly recommend for anyone looking for practice problems when learning a new language or whatever.
71.
▲
by
Imnimo
11mo ago
According to the "best days" link in the article, November 7th is the best day to cut your hair because the moon phase and zodiac will lead to slower hair growth if you cut it today. I am amazed this publication made it this far.
72.
▲
by
Imnimo
11mo ago
And still today we spend a great deal of effort trying to make our randomly-sampled LLM outputs reproducibly deterministic: https://thinkingmachines.ai/blog/defeating-nondeterminism-in...
73.
▲
by
Imnimo
1y ago
As with any quantum computing news, I will wait for Scott Aaronson to tell me what to think about this.
74.
▲
by
Imnimo
1y ago
The trouble is Karpathy already speaks at 1.5x speed.
75.
▲
by
Imnimo
1y ago
>What takes the long amount of time and the way to think about it is that it’s a march of nines. Every single nine is a constant amount of work. Every single nine is the same amount of work. When you get a demo and something works 90% of
76.
▲
by
Imnimo
1y ago
I feel like a danger with this sort of thing is that the capability of the system to use the right skill is limited by the little blurb you give about what the skill is for. Contrast with the way a human learns skills - as we gain experienc
77.
▲
by
Imnimo
1y ago
This is basically Ouija board for LLMs. You're not making it more true, you're making it sound more like what you want to hear.
78.
▲
by
Imnimo
1y ago
I'm curious whether this is work that was specifically begun under the "superintelligence" umbrella, or if it's just that the people who were working on it had been shifted to the Superintelligence team by the time they
79.
▲
by
Imnimo
1y ago
I would say its the applicants who seem to know which ones are good.
80.
▲
by
Imnimo
1y ago
Suppose that a school takes those people and fails to get half of them to graduation? Have they been given a leg up? Is that a "great" outcome?
81.
▲
by
Imnimo
1y ago
If the mark of a great school were producing a lot of degrees, what would WKU's ~50% graduation rate mean?
82.
▲
by
Imnimo
1y ago
I would expect a great school to be appealing to potential students, and therefore attract a lot of them. I would also expect a great school to have high academic standards that not all applicants meet.
83.
▲
by
Imnimo
1y ago
In my book a great school would likely have a low-ish acceptance rate. And so they could (even though they may not be happy about it) absorb some amount of declining applications by adjusting their acceptance bar. WKU's acceptance rate
84.
▲
by
Imnimo
1y ago
I only know the story from word-of-mouth. My understanding is that the submitted paper wasn't written/presented well, and reviewers had trouble understanding the significance of what was being proposed. But take that story with a
85.
▲
by
Imnimo
1y ago
An interesting bit of trivia about the Burrows-Wheeler transform is that it was rejected when submitted to a conference for publication, and so citations just point to a DEC technical report, rather than a journal or conference article.
86.
▲
by
Imnimo
1y ago
>In a world where there’s enough AI capability to process the entire web and rewrite every page to remove something, the cost of “changing history” is much reduced, so we can expect more of it. I gotta be honest, this scenario is not a c
87.
▲
by
Imnimo
1y ago
I'm just not sure that I would trust that the view/description of the item ChatGPT shows me and the thing I'm actually agreeing to buy are the same thing.
88.
▲
by
Imnimo
1y ago
We can look at what he did when he had the opportunity to help reduce it, for one.
89.
▲
by
Imnimo
1y ago
It's very hard for me to envision something I would use this for. None of the examples in the post seem like something a real person would do.
90.
▲
by
Imnimo
1y ago
Sorry, I realized I didn't quite write what I meant to. I didn't intend to say that LLMs are non-Markovian from a theoretical standpoint. I meant to say that the language generation task is non-Markovian from a theoretical sta
More ›