Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sigmar
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
sigmar
6d ago
>If I make my code available, AI can be used to more easily discover vulnerabilities to abuse my coding errors, Keeping your source code private doesn't save you. These LLMs are really good at decompilation efforts. I've seen a
2.
▲
by
sigmar
8d ago
>I’ve emailed some of the mathematicians with a few proposed typo fixes, and I got confirmation that at least a few of those fixes seemed real. However, some of the problems that weren’t backed by Lean also turned out to be misunderstand
3.
▲
by
sigmar
10d ago
yeah, I can see it as useful in customer service, but they're adding these sponsored agents directly in the chatgpt app. the word "ad" or "sponsored" isn't visible after you click "chat with us" at 0:
4.
▲
by
sigmar
10d ago
I'm definitely more pro ad than the average HN user (I think they occasionally show me products I want to buy), but this seems bad and one step closer to "the LLM is biased by openai's financial interests." In the exampl
5.
▲
by
sigmar
12d ago
that says only that the gross margin calculation excludes profit sharing and training. You should read it more carefully edit: def not gaap profitable or they would have said that to investors. and their stock-based comp is surely astronomi
6.
▲
by
sigmar
12d ago
>This is reportedly a sort of "Enron" accounting which excludes some really big expenses like revenue sharing, the cost of model training and hardware deploymments source? this seems false. reportedly the adjusted profitability
7.
▲
by
sigmar
12d ago
>lists a few methods that are quite similar to what’s proposed in the comments under that post, except for those two comments. Is today the first time you've heard of a book cipher? Those blog comments didn't provide much progr
8.
▲
by
sigmar
13d ago
>Just remove hacking (bio-weapon, etc.) data from the training dataset and you're done. Reasoning about how to write secure software uses the same knowledge as reasoning about how to break/hack it.
9.
▲
by
sigmar
13d ago
Lots of private benchmarks already exist, where you have to trust the tester (ex Artificial Analysis, Arc-agi).
10.
▲
by
sigmar
14d ago
Sure, if you remove almost all of my comment my argument disappears. To spell things out for people that don't know about the topic: Zitron is neither an expert on the topic, nor a credible source of information: https://tec
11.
▲
by
sigmar
14d ago
>To understand the truth about the Hugging Face hack, you could do a lot worse than to listen to Ed Zitron and Cal Newport's recent podcast conversation Lol, okay... The crux of this piece is Doctorow saying that the hack was just a
12.
▲
by
sigmar
15d ago
>or there is a more deliberative approach to assigning credit than who was "first" to solve some problem it feels like an unintended consequence of the millennium prize is that people view the [last contributor to the solution]
13.
▲
by
sigmar
16d ago
Were the agents ever tasked with algorithm improvements? Post just says he didn't find any ("report essentially no algorithm advancements"). These LLMs are useful for optimization tasks where they can attempt a change and the
14.
▲
by
sigmar
17d ago
>We have a lot of physics based simulation tools, but they tend to focus on small subsets of the full design problem and they make limiting approximations Do you think it is possible that better math will lead to better physics models?
15.
▲
by
sigmar
22d ago
>The speed with which we were able to produce this proof demonstrates that it is now possible to formalize large swaths of mathematics, which may both catch errors in the common body of mathematical proofs and reduce the burden of refere
16.
▲
by
sigmar
29d ago
https://en.wikipedia.org/wiki/Attempted_assassination_of_Don... what evidence is there that this registered republican was "far left"?
17.
▲
by
sigmar
2mo ago
>We are currently working closely with our inference partners and open-source maintainers to align the technical details and ensure the model can be reliably deployed across the ecosystem. The full model weights will be released by July
18.
▲
by
sigmar
2mo ago
Most of what an LLM does "could have" been done by a human if you throw enough human hours at it. But the reality in this circumstance is that a new tool helped find this leak. Saying this could have happened in a "non LLM wo
19.
▲
by
sigmar
2mo ago
>Evaluators validate each tag individually — for example, protein, preparation, or health, individually rather than judging the item as a whole. Am I reading this right that the jury is multiple LLMs each iterating through each tag and v
20.
▲
by
sigmar
3mo ago
>I think it's reasonable for people to say, hey, if you're going to trash the reputation of Zig (in a pretend-objective way) what specifically is this referring to? Not aware of any comments from Anthropic on this topic.
21.
▲
by
sigmar
3mo ago
To me, this addendum makes it worse. Making small edits to a post like this makes it seem like you're doubling down on the original resentful points, especially with all the new justifications like "a trillion dollar company fired
22.
▲
by
sigmar
3mo ago
>One, this constant bullshit about some window closing, or the perpetual underclass, or falling hopelessly behind. This is negative valence hype, not only is it not true, it’s mostly designed to make you feel bad about yourself and move
23.
▲
by
sigmar
3mo ago
Proven to be safe? Do you want a randomized trial? Do you demand that of every math tutorial video that goes up on YouTube?
24.
▲
by
sigmar
3mo ago
How many years do you use them?
25.
▲
by
sigmar
3mo ago
They don't make any specific claims about what conditions it will diagnose. At 16:30, he says they are only initially doing "body composition" because anything more would add 9+ months to the deployment timeline. I assume the
26.
▲
by
sigmar
3mo ago
that's an obsequious Altman, not their model being banned for being too good
27.
▲
by
sigmar
3mo ago
I visited CERN last July. Was lucky enough to get into a group tour. The tour guide was a postdoc researcher who said the only times that public tours are allowed to take an elevator down is during long shutdowns. So while they do this work
28.
▲
by
sigmar
3mo ago
>publish these incredible papers explaining how they achieved their gains - something the American labs no longer do unfortunately. Google is still releasing a lot of llm architecture research. They introduced speculative decoding of LLM
29.
▲
by
sigmar
3mo ago
The ATF was created by an act of congress. https://en.wikipedia.org/wiki/Gun_Control_Act_of_1968
30.
▲
by
sigmar
3mo ago
>the language in the docs is awfully indirect. writes this^ and then proceeds to highlight a bold title from the docs that says "summarized thinking" that explains things clearly in the first sentence. lol
More ›