Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
andy99
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
11 ms
·
181.
▲
by
andy99
3mo ago
I think it’s just as well he was opinionated and realized it wasn’t going to work. Ironically, if this was actually the start of the job, I think bailing might have made more sense. Given that it was just a one week work trial, I feel like
182.
▲
by
andy99
3mo ago
> every recruiter I've worked with has been shit That’s a feature, not a bug. It’s an obviously solvable problem but companies don’t solve it because incentives push them to this model. There are of course exceptions where recruitme
183.
▲
by
andy99
3mo ago
LLM writing style is trained in by data labellers, it’s not just emergent behavior from being trained on internet texts.
184.
▲
by
andy99
3mo ago
This isn’t just em-dashes—it's the empty phrase that includes both whatever the contrastive construction is called and an “Honestly”. It might have been human written but the density of LLM flags is undeniable.
185.
▲
by
andy99
3mo ago
Can you explain that please? (Not the syllogism, the premise)
186.
▲
by
andy99
3mo ago
I also saw the tells but found it direct enough that it wasn’t really a concern. LLM writing style is a good signal that something is slop and should be ignored but isn’t exactly causal... it would be an interesting exercise to try and writ
187.
▲
by
andy99
3mo ago
Not to argue the point but that statement isn’t logical, look at all the complaints about restaurants. Publicly complaining about something doesn’t require it be a monopoly.
188.
▲
by
andy99
3mo ago
Anthropic injects text into the conversation triggered by certain conversation topics. This happened to me in relation to some red-teaming related discussion that was adjacent to something “sensitive”, I think sex, and Claude got confused a
189.
▲
by
andy99
3mo ago
Worth looking at https://www.anthropic.com/engineering/a-postmortem-of-three-... They can “go insane” but it seems often to be infra related as opposed to anything one would consider hallucination. Smaller models will
190.
▲
by
andy99
3mo ago
Interesting to see the claudeslop reply as the first comment to the gh post and the reaction to it.
191.
▲
by
andy99
3mo ago
Not sure the relevance of this comment, but normally if someone built a classifier that bad they’d be fired. Anthropic obviously thinks they have some monopoly power they can use to foist garbage on consumers, I think they don’t.
192.
▲
by
andy99
3mo ago
I realize hallucination has no precise definition but this doesn’t sound at all like anything I’ve ever heard called hallucination. Hallucination is usually plausible wrong answers or made up info that ends up fitting the most likely respon
193.
▲
by
andy99
3mo ago
The people that are homeless now would have been institutionalized or dead back then.
194.
▲
by
andy99
3mo ago
Read “The Gervais Principle” https://ribbonfarm.com/2009/10/07/the-gervais-principle-or-t... I think it explains everything, most companies are optimizing for “confused” - a class within the framework of peop
195.
▲
by
andy99
3mo ago
I worked in a grocery store meat department in high school. I wasn’t a butcher, my main job was wrapping the meat for the counter, weighing and pricing it, but over time I learned to do basic meat cutting, still nowhere near a full butcher
196.
▲
by
andy99
3mo ago
You can run one on a cloud provider. You’re correct that intelligence orgs probably still can access them, but if you’re that high value of a target then you have bigger problems and / or can afford to build an air gapped system or wha
197.
▲
by
andy99
3mo ago
It’s not no reason. At a fundamental level I don’t trust the companies any differently. But at a professional level, nobody is going to question my using Claude or OpenAI in a professional capacity - to work on customer projects, analyze th
198.
▲
by
andy99
3mo ago
I’ve never had a refusal coding, and in some areas (AI red teaming specifically) I’ve found it quite good at recognizing and discussing “white hat” stuff that in the past I think would have got refusals. But when there was the Hantavirus th
199.
▲
by
andy99
3mo ago
Do you guys use it through open router? Do you have any concerns about how the data you send is being intercepted? Not that I trust Anthropic but it’s widely agreed that it’s kosher to use them for commercial work, I can’t see comfortably s
200.
▲
by
andy99
3mo ago
All the discussion this week have been about GLM, Qwen, etc. Both over 1000 comments in the last couple days. https://news.ycombinator.com/item?id=48709670 https://news.ycombinator.com/item?id=48721903 Of c
201.
▲
by
andy99
3mo ago
By “abusive” they probably mean “doesn’t let us track them”. I’m surprised they’ve kept old.reddit.com going this long, it’s actually an un-enshittified version of the product, which is basically unheard of.
202.
▲
by
andy99
3mo ago
That was with the MTP version
203.
▲
by
andy99
3mo ago
I get ~55 Tok/s on my framework desktop with the 35B A3B q8 model, and so far am also very happy with the coding performance.
204.
▲
by
andy99
3mo ago
I wondered if the whole thing was just a vibe coded weekend project. It’s hard to gauge the seriousness and level of effort but without some evidence to the contrary I’m guessing it’s all LLM generated.
205.
▲
by
andy99
3mo ago
Right, but is there any evidence of intelligence behind any of these (government) decisions? It’s just regulatory capture + marketing (plus some people living out an imaginary fantasy that they’re in Neuromancer or something), absolutely no
206.
▲
by
andy99
4mo ago
Other than maybe some in-the-moment cybersec wrappers, is this really true? Does anyone think a startup with a good product is going to be materially disadvantaged by not having access to an incrementally better security focused LLM release
207.
▲
by
andy99
4mo ago
> Chinese labs must entirely retool from harvesting frontier model data to producing the data systems and efforts to produce novel data Even if your characterization is accurate, they could do this tomorrow and are not so myopic that the
208.
▲
by
andy99
4mo ago
In the early days of the LLM era, there was lots of talk about how big incumbents, in particular google would be disadvantaged relative to “startups” like OpenAI because of their valuable legacy businesses that could be destroyed if somethi
209.
▲
by
andy99
4mo ago
How compatible is never replying with the threat model you are trying to avoid? Attack success is probably more likely when the attacker can iterate based on replies or engage in multi-turn conversations. Here they’re just taking stabs in t
210.
▲
by
andy99
4mo ago
I think they may have overplayed their hand so to speak. The end consequence is that their best model isn’t available right now, people are exploring alternatives, and realizing they work fine. It’s such a fast paced and competitive industr
More ›