Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
sheepscreek
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
31.
▲
by
sheepscreek
1mo ago
I thought the same. But why claim something so shocking when it can easily be discredited and puts your reputation at risk? If they’re claiming Opus 4.6 level, I expect it to at least match Sonnet 4.6.
32.
▲
by
sheepscreek
1mo ago
Better than Opus 4.6 at computer use? Comparable with it for SWE? Am I reading this right? I’ve heard rumours about AI shops optimizing for benchmarks. I also don’t think Qwen/Alibaba would be crazy enough to claim something unless the
33.
▲
by
sheepscreek
1mo ago
I’ve read a few different accounts, including OpenAI’s own admission, that Terra Medium or higher will likely produce better results than Luna xhigh and cost about the same or less.
34.
▲
by
sheepscreek
1mo ago
Yes and AFAIK any fil-C program needs to be linked against fil-C compiled libraries, such as libc/musl/etc. I too would like to know what other challenges exist in making this process more automated.
35.
▲
by
sheepscreek
1mo ago
Understand the problem and the solution broadly. I don’t think it’s reasonable or sustainable for humans to understand every line of code written by bots, we could soon be outnumbered by the number of active agents writing code. The main ch
36.
▲
by
sheepscreek
1mo ago
I don’t get the appeal of Lovable. How is it different from bolt.new, or Figma’s AI feature, of GitHub Spark, or the countless others? Gaining traction or having a large marketshare cannot be the USP.
37.
▲
by
sheepscreek
2mo ago
It could be that they’re pitching Meta Muse 1.2 against Terra and Opus level models. They probably consider Sol to be a level above, along with Fable.
38.
▲
by
sheepscreek
2mo ago
I couldn’t help but feel a bit emotional reading this. Elise was clearly an incredible human. Thank you Stephen for sharing her story and your love for her. I cannot imagine a better way to have spent those 36 years. Rest in peace Elise.
39.
▲
by
sheepscreek
2mo ago
I was super excited to read this but lost the plot when I go to > Qwen3.8-Max was asked to create the oh-my-cli project from scratch and, over a 10+ day long-horizon autonomous coding run 10+ days of building what exactly? Is that a shel
40.
▲
by
sheepscreek
2mo ago
I didn’t imply it was novel.
41.
▲
by
sheepscreek
2mo ago
High quality research - my mind is blown by the concept of replacing “reasoning” with “…” (literally) with no effect on actual output.
42.
▲
by
sheepscreek
2mo ago
It gives every employee their own personal assistant (AI agent). It learns their preferences, and more importantly, can access whatever they can.
43.
▲
by
sheepscreek
2mo ago
I don’t think this is true if a disclaimer is provided. Most investment advice videos on YouTube make this very clear, but it doesn’t stop them from sharing advice. I suppose there is some precedent for this in the American justice system.
44.
▲
by
sheepscreek
2mo ago
You're good. Dude that topped Meta's tokenmaxxxing board before it was shut down used 265 billion tokens in a month. I kid you not.
45.
▲
by
sheepscreek
2mo ago
I really hope the lawsuit does not end in OpenAI tightening the “safety” and “security” of their models into oblivion. I hope they understand that a lobotomized model is as good as my Casio watch. I want my models to give it to me straight
46.
▲
by
sheepscreek
2mo ago
Hey Thariq. Don't dwell too much on the negativity here. It's a hard act to balance and I'm sure the frightening pace has only made your job harder. I think there might be a case for multiple CC release streams, same as Chrom
47.
▲
by
sheepscreek
2mo ago
AFAIK there’s no ZDR with Claude models accessed directly via Anthropic. You’d have to go through either Google Vertex, Azure or AWS for true ZDR (at least legally/on paper).
48.
▲
by
sheepscreek
2mo ago
I’ve had great luck building SwiftUI apps with GPT-5.5 (now GPT-5.6 Sol), Opus 4.8 and Fable 5. I’m just offering another data point, not suggesting on the effectiveness of DSLs for this. One line of thinking can be that frontier models are
49.
▲
by
sheepscreek
2mo ago
Agreed. I think we’re entering an era where some level of specialization for general LLMs is a good thing. Particularly between tuning for agentic use cases (where you want agency with a ton of guardrails and control) and writing which is m
50.
▲
by
sheepscreek
3mo ago
I wonder if they are double-counting Anthropic's leased capacity from SpaceX under SpaceX again.
51.
▲
by
sheepscreek
3mo ago
I don't think it even matters. Because noone will continue to use an LLM that doesn't work well for them, whether or not it has a good bench result. So for their own sake, the correct representation can actually win them some loya
52.
▲
by
sheepscreek
3mo ago
Moneyshot: > The three bacterial strains that successfully induced tumor regression (E. americana, C. portucalensis, and E. ludwigii) were all identified as facultative anaerobic bacteria. > This finding is consistent with established
53.
▲
by
sheepscreek
3mo ago
Yes - they create new/cheaper products for a different consumer but not at the cost of their margins. Vision Pro may have been the only device in recent history that likely had slim margins (if any).
54.
▲
by
sheepscreek
3mo ago
> It's easy to blame datacenters, but there are many factors at play here. Henrico County currently has 37 data-centres with ~2 gigawatts capacity (expected to reach 3 gigawatts). Apparently, 1 MW can power approximately 834 homes a
55.
▲
by
sheepscreek
3mo ago
I’m a complete layman to this field, but what the article did say was they’re hopeful that AI/ML can help develop a model that can pull out information such as the scattering caused by RBCs (which is present in the large volume of data
56.
▲
by
sheepscreek
3mo ago
The message I’m getting is that Apple will never compromise on its healthy margins. If something becomes basically unaffordable for their target market, they’d cut the production and even discontinue the product, than take a hit on margins.
57.
▲
by
sheepscreek
3mo ago
Speculating here - “effectively” cooling the CPU and GPU materially using this technique at datacenter scale may have never been done. Those things than run hot, easily crossing 100C. So the loop is doing a lot of work to keep them stable a
58.
▲
by
sheepscreek
3mo ago
I think it’s more than that. Piecing together the perspective of a few commentators in this post - it’s plausible Anthropic is trying to shift the narrative from US vs. Rest of the world to US vs. China. In other words, they want to sell Fa
59.
▲
by
sheepscreek
3mo ago
So I think the takeaway here is, this is a super fast companion model to larger models, that reasons quickly. Perhaps this technique can be used to train a highly optimized reasoning "expert" in MoEs.
60.
▲
by
sheepscreek
3mo ago
The initial motivation for this was likely to thwart any competition. Already Anthropic has accused some companies of organized distillation efforts at a massive scale. Back when I used antigravity, it used to show the reasoning intact - at
More ›