Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
redox99
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
redox99
5d ago
Renting datacenters is their mission, now on earth and later in space (assuming they deliver).
2.
▲
by
redox99
5d ago
Not surprising considering Grok 4.7 is a 2T model, so Sol/Opus class, not Astra/Fable class.
3.
▲
by
redox99
5d ago
That's outdated and doesn't fully include the multiple billion per month contracts. Anthropic: 1.25B/month Google: 0.92B/month Unnamed customer starting in december: 1.1B/month Starlink monthly revenue is ~1.5B/
4.
▲
by
redox99
5d ago
xAI is the most profitable part of SpaceX by far.
5.
▲
by
redox99
5d ago
A dense 27B doesn't really make sense for the Mac. A MoE makes way more sense when you have modest bandwidth but lots of memory.
6.
▲
by
redox99
5d ago
I mean tokens for internal use.
7.
▲
by
redox99
5d ago
Yeah it's a bit ambiguous, but in any case I'd think simply using that idle capacity to generate tokens would be more valuable than finding a basically meaningless number.
8.
▲
by
redox99
5d ago
You're underestimating how good models are. I'm pretty sure you could in fact prompt that and have it work.
9.
▲
by
redox99
5d ago
> To what end? The biggest benefit of tinkering with 40-year-old hardware is the experience of doing so. The biggest benefit to you. Other people, most people in fact, just like the game and want to enjoy it in new ways. They don't
10.
▲
by
redox99
5d ago
> the whole point of the endeavor is the process of getting that knowledge and becoming an expert in the system There's nothing stopping you from making an emulator without using AI or whichever way you think is best for your own le
11.
▲
by
redox99
6d ago
Quite bearish on Anthropic if they had nothing better to do with 2048 GPUs for 10 days than finding an RSA number with already existing algorithms.
12.
▲
by
redox99
7d ago
It's pretty sad when you realize the average person is just kind of... unaware. They are fine with whatever, they don't really notice most stuff that goes around them.
13.
▲
by
redox99
8d ago
It's a monorepo and they're at over 1 million PRs. There's surely some juicy stuff there.
14.
▲
by
redox99
8d ago
It's crazy that we still rely on these unsafe C dependencies, in an era where migrating code to Rust (or other languages) is so easy. There's really no excuse.
15.
▲
by
redox99
8d ago
I didn't claim you wouldn't discover things in pure math that would decades later turn out to be useful in other fields. I'm sure you would. My claim was that the ROI is lower compared to applied math, which matters when fund
16.
▲
by
redox99
8d ago
I think most of the long hanging fruit has been discovered, and even then you'll have better ROI focusing funding on applied math instead of pure math.
17.
▲
by
redox99
8d ago
You can trivially run 131k on 24GB 4bit, and there are repos with tweaks that allow you to get the full 262k but idk if there's degradation with their approach.
18.
▲
by
redox99
8d ago
I tried their WebGPU version and it immediately started looping. Yeah "near lossless" my ass. Plus the reasoning that it looped on was clearly wrong and unlike the non quantized 27B
19.
▲
by
redox99
10d ago
Not everyone uses laptops. I can run Qwen 3.8 27B (which is a REALLY capable model) in the background coding for me while I'm simultaneously browsing the web and playing VALORANT without any performance impact, and that's on a 6 y
20.
▲
by
redox99
10d ago
> But I also think that the state of the art in small LLM and user device capabilities aren't there yet to put a "good enough to be actually useful" local-only LLM as a prepackaged thing in a mass market distributed browse
21.
▲
by
redox99
11d ago
Targeting byte code or asm instead of high level would be silly for everyday tasks. You blow up the number of tokens, reduce your effective context, and there's just more places for it to make a mistake, which most likely won't be
22.
▲
by
redox99
11d ago
I'm so happy for the two used 3090s I bought for $500 each after Ethereum mining ended. I even saw them for like $430 at some point lol.
23.
▲
by
redox99
11d ago
The math is wrong, the tok/s is at least 2x that, at least with MTP and Q8 KV which you should always use. And the default tokens a day is ridiculously low at least for coding. Having said that, it will never pay for itself. A simpler
24.
▲
by
redox99
13d ago
Terminal bench 4 is good largely because it's recent so it hasn't been benchmaxxed yet. It's more of a sysadmin/devops benchmark than a coding benchmark though, but still a decent proxy. https://artificialanal
25.
▲
by
redox99
14d ago
They may allocate different number of resources every year based on market conditions but they'll never give up on gaming, that would be extremely silly.
26.
▲
by
redox99
14d ago
I think in a few decades when there are more humanoid robots than humans we'll likely have skynet, so I'm pessimistic in a way lol. But in the near future and on a personal level I think we'll need to adapt and pivot but we&#
27.
▲
by
redox99
14d ago
I think it will be like how a lot of people know how to code in python but have zero understanding of assembly or how a cpu works. As a researcher you'll accept there's this low level stuff that if you want you can dig into but i
28.
▲
by
redox99
14d ago
Of course you're not going to get rich with the kind of software that LLMs can one shot these days. But that kind of software like to-do lists or basic CRUD have been saturated for over a decade, way before LLMs. People overestimate ho
29.
▲
by
redox99
14d ago
This has always happened, way before AI. You'd spend months or years building and growing your business, and then Google would release a feature or product that would kill your business overnight because they can throw way more money a
30.
▲
by
redox99
14d ago
> Many graduate students (I know) are having a crisis if any of their research worth it? If AI can (or will) do everything, what's the point of doing experiments and all? This will eventually deter a whole generation of curious mind
More ›