Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
pants2
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
91.
▲
by
pants2
4mo ago
You're comparing the highest tier Claude subscription to something Qwen3.5-122B-A10B running locally, apples to oranges. If you compare to a smarter US model like Grok 4.3, $1400 will pay for 560M output tokens, which at ~25 t/s l
92.
▲
by
pants2
4mo ago
You'll need to spend at least $20K on a workstation that can run DS4 Flash. It would take ages to reach that much in token spend at the speeds it runs at, and if you factor electricity costs you will likely never break even vs using AP
93.
▲
by
pants2
4mo ago
Not similar. DeepInfra[1] has DS4 Pro pricing at $1.30/$2.60 which is 3X the Deepseek[2] (Chinese) hosting at $0.435/$0.87. DeepInfra is also very slow at 37 t/s and uses an FP4 quant[3], so intelligence will be degraded slig
94.
▲
by
pants2
4mo ago
Yes, because of the limits on tok/s, and you have to compare apples to apples, not Gemma 27B to Opus 4.7.
95.
▲
by
pants2
4mo ago
Running them locally is cool and has privacy/autonomy benefits, but you can't really make a value case for it. Guaranteed if you run the math you will never run enough inference to pay off your hardware vs buying tokens. Last time
96.
▲
by
pants2
4mo ago
Yes, I think this has become their competitive edge to stay relevant and retain customers. If a lab falls behind the frontier for too long, they will lose customers to other models. Google, DeepSeek, and XAI have all released frontier model
97.
▲
by
pants2
4mo ago
Polymarket says not likely until the end of June. Maybe some money to be made? https://polymarket.com/event/gpt-5pt6-released-by
98.
▲
by
pants2
4mo ago
The Chinese models are only cheap on subsidized Chinese hosting. I have yet to find a USA-hosted Chinese model with a very clear value advantage over US models.
99.
▲
by
pants2
4mo ago
I would love if Apple enforced that rule, but they certainly don't
100.
▲
by
pants2
4mo ago
I definitely run all my emails through an LLM filter and wish I could do the same for push notifications!
101.
▲
by
pants2
4mo ago
The biggest problem are apps that do both. For example, I want Uber to notify me when my driver has arrived, but I don't want it to notify me when they have a special 10% discount on my next 5 rides. It's not straightforward to bl
102.
▲
by
pants2
5mo ago
I have been a Kagi subscriber for years but I do increasingly find myself using Google. It's not good at local searches or news searches. It's also not good at showing quick context like Google's knowledge graph, especially w
103.
▲
by
pants2
5mo ago
haha you got me there, can hardly even consider those as roads!
104.
▲
by
pants2
5mo ago
They have done a lot of testing in Pittsburgh which has some of the craziest roads and intersections anywhere, so I'd assume yes
105.
▲
by
pants2
5mo ago
That would take ages!
106.
▲
by
pants2
5mo ago
In a just world all companies would be taxed on their overall impact and not just revenue. Coca Cola would be taxed for their contribution to obesity and plastic waste. Exxon would be taxed for their emissions. Meta would be taxed for its h
107.
▲
by
pants2
5mo ago
Yeah the Google AI results are more dangerous than ChatGPT, not only because it uses a smaller model but because Google's knowledge graph used to deliver very accurate and authoritative information but now that's been replaced by
108.
▲
by
pants2
5mo ago
> Last month, Mayor Carmella Mantello, flanked by officers in blue, accused the city council of “defunding” the police and declared a state of emergency to keep the cameras running, a designation usually reserved for floods and blizzards
109.
▲
by
pants2
5mo ago
I used an M1 Air for developing iPhone Apps for years and it worked great. Not "wow" fast but I never had a complaint about it.
110.
▲
by
pants2
5mo ago
Criminals posting evidence of their crimes online certainly makes the prosecutor's job easy. I wonder what's going through his head though. Did he just not expect to get charged? Or just wanted to live life to it's fullest kn
111.
▲
by
pants2
5mo ago
You might be surprised how good cloud gaming has gotten. I play AAA games at max settings on my MacBook Pro through GeForce Now, and with fiber internet it's nearly indistinguishable from native.
112.
▲
by
pants2
5mo ago
I think the real crux of the moat is model intelligence. I'd bet that most of the money being spent on inference is on the top few models (today Opus-4.7 and GPT-5.5) from people and companies that benefit from using the best models. T
113.
▲
by
pants2
5mo ago
- Right, R1 affected markets because the market originally believed your theory, but it doesn't any more, which is why V4 didn't move markets at all. - Sure, you can use US infra providers. Together.ai is a good US provider but th
114.
▲
by
pants2
5mo ago
DeepSeek r1 affected markets because for a little while people bought this, but it's not true for so many reasons. Sending data to China is out of the question for every American Enterprise. OAI and Anthropic have rich product suites a
115.
▲
by
pants2
5mo ago
It's incredible how far behind Gemini has gotten, both the product and the model. Even the ChatGPT plugin for Google Sheets blows away the native Gemini integration. Everyone thought Google was pulling ahead with Gemini 3. For a minute
116.
▲
by
pants2
5mo ago
When will countries start treating cyberattacks as an act of war? If the North Korean military came to America and robbed fort Knox of $200M in gold there would be retribution. But hack an American company for the same amount and the feds d
117.
▲
by
pants2
5mo ago
Folding my laundry
118.
▲
by
pants2
5mo ago
I've never in my life seen a useful product tour. They're always blatantly obvious like "THIS IS THE SEARCH BAR. USE IT TO FIND CONTENT ACROSS OUR PRODUCTS AND SERVICES." The best UX is using obvious and standard design,
119.
▲
by
pants2
5mo ago
Good point. I feel like this does a disservice to ChatGPT -- IIRC even the free tier of Claude points you to Sonnet 4.6 by default, which is magnitudes better than 5.3-instant which has been the default in ChatGPT. Hence most users will imm
120.
▲
by
pants2
5mo ago
Good question. I find that GPT-5.5 thinking is very good at not thinking for simple questions, so much so that I've never had the need to use the instant model even for quick Q&A. I'm assuming the instant model, then, is an
More ›