Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mchusma
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
61.
▲
by
mchusma
3mo ago
Maybe, but they said they have “started” the Gemini 4 pretrain. So not having done any significant pretrain in a year or so seems odd to me.
62.
▲
by
mchusma
3mo ago
3.6 Flash would be a great model at 3.0 flash pricing. At this pricing, its thoroughly trounced by about 10 models on cost/performance including Grok 4.5. 3.5 Flash-ite would be a great model at 2.5 flash-lite pricing, as is, its troun
63.
▲
by
mchusma
3mo ago
The issue is not investment at all, we want more investment into housing. But you are absolutely correct you cannot both want housing to be a good investment (increase in value faster than inflation) and an affordable (drop in price lower t
64.
▲
by
mchusma
3mo ago
I am generally for a light government touch, but I think excessive light and sound are forms of pollution and should be generally regulated in a not-overly burdensome way (give people fix it tickets etc). Policing excessive light, and exces
65.
▲
by
mchusma
3mo ago
I counted the other day and there were at least 12 different providers with "better than Opus 4.5 performance" on Artificial Analysis, Opus 4.5 being Anthropic's December release that many say kicked off the latest accelerati
66.
▲
by
mchusma
3mo ago
I’m not categorically against bans of this kind but I’m pretty sure everyone knows these are fake. I am worried about freedom of speech slippery slope issues, and also San Francisco (a relatively small eccentric city) trying to dictate US p
67.
▲
by
mchusma
3mo ago
I actually think a reasonable strategy of Google is to focus on biotech, where the demand for compute is infinite, and focus on cost effective solutions with really good search for other verticals (eg legal, accounting, etc), and offer loss
68.
▲
by
mchusma
3mo ago
This paints everyone as right or left, which I don’t find accurate or helpful. I think other labels even if they meant the same category, to be more helpful.
69.
▲
by
mchusma
3mo ago
I don’t know if it’s a great business model but it makes perfect sense to me. Open models when fine tuned are capable at better than frontier performance at a fraction of the price for many (probably most) domain specific tasks. If companie
70.
▲
by
mchusma
3mo ago
Non paywalled alternative: https://www.theverge.com/science/965849/spotify-founder-ek-s... This looks similar in concept to the Midjourney health, but much more about multiple modalities (different surface scanner
71.
▲
Neko Health Raises $700M
(nytimes.com)
3 points
by
mchusma
3mo ago
|
1 comments
72.
▲
by
mchusma
3mo ago
I do think this would be interesting if they made these easy to finetune, as I do think this level of intelligence is likely sufficient for many applications and could be extremely cheap to run.
73.
▲
by
mchusma
3mo ago
Incentivizing usage during peak times makes total sense, but if price swings are this wild, how are grid scale batteries not highly economical? My rough ballpark math was that you need roughly 20 kilowatts of battery storage to make this is
74.
▲
by
mchusma
3mo ago
I would say that overall there are pros and cons to this, I really want to be allowed to use agents on my behalf, and don't want to see sites prevent me from doing this. On the other hand, I do recognize there are cases when its good&#
75.
▲
by
mchusma
3mo ago
I stayed in downtown LA recently and looked like the set from the walking dead. Literally blocks of people wandering in traffic. I guess you could argue you definitely don't need flock cameras to see the problem, but also I don't
76.
▲
by
mchusma
3mo ago
I agree with parent, the full quote is: "The whole thing relies on donations to keep it afloat, which is really what tax dollars are for." I think this is a great site, love what they are doing, and support them (including a liter
77.
▲
by
mchusma
3mo ago
I like the linked Scott Alexander post, but I also genuinely wonder what is the rate of change on these tests? The linked test Prenuvo has competition from Ezra + Function and others. It this drops from $2k to $500 over time, it makes it lo
78.
▲
by
mchusma
3mo ago
I was similarly confused. Saying a MRI is the equivalent of stopping smoking for 1 year earlier, or driving a motorcycle 10,000km less seems actually really good! Go MRIs! As another point, most of the negative costs of getting full body sc
79.
▲
by
mchusma
3mo ago
This was my favorite as well.
80.
▲
by
mchusma
3mo ago
I will plug Willow for mac recording. IMO it's basically to me a "better than perfect transcription" as it cleans things up and is almost instant. I liked Superwhisper but switched to Willow as it was a big difference. Its so
81.
▲
by
mchusma
3mo ago
I think the content here is not controversial EXCEPT that it sounds way too doomerish. The only mention of positive effects of AI is "It could bring...major gains in living standards." It really should say "In the US, social
82.
▲
by
mchusma
3mo ago
After seeing studies like this, and how the shingles vaccine reduces dimensia, I have become increasingly convinced that it’s bad to get almost any disease, even transiently. I used to think that it was kind of good to train your immune sys
83.
▲
by
mchusma
3mo ago
This is great. Federally subsidized loans is directly (not solely) responsible for rapid inflation of college costs in the US. Anything to limit its use is a good thing. I’d argue that this test would be better expanded to actually having a
84.
▲
by
mchusma
3mo ago
I agree with parent, Meta has been at this a long time and its only because they have recently fallen off that they pushed this "oh give us credit its really a new org" thing. Basically, if you can't actually "win"
85.
▲
by
mchusma
3mo ago
This not being available on Openrouter really makes it hard to test. I was going to compare vs Grok 4.5 and GPT-5.6 Luna, but I don't want to deal with signing up for Meta for it unless it checks out. Please Meta make this available.
86.
▲
by
mchusma
3mo ago
Looks like a great set of models, but there are about 20 different thinking/model levels here in this family and they are very complex to pick the right one for the task E.g. for GeneBench Pro, it looks like you would always use GPT-5.
87.
▲
by
mchusma
3mo ago
Yeah, this is most directly comparable to xAI Grok 4.5. In both cases, directionally "opus level intelligence for haiku prices" which is a really big deal for application developers who want to include models like this in their ap
88.
▲
by
mchusma
3mo ago
Great model, very nice. Opus class performance at Haiku level pricing (or cheaper with the token efficiency). This seems like a GLM-5.2 killer and this is what Sonnet 5 should have been. This is a model I could really see used inside applic
89.
▲
by
mchusma
3mo ago
I agree. Gemini actually is pretty good for isolated components too. But fable is much better at design than opus or gpt5.5. I have not seen as much difference elsewhere, but definitely design fable is great.
90.
▲
by
mchusma
3mo ago
I think scaffolds and the app layer are really the two big things needed for the deployment of AI in most use cases. In general, my company says for a given problem, we prefer deterministic software as the solution first, followed by LLMs,
More ›