Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
beering
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
7 ms
·
151.
▲
by
beering
6mo ago
people have been complaining about this since GPT-4 and have never been able to provide any evidence (even though they have all their old conversations in their chat history). I think it’s simply new model shininess turning into raised expe
152.
▲
by
beering
6mo ago
Word on the street is that Opus is much much larger of a model than GPT-5.4 and that’s why the rate limits on Codex are so much more generous. But I guess you could also just switch to Sonnet or Haiku in Claude Code?
153.
▲
by
beering
6mo ago
Claude Code is not a platform and you’re not meant to be building on it. Netflix is also not a platform and you shouldn’t be running code (open source or not) to mass download Netflix movies either.
154.
▲
by
beering
6mo ago
They have so much mindshare right now that they can’t lose, and the number of users that use opencode and would be affected is miniscule—-on the level of complaining about your online bank not supporting Konqueror.
155.
▲
by
beering
6mo ago
Mainly OpenSCAD is not a BRep modeling tool! It is not on the same level of power as CAD tools with a BRep kernel and this especially shows when you want to do a fillet over an arbitrary edge. Unfortunately these kernels are hard to make an
156.
▲
by
beering
6mo ago
They already see everything I’m doing because I send my prompts to them. What “workaround” are you referring to?
157.
▲
by
beering
6mo ago
So are you able to get free inference now that you decrypted this?
158.
▲
by
beering
7mo ago
This is 100% crank “science” that has wrapped up a banal finding in big words and LaTeX. The claim is roughly as exciting as, “some computer programs print nothing to stdout.” The output shown is not “null” or “void”. It is the empty string
159.
▲
by
beering
8mo ago
Doesn’t basic airplane autopilot just maintain flight level, speed, and heading? What are some other things it can do?
160.
▲
by
beering
8mo ago
Simultaneously, if you hire human translators, you are likely to get machine translations. Maybe not often or overtly, but the translation industry has not been healthy for a while.
161.
▲
by
beering
8mo ago
Is “Start Me Up” the song that goes, “you make a grown man cry”?
162.
▲
by
beering
9mo ago
I think Google has already shown that in the long run, people accept ads and prefer them to paying a subscription fee. If that weren’t true, then YouTube Premium would have double-digit % of youtube users and Kagi Search would be huge.
163.
▲
by
beering
10mo ago
That may be true, but you can’t compare average GenAI with the best humans because there are many reasons the human output is low quality: budget, timelines, oversights, not having the best artists, etc. Very few games use the best human ar
164.
▲
by
beering
10mo ago
Model capability improvements are very uneven. Changes between one model and the next tend to benefit certain areas substantially without moving the needle on others. You see this across all frontier labs’ model releases. Also the version n
165.
▲
by
beering
1y ago
Note that this is not relevant for reasoning models, since they will think about the problem in whatever order it wants to before outputting the answer. Since it can “refer” back to its thinking when outputting the final answer, the output
166.
▲
by
beering
1y ago
This article spent a lot of words to say very little. Specifically, it doesn’t really say why working towards AGI doesn’t bring advancements to “practical” applications and why the gazillion AI startups out there won’t either. Instead, we n
167.
▲
by
beering
1y ago
That’s an interesting point. It’s not hard to imagine that LLMs are much more intelligent in areas where humans hit architectural limitations. Processing tokens seems to be a struggle for humans (look at how few animals do it overall, too),
168.
▲
by
beering
1y ago
The past few years I’ve been hearing crazy stories of workarounds and scripts to deal with all these new features in Windows. Isn’t that what was preventing people from using Linux? Replacing utilman.exe with cmd.exe is not something a norm
169.
▲
by
beering
1y ago
That is what happened in the 18,000 water cups video. It was presented as a way to avoid the ai and get a human on the other end.
170.
▲
by
beering
1y ago
> Does "career development" just mean "more money"? Big companies means more opportunities to lead bugger project. At a big company, it’s not uncommon to in-house what would’ve been an entire startup’s product. And de
171.
▲
by
beering
1y ago
I think text-davinci-001 is GPT-3 and original ChatGPT was GPT-3.5 which was left out.
172.
▲
by
beering
1y ago
GPT-4 is very different from the latest GPT-4o in tone. Users are not asking for the direct no-fluff GPT-4. They want the GPT-4o that praises you for being brilliant, then claims it will be “brutally honest” before stating some mundane take
173.
▲
by
beering
1y ago
It feels crazy to keep arguing about LLMs being able to do this or that, but not mention the specific model? The post author only mentions the IMO gold-medal model. And your post could be about anything. Am I to believe that the two of you
174.
▲
by
beering
1y ago
The 5 seconds delay is probably due to reasoning. Maybe try setting it to minimal? If your use case isn’t complex maybe reasoning is overkill and gpt-4.1 would suffice.
175.
▲
by
beering
1y ago
Because people will switch. It’s trivial to go to old conversations in your history and try those prompts again and see if chatgpt used to be smarter.
176.
▲
by
beering
1y ago
I think you got some different things mixed up. the deprecation is for chatgpt. (but i think Pro users can still use the old models)
177.
▲
by
beering
1y ago
Curious to know how the different models compare for you for doing math. Heard o4-mini is really good at math but haven’t tried o3-pro much.
178.
▲
by
beering
1y ago
Do you think that is absurd because OpenAI is overvalued? Or because Stripe is overvalued? Or one of them is undervalued?
179.
▲
by
beering
1y ago
It’s exciting because nearly all humans have 0% chance of throwing the rock into the bucket, and most people believed a rock-into-bucket-thrower machine is impossible. So even an inefficient rock-into-bucket-thrower is impressive. But the b
180.
▲
by
beering
1y ago
What do you mean by “pure language model”? The reasoning step is still just the LLM spitting out tokens and this was confirmed by Deepseek replicating the o models. There’s not also a proof verifier or something similar running alongside it
More ›