Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
gizmodo59
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
gizmodo59
5d ago
You can pay for whatever Google service and they still collect this data. Same with meta subscription.
2.
▲
by
gizmodo59
6d ago
You don’t need AI to solve this. Just lots of compute.
3.
▲
by
gizmodo59
10d ago
codex (their app) is pretty good and lots of banked resets, luna is cost effective and Astra seems better than Fable for many tasks. Most important one is less refusals and I can use it the way I want without the fear of getting banned. I d
4.
▲
by
gizmodo59
10d ago
why cant models make a tool call to stockfish? its like saying model can't execute python for complex math calculations
5.
▲
by
gizmodo59
11d ago
this is published july 27th. I know HN want's to pile on any negative news these days but this is very old in today's world. Please update the title
6.
▲
by
gizmodo59
16d ago
The fame he got, the stamp of a frontier lab in the resume - its worth more than working in many other companies for years. Plus, not everyone thinks working more is needed once they achieve a certain number.
7.
▲
by
gizmodo59
17d ago
This is a lame argument. Questioning his financial incentive is very legit. Why should we trust him?
8.
▲
by
gizmodo59
17d ago
And Anthropic are the good guys? They are talking insane stuff these days. May be we can trust zuck after all
9.
▲
by
gizmodo59
18d ago
No offense whatsoever. But this whole line "but it won't replace the real experience and taste required for larger projects" is not how I see it. I wish people are more humble at this juncture of AI progress. This is not to s
10.
▲
by
gizmodo59
20d ago
If you haven’t tried computer use with Astra with codex I highly highly recommend it. Just like how gpt 4 and agentic coding with cc. This thing is the most exciting stuff I’ve seen in a while. And then all the blender, cad stuff is cherry
11.
▲
by
gizmodo59
21d ago
I don't agree. The main issue with their scoring/methodology is that the numbers make it seem like 5-6 models have little to no difference when in fact there is a significant difference between fable and opus and sol and astra for
12.
▲
by
gizmodo59
23d ago
99 on arc agi 3 is insane. The arc agi committee were so proud of creating a benchmark they thought will take forever to saturate.
13.
▲
by
gizmodo59
26d ago
Side note.. Fable just rejected this. GLM 5.3 did without questioning me. 5.6 sol did it beautifully.
14.
▲
by
gizmodo59
1mo ago
I somehow find it better to give 2 frontier model companies 100-200/month than dropping 10 grand on a hardware that will get old in no time with bad TPS. I really want to have a fully local model but seems like one more generation wait
15.
▲
by
gizmodo59
1mo ago
I’m referring to Fable vs 5.6 Sol. Opus 5 being bad is universal at this point.
16.
▲
by
gizmodo59
1mo ago
I second the parent comment. 5.6 sol xhigh is not only better than fable I can also run it forever without worrying about limits. The frontend has gotten much better too with the plugins that come with codex.
17.
▲
GPT 5.6 Cyber
(openai.com)
132 points
by
gizmodo59
2mo ago
|
71 comments
18.
▲
by
gizmodo59
2mo ago
I will believe there is no moat when the revenues for Anthropic is not 70B. It seems like people want to throw away money and they don’t like switching
19.
▲
by
gizmodo59
2mo ago
It will be comparable to Luna then.
20.
▲
by
gizmodo59
2mo ago
how does this compare with https://developers.openai.com/api/docs/models/omni-moderatio... As for use cases, obviously we can't fully rely on non-deterministic capability for sensitive things but a small
21.
▲
by
gizmodo59
2mo ago
If your comment is referring to situational awareness, its due to 4x leverage. leverage is always risky. AI/semis are still doing extremely well (over last 2 years) despite the recent dip
22.
▲
by
gizmodo59
2mo ago
Also: https://www.cnet.com/tech/tech-industry/apple-google-others-...
23.
▲
by
gizmodo59
2mo ago
The web and connecting to other services is very important for almost all of my use cases. While I believe we are going to get better and faster models, the web index is certainly not downloadable and maintainable for 99.99% of the folks wh
24.
▲
by
gizmodo59
2mo ago
Moat is not the harness. Harness itself is temporary until the models get better and slowly the code in harness will go down. Note that the biggest GPU providers in the world are the hyper scalers and even they couldn’t allocate more if you
25.
▲
by
gizmodo59
2mo ago
That’s a very narrow view. So non of the fields of science and engineering matter but just cancer?
26.
▲
by
gizmodo59
2mo ago
It’s also very very divided (x companies, oss vs not and other interests)
27.
▲
by
gizmodo59
2mo ago
People still believe that the model intelligence has plateaud and I feel like they live under the rocks. I get the insane capex spend, I get the hype, I even get certain company will go under but to ignore the capability jump in a year is a
28.
▲
by
gizmodo59
2mo ago
Yet another "benchmark to promote their own harness"
29.
▲
by
gizmodo59
2mo ago
When would I use this over the plugin in codex? Which I think can be invoked from cli as well
30.
▲
by
gizmodo59
2mo ago
Because lack of talent and organizational disfunction matters a lot more than you think. The reason why OAI and Ant are always at the top is because of this and I’d say compute is third on the list.
More ›