Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
enraged_camel
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
31.
▲
by
enraged_camel
18d ago
Because OpenAI says so, obviously!
32.
▲
by
enraged_camel
18d ago
>> I'm not sure how any of this provides evidence that OpenAI took any of their work. Sorry, but the burden of proof lies in the other direction: OpenAI needs to definitively prove that their agents did not look at the existing
33.
▲
by
enraged_camel
19d ago
No, not really. Yesterday I asked Opus if it can read the logs from the sessions I have on the local ChatGPT app. It looked around and said no, that’s not possible. I said “what about these jsonl files in this folder?” It read them and said
34.
▲
by
enraged_camel
20d ago
I’ve used Astra for the past day and a half. My layperson’s review is that it is impressive at computer use and 3D reasoning, and fails in similar ways to 5.6 Sol at similar rates when it comes to coding. I have no idea how it scored so hig
35.
▲
by
enraged_camel
20d ago
You are falling for selection bias. For every person using Astra to create a game from some random idea and sharing the impressive result, there are an unknown number who have tried the same thing and gave up in frustration.
36.
▲
by
enraged_camel
20d ago
There is something deeply wrong with Astra. I can’t quite put my finger on it. On the one hand it is a lot more knowledgeable, which makes sense since it’s a larger model. On the other hand that knowledge doesn’t reliably translate to intel
37.
▲
by
enraged_camel
21d ago
Yeah their support is non-existent. It's actually mind-blowing that so many people use it.
38.
▲
OpenAI boosts Astra's eval metrics, and continues to change others
(fortune.com)
5 points
by
enraged_camel
21d ago
|
0 comments
39.
▲
by
enraged_camel
22d ago
This whole thing is an absolute disaster honestly, and yes it is being downplayed and hand-waved away. Since March, so many people have mocked Anthropic for their approach to Mythos release, claimed it was all marketing, accused them of hol
40.
▲
by
enraged_camel
22d ago
>>> The demos of Fable/GPT-6 are impressive, but "real AGI" should act more like a collaborator than either a peon or overachiever. I don't really agree. The thing that makes Fable feel like an actual collaborat
41.
▲
by
enraged_camel
22d ago
>> I'm genuinely so confused when people say this with a straight face. Are you talking about coding? Desktop use? Prose? Or something else? Same. It makes me wonder what types of things the person must be working on.
42.
▲
by
enraged_camel
22d ago
I can't speak for others but I have a feeling you're in the very small minority with this take. You could say Sol is faster and cheaper and that's true. Outperforms Fable? Impossible to believe without hard evidence.
43.
▲
by
enraged_camel
23d ago
They are desperately, desperately trying to make a name for themselves as the lab that first created AGI, because Anthropic's IPO is just around the corner.
44.
▲
by
enraged_camel
23d ago
They used a custom harness. It's not a one-to-one comparison.
45.
▲
by
enraged_camel
23d ago
Yep. Incredibly misleading. Although it is not surprising at this point. They are desperate and will do anything to undermine Anthropic's upcoming IPO.
46.
▲
by
enraged_camel
23d ago
Exact same thing happed to me. I gave it a small/medium-sized ticket, walked away, came back to a 25,000 LoC monstrosity that both Fable and another 5.6 Sol agent said is 98% useless and should be thrown away.
47.
▲
by
enraged_camel
23d ago
Paywalled.
48.
▲
by
enraged_camel
23d ago
>> The way they keep spinning off these niche offerings tells you one of 2 things, it's either working pretty well, or it is failing very, very badly. You posited two extremes, with no evidence for either, and with no other possi
49.
▲
by
enraged_camel
23d ago
>> Perhaps switching costs are greater than some would believe. I haven't switched because there's nothing to switch to that is anywhere as good. I've been making dedicated attempts at using Sol but it falls short, desp
50.
▲
by
enraged_camel
23d ago
There are so many deep flaws with this article that it is difficult to believe it was written by a seasoned economics professor. I can only conclude that he's pushing some sort of agenda or is otherwise politically motivated. 1. The co
51.
▲
by
enraged_camel
23d ago
I disagree. As the saying goes: never attribute to malice what can be adequately explained by incompetence. That goes for both OpenAI and HF (but mostly the former, as the latter was the victim).
52.
▲
by
enraged_camel
24d ago
You are either seriously naive, or have never worked on any large and complex legacy codebase.
53.
▲
by
enraged_camel
24d ago
Yes they are. The OP doesn't appear to know what they are talking about. Fable can absolutely be used to develop applications. It's just that for security stuff I use Opus 5. Which is fine for most use cases.
54.
▲
by
enraged_camel
24d ago
The thing that blows my mind about it is that they straight up copied Claude Code and tried to frame it as some sort of incredible accomplishment. There are no new ideas, no insightful UI/UX paradigms, no killer features. They say mone
55.
▲
by
enraged_camel
24d ago
From the article: "We plan to make Astra available soon, but access to its most advanced cybersecurity capabilities will be more limited. Advanced cybersecurity work will initially be available to a group of testers, with access throug
56.
▲
by
enraged_camel
24d ago
What predictions have Ed made that were correct? And how do they weigh against the incorrect ones, in terms of both quantity and quality? That's relevant because if you make a thousand predictions, a few of them might turn out to be tr
57.
▲
by
enraged_camel
24d ago
If you find derision and extreme cynicism "soothing", the problem may be with you, rather than those who are optimistic about AI.
58.
▲
by
enraged_camel
24d ago
>> I think it's fair to call his specific predictions "early." I'm saying his timelines are too short. That means his predictions were either wrong, or meaningless.
59.
▲
by
enraged_camel
24d ago
In this case, it is. Zitron does not possess even the most basic knowledge about the issues and practices he complains about. For example: https://www.crisesnotes.com/sigh-no-ed-zitron-ai-bond-issuan...
60.
▲
by
enraged_camel
25d ago
In a way, this is the only benchmark I care about now. :)
More ›