Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
user43928
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
61.
▲
by
user43928
14d ago
If China for some reason agrees that would probably be enough for the time being. Who else should develop AGI, Mistral? Maybe in a decade.
62.
▲
by
user43928
14d ago
Didn't they walk that one back eventually?
63.
▲
by
user43928
14d ago
Internally Mythos has been available in February. The labs have been holding their best models back for a while it seems like.
64.
▲
by
user43928
14d ago
Yes, and a lot of QA testing.
65.
▲
by
user43928
14d ago
Not the case here. When I was young I dabbled in reverse engineering and hacking games for example. I don't think I am less of a 'hacker' or programmer than the next person on HN.
66.
▲
by
user43928
14d ago
I spent hundreds of hours on this, so there is still some friction. If I told you the exact market and use case, I'd still have a decent head start, but indeed, someone without a job could invest more time than I do. People in a cheape
67.
▲
by
user43928
14d ago
I said I'm hoping to complete it this month and I already shared some details. Should I post my market and keyword research too, here on this site for software developers and entrepreneurs?
68.
▲
by
user43928
14d ago
I'm not going to share exact details about my paid app, but you could think media management, with a lot of extra features, an in-app browser with custom controls, and multiple AI powered features with on device processing.
69.
▲
by
user43928
14d ago
I'd like to think it comes down to my expertise, but I don't believe I did a lot of designing. Sure, I've given a lot of inputs over the time, probably I've steered it to solutions that work. Maybe it was also helpful th
70.
▲
by
user43928
14d ago
I think selling software that solves your own problems is great and beats trying to develop for customers you do not understand.
71.
▲
by
user43928
14d ago
I liked building apps and helpful tools before AI came around. I've been doing so for 15 years, 10 of them professionally. Today, I like it even better. It's never been so fun. What would have taken days of work, if it was feasibl
72.
▲
by
user43928
14d ago
I have a more optimistic outlook on the abilities of our governments as well as the motivations of rich people. I believe we will distribute the fruits of AI's labor at the very least to the extend where everyone can live comfortably.
73.
▲
by
user43928
14d ago
Out of 3.5 billion employed people globally, only 10% are professionals. Do you really think making labor obsolete and giving humans back their time is such a bad thing?
74.
▲
by
user43928
14d ago
In December 2024 o3 scored 87.5% on ARC-AGI-1 and cost $4560 per task. DeepSeek V4 Flash 0731 scores 89% and costs $0.02 per task. If we apply the same factor to the guesstimated API price of $20M for this problem, we arrive at $57. Real co
75.
▲
by
user43928
15d ago
Disagree. I have 200k LOC now plus 100k in tests, and it is still performing like it was four months ago when I started to seriously use AI. If anything, it works more reliably today with the smarter models.
76.
▲
by
user43928
15d ago
Have you worked with Opus 5? Its documentation about what the code does not do could fill whole books. UI copy being full of slop explaining what the software does not do is another problem. I am not convinced that a lack of negative test c
77.
▲
by
user43928
15d ago
That seems like comparing the internet to jQuery.
78.
▲
by
user43928
15d ago
Bubeck made no such demands, particularly not for removing Levent from authorship of his own work. What they discussed was that Bubeck felt it would be inappropriate for an Anthropic employee to author the proposed rewrite of OpenAI's
79.
▲
by
user43928
15d ago
Can you specify what leverage you think Brubeck has over an independent professor's career in order to make threats? I see none, and consequently Brubeck's explanation makes more sense to me. I understand he meant these words, whi
80.
▲
by
user43928
15d ago
The comment you replied to quoted "no user inputs after July 3rd" with no restriction to Buckmaster or Codex. Obviously the result of OpenAI's investigation was that no usage data has interacted with the system after that dat
81.
▲
by
user43928
15d ago
The browser, iOS, and Android all use a main thread separate from the thread responsible for scrolling animations. However, I think it's true that controlling threading in order to perform gnarly work in a separate thread is more ergon
82.
▲
by
user43928
15d ago
Storing private secrets in your public client is easy to avoid for anyone halfway competent. We are all professionals here. Turn on the secrets scan in GitLab, and put in your release checklist to have the AI audit the usage of secrets in y
83.
▲
by
user43928
15d ago
This is not as obvious as many naively believe. It depends on how hard to maintain the code added is, how likely it needs to change in the future, and most importantly on the cost. If reviewing and manually improving the code takes hours, t
84.
▲
by
user43928
15d ago
Yes. The key point being that this concerns a new paper about OpenAI's result rather than the paper Buckmaster and Alpöge were working on.
85.
▲
by
user43928
15d ago
They specifically say in the post how React Native was the correct decision and that it worked well for them. Now it's a different situation as implementation has become incredibly cheap.
86.
▲
by
user43928
15d ago
The data wasn't used, it just does not line up with the time frame. And for the usage data they do use, when you leave the relevant 'Help improve the model' toggle on, everyone working at the labs says it isn't used in t
87.
▲
by
user43928
15d ago
That's also incorrect. My understanding is that they asked the independent researcher to improve OpenAI's AI generated proof and be the lead author of the paper to publish OpenAI's result. This is the paper where they did not
88.
▲
by
user43928
15d ago
How about the fact that it almost certainly did not happen? I read today that OpenAI after investigation was able to categorically rule out that usage data from before beginning of July could have affected the system that was used.
89.
▲
by
user43928
16d ago
I could buy four of them for ~20k. That's like four years of ChatGPT + Claude subscription. Eight years if only ChatGPT, or sixteen years of the Pro 5x subscription.
90.
▲
by
user43928
16d ago
That's 1/3 of GPT 5.6 Luna. It seems rather close to me. But great that we have a new leader in performance/price in that segment.
More ›