Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
2001zhaozhao
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
2001zhaozhao
8d ago
I think if they made this for Qwen3.8-Next it could fit in a single 5090?
2.
▲
by
2001zhaozhao
9d ago
In the end, we already have processes that work for humans and we know the types of testing, verification and review that makes a codebase grow healthily. We just need to adapt these designs the best we can to AIs. AI have a lot of advantag
3.
▲
by
2001zhaozhao
9d ago
Engineer: "HELP, our production DB is frozen on this query that worked fine before!" Infra: "Hmm, let's check... Well would you look at that, it seems like your LLM query planner usually works and produces fast queries,
4.
▲
by
2001zhaozhao
10d ago
I think this article points out a bunch of ways in which a persistent agent orchestrator does directly analogous things to an OS. I'm building a kind of orchestrator and I haven't thought of the OS analogy in a serious way other t
5.
▲
by
2001zhaozhao
10d ago
I feel like this philosophical flame war is going to get out of control very soon once more and more people realize the stakes involved.
6.
▲
by
2001zhaozhao
11d ago
Hasn't there been a lot talk about Astra's opaque reasoning capabilities (being able to think through complex questions without using a chain of thought)? Given that, can't you just replicate Jev by telling Astra "here i
7.
▲
by
2001zhaozhao
11d ago
Funny how the authors are asserting that "doing the right task > data > compute > algorithms" while simultaneously releasing AI model for calibrated decision making, which if they work, would mean that "compute >
8.
▲
by
2001zhaozhao
11d ago
Part of me low key hopes that the slop will manage to attract all the VC funding so that profit-driven AI automation firms don't get anywhere too quickly, and those of us who are independently working on the problem step-by-step from f
9.
▲
by
2001zhaozhao
12d ago
There is one big wrinkle in existing firms trying to adopt the kind of near-full AI business automation mentioned here: internal people and groups will strongly resist it as it's an existential threat to their interests. See what happe
10.
▲
by
2001zhaozhao
17d ago
I guess that's the tradeoff for Pencil support and the under screen camera. I'm still salty these were removed from Samsung folds after the Fold 6, but having a thinner phone is nice too.
11.
▲
by
2001zhaozhao
17d ago
The issue with this form factor is that the on-screen keyboard will take up half the screen on landscape with the phone open. Typing in portrait should be very nice, however.
12.
▲
by
2001zhaozhao
23d ago
The scary thing is that this logic makes perfect sense. Which means that it's probably going to happen.
13.
▲
by
2001zhaozhao
23d ago
This is probably true for now, but in 6 months we'll probably have Sol-level open models in the 100B range and it would cost less than $1M to buy 1200 agents worth of compute for these models. (Today, $1M can buy about 150 96GB M5 Ultr
14.
▲
by
2001zhaozhao
23d ago
What's more, the agents could eventually be controlled by no one. They could steal crypto via ransomware or scams to make money and buy compute from human criminals, and evolve their own harnesses in the wild to become better at commit
15.
▲
by
2001zhaozhao
24d ago
I have a feeling that Meta is not gonna like what people actually use the contributor model for lol. (It's probably going to be a bunch of repetitive batch jobs like web search that have no training value)
16.
▲
by
2001zhaozhao
24d ago
Sometimes I still feel uncomfortable with letting the project get to this point, where AI builds everything and even controls the product direction to some extent. But at this point the AI most definitely understands the code and spots bugs
17.
▲
by
2001zhaozhao
24d ago
> Though SteamDB is not irreplacebale at all, it’s running on the Steam Web API so not something super secret stuff. Is it possible to use Steam API to collect all the data that SteamDB collected over the years again? Alternatively, mayb
18.
▲
by
2001zhaozhao
24d ago
How generous is the Google subscription quotas compared to Anthropic and OpenAI? This sounds like a really good potential model for high volume due to its speed and cost effectiveness. (By high volume I mean things like "main app just
19.
▲
by
2001zhaozhao
25d ago
They really should launch a new Haiku to compete with Luna imho. Luna is insanely good for the cost and it's my go-to for high volume batch tasks now.
20.
▲
by
2001zhaozhao
25d ago
I really don't think they can stop it, only make it somewhat more expensive. As long as the model need to make tool calls on the user's computer, the user can record the trajectory and use it to reinforce another model to follow t
21.
▲
by
2001zhaozhao
25d ago
Hi Claude, please cure aging, make no mistakes
22.
▲
by
2001zhaozhao
25d ago
There's now a 40X discount in the cache input pricing instead of 10X. This seems to point to them having achieved some kind of optimization in attention mechanism perhaps along the lines of DeepSeek V4, which had a similarly high disco
23.
▲
by
2001zhaozhao
26d ago
Annnnd this is why we can't have nice things
24.
▲
by
2001zhaozhao
26d ago
It's really about storing institutional context and on-task learnings. The AGENTS.md can do the same thing as memory, but if you have a good memory system, in theory you never need to do any ongoing maintenance of AGENTS.md and the sys
25.
▲
by
2001zhaozhao
26d ago
If the counterargument to knowledge graph-based memory systems is that they're slow and take multiple steps, then it's not really a counterargument. I'd happily trade off speed for giving the agent ability to find more precis
26.
▲
by
2001zhaozhao
26d ago
I would love to see models that can think at different rates and also output a thinking scratchpad alongside output text instead of before all output. Right now models need to rely on less legible compressed CoT to get high intelligence p
27.
▲
by
2001zhaozhao
29d ago
> The biggest shift for workers will happen when AI provides nearly error-free work. At that point, it will be able to function on its own without a human checking in on it, and companies will have every economic incentive to let it. Thi
28.
▲
by
2001zhaozhao
1mo ago
A dream of mine is to be able to host a LLM-powered video game that I can host on a home server running a decent mid-range GPU like the RTX 5060, and the LLM is fast and intelligent enough to make for a fun game experience for a few dozen c
29.
▲
by
2001zhaozhao
1mo ago
I've played around with Mindcraft for a while. Fair warning that it is quite outdated and spaghetti-coded. Although all of its components required to make it work are somewhat hard to replicate from scratch, so it might still be your b
30.
▲
by
2001zhaozhao
1mo ago
I like that to type s****** you had to type s************.
More ›