Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
kgeist
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
kgeist
2mo ago
>avoid extra detail when it does not help I wonder if they actually do it to optimize inference. I maintain a corporate AI server and one of the tricks to reduce the load was to modify the system prompt to be as terse as possible so the
32.
▲
by
kgeist
2mo ago
Transformers lack recursion and are limited by the network's fixed depth, so "reasoning", IMHO, is basically a way to emulate deeper recursion. As we go through the layers, concepts are pattern-matched and refined, but at som
33.
▲
by
kgeist
2mo ago
In my experience, almost all the problems that microservices advertise solving can also be solved with a modular monolith plus some tooling to enforce certain rules (say, one module shouldn't be able to peek into another module's
34.
▲
by
kgeist
3mo ago
The knowledge is lossy, and code generation itself is non-deterministic (temperature), so the operator-tree executor must be interference from other DB implementations, because it's uncommon to have a bytecode interpreter
35.
▲
by
kgeist
3mo ago
We've shipped some code generated by Qwen3.6 27B to production (under OpenCode). It lacks the breadth of knowledge of models like Opus, but if a change is fully inferable from the prompt and the surrounding code, it works very well. It
36.
▲
by
kgeist
3mo ago
Even if no Rust code for it was seen during training, an LLM can trivially transpile SQLite's C codebase to Rust on the fly. For example, I just asked ChatGPT to write John Carmack's famous Fast Inverse Square Root algorithm in Er
37.
▲
by
kgeist
3mo ago
I reproduced it in Sqlite with short-lived reads/writes though. Other DBMSes seem to not have this issue (IIRC MySQL will block a write if WAL falls behind)
38.
▲
by
kgeist
3mo ago
On the page you linked: >However, if a database has many concurrent overlapping readers and there is always at least one active reader, then no checkpoints will be able to complete and hence the WAL file will grow without bound. >This
39.
▲
by
kgeist
3mo ago
They use WAL in SQLite. If I continuously perform reads/writes so that they overlap with no gaps, I can make their VM go down because SQLite will not have time to initiate a checkpoint to trim the WAL file. SQLite waits for a time wind
40.
▲
by
kgeist
3mo ago
Ollama uses 4 bit quants and a very short context window by default. It can easily break on anything more complex than a simple chat.
41.
▲
by
kgeist
3mo ago
Qwen3.6 below Q8 often can't exit a reasoning loop (until it hits max output token count), forgets to insert a tool call, often mistakenly inserts them inside the thinking block... It's still usable though.
42.
▲
by
kgeist
3mo ago
How was qwen3.6 launched? The thing is, everyone has their own variant of "qwen3.6 27b" depending on the launch parameters, ranging from "SOTA in its class" to "completely broken"
43.
▲
by
kgeist
3mo ago
VPN, accessible only from inside the corporate network
44.
▲
by
kgeist
3mo ago
Yes, Rocket.Chat
45.
▲
by
kgeist
3mo ago
We've been self-hosting GitLab for about a year now, and I don't remember it ever going down or being unavailable. We self-host almost everything else too (except for online meetings), and it's all been pretty stable as well.
46.
▲
by
kgeist
3mo ago
It converges to "almost deterministic" on highly predictable outputs (i.e. code) with the right sampling params (say, you only sample the most probable token without randomness/high temperature) and with self-correction loops
47.
▲
by
kgeist
3mo ago
>they turned it into something unreadable Did you compare the code before/after? It's a mechanical line-by-line port, and most of the code is identical to the old version, just with Rust syntax. They have an example in the blog
48.
▲
by
kgeist
3mo ago
On artificialanalysis.ai, Kimi 2.7 Code is way worse than GLM 5.2 at everything (general intelligence, coding, agentic tasks). But here, both Kimi 2.7 and its derivative SWE-1.7 are ahead of GLM 5.2. This tells me the benchmarks they use ar
49.
▲
by
kgeist
3mo ago
>Pragmatically, often users without new browsers and OSses are not the best clients Hmm, it could be fat enterprise clients with locked-down software versions (legacy, security etc.) That's where most of the money is, isn't it?
50.
▲
by
kgeist
3mo ago
Judging by the examples, if I understand it correctly, J-space supports higher-order logical / multihop transformations, but it is limited in size because of the limited network depth (max number of layers). When we emulate "reaso
51.
▲
by
kgeist
3mo ago
Probably an instance of: "The Reversal Curse: LLMs trained on "A is B" fail to learn "B is A" https://arxiv.org/abs/2309.12288
52.
▲
by
kgeist
3mo ago
Fable 5 was released on June 9 and removed on June 12. GLM-5.2 was released on June 13. It would be an amazing feat to make a model SOTA in just 3 days but I highly doubt it. It's more like z.ai released an existing checkpoint earlier
53.
▲
by
kgeist
3mo ago
>$40k gets you almost-Opus GLM 5.2 is "almost Opus," and it needs at least 8xH200s for comfortable inference (so it's closer to $400k than $40k). They suggest using this modified model: >A REAP-pruned (≈22% of experts r
54.
▲
by
kgeist
3mo ago
I run a corporate AI server and coding peak hours here are 1PM-5PM judging by AI usage stats. My guess is that people spend 9AM-12PM in meetings and at lunch, and the actual coding starts around 1 PM.
55.
▲
by
kgeist
3mo ago
From the perspective of LLM inference, you currently mostly care about: - Memory bandwidth; BUT the requirements are currently capped because models have stopped growing at around 1-1.5 trillion parameters for quite a while now. You only ne
56.
▲
by
kgeist
3mo ago
Yeah people don't realize these "toy models" now completely destroy gpt-4o on most tasks, and no one called gpt-4o a toy model back in the day... It was OpenAI's flagship model from 2024 to 2025.
57.
▲
by
kgeist
3mo ago
Some countries and jurisdictions still have laws that allow for the involuntary confinement of tuberculosis patients, I guess dating back to the times when tuberculosis was rampant in those countries? And most professionals seem to be okay
58.
▲
by
kgeist
3mo ago
See my answer in this same subthread. I was perplexed myself as to why I was diagnosed based on just one radiology report. But the moral of my story is that you can always try to obtain a second opinion from another doctor. I'm not say
59.
▲
by
kgeist
3mo ago
>never read or send .env, .env.*, .pem, id_, .aws/, .ssh/. A think a better practice is to not store those things in the repository folder in the first place.
60.
▲
by
kgeist
3mo ago
I have a memory of myself lying motionless for what felt like forever and staring at a wallpaper featuring Disney characters. Years later I found photos from when I was around 1 year old and the home had those wallpapers (we stayed there fo
More ›