Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
raylad
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
raylad
7d ago
My take away from this is that perhaps it’s better to not compact and instead to start new sessions whenever possible. I tell the agents to write a HANDOVER.md file that contains sufficient information to resume but not any extraneous infor
2.
▲
by
raylad
8d ago
I just checked that model (ISTA-DASLab/Qwen3.8-27B-GSQ-RCO-GGUF:IQ3_S) and it does much much worse on the "Please recite Jabberwocky" test than the original bf16 does. The bf16 only misses "snicker-snack" and this q
3.
▲
by
raylad
8d ago
How does that compare with the bf16 version? For my "Please recite Jabberwocky" test the bf16 almost passes but the ternary and even fp8 versions fail badly.
4.
▲
by
raylad
10d ago
Wait till they start mixing every language in one sentence, and using poetic allusions that only scholars of each language would know.
5.
▲
by
raylad
12d ago
Sounds great except it’s too tall to get under the bed.
6.
▲
by
raylad
1mo ago
In 5 years the models will probably be so much better and more compact that phones will be running models equivalent at least to Opus 4.6 if not Fable, at least within the areas they are tuned for (which probably won't include coding).
7.
▲
by
raylad
2mo ago
It seems to fail. I sent the prompt: “ 5° warmer” And it said: “ setting the temperature to 5°F”
8.
▲
by
raylad
2mo ago
As a child, I lived in Firenze (Florence, Italy) and had a pet cricket. All the kids had them, in little bamboo cages, for the Festival of the Cricket: https://www.italymagazine.com/forums/do-see/6010-festival-cr..
9.
▲
by
raylad
2mo ago
Love it! But 10 is a bit of a time commitment. I feel like 5 would be ideal for a daily diversion.
10.
▲
by
raylad
2mo ago
Anything you ever do with any non-locally-hosted model always "uploads code" to the inference provider because that's how it works: the model uses tools to inspect the code, the result of the tool use is sent in an API call t
11.
▲
by
raylad
2mo ago
This got a downvote and I understand why: because I didn't describe the test, which is to ask it "Please recite Jabberwocky". This is actually difficult because there are so many invented words in the poem which have extremel
12.
▲
by
raylad
2mo ago
Not impressed. It fails the "Jabberwocky" test.
13.
▲
by
raylad
3mo ago
Looks very cool. I would like to try it, but don't want to use API billing. OpenAI I think would allow it to use account login. Would you support that?
14.
▲
by
raylad
4mo ago
Possibly a deliberate strategy by the Chinese to undermine the US AI industry, data centers, and basically everything that’s powering the economy. Just like they did with the US steel industry in the 80s.
15.
▲
by
raylad
4mo ago
I wonder if this violates noncompete/no reverse engineering clauses in many or some SAAS agreements?
16.
▲
by
raylad
5mo ago
Seems very cool, but also opens an additional attack vector into your development machine or personal laptop as the case might be.
17.
▲
by
raylad
5mo ago
So all these sites like EquityZen or Forge Global or Hiive etc. who sold "shares" in these companies through SPVs are now going to lose their customer's money? How would that all play out?
18.
▲
by
raylad
5mo ago
The Pedometer functionality is free, and I’ve been using it for many years just because its display is pleasant. The map tracking features cost $29.95 a year.
19.
▲
by
raylad
5mo ago
The article complains about the keyboard layout. And they are probably right. But if you are going to use a VT-100 keyboard you might as well try an editor actually designed for that keyboard, which I remember really loving at the time. KED
20.
▲
by
raylad
5mo ago
Actual write up: https://www.arimlabs.ai/writing/loss-of-control
21.
▲
by
raylad
5mo ago
I haven’t noticed this sort of behavior with opus 4.6, but the first time I used 4.7 it decided to “simplify“ an existing piece of functionality rather than fixing it, which of course made it completely unusable.
22.
▲
by
raylad
5mo ago
That was more than one task. It was 3. I also had Opus 4.7 and Opus 4.6 do audits of a very long document using identical prompts. I then had Codex 5.4 compare the audits. Codex found that 4.6 did a far better job and 4.7 had missed things
23.
▲
by
raylad
5mo ago
I am using 4.7 with the default extra high thinking, and it is clearly very stupid. It's worse than old Sonnet 4.5. I had it suggest some parameters for BCFtools and it suggested parameters that would do the opposite of what I wanted t
24.
▲
by
raylad
6mo ago
One explanation is that the ones who quit believed it was a real experiment and decided they wanted no part of it. The ones who were "obedient" figured out it was fake and treated it as a game.
25.
▲
by
raylad
6mo ago
N-acetylglucosamine actually is effective against MS. Article: https://www.ucihealth.org/about-us/news/2023/09/multiple-scl... Paper: "N-acetylglucosamine inhibits inflammation and neurodegeneration
26.
▲
by
raylad
6mo ago
One question is whether the participants really believed they were giving shocks to the "learners". In college I participated in a number of psychological studies that were similarly deceptive, where one of the other participants
27.
▲
by
raylad
7mo ago
So the real site is https://nanoclaw.dev (putting this here for the search engines to see)
28.
▲
by
raylad
8mo ago
Also why did they get rid of select all? Is there any excuse for that?
29.
▲
by
raylad
8mo ago
"DON'T use pure white or pure black..." This is something I hate: gray text. Designers love it but it is often very illegible because of inadequate contrast.
30.
▲
by
raylad
9mo ago
MANTIC_IGNORE_PATTERNS seems not to be implemented, or am I missing something?
More ›