Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
thot_experiment
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
91.
▲
by
thot_experiment
5mo ago
the kids are alright
92.
▲
by
thot_experiment
5mo ago
I've wanted a reliable tool like this for the longest time because I'd like to be able to draw arbitrary poses off my phone if I'm just sketching on the train and I'm having trouble visualizing how a shadow would fall or
93.
▲
by
thot_experiment
5mo ago
idk, my 10 year old makerbot 2 has been pretty reliable, ever since Prusa slicer came out and I tuned a profile for it maybe 6 years ago it's been spitting out quick dimensionally accurate prints. i use it all the time, probably go thr
94.
▲
by
thot_experiment
5mo ago
I can't tell if this is sarcastic or not, but it seems insane to let the AI write the tests. AI can't be held accountable, it shouldn't be writing the tests that determine whether car systems function correctly.
95.
▲
by
thot_experiment
5mo ago
Yo, MTP for Qwen is sick, thank you! Your work is invaluable.
96.
▲
by
thot_experiment
5mo ago
Also you can feed it ALL of your data willy nilly without ever worrying about safety because you can just do it with the LAN cable unplugged, for applications that demand data hygiene it's a cheat code that guarantees safety without an
97.
▲
by
thot_experiment
5mo ago
Yeah, thanks, though I think local models are at least a Cessna, which while being nothing like an F-35 can fly.
98.
▲
by
thot_experiment
5mo ago
Maybe a skill issue but they both feel about the same and the MoE is 3x faster so I barely use the dense model.
99.
▲
by
thot_experiment
5mo ago
This is probably a precision thing, I think there's a really big difference in long running tasks between q4 and q6.
100.
▲
by
thot_experiment
5mo ago
Sorry, "essentially useless in the context of local model availability". It's a fine model but it's tier of inference is fully fungible.
101.
▲
by
thot_experiment
5mo ago
It's more complex than that, I think the reality is that there's a lot of code that's just not that deep bro. I have some purely personal projects that have components that I don't understand anymore, I wrote that shit b
102.
▲
by
thot_experiment
5mo ago
Overall using screentime as the metric, derived from some imperfect logging and vibes it's about 50% OpenCode 15% Continue 15% my homebrew bullshit 13% Claude Code and 7% Cline. I've been deep on agentic stuff lately (1.3wks aka 3
103.
▲
by
thot_experiment
5mo ago
No, it isn't. I am saying that the set of tasks that can be completed by Opus 4.7 has a surprisingly large overlap with the set of tasks that can be completed by Gemma 31B. It is meaningfully equivalent in many cases. (of course if i&#
104.
▲
by
thot_experiment
5mo ago
Benchmarks only give you the roughest idea of how models compare in real world use. They're essentially useless beyond maybe classifying models into a few buckets. The only way you gain an understanding of something as complex as how a
105.
▲
by
thot_experiment
5mo ago
I 100% agree with your philosophy but I wanna note that I genuinely find Gemma 4 31b to be better than Sonnet. To be clear, this makes NO sense to me, so I'm probably just high and making stuff up or just biased by a small sample siz
106.
▲
by
thot_experiment
5mo ago
False. The absolute capability is irrelevant, with the proper harness 31b is more than adequate for a very large portion of the tasks I ask AI to do. The metric isn't how good the model is at Erdos Problems, it's how reliably it c
107.
▲
by
thot_experiment
5mo ago
No, exactly the opposite actually. Qwen3.6 is too imprecise for long running agentic tasks. It doesn't have the same ability to check itself as Gemma does in my testing. I keep Qwen MoE in vram by default because there are tons of task
108.
▲
by
thot_experiment
5mo ago
Flat wrong. Q6 Gemma 31b feels a lot like opus 4.5 to me when run in a harness so it can retrieve information and ground itself. The gap is not that big for a lot of usecases. Qwen MoE is fast as fuck locally for things that are oneshottabl
109.
▲
by
thot_experiment
5mo ago
Depending on your laptop, if your laptop is a Strix Halo or a Macbook with a decent amount of ram, that day they arrived is about 6 months ago, and today if you can run Gemma 31b, you're golden for your basic workslop code. You can d
110.
▲
by
thot_experiment
5mo ago
Re-posting this from a buried comment for visibility because it's just so fucking impressive to me. I went to the store to buy mixers and while I was out Gemma 4 31b got pretty far along with reverse engineering the bluetooth protocol
111.
▲
by
thot_experiment
5mo ago
It may surprise you but over thousands of hours I have actually gathered more than one sample. EDIT: Here's another sample for ya. I went to the store to buy mixers and while I was out Gemma 4 31b got pretty far along with reverse engi
112.
▲
by
thot_experiment
5mo ago
Very different from my experience, Gemma 31b just solved a physics problem Opus 4.7 gave up on. I definitely don't think they're equivalent in general, Opus for sure is way smarter and way more likely to get things right on the ed
113.
▲
by
thot_experiment
5mo ago
Gemma 4 IS good, I've literally had it get a thing right that Opus 4.7 missed, the edges are ragged and I'm reliably finding usecases where it's basically equivalent. Ultimately the metric is "what can I RELY on it t
114.
▲
by
thot_experiment
5mo ago
It's a figurative 900lbs.
115.
▲
by
thot_experiment
5mo ago
I've stopped paying for software outside of games almost entirely. SaaS is a universally terrible UX and it's impossible to actually purchase software anymore. Especially with local LLMs around to smooth out all the rough edges wh
116.
▲
by
thot_experiment
5mo ago
What are you talking about. A. those systems don't make sure you partake responsibly, you even admit it yourself with the claim that legalization reduced popularity. and B. it's like so so so much better than alcohol or tobacco. A
117.
▲
by
thot_experiment
5mo ago
Speaking of, LTT posted a video about DDR pad, which triggered the sleeper cell programming of my youth and I opened up StepMania to play a few rounds. I was shutting down the program and I noticed the build info in the corner. 6-19-2005 My
118.
▲
by
thot_experiment
5mo ago
Wow, I think they finally fixed the awful lag just in time for me to have completely moved my practice to https://graphite.art/
119.
▲
by
thot_experiment
5mo ago
I am gonna check my sources better next time lmao, sorry!
120.
▲
by
thot_experiment
5mo ago
In what particular way? I've been using Typescript a lot more recently (unfortunately XD) and I've found the native experience in Node to be totally fine.
More ›