Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
Phemist
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
by
Phemist
6d ago
Granted, this one is pretty funny.
2.
▲
by
Phemist
6d ago
I would think the intelligence aspect is a bit hard to define, but to my mind (having called LLMs tools before), the main utility of a tool is reliability. Given a certain world state (including a tool's internal state), its effects ba
3.
▲
by
Phemist
6d ago
Anything that touches on US geopolitical/immigration policy is bound to bring out the russian troll army.
4.
▲
by
Phemist
7d ago
Where are you from? It's weird to me that you on the one hand defend the US, but on the other DID notice all these armed guards everywhere you went in Europe. If you don't feel it's different, then I don't think you shou
5.
▲
by
Phemist
7d ago
Nothing wrong indeed with the sticker that says "I value my life more than your truck". Unfortunately that is not the sticker the OP was talking about.
6.
▲
by
Phemist
7d ago
I've been using Playwright a lot and also hit on the friction between playwright's web testing scope (e.g. no 1st class support for "muted" start-up of browsers, jeez) and usage for task automation all the time. This loo
7.
▲
by
Phemist
8d ago
The secrecy will be volumetric. Im sure it will continue to be possible for bullshit to out-generate information. With bullshit being redefined to true, but informationally/conceptually useless proofs.
8.
▲
by
Phemist
12d ago
Classic mistake. Tell the agent to highlight this code, but dont give it any actual code. Agent hacks its own gem to find the code to highlight.
9.
▲
by
Phemist
14d ago
Also debunked in the Gamer Nexus video. The TV will scan for public wifis and connect. Watching the video, it also seemed like a "missed opportunity" for LG to not create a mesh network of LG devices that would help eachother conn
10.
▲
by
Phemist
17d ago
It definitely is solvable though. Data versioning is a thing and it can work quite transparently to the mutations done on the data. To not know who made and who approved a set of mutations on data can easily become equally as mind-blowingly
11.
▲
by
Phemist
19d ago
Yeah, vanilla MacOS is super focused on the set of "all windows of a given Application" as the useful "unit of work". Whereas in real work, you actually handle individual windows of a set of applications (e.g. firefox wi
12.
▲
by
Phemist
22d ago
My knowledge is a few years outdated by now, but I remember digging into this and realizing that most of the chinese open-source libs were license-washing software. E.g. PaddleOCR is licensed under Apache 2.0, a very permissive license, how
13.
▲
by
Phemist
22d ago
Yes, but expect those prison sentences will be more popular than you might expect. The current admin expends an insane amount of energy pumping the size of their supposed fanbase.
14.
▲
by
Phemist
23d ago
What is the intuition. Higher quality turns due to more reasoning results in significantly fewer turns taken?
15.
▲
by
Phemist
23d ago
Potentially because they run from the same colossus DCs?
16.
▲
by
Phemist
26d ago
This default-to-auto-mode and the misleading marketing is begging for a class action once damages accumulate. Especially considering the Auto Mode even can actively prevent the clean-up!
17.
▲
by
Phemist
29d ago
With every newly released open weight model, the clock on the issues you describe is reset. I can see a marketplace arising for paid updates to common lines of open weight models, which will incentivize those with the hardware to train to f
18.
▲
by
Phemist
29d ago
> They didn't leave my computer I guess, I just showed me the output of some `ls` commands Not sure exactly how Zed works, but wouldn't the results of the `ls` tool call be fed back into claude?
19.
▲
by
Phemist
1mo ago
I am not arguing that there are perhaps other models that can run at the same quality, can coordinate between the different modalities, but are way less power hungry. My point is exactly about the comparison between the token output of the
20.
▲
by
Phemist
1mo ago
They are already working on it. https://huggingface.co/unsloth/Qwen3.8-Flash-Next-GGUF https://unsloth.ai/docs/models/qwen3.8-next > You will need at least 75 GB of RAM or unified memory t
21.
▲
by
Phemist
1mo ago
They are not. If the robot speech is a tool call, then for a fair comparison we need to take the tool call scaffolding (and probably the reasoning too) into account. So rather than a sentence of 10 tokens worth of speech being the output, t
22.
▲
by
Phemist
1mo ago
It seems that e.g. OpenAI is mostly power constrained. To go to the middle of nowhere would mean also basically no power grid to speak off?
23.
▲
by
Phemist
1mo ago
> I don’t see how tokens can’t produce speech or track metabolic needs. It probably could, but the point is this would require additional tokens, blowing up the comparison. The token output of LLMs and "token output" of speech
24.
▲
by
Phemist
1mo ago
The 20W number includes EVERYTHING else the brain does. The chips/models are literally only producing tokens. Let's see an LLM drive a robot harness and have the robot produce speech, as well as move through 3D space, keep track o
25.
▲
by
Phemist
1mo ago
RAM and SSD in apple gear has always been way over-priced. There was a short blessed period in March where the M5 Max macbook pro was out, but the general 30% price hike had not yet happened. In this period, given the insane inflated RAM pr
26.
▲
by
Phemist
1mo ago
Single turn set-ups may work like this. You control the thing you input, the LLM outputs something and then nothing happens further for that specific context. (Simple question/answer style interactions..) (Multi-turn) tool calling set-
27.
▲
by
Phemist
1mo ago
Maybe detecting watermarked text is a skill to be attained. Not allowing proper feedback and training will not allow people to notice the difference on time, thus ruining the data? Best practice is to allow a number (scaled based on complex
28.
▲
by
Phemist
1mo ago
I had Kagi configured to rewrite links old.reddit.com. Looks like I need to update it. Edit: Actually I think this was part of their documentation to explain the intended use of that feature. Too bad this will now no longer work! ( https:&#
29.
▲
by
Phemist
1mo ago
On the same day that OpenAI cuts per token cost by 50% on GPT5.6 Sol. :')
30.
▲
by
Phemist
1mo ago
Ok. I dont have a good feeling for the actual completion distributions. The noise sounds problematic. I can imagine it relates to the size of the context used for hashing. You want this as long as possible so that the entropy is higher, but
More ›