Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
extr
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
8 ms
·
91.
▲
by
extr
7mo ago
Don't know what to tell you. Sounds like you're holding it wrong. Based on the current state of things I would try to get better at holding it the right way.
92.
▲
by
extr
7mo ago
You probably just don't have the hang of it yet. It's very good but it's not a mind reader and if you have something specific you want, it's best to just articulate that exactly as best you can ("I want a test harne
93.
▲
by
extr
7mo ago
Hard to read due to LLM generated prose.
94.
▲
by
extr
8mo ago
I rarely dream either way (unless I start focusing on that specifically, then my recall will improve quickly). When I was younger and would go to bed severely stoned I would wake up groggy and lethargic - clearly not optimal sleep. On 3-4%
95.
▲
by
extr
8mo ago
This is and has always been trivially configurable. Just put `Task` as a disallowed tool.
96.
▲
by
extr
8mo ago
Part of the issue with legal weed is it's much like if all alcohol was sold as minorly different varieties of Everclear at 150+ ABV, and brands primary boast was just how potent and alcoholic their mix is. It doesn't encourage app
97.
▲
by
extr
8mo ago
You get what you pay for imo.
98.
▲
by
extr
8mo ago
My answer was (for which it did zero thinking and answered near-instantaneously): "Drive. You're going there to use water and machinery that require the car to be present. The question answers itself." I tried it 3 more times
99.
▲
by
extr
8mo ago
thanks dude. you are living my worst nightmare which is that my ultra cool tech demo i made for cracked engineers on the bleeding edge with 128GB ram apple silicon using frontier AI gets adopted by everyone in the world and becomes load bea
100.
▲
by
extr
8mo ago
FWIW I mentioned this in the thread (I am the guy in the big GH issue who actually used verbose mode and gave specific likes/dislikes), but I find it frustrating that ctrl+o still seems to truncate at strange boundaries. I am looking a
101.
▲
by
extr
8mo ago
I tried this today. It's good - but it was significantly less focused and reliable than Opus 4.5 at implementing some mostly-fleshed-out specs I had lying around for some needed modifications to an enterprise TS node/express servi
102.
▲
by
extr
9mo ago
Had this problem awhile ago of my zsh startup being slow. Just opened claude code and told it to benchmark my shell start and then optimize it. Took like 5 minutes and now it's ultra fast. Hardly any idea what it did exactly but worked
103.
▲
by
extr
10mo ago
I think people fool themselves with this kind of thing a lot. You debug some issue with your GH actions yaml file for 45 minutes and think you "learned something", but when are you going to run into that specific gotcha again? In
104.
▲
by
extr
10mo ago
I think Claude is more practically minded. I find that OAI models in general default to the most technically correct, expensive (in terms of LoC implementation cost, possible future maintenance burden, etc) solution. Whereas Claude will tak
105.
▲
by
extr
10mo ago
Are those responses really "better"? Having the LLM tell you you're wrong can mean different things. Your system prompt makes it more direct and less polite, but that's very different from challenging the frame of your q
106.
▲
by
extr
10mo ago
Hard to believe you could be so misinformed. Anthropic is not far behind OAI on revenue and has a much more stable position with most of it coming from enterprise/business customers.
107.
▲
by
extr
10mo ago
Oh yeah I forgot the biggest one. Claude fucking code. Lol
108.
▲
by
extr
10mo ago
It’s crazy how Anthropic keeps coming up with sticky “so simple it seems obvious” product innovations and OpenAI plays catch up. MCP is barely a protocol. Skills are just md files. But they seem to have a knack for framing things in a way t
109.
▲
by
extr
10mo ago
??? Closed US frontier models are vastly more effective than anything OSS right now, the reason they didn’t compare is because they’re a different weight class (and therefore product) and it’s a bit unfair. We’re actually at a unique point
110.
▲
by
extr
11mo ago
Yeah data.table is just about the best-in-class tool/package for true high-throughput "live" data analysis. Dplyr is great if you are learning the ropes, or want to write something that your colleagues with less experience ca
111.
▲
by
extr
11mo ago
Super disappointing there effectively doesn’t exist an “open” competitor in this space that’s close to parity with Cursor/supermaven. Although I wouldn’t have guessed the product category would get out-competed by agentic AI agents wri
112.
▲
by
extr
11mo ago
This is/was a great trick for improving accuracy of small model + structured output. Kind of an old-fashoined Chain of Thought type of thing. Eg: I used this before with structured outputs in Gemini Flash 2.0 to significantly improve t
113.
▲
by
extr
1y ago
I am a SWE myself and use LLMs to write ~100% of my code. That does not mean I fire and forget multiplexed codex instances. Many times I step through and approve every edit. Even if it was nothing but a glorified stenographer - there are su
114.
▲
by
extr
1y ago
> Humans generally don't do that when their goal is clear and aligned (hence deterministic). Look at the language you're using here. Humans "generally" make less of these kinds of errors. "Generally". That i
115.
▲
by
extr
1y ago
I'm taking a much weaker position than the respondent: LLMs are useful for many classes of problem that do not require zero shot perfect accuracy. They are useful in contexts where the cost of building scaffolding around them to get th
116.
▲
by
extr
1y ago
All processes in reality, everywhere, are probablistic. The entire reason "engineering" is not the same as theoretical mathematics is about managing these probabilities to an acceptable level for the task you're trying to per
117.
▲
by
extr
1y ago
Have you never used one to hunt down an obscure bug and found the answer quicker than you likely would have yourself?
118.
▲
by
extr
1y ago
I would completely disagree. I use LLMs daily for coding. They are quite far from AGI and it does not appear they are replacing Senior or Staff Engineers any time soon. But they are incredible machines that are perfectly capable of performi
119.
▲
by
extr
1y ago
Is this just a feeling you have or is this downstream of actual use cases you've applied AI to observed and measured reliability on?
120.
▲
by
extr
1y ago
Based on the comments here, it's surprisingly anything in society works at all. I didn't realize the bar was "everything perfect every time, perfectly flexible and adaptable". What a joy some of these folks must be to wo
More ›