Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
extr
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
extr
2mo ago
Seems pretty reasonable.
32.
▲
by
extr
3mo ago
I think DeepSWE is flawed in a different way: the tasks look like someone took a bunch of big highly technical PRs they found really well done, and inverted it into specs for agents to autistically execute. This is not really how people use
33.
▲
by
extr
3mo ago
[Future voice]
34.
▲
by
extr
3mo ago
I'm skeptical any of that matters at all if at some point AI is perceived by the government to be a true existential risk to public welfare.
35.
▲
by
extr
3mo ago
I don't know that I want to stop such a thing. It's good that nerve gas is banned. I don't want random people having access to easy-to-follow instructions to make COVID-29.
36.
▲
by
extr
3mo ago
They are not going to let open weights models with zero restrictions exist dude. They will be regulated like guns, or probably closer to nerve gas or enriched uranium.
37.
▲
by
extr
4mo ago
I think they were totally correct in spirit. But RE: details about them giving access to an SK corp with possible Chinese ties. Of course that raises eyebrows in the USG, justifiably. Sloppy work from Anthropic.
38.
▲
by
extr
4mo ago
To me they are genuinely trying to walk a tough line - they legitimately believe that they need to warn the public and make a lot of noise so society can try to adapt to this technology. OTOH no adaption (good or bad) can take place if the
39.
▲
by
extr
4mo ago
I mean obviously they're correct but also the complaints of the administration aren't totally without basis. - They're obviously being targeted politically because they refuse to kiss the ring, vibes, whatever you want to cal
40.
▲
by
extr
4mo ago
Really funny to describe OpenAI/Anthropic as a "SaaS"
41.
▲
by
extr
4mo ago
This isn't even about cyber attacks. This is just LLM development which is increasingly just called software development. And at least for cyber it says "Sorry I can't help with that"!
42.
▲
by
extr
4mo ago
I'm a big fan of Anthropic. Just check my post history. I've been accused of working there. But this is complete bullshit and they need to get real. Silent sandbagging is not acceptable, especially given they've shown with th
43.
▲
by
extr
4mo ago
Interesting it's in python!
44.
▲
by
extr
4mo ago
The points in this article don't really land for me. They are mostly critiques of particular MCP implementations rather than the modality itself. My impression right now: - MCPs are great for stateless, mostly read-only interactions wi
45.
▲
by
extr
4mo ago
Damn already there in 154. Thank you man.
46.
▲
by
extr
4mo ago
IMO they have all been clean and noticeable upgrades over their predecessors. Opus 4.7 in particular was a solid jump in capabilities.
47.
▲
by
extr
4mo ago
Are you thinking of the /effort level in Claude Code? I would just go with xhigh as a reasonable default. Most important thing in prompting is specifying what "done" and "success" looks like to you. Ask Claude to he
48.
▲
by
extr
4mo ago
I don't even bother looking at the code until I've run a code review pass on it. Why waste my time with trivial bug fixes? I find the best way to spend time right now is like: - Defining the issue/ticket, what "success&q
49.
▲
by
extr
4mo ago
The advantage is that /code-review supplies a structured idea of how to review and what that process should look like and then launches independent subagents to approach the issue from multiple angles. It's analogous to how in the
50.
▲
by
extr
4mo ago
Hey Boris, some feedback. I like the new /code-review skill but was disappointed you guys removed /simplify because I quite liked the focus on finding code reuse/efficiency opportunities. I see now in 2.1.152 you added those
51.
▲
by
extr
4mo ago
For a lot (most) of what we do with programming, the process actually doesn't matter. I understand you are a real ass dude who is in this shit for the love of the game. I respect that. You are a true artisan and exist in a kind of rari
52.
▲
by
extr
5mo ago
ok thank you for this anecdote
53.
▲
by
extr
5mo ago
Did you RTFA?
54.
▲
by
extr
5mo ago
What is the OP talking about. $/unit intelligence is going down rapidly. You can achieve what would have been considered miracles in 2022 with < $10.
55.
▲
by
extr
5mo ago
Pangram is highly reliable.
56.
▲
by
extr
5mo ago
If you do not live in one of the handful of areas in the US with public transportation infrastructure and also do not own a car you are an extreme outlier. Likewise, if you do not use AI tools to code, outside of some highly niche and speci
57.
▲
by
extr
5mo ago
If they want to continue that practice that's fine but they will quickly find that it severely limits their options for participating in the software industry.
58.
▲
by
extr
5mo ago
You are picking at the analogy rather than engaging with the point. In the US, excluding areas with substantial public transportation infrastructure (realistically just a few major cities), car ownership is nearly universal. You can choose
59.
▲
by
extr
5mo ago
You're welcome to feel that way but it's a luxury belief. In reality, outside of a few (one?) major city in the US with public transportation infrastructure, you need a car. 92% of people own a car, higher if you exclude the dense
60.
▲
by
extr
5mo ago
This is like reading an article "I Don't Drive Cars" that goes on like - They're too expensive - My buddy's 1995 Accord breaks down a lot - Walking is healthier, plus you can stop and smell the roses - I enjoy carin
More ›