Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
mckennameyer
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
4 ms
·
1.
▲
AI takes people at their word
(futuresearch.ai)
5 points
by
mckennameyer
4mo ago
|
1 comments
2.
▲
by
mckennameyer
4mo ago
Claude's great at reading what people say, but surprisingly bad at recognizing when a politician's stance is just the first signal in a negotiation.
3.
▲
by
mckennameyer
4mo ago
You're right, just updated. Original title took one framing from the back half of the post (3 update cycles that can loosely be called the "ChatGPT era, then xAI/Meta/Gemini era, then Anthropic era"), but definitely
4.
▲
How long until AI automates all cognitive labor?
(futuresearch.ai)
45 points
by
mckennameyer
4mo ago
|
82 comments
5.
▲
Apple's Plan to Power Siri with ChatGPT Was a Predictable Failure
(futuresearch.ai)
3 points
by
mckennameyer
5mo ago
|
0 comments
6.
▲
Claude is that gullible friend who takes everyone at their word
(futuresearch.ai)
1 points
by
mckennameyer
5mo ago
|
0 comments
7.
▲
by
mckennameyer
6mo ago
So basically the attacker and the dev who caught it were probably using the same tools if the malware was AI-generated (hence the fork bomb bug), and the investigation was AI-assisted (hence the speed). Less "tip of the iceberg" a
8.
▲
by
mckennameyer
7mo ago
It seems like a marketing play to seize on the protein movement. What will they do when fiber becomes the next craze?
9.
▲
by
mckennameyer
7mo ago
Can definitely relate. I think forcing myself to conduct 1 session at a time feels so difficult not only from an efficiency POV but from an attention standpoint. Waiting for a session to finish, being alone with my thoughts... we're fa
10.
▲
by
mckennameyer
7mo ago
For anyone following the Chalamet drama... next you'll have to look into how many times a best actor frontrunner has lost thanks to their ego last week of the race!
11.
▲
by
mckennameyer
7mo ago
Do you think reasoning and behavioral effort should be separate knobs, or is bundling them the right call?
12.
▲
by
mckennameyer
7mo ago
Aren't vibe PRs way more likely to get abandoned? Sure they reduce reviewer load, but then everyone feels less urgency to do a human review after. Do you think the skill is making that better or worse?
13.
▲
by
mckennameyer
8mo ago
We tested GPT-5 and Gemini Flash 3 at low, medium, and high effort on 169 instances with human-verified answers, scored against a frozen offline web corpus using Deep Research Bench. High effort consistently scored worse than lower thinking
14.
▲
by
mckennameyer
9mo ago
Interesting approach with the cascade. How do you decide when to escalate from fuzzy matching to LLM?
15.
▲
by
mckennameyer
11mo ago
Super interesting direction. I've been pretty skeptical of “AI for stock picking” for the same reasons you mention. Curious how you handle the challenge of companies pivoting into new business areas that don't have historical prec