Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
recitedropper
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
61.
▲
by
recitedropper
10mo ago
I think in this case, tokenization and percpetion are somewhat analogous. I think it is probably the case our current tokenization schemes are really simplistic compared to what nature is working with. If you allow the analogy.
62.
▲
by
recitedropper
10mo ago
Who wants to bet they benchmaxxed ARC-AGI-2? Nothing in their release implies they found some sort of "secret sauce" that justifies the jump. Maybe they are keeping that itself secret, but more likely they probably just have had h
63.
▲
by
recitedropper
10mo ago
That's a good point, although given I'd never seen this rule I question if it is commonly known enough that it is actually the reason I'm being downvoted. Do you not think what has happened today is suspicious? The Gemini 3 p
64.
▲
by
recitedropper
10mo ago
Not sure if this is agreeing or disagreeing with there being astroturfing. But I'd reckon that the negative sentiments at the top, combined with that there are over eight Gemini 3 posts on the front page recently, is good evidence of m
65.
▲
by
recitedropper
10mo ago
This is the million dollar question. I'm not qualified to answer it, and I don't really think anyone out there has the answer yet. My armchair take would be that watt usage probably isn't a good proxy for computational comple
66.
▲
by
recitedropper
10mo ago
The amount of capital that rests on releases like these is insane. The incentive is just too high to not manipulate places like HN, which have a surprising amount of sway with tech industry sentiment. Edit: Check out how my claim against as
67.
▲
by
recitedropper
10mo ago
Perception seems to be one of the main constraints on LLMs that not much progress has been made on. Perhaps not surprising, given perception is something evolution has worked on since the inception of life itself. Likely much, much more exp
68.
▲
by
recitedropper
10mo ago
Nice, without this thread I would never have known Gemini 3 released today. Going to download Gemini CLI right now™ and see how it performs™ against Cursor, Claude Code, Aider, OpenCode, Droid, Warp, Devin, and ForgeCode.
69.
▲
by
recitedropper
10mo ago
I'm pretty sure they mention in their various TOSes that they don't train on user data in places like Gmail. That said, LLMs are the most data-greedy technology of all time, and it wouldn't surprise me that companies building
70.
▲
by
recitedropper
10mo ago
They claim AI overviews as having "2 billion users" in the sentences prior. They are clearly trying as hard as possible to show the "best" numbers.
71.
▲
by
recitedropper
10mo ago
Perhaps I shouldn't have implied an expectation of lots of explicit mentions of "AGI". It is more the general sentiments being expressed, and the extent to which critical takes seem to be quickly buried. I'm totally open
72.
▲
by
recitedropper
10mo ago
I definitely believe it--I'm not a total AI hater. The jump on the screen usage benchmark is really exciting in that it might substantially help computer-use agentic workflows. That said, I think there is too much a pattern with recent
73.
▲
by
recitedropper
10mo ago
"Since then, it’s been incredible to see how much people love it. AI Overviews now have 2 billion users every month." Cringe. To get to 2 billion a month they must be counting anyone who sees an AI overview as a user. They should
74.
▲
by
recitedropper
10mo ago
Peek the other threads.
75.
▲
by
recitedropper
10mo ago
I'm primarily reacting to the other threads, like the one that leaked the system card early. And, perhaps unfairly, Twitter as well.
76.
▲
by
recitedropper
10mo ago
Inevitable... certainly more so than AGI :)
77.
▲
by
recitedropper
11mo ago
For the most part I think we agree. There is a lot of uncertainty around the mechanics of consciousness, a lot of reasons to doubt the existence of those mechanics in current AI, and a lot of failed endeavors to use biological mimicry to im
78.
▲
by
recitedropper
11mo ago
I hope your generous interpretation is right... I can't really tell what's going on with Anthropic's theater either. They definitely seem like they are vigilant of bad outcomes, going as far as to publish their own economic i
79.
▲
by
recitedropper
11mo ago
I am empathetic to arguments against consciounsess being computational. Definitely strange to imagine an algorithm played out on trillions of abacuses being conscious. That said, I don't think it is a sufficient appeal to entirely disc
80.
▲
by
recitedropper
11mo ago
In this thought experiment, I am considering artificial life genuine. I would agree that there could be productive outlets for our selfish impulses if there was something that mimicked their targets without consciousness to experience the e
81.
▲
by
recitedropper
11mo ago
For the record, I'm agnostic to whether or not consciousnses is possible upon silica. I think it is pretty safe to say though that it likely is an emergent property of specifically-configured complex systems, and humanlike intelligence
82.
▲
by
recitedropper
11mo ago
Repo seems legit, and some of the ideas are pretty novel. As always though, we'll have to see how it scales. A lot of interesting architectures have failed the GPT3+ scale test. As a sidenote--does anyone really think human-like intell
83.
▲
by
recitedropper
1y ago
Hilariously disingenuous.
84.
▲
by
recitedropper
3y ago
I appreciate the clarification, although I was linking this as clearly ConceptARC has a number of tasks where GPT-4 fails badly compared to humans. On "Extend To Boundary" category GPT-4 scores 0.2 and humans score 0.93 -- I'
85.
▲
by
recitedropper
3y ago
https://arxiv.org/pdf/2311.09247.pdf