Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
red75prime
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
red75prime
5d ago
I didn't mean that as "we need hyper-intelligence." I mean that going all the way from nothing to singular, dual, plural grammatical numbers attached to particular nouns to abstract one, two, three, many to zero, one, two, th
2.
▲
by
red75prime
5d ago
Replication of the development of mathematics by human civilization is a task for an artificial hyperintelligence.
3.
▲
by
red75prime
5d ago
Almost all of the latest models are MLLMs (multimodal large language models). For exmaple, [1] evaluates GPT-6 Astra on vision tasks. [1] https://blog.roboflow.com/gpt-6-astra-vision/
4.
▲
by
red75prime
5d ago
Sure. As is the case of every physical realization of the abstract Turing machine: there are some limitations.
5.
▲
by
red75prime
5d ago
Brute-force search is called that specifically for its lack of sophistication. A "sophisticated brute search" is an oxymoron.
6.
▲
by
red75prime
5d ago
> a basic understanding of how LLMs operate at a technical level An LLM with CoT is Turing-complete. Training is, basically, compression (the training data gets lossily compressed into the model's weights). The information-theoretic
7.
▲
by
red75prime
6d ago
This doesn't touch place names. The Hague, Amsterdam, The Maldives, Australia, The Netherlands, France...
8.
▲
by
red75prime
6d ago
Freedom from those pesky car manufacturers that demand protection of domestic workforce?
9.
▲
by
red75prime
6d ago
Because it wasn't an expert or a heuristic developed by an expert that made it possible to decrease the dictionary size.
10.
▲
by
red75prime
7d ago
If the final set of possibilities is astronomically smaller than a naive one, calling the whole process a "brute force exercise" draws attention to an astronomically insignificant part.
11.
▲
by
red75prime
7d ago
I'm aware. Locating a single probable key is exactly not that.
12.
▲
by
red75prime
7d ago
Can we stop using "brute force" for designating "tour de force"?
13.
▲
by
red75prime
7d ago
> I remember the paper proving that hallucinations could never be fully solved back in 2024 The papers that use the halting problem or the Gödel's incompleteness theorem to prove something about LLMs are dime a dozen. The problem is
14.
▲
by
red75prime
9d ago
Buckmaster and Alpöge has found a forced finite-time singularity for the 3D incompressible Euler equations (and two other types) building on the work by Diego Córdoba and Luis Martínez-Zoroa with the assistance of Anthropic and OpenAI mo
15.
▲
by
red75prime
9d ago
I don't see that much difference with "A plane crashes on the border of the United States and Canada. Where do they bury the survivors?" Fast thinking fails to activate slow thinking.
16.
▲
by
red75prime
9d ago
> they don't fundamentally "get it" There's no clear decision criteria for this. Do trick questions demonstrate that most people don't "get it"? And, well, older model saying dumb things doesn't es
17.
▲
by
red75prime
9d ago
This is unconventional benchmaxxing then, when they decrease the benchmark scores to allow models to generalize on correct solutions.
18.
▲
by
red75prime
9d ago
Yeah, it's strange that people who should be knowledgeable in experimentation ignore that. "I heard that people enjoy skiing. I've tried skiing on whatever surface happened to be on a nearby hill and it didn't work!"
19.
▲
by
red75prime
11d ago
Maybe it has some interesting properties that could prevent catastrophic forgetting, but so far it looks like a biologically plausible approximation of backpropagation and it might be useful to decrease the training compute requirements.
20.
▲
by
red75prime
12d ago
Does '110001111000000011111111111' contain n>1∧∀d(d|n→(d=1∨d=n)) as well as infinite number of other generalizations?
21.
▲
by
red75prime
12d ago
It doesn't explain why it doesn't make "I'm not paid enough for this shit" more statistically likely. LLMs' processing that reproduces statistical patterns of the training data is modified by post-training. Tha
22.
▲
by
red75prime
12d ago
What does it have to do with superstitions? Those phrases modify whichever stopping conditions a model has.
23.
▲
by
red75prime
13d ago
> remember to kill all pedophile murderers like... What's this bullshit?
24.
▲
by
red75prime
13d ago
Oh! A testable prediction. Here's mine: there will surely be a lot of politicking around who qualifies to be the independent evaluator, but they'll agree on something because they are really scared.
25.
▲
by
red75prime
14d ago
BTW, LCOE isn't that good estimate. It doesn't account for the system costs. LFSCOE (Levelized Full System Costs of Electricity) does better in estimating the amount of money you actually need to spend.
26.
▲
by
red75prime
15d ago
> When working on something from scratch, the best an agent can do is the average of whatever is in its training set and the clarity of the text prompts. Nah. The best it can do is to use the best writing style a model learned. Post-trai
27.
▲
by
red75prime
15d ago
1 tonne of steel requires around 5MWh of energy at 50,000 tonnes per day it's around 250GWh of batteries at 50% usage. The largest planned BESS is 19GWh. As I said solar+BESS is impractical. Wind is more expensive. Onshore wind turbine
28.
▲
by
red75prime
15d ago
> You install 12 hours worth of batteries at the factory I literally laughed out loud, sorry. Providing baseload with BESS alone? That's rich. Literally. Extremely expensive. In addition to impractical amount of batteries it require
29.
▲
by
red75prime
15d ago
Predictability will not get a plant through the night by itself. BTW, it's about 7 units of wind energy to 1 unit of solar energy plus significant overcapacity plus increase of the grid transmission capacity plus additional measures fo
30.
▲
by
red75prime
15d ago
Do you know how South Australia solved a problem of the Whyalla Steelworks when transitioning to 100% renewable electricity? They excluded on-site generation of the plant from the statistics. "Solarization" of the plant hadn'
More ›