Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
p1esk
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
31.
▲
by
p1esk
1mo ago
Strictly speaking, all we need is them improving AI research.
32.
▲
by
p1esk
1mo ago
The same happens if you ask Opus 5 to "think again"
33.
▲
by
p1esk
2mo ago
Google is not able to currently produce a frontier model despite all the capital.
34.
▲
by
p1esk
2mo ago
If Siri is using a 3T model in high reasoning mode to answer your question you will.
35.
▲
by
p1esk
2mo ago
I stopped writing code completely about a year ago, stopped reading code completely about 8 months ago, and I feel like I stopped thinking hard about anything at work about 6 months ago. To clarify - I still produce code for a living - all
36.
▲
by
p1esk
2mo ago
Is there any degradation with INT8 weights quantization? Why would anyone want to apply ConvRot to do 8 bit weights? Note the paper [1] focuses on 4 bit weights and 4 bit activations (W4A4) quant scheme - a much more challenging goal. My un
37.
▲
by
p1esk
2mo ago
Zero. These were open problems.
38.
▲
by
p1esk
2mo ago
This blog post
39.
▲
by
p1esk
2mo ago
It’s still at GPT-1 level, but GPT-2 moment feels imminent.
40.
▲
by
p1esk
2mo ago
replace they key-query-value mechanic by just dropping it while making the entire context the latent space. What do you mean by this? Like concatenating all token embeddings into one large vector?
41.
▲
by
p1esk
2mo ago
I wonder how difficult it would be to convert these to 3D
42.
▲
by
p1esk
2mo ago
I’m guessing other models would probably stop in this situation and ask the user for instructions.
43.
▲
by
p1esk
2mo ago
I think they meant that Opus 5 had to find and set up a vision model to process the image
44.
▲
by
p1esk
3mo ago
It’s nice to be rich I guess
45.
▲
by
p1esk
3mo ago
where do you see "twice the memory bandwidth"?
46.
▲
by
p1esk
3mo ago
There’s noticeable accuracy degradation when they switched from fp8 to mxfp4
47.
▲
by
p1esk
3mo ago
That’s how you get skills
48.
▲
by
p1esk
3mo ago
Do a decentralized p2p one
49.
▲
by
p1esk
3mo ago
We’re experiencing gpt-2 moment in robotics now. This means in about 2-3 years they will do useful work (cooking, repairs, cleaning, etc).
50.
▲
by
p1esk
3mo ago
It’s got to be similar to Fable, which I experienced for 3 days, and which impressed me (compared to Opus 3.8)
51.
▲
by
p1esk
3mo ago
VLAs are new LLMs. Give them 5 years to develop. But even good old LLMs are still improving every six months.
52.
▲
by
p1esk
3mo ago
Unfortunately if your manager thinks something is your problem, it becomes your problem.
53.
▲
by
p1esk
3mo ago
You realize llms as a field is barely 5 years old? Give it at least another 5.
54.
▲
by
p1esk
3mo ago
Many people here spent a lot more than $300 on headphones long before AirPods appeared.
55.
▲
by
p1esk
3mo ago
100×10^15 km
56.
▲
by
p1esk
4mo ago
Do you feel that recent advances in AI can speed up such rare disease research?
57.
▲
by
p1esk
4mo ago
I’d also like to read about your experience.
58.
▲
by
p1esk
4mo ago
we are very far away from curing death That’s fine - we just need to find a way to slow aging and wait until science advances. Strictly speaking we just need to find a way to keep our brains alive and stimulated, not the whole body.
59.
▲
Nvidia RTX Spark Laptops
(nvidianews.nvidia.com)
8 points
by
p1esk
4mo ago
|
0 comments
60.
▲
by
p1esk
4mo ago
Where do I sign up?
More ›