Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
GaggiX
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
5 ms
·
31.
▲
by
GaggiX
3mo ago
I would be more interested about terrorists organization like Al-Shabaab that at least control many towns. Does Boko Haram and ISWAP even control a single town or they just control a few villages in Lake Chad and in the Sambisa forest? Also
32.
▲
by
GaggiX
3mo ago
Grok did not render anything, they had to prompt it again.
33.
▲
Explaining Attention with Program Synthesis
(arxiv.org)
2 points
by
GaggiX
3mo ago
|
0 comments
34.
▲
by
GaggiX
3mo ago
Someone should try it while sleeping and see if anything is related to a dream.
35.
▲
by
GaggiX
4mo ago
Popular open models on Openrouter have dozens of providers.
36.
▲
by
GaggiX
4mo ago
I guess we can start a chain of comment listing their favorite ones, I haven't tried many but "wobble stack" is pretty cool and simple, "Starfall Defense" is a much more complex game but I love tower defense.
37.
▲
by
GaggiX
4mo ago
One of the many therapies that are being developed so that you can survive longer even with the most lethal tumours.
38.
▲
by
GaggiX
4mo ago
Chinese models have become really good and cheap. MiMo V2.5 Pro, Kimi K2.7-code, Minimax M3 etc
39.
▲
by
GaggiX
4mo ago
This is just an Ad, I thought it was a new product from Kagi.
40.
▲
by
GaggiX
4mo ago
I know nothing about this field, but I imagine the actual problem is how do you deliver the Cas12a2 protein to each individual cancer cell compare to a viral gene therapy?
41.
▲
by
GaggiX
4mo ago
Codex also works, before Opus 4.8 and Fable it wasn't very clear who had the best agentic model.
42.
▲
by
GaggiX
4mo ago
Well with a standard autoregressive model you can generate for example 256 tokens at once if you have 256 users, with this approach you can generate 256 tokens for a single user but you need several forward steps. So the diffusion process t
43.
▲
by
GaggiX
4mo ago
$0.435/$0.87 for the standard speed, this one should be 3 times that.
44.
▲
by
GaggiX
4mo ago
Not to be confused with Flash Attention. What's novel here is the extremely small KV cache memory usage per long context windows, like 0.77GB with 512K, a 90% memory usage reduction compare to the already really small KV cache memory u
45.
▲
FlashMemory-DeepSeek-V4
(arxiv.org)
1 points
by
GaggiX
4mo ago
|
1 comments
46.
▲
by
GaggiX
4mo ago
>I’ll take a few f bombs and the truth. Don't want to ruin it but go read some old posts from the author about AI, the tone is the same and he is very much wrong.
47.
▲
by
GaggiX
4mo ago
If MiMo v2.5 Pro can run at >1000tk/s on GPUs then I will soon expect the same from OpenAI/Anthropic/Google.
48.
▲
Ideogram 4.0 Technical Details: Open model at the forefront of design
(ideogram.ai)
3 points
by
GaggiX
4mo ago
|
0 comments
49.
▲
by
GaggiX
4mo ago
> That's technically encoding Isn't that just projecting the patches into the d_model size vectors that the models takes? >I am assuming that involves of quantization 12B model in 16GB seems very reasonable to me, int8 is to
50.
▲
by
GaggiX
4mo ago
>Even though it feels vibed Where does it feel vibed?
51.
▲
by
GaggiX
4mo ago
Is opus being used for the audio or it's not the solution for extreme lossy compression?
52.
▲
by
GaggiX
4mo ago
I love this, hope to see a AV2 version at 8MB
53.
▲
by
GaggiX
4mo ago
I would love to see comparisons with AV1 on very low bitrates.
54.
▲
by
GaggiX
4mo ago
The giant Umarell in the background is a nice piece of furniture. Edit: I noticed later it was in Milan, I guess it makes perfect sense.
55.
▲
by
GaggiX
4mo ago
Given your definition what's the difference between AGI and superintelligence? AGI should at least match, not surpass humans in every cognitive task.
56.
▲
Snowboard Kids 2 is 100% Decompiled
(blog.chrislewis.au)
284 points
by
GaggiX
5mo ago
|
109 comments
57.
▲
Nvidia: Nemotron Labs Diffusion 14B
(huggingface.co)
3 points
by
GaggiX
5mo ago
|
0 comments
58.
▲
by
GaggiX
5mo ago
I wonder does it mean that ublock origin has anti-anti-adblock functionality? (My guess is yes but I wanted to take the opportunity to spell that word)
59.
▲
by
GaggiX
5mo ago
If don't want to spend 1.5$/9$ for the lastest model then yes use a cheaper model, DeepSeek V4 Flash is 0.11$/0.22$ on OpenRouter and it's more capable than the most expensive model a year ago. Models have never been so
60.
▲
Cerebras Brings Trillion Parameter Inference to Enterprises with Kimi K2.6
(cerebras.ai)
2 points
by
GaggiX
5mo ago
|
0 comments
More ›