Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
yberreby
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
6 ms
·
1.
▲
CanViT: Toward Active-Vision Foundation Models
(huggingface.co)
3 points
by
yberreby
6mo ago
|
0 comments
2.
▲
by
yberreby
6mo ago
Wouldn't be a very good look if they did anything else.
3.
▲
by
yberreby
7mo ago
The Houthis have been doing a lot of shipping lane disruption, recently. They have sunk several ships. Iran's Islamic regime has provided material and monetary support to the Houthis. Crippling their capabilities aligns with the goal o
4.
▲
by
yberreby
7mo ago
This applies to any US company. Have we forgotten everything we learned in 2012? If your data is shared with Google, Anthropic, Meta, Amazon, or any of their US competitors, it is within reach of the NSA. Whether or not a company provides s
5.
▲
by
yberreby
7mo ago
Sure, but aren't most people running the *Claw projects using cloud inference?
6.
▲
by
yberreby
8mo ago
It's up: https://yberreby.com/posts/hands-free-claude-code/
7.
▲
Hands-Free Claude Code with the Agent SDK
(yberreby.com)
4 points
by
yberreby
8mo ago
|
2 comments
8.
▲
by
yberreby
8mo ago
Since this has garnered some interest, I definitely will sit down and write a blog post when I have a little bit of time. I have upgraded the setup since that post a few days ago, and keep doing so continuously; it's always running in
9.
▲
by
yberreby
8mo ago
That's a fair point, and I had the exact same thought while building this. I had previously resisted the urge of integrating Claude Code with e.g. ntfy.sh for this reason. But in practice, this works for me. I end up being less likely
10.
▲
by
yberreby
8mo ago
Looks like a nice library, thanks for sharing! I know Telegram bots are very popular and that the API story is quite nice, but I have tended to avoid Telegram. My preference would be to go through Signal. I just started looking into my opti
11.
▲
by
yberreby
8mo ago
I'm in the process of migrating from my first POC's disgusting mess of vibe-coded Python to a cleaner (and shareable) Rust architecture. It's going well but I will wait for it to stabilize a bit before sharing. The main non-t
12.
▲
by
yberreby
8mo ago
Watching the OpenClaw/Molbot craze has been entertaining. I wouldn't use it - too much code, changing too quickly, with too little regard for security - but it has inspired me. I often have ideas while cleaning around, cooking, et
13.
▲
by
yberreby
9mo ago
https://yberreby.com
14.
▲
by
yberreby
10mo ago
Would you share some additional details? CPU, amount of unified memory / VRAM? Tok/s with those?
15.
▲
by
yberreby
10mo ago
Based on what works elsewhere in deep learning, I see no reason why you couldn't train once with a randomized number of experts, then set that number during inference based on your desired compute-accuracy tradeoff. I would expect that
16.
▲
by
yberreby
10mo ago
> I.e you can automate things like checking for memory freeing. Or, if you don't need to use C (e.g. for FFI or platform compatibility reasons), you could use a language with a compiler that does it for you.
17.
▲
Drawing with Chaos
(yberreby.com)
6 points
by
yberreby
10mo ago
|
0 comments
18.
▲
by
yberreby
10mo ago
Yes, this is part of the French prépa/CPGE system, which is the "standard" way for students to enter elite engineering schools. You do your first 2-3 years of undergrad in prépa. Source: I did prépa.
19.
▲
by
yberreby
11mo ago
Delightfully evil.
20.
▲
by
yberreby
11mo ago
I have seen it noticed, called out in the talk page, and not rectified.
21.
▲
by
yberreby
11mo ago
I encourage you to look up the "bro" in question. He's a Fields medalist.
22.
▲
by
yberreby
11mo ago
That is also the case on Wikipedia, though. And it's not always trivial to rectify.
23.
▲
by
yberreby
11mo ago
I've seen this approach applied to spectrograms. Convolutions do make enough sense there.
24.
▲
by
yberreby
11mo ago
JAX code usually ends up being way faster than equivalent torch code for me, even with torch.compile. There are common performance killers, though. Notably, using Python control flow (if statements, loops) instead of jax.lax primitives (whe
25.
▲
A Recipe for Training Neural Networks (2019)
(karpathy.github.io)
2 points
by
yberreby
1y ago
|
0 comments
26.
▲
How to Scale
(howtoscalenn.github.io)
2 points
by
yberreby
1y ago
|
0 comments
27.
▲
by
yberreby
1y ago
If C ~ D^2, then D ~ sqrt(C). In other words, the required amount of data scales with the square root of the compute. The square root of 2 ~= 1.414. If you double the compute, you need roughly 1.414 times more data.
28.
▲
by
yberreby
1y ago
I'm curious who among the three you think is "outright fraudulent."
29.
▲
by
yberreby
1y ago
It took me a second to realize you were talking about prompting a LLM . This is fundamentally different from what the parent is doing. "AI" is so much more than "talking to a pretrained LLM."
30.
▲
by
yberreby
1y ago
That shouldn't happen? Normally, the button is grayed out while optimization is running, but after it converges or reaches the max number of steps, you can change the parameters and restart. You may want to lower the max number of step
More ›