Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
jg8610
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
Foundation Models for Atoms not Bits
(rootnodes.substack.com)
1 points
by
jg8610
4y ago
|
0 comments
2.
▲
by
jg8610
10y ago
So interestingly, SGD has a nice intuitive explanation for why it is better than GD. If you compute the gradient step for all data, you're expending computational power on redundant data. You're going to get to the minimum with fe
3.
▲
by
jg8610
10y ago
It's good to see people write up their experiments, it's useful for the rest of us to test how we understand neural nets. I think there are a few mistake in your maths though. You can learn a 1-1 discrete mapping through a single