Y
HN Search
Hacker News Search
new
|
comments
|
top
|
jobs
starzmustdie
searching PlanetScale…
1.
▲
2.
▲
3.
▲
4.
▲
5.
▲
6.
▲
3 ms
·
1.
▲
by
starzmustdie
6mo ago
> (2012): Introducing Google Drive... yes, really [1] What a difference in announcing products between then and now [1] https://googleblog.blogspot.com/2012/04/introducing-google-d...
2.
▲
Show HN: #1 On This Day
(onthisday-theta.vercel.app)
18 points
by
starzmustdie
6mo ago
|
1 comments
3.
▲
A minimal hackable implementation of policy gradients (GRPO, PPO, REINFORCE)
(github.com)
1 points
by
starzmustdie
9mo ago
|
0 comments
4.
▲
by
starzmustdie
1y ago
GitHub: https://github.com/open-thought/reasoning-gym
5.
▲
Reasoning Gym: Procedural Dataset Generation for Reinforcement Learning
(github.com)
1 points
by
starzmustdie
1y ago
|
0 comments
6.
▲
by
starzmustdie
2y ago
Reasoning Gym ( https://github.com/open-thought/reasoning-gym ) A library that procedurally generates datasets for training reasoning models (like o1/r1) with verifiable rewards.
7.
▲
Show HN: Word Game Bench – evaluating language models on word puzzles
(wordgamebench.github.io)
1 points
by
starzmustdie
2y ago
|
0 comments
8.
▲
Show HN: Answers to Chip Huyen's ML Interview Questions
(github.com)
3 points
by
starzmustdie
3y ago
|
0 comments