3 ms·
These are simply benchmaxxed versions of either Qwen or Gemma 4.
by S0y 3mo ago
These are simply benchmaxxed versions of either Qwen or Gemma 4.
- jorisw 3mo agoCitation needed
- S0y 3mo agoSure. https://deep-reinforce.com/ornith_1_0.html https://deep-reinforce.com/ornith_1_0.html >Built on top of pretrained Gemma 4 and Qwen 3.5, it achieves state-of-the-art performance among open-source models of comparable size on coding benchmarks. >Ornith-1.0 is a self-improving training framework. Instead of relying on human-designed harnesses to drive solution generation in RL, Ornith-1.0 learns to generate both solution rollouts and the task-specific harnesses that guide those rollouts.
- 2001zhaozhao 3mo agoIf so, it's impressive they managed to benchmaxx Qwen even further than it's already benchmaxxed.
- v3ss0n 3mo agoNah , they just put graphs with different color prioritizing themselves.