4 ms·Evolution Strategies at Scale: LLM Fine-Tuning Beyond Reinforcement Learning2 points by che_shr_cat 1y ago