2 ms·WebRL: Training LLM Web Agents via Self-Evolving Online Reinforcement Learning23 points by theredsix 2y agoHellsMaddy 2y agoRepo seems to be here: https://github.com/THUDM/WebRL https://github.com/THUDM/WebRL