WebRL: Training LLM Web Agents via Self-Evolving Online Reinforcement Learning (arxiv.org) 23 points by theredsix 1y ago ↗ HN
1 comment
[ 1.0 ms ] story [ 17.1 ms ] thread