Pinned
Software agents can self-improve via self-play RL
Introducing Self-play SWE-RL (SSR): training a single LLM agent to self-play between bug-injection and bug-repair, grounded in real-world repositories, no human-labeled issues or tests. 🧵
Sign up now to get your own personalized timeline!
By signing up, you agree to the Terms of Service and Privacy Policy, including Cookie Use.
