Why RL Environments are all you need for building agents
RL Environments are used by frontier labs to train their latest models and agents. In this talk, I will explain why RL environments are fundamental to all stages of building agents: from evaluations and prompt optimization to post-training agents. I will show concrete industry examples of how practitioners use RL environments to evaluate their agents, and to understand how their agents fail. Lastly, I will also talk about how one can set up data flywheels which helps you figure out how to improve your agents on your specific data and use-cases.
About the speaker
Mahesh Sathiamoorthy:
Mahesh Sathiamoorthy is the co-founder and CEO of Bespoke Labs, a data research lab that has raised $40M in Seed and Series A. Previously he was at Google DeepMind where he published seminal work in the intersection of LLMs and recommender systems, which are now being adopted widely in the recommender systems industry across the world.
About
TestMu Conf
Testμ (TestMu) is the world’s largest virtual conference on agentic engineering and quality, built by the community, for the community. As AI reshapes how we build, test, and ship software, Testμ Conf is where you connect, grow, and lead: agentic workflows, autonomous quality, battle-tested AI playbooks, hands-on workshops, and the engineering culture driving it all.