The Agent Quality Loop: Build, Evaluate, Trace, Score, and Learn
Traditional systems follow predictable execution paths, while agentic systems are non-deterministic, reason across multiple steps, interact with tools, and produce different outcomes for the same request. Their technology stack and failure modes are fundamentally different, requiring quality practices that go beyond conventional test automation. This session presents a continuous evaluation approach spanning development, CI pipelines, production deployment, end-to-end tracing, live scoring, and feedback-driven dataset creation. You will learn how to evaluate agent behavior, detect unexpected failures, establish measurable controls, and create a continuous improvement loop that strengthens reliability, observability, and release confidence.
Key Takeaways:
Understand why agentic systems' technology stack and failure modes require quality practices beyond conventional test automation.
Learn a continuous evaluation approach spanning development, CI pipelines, production deployment, end-to-end tracing, live scoring, and feedback-driven dataset creation.
Discover how to evaluate agent behavior, detect unexpected failures, and establish measurable controls.
Create a continuous improvement loop that strengthens reliability, observability, and release confidence.
About the speaker
Rakesh Sukla:
Rakesh Sukla is an accomplished engineering leader and domain expert in Quality Automation, Cloud-Orchestrated CI/CD, and Developer Productivity, with a proven record of driving innovation at scale. He has consistently served as a foundational engineer across organizations, architecting enterprise-grade automation platforms and orchestrating pipelines from the ground up. He is recognized for improving software quality, accelerating release velocity, and reducing operational costs by building high-impact systems and scaling high-performing global teams.
About
TestMu Conf
Testμ (TestMu) is the world’s largest virtual conference on agentic engineering and quality, built by the community, for the community. As AI reshapes how we build, test, and ship software, Testμ Conf is where you connect, grow, and lead: agentic workflows, autonomous quality, battle-tested AI playbooks, hands-on workshops, and the engineering culture driving it all.