You Can't assertEquals an Agent: A Tester's Guide to Agentic Quality | TestMu 2026
The first time I tried to write a test for an agentic workflow, I reached for assertEquals — and immediately realized how useless it was. The output was different every run. The tool calls happened in a different order. The reasoning path shifted mid-conversation. Everything I knew about testing said this system was broken. But it was working exactly as designed.
This is not a theoretical problem anymore. Teams across the industry are shipping agents that plan, reason, invoke tools, and make autonomous decisions in production. But testing practices have not kept up — most teams rely on happy-path prompts and manual spot checks because the traditional playbook was never built for systems that think for themselves.
In this talk, I will present a practical testing playbook built from real-world experience of breaking, debugging, and hardening agentic workflows. I will introduce a layered mental model that decomposes any agentic workflow into distinct testable components — giving your team a shared vocabulary to move from “where do we even start” to a structured strategy.
Then I will open a live demo. Using an open-source orchestration framework with a visual graph interface, I will walk through a working agent and systematically break each layer on screen — showing subtle, dangerous failure modes and the concrete testing techniques that catch them. You will walk away with a testing mental model, a working demo you can clone, and the confidence to stop guessing and start testing agents with intention.
A five-layer mental model for decomposing any agentic workflow into testable components — reasoning, tool use, memory, orchestration, and output quality.
Practical testing techniques for non-deterministic systems including trajectory testing, LLM-as-judge evaluation, chaos injection, and semantic regression suites.
A clear framework for deciding what to automate, what needs human review, and where to draw the line between flaky and broken.

TestMu Conf
Testμ(TestMu) Conference is TestMu AI’s (Formerly LambdaTest) annual flagship event, one of the world’s largest virtual software testing conferences dedicated to decoding the future of testing and development. Built by the community, for the community, it’s a space where you’re at the center, connecting, learning, and leading together. From deep-dive sessions on emerging trends in engineering, testing, and DevOps, to hands-on workshops and inspiring culture-driven talks, every experience is designed to keep you at the heart of the conversation.

From AI Assistants to AI Coworkers: How Engineering Teams Ship Faster with Enterprise Context
TestMu 2026
Keynote: Beyond Benchmarks - Evaluating Agents Against What They Are Actually Supposed to Do
TestMu 2026
Panel Discussion: Money Moves at Machine Speed - Trust, Risk, and Quality in Agentic Finance
TestMu 2026
From Load Testing to Reliability Engineering: Making Performance Testing Predict Production Behavior
TestMu 2026
Panel Discussion: Who Tests the Machines? QE Leaders on Quality in the Age of AI-Written Code
TestMu 2026
Fireside Chat: The Economics of AI Agents: How Startups Are Rethinking Value and Monetization
TestMu 2026
Panel Discussion: Mission-Critical Priorities in Quality Engineering: The Leader's Playbook
TestMu 2026