Cloud infrastructure was built around a request whose cost you can estimate before it runs. You size timeouts to it, forecast capacity from its average, rate-limit against it, and bill by it. Agent workloads break that assumption at the source: an agent decides what work to do while it executes, so the count of tool calls, model invocations, and retries is only settled once the run is finished.
This session traces what fails when unpredictable work lands on infrastructure designed for predictable requests. Timeouts fire in the middle of legitimate long runs and surface as silent failures. Capacity plans built on averages miss a heavy tail that a few runs can dominate. Per-request billing and rate limits drift away from the resources actually being consumed.
The larger argument is that the request is the wrong thing to price and bound. The talk proposes a different unit of accounting, closer to the run and its budget, and covers what teams can do about timeouts, forecasting, and cost controls today while the platforms catch up.
The request assumes a knowable cost; agent runs discover their work mid-execution, so that assumption quietly breaks everything built on it.
Timeouts, capacity forecasts, and per-request billing all fail on agent workloads because they're tuned to an average that a heavy tail ignores.
Stop pricing and bounding the request; move to the run and its budget, and here's what you can do about it today.

TestMu Conf
Testμ(TestMu) Conference is TestMu AI’s (Formerly LambdaTest) annual flagship event, one of the world’s largest virtual software testing conferences dedicated to decoding the future of testing and development. Built by the community, for the community, it’s a space where you’re at the center, connecting, learning, and leading together. From deep-dive sessions on emerging trends in engineering, testing, and DevOps, to hands-on workshops and inspiring culture-driven talks, every experience is designed to keep you at the heart of the conversation.

From AI Assistants to AI Coworkers: How Engineering Teams Ship Faster with Enterprise Context
TestMu 2026
Keynote: Beyond Benchmarks - Evaluating Agents Against What They Are Actually Supposed to Do
TestMu 2026
Panel Discussion: Money Moves at Machine Speed - Trust, Risk, and Quality in Agentic Finance
TestMu 2026
From Load Testing to Reliability Engineering: Making Performance Testing Predict Production Behavior
TestMu 2026
Panel Discussion: Who Tests the Machines? QE Leaders on Quality in the Age of AI-Written Code
TestMu 2026
Fireside Chat: The Economics of AI Agents: How Startups Are Rethinking Value and Monetization
TestMu 2026
Panel Discussion: Mission-Critical Priorities in Quality Engineering: The Leader's Playbook
TestMu 2026