TestMu Conf 2026
Ship Faster. Test SmarterJoin Now
Ship Faster. Test SmarterJoin Now
SESSION

Beyond the Hype: ML-Driven Test Intelligence at Scale - What Works, What Fails, and Why It Matters

AUG 21, 202606:30 - 07:00 AM (PT)30 MINS

The promise of ML-driven test intelligence is compelling: faster feedback loops, smarter test selection, and anomaly detection that catches what traditional automation misses. But the gap between that promise and production reality is where most teams quietly struggle and rarely talk about it publicly. This session closes that gap.

Drawing from hands-on production experience integrating machine learning into an enterprise QA pipeline inside one of the largest financial institutions in the United States, this talk delivers an unfiltered account of what ML-augmented testing looks like when the stakes are real — regulated environments, high transaction volumes, zero tolerance for silent failures, and teams that still need to ship on time.

We'll walk through a dual-layer ML architecture: a gradient boosting model (XGBoost + scikit-learn) for intelligent test selection, and an LSTM autoencoder (TensorFlow/PyTorch) for post-deployment anomaly detection. Not as a vendor showcase but as a case study in failure, iteration, and eventual production stability. You'll see where the models performed beyond expectations, cutting test execution time dramatically while maintaining defect detection precision. You'll also see where they failed, including a production payment bug that passed ML-assisted test selection cleanly and only surfaced through a human engineer's judgment. That failure became the foundation of a human-veto policy that now governs every ML recommendation in the pipeline.

This talk also tackles the organizational side that most technical sessions ignore: how QA roles evolve when ML enters the pipeline, how to govern model recommendations without creating bottleneck processes, and how to build team trust in a system that sometimes says "skip this test." Critically, this session addresses the security dimension, specifically prompt injection risks in AI-augmented pipelines and why testing AI-integrated systems requires a fundamentally different threat model than testing traditional software.

Whether you are exploring ML for test optimization, mid-implementation and hitting friction, or evaluating whether the investment is worth it, this session gives you the real data, the real failures, and four reusable human-ML collaboration patterns you can take back to your team on day one.

Key Takeaways:

  • Takeaway

    ML test selection works with guardrails: gradient boosting models reduce execution time significantly but require a human-veto policy to catch edge cases the model structurally cannot see.

  • Takeaway

    Anomaly detection is post-deployment QA's most underused lever: LSTM autoencoders catch behavioral drift after release that no pre-deploy test suite will ever surface.

  • Takeaway

    Your biggest ML risk isn't accuracy, it's trust: explainability is not optional if teams are to neither override recommendations arbitrarily nor follow them blindly.

  • Takeaway

    AI-integrated pipelines need security testing, not just functional testing: prompt injection is a real attack surface your QA strategy must account for before production.

  • Takeaway

    QA roles don't disappear with ML, they evolve: the engineers who thrive shift from writing test cases to governing model behavior, validating training data, and owning the human-ML boundary.

About the speaker

Tanvi Mittal:

Tanvi Mittal is a Test Automation Lead and Software Quality Engineering professional with 15+ years of enterprise experience, currently driving AI-augmented testing strategy at U.S. Bank (U.S. Bancorp), one of the largest financial institutions in the United States. She architects large-scale automation frameworks across React and Angular ecosystems, integrates ML-driven test intelligence into production pipelines, and embeds quality engineering deep into DevOps and CI/CD workflows in highly regulated financial environments. An IEEE Senior Member, Tanvi has published peer-reviewed research on ethical AI governance for federated multi-agent systems and serves as a Technical Program Committee reviewer for ITU Kaleidoscope 2026. She is the creator of two open-source tools: LogMiner-QA, a privacy-preserving, AI-powered library for production-log-driven test generation, and PromptArmor, a prompt injection testing library covered under a USPTO provisional patent. Beyond her engineering work, Tanvi is the founder of HerNextTech, a community supporting early-career women in technology, the founding leader of BrowserStack's Cincinnati Chapter, and an active contributor to IEEE WIE.

TESTMU-CONF 2026

GET YOUR FREE BOARDING PASS

I agree to TestMu AI's Privacy Policy, Conference Terms and Conditions.

About
TestMu Conf

Testμ (TestMu) is the world’s largest virtual conference on agentic engineering and quality, built by the community, for the community. As AI reshapes how we build, test, and ship software, Testμ Conf is where you connect, grow, and lead: agentic workflows, autonomous quality, battle-tested AI playbooks, hands-on workshops, and the engineering culture driving it all.

More Sessions

Join the builders, testers, and innovators shaping the next generation of web experiences.
Testμ Conf 2026 is where they meet.

Register Now