Before you leave...
Take 20% off your first order
20% off
Enter the code below at checkout to get 20% off your first order
Discover summer reading lists for all ages & interests!
Find Your Next Read
AI makes software faster to create and harder to trust.
Testing AI is a practical operating manual for shipping AI systems with confidence. It is written for developers, testers, QA leaders, AI builders, product engineers, architects, technical leaders, and executives who need evidence before putting AI-generated code, agents, chatbots, search systems, RAG workflows, or model-backed products into production.
This book moves beyond one-off demos and simple pass/fail testing. It shows how to measure behavior under uncertainty, design useful evals, use human and LLM judges responsibly, reason about sampling and confidence intervals, test generated code and tool-using agents, monitor production behavior, and decide when to ship, hold, canary, shadow, or roll back.
Inside, you will learn how to:
Along the way, Jason Arbon shares hard-earned lessons from Microsoft, Bing, Google, Chrome, test.ai, and testers.ai, including production failures, misleading metrics, flaky automation, biased labels, broken rollbacks, and the strange new problems created when AI starts helping test AI.
Testing is the doorway. Confidence engineering is the destination.
Thanks for subscribing!
This email has been registered!
Take 20% off your first order
Enter the code below at checkout to get 20% off your first order