- Home
- Computers
- Data Processing
- AI Model Evaluation
AI Model Evaluation
List Price:
$59.99
| Expected release date is Dec 29th 2026 |
- Availability: Confirm prior to ordering
- Branding: minimum 50 pieces (add’l costs below)
- Check Freight Rates (branded products only)
Branding Options (v), Availability & Lead Times
- 1-Color Imprint: $2.00 ea.
- Promo-Page Insert: $2.50 ea. (full-color printed, single-sided page)
- Belly-Band Wrap: $2.50 ea. (full-color printed)
- Set-Up Charge: $45 per decoration
- Availability: Product availability changes daily, so please confirm your quantity is available prior to placing an order.
- Branded Products: allow 10 business days from proof approval for production. Branding options may be limited or unavailable based on product design or cover artwork.
- Unbranded Products: allow 3-5 business days for shipping. All Unbranded items receive FREE ground shipping in the US. Inquire for international shipping.
- RETURNS/CANCELLATIONS: All orders, branded or unbranded, are NON-CANCELLABLE and NON-RETURNABLE once a purchase order has been received.
Product Details
Author:
Leemay Nassery
Format:
Paperback
Pages:
250
Publisher:
Manning (December 29, 2026)
Imprint:
Manning
Release Date:
December 29, 2026
Language:
English
ISBN-13:
9781633435674
ISBN-10:
1633435679
Weight:
10.56oz
Dimensions:
7.375" x 9.25"
File:
Eloquence-SimonSchuster_08192026_P10502905_onix30-20260819.xml
List Price:
$59.99
Pub Discount:
37
As low as:
$56.99
Publisher Identifier:
P-SS
Discount Code:
H
Folder:
Eloquence
Overview
De-risk AI models, validate real-world performance, and align output with product goals.
Before you trust critical business systems to an AI model, you need to answer a few questions. Will it be fast enough? Will the system satisfy user expectations? Is it safe? Can you trust the output? This book will help you answer these questions and more before you roll out an AI system—and make sure it runs smoothly after you deploy.
In AI Model Evaluation you’ll learn how to:
About the book
AI Model Evaluation teaches you how to effectively evaluate and assess machine learning models for better scaling and integration into production systems. Each chapter tackles a different evaluation method. You'll start with offline evaluations, then move into live A/B tests, shadow traffic deployments, qualitative evaluations, and LLM-based feedback loops. You’ll learn how to evaluate both model behavior and engineering system performance, with a hands-on example grounded in a movie recommendation engine.
About the reader
For practitioners with experience in machine learning, data science, or software engineering. Familiarity with Python is recommended.
About the author
Leemay Nassery is an engineering leader specializing in experimentation and personalization. With a notable track record that includes evolving Spotify's A/B testing strategy for the Homepage, launching Comcast's For You page, and establishing data warehousing teams at Etsy, she firmly believes that the key to innovation at any company is the ability to experiment effectively.
Before you trust critical business systems to an AI model, you need to answer a few questions. Will it be fast enough? Will the system satisfy user expectations? Is it safe? Can you trust the output? This book will help you answer these questions and more before you roll out an AI system—and make sure it runs smoothly after you deploy.
In AI Model Evaluation you’ll learn how to:
- Build diagnostic offline evaluations that uncover model behavior
- Use shadow traffic to simulate production conditions
- Design A/B tests that validate model impact on key product metrics
- Spot nuanced failures with human-in-the-loop feedback
- Use LLMs as automated judges to scale your evaluation pipeline
About the book
AI Model Evaluation teaches you how to effectively evaluate and assess machine learning models for better scaling and integration into production systems. Each chapter tackles a different evaluation method. You'll start with offline evaluations, then move into live A/B tests, shadow traffic deployments, qualitative evaluations, and LLM-based feedback loops. You’ll learn how to evaluate both model behavior and engineering system performance, with a hands-on example grounded in a movie recommendation engine.
About the reader
For practitioners with experience in machine learning, data science, or software engineering. Familiarity with Python is recommended.
About the author
Leemay Nassery is an engineering leader specializing in experimentation and personalization. With a notable track record that includes evolving Spotify's A/B testing strategy for the Homepage, launching Comcast's For You page, and establishing data warehousing teams at Etsy, she firmly believes that the key to innovation at any company is the ability to experiment effectively.









