• Fri frakt över 249 kr
  • •
  • Snabba leveranser
  • •
  • Billiga böcker
Kundservice

Du är på sajten för privatpersoner.

Företag, bibliotek eller offentlig verksamhet?

Du handlar på classic.bokus.com, där alla dina funktioner finns intakta.
Till classic.bokus.com
Bokus logotyp. Gå till startsidan.
  • Erbjudanden
  • Nyheter
  • Student
  • Topplistor
  • Barn & ungdom
  • Bokus Play
  • E-böcker
  • Pocketböcker
  • Spel & pussel

10% rabatt på allt med kod NYSTART10 →

Sidfot

Mina sidor

    Hjälp

    • Kundservice
    • Vanliga frågor och svar
    • Frakt och leverans
    • Retur vid ångerrätt
    • Reklamera vara
    • Betalning
    • Köpvillkor
    • Allmänna villkor
    • Information om webbplatsens tillgänglighet

    Om Bokus

    • Om oss
    • Pressrum
    • För studenter
    • För företag
    • För bibliotek och offentlig verksamhet
    • För leverantörer
    • Hållbarhet

    Populärt

    • Aktuella erbjudanden
    • Presentkort
    • Studentlitteratur
    • Nya böcker
    • Topplistor
    • Signerade böcker
    • Engelska böcker

    Inspiration

    • Boktips
    • BookTok
    • Populära bokserier
    • Barnbokskaraktärer
    • Populära författare
    Logotyp för Bokus
    Följ oss på Facebook (extern länk)Följ oss på Instagram (extern länk)Följ oss på YouTube (extern länk)Följ oss på TikTok (extern länk)
    bokus @ CookiesAnpassa cookiesIntegritetspolicyKöpvillkor
    Till Citymail hemsida (extern länk)Till Budbee hemsida (extern länk)Till Postnord hemsida (extern länk)Till Schenker hemsida (extern länk)Till Early Bird hemsida (extern länk)Till Walleys hemsida (extern länk)
    1. Ekonomi och Ledarskap
    2. Företagsekonomi
    3. Affärsförhandlingar

    AI Model Evaluation

    AvLeemay Nassery

    Häftad, Engelska, 2026

    Del i serien Manning Publications

    428 kr

    Kommande

    Beskrivning

    Before you trust critical business systems to an AI model, you need to answer a few questions. Will it be fast enough? Will the system satisfy user expectations? Is it safe and can you trust the output? This book will help you answer these questions before you roll out an AI system, and make sure it runs smoothly after you deploy. 

    • Learn simple ways to test how your model behaves before it is used in real systems. 
    • Try your model with real data to see how it performs in real situations. 
    • Design A/B tests that validate model impact on key product metrics. 
    • Spot nuanced failures with human-in-the-loop feedback and qualitative evaluations. 
    • Use LLMs to help review and test models more quickly. 

    AI Model Evaluation teaches you how to effectively evaluate and assess machine learning models for better scaling and integration. Each chapter looks at a different way to test a model, starting with offline evaluations and moving into live A/B tests, shadow traffic deployments and LLM-based feedback loops. The book uses a hands-on example grounded in a movie recommendation engine. 

    After reading this book, you will be able to evaluate both model behaviour and engineering system performance. You will have the tools to ensure your AI systems are effective and reliable in production. This book is for practitioners with experience in machine learning, data science, or software engineering. 

    Produktinformation

    • Utgivningsdatum:2026-11-04
    • Mått:117 x 293 x 13 mm
    • Vikt:356 g
    • Format:Häftad
    • Språk:Engelska
    • Serie:Manning Publications
    • Antal sidor:250
    • Upplaga:1
    • Förlag:Pearson Education
    • ISBN:9781633435674

    Utforska kategorier

    • Affärsförhandlingar inom Ekonomi och Ledarskap
    • Företagsekonomi inom Ekonomi och Ledarskap
    • Affärsstrategi inom Ekonomi och Ledarskap

    Mer om författaren

    Leemay Nassery is an engineering leader specialising in experimentation and personalisation. She is known for improving A/B testing approaches at major companies. With experience at Spotify, Comcast, and Etsy, Leemay brings strong practical knowledge of building effective testing and evaluation systems to her writing. She turns her belief in careful testing into a practical guide that helps people try new ideas with confidence.

    Innehållsförteckning

    • PART 1: AI MODEL OFFLINE EVALUATIONS 1 SETTING THE STAGE FOR OFFLINE EVALUATIONS 2 ANATOMY OF AN OFFLINE EVALUATION 3 USING OFFLINE EVALUATIONS AS DIAGNOSTICS 4 ENGINEERING SYSTEM PERFORMANCE EVALUATIONS 5 COUNTERFACTUAL EVALUATIONS PART 2: AI MODEL ONLINE EVALUATIONS 6 EVALUATING MODELS IN AN A/B TEST 7 FROM OFFLINE EVALUATION TO LIVE EXPERIMENT 8 PITFALLS OF ONLINE METRICS PART 3: LLM-AS-A-JUDGE 9 LLM-AS-A-JUDGE FUNDAMENTS 10 DESIGN PATTERNS FOR LLM-AS-A-JUDGE PART 4: HUMAN-IN-THE-LOOP OR QUALITATIVE EVALUATIONS FOR AI MODELS 11 WHY HUMAN EVALUATION MATTERS 12 HUMAN-IN-THE-LOOP EVALUATION TECHNIQUES 13 TRUST, SAFETY, AND RED TEAMING 14 OPERATIONALIZING HUMAN FEEDBACK