9798270714826 - building robust ai evals: proven strategies for testing, monitoring, and improving llm performance: 6 di primeaux, henry v. (12 risultati)
Lingua: Inglese
Editore: Independently published, 2025
- Brossura
Da: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 19,33
EUR 2,32 spedizioneSpedito in U.S.A.Quantità: Più di 20 disponibili
Condizione: New.
Lingua: Inglese
Editore: Independently published, 2025
- Brossura
Da: GreatBookPrices, Columbia, MD, U.S.A.GreatBookPrices
Contatta il venditoreVenditore con 5 stelleCondizione: Usato - Come nuovo
EUR 20,98
EUR 2,32 spedizioneSpedito in U.S.A.Quantità: Più di 20 disponibili
Condizione: As New. Unread book in perfect condition.
- Altre immagini
Lingua: Inglese
Editore: Independently Published, 2025
- Brossura
Da: Rarewaves USA, OSWEGO, IL, U.S.A.Rarewaves USA
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 23,73
Spedizione gratuitaSpedito in U.S.A.Quantità: Più di 20 disponibili
Paperback. Condizione: New.
Lingua: Inglese
Editore: Amazon Digital Services LLC - Kdp, 2025
- Brossura
Da: PBShop.store UK, Fairford, GLOS, Regno UnitoPBShop.store UK
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 21,70
EUR 4,87 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
PAP. Condizione: New. New Book. Shipped from UK. Established seller since 2000.
- Altre immagini
Lingua: Inglese
Editore: Independently Published, 2025
- Brossura
Da: Rarewaves.com USA, London, LONDO, Regno UnitoRarewaves.com USA
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 28,78
Spedizione gratuitaSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
Paperback. Condizione: New.
Lingua: Inglese
Editore: Independently published, 2025
- Brossura
Da: GreatBookPricesUK, Woodford Green, Regno UnitoGreatBookPricesUK
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 21,69
EUR 17,57 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
Condizione: New.
Lingua: Inglese
Editore: Independently published, 2025
- Brossura
Da: GreatBookPricesUK, Woodford Green, Regno UnitoGreatBookPricesUK
Contatta il venditoreVenditore con 5 stelleCondizione: Usato - Come nuovo
EUR 22,81
EUR 17,57 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
Condizione: As New. Unread book in perfect condition.
- Altre immagini
Lingua: Inglese
Editore: Independently Published, 2025
- Brossura
Da: Rarewaves USA United, OSWEGO, IL, U.S.A.Rarewaves USA United
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 25,11
EUR 43,95 spedizioneSpedito in U.S.A.Quantità: Più di 20 disponibili
Paperback. Condizione: New.
Lingua: Inglese
Editore: Independently published, 2025
- Brossura
- Print on Demand
Da: California Books, Miami, FL, U.S.A.California Books
Contatta il venditoreVenditore con 4 stelleCondizione: Nuovo
EUR 21,73
Spedizione gratuitaSpedito in U.S.A.Quantità: Più di 20 disponibili
Condizione: New. Print on Demand.
- Altre immagini
Lingua: Inglese
Editore: Independently Published, 2025
- Brossura
Da: Rarewaves.com UK, London, Regno UnitoRarewaves.com UK
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 25,65
EUR 76,12 spedizioneSpedito da Regno Unito a U.S.A.Quantità: Più di 20 disponibili
Paperback. Condizione: New.
Lingua: Inglese
Editore: Independently Published, 2025
- Brossura
- Print on Demand
Da: Grand Eagle Retail, Bensenville, IL, U.S.A.Grand Eagle Retail
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 24,07
Spedizione gratuitaSpedito in U.S.A.Quantità: 1 disponibili
Paperback. Condizione: new. Paperback. Building Robust AI Evals: Proven Strategies for Testing, Monitoring, and Improving LLM PerformanceAre your AI models truly performing as intended, or are hidden failures silently undermining their reliability? In an era where large language models power critical business operations, custome…r interactions, and research breakthroughs, rigorous evaluation is not optional-it's essential. "Building Robust AI Evals" provides a comprehensive, hands-on blueprint for testing, monitoring, and improving LLM performance across real-world applications.This book offers practical, actionable strategies for designing evaluation pipelines that are scalable, repeatable, and aligned with both business and technical goals. From defining meaningful metrics and curating high-quality datasets to implementing automated and human-in-the-loop evaluation workflows, you will learn how to ensure your AI systems are not only accurate but safe, reliable, and compliant.Inside, you will discover how to: Design effective evaluation frameworks that align with business objectives and technical requirements.Implement core and advanced metrics for LLMs, including semantic similarity, multi-step reasoning, and multi-modal assessment.Build modular, automated evaluation pipelines with logging, monitoring, and regression testing for scalable deployments.Detect data drift, concept drift, and performance anomalies in production, and trigger timely retraining and re-evaluation.Integrate safety, fairness, and compliance checks into all stages of evaluation, ensuring ethical and reliable model behavior.Leverage human-in-the-loop and multi-evaluator strategies to capture nuanced model performance beyond automated metrics.Scale evaluation practices across teams and projects while maintaining governance, traceability, and knowledge transfer.Whether you are an AI engineer, data scientist, or machine learning practitioner responsible for deploying large language models, this book equips you with the tools and frameworks to implement evaluation processes that are actionable, auditable, and robust. By following the techniques in this guide, you will reduce risk, improve model reliability, and gain confidence in the real-world performance of your AI systems. This item is printed on demand. Shipping may be from multiple locations in the US or from the UK, depending on stock availability.
Lingua: Inglese
Editore: Independently Published, 2025
- Brossura
- Print on Demand
Da: CitiRetail, Stevenage, Regno UnitoCitiRetail
Contatta il venditoreVenditore con 5 stelleCondizione: Nuovo
EUR 25,32
EUR 43,33 spedizioneSpedito da Regno Unito a U.S.A.Quantità: 1 disponibili
Paperback. Condizione: new. Paperback. Building Robust AI Evals: Proven Strategies for Testing, Monitoring, and Improving LLM PerformanceAre your AI models truly performing as intended, or are hidden failures silently undermining their reliability? In an era where large language models power critical business operations, custome…r interactions, and research breakthroughs, rigorous evaluation is not optional-it's essential. "Building Robust AI Evals" provides a comprehensive, hands-on blueprint for testing, monitoring, and improving LLM performance across real-world applications.This book offers practical, actionable strategies for designing evaluation pipelines that are scalable, repeatable, and aligned with both business and technical goals. From defining meaningful metrics and curating high-quality datasets to implementing automated and human-in-the-loop evaluation workflows, you will learn how to ensure your AI systems are not only accurate but safe, reliable, and compliant.Inside, you will discover how to: Design effective evaluation frameworks that align with business objectives and technical requirements.Implement core and advanced metrics for LLMs, including semantic similarity, multi-step reasoning, and multi-modal assessment.Build modular, automated evaluation pipelines with logging, monitoring, and regression testing for scalable deployments.Detect data drift, concept drift, and performance anomalies in production, and trigger timely retraining and re-evaluation.Integrate safety, fairness, and compliance checks into all stages of evaluation, ensuring ethical and reliable model behavior.Leverage human-in-the-loop and multi-evaluator strategies to capture nuanced model performance beyond automated metrics.Scale evaluation practices across teams and projects while maintaining governance, traceability, and knowledge transfer.Whether you are an AI engineer, data scientist, or machine learning practitioner responsible for deploying large language models, this book equips you with the tools and frameworks to implement evaluation processes that are actionable, auditable, and robust. By following the techniques in this guide, you will reduce risk, improve model reliability, and gain confidence in the real-world performance of your AI systems. This item is printed on demand. Shipping may be from our UK warehouse or from our Australian or US warehouses, depending on stock availability.




