Evals For Ai Engineers

Prijzen vanaf
69,99

Uitgelicht

VERGELIJK ALLE AANBIEDERS (3)

Beschrijving

Bol Stop using guesswork to find out how your AI applications are performing. Evals for AI Engineers equips you with the proven tools and processes required to systematically test, measure, and enhance the reliability of AI applications, especially those using LLMs. Written by AI engineers with extensive experience in real-world consulting (across 35+ AI products) and cutting-edge research, this practical resource will help you move from assumptions to robust, data-driven evaluation. Ideal for software engineers, technical product managers, and technical leads, this hands-on guide dives into techniques like error analysis, synthetic data generation, automated LLM-as-a-judge systems, production monitoring, and cost optimization. You'll learn how to debug LLM behavior, design test suites based on synthetic and real data, and build data flywheels that improve over time. Whether you're starting without user data or scaling a production system, you'll gain the skills to build AI you can trust—with processes that are repeatable, measurable, and aligned with real-world outcomes. Run systematic error analyses to uncover, categorize, and prioritize failure modesBuild, implement, and automate evaluation pipelines using code-based and LLM-based metricsOptimize AI performance and costs through smart evaluation and feedback loopsApply key principles and techniques for monitoring AI applications in production

Vergelijk aanbieders (3)

Shop
Prijs
Verzendkosten
Totale prijs
69,99
Gratis
69,99
Naar shop
Gratis Shipping Costs
80,50
Gratis
80,50
Naar shop
Gratis Shipping Costs
80,50
Gratis
80,50
Naar shop
Gratis Shipping Costs
Beschrijving (2)
Bol

Stop using guesswork to find out how your AI applications are performing. Evals for AI Engineers equips you with the proven tools and processes required to systematically test, measure, and enhance the reliability of AI applications, especially those using LLMs. Written by AI engineers with extensive experience in real-world consulting (across 35+ AI products) and cutting-edge research, this practical resource will help you move from assumptions to robust, data-driven evaluation. Ideal for software engineers, technical product managers, and technical leads, this hands-on guide dives into techniques like error analysis, synthetic data generation, automated LLM-as-a-judge systems, production monitoring, and cost optimization. You'll learn how to debug LLM behavior, design test suites based on synthetic and real data, and build data flywheels that improve over time. Whether you're starting without user data or scaling a production system, you'll gain the skills to build AI you can trust—with processes that are repeatable, measurable, and aligned with real-world outcomes. Run systematic error analyses to uncover, categorize, and prioritize failure modesBuild, implement, and automate evaluation pipelines using code-based and LLM-based metricsOptimize AI performance and costs through smart evaluation and feedback loopsApply key principles and techniques for monitoring AI applications in production

Amazon

Pagina's: 225, Paperback, O'Reilly Media


Productspecificaties

Merk O'reilly
EAN
  • 9798341660724
Maat


Prijshistorie

* Prijshistorie bevat geen data van Amazon, Amazon Marketplace.

Prijzen voor het laatst bijgewerkt op:

Uitgelichte Keuze
69,99
Naar shop