LIBRISTO
LIBROAMANTO
obvezno
Pridružite se zajednici ljubitelja knjige iz cijelog svijeta i ostvarite mnoštvo pogodnosti. Izradite besplatni račun
0
Besplatna dostava Overseas kurirskom službom iznad 69.99 €
DPD kurir 3.99 DPD točka 3.49 GLS Kurir 4.99 GLS paketomat 3.99 Hrvatska pošta 4.99 Dostava Overseas 4.99 Box Now 4.49

Besplatna dostava putem Box Now paketomata i Overseas kurirske službe iznad 69,99 €.

Testing AI

Engineering Confidence in Non-Deterministic Systems: Practical Strategies for Evaluating LLMs, AI Agents and Generative AI with Evals, Automated Testing, Reliability, Observability and Production-Ready Quality Engineering

Jezik EngleskiEngleski
Knjiga Meki uvez
Knjiga Testing AI Erma K. Derossett
Libristo kod: 53615338
Nakladnici Independently published, kolovoz 2026
How do you test software when the same input does not always produce the same output?Traditional sof... Cijeli opis
? points 119 b Novo Novo
49.16
Vanjske zalihe Šaljemo za 14-21 dana

Do 30 dana za povrat

How do you test software when the same input does not always produce the same output?

Traditional software testing depends on predictable behavior: provide an input, compare the result with an expected output, and determine whether the system passes or fails. AI changes that equation.

Large language models can generate multiple valid responses. AI agents may take different paths toward the same objective. Retrieval systems depend on changing context, and model or prompt updates can improve one capability while quietly degrading another. An AI application that performs brilliantly in a demonstration can still fail when confronted with real users, edge cases, unexpected inputs, and production workloads.

Testing AI: Engineering Confidence in Non-Deterministic Systems is a practical guide to tackling this new generation of quality-engineering challenges.

Designed for software engineers, QA professionals, AI/ML engineers, developers, technical leaders, and teams building AI-powered products, this book shows how to move beyond traditional pass/fail testing and develop systematic methods for evaluating the quality, reliability, and performance of non-deterministic systems.

Inside, You'll Discover How To:
  • Design meaningful evaluations for LLM-powered applications
  • Define and measure quality when outputs are probabilistic
  • Build repeatable automated evaluation pipelines
  • Create representative datasets and effective test cases
  • Combine deterministic checks, human evaluation, and model-based evaluation
  • Test prompts, structured outputs, tool use, and multi-step workflows
  • Evaluate retrieval-augmented generation and grounded responses
  • Test AI agents for task completion and tool interactions
  • Detect hallucinations, regressions, inconsistencies, and unexpected behavior
  • Establish useful metrics, baselines, thresholds, and acceptance criteria
  • Integrate AI evaluations into development and CI/CD workflows
  • Use observability to understand production behavior
  • Monitor quality degradation and emerging failure patterns
  • Turn production feedback into stronger evaluation suites
Move Beyond "Does It Work?"

AI quality cannot always be reduced to one correct answer. Effective testing must account for variation, context, factuality, relevance, robustness, task completion, latency, cost, and the consequences of failure.

This guide helps you build an engineering approach to that uncertainty. Instead of treating evaluation as a final checkpoint, you'll learn how to incorporate testing throughout the AI development lifecycle-from early experimentation and prompt changes to regression testing, deployment, monitoring, and continuous improvement.

Build AI Systems with Greater Confidence

A compelling prototype is only the beginning. Production AI must perform across unpredictable interactions while models, prompts, retrieval pipelines, tools, data, and user behavior continue to evolve.

Whether you're building an LLM application, RAG pipeline, AI assistant, agentic workflow, or generative AI product, this book provides practical strategies for turning uncertain behavior into measurable engineering evidence.

Stop relying on impressive demos and intuition alone. Build evaluation, testing, and observability practices that help reveal when your AI works, where it fails, and whether it is ready for production.

Get your copy of Testing AI today and start engineering confidence into every stage of your AI development lifecycle.

Glumica & Poliglotkinja
EWA KASP za
Pusti video
Ewa Kasp
Libristo ima najveći izbor literature na stranim jezicima. Zato svoje knjige kupujem ovdje.

Informacije o knjizi

Puni naziv Testing AI
Jezik Engleski
Uvez Knjiga - Meki uvez
Datum izdanja 2026
Broj stranica 486
EAN 9798194306510
Libristo kod 53615338
Težina 1117
Dimenzije 216 x 280 x 25
Poklonite ovu knjigu još danas
To je jednostavno
1 Dodajte knjigu u košaricu i odaberite isporuku kao poklon 2 Zauzvrat ćemo vam poslati kupon 3 Knjiga dolazi na adresu poklonoprimca

Prijava

Prijavite se na svoj račun. Još nemate Libristo račun? Otvorite ga odmah!

 
obvezno
obvezno

Nemate račun? Ostvarite pogodnosti uz Libristo račun!

Sve ćete imati pod kontrolom uz Libristo račun.

Otvoriti Libristo račun