LLM Evaluation and Alignment, the Foundational Ideas
Autor Han Leeen Limba Engleză Paperback – 24 noi 2026
It’s a fundamental truth that all software—even AI systems—is broken. AI engineers who can diagnose faults and refine systems to align with business needs are in high demand. This book expands the foundational research into judging and adapting AI systems into a collection of practical techniques you can use on the job. As you trace the progression from surface-level text matching to semantic similarity to judgment-based evaluation, you’ll build the mental models necessary to choose the right metrics, detect failure modes, and close the loop from evaluation to alignment.
> evaluate > analysis > align cycle, you’ll start making more informed tradeoffs and expertly balancing helpfulness, safety, and brand voice in your models.
What's inside
• BLEU, ROUGE, BERTScore, COMET, and LLM-as-a-judge methods
• Detecting and quantifying hallucinations
• Aligning AI with RLHF, constitutional AI, and red teaming
• Timeless best practices that will apply as models evolve
About the reader
For AI engineers and LLM practitioners. No prior knowledge of NLP metrics, reinforcement learning, or alignment research is required.
About the author
Han Lee has spent more than a decade applying cutting-edge research on large-scale AI and machine learning systems into production-grade products. A Senior Director of Data and AI at Moody’s, he leads teams that ship generative AI applications and has daily, hands-on exposure to safety-critical evaluation pipelines.
Preț: 349.51 lei
Preț vechi: 436.88 lei
-20% Precomandă
Puncte Express: 524
Carte nepublicată încă
Livrare prin curier în România Precomanda se expediază când titlul devine disponibil.
Transport gratuit de la 400.00 lei Plată online sau ramburs, în funcție de opțiunile comenzii.
Retur gratuit în 14 zile Comandă securizată și suport în română.
Doresc să fiu notificat când acest titlu va fi disponibil:
Se trimite...
Specificații
Notă biografică
Han Lee has spent more than a decade applying cutting-edge research on large-scale AI and machine learning systems into production-grade products. A Senior Director of Data and AI at Moody’s, he leads teams that ship generative AI applications and has daily, hands-on exposure to safety-critical evaluation pipelines.