Hands-On RAG for Production
Shipping & Delivery
Our Delivery Time Frames Explained
2-4 Working Days: Available in-stock
14-28 Working Days: On Backorder
Will Deliver When Available: On Pre-Order or Reprinting
We ship your order once all items have arrived at our warehouse and are processed. Need those 2-4 day shipping items sooner? Just place a separate order for them!
Product details
- ISBN 9798341621718
- Dimensions: 178 x 232mm
- Publication Date: 30 Jun 2026
- Publisher: O'Reilly Media
- Publication City/Country: US
- Product Form: Paperback
Retrieval-augmented generation (RAG) is the go-to strategy for integrating large language models with your organization's unique knowledge. However, the market is full of RAG pipelines and components, making it hard to choose the right solution for your enterprise's needs. This book simplifies the process, offering a comprehensive road map to building, refining, and scaling production-grade RAG applications.
Authors Ofer Mendelevitch and Forrest Bao guide you through every phase of development, from data ingestion, embeddings, and vector search to advanced techniques like agentic RAG, multimodal RAG, and GraphRAG. Engineers and architects will learn how to tackle the challenges they'll encounter when building RAG applications at enterprise scale: ensuring high accuracy with minimal hallucinations, maintaining low-latency performance, safeguarding data privacy, and providing transparent, explainable responses among them.
- Determine whether to build RAG yourself or deploy a RAG-as-a-service platform
- Build a basic RAG stack that maximizes performance and cost-effectiveness
- Measure key metrics such as hallucinations, response quality, latency, and cost
- Address challenges in enterprise deployment, such as compliance with data security and privacy requirements, explainability, and prompt design
- Implement advanced techniques such as multimodal RAG, agentic RAG, and GraphRAG
