Playbooks
Tutorials
Long-form guides optimized for engineers shipping GenAI features responsibly.
13 min read
Embedding Drift Monitoring in Production for Healthcare Applications
This tutorial covers the implementation of embedding drift monitoring in production systems for healthcare applications, ensuring model accuracy over time. Prerequisites include knowledge of machine learning models and monitoring techniques.
14 min read
Implementing Shadow Traffic for Safe Model Rollouts in E-commerce
This tutorial explains how to implement shadow traffic to test new models in an e-commerce environment without affecting live traffic. Prerequisites include knowledge of machine learning deployment and monitoring practices.
15 min read
Implementing PII Handling in Retrieval Pipelines for Financial Services
This tutorial guides you through implementing PII handling in retrieval pipelines specifically tailored for financial services, ensuring compliance with regulations like GDPR and CCPA. Prerequisites include familiarity with data privacy laws and experience in building retrieval systems.
14 min read
Structured Outputs vs JSON Mode Tradeoffs in E-commerce
This tutorial examines the trade-offs between structured outputs and JSON mode in e-commerce applications. Prerequisites include familiarity with data formats and experience in e-commerce platforms.
16 min read
Golden-Set Design for RAG Faithfulness in Financial Services
This tutorial discusses the design of golden sets to ensure the faithfulness of retrieval-augmented generation (RAG) systems in financial services. Prerequisites include experience with RAG systems and access to financial datasets.
15 min read
Reducing Hallucinations with Citation Constraints in Academic Research
This tutorial explores how to effectively implement citation constraints to minimize hallucinations in academic research models. Prerequisites include familiarity with natural language processing (NLP) and access to a research dataset.
22 min read
Embedding Drift Monitoring in Production for Financial Services
This tutorial focuses on techniques for monitoring embedding drift in production environments specifically tailored for financial services. Prerequisites include understanding of machine learning embeddings and production systems.
18 min read
Shadow Traffic for Safe Model Rollouts in E-commerce Platforms
This tutorial explains how to implement shadow traffic techniques for safely rolling out new machine learning models in e-commerce applications. Prerequisites include knowledge of machine learning deployment and A/B testing.
20 min read
Evaluating Tool-Calling Reliability Under Load in IT Support Systems
This tutorial focuses on assessing the reliability of tool-calling mechanisms in IT support systems during peak loads. Prerequisites include familiarity with load testing and IT support workflows.
15 min read
Ensuring PII Handling in Retrieval Pipelines for Healthcare Applications
This tutorial covers best practices for managing Personally Identifiable Information (PII) in retrieval pipelines within healthcare settings, ensuring compliance with regulations like HIPAA. Prerequisites include understanding of data privacy laws and basic retrieval pipeline concepts.
18 min read
Evaluating Tool-Calling Reliability Under Load in IT Support
This tutorial provides a framework for assessing the reliability of tool-calling in RAG systems under high load conditions, specifically for IT support applications. It requires knowledge of system performance metrics and load testing methodologies.
16 min read
Ensuring PII Handling in RAG Pipelines for Legal Firms
This tutorial focuses on best practices for handling Personally Identifiable Information (PII) in RAG pipelines within legal firms. It requires knowledge of legal compliance and data protection standards.
14 min read
Implementing Cost Controls in RAG: Batching vs Streaming Tokens in Financial Services
This tutorial explores the cost implications of batching versus streaming token usage in RAG systems for financial services. It requires familiarity with RAG tokenization and financial data processing.
12 min read
Comparing Structured Outputs and JSON Mode for RAG in E-commerce
This tutorial examines the trade-offs between structured outputs and JSON mode in RAG systems tailored for e-commerce applications. It requires a basic understanding of RAG and JSON data formats.
15 min read
Optimizing Golden-Set Design for RAG in Healthcare Applications
This tutorial covers the design of golden sets for ensuring RAG (Retrieval-Augmented Generation) faithfulness in healthcare applications. It requires an understanding of RAG principles and access to domain-specific datasets.
10 min read
Multimodal Prompts for Document QA in Legal Settings
Using multimodal prompts can improve document question answering (QA) in legal contexts. Prerequisites include access to relevant legal documents and a model capable of processing multimodal inputs.
11 min read
Reducing Hallucinations with Citation Constraints in Research Models
Implementing citation constraints can significantly reduce hallucinations in research-oriented models. Prerequisites include a robust database of citations and a model capable of handling constraints.
9 min read
Latency Budgets for Streaming Chat UX in Customer Support
Establishing latency budgets can enhance the user experience in customer support chat applications. Prerequisites include understanding user expectations and system capabilities.
12 min read
Embedding Drift Monitoring in Financial Services
Monitoring embedding drift is crucial for financial services to ensure model accuracy over time. Prerequisites include a data pipeline that captures embeddings and a monitoring framework.
10 min read
Shadow Traffic for Safe Model Rollouts in E-commerce
Implementing shadow traffic allows e-commerce platforms to test new models against live traffic without affecting user experience. Prerequisites include a robust logging mechanism and a dual model setup.
10 min read
Offline vs Online Evaluation Frequency
This tutorial explores the differences between offline and online evaluation methods for machine learning models, focusing on their respective benefits and drawbacks. Prerequisites include a basic understanding of machine learning evaluation metrics and experience with model deployment.
20 min read
Pgvector Index Tuning (HNSW vs IVF)
Learn how to tune pgvector indexes using HNSW and IVF algorithms for optimal performance. Prerequisites include familiarity with PostgreSQL and vector databases.
11 min read
Multimodal Prompts for Document QA
Explore how to create effective multimodal prompts for document question answering (QA) systems. Prerequisites include understanding of multimodal models and QA frameworks.
18 min read
SLI/SLO for Generative Endpoints
Establishing Service Level Indicators (SLIs) and Service Level Objectives (SLOs) for generative endpoints is crucial for maintaining quality and reliability. This tutorial outlines how to define and implement SLIs/SLOs effectively.