10:15
LLMOps #1 — What Is LLMOps? Why Shipping an LLM Feature Isn't Like Shipping Software
LLMwork
10:06
LLMOps #2 — The LLM App Lifecycle: The Gap Between a Demo and a Product
10:35
LLMOps #3 — Nondeterminism: Why the Same Question Gives You a Different Answer
10:04
LLMOps #4 — Prompt Management: Treating Prompts Like Code
11:22
LLMOps #5 — Evaluation, Part 1: How Do You Know Your LLM System Is Any Good?
10:46
LLMOps #6 — LLM-as-a-Judge: How to Judge the Judge (Evaluation, Part 2)
11:18
LLMOps #7 — Observability and Tracing: Seeing Inside a Single Request
10:37
LLMOps #8 — RAG in Production, Part 1: The Pipeline Behind One Answer
10:53
LLMOps #9 — RAG in Production, Part 2: Retrieval Quality
10:41
LLMOps #10 — Keeping Knowledge Fresh: Sync, Staleness, Deletion, and Permissions in RAG
10:28
LLMOps #11 — Fine-Tune, RAG, or Just a Better Prompt: Making the Decision Properly
10:02
LLMOps #12 — Serving and Latency: Where the Seconds Actually Go in an LLM Request
10:34
LLMOps #13 — Cost, Caching, and Model Routing: Making LLMs Affordable at Scale
10:21
LLMOps #14 — LLM Guardrails: Input Checks, Output Checks, and Failing Safely
LLMOps #15 — Agent Ops: Operating a System That Takes Actions
11:06
LLMOps #16 — Migrations, Incidents, and the Feedback Flywheel: Closing the Loop in LLM Ops