How to Build End-to-End LLM Observability in FastAPI with OpenTelemetry
This article shows how to build end-to-end, code-first LLM observability in a FastAPI application using the OpenTelemetry Python SDK. Instead of relying on vendor-specific agents or opaque SDKs, we wi
Jessica Patel This article shows how to build end-to-end, code-first LLM observability in a FastAPI application using the OpenTelemetry Python SDK. Instead of relying on vendor-specific agents or opaque SDKs, we will manually design traces, spans, and semantic attributes that capture the full lifecycle of an LLM-powered request. Table of Contents Introduction Prerequisites and Technical Context Why LLM Observability Is Fundamentally Different Reference Architecture: A Traceable RAG Request Reference Architecture Explained Why This Design Is Better Than Simpler Alternatives LLM Models That Work
Explore this link on the map →saved by
related reading
- Building an LLM evaluation framework: best practices | Datadogdatadoghq.com
- LLM evaluation: a beginner's guideevidentlyai.com
- A pragmatic guide to LLM evals for devsnewsletter.pragmaticengineer.com
- Langfuselangfuse.com
- Patterns for Building LLM-based Systems & Productseugeneyan.com
- GitHub - traceloop/openllmetry: Open-source observability for your GenAI or LLM application, based on OpenTelemetry · GitHubgithub.com
- Agent Observability and Tracingarize.com
- Building Effective AI Agents \ Anthropicanthropic.com
- State of AI 2025: 100T Token LLM Usage Study | OpenRouteropenrouter.ai
- Building Effective AI Agents \ Anthropicanthropic.com
- 2025: The year in LLMssimonwillison.net
- Your AI Product Needs Evals – Hamel's Blog - Hamel Husainhamel.dev