tracing
6 talks
The Future of Evals: From LLM as a Judge to Agent as a Judge
Citation Needed: Provenance for LLM-Built Knowledge Graphs
Learned Execution Graphs for Anomaly Detection & Drift in APIs
WTF Is the Context Layer? The Missing Infrastructure for Production Agents
The Agentic AI Engineer
LLM Observability, Evaluation, Experimentation Platform