Skip to main content
Every capability shares this vocabulary. Each term is defined once, in the table for the area that owns it - skim the tables, then follow the links for the full guides.
entity-relationship style diagram. Left: Agent emits Traces (with child Spans and Tool calls), Traces group into Sessions. Middle: Patterns and LLM judge scorers read Traces/Sessions and raise Signals. Right: Datasets (Questions + criteria) drive Evaluation Runs producing Results with ratings. Bottom: Prompts and Tool Schemas registries receive evidence from Signals and Results.

How the core objects relate: traces group into sessions; scorers raise signals; datasets drive eval runs; findings become validated registry proposals.

Tracing

Monitoring

Evaluation

CI/CD

Insights and Improve

The old Improve tab merged into Insights, which holds three views: Suggestions, Dataset coverage, and Auto-improve (confirmed failures become an improvement report that becomes a code fix - see Auto-improve). The Prompts and Tools & MCPs registries live under the sidebar’s Manage section.