A trace groups related requests into one workflow, so you can follow a multi-step agent loop or RAG pipeline end to end instead of staring at isolated calls. Infere lets you set a trace ID with a header and see every request, cost, and latency under that trace.

Why it matters

Agent workflows and RAG pipelines span many model calls. Without grouping, debugging means correlating isolated requests by hand.

How it works

A trace ID is attached to each request in a workflow, usually through a header. Every request, span, cost, and latency under that ID rolls up into one view, so you can follow a single user action end to end.

Example

Set X-Trace-Id on each step of an agent loop and read the whole run as one waterfall instead of a dozen separate calls.

← Back to the full glossary

Put the platform behind the terms

Route, evaluate, and monitor every AI request from one OpenAI-compatible platform.

Start Free → Explore the Features