See exactly what every run costs.
Tokens and dollars attributed to the step that emitted them. Roll up by agent, build, deployment, model vendor. Catch cost drift before the invoice arrives.
How it works.
Every LLM span carries token counts and the model used. Trefur multiplies tokens by current pricing per model, per vendor, to produce a dollar figure on each span. Sum the spans in a trace, you get cost per run. Sum the runs in a window, you get cost per agent, per build, per deployment.
Pricing tables are kept current for OpenAI, Anthropic, Bedrock, Azure OpenAI, Gemini, Groq, and MiniMax. Custom model pricing for self-hosted or negotiated rates is available on request.
Cost is graphed live. Filter by agent, build SHA, environment, or any tag you emit. Set an envelope per agent — Trefur alerts when cost-per-run drifts above it, with the worst-offending traces already linked.
What you can do with it.
What the cost view looks like.
Works with.
Catch the next cost spike before finance does.
Free tier. No card required. First trace in under five minutes.