Timing, without the blind spot.
See latency and status alongside each request. Find the slow call rather than blaming the whole agent.
Agent observability
Follow the request from agent to provider. See its tokens, timing, owner, and catalog cost—without piecing the story together after the fact.
Powered by Caveman Platform · in private development
A total is not an explanation.
A model call is one part of an agent workflow. Select a step below to inspect the request path and the facts recorded along the way.
Record model, reported usage, cache tokens, status, and latency.
48–1,248 msSee latency and status alongside each request. Find the slow call rather than blaming the whole agent.
Record input, output, and cached tokens from the provider. Preserve the basis of the calculation.
Group requests by workflow and agent. Missing attribution remains visible.
Explore a sample ledger. Switch between agents and workflows, then select a request to inspect the cost behind it.
Sample records only. Actual subtotals use public-catalog prices and provider-reported usage, not your negotiated provider invoice.
Investigate related work, repeated paths, and recurring failure themes. Move from one expensive request to a case worth testing.
The same artifact gets read and verified twice. Inspect the traces before deciding whether either check is redundant.
Content-based analysis depends on your capture settings and available evidence. Semantic themes and lexical traffic clusters have different bases; the product keeps that distinction visible.
Missing price is not a zero-dollar call.
An unpriced model stays unpriced. Missing catalog data never quietly becomes $0.
Give a request an owner—or leave it marked unattributed until you know.
Keep the request record close to the calculation. Estimates and verified savings remain separate.
Open the request, inspect its timeline, and follow the work it belongs to. The record is already there when the question arrives.
Full payload visibility depends on your capture and retention settings. Request metadata and usage are distinct from captured content.
Interactive product UI · fictional sample data · Open full demo ↗
Make every token count.
Connect your existing LiteLLM gateway through OpenTelemetry, or use Caveman’s gateway. Caveman Platform is in private development.