See everything, per agent.
Every LLM request flows through the proxy with full attribution. Cost, model, task type, tokens, latency, all live. The per-agent breakdown uses the system-prompt fingerprint, so there is no annotation work required.
- Per-agent and per-model cost tracking
- Cache-aware accounting (Anthropic prompt caching)
- Full request ledger, exportable
- Local retention, no account, no cap on history
