Follow the session, not the invoice.
Reconstruct working windows from request timing and metadata. See context growth, model changes, cache behavior, and errors turn by turn.
Token intelligence for LLM work
TknScope sits between your tools and model providers, turning every request into clear, session-level usage intelligence—without changing the response your client receives.
The product
Usage data is only useful when it retains the shape of the work that produced it.
Reconstruct working windows from request timing and metadata. See context growth, model changes, cache behavior, and errors turn by turn.
A shared heatmap shows when and where token space lights up—without another dashboard for every provider.
Group by provider, model, session, key, endpoint, or time. Compare the same signal from the angle your question needs.
How it works
Point your existing client at the TknScope proxy.
Auth, request bodies, and streamed responses stay intact.
Provider usage becomes sessions, heatmaps, and breakdowns.
Privacy boundary
Prompts, completions, and API keys never become logs or metric labels. TknScope records the usage facts the provider returns. Content capture is a separate, explicit gate—and it ships off.
capture_content = offpersist_api_keys = neverlog_prompt_text = neverusage_counts = observedTknScope