TALYVOR LENS

v0.1.0 AI token intelligence
⚠ Unable to reach Lens API. Configure your API key (header Authorization: Bearer <key>) to view live data.

Summary

Total Spend
last 30 days
Cache Hit Rate
Total Requests
Avg Cost / Request

Spend by Model

Model Requests Input Output Cost % of total
Loading…

Cache Performance

Hit Rate
Estimated Savings
Pattern Hash Hits Tokens Saved
Loading…

Circuit Breakers

Loading…

Local Model Status

Loading…

Local model endpoints

Multi-endpoint registry powered by internal/localrouter. Health checks run every 30 seconds; load + latency stats feed the routing strategies (round-robin, least-loaded, lowest-latency, priority). Configure at boot via LENS_LOCAL_ENDPOINTS or manage at runtime:

Workspaces

Loading…

Anomalies (over time)

Temporal — a unit's spend vs its own recent history (the cross-sectional peer view is Cost outliers below).

Loading…

Model capabilities multimodal

Which models can serve which modalities. A multimodal request is routed to a capable model (or fails fast with a clear error) — never silently answered from text. Unknown models are treated as text-only.

Loading…

Model catalog single source

The authoritative model registry — provider, pricing ($/1M tokens), capabilities, and context limits. Cost attribution, capability routing, and introspection all read from here.

Loading…

Git attribution

Workspace-scoped Git rollups (branch, PR, commit, author, repo) recorded automatically when callers send the X-Talyvor-* headers. Per-workspace dashboards consume these endpoints directly: