What problem does it solve?
Sticker-price comparisons between LLM providers ignore cache-read and cache-write pricing, so a provider with cheaper per-token rates can actually cost more on cache-heavy workloads. This Skill prices your actual month of LangWatch usage, with its real input/output/cache mix, under each candidate provider.
Core Features & Use Cases
- Real Usage Export: Exports your actual token mix per model via the LangWatch CLI, including the cache read/write split from session events, scoped by window and origin.
- Current Price Fetching: Fetches live provider price pages at analysis time and records cache pricing shape (write premium, storage-by-time, or no cache).
- Repricing with Sensitivity: Reprices the same usage under each candidate with direct, no-cache-degradation, and cache-hit sensitivity rows, then writes a self-contained HTML report.
- Use Case: Ask whether switching your coding-agent workload from one provider to another would save money, and get a report showing your $X actual spend versus $Y repriced, with the cache assumptions stated.
Quick Start
Ask the assistant to check whether another model provider would be cheaper for your LangWatch usage over the last 30 days.