What problem does it solve?
This Skill addresses silent token waste that leads to unexpected rate limit errors during AI coding sessions, where long conversations, incorrect model choices, verbose output, and unnecessary full-file processing burn through context capacity without warning. Most assistants waste additional tokens explaining how to save tokens instead of acting immediately under pressure, worsening the problem right when resource conservation is critical.
Core Features & Use Cases
- Auto-activation on emergency red flags: Triggers immediately when any one of 7 symptoms is present (rate limit warnings, 40%+ context usage, 5+ loaded MCP plugins, Opus used for simple tasks, etc.) using OR logic so no multiple symptoms are required to act.
- Immediate, permission-free corrective actions: Follows a pre-defined action matrix to compact context, switch to appropriate models, disable unused MCPs, request text excerpts instead of full file ingestion, and refuse inefficient sub-agent usage without asking for confirmation.
- Strict terse response rules: Enforces under-100-word responses under pressure, eliminating markdown sections, reasoning blocks, tables, and unnecessary justifications that waste tokens.
- Use case example: If you are running a CRUD refactor with Opus and hit a rate limit, the skill immediately switches you to Sonnet and begins work without lengthy justification, saving thousands of tokens.
Quick Start
Mention that you are receiving rate limit warnings or ask Claude to complete a task without losing context to activate the skill's automatic emergency token-saving workflow.