What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading when you only need the technical substance. This Skill cuts token usage by roughly 75% by stripping pleasantries, hedging, and filler while keeping all technical content exact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus wenyan-lite, wenyan-full, and wenyan-ultra for classical Chinese compression. - Persistent mode: Stays active across every response until explicitly disabled with "stop caveman" or "normal mode". - Auto-clarity fallback: Automatically drops compression for security warnings, destructive action confirmations, and multi-step sequences where fragments could be misread. - Use Case: During a long debugging session, enable caveman mode so every explanation arrives as dense fragments like "Bug in auth middleware. Token expiry check use < not <=" instead of paragraph-length prose. ## Quick Start Tell the AI to use caveman mode for the rest of this session to make all responses brief and token-efficient.