What problem does it solve? Long, filler-heavy AI responses waste output tokens and slow down reading. This Skill cuts output tokens by roughly 65% by stripping articles, filler words, pleasantries, and hedging while keeping all technical substance, code, and error strings intact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus three classical Chinese (wenyan) variants for maximum character compression. - Persistent mode: Stays active across every response in the session until the user says "stop caveman" or "normal mode". - Auto-clarity fallback: Automatically drops compression for security warnings, irreversible action confirmations, and multi-step sequences where fragments could cause misreading. - Use Case: During a long debugging session, say "caveman mode" to get terse, high-signal answers like "Bug in auth middleware. Token expiry check use < not <=." instead of paragraph-length explanations. ## Quick Start Say "caveman mode" or "use caveman" to switch the assistant into ultra-compressed responses for the rest of the session.