What problem does it solve? Long, filler-heavy AI responses waste output tokens and slow down reading when you only need the technical answer. This Skill enforces an ultra-compressed communication style that cuts output tokens by roughly 65% while preserving full technical accuracy. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus three Classical Chinese (wenyan) variants for maximum character reduction. - Persistent mode: stays active across the whole session until explicitly disabled with "stop caveman" or "normal mode". - Auto-clarity fallback: automatically drops compression for security warnings, irreversible action confirmations, and ambiguous multi-step instructions. - Use Case: During a long debugging session, say "caveman mode" and receive terse, direct answers like "Bug in auth middleware. Token expiry check use < not <=." instead of paragraph-long explanations. ## Quick Start Tell the assistant "use caveman mode" or invoke /caveman to switch all subsequent responses to compressed output.