What problem does it solve? Long, verbose AI responses waste tokens and slow down reading. This Skill cuts output token usage by roughly 65-75% by stripping filler, articles, and pleasantries while keeping every technical detail, code block, and error string exact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus three Classical Chinese (文言文) variants for maximum character reduction. - Persistent session mode: once activated, the compressed style applies to every response until explicitly stopped. - Auto-clarity fallback: automatically reverts to normal prose for security warnings, irreversible-action confirmations, and ambiguous multi-step instructions. - Use Case: During a long debugging session, enable full mode so every explanation arrives as terse fragments like "Inline obj prop → new ref → re-render. useMemo.", cutting reading time without losing precision. ## Quick Start Ask the AI to talk like caveman or invoke /caveman to start receiving ultra-compressed responses for the rest of the session.