caveman

Compress AI model responses by removing non-essential linguistic elements while preserving technical accuracy and code structure.

Updated Jul 6, 2026
One-click install
npx skills add https://github.com/shirulot/codex-skill --skill caveman-shirulot
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: caveman
Source: https://github.com/shirulot/codex-skill/tree/main/caveman
Command: npx skills add https://github.com/shirulot/codex-skill --skill caveman-shirulot

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill addresses high token consumption by stripping away conversational filler, pleasantries, and redundant prose while maintaining full technical accuracy and code integrity.

Core Features & Use Cases

  • Multi-Level Compression: Offers six distinct intensity levels ranging from professional-tight to extreme classical Chinese compression.
  • Context-Aware Persistence: Maintains the chosen compression style across the entire session until explicitly stopped.
  • Use Case: Use this when working with long-context technical tasks where you need to minimize output tokens without losing critical code blocks, error strings, or technical logic.

Quick Start

Activate the caveman mode by typing /caveman to immediately begin receiving ultra-compressed responses.

Frequently Asked Questions about caveman

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I reduce token consumption in AI responses for technical tasks?

To reduce token consumption in AI responses, you can use compression techniques that strip away conversational filler and redundant prose while preserving technical accuracy and code structure. This ensures long-context tasks use fewer output tokens without losing critical logic.

Does compressed communication remove code blocks and error strings from output?

Compressed communication does not remove code blocks or error strings from output. The process specifically targets non-essential linguistic elements and pleasantries, maintaining full technical accuracy and code integrity throughout the response.

How do I start compressing AI model responses using the caveman skill?

To start compressing AI model responses, activate the caveman mode by typing /caveman. This immediately initiates ultra-compressed responses that persist across the entire session until explicitly stopped.

What compression intensity levels are available for minimizing output tokens?

Available compression intensity levels for minimizing output tokens include lite, full, ultra, and classical Chinese variants. These six distinct levels range from professional-tight to extreme classical Chinese compression to suit different needs.

When should I not use ultra-compressed communication for AI responses?

You should not use ultra-compressed communication during security-sensitive or high-risk technical operations. The system implements automatic clarity triggers that revert to normal prose in these scenarios to ensure safe and clear execution.

Will my chosen compression style stay active for the entire session?

Your chosen compression style will stay active for the entire session. It features context-aware persistence that maintains the selected compression level across multiple interactions until you explicitly stop the mode.