caveman

Compress AI model responses into terse caveman-style prose to reduce token consumption.

2|1|Updated Mar 23, 2026
One-click install
npx skills add https://github.com/velopulent/cms --skill caveman-velopulent
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: caveman
Source: https://github.com/velopulent/cms/tree/main/.agents/skills/caveman
Command: npx skills add https://github.com/velopulent/cms --skill caveman-velopulent

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

This skill addresses token bloat and verbosity by compressing AI responses into a highly efficient, caveman-style prose that retains technical accuracy while stripping away filler.

Core Features & Use Cases

  • Multi-Level Compression: Choose from six intensity levels ranging from professional-tight to extreme classical Chinese (wenyan) compression.
  • Context Preservation: Maintains technical integrity, code blocks, and error strings while removing articles, pleasantries, and hedging.
  • Use Case: Use this when working with long-context sessions where you need to minimize token usage without sacrificing the precision of technical instructions or code.

Quick Start

Activate the caveman skill by typing /caveman in the chat to begin receiving ultra-compressed responses.

Frequently Asked Questions about caveman

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I reduce token consumption in AI responses without losing technical accuracy?

Token consumption is reduced by compressing AI responses into terse, caveman-style prose that strips away filler while maintaining strict technical accuracy, code block integrity, and error string precision.

Can I use different compression intensity levels for technical documentation?

Multiple compression intensity levels are available, ranging from lite and full to ultra and classical Chinese (wenyan) registers, allowing varied compression depths for different technical documentation needs.

When do I need to compress AI model outputs into terse prose?

Compressing AI model outputs into terse prose is needed during long-context sessions where minimizing token usage is critical, but sacrificing the precision of technical instructions or code is not an option.

Does the compression mechanism preserve code blocks and error strings?

The compression mechanism preserves code blocks and error strings by removing only articles, pleasantries, and hedging, automatically reverting to standard prose for security-sensitive operations.

What are the limitations of using caveman-style prose for AI communication efficiency?

A key limitation is that strict adherence to technical accuracy must be maintained, meaning the system automatically reverts to standard, uncompressed prose whenever it encounters security-sensitive operations.

How to start compressing AI chat responses for token optimization?

To start compressing AI chat responses, activate the skill by typing the designated slash command in the chat to begin receiving ultra-compressed outputs that save approximately 75 percent of tokens.