caveman

Reduce AI conversation token usage with terse technical responses.

1|4|Updated Apr 9, 2026
One-click install
npx skills add https://github.com/rudometkin/GenAI-intensive --skill caveman-rudometkin
Or copy as Structured Prompt for Agent
Please help me install this Agent Skill.
Skill: caveman
Source: https://github.com/rudometkin/GenAI-intensive/tree/main/projects/DemyanovaDarya/Skills/caveman
Command: npx skills add https://github.com/rudometkin/GenAI-intensive --skill caveman-rudometkin

SYSTEM DOCUMENTATION & REQUIREMENTS

What problem does it solve?

Ultra-compressed communication mode reduces token usage while preserving technical accuracy. Supports intensity levels: lite, full (default), ultra, wenyan-lite, wenyan-full, wenyan-ultra. Use when brevity and accuracy are both needed in technical chats.

Core Features & Use Cases

  • Token-efficient dialogue: terse, precise responses with minimal fluff.
  • Multiple intensity levels: lite, full, ultra, wenyan-lite, wenyan-full, wenyan-ultra.
  • Auto-trigger and switching: auto-triggers when token efficiency is requested; manual toggle via /caveman.

Quick Start

Enable caveman mode by saying caveman or /caveman; responses remain terse with full technical content.

Frequently Asked Questions about caveman

High-intent search queries and answers about installing and using this skill.

FAQPage Schema
How do I reduce token usage in AI conversations during code reviews?

To reduce token usage in AI conversations, enable a terse communication mode that enforces restrictions on articles, filler, and pleasantries while preserving technical accuracy. This mode ensures responses remain precise without wasting context window space.

What is the best way to maintain technical accuracy while using fewer tokens for debugging?

The best way to maintain technical accuracy while using fewer tokens is to switch to a compressed communication mode. It strips filler and hedging from responses, ensuring fast-paced debugging sessions remain precise and token-efficient.

Can I adjust the intensity of token-efficient prompting in technical brainstorming sessions?

Yes, you can adjust the intensity of token-efficient prompting using multiple levels: lite, full, ultra, wenyan-lite, wenyan-full, and wenyan-ultra. These levels allow you to control the compression depth during technical brainstorming sessions.

How do I manually toggle a compressed dialogue mode in my engineering workflow?

You can manually toggle compressed dialogue mode in your engineering workflow by using the specific switching command. This auto-trigger function activates when token efficiency is requested, immediately enforcing terse and precise technical responses.

When should I not use ultra-compressed communication for technical chats?

You should not use ultra-compressed communication for technical chats when pleasantries, conversational context, or detailed explanations are required. The mode enforces strict brevity by restricting filler and articles, which may hinder non-technical readability.