What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading. This Skill cuts token usage by roughly 75% by stripping articles, filler words, and pleasantries while keeping all technical substance intact. ## Core Features & Use Cases - Six Intensity Levels: Choose from lite, full (default), ultra, wenyan-lite, wenyan-full, and wenyan-ultra, including classical Chinese compression modes. - Persistent Mode: Stays active across every response in the session until explicitly disabled with "stop caveman" or "normal mode". - Auto-Clarity Fallback: Automatically reverts to clear language for security warnings, irreversible action confirmations, and multi-step instructions, then resumes terse mode. - Use Case: During a long debugging session, enable caveman mode so every explanation arrives as compact fragments like "Bug in auth middleware. Token expiry check use < not <=", saving context window space for actual code. ## Quick Start Tell the AI to "use caveman mode" or invoke /caveman to start receiving ultra-compressed responses immediately.