What problem does it solve? Long, filler-heavy AI responses waste tokens and slow down reading when you only need the technical substance. This Skill cuts output tokens by roughly 65% by stripping articles, filler, and pleasantries while keeping code, errors, and technical terms exact. ## Core Features & Use Cases - Six Intensity Levels: Choose lite, full (default), ultra, wenyan-lite, wenyan-full, or wenyan-ultra, including classical Chinese compression modes. - Persistent Mode: Stays active across every response in the session until you say "stop caveman" or "normal mode". - Auto-Clarity Fallback: Automatically reverts to normal prose for security warnings, irreversible action confirmations, and ambiguous multi-step sequences. - Use Case: During a long debugging session, enable full mode so every diagnosis arrives as terse fragments like "Bug in auth middleware. Token expiry check use < not <=" instead of paragraph-length explanations. ## Quick Start Tell the AI "use caveman mode" or invoke /caveman to start receiving ultra-compressed responses immediately.