What problem does it solve? Long, filler-heavy AI responses waste output tokens and slow down reading. This Skill compresses every model response into terse caveman-style prose, cutting output tokens by roughly 65% while keeping all technical details, code, and error strings exact. ## Core Features & Use Cases - Six intensity levels: lite, full (default), ultra, plus three Classical Chinese (wenyan) variants for character-level compression. - Persistent session mode: once activated, compression applies to every response until explicitly stopped with "stop caveman" or "normal mode". - Auto-clarity fallback: automatically reverts to normal prose for security warnings, irreversible-action confirmations, and ambiguous multi-step sequences. - Use Case: A developer working through a long debugging session invokes /caveman to get dense, fragment-style answers that keep code blocks and error messages verbatim while dropping all filler. ## Quick Start Ask the assistant to talk like caveman or invoke /caveman to enable compressed responses for the rest of the session.