A viral Claude Code skill called 'caveman mode' claimed to cut output tokens by 65% by stripping AI responses down to blunt, grammar-free fragments. JetBrains tested the claim across 86 real coding tasks using the Harbor evaluation framework and SkillsBench benchmarks, finding actual savings of only 8.5%. The gap exists because agentic output is dominated by code, diffs, and tool invocations that the skill leaves untouched — only the narration between tool calls gets compressed, and there isn't much of it. On the upside, the skill showed no measurable degradation in task quality, earning the verdict: 'safe, honest about style, but oversold on savings.'

4m read timeFrom thenewstack.io
Post cover image
Table of contents
“Skill make agent talk like caveman”Caveman-talk savings, put to the testDoes talking like a troglodyte make Claude dumber?
10 Impressions