AI News Today | Julian Goldie Podcast

Cut 65–69% of Tokens with Claude Code (and Any AI Agent) Using Caveman

The script explains a free, open-source “Caveman” skill that reduces output tokens for Claude Code and other agents (Codex, Gemini, Cursor, Windsurf, Cline, Copilot, and more) by making responses short, blunt, and direct while keeping code, commands, file paths, and error messages unchanged. Installed via a one-line command, it drops a rules text file into each agent’s skills/plugin folder so the agent reads it at the start of every chat. Tests on Fable 5 using five real prompts showed about 69% fewer output tokens (e.g., 1,349 down to 324) and about 37% lower total cost while keeping answers correct; it can be toggled with /caveman and set to light/full/ultra, plus optional tools like Commit/Review/Compress. The video also promotes using Caveman across an agent operating system for compounding savings and mentions the AI Profit Ballroom community and training.

00:00 Cut Tokens With Caveman
01:16 Why Token Costs Hurt
01:58 How Caveman Works
02:21 Rules And Safety
03:06 Does Shorter Mean Worse
04:00 Install Across Agents
05:03 Real Token Savings Tests
07:29 Old Replies Vs New
08:49 Modes And Commands
09:47 Agent OS And Extras
10:30 Join The Community
11:29 Wrap Up

Creators and Guests

Host
Julian Goldie
Founder of AI Profit Boardroom and your daily guide to the AI revolution. I break down the biggest AI news, agent updates, and breakthroughs — fast, clear, and no hype.

What is AI News Today | Julian Goldie Podcast?

Latest Podcast