









Caveman is an AI efficiency operating stack that reduces the cost and context size of agent-native development. It provides an MIT-licensed skill, a free local proxy, an Agent SDK, and planned cloud and enterprise services. Caveman compresses logs, JSON, code, diffs, tables, tool outputs, and files before provider calls while preserving recoverable original bytes. It also supports caching, eval-gated model routing, token usage visibility, savings estimates, and verification workflows. The local wrap works with users' own provider keys and does not require a Caveman account.
Install the free Claude Code skill with the provided shell command, or install Caveman Code and Cavemem with npm. For local proxy optimization, run an existing agent through the Caveman wrapper, such as `caveman claude`. Users can optionally create a free account for cloud synchronization and dashboard access, or join the waitlist for hosted gateway and team features.

Free
$0
One-seat local wrap, MIT skill and extension, local inferred savings, optional free account and cloud sync, and token-count telemetry only.
Indie
$29 per month
One seat with the local wrap and hosted gateway, synced savings dashboard, 50 million optimized tokens per week, and no gainshare.
Team
$349 per month
10 seats included, additional seats at $29 each, eval-gated rollout, automatic rollback, receipt export, Ed25519 verification, projects, and $0.75 per million tokens beyond the plan.
Enterprise
Custom
Planned platform floor plus gainshare on verified savings, with SSO, RBAC, audit logs, on-premise or BYOC deployment, OEM embedding, and planned provider-invoice reconciliation.



29.27%
Social Listening