🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman
- Updated
Jul 3, 2026 - JavaScript
🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman
Cross-agent persistent memory for coding assistants. Stored compressed. Retrieved fast. Local by default.
genshijin 原始人 🗿| Claude Code / Codex等AIエージェント 向け超圧縮コミュニケーションスキル。caveman の日本語版をベースに、日本語特有の冗長表現に最適化。
LoRA fine-tune Gemma 4 31B to speak caveman-mode natively. Style: github.com/JuliusBrussee/caveman
We have caveman system prompts and skills for AI models to reduce token use, why not try bake it into the model itself with fine-tuning?
OpenCode package for Caveman: terse AI responses, slash commands, compact reviews, commit messages, and markdown memory compression.
Run many Codex & Claude agents in parallel without them overwriting each other. Isolated worktrees, file locks, PR-only merges. Auto-wires Oh My Codex, Oh My Claude, OpenSpec, and Caveman in every worktree.
Multi-agent skill for faster workflows
Caveman output style for Claude Code: 40% fewer output tokens, always-on formatting
🔥 Save up to 96% tokens — more than Caveman (65%) or RTK (80%). AI coding agent: chat, map, edit, multi-agent. Single Rust binary.
Multi-agent orchestration plugin for OpenCode. 10 configurable model slots, 30 hooks, memory search, LSP diagnostics, agent-browser integration, completion controller, auto dream/distill memory consolidation. One prompt to set up and run.
websocket-driver caveman chat example.
Caveman prompting, measured. A two-channel evaluation protocol scoring what input and output compression actually cost LLMs in dollars, accuracy, and surface-text fidelity across seven models and five benchmarks.
Plugin для Claude Code що вмикає печерний режим — видаляє воду, зберігає суть. −65–90% токенів без втрати точності. Побудований для українських розробників, розуміє обидві мови.
One command turns any repo into a governed, AI-driven delivery pipeline for Claude Code: Beads task memory, code graph + semantic RAG, agents, Ollama, team sync, token savings.
Universal AI skill router — auto-detects best installed skill per prompt + activates caveman mode for ~75% token reduction. Works with Claude, Codex, Cursor, OpenCode, Gemini CLI.