Trim your Claude Code prompts locally — your code never sent to us
tokenolo runs a local proxy that strips redundant filler from your prompts before Claude Code sends them - removal-only, cache-safe, your own key. Tool output, files and history pass through untouched.
npm install -g tokenoloInvisible by design.
tokenolo is a middleware layer — not a new model and not a new workflow. It wraps your agent behind a local proxy, trims the filler from each message on its way out, then steps aside.
Run one command
tokenolo wrap claudestarts a local proxy on 127.0.0.1 and launches Claude Code pointed at it.
Work like you always do
Type prompts, run tools, edit files — nothing about your workflow changes.
tokenolo trims the filler inline
On the way out, each message is stripped of redundant filler — removal-only, on your machine, under your own key. Files, tool output and history pass through untouched.
Also works with Codex: tokenolo wrap codex
Before and after.
The filler trim is live today. Context and tool-output pruning are on the roadmap — shown here with example figures, not measured results.
please, could you update the readme when you get a chance — thanks!
update the readme
Removal-only · meaning preserved · the “after” is always a subsequence
$ ls -la node_modules/.bin total 4128 -rwxr-xr-x 1 u staff 384 acorn -rwxr-xr-x 1 u staff 384 acorn-walk … 212 more entries …
# 215 bin entries — relevant to this task: next eslint tsc # full listing pruned
Same next step chosen · 1,402 → 38 tokens · example
# context assembled by the agent system prompt + open files + history 42 files attached · ≈ 1,842 tokens
# pruned to what this task reads 9 relevant files kept · 33 dropped ≈ 612 tokens
Same answer · 1,842 → 612 tokens · example
Simple pricing. Scale on wraps.
Every plan has the same features - they differ only by your daily wrap limit and rate. Pick a plan, raise the limit when you grow.
- 500 wraps / day per seat
- 120 requests / min
- Every feature, no gates
- Unlimited wraps per seat
- 300 requests / min
- Every feature, no gates
Custom wrap + rate limits, SSO / SAML & audit logs, and dedicated support with an SLA. Built for teams running tokenolo at scale.
Questions? Answered.
What is tokenolo?
Which coding agents does it work with?
Do I have to change my workflow?
How much does it save?
What's on the roadmap?
How much does it cost?
How do I get started?
Your next run starts lean.
Drop tokenolo in, keep every command you already run, and watch the token bill fall - same model, comparable results.