Tokenade
Local proxy that compacts what a coding agent sends to the model
Tokenade is a CLI that sits between a coding agent and the model provider, rewriting requests to cut token usage before they are sent. It ships command-specific compactors and an MCP proxy wrapper, so it works with agents that speak MCP without changes to the agent itself.
It runs locally rather than as a hosted service, so prompts do not leave the machine on their way through it.
Pricing is by tokens saved rather than per call: a free tier up to 10M saved per month, then Pro at $24.90/mo up to 100M, with $0.30 per million beyond that (verified August 2026). Distributed on npm as @tokenade/cli. Early-stage and closed source despite the public repository, which holds the distribution wrapper rather than the implementation.
Pricing: Monthly subscriptions
Tokenade Alternatives
Explore 35 products in the Frameworks & Stacks category. View all Tokenade alternatives.
Anycloud
CLI and Python SDK for running AI jobs, services and VMs across your own AWS, Azure, GCP, Lambda and Vast accounts
vLLM
High-throughput LLM inference engine with PagedAttention for efficient GPU memory usage
Jan
Open-source desktop app for running LLMs locally with a clean GUI
Ollama
Run large language models locally with a single command
Work on Tokenade? Feature it at the top of Frameworks & Stacks.
Is your product missing?