Best alternative to squeez
Tokenade is the best alternative to squeez — Universal token-optimization engine for AI coding agents — native hooks for 18 agents combine output filtering, semantic code search, skeleton compression, sandboxed execution, MCP proxying and a live savings dashboard in a single dependency-free binary.
Get TokenadeOutput Filtering
Format-aware compactors cover git, cargo, kubectl, terraform, docker and more — 60–99% reduction on the noisiest commands. Command rewriting further trims source-side before the shell even runs.
Not available
Semantic Code Search
Finds the most relevant files for a task and sends only those to the model, instead of the whole repo. Runs fully on-device with no external vector database and no model downloads — fast even on large codebases.
Not available
Skeleton Compression
Signatures-only view of source files, YAML, Markdown and Terraform — −64% on file reads while preserving every top-level declaration. Stacks on top of output filtering for maximum savings.
Not available
Third-Party MCP Optimization
tokenade mcp-proxy wraps any third-party MCP server's launch command in the agent's MCP config, so every tool result (verbose JSON, logs, console output) is folded on the way back — set once, not per call. Image results pass through untouched.
Not available
Mechanism Breadth
The only tool combining output filtering + semantic search + skeleton compression + sandbox execution + MCP proxying + secret redaction + content-addressed cache in a single binary. On the open THOL benchmark it is the only one of the twelve tools tested that measurably cuts costs: 23% cheaper than running no tool at all, and 39% cheaper on long sessions.
Not available
Setup & Installation
npm install -g @tokenade/cli then tokenade install — native hooks auto-detected for 18 agents (Claude Code, Cursor, Codex, Gemini CLI, Copilot, Windsurf and more). Works without an account: 10M tokens offered per machine to try. Not yet on crates.io or Homebrew.
Not available
Savings Dashboard
tokenade dashboard shows measured savings, per-command and per-project breakdown, and framework-detection status. Local logs rotate automatically with built-in secret redaction.
Not available
Pricing
Freemium: works without an account (10M tokens offered per machine to try); a free account raises that to 10M tokens saved per month on unlimited machines. Pro at $24.90/mo excl. tax (19,90 € incl. tax) covers 100M/month, then optional pay-as-you-save at $0.30 excl. tax (0,20 € incl. tax) per extra million saved. Enterprise on quote ([email protected]).
Not available
squeez at a glance
squeez starts at Free (open source). Rust hook-based output compressor: smart filtering, cross-call deduplication, log-template compaction and relevance-aware truncation applied to shell output and file reads before they enter context — with the folded bytes kept in local content-addressed storage so the agent can retrieve the exact original.
Pros
- Compression is reversible: the agent can pull back exact bytes, so folding never destroys information
- Fully local with zero runtime dependencies — no data leaves the machine
- Cross-call deduplication catches output repeated across a whole session, not just within one call
- Supports several agent CLIs from a single install
Cons
- No measurable saving on the open THOL benchmark
- By its own documentation it cannot compress sub-agent return values, top-level user prompts, or skill files loaded at session start
- No semantic search and no repository index — it compresses what passes through, it does not decide what to read
- Support is uneven across host CLIs; some lack a clean injection point for file reads
Ready to cut costs with Tokenade?
Join the teams that already chose Tokenade over squeez.
Get Tokenade