
SuperCompress
io.github.Supercompressv0.5.38Updated Oct 3, 2026
Compress coding-agent context ~64%. Hosted Neural Keep MCP for Cursor/Claude/Codex.
Overview
Compresses bulky coding-agent context against the current question to cut LLM input tokens, exposing compression, account linking, and usage tools over MCP.
- What it does
- SuperCompress scores large context such as files, logs, tool dumps, and pastes against the current question and drops filler while keeping evidence-critical lines in their original wording. It exposes MCP tools including compress_context, connect_account, and usage_summary, and can also install hooks and MCP configuration across many coding harnesses. An optional localhost proxy can rewrite base URLs for OpenAI- and Anthropic-compatible clients.
- When to use it
- Worth adding when a coding agent regularly sends large files, logs, or tool output and you want fewer input tokens without losing answer-critical content. It suits Cursor, Claude Code, Codex, and similar harnesses that support MCP or hooks.
- Requirements
- Node.js 18+ for the npm package supercompress-proxy, run via npx or installed globally. A SuperCompress account is linked through the setup command or OAuth/device-link; the optional API key is passed as SUPERCOMPRESS_API_KEY. Network access to the SuperCompress API is required, and the optional local proxy listens on localhost:8080.
Installation
In SourceWeft
- Open SuperCompress in the dashboard and add it to a workspace.
- Enable the server for the chats that should use its tools.
Web executable via Streamable HTTP. Remote servers run from the web runtime once configured in a workspace.
Other MCP clients
Add this to your client's mcpServers config.
{
"mcpServers": {
"supercompress": {
"type": "http",
"url": "https://api.supercompress.dev/api/mcp"
}
}
}README
SuperCompress
SuperCompress v2 — cut ~64% of LLM input tokens for coding agents, without losing the answer.
~400M Neural Keep on the SuperCompress API scores bulky context (files, logs, tool dumps, pastes) against the current question and drops the rest. Your ask stays intact. This package wires Cursor / Claude Code / Codex / 60+ harnesses into that API via MCP + hooks.
Website · Benchmarks · Playground · Docs
Install
Requires Node.js 18+.
Quick start (recommended)
One command links your account and auto-adds MCP + hooks across 60+ harnesses (25+ get a native auto MCP plugin — Cursor, Claude Code, Codex, Goose, Zed, OpenCode, fx, Hermes, OpenClaw, Grok, Gemini, Windsurf, Continue, and more):
Then restart your agent so integrations reload. That’s it.
Re-detect later (new agent installed, etc.):
Any other agent (custom harness, closed-source, DIY):
That prints a stdio MCP snippet and writes an Agent Plugins 1.0 pack you can drop into any compatible client.
Benchmarks
Same keep-budget (35% of tokens kept). Who still has the answer?
Full methodology and charts: supercompress.dev/benchmarks
How it works
- You ask a question (never rewritten).
- Large context is scored against that question.
- Evidence-critical lines stay in original wording; filler drops.
- You pay for fewer input tokens.
Commands
Optional localhost API proxy (base-URL rewrite) if you explicitly need it:
Then point OpenAI/Anthropic-compatible clients at http://localhost:8080/v1.
MCP
setup / plugin registers the MCP server on every detected host. You can also run it directly:
Manual registration:
Account & pricing
Launch promo: 5M tokens/month free, then $0.10 / 1M from the dashboard.
Privacy
Hooks / MCP run on your machine. Provider API keys stay with your agent. Context text is sent to the SuperCompress API so the hosted compiler can compress it.
More
- Coding agents: https://docs.supercompress.dev/coding-agents
- HTTP / Python API: https://docs.supercompress.dev/quickstart
- Source: https://github.com/Supercompress/Supercompress
License
MIT — see LICENSE.
Source: packages/proxy/README.md at commit 08a34db
Tools
0Version history
1- v0.5.38LatestOct 3, 2026

