v1.4.0 · MIT
Context-as-image compression proxy for LLMs: renders bulky context (system prompt, tool docs, history) as dense PNG pages with exact per-provider billing math (Anthropic/OpenAI/Gemini). Node and Cloudflare Workers. Part of the OmniRoute family.
v2.7.2 · Apache-2.0
Content-aware output compression for AI coding assistants. 36 specialized processors cut CLI output tokens by 60-99% (git, pytest, npm, terraform, kubectl, docker, and more) without losing errors, diffs, or stack traces.
v5.7.0 · MIT
Measure token savings per AI coding agent, optimize context, and share a live local knowledge graph across 16 CLI clients.
v4.9.2 · MIT
Claude Code plugin that tracks token usage, identifies wasted context, and saves 30-50% on API costs. Heatmaps, ROI reports, budget alerts, efficiency scores, git-aware suggestions — all local, zero config.
v0.2.3 · MIT
Claude Code usage governor: compact professional output, context slimming, tool-output filtering, telemetry, and drift guardrails.
v0.8.2 · MIT
Context compression plugin for Claude Code. Automatically trims large tool output—JSON, YAML, stack traces, and logs—before it enters the context window.
— · MIT
Cut context bloat in your AI-agent stack: find and safely prune unused skills, MCP servers and subagents from real transcript evidence
v1.2.0 · MIT
Pith is the hook that makes Claude Code sessions last 3x longer.
v2.4.1 · Apache-2.0
45% cost reduction measured. The only Claude Code plugin built from CC source analysis — cache expiry prevention, SubTask auto-delegation, zero-cost context restoration, real-time dashboard. Max Plan + API pay-per-use.
v0.5.4 · MIT
Graph memory for AI agents — decisions, context, and session history that survive across every conversation. Works with any LLM.
— · Apache-2.0
ANOLISA (Agentic Nexus Operating Layer & Interface System Architecture) | Agentic OS with runtime, security, observability, and Tokenless response compression for lower token usage and cost.
— · MIT
Context engineering for AI agents. ~80% fewer tokens. Fix tool overload. Skills and memory with in-process BM25 and semantic retrieval. Progressive Disclosure. No vector DB.
— · Apache-2.0
Control what your AI can see. LeanCTX (Lean Context) is the context intelligence layer for AI agents — one local Rust binary that decides what they read, remembers what they learn, guards what they touch, and proves what they save. 60–90% fewer tokens as the receipt. 76 MCP tools, 30+ agents, local-first.
v3.0.0 · MIT
MCP server that lets Claude Code delegate heavy-token tasks to DeepSeek, Kimi, GLM, Qwen, Grok, or any OpenAI-compatible model. Claude orchestrates; the delegate does the heavy lifting. Zero dependencies.
— · no license
Token-efficient Claude Code workspace with parallel agents and persistent memory. Research → Plan → Implement → Validate workflow.
v0.52.55 · AGPL-3.0-only
Installer for the Wizard-AI environment: AI CLI tools and Claude Code skills (graphify, llmlingua, flashrank, markitdown and more). Clones the repo and runs the platform setup script.
— · no license
Give your Claude Code Agent Teams a memory. Auto-injects role-specific context into every new teammate — your team never starts blind again.
— · no license
Claude Code plugin that offloads large outputs to filesystem and retrieves when required.