Token Saver 中文版 — 压缩 AI 常驻上下文与 token 成本

    1

    Reduce AI costs and prevent context overflow by auditing resident instructions and throttling runtime token usage.

    $6.99

    Secure checkout via Stripe

    30-day refund guarantee

    Converts to your local currency at checkout

    0 installsSecurity scanned

    Works with the AI tools you already use

    Claude CodeClaude CodeCursorCursorCodex CLICodex CLIGitHub CopilotGitHub CopilotGemini CLIGemini CLI+20 more

    Token Saver 中文版 — 压缩 AI 常驻上下文与 token 成本

    Example session with this skill installed

    My CLAUDE.md is getting huge and I'm hitting context limits. Can you audit it and suggest ways to save tokens?

    • Read your context and instructions
    • Compiled the token saver
    • Generated the document

    I have audited your resident context files using the token_meter.py script. Your current configuration is approaching the limits for the Codex harness.

    Resident Context Audit

    Current Harness: Codex
    Budget: 32,768 bytes

    FileBytesEstimated TokensBudget %Risk
    CLAUDE.md24,150~6,10073.7%High
    AGENTS.md4,200~1,05012.8%Low
    Total28,350~7,15086.5%Warning

    Estimated after optimization: 28,350 → 12,400 bytes.

    Proposed Modifications

    CLAUDE.md (Lines 80-150): Move the detailed React component styling guide to docs/style-guide.md. Replace with a single-line reference.

    • Impact: -8,400 bytes.
    1. CLAUDE.md (Lines 200-230): Remove redundant environment setup commands already covered by the setup MCP tool.
      • Impact: -1,200 bytes.
    2. AGENTS.md: Consolidate multiple project-specific rules into concise imperative statements.
      • Impact: -6,350 bytes.

    Next steps

    • Run python3 scripts/token_meter.py count --path CLAUDE.md to verify these specific segments.
    • Review the suggested diffs; say "Apply CLAUDE.md diff" to proceed with the migration to reference files.
    • Consider starting a new session to clear the accumulated history now that the base context is leaner.

    token-saver-ai-token.pdf

    PDF · document

    Generated

    Example file from a real run - the skill writes it into your workspace.

    Connects securely to your tools. The creator never sees your data.

    What you get

    Reduce per-message token costs by 40-60% through precise file reading.Prevent silent truncation of project rules in Codex and DeepSeek Harness.Clean up bloated instruction files without losing functional requirements.Monitor real-time token usage with precision measurement scripts.

    About this skill

    The problem

    Large project files and bloated instruction documents (CLAUDE.md/AGENTS.md) drain context windows and drive up API costs. Developers often pay for thousands of tokens of redundant history or unread logs that never needed to be in the prompt.

    What it does

    • Enforces "Read-on-Demand" patterns, using grep and line offsets to avoid reading entire large source files.
    • Cuts log bloat by redirecting command outputs to disk and only extracting relevant error traces or tail segments.
    • Audits resident context files (AGENTS.md, CLAUDE.md) against specific harness budgets like Codex (32KB) and DeepSeek (64KB).
    • Generates diff-based compression proposals to move verbose instructions into reference files while keeping active context lean.
    • Calculates real-time token usage via a standalone Python script, using tiktoken for precision or weighted estimation as a fallback.

    Frameworks & tools

    Compatible with Claude Code, Codex, DeepSeek Harness, and WorkBuddy. Uses Python 3.10+ standard library, grep, and optional tiktoken.

    Why this beats prompting it yourself

    Generic "be concise" prompts fail because LLMs naturally want to read more to be sure. This skill enforces technical constraints like shell-level piping, file indexing, and specific byte-limit audits that a standard prompt cannot consistently execute across long sessions.

    Use cases

    • Reducing session costs when debugging long stack traces or build logs.
    • Optimizing CLAUDE.md or AGENTS.md files that have grown past harness budget limits.
    • Detecting session fatigue to signal when to start a fresh conversation.
    • Auditing MCP tool definitions and skill descriptions for hidden context overhead.

    Known limitations

    Does not modify session history files. Resident file optimizations require manual user confirmation of generated diffs. Precision counting requires tiktoken installation.

    How to install

    Works the same in every agent - Claude, Cursor, Codex, Copilot and 20+ more.

    ~30 seconds
    1. 1

      Download the ZIP

      Free skills download straight away. Paid skills unlock right after purchase.

    2. 2

      Unzip into your skills folder

      Every agent reads skills from one folder on your machine. Drop the unzipped folder in there.

    3. 3

      Ask your agent to use it

      Restart the agent if it was already running. It picks the skill up automatically - no config needed.

    Skills folder by agent

    Click the path to copy it. Create the folder if it does not exist yet.

    Reviews

    No reviews yet

    Be one of the first to try it. Every listed skill passes our trust checks below.

    Security scanned

    Passed our 8-point scan before listing

    Fresh listing

    Recently published to Agensi

    30-day refund

    Not a fit? Get your money back

    Trust & safety

    Security scanned

    Verified clean 25 days ago

    • Passed all security checks, Safe to install

    Listed25 days ago

    What's inside

    Frequently Asked Questions