Skip to content
| Marketplace
Sign in
Visual Studio Code>Programming Languages>TokenSculpt — AI Token & Cost Optimizer (Copilot & Claude)New to Visual Studio Code? Get it now.
TokenSculpt — AI Token & Cost Optimizer (Copilot & Claude)

TokenSculpt — AI Token & Cost Optimizer (Copilot & Claude)

Sarfudheen JAINULABUDHEEN

| (0) | Free
Save 60–90% AI token & LLM costs across GitHub Copilot, Claude Code, and Codex. Zero-latency context pruning, AST skeletons, semantic cache, and real-time savings tracking.
Installation
Launch VS Code Quick Open (Ctrl+P), paste the following command, and press enter.
Copied to clipboard
More Info

TokenSculpt — AI Token & Cost Optimizer (Copilot & Claude)

TokenSculpt Icon

Universal AI Token & Cost Optimization Platform for Developers, Teams, and Enterprises.
Compatible with Google Antigravity IDE, GitHub Copilot, Cursor, Windsurf, Claude Code, and Codex.

Full Installation Guide • 20 Optimization Features • Status Bar Hub • Savings Dashboard • MCP Integration • Commands • Configuration


🌟 Why TokenSculpt?

AI coding assistants (Copilot, Claude, Gemini, GPT-4o) frequently consume tens of thousands of redundant tokens by re-reading build bundles, lockfiles, verbose test logs, repetitive system instructions, and conversational filler.

TokenSculpt acts as a zero-latency, 100% on-device efficiency layer that enforces intelligent caching, AST structural pruning, reversible context compression, context noise exclusion, and unified diff editing across all your favorite IDEs.

🏆 Key Benchmarked Gains:

  • ~75–95% context reduction during file navigation using AST Signatures (skeleton_view).
  • ~92% output token savings by enforcing unified diff patches instead of full 500-line file rewrites.
  • 60–95% context reduction on massive JSON dumps and tool logs using Headroom Reversible CCR (headroom_compress).
  • 100% free answer reuse (<2ms latency) via local on-disk semantic caching (.aicache/).
  • ~500,000+ tokens blocked per workspace scan by automatically excluding dist/, lockfiles, and minified bundles.
  • 30–50% prompt compression with the built-in Adaptive Context Pruner (prune_context).
  • Persistent Session & Lifetime Savings Tracking across IDE restarts with status bar metrics and session archival.

⚡ The 20 Optimization Features

TokenSculpt comes equipped with 20 modular optimizations, active right out of the box:

# Category Optimization Feature How It Works Measured Gain
1 Code Search CodeGraph Pre-Indexing Uses AST symbol graphs to find files instead of wide grep scans ~97% savings vs multi-file grep
2 Terminal CLI Output Compression Strips build spinners, git noise, and verbose logs via RTK/inline filters 60–90% terminal log reduction
3 Prompt Filter Concise AI Responses Strips conversational filler, apologies, and greetings from AI responses ~35% output token reduction
4 Session Context Compaction Prunes stale multi-turn history and drops redundant tool output Prevents context window bloat
5 Disk Cache Semantic Cache Local semantic cache serving repeat answers instantly at $0.00 cost 100% savings on repeated queries
6 AST Parser AST Skeletons Extracts types, classes, and function signatures without full bodies 75–95% file inspection savings
7 Exclusions Smart Context Exclusions Auto-excludes dist/**, package-lock.json, and minified assets ~500k tokens blocked per scan
8 Patch Editing Diff-Only Output Outputs targeted diff hunks (±3 lines) rather than rewriting full files ~92% reduction on code changes
9 Safety Loop Guardrails Halts runaway retry loops after 3 consecutive autonomous failures Prevents runaway credit burn
10 Routing Smart Model Routing Routes routine tasks (formatting, typos, minor edits) to lighter models Up to 99% cost reduction
11 Git Scope Git Diff Scoping Scopes code reviews and unit test generation strictly to git diff lines ~85% context reduction
12 Cloud Cache Prompt Prefix Caching Normalizes prompt prefixes and sinks volatile tokens to suffix (align_prefix_cache) Up to 90% cloud discounts
13 Minifier License Header Stripper Strips copyright & license preambles (safe mode); preserves code comments 15–30% context reduction
14 Test Runner Test Failure Isolator Captures failing assertions and stack traces while stripping passing suites ~95% test log compression
15 Range Slicer Windowed Range Slicing Restricts large file reads to 100-line windows around symbol targets ~80% reduction on file lookups
16 Editor Scope Inline Chat Scope Lock Constrains inline editor chat strictly to selected lines and dependencies ~85% prompt reduction
17 Rules .copilotignore Generator Maintains project-level .copilotignore rules to block build noise Blocks secrets and artifacts
18 Session Cache Edit Session Awareness Treats files open in multi-file edit sessions as already loaded Avoids redundant tool re-reads
19 Monitor Context Saturation Monitor Proactively suggests a fresh chat thread when conversation exceeds 40 turns Prevents quality degradation
20 CCR Engine Headroom Reversible CCR Reversible context compression & SmartCrusher for massive JSON/logs with local retrieval 60–95% context reduction

[!TIP] Directive Compaction (Prompt ROI Optimization): While all 20 strategies remain individually configurable in TokenSculpt's engine, generated instruction files (AGENTS.md, CLAUDE.md, .github/copilot-instructions.md) output a streamlined, consolidated directive set (~35% smaller prompt footprint). Native editor behaviors (inline chat scoping, edit-session awareness, context saturation) are delegated to the host IDE, and complementary directives (range slicing into AST skeletons, .copilotignore into context exclusions) are merged to maximize cloud KV-cache efficiency and eliminate prompt overhead.


🚀 Step-by-Step Installation Guide

For detailed step-by-step setup, companion tool installation, and team configuration, see the Full Installation Guide.

📥 Download VSIX

Download the latest .vsix binary from GitHub Releases (under Assets):

👉 Download Latest VSIX Release

(Or download specific versions from All Releases.)

⚡ Quick Install (VSIX)

Once downloaded, install the .vsix package into your editor:

1. Google Antigravity IDE

  1. Open Antigravity IDE.
  2. Press Ctrl+Shift+P (or Cmd+Shift+P on macOS) → Extensions: Install from VSIX....
  3. Select tokensculpt-1.0.17.vsix.
  4. Reload the window (Ctrl+Shift+P → Developer: Reload Window).

Optional Direct Sync:

New-Item -ItemType Directory -Force -Path "$HOME\.antigravity-ide\extensions\tokensculpt"
Copy-Item -Path ".\dist\*" -Destination "$HOME\.antigravity-ide\extensions\tokensculpt\dist\" -Recurse -Force
Copy-Item -Path ".\package.json" -Destination "$HOME\.antigravity-ide\extensions\tokensculpt\package.json" -Force

2. Visual Studio Code & VS Code Insiders

code --install-extension tokensculpt-1.0.17.vsix

Or via Extensions view (Ctrl+Shift+X) → ... menu → Install from VSIX....

3. Cursor & Windsurf

cursor --install-extension tokensculpt-1.0.17.vsix
# or
windsurf --install-extension tokensculpt-1.0.17.vsix

4. Claude Code CLI

TokenSculpt automatically configures Claude Code's MCP settings in ~/.claude.json. Alternatively, point directly to the built-in MCP server:

{
  "mcpServers": {
    "token-cache": {
      "command": "node",
      "args": ["/path/to/tokensculpt/dist/cache-server.js", "/path/to/workspace"]
    }
  }
}

🛡️ Status Bar Master Hub & Lifetime Tracking

TokenSculpt features a consolidated status bar item displaying live savings metrics:

$(shield) TS: Full · 520k ↓ · $7.80

Highlights:

  • Dual Tracking: Tracks both Active Session Savings and Persistent Lifetime Savings stored safely on disk in .aicache/lifetime-savings.json.
  • 1-Click Master Hub: Clicking the status bar badge opens the TokenSculpt Hub QuickPick with instant access to:
    • Switching optimization profiles (full, debug, planning, review, custom)
    • 1-click toggling of any individual strategy (TokenSculpt: Toggle Feature)
    • Viewing live session breakdown & starting a fresh session (TokenSculpt: Start New Session)
    • Opening the full visual Savings Dashboard
    • Completely deactivating or reactivating TokenSculpt
    • Resetting all historical tracking data

🔌 Multi-Editor MCP Auto-Configuration

TokenSculpt automatically configures and maintains Model Context Protocol (MCP) integrations across your environments on activation:

  • VS Code Copilot: Configures mcp.servers and github.copilot.chat.mcp.servers in .vscode/settings.json.
  • Antigravity IDE: Configures .agents/mcp_config.json.
  • Claude Code: Configures ~/.claude.json.

Configured On-Device Servers (100% Local):

  1. token-cache (Built-in): On-device semantic cache (cache_lookup, cache_store), AST signature extraction (skeleton_view), adaptive context pruning (prune_context), prompt KV-cache alignment (align_prefix_cache).
  2. headroom (Reversible CCR): Local reversible context compression and SmartCrusher (headroom_compress, headroom_retrieve) with air-gapped zero-telemetry enforcement.
  3. codegraph (Semantic Graph): Local AST symbol graph query integration (codegraph_explore).

[!NOTE] All configured MCP servers execute 100% locally via stdio on your machine. TokenSculpt makes zero external network calls and sends no code or queries to third-party cloud services.

Run Ctrl+Shift+P → TokenSculpt: Configure MCP Servers to re-run auto-discovery and configuration at any time.


🔴 Real-Time Savings Dashboard & Activity Log

TokenSculpt features a modern Glassmorphism Savings Dashboard that tracks your token savings and estimated spend reduction in real time:

  • Tokens Saved Counter: Live session and lifetime savings aggregated across AST skeletons, semantic caching, exclusions, and diff modifications.
  • Estimated Spend Reduction ($): Calculated using active model rates (Gemini Flash, Haiku, GPT-4o, Claude Sonnet/Opus).
  • Live Activity Log: Every single optimization event tracked with timestamps, target file, exact tokens saved, and mechanism.
  • Visual Donut & Progress Metrics: Clean visual health bars and efficiency gauges for all 20 optimization features.
  • Context Exclusion Manager: Inspect and toggle ignored directories, lockfiles, and custom patterns directly from the UI.

How to Access:

  • Press Ctrl+Shift+P → TokenSculpt: Open Savings Dashboard
  • Or click the $(shield) TS: badge in the bottom status bar!

⌨️ Commands & Keybindings

All commands are registered under tokensculpt.* with backward compatibility aliases for tokenshield.*:

Command Title Command ID Description
TokenSculpt: Hub tokensculpt.hub Open Master Hub QuickPick for fast profile switching, feature toggles, and status
TokenSculpt: Open Savings Dashboard tokensculpt.dashboard Open visual savings dashboard with live activity log and metric donuts
TokenSculpt: Toggle Feature (1-Click Switch) tokensculpt.toggleFeature Interactive QuickPick to instantly toggle any of the 20 optimization strategies
TokenSculpt: Switch Profile tokensculpt.switchProfile Switch between full, debug, planning, review, and custom
TokenSculpt: Toggle On/Off tokensculpt.toggle Globally pause or resume TokenSculpt optimizations
TokenSculpt: Deactivate Completely tokensculpt.deactivateCompletely Cleanly strip all TokenSculpt directives from .github, CLAUDE.md, AGENTS.md
TokenSculpt: Reactivate All Optimizations tokensculpt.reactivate Re-inject all optimization directives and re-enable active strategies
TokenSculpt: Start New Session tokensculpt.newSession Archive current session savings to lifetime store and reset active counters
TokenSculpt: Reset Complete Data (Wipe All History) tokensculpt.resetAllData Wipe all lifetime history, archived sessions, and event logs from disk
TokenSculpt: Edit Context Exclusions tokensculpt.exclusions Interactive picker for build folder, lockfile, and bundle exclusions
TokenSculpt: Configure MCP Servers tokensculpt.configureMcp Auto-discover and configure MCP servers for VS Code, Claude, and Antigravity
TokenSculpt: Run Health Check tokensculpt.healthCheck Live diagnostic validation across all 20 optimization features and tools
TokenSculpt: Flush Semantic Cache tokensculpt.flushCache Clears local disk cache in .aicache/semantic-cache.json
TokenSculpt: Reindex Code Graph tokensculpt.reindex Triggers a fresh AST symbol graph index in .codegraph/
TokenSculpt: Validate Code Graph tokensculpt.validateGraph Verify integrity of .codegraph/ semantic symbol index
TokenSculpt: Manage Graph Projects tokensculpt.manageProjects Manage multi-root CodeGraph project indices
TokenSculpt: Export Savings Report tokensculpt.exportReport Exports clean savings reports in Markdown, JSON, or CSV format
TokenSculpt: Regenerate Directives tokensculpt.regenerate Updates AI instruction files (AGENTS.md, .github/, CLAUDE.md, .codex/)
TokenSculpt: Export Directives to Repo tokensculpt.exportToRepo Promote local .vscode/ instructions to repository-tracked .github/ files
TokenSculpt: Initialize Project tokensculpt.init Analyzes workspace stack and generates customized instruction files
TokenSculpt: Setup CLI Tools tokensculpt.setupTools Verifies and configures CLI acceleration tools (RTK, CodeGraph, Headroom)
TokenSculpt: Prune & Copy to Clipboard tokensculpt.pruneAndCopy Compresses active file or selection (30–50% savings) to clipboard
TokenSculpt: Compress Git Diff tokensculpt.compressDiff Compresses current git diff for token-efficient PR reviews

⚙️ Configuration Reference (settings.json)

Configure TokenSculpt via .vscode/settings.json or user settings under the tokensculpt namespace (with automatic fallback to tokenshield):

{
  "tokensculpt.enabled": true,
  "tokensculpt.profile": "full",
  "tokensculpt.autoApply": true,
  "tokensculpt.useVscodeStorage": true,
  "tokensculpt.githubInstructions": true,
  "tokensculpt.githubStructureMode": "auto",
  "tokensculpt.generateAgentFiles": false,
  "tokensculpt.configureMcpOnActivation": true,
  "tokensculpt.telemetry.enabled": true,
  "tokensculpt.verbosityLevel": "full",
  "tokensculpt.commentStrippingMode": "headers-only",
  "tokensculpt.activeStrategies": {
    "codeGraph": true,
    "outputCompression": true,
    "verbosityControl": true,
    "sessionManagement": true,
    "semanticCache": true,
    "astSkeleton": true,
    "contextExclusion": true,
    "diffOnlyOutput": true,
    "agentGuardrails": true,
    "smartModelRouting": true,
    "gitDiffContext": true,
    "kvCacheAlignment": true,
    "commentStripper": true,
    "testFailureIsolator": true,
    "rangeSlicing": true,
    "inlineChatScopePinning": true,
    "copilotIgnoreGeneration": true,
    "copilotEditsAwareness": true,
    "threadResetTrigger": true,
    "headroomCompression": true
  },
  "tokensculpt.guardrails": {
    "maxRetries": 3,
    "maxFilesPerTask": 10,
    "maxFileReads": 2
  },
  "tokensculpt.pricing": {
    "flagship": { "inputPerMillion": 15.0, "outputPerMillion": 75.0 },
    "standard": { "inputPerMillion": 3.0, "outputPerMillion": 15.0 },
    "lightweight": { "inputPerMillion": 0.15, "outputPerMillion": 0.60 }
  }
}

🔒 100% Security & Privacy Guarantee

  • Zero Cloud Leakage: 100% on-device local execution. No source code or telemetry ever leaves your machine.
  • Offline & Air-Gapped Friendly: Fully functional without internet connectivity.
  • Zero Repo Clutter: Local configuration stored in .vscode/ by default — never touches git tracking unless explicitly exported.

📄 License

MIT License. Built for developers and modern software engineering teams.

  • Contact us
  • Jobs
  • Privacy
  • Manage cookies
  • Terms of use
  • Trademarks
  • Your Privacy Choices
  • Consumer Health Privacy
© 2026 Microsoft