Skip to content

Repository files navigation

⚡ ZipAI

Ultra-dense token optimizer for LLM agents — maximize prompt caching, prune input, and compress output.

Version License: MIT Security Policy GitHub stars Last commit Issues PRs Platform: Any Agent

Framework-agnostic markdown rules for Claude Code, Antigravity, Gemini CLI, Cline, Roo Code, Hermes & more.


If you find ZipAI helpful for reducing LLM API token costs, please consider giving it a star on GitHub!


✨ Features

  • 🧹 Zero Filler — strips conversational padding, enforces telegraphic grammar.
  • 💾 Prompt Caching — static-first ordering preserves prefix cache hits (up to 90%).
  • ✂️ Input Pruning — log compression, AST skeletal inspection, JSON/YAML minification.
  • 🔬 Surgical Output — localized diffs only, no full-file reprints.
  • 🧠 Reasoning Budget — adaptive Chain-of-Thought depth based on task complexity.
  • 📐 Schema Enforcement — JSON Schema replaces verbose format instructions.
  • 🗜️ Semantic Compression — lossless context reduction (50–90% savings).
  • 🧭 ARM-style Reasoning — task-aware reasoning mode selection.
  • KV Cache Optimization — structured inter-tool communication, dynamic eviction.
  • 📊 TAAC — domain-specific compression ratios (code vs math vs NL).

📊 Token Savings Benchmarks

Category Before After Reduction
Log Files (traceback + context) ~4,500 ~225 -95%
MCP Payloads (minified JSON) ~1,800 ~180 -90%
Code Inspection (AST + lines) ~6,200 ~480 -92%
Session Context (static-first) 30,000 ~3,000 -90%
Agent Generation (telegraphic) ~450 ~95 -79%

🚀 Quick Start (under 2 min)

Install the rules into your AI agent. Pick the path for your tool:

# Antigravity IDE (Gemini Agent Skill)
mkdir -p ~/.gemini/config/skills/zipai
cp SKILL.md ~/.gemini/config/skills/zipai/SKILL.md

# Claude Code CLI
cp CLAUDE.md /path/to/your/project/CLAUDE.md

Restart your agent — it now applies ZipAI rules automatically. ✅

📦 Installation

Antigravity IDE (Gemini Agent Skill)

mkdir -p ~/.gemini/config/skills/zipai
cp SKILL.md ~/.gemini/config/skills/zipai/SKILL.md

The IDE lists zipai-optimizer as an available skill and applies the rules automatically.

Claude Code CLI

cp CLAUDE.md /path/to/your/project/CLAUDE.md

Claude Code reads this file at startup and applies the rules for the workspace session.

VS Code Agent Extensions (Cline, Roo Code, etc.)

mkdir -p /path/to/your/project/.agents/rules/
cp .agents/rules/zipai.md /path/to/your/project/.agents/rules/zipai.md

Gemini CLI

cp SKILL.md /path/to/your/project/GEMINI.md

Hermes Agent & Custom AI Assistants

These rules are framework-agnostic markdown. Load SKILL.md into any system-instructions template:

  • Hermes Agent: append to the agent's system instructions profile or config.yaml.
  • Custom loops: reference SKILL.md directly in system instructions.

🛠️ Usage

ZipAI is a ruleset, not a runtime — once installed, your agent self-applies the optimizations. Verify behavior by prompting normally and observing denser, cache-friendly output:

# In your agent session
"Summarize this 50-line log and propose a fix."   # expect ~225 tokens, not ~4,500

Tip: CLAUDE.md and .agents/rules/zipai.md are mirrors of SKILL.md. Update SKILL.md first, then propagate.

🏗️ Architecture

ZipAI compresses tokens at every stage of the agent loop:

flowchart LR
  A[Raw Prompt / Output] --> B[Zero Filler]
  B --> C[Prompt Caching]
  C --> D[Input Pruning]
  D --> E[Semantic Compression]
  E --> F[Ultra-Dense Tokens]
  F --> G[LLM Context + Cache]
Loading

📁 File Structure

ZipAI/
├── SKILL.md                    # Canonical rules (v15.0) — single source of truth
├── CLAUDE.md                   # Mirror for Claude Code auto-detection
├── .agents/rules/zipai.md     # Mirror for VS Code agent extensions
├── README.md                   # This file
├── LICENSE                     # MIT License
├── .pre-commit-config.yaml    # Git hooks
├── .gitattributes              # Git merge strategies
└── .github/
    ├── scripts/ai-pr-reviewer.cjs
    └── workflows/
        ├── ai-pr-reviewer.yml
        └── auto-resolve-pr-conflicts.yml

⚠️ Limitations

  • Brainstorming: disable during creative / open-ended design phases.
  • Grep Blindness: key context may fall outside filter boundaries.
  • Overshadowing: aggressive pruning may drop micro-variables in long sessions.
  • Math Fragility: numerical reasoning chains resist aggressive compression.

🌟 Stargazers & Community

If ZipAI helped optimize your agent workflows or slashed your API bill, leave a star ⭐ to help other developers discover it!

Star History Chart


🔒 Security

ZipAI ships markdown rules only — no executable code. See SECURITY.md for the vulnerability-reporting policy.


📜 License

This project is licensed under the MIT License.

About

⚡ Ultra-dense token optimizer for LLM coding agents (Claude Code, Antigravity, Cline, Roo) — prompt caching, AST pruning & minified payloads.

Topics

Resources

Code of conduct

Contributing

Security policy

Stars

1 star

Watchers

0 watching

Forks

Releases

Packages

Contributors

Languages