🪨 why use many token when few token do trick — Claude Code skill that cuts 65% of tokens by talking like caveman
Local proxy and assistant for coding agents
CliGate combines a resident assistant with a local model proxy. The assistant handles tasks, memory, tools, MCP, skills, and channel workflows, while the proxy routes Claude Code, Codex CLI, Gemini CLI, OpenClaw, and API clients through one localhost endpoint.
Builders who want one local place for agent tasks, model routing, credentials, and logs.
You can route multiple agent tools through one local control plane and let a private assistant keep working in the background.
What it does
Resident assistant
Runs a private assistant with task records, memory, approvals, follow-ups, and resumable work.
Model proxy
Provides local Anthropic, OpenAI, Codex, and Gemini-compatible endpoints for agent tools and API clients.
Account and key routing
Manages account pools, API key pools, routing priority, model mapping, and app-level bindings.
Tool and channel execution
Uses skills, MCP, shell and file tools, scheduled tasks, desktop automation, and channels like Telegram, Feishu, and DingTalk.
Observability
Shows usage, pricing, request logs, live log streaming, and API exploration in the dashboard.
Local setup helpers
Includes one-click config paths for Claude Code, Codex CLI, Gemini CLI, and OpenClaw.
How to get it
- 1Run
npx cligate@latest start
- 2Or install globally
npm install -g cligate cligate start
- 3Claude Code
export ANTHROPIC_BASE_URL=http://localhost:8081 export ANTHROPIC_API_KEY=any-key claude
README
CliGate

CliGate is a local AI control plane built around two core capabilities:
- Assistant: a resident private assistant that stays available in the background, understands user tasks, remembers context, uses tools, schedules work, operates channels, and can execute actions through MCP, skills, desktop automation, shell/file tools, or delegated runtimes.
- Model Proxy: one local API and model-routing layer for Claude Code, Codex CLI, Gemini CLI, OpenClaw, and API-compatible clients, with unified accounts, API keys, model mapping, logs, usage, and cost visibility.
It keeps both layers local-first on localhost: the assistant acts like a personal operator for real tasks, while the proxy owns provider access, routing, credentials, unified model names, and observability.
Why CliGate
- One local dashboard for a resident private assistant and unified model routing
- Assistant tasks with memory, approvals, follow-ups, scheduled work, channels, and tool execution
- Account pools, API keys, local runtimes, app routing, and model mapping in one proxy layer
- Channels for Telegram, Feishu, and DingTalk workflows
- Local-first deployment without a hosted relay
What It Includes
Assistant
- Dashboard chat and Assistant Tasks for personal task execution
- A persistent assistant agent with task records, memory, policies, approvals, and resumable executions
- Tool execution through skills, MCP, shell/file tools, scheduled tasks, channels, and optional desktop automation
- Optional delegation to Codex / Claude Code runtime sessions when a task needs an external coding agent
- Telegram, Feishu, and DingTalk channel workflows
Model Proxy
- Anthropic Messages, OpenAI Chat Completions, OpenAI Responses, Codex, and Gemini-compatible endpoints
- One-click configuration for Claude Code, Codex CLI, Gemini CLI, and OpenClaw
- ChatGPT, Claude, and Antigravity account pools
- API key pools for OpenAI, Azure OpenAI, Anthropic, Gemini, Vertex AI, MiniMax, Moonshot, ZhipuAI, DeepSeek, Qwen, and OpenRouter
- Routing priority, app-level bindings, provider model mapping, and free-model routing
- Optional local model routing through Ollama-style runtimes
Observability and operations
- Usage and pricing views
- Request logs and live log streaming
- API explorer
- Tool installer and CLI config helpers
- Resources catalog for free/trial model providers
Quick Start
1. Start CliGate
npx cligate@latest start
Or install globally:
npm install -g cligate
cligate start
Or use a desktop release package:
- Download the installer or app package for your platform from Releases.
- Install or open the package, then launch
CliGate. - CliGate will start the local service and open the desktop window automatically.
Default dashboard:
http://localhost:8081
2. Add at least one working credential
Use the dashboard:
Accountsfor ChatGPT / Claude / AntigravityAPI Keysfor provider keysLocal Modelsfor on-device runtimes
3. Choose your first path
For Assistant use, open Chat or Assistant Tasks and tell the private assistant what you want done.
For Model Proxy use, point a CLI tool or API-compatible client to CliGate.
Claude Code:
export ANTHROPIC_BASE_URL=http://localhost:8081
export ANTHROPIC_API_KEY=any-key
claude
Codex CLI:
# ~/.codex/config.toml
chatgpt_base_url = "http://localhost:8081/backend-api/"
openai_base_url = "http://localhost:8081"
Gemini CLI and OpenClaw can also be configured from the dashboard.
User Paths
Assistant users
Use Chat, Assistant Tasks, Conversation Records, Scheduled, Skills, MCP, and channels to ask the resident assistant to execute real tasks, remember context, use tools, send follow-ups, and keep working in the background.
Model Proxy users
Start the service, add one credential, run one-click config, and send your first proxied request from Claude Code, Codex CLI, Gemini CLI, OpenClaw, or an API-compatible client.
Dashboard operators
Use the dashboard to manage accounts, API keys, routing priority, model mapping, local runtimes, pricing, request logs, usage, channel settings, skills, MCP, and desktop-agent settings.
Documentation
Start here if you want the shortest path to the right document:
- Documentation Hub
- Product Manual (English)
- Product Manual (Chinese)
- Architecture
- API Reference
- App Routing
- Accounts
- OpenClaw Integration
- Screenshot Guide
- Release Guide
- Community
- Contributing
- Security
- Support
- Changelog
After the server starts, a lightweight product guide is also available at:
http://localhost:8081/manual/http://localhost:8081/resources/
Local Architecture
Assistant Surfaces
Web Chat / Assistant Tasks / Telegram / Feishu / DingTalk / Scheduled Tasks
|
v
Private Assistant and Tools
Memory / Policies / Skills / MCP / Desktop Agent / Shell + File Tools / Optional Codex + Claude Code Delegation
|
v
CliGate Local Control Plane (localhost:8081)
|
+--> Model Proxy
| - Protocol translation
| - Account and API key routing
| - App-level bindings and model mapping
| - Local model routing
|
v
Upstream Providers and Local Runtimes
OpenAI / Anthropic / Gemini / Vertex AI / Kilo / Ollama / others
API Surface
| Endpoint | Use |
|---|---|
POST /v1/messages | Anthropic Messages proxy |
POST /v1/chat/completions | OpenAI Chat Completions proxy |
POST /v1/responses | OpenAI Responses proxy |
POST /backend-api/codex/responses | Codex internal compatibility |
POST /v1beta/models/* | Gemini CLI proxy |
GET /api/agent-runtimes/providers | Runtime provider catalog |
GET /api/agent-channels/conversations | Channel conversation records |
GET /api/assistant/tasks | Assistant task records |
GET /api/assistant/mcp/servers | MCP server management |
GET /api/assistant/skills | Assistant skill management |
GET /api/desktop-agent/status | Desktop-agent status |
GET /api/local-runtimes | Local runtime status |
GET /api/resources | Resource catalog |
GET /health | Health and version |
See docs/API.md for more detail.
Community
If you plan to contribute, read CONTRIBUTING.md before opening a pull request.
License
This project is licensed under AGPL-3.0.
Disclaimer
CliGate is an independent open-source project and is not affiliated with Anthropic, OpenAI, Google, or other upstream providers.
Files in the repo
- .claude
- .github
- .idea
- bin
- design
- docs
- images
- native
- public
- scripts
- src
- tests
- .gitignore
- .npmignore
- CHANGELOG.md
- CONTRIBUTING.md
- electron-main.cjs
- electron-mascot-preload.cjs
- LICENSE
- package-lock.json
- package.json
- README_CN.md
- README.md
- SECURITY.md
- SUPPORT.md
Discussion (0)
Ask about usage, or say what you built with itSign in to join the discussion.
No comments yet. Be the first to say what this is good for.
More tools
The best-benchmarked open-source AI memory system. And it's free.
Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription. Available on desktop, mobile and remote runtime.

A cross-platform desktop All-in-One assistant for Claude Code, Codex, OpenCode, OpenClaw, Grok Build & Hermes Agent. Only official website: ccswitch.io
Never stop coding. Free MIT AI gateway: one endpoint, 352 providers (150+ free), 1200+ models Kimi, Claude, GPT, Gemini, GLM, DeepSeek, MiniMax. Works with Claude Code, Codex, Cursor, OpenCode, Cline & Copilot. Quota-aware auto-fallback, RTK+Caveman compression saves 15-95% tokens, MCP/A2A, Desktop/PWA. Built by 550+ contributors
Compress tool outputs, logs, files, and RAG chunks before they reach the LLM. 20% fewer tokens for coding agents, 60-95% fewer tokens for JSON, same answers. Library, proxy, MCP server.