decolua/9routerPublic

Unlimited FREE AI coding. Connect Claude Code, Codex, Cursor, Cline, Copilot, Antigravity to FREE Claude/GPT/Gemini via 40+ providers. Auto-fallback, RTK -40% tokens, never hit limits.

AI summary: An AI coding gateway connecting multiple tools to 40+ providers with token-saving features.

Stars
30.3K
+96 today
Forks
5.8K
Watchers
160
Open issues
1.3K
Open PRs
1.1K
Contributors
~354
Commits
1.4K
Branches
3

JavaScriptMITCreated Jan 5, 2026Last push 3d agoLatest release v0.5.35+451 stars this week+3.2K this month

Quick answers

What is 9router?
An AI coding gateway connecting multiple tools to 40+ providers with token-saving features.
What does 9router do?
9Router acts as an intelligent proxy and load balancer for AI-assisted coding tools like Cursor, Copilot, and Claude Code. It connects these clients to over 40 distinct AI model providers, offering automatic fallback if a primary provider fails. The platform includes token-saving features, reducing token usage by 20% to 40% using Request Token Knockdown (RTK) techniques. It effectively centralizes billing and model management for individual developers and teams. By operating as a local proxy, it mitigates rate limits, automates failovers to cheaper models, and standardizes disparate LLM protocols into a single access layer.
Who is 9router for?
This gateway is for software developers, AI enthusiasts, and engineering teams using multiple AI coding assistants. It is especially useful for those seeking to reduce API costs or avoid rate limits.
How do I get started with 9router?
npm install -g 9router && 9router
How popular is 9router on GitHub?
decolua/9router has 30,281 stars and 5,758 forks on GitHub, and gained 451 stars in the last 7 days.
What license does 9router use?
decolua/9router is released under the MIT license.

Star history

since Jul 29, 2026
010K20K30KJul 2026Aug 2026Sep 2026Oct 2026
30.3K stars as of Oct 4, 2026. Measured daily since Jul 29, 2026; GitHub no longer exposes earlier star timestamps.

Contribution activity

commits per day, last 52 weeks
OctNovDecJanFebMarAprMayJunJulAugSepOctMonWedFri2025-10-11: 0 commits2025-10-12: 0 commits2025-10-13: 0 commits2025-10-14: 0 commits2025-10-15: 0 commits2025-10-16: 0 commits2025-10-17: 0 commits2025-10-18: 0 commits2025-10-19: 0 commits2025-10-20: 0 commits2025-10-21: 0 commits2025-10-22: 0 commits2025-10-23: 0 commits2025-10-24: 0 commits2025-10-25: 0 commits2025-10-26: 0 commits2025-10-27: 0 commits2025-10-28: 0 commits2025-10-29: 0 commits2025-10-30: 0 commits2025-10-31: 0 commits2025-11-01: 0 commits2025-11-02: 0 commits2025-11-03: 0 commits2025-11-04: 0 commits2025-11-05: 0 commits2025-11-06: 0 commits2025-11-07: 0 commits2025-11-09: 0 commits2025-11-10: 0 commits2025-11-11: 0 commits2025-11-12: 0 commits2025-11-13: 0 commits2025-11-14: 0 commits2025-11-15: 0 commits2025-11-16: 0 commits2025-11-17: 0 commits2025-11-18: 0 commits2025-11-19: 0 commits2025-11-20: 0 commits2025-11-21: 0 commits2025-11-22: 0 commits2025-11-23: 0 commits2025-11-24: 0 commits2025-11-25: 0 commits2025-11-26: 0 commits2025-11-27: 0 commits2025-11-28: 0 commits2025-11-29: 0 commits2025-11-30: 0 commits2025-12-01: 0 commits2025-12-02: 0 commits2025-12-03: 0 commits2025-12-04: 0 commits2025-12-05: 0 commits2025-12-06: 0 commits2025-12-07: 0 commits2025-12-08: 0 commits2025-12-09: 0 commits2025-12-10: 0 commits2025-12-11: 0 commits2025-12-12: 0 commits2025-12-13: 0 commits2025-12-14: 0 commits2025-12-15: 0 commits2025-12-16: 0 commits2025-12-17: 0 commits2025-12-18: 0 commits2025-12-19: 0 commits2025-12-20: 0 commits2025-12-21: 0 commits2025-12-22: 0 commits2025-12-23: 0 commits2025-12-24: 0 commits2025-12-25: 0 commits2025-12-26: 0 commits2025-12-27: 0 commits2025-12-28: 0 commits2025-12-29: 0 commits2025-12-30: 0 commits2025-12-31: 0 commits2026-01-01: 0 commits2026-01-02: 0 commits2026-01-03: 0 commits2026-01-04: 0 commits2026-01-05: 8 commits2026-01-06: 10 commits2026-01-07: 1 commit2026-01-08: 0 commits2026-01-09: 4 commits2026-01-10: 0 commits2026-01-11: 1 commit2026-01-12: 8 commits2026-01-13: 2 commits2026-01-14: 2 commits2026-01-15: 2 commits2026-01-16: 3 commits2026-01-17: 0 commits2026-01-18: 2 commits2026-01-19: 4 commits2026-01-20: 2 commits2026-01-21: 1 commit2026-01-22: 0 commits2026-01-23: 1 commit2026-01-24: 0 commits2026-01-25: 0 commits2026-01-26: 0 commits2026-01-27: 3 commits2026-01-28: 2 commits2026-01-29: 3 commits2026-01-30: 0 commits2026-01-31: 4 commits2026-02-01: 0 commits2026-02-02: 8 commits2026-02-03: 5 commits2026-02-04: 3 commits2026-02-05: 5 commits2026-02-06: 18 commits2026-02-07: 3 commits2026-02-08: 3 commits2026-02-09: 8 commits2026-02-10: 7 commits2026-02-11: 9 commits2026-02-12: 1 commit2026-02-13: 2 commits2026-02-14: 1 commit2026-02-15: 6 commits2026-02-16: 2 commits2026-02-17: 0 commits2026-02-18: 5 commits2026-02-19: 2 commits2026-02-20: 11 commits2026-02-21: 6 commits2026-02-22: 4 commits2026-02-23: 1 commit2026-02-24: 0 commits2026-02-25: 7 commits2026-02-26: 2 commits2026-02-27: 6 commits2026-02-28: 6 commits2026-03-01: 4 commits2026-03-02: 3 commits2026-03-03: 7 commits2026-03-04: 0 commits2026-03-05: 5 commits2026-03-06: 16 commits2026-03-07: 1 commit2026-03-08: 0 commits2026-03-09: 13 commits2026-03-10: 3 commits2026-03-11: 8 commits2026-03-12: 10 commits2026-03-13: 6 commits2026-03-14: 9 commits2026-03-15: 0 commits2026-03-16: 2 commits2026-03-17: 6 commits2026-03-18: 1 commit2026-03-19: 4 commits2026-03-20: 2 commits2026-03-21: 0 commits2026-03-22: 7 commits2026-03-23: 16 commits2026-03-24: 0 commits2026-03-25: 1 commit2026-03-26: 9 commits2026-03-27: 7 commits2026-03-28: 1 commit2026-03-29: 0 commits2026-03-30: 6 commits2026-03-31: 3 commits2026-04-01: 1 commit2026-04-02: 0 commits2026-04-03: 5 commits2026-04-04: 6 commits2026-04-05: 4 commits2026-04-06: 6 commits2026-04-07: 2 commits2026-04-08: 4 commits2026-04-09: 2 commits2026-04-10: 8 commits2026-04-11: 2 commits2026-04-12: 0 commits2026-04-13: 10 commits2026-04-14: 7 commits2026-04-15: 4 commits2026-04-16: 1 commit2026-04-17: 18 commits2026-04-18: 0 commits2026-04-19: 0 commits2026-04-20: 0 commits2026-04-21: 3 commits2026-04-22: 13 commits2026-04-23: 4 commits2026-04-24: 8 commits2026-04-25: 5 commits2026-04-26: 2 commits2026-04-27: 0 commits2026-04-28: 9 commits2026-04-29: 4 commits2026-04-30: 2 commits2026-05-01: 8 commits2026-05-02: 0 commits2026-05-03: 19 commits2026-05-04: 3 commits2026-05-05: 3 commits2026-05-06: 0 commits2026-05-07: 12 commits2026-05-08: 0 commits2026-05-09: 12 commits2026-05-10: 4 commits2026-05-11: 7 commits2026-05-12: 9 commits2026-05-13: 13 commits2026-05-14: 10 commits2026-05-15: 13 commits2026-05-16: 10 commits2026-05-17: 7 commits2026-05-18: 10 commits2026-05-19: 0 commits2026-05-20: 3 commits2026-05-21: 6 commits2026-05-22: 0 commits2026-05-23: 12 commits2026-05-24: 0 commits2026-05-25: 1 commit2026-05-26: 9 commits2026-05-27: 0 commits2026-05-28: 0 commits2026-05-29: 6 commits2026-05-30: 0 commits2026-05-31: 1 commit2026-06-01: 0 commits2026-06-02: 3 commits2026-06-03: 1 commit2026-06-04: 0 commits2026-06-05: 0 commits2026-06-06: 23 commits2026-06-07: 0 commits2026-06-08: 15 commits2026-06-09: 0 commits2026-06-10: 0 commits2026-06-11: 0 commits2026-06-12: 0 commits2026-06-13: 45 commits2026-06-14: 7 commits2026-06-15: 3 commits2026-06-16: 1 commit2026-06-17: 15 commits2026-06-18: 7 commits2026-06-19: 3 commits2026-06-20: 12 commits2026-06-21: 18 commits2026-06-22: 0 commits2026-06-23: 0 commits2026-06-24: 0 commits2026-06-25: 0 commits2026-06-26: 33 commits2026-06-27: 0 commits2026-06-28: 1 commit2026-06-29: 16 commits2026-06-30: 0 commits2026-07-01: 3 commits2026-07-02: 0 commits2026-07-03: 14 commits2026-07-04: 1 commit2026-07-05: 6 commits2026-07-06: 0 commits2026-07-07: 9 commits2026-07-08: 1 commit2026-07-09: 6 commits2026-07-10: 26 commits2026-07-11: 0 commits2026-07-12: 0 commits2026-07-13: 2 commits2026-07-14: 0 commits2026-07-15: 4 commits2026-07-16: 18 commits2026-07-17: 2 commits2026-07-18: 0 commits2026-07-19: 7 commits2026-07-20: 7 commits2026-07-21: 0 commits2026-07-22: 0 commits2026-07-23: 8 commits2026-07-24: 0 commits2026-07-25: 3 commits2026-07-26: 3 commits2026-07-27: 0 commits2026-07-28: 0 commits2026-07-29: 19 commits2026-07-30: 1 commit2026-07-31: 0 commits2026-08-01: 1 commit2026-08-02: 1 commit2026-08-03: 0 commits2026-08-04: 0 commits2026-08-05: 39 commits2026-08-06: 0 commits2026-08-07: 0 commits2026-08-08: 0 commits2026-08-09: 0 commits2026-08-10: 0 commits2026-08-11: 0 commits2026-08-12: 0 commits2026-08-13: 19 commits2026-08-14: 12 commits2026-08-15: 0 commits2026-08-16: 1 commit2026-08-17: 0 commits2026-08-18: 0 commits2026-08-19: 0 commits2026-08-20: 0 commits2026-08-21: 0 commits2026-08-22: 0 commits2026-08-23: 0 commits2026-08-24: 0 commits2026-08-25: 1 commit2026-08-26: 0 commits2026-08-27: 16 commits2026-08-28: 22 commits2026-08-29: 1 commit2026-08-30: 0 commits2026-08-31: 0 commits2026-09-01: 1 commit2026-09-02: 9 commits2026-09-03: 24 commits2026-09-04: 1 commit2026-09-05: 15 commits2026-09-06: 0 commits2026-09-07: 0 commits2026-09-08: 0 commits2026-09-09: 6 commits2026-09-10: 21 commits2026-09-11: 2 commits2026-09-12: 0 commits2026-09-13: 0 commits2026-09-14: 0 commits2026-09-15: 0 commits2026-09-16: 3 commits2026-09-17: 17 commits2026-09-18: 11 commits2026-09-19: 4 commits2026-09-20: 0 commits2026-09-21: 11 commits2026-09-22: 21 commits2026-09-23: 15 commits2026-09-24: 0 commits2026-09-25: 0 commits2026-09-26: 27 commits2026-09-27: 1 commit2026-09-28: 23 commits2026-09-29: 0 commits2026-09-30: 0 commits2026-10-01: 17 commits2026-10-02: 0 commits2026-10-03: 0 commits2026-10-04: 0 commits2026-10-05: 0 commits2026-10-06: 0 commits2026-10-07: 0 commits2026-10-08: 0 commits2026-10-09: 0 commits2026-10-10: 0 commits
1,337 commits in the last yearLessMore

Signals and awards

derived from tracked data
  • Widely adopted

    30,281 stars

  • Very active

    1,337 commits in 52 weeks

  • Community-driven

    ~354 contributors

  • Outside contributions

    89% of recent commits from the community

  • Permissive license

    MIT

  • Continuous integration

    Automated checks passing

  • Repeat trending

    9 trending appearances

What 9router does

9Router acts as an intelligent proxy and load balancer for AI-assisted coding tools like Cursor, Copilot, and Claude Code. It connects these clients to over 40 distinct AI model providers, offering automatic fallback if a primary provider fails. The platform includes token-saving features, reducing token usage by 20% to 40% using Request Token Knockdown (RTK) techniques. It effectively centralizes billing and model management for individual developers and teams. By operating as a local proxy, it mitigates rate limits, automates failovers to cheaper models, and standardizes disparate LLM protocols into a single access layer.

This gateway is for software developers, AI enthusiasts, and engineering teams using multiple AI coding assistants. It is especially useful for those seeking to reduce API costs or avoid rate limits.

  • Broad client support: Seamlessly integrates with Claude Code, Cursor, Copilot, and other major AI coding assistants.
  • Extensive provider network: Connects to more than 40 AI providers, ensuring redundancy and model variety.
  • Automatic fallback routing: Reroutes failed API calls to alternative, often free, providers to prevent workflow interruption.
  • Token saving algorithm: Reduces token consumption by up to 40% through intelligent Request Token Knockdown.
  • Centralized dashboard: Provides a unified interface to monitor usage, manage keys, and configure routing rules locally.
  • Multi-account load balancing: Maximizes API quota utility by round-robining requests across multiple provider accounts transparently.

Where teams use it

Uninterrupted AI coding

Allows developers to continue using AI features in their IDE even if their primary LLM provider experiences an outage.

Cost reduction for teams

Leverages the RTK algorithm to minimize token usage and route non-critical requests to cheaper or free models.

Multi-model experimentation

Enables developers to quickly switch between Claude, GPT, and Gemini for the same coding task without changing IDE settings.

Bypassing rate limits

Distributes requests across multiple free-tier API accounts to avoid hitting provider-specific rate limits.

Getting started: npm install -g 9router && 9router

README

master branch
9Router Dashboard

9Router - FREE AI Router & Token Saver

Never stop coding. Save 20-40% tokens with RTK + auto-fallback to FREE & cheap AI models.

Connect All AI Code Tools (Claude Code, Cursor, Antigravity, Copilot, Codex, Gemini, OpenCode, Cline, OpenClaw...) to 40+ AI Providers & 100+ Models.

npm Downloads Docker Pulls GHCR License

decolua%2F9router | Trendshift

🚀 Quick Start • 💡 Features • 📖 Setup • 🌐 Website

🇧🇷 Português (Brasil) • 🇻🇳 Tiếng Việt • 🇨🇳 中文 • 🇯🇵 日本語 • 🇷🇺 Русский • 🇹🇭 ไทย • 🇮🇷 فارسی • 🇮🇩 Indonesia • 🇪🇸 Español • 🇫🇷 Français


🤔 Why 9Router?

Stop wasting money, tokens and hitting limits:

  • ❌ Subscription quota expires unused every month
  • ❌ Rate limits stop you mid-coding
  • ❌ Tool outputs (git diff, grep, ls...) burn tokens fast
  • ❌ Expensive APIs ($20-50/month per provider)
  • ❌ Manual switching between providers

9Router solves this:

  • ✅ RTK Token Saver - Auto-compress tool_result content, save 20-40% tokens per request
  • ✅ Maximize subscriptions - Track quota, use every bit before reset
  • ✅ Auto fallback - Subscription → Cheap → Free, zero downtime
  • ✅ Multi-account - Round-robin between accounts per provider
  • ✅ Universal - Works with Claude Code, Codex, Cursor, Cline, any CLI tool

🔄 How It Works

┌─────────────┐
│  Your CLI   │  (Claude Code, Codex, OpenClaw, Cursor, Cline...)
│   Tool      │
└──────┬──────┘
       │ http://localhost:20128/v1
       ↓
┌─────────────────────────────────────────────┐
│           9Router (Smart Router)            │
│  • RTK Token Saver (cut tool_result tokens) │
│  • Format translation (OpenAI ↔ Claude)     │
│  • Quota tracking                           │
│  • Auto token refresh                       │
└──────┬──────────────────────────────────────┘
       │
       ├─→ [Tier 1: SUBSCRIPTION] Claude Code, Codex, GitHub Copilot
       │   ↓ quota exhausted
       ├─→ [Tier 2: CHEAP] GLM ($0.6/1M), MiniMax ($0.2/1M)
       │   ↓ budget limit
       └─→ [Tier 3: FREE] Kiro, OpenCode Free, Vertex ($300 credits)

Result: Never stop coding, minimal cost + 20-40% token savings via RTK

⚡ Quick Start

1. Install globally:

npm install -g 9router
9router

🎉 Dashboard opens at http://localhost:20128

2. Connect a FREE provider (no signup needed):

Dashboard → Providers → Connect Kiro AI (~50 credits/month free: Claude 4.5 + GLM-5 + MiniMax) or OpenCode Free (no auth) → Done!

3. Use in your CLI tool:

Claude Code/Codex/OpenClaw/Cursor/Cline Settings:
  Endpoint: http://localhost:20128/v1
  API Key: [copy from dashboard]
  Model: kr/claude-sonnet-4.5

That's it! Start coding with FREE AI models.

Alternative: run from source (this repository):

This repository package is private (9router-app), so source/Docker execution is the expected local development path.

cp .env.example .env
npm install
PORT=20128 NEXT_PUBLIC_BASE_URL=http://localhost:20128 npm run dev

Production mode:

# Create Temporary Memory For Build
sudo fallocate -l 2G /swapfile_temp
sudo chmod 600 /swapfile_temp
sudo mkswap /swapfile_temp
sudo swapon /swapfile_temp

export MAKEFLAGS="-j1"
export DLIB_NO_GUI_SUPPORT=1
export CFLAGS="-mno-avx"

npm run build

# Clear temporary swap
sudo swapoff /swapfile_temp
sudo rm /swapfile_temp

PORT=20128 HOSTNAME=0.0.0.0 NEXT_PUBLIC_BASE_URL=http://localhost:20128 npm run start

Default URLs:

  • Dashboard: http://localhost:20128/dashboard
  • OpenAI-compatible API: http://localhost:20128/v1

Video Guides

Tiết kiệm chi phí LLM với 9Router
🇻🇳 Tiếng Việt
Tiết kiệm chi phí LLM cho OpenClaw với 9Router
by Mì AI
9Router + Claude Code FREE Unlimited Setup
🇵🇰 اردو / हिन्दी
9Router + Claude Code FREE Unlimited Setup
by Build AI With Hamid
9Router Setup Tutorial
🇺🇸 English
9Router + Claude Code FREE Setup
by Build AI With Hamid
9Router Setup Tutorial
🇺🇸 English
9Router + Claude Code FREE Setup
by Build AI With Hamid
Claude Code FREE Forever
🇺🇸 English
Claude Code FREE Forever — Unlimited Models
by Build AI With Hamid
Claude CLI Free Setup
🇺🇸 English
Claude CLI Free Setup with 9Router 🚀
by CodeVerse Soban
Cài đặt OpenClaw Free A-Z
🇻🇳 Tiếng Việt
Cài Đặt OpenClaw Free Từ A-Z + 9Router
by Mai Gia
FREE OpenClaw with Claude Opus
🇺🇸 English
FREE OpenClaw + Claude Opus 4.6
by Build AI With Hamid
Claude CLI Free Setup
🇮🇩 Indonesia
Koding 24 Jam Anti Rate Limit! Hemat Token AI 65% | Tutorial Quick Setup 9Router 🚀
by Krisswuh
Cara Deploy 9Router di Hugging Face GRATIS Non-Stop! | Alternatif VPS RAM 16GB
🇮🇩 Indonesia
Cara Deploy 9Router di Hugging Face GRATIS Non-Stop! | Alternatif VPS RAM 16GB
by Krisswuh
این شکلی از هر API ای استفاده کن برای هوش مصنوعی
🇮🇷 Persian-فارسی
این شکلی از هر API ای استفاده کن برای هوش مصنوعی
by Matin SenPai
Hướng Dẫn Setup OpenClaw + 9Router: Tạo Bot Zalo AI Tự Động Từ A-Z
🇻🇳 Tiếng Việt
Hướng Dẫn Setup OpenClaw + 9Router: Tạo Bot Zalo AI Tự Động Từ A-Z
by tuanminhhole
Bye Limit! Cara Bikin Sistem 'AI Unlimited' 100% Gratis Dengan 9Router!
🇮🇩 Indonesia
Bye Limit! Cara Bikin Sistem "AI Unlimited" 100% Gratis Dengan 9Router!
by neptiver

🎬 Made a video about 9Router? Submit a Pull Request adding your video to this section — we'll merge it!


🛠️ Supported CLI Tools

9Router works seamlessly with all major AI coding tools:

Claude Code
Claude-Code
OpenClaw
OpenClaw
Codex
Codex
OpenCode
OpenCode
Cursor
Cursor
Antigravity
Antigravity
Cline
Cline
Continue
Continue
Droid
Droid
Roo
Roo
Copilot
Copilot
Kilo Code
Kilo Code
OpenDesign
OpenDesign
jcode
jcode
Grok Build
Grok Build
Devin CLI
Devin CLI
DeepSeek TUI
DeepSeek TUI
Qwen Code
Qwen Code

🌐 Supported Providers

🔐 OAuth Providers

Claude Code
Claude-Code
Antigravity
Antigravity
Codex
Codex
GitHub
GitHub
Cursor
Cursor
Kimchi
Kimchi

🆓 Free Providers

Kiro
Kiro AI
Claude 4.5 + GLM-5 + MiniMax
50 credits/month free
OpenCode Free
OpenCode Free
No auth • Auto-fetch models
Free (model list varies)
Vertex AI
Vertex AI
Gemini 3 Pro + GLM-5 + DeepSeek
$300 credits free

Note: iFlow, Qwen Code and Gemini CLI free tiers were discontinued in 2026. Use Kiro / OpenCode Free / Vertex instead.

Kiro AI moved to a paid model in Sep 2025 — the free tier is now capped at 50 credits/month (plus 500 trial credits for new accounts in the first 30 days). Paid tiers: Pro $20/mo (1,000 credits), Pro+ $40/mo (2,000), Pro Max $100/mo (5,000), Power $200/mo (10,000). OpenCode Free model list fluctuates over time (some models free only for limited promos) — subject to change without notice. Vertex AI: the $300 free credit for new GCP accounts is still valid, but since Mar 2026 the Gemini API endpoint no longer consumes these credits — call the Vertex AI Studio endpoint instead.

🔑 API Key Providers (40+)

OpenRouter
OpenRouter
GLM
GLM
Kimi
Kimi
MiniMax
MiniMax
OpenAI
OpenAI
Anthropic
Anthropic
Gemini
Gemini
DeepSeek
DeepSeek
Groq
Groq
xAI
xAI
Mistral
Mistral
Perplexity
Perplexity
Together
Together AI
Fireworks
Fireworks
Cerebras
Cerebras
Cohere
Cohere
NVIDIA
NVIDIA
SiliconFlow
SiliconFlow

...and 20+ more providers including Nebius, Chutes, Hyperbolic, and custom OpenAI/Anthropic compatible endpoints

🏠 Self-hosted Providers

For speech and embeddings served from your own machine — whisper.cpp, faster-whisper, Speaches, Kokoro-FastAPI, openedai-speech, llama.cpp/llama-server, vLLM, Infinity, text-embeddings-inference, or anything else that speaks the OpenAI shape.

Provider Endpoint used Typical server
Self-hosted STT /v1/audio/transcriptions whisper.cpp, faster-whisper
Self-hosted TTS /v1/audio/speech Kokoro-FastAPI, openedai-speech
Self-hosted Embedding /v1/embeddings llama-server, vLLM, Infinity

Every other speech provider is a named cloud service with a fixed endpoint. These three read their address from each connection, so one provider can front several machines and load-balance across them like any other.

Set it on the connection as providerSpecificData.baseUrl:

Provider Give it Result
Self-hosted STT the full URL — http://host:8080/v1/audio/transcriptions used as-is
Self-hosted TTS the server root — http://host:8880 + /v1/audio/speech
Self-hosted Embedding the OpenAI base, /v1 included — http://host:8080/v1 + /embeddings

Mind the /v1 on embeddings. The adapter appends /embeddings, so http://host:8080 resolves to http://host:8080/embeddings and misses the OpenAI route — llama-server answers 501. Give it the same base URL an OpenAI client would use. A full .../v1/embeddings is also accepted, so a value pasted from a curl example works too.

The API key is not checked by most local servers, but the field must be non-empty: it is what gives the connection a credentials record, and baseUrl lives there. Any placeholder works.

Self-hosted Embedding has no cloud fallback by design — a connection saved without a baseUrl is reported as a configuration error rather than quietly falling back to api.openai.com, which would send your input text and API key to a third party under a provider named "Self-hosted".


💡 Key Features

Feature What It Does Why It Matters
🚀 RTK Token Saver (RTK ⭐40K) Compress tool outputs (git diff, grep, ls, tree...) before sending to LLM Save 20-40% input tokens per request
🧠 Headroom Token Saver (Headroom) Optional external /v1/compress proxy before provider routing Save more context tokens without changing clients
🪨 Caveman Mode (Caveman ⭐52K) Inject caveman-speak prompt → LLM replies terse, technical substance preserved Save up to 65% output tokens
🐴 Ponytail (Ponytail) Inject "lazy senior dev" prompt → LLM writes minimal, YAGNI-first code (Lite/Full/Ultra) Fewer output tokens, less refactoring
🎯 Smart 3-Tier Fallback Auto-route: Subscription → Cheap → Free Never stop coding, zero downtime
📊 Real-Time Quota Tracking Live token count + reset countdown Maximize subscription value
🔄 Format Translation OpenAI ↔ Claude ↔ Gemini ↔ Cursor ↔ Kiro ↔ Vertex Works with any CLI tool
👥 Multi-Account Support Multiple accounts per provider Load balancing + redundancy
🔄 Auto Token Refresh OAuth tokens refresh automatically No manual re-login needed
🎨 Custom Combos Create unlimited model combinations Tailor fallback to your needs
📝 Request Logging Debug mode with full request/response logs Troubleshoot issues easily
💾 Cloud Sync Sync config across devices Same setup everywhere
📊 Usage Analytics Track tokens, cost, trends over time Optimize spending
🌐 Deploy Anywhere Localhost, VPS, Docker, Cloudflare Workers Flexible deployment options

Set X-9Router-Token-Saver: off to bypass all token savers for one chat request.

📖 Feature Details

🚀 RTK Token Saver

Tool outputs (git diff, grep, find, ls, tree, log dumps...) often eat 30-50% of your prompt budget. RTK detects them and applies smart, lossless compression before the request hits the LLM:

  • Filters: git-diff, git-status, grep, find, ls, tree, dedup-log, smart-truncate, read-numbered, search-list
  • Auto-detect: No config needed — RTK peeks the first 1KB of each tool_result and picks the right filter.
  • Safe by design: If a filter fails, throws, or makes output bigger, RTK silently keeps the original text. Errors never break your request.
  • Universal: Works across all formats (OpenAI, Claude, Gemini, Cursor, Kiro, OpenAI Responses) because it runs before any format translation.
  • Default ON: Toggle anytime in Dashboard → Endpoint settings.
Without RTK: 47K tokens sent to LLM
With RTK:    28K tokens sent to LLM   (40% saved · same context · same answer)

🧠 Headroom Token Saver

Headroom is optional and runs separately. 9Router calls Headroom's local /v1/compress endpoint, then keeps normal routing, fallback, auth, and usage tracking:

Client → 9Router → Headroom /v1/compress → 9Router → provider

Local setup:

pip install "headroom-ai[proxy]"
headroom proxy --port 8787

Enable in Dashboard → Endpoint → Token Saver → Headroom. Default URL: http://localhost:8787.

Docker examples:

# Headroom service in same Docker network
http://headroom:8787

# Headroom running on host machine
http://host.docker.internal:8787

If Headroom is down or returns an error, 9Router fails open and sends the original request.

🐴 Ponytail (Lazy Senior Dev)

Ponytail injects a "lazy senior dev" system prompt into every request, biasing the LLM toward minimal, YAGNI-first code — deletion over addition, stdlib over new deps, one-liners over abstractions. Adapted from DietrichGebert/ponytail.

  • Lite — Build what's asked, name the lazier alternative.
  • Full — YAGNI ladder enforced: stdlib → native → existing deps → one-liner → minimal code.
  • Ultra — YAGNI extremist: deletion first, ship the one-liner, challenge the rest of the requirement in the same response.
Without Ponytail: verbose code, extra abstractions, "just in case" scaffolding
With Ponytail:    shortest working diff, no unrequested abstractions, fewer tokens

Never trades away: input validation, error handling that prevents data loss, security, accessibility, or anything explicitly requested. Enable in Dashboard → Endpoint → Ponytail. Stacks with Caveman (output terseness) and RTK (input compression).

🎯 Smart 3-Tier Fallback

Create combos with automatic fallback:

Combo: "my-coding-stack"
  1. cc/claude-opus-4-6        (your subscription)
  2. glm/glm-4.7               (cheap backup, $0.6/1M)
  3. if/kimi-k2-thinking       (free fallback)

→ Auto switches when quota runs out or errors occur

📊 Real-Time Quota Tracking

  • Token consumption per provider
  • Reset countdown (5-hour, daily, weekly)
  • Cost estimation for paid tiers
  • Monthly spending reports

🔄 Format Translation

Seamless translation between formats:

  • OpenAI ↔ Claude ↔ Gemini ↔ Cursor ↔ Kiro ↔ Vertex ↔ Antigravity ↔ Ollama ↔ OpenAI Responses
  • Your CLI tool sends OpenAI format → 9Router translates → Provider receives native format
  • Works with any tool that supports custom OpenAI endpoints

👥 Multi-Account Support

  • Add multiple accounts per provider
  • Auto round-robin or priority-based routing
  • Fallback to next account when one hits quota

🔄 Auto Token Refresh

  • OAuth tokens automatically refresh before expiration
  • No manual re-authentication needed
  • Seamless experience across all providers

🎨 Custom Combos

  • Create unlimited model combinations
  • Mix subscription, cheap, and free tiers
  • Name your combos for easy access
  • Share combos across devices with Cloud Sync

📝 Request Logging

  • Enable debug mode for full request/response logs
  • Track API calls, headers, and payloads
  • Troubleshoot integration issues
  • Export logs for analysis

💾 Cloud Sync

  • Sync providers, combos, and settings across devices
  • Automatic background sync
  • Secure encrypted storage
  • Access your setup from anywhere
Cloud Runtime Notes
  • Prefer server-side cloud variables in production:
    • BASE_URL (internal callback URL used by sync scheduler)
    • CLOUD_URL (cloud sync endpoint base)
  • NEXT_PUBLIC_BASE_URL and NEXT_PUBLIC_CLOUD_URL are still supported for compatibility/UI, but server runtime now prioritizes BASE_URL/CLOUD_URL.
  • Cloud sync requests now use timeout + fail-fast behavior to avoid UI hanging when cloud DNS/network is unavailable.

📊 Usage Analytics

  • Track token usage per provider and model
  • Cost estimation and spending trends
  • Monthly reports and insights
  • Optimize your AI spending

💡 IMPORTANT - Understanding Dashboard Costs:

The "cost" displayed in Usage Analytics is for tracking and comparison purposes only. 9Router itself never charges you anything. You only pay providers directly (if using paid services).

Example: If your dashboard shows "$290 total cost" while using Kiro free models, this represents what you would have paid using paid APIs directly. Your actual cost = $0 (Kiro free tier: ~50 credits/mo).

Think of it as a "savings tracker" showing how much you're saving by using free models or routing through 9Router!

🌐 Deploy Anywhere

  • 💻 Localhost - Default, works offline
  • ☁️ VPS/Cloud - Share across devices
  • 🐳 Docker - One-command deployment
  • 🚀 Cloudflare Workers - Global edge network

💰 Pricing at a Glance

Tier Provider Cost Quota Reset Best For
🚀 TOKEN SAVER RTK (built-in) FREE Always on Save 20-40% tokens on EVERY request
💳 SUBSCRIPTION Claude Code (Pro/Max) $20-200/mo 5h + weekly Already subscribed
Codex (Plus/Pro) $20-200/mo 5h + weekly OpenAI users
GitHub Copilot $10-19/mo Monthly GitHub users
Cursor IDE $20/mo Monthly Cursor users
💰 CHEAP GLM-5.1 / GLM-4.7 $0.6/1M Daily 10AM Budget backup
MiniMax M2.7 $0.2/1M 5-hour rolling Cheapest option
Kimi K2.5 $9/mo flat 10M tokens/mo Predictable cost
🆓 FREE Kiro AI $0 50 credits/mo Claude 4.5 + GLM-5 + MiniMax free (paid tiers above)
OpenCode Free $0 Varies* No auth, auto-fetch models (list changes over time)
Vertex AI $300 credits New GCP accounts Gemini 3 Pro + DeepSeek + GLM-5 (use Vertex AI Studio endpoint for free credits)

💡 Pro Tip: RTK + Kiro AI + OpenCode Free combo = $0 cost + 20-40% token savings!


📊 Understanding 9Router Costs & Billing

9Router Billing Reality:

✅ 9Router software = FREE forever (open source, never charges)
✅ Dashboard "costs" = Display/tracking only (not actual bills)
✅ You pay providers directly (subscriptions or API fees)
✅ FREE providers stay FREE (Kiro ~50 credits/mo, OpenCode Free, Vertex $300 credits = $0 within free-tier limits) — note iFlow/Qwen/Gemini CLI free tiers were discontinued in 2026 ❌ 9Router never sends invoices or charges your card

How Cost Display Works:

The dashboard shows estimated costs as if you were using paid APIs directly. This is not billing - it's a comparison tool to show your savings.

Example Scenario:

Dashboard Display:
• Total Requests: 1,662
• Total Tokens: 47M
• Display Cost: $290

Reality Check:
• Provider: Kiro (free tier: ~50 credits/mo)
• Actual Payment: $0.00
• What $290 Means: Amount you SAVED by using free models!

Payment Rules:

  • Subscription providers (Claude Code, Codex): Pay them directly via their websites
  • Cheap providers (GLM, MiniMax): Pay them directly, 9Router just routes
  • FREE providers (iFlow, Kiro, Qwen): Genuinely free forever, no hidden charges
  • 9Router: Never charges anything, ever

🎯 Use Cases

Case 1: "I have Claude Pro subscription"

Problem: Quota expires unused, rate limits during heavy coding

Solution:

Combo: "maximize-claude"
  1. cc/claude-opus-4-7        (use subscription fully)
  2. glm/glm-5.1               (cheap backup when quota out)
  3. kr/claude-sonnet-4.5      (free emergency fallback)

Monthly cost: $20 (subscription) + ~$5 (backup) = $25 total
vs. $20 + hitting limits = frustration

Case 2: "I want zero cost"

Problem: Can't afford subscriptions, need reliable AI coding

Solution:

Combo: "free-forever"
  1. kr/claude-sonnet-4.5      (Claude 4.5 free via Kiro, ~50 credits/mo)
  2. kr/glm-5                  (GLM-5 free via Kiro)
  3. oc/<auto>                 (OpenCode Free, no auth)

Monthly cost: $0
Quality: Production-ready models + RTK saves 20-40% tokens

Case 3: "I need 24/7 coding, no interruptions"

Problem: Deadlines, can't afford downtime

Solution:

Combo: "always-on"
  1. cc/claude-opus-4-7        (best quality)
  2. cx/gpt-5.5                (second subscription)
  3. glm/glm-5.1               (cheap, resets daily)
  4. minimax/MiniMax-M2.7      (cheapest, 5h reset)
  5. kr/claude-sonnet-4.5      (free via Kiro, ~50 credits/mo)

Result: 5 layers of fallback = zero downtime
Monthly cost: $20-200 (subscriptions) + $10-20 (backup)

Case 4: "I want FREE AI in OpenClaw"

Problem: Need AI assistant in messaging apps (WhatsApp, Telegram, Slack...), completely free

Solution:

Combo: "openclaw-free"
  1. kr/claude-sonnet-4.5      (Claude 4.5 free)
  2. kr/glm-5                  (GLM-5 free)
  3. kr/MiniMax-M2.5           (MiniMax free)

Monthly cost: $0
Access via: WhatsApp, Telegram, Slack, Discord, iMessage, Signal...

❓ Frequently Asked Questions

📊 Why does my dashboard show high costs?

The dashboard tracks your token usage and displays estimated costs as if you were using paid APIs directly. This is not actual billing - it's a reference to show how much you're saving by using free models or existing subscriptions through 9Router.

Example:

  • Dashboard shows: "$290 total cost"
  • Reality: You're using Kiro free models (~50 credits/mo)
  • Your actual cost: $0.00
  • What $290 means: Amount you saved by using free models instead of paid APIs!

The cost display is a "savings tracker" to help you understand your usage patterns and optimization opportunities.

💳 Will I be charged by 9Router?

No. 9Router is free, open-source software that runs on your own computer. It never charges you anything.

You only pay:

  • ✅ Subscription providers (Claude Code $20/mo, Codex $20-200/mo) → Pay them directly on their websites
  • ✅ Cheap providers (GLM, MiniMax) → Pay them directly, 9Router just routes your requests
  • ❌ 9Router itself → Never charges anything, ever

9Router is a local proxy/router. It doesn't have your credit card, can't send invoices, and has no billing system. It's completely free software.

🆓 Are FREE providers really unlimited?

Mostly! The current FREE providers (Kiro, OpenCode Free, Vertex) are genuinely free, but free tiers have limits:

These are free services offered by those respective companies:

  • Kiro AI: ~50 credits/month free (plus 500 trial credits for new accounts in the first 30 days) via AWS Builder ID / Google / GitHub OAuth. Paid tiers available above that.
  • OpenCode Free: No-auth passthrough proxy, models auto-fetched from opencode.ai/zen/v1/models. The free model list fluctuates over time (some models free only for limited promos) — subject to change without notice.
  • Vertex AI: $300 free credits for new Google Cloud accounts (90 days). Since Mar 2026 the Gemini API endpoint no longer consumes these credits — use the Vertex AI Studio endpoint instead.

9Router just routes your requests to them - there's no "catch" or future billing from 9Router itself. They're truly free services, and 9Router makes them easy to use with fallback support.

Discontinued free tiers (no longer recommended):

  • ❌ iFlow: Was free unlimited, now changed to paid (2026)
  • ❌ Qwen Code: Free OAuth tier fully discontinued by Alibaba on 2026-04-15
  • ❌ Gemini CLI: Service fully shut down by Google on 2026-06-18 (replaced by the closed-source Antigravity CLI). Discontinued — do not use.
💰 How do I minimize my actual AI costs?

Free-First Strategy:

  1. Start with 100% free combo:

    1. kr/glm-5 (GLM-5 free via Kiro, ~50 credits/mo)
    2. OpenCode Free models (no auth, auto-fetched)
    3. Vertex AI Gemini 3 Pro (using the Vertex AI Studio endpoint with $300 credits)
    

    Cost: $0/month (within Kiro's free credit cap; OpenCode/Vertex subject to their free-tier limits)

  2. Add cheap backup only if you need it:

    4. glm/glm-4.7 ($0.6/1M tokens)
    

    Additional cost: Only pay for what you actually use

  3. Use subscription providers last:

    • Only if you already have them
    • 9Router helps maximize their value through quota tracking

Result: Most users can operate at $0/month using only free tiers!

📈 What if my usage suddenly spikes?

9Router's smart fallback prevents surprise charges:

Scenario: You're on a coding sprint and blow through your quotas

Without 9Router:

  • ❌ Hit rate limit → Work stops → Frustration
  • ❌ Or: Accidentally rack up huge API bills

With 9Router:

  • ✅ Subscription hits limit → Auto-fallback to cheap tier
  • ✅ Cheap tier gets expensive → Auto-fallback to free tier
  • ✅ Never stop coding → Predictable costs

You're in control: Set spending limits per provider in dashboard, and 9Router respects them.


📖 Setup Guide

🔐 Subscription Providers (Maximize Value)

Claude Code (Pro/Max)

Dashboard → Providers → Connect Claude Code
→ OAuth login → Auto token refresh
→ 5-hour + weekly quota tracking

Models:
  cc/claude-opus-4-7
  cc/claude-opus-4-6
  cc/claude-sonnet-4-6
  cc/claude-haiku-4-5-20251001

Pro Tip: Use Opus for complex tasks, Sonnet for speed. 9Router tracks quota per model!

OpenAI Codex (Plus/Pro)

Dashboard → Providers → Connect Codex
→ OAuth login (port 1455)
→ 5-hour + weekly reset

Models:
  cx/gpt-5.5
  cx/gpt-5.4
  cx/gpt-5.3-codex
  cx/gpt-5.2-codex

GitHub Copilot

Dashboard → Providers → Connect GitHub
→ OAuth via GitHub
→ Monthly reset (1st of month)

Models:
  gh/gpt-5.4
  gh/claude-opus-4.7
  gh/claude-sonnet-4.6
  gh/gemini-3.1-pro-preview
  gh/grok-code-fast-1

Cursor IDE

Dashboard → Providers → Connect Cursor
→ OAuth login
→ Monthly subscription

Models:
  cu/claude-4.6-opus-max
  cu/claude-4.5-sonnet-thinking
  cu/gpt-5.3-codex
💰 Cheap Providers (Backup)

GLM-5.1 / GLM-4.7 (Daily reset, $0.6/1M)

  1. Sign up: Zhipu AI
  2. Get API key from Coding Plan
  3. Dashboard → Add API Key:
    • Provider: glm
    • API Key: your-key

Use: glm/glm-5.1, glm/glm-5, glm/glm-4.7

Pro Tip: Coding Plan offers 3× quota at 1/7 cost! Reset daily 10:00 AM.

MiniMax M2.7 (5h reset, $0.20/1M)

  1. Sign up: MiniMax
  2. Get API key
  3. Dashboard → Add API Key

Use: minimax/MiniMax-M2.7, minimax/MiniMax-M2.5

Pro Tip: Cheapest option for long context (1M tokens)!

Kimi K2.5 ($9/month flat)

  1. Subscribe: Moonshot AI
  2. Get API key
  3. Dashboard → Add API Key

Use: kimi/kimi-k2.5, kimi/kimi-k2.5-thinking

Pro Tip: Fixed $9/month for 10M tokens = $0.90/1M effective cost!

🆓 FREE Providers (Recommended)

Kiro AI (Claude 4.5 + GLM-5 + MiniMax FREE)

Dashboard → Connect Kiro
→ AWS Builder ID, AWS IAM Identity Center, Google, or GitHub
→ Unlimited usage

Models:
  kr/claude-sonnet-4.5
  kr/claude-haiku-4.5
  kr/glm-5
  kr/MiniMax-M2.5
  kr/qwen3-coder-next
  kr/deepseek-3.2

Pro Tip: Best free option for Claude. No API key, no payment, fully unlimited.

OpenCode Free (No auth, auto-fetch models)

Dashboard → Connect OpenCode Free
→ No login required (passthrough proxy)
→ Models auto-fetched from opencode.ai/zen/v1/models

Pro Tip: Fastest setup. Just connect and start coding.

Vertex AI ($300 free credits for new GCP accounts)

Dashboard → Connect Vertex AI
→ Upload Google Cloud Service Account JSON
→ Enable Vertex AI API in your GCP project

Models:
  vertex/gemini-3.1-pro-preview
  vertex/gemini-3-flash-preview
  vertex/gemini-2.5-flash

Vertex Partner (Anthropic / DeepSeek / GLM / Qwen via Vertex):
  vertex-partner/glm-5-maas
  vertex-partner/deepseek-v3.2-maas
  vertex-partner/qwen3-next-80b-a3b-thinking-maas

Pro Tip: New Google Cloud accounts get $300 credits free for 90 days. Plenty for daily coding.

🎨 Create Combos

Example 1: Maximize Subscription → Cheap Backup

Dashboard → Combos → Create New

Name: premium-coding
Models:
  1. cc/claude-opus-4-7 (Subscription primary)
  2. glm/glm-5.1 (Cheap backup, $0.6/1M)
  3. minimax/MiniMax-M2.7 (Cheapest fallback, $0.20/1M)

Use in CLI: premium-coding

Monthly cost example (100M tokens):
  80M via Claude (subscription): $0 extra
  15M via GLM: $9
  5M via MiniMax: $1
  Total: $10 + your subscription

Example 2: Free-Only (Zero Cost)

Name: free-combo
Models:
  1. kr/claude-sonnet-4.5 (Claude 4.5 free via Kiro, ~50 credits/mo)
  2. kr/glm-5 (GLM-5 free via Kiro)
  3. vertex/gemini-3.1-pro-preview ($300 free credits)

Cost: $0 forever (+ 20-40% token savings via RTK)!
🔧 CLI Integration

Cursor IDE

Settings → Models → Advanced:
  OpenAI API Base URL: http://localhost:20128/v1
  OpenAI API Key: [from 9router dashboard]
  Model: cc/claude-opus-4-7

Or use combo: premium-coding

Claude Code

Edit ~/.claude/config.json:

{
  "anthropic_api_base": "http://localhost:20128/v1",
  "anthropic_api_key": "your-9router-api-key"
}

Codex CLI

export OPENAI_BASE_URL="http://localhost:20128"
export OPENAI_API_KEY="your-9router-api-key"

codex "your prompt"

OpenClaw

Option 1 — Dashboard (recommended):

Dashboard → CLI Tools → OpenClaw → Select Model → Apply

Option 2 — Manual: Edit ~/.openclaw/openclaw.json:

{
  "agents": {
    "defaults": {
      "model": {
        "primary": "9router/kr/claude-sonnet-4.5"
      }
    }
  },
  "models": {
    "providers": {
      "9router": {
        "baseUrl": "http://127.0.0.1:20128/v1",
        "apiKey": "sk_9router",
        "api": "openai-completions",
        "models": [
          {
            "id": "kr/claude-sonnet-4.5",
            "name": "Claude Sonnet 4.5 (Kiro Free)"
          }
        ]
      }
    }
  }
}

Note: OpenClaw only works with local 9Router. Use 127.0.0.1 instead of localhost to avoid IPv6 resolution issues.

Cline / Continue / RooCode

Provider: OpenAI Compatible
Base URL: http://localhost:20128/v1
API Key: [from dashboard]
Model: cc/claude-opus-4-7
🚀 Deployment

VPS Deployment

# Clone and install
git clone https://github.com/decolua/9router.git
cd 9router
npm install
npm run build

# Configure
export JWT_SECRET="your-secure-secret-change-this"
export INITIAL_PASSWORD="your-password"
export DATA_DIR="/var/lib/9router"
export PORT="20128"
export HOSTNAME="0.0.0.0"
export NODE_ENV="production"
export NEXT_PUBLIC_BASE_URL="http://localhost:20128"
export NEXT_PUBLIC_CLOUD_URL="https://9router.com"
export API_KEY_SECRET="endpoint-proxy-api-key-secret"
export MACHINE_ID_SALT="endpoint-proxy-salt"

# Start
npm run start

# Or use PM2
npm install -g pm2
pm2 start npm --name 9router -- start
pm2 save
pm2 startup

Docker

Published images (multi-platform linux/amd64 + linux/arm64):

Quick start (use published image):

docker run -d \
  --name 9router \
  -p 20128:20128 \
  -v "$HOME/.9router:/app/data" \
  -e DATA_DIR=/app/data \
  decolua/9router:latest

(README truncated)

View on GitHub

Recent activity

commits and pull requests

Releases and announcements

73 total
  1. Features: - xAI: Grok Imagine video generation (/v1/videos) + CLI - CLI tools: Grok Build setup — writes [model.9router] to ~/.grok/config.toml - GitHub Copilot: route Claude models through Copilot native /v1/messages - Kiro: add GPT-5.6 model family (#2596) - RTK: X-9Router-Token-Saver header to bypass token savers per request - Providers: quota visibility settings - Translator: drop temperature for all Claude models - i18n: Thai (th) + Persian (fa) translations / README Fixes: - Providers: bulk-add API keys no longer overwrite existing keys (gap-fill Key N) - Anthropic: lowercase anthropic-version header to prevent duplication on /v1/messages - Alicode-intl: use DashScope compatible-mode endpoint so standard keys work - Grok CLI: align Grok Build with current subscription protocol (#2590) - Grok CLI: surface expiresAt so proactive token refresh fires (#2546) - Kiro: improve direct session cache reuse - Models: populate capabilities for live-catalog LLM models - Models: list compatible provider models in /v1/models - Thinking: send explicit thinking:{type:adaptive} alongside output_config.effort - Translator: strip client_metadata when converting openai-responses to openai Impro

  2. Features: - Thinking: per-model thinking level picker on provider page, appends (level) suffix to copied model names for forced reasoning effort across all formats (openai, claude, gemini, deepseek, kimi, qwen, zai, minimax, hunyuan, step) - RTK: add JS-native git-log filter (#2423) - Caveman: add targeted upstream-aligned style rules (#2424) - i18n: add Farsi (fa) language support (#2385) Fixes: - Thinking: strip (level) suffix from upstream body.model so providers no longer reject requests - Translator: preserve developer instructions in openai-responses conversion (#2434) - count_tokens: count structured Anthropic blocks (#2419) - Volcengine-ark: clamp GLM-5 max_tokens to model output ceiling (#2428) - Kimi: normalize reasoning_effort to backend enum (#2427) - Claude: reconcile max_tokens vs thinking budget and lift per-model ceiling (#2381) - Kiro: deliver system prompt natively, add Opus 4.5/4.7/4.8, tolerate dash version ids (#2366) - Headroom: proxy dashboard through app (#2372) - MITM: recover from stale lock file on server start Docs: - README: swap in Vietnamese tutorial video; add English and Urdu/Hindi tutorials (#2305) - Add CLAUDE.md guidance for Claude Code (#2354)

  3. Features: - Usage: track cached tokens + correct input/output/cache cost (#2209) - Codex: show reset credit expiry details (#2290) - NVIDIA: add new models and capabilities - ClinePass: add provider support Fixes: - Usage: dedupe streaming request-details log entries - Claude: drop foreign thinking signatures in passthrough - Prevent non-SSE stream pipe crash and cross-IdP account overwrites (#2244) - Kiro: route IdC auth to regional CodeWhisperer surface (#2297) - Kiro: add Claude Sonnet 5 model support (#2264) - Xiaomi-tokenplan: region selector, key validation, multi-connection (#2251) - Translator: strict Anthropic content block compliance (#2225) - Kimchi: strip reasoning_content echo to bound multi-turn input tokens - Kimchi: bump User-Agent to kimchi/0.1.40 (#2256) - Codebuddy-cn: strip empty tool_calls arrays to preserve reasoning - Antigravity: preserve Claude tool delta index (#2223) - MITM: generate root CA on server startup (#2228)

  4. Features: - Add token-saver dashboard page — decolua - Add bulk delete for provider connections — teddytkz - Resolve GitHub Copilot model catalog from upstream — caiqinzhou - Add Venice AI provider — Brokenc0de - Add Kiro external_idp import for Microsoft SSO (CLIProxyAPI) — Stevanus Pangau - Overhaul Blackbox provider catalog + WebUI test support — suryacagur Fixes: - Provider thinking compatibility (DeepSeek/Gemini) — Mink Nguyen - Stop double-counting streaming usage at source — decolua - Usage logging dedupe to reduce stats churn — Mink Nguyen - Prevent non-JSON SSE lines / duplicate [DONE] from breaking clients (PR #2046) — qianze - Resolve Gemini TTS models from catalog — nguyenha935 - Support Kiro IDC (organization) token import — quanturbo - Preserve forced streaming for JSON clients (#2031) — Joseph Yaksich - Preserve Responses text format (Codex) — tenglong - Support Gemini native TTS generateContent endpoint — nguyenha935 - Add missing zh-CN endpoint key label (i18n) — weimaozhen - CodeBuddy: only send reasoning params when client requests reasoning (#2071) — Rex - Show custom provider models in combo picker — Sapto - Docker: add docker-compose.yml with headroom enabled

  5. Features: - Antigravity: native image generation support (image models tagged kind:image, shown in media-providers UI) - CodeBuddy CN: API key auth + credit quota tracker - CodeBuddy CN: short model prefix alias cbcn Fixes: - MiniMax-M3: enable vision capability - Headroom: support Docker sidecar proxy - Antigravity: image executor fixes - mimo-free: Chrome User-Agent rotation to bypass anti-abuse gate - cloudflare-ai: flatten content-part arrays to string to avoid oneOf 400 (#1926) - Translator: normalize tools to Anthropic-native shape for non-Anthropic providers - CLI: handle Next.js 16 nested standalone output path (#1940) - Codex: preserve custom tools during request normalization - next.config: add new route for responses endpoint to API

Code frequency

additions and deletions
+51.2K-51.2KWeek of 2026-01-04: +33,884 linesWeek of 2026-01-04: -2,643 linesWeek of 2026-01-11: +3,775 linesWeek of 2026-01-11: -7,609 linesWeek of 2026-01-18: +1,775 linesWeek of 2026-01-18: -744 linesWeek of 2026-01-25: +4,103 linesWeek of 2026-01-25: -1,010 linesWeek of 2026-02-01: +9,680 linesWeek of 2026-02-01: -2,772 linesWeek of 2026-02-08: +8,595 linesWeek of 2026-02-08: -2,288 linesWeek of 2026-02-15: +8,077 linesWeek of 2026-02-15: -3,630 linesWeek of 2026-02-22: +7,581 linesWeek of 2026-02-22: -2,769 linesWeek of 2026-03-01: +9,343 linesWeek of 2026-03-01: -2,083 linesWeek of 2026-03-08: +5,767 linesWeek of 2026-03-08: -2,696 linesWeek of 2026-03-15: +7,488 linesWeek of 2026-03-15: -776 linesWeek of 2026-03-22: +5,118 linesWeek of 2026-03-22: -653 linesWeek of 2026-03-29: +4,363 linesWeek of 2026-03-29: -1,815 linesWeek of 2026-04-05: +4,888 linesWeek of 2026-04-05: -1,241 linesWeek of 2026-04-12: +5,090 linesWeek of 2026-04-12: -3,319 linesWeek of 2026-04-19: +7,642 linesWeek of 2026-04-19: -571 linesWeek of 2026-04-26: +5,790 linesWeek of 2026-04-26: -1,358 linesWeek of 2026-05-03: +20,239 linesWeek of 2026-05-03: -8,159 linesWeek of 2026-05-10: +51,237 linesWeek of 2026-05-10: -5,605 linesWeek of 2026-05-17: +6,695 linesWeek of 2026-05-17: -1,442 linesWeek of 2026-05-24: +3,623 linesWeek of 2026-05-24: -2,188 linesWeek of 2026-05-31: +5,631 linesWeek of 2026-05-31: -1,249 linesWeek of 2026-06-07: +13,539 linesWeek of 2026-06-07: -5,480 linesWeek of 2026-06-14: +23,037 linesWeek of 2026-06-14: -17,635 linesWeek of 2026-06-21: +4,713 linesWeek of 2026-06-21: -709 linesWeek of 2026-06-28: +4,382 linesWeek of 2026-06-28: -610 linesWeek of 2026-07-05: +9,481 linesWeek of 2026-07-05: -1,391 linesWeek of 2026-07-12: +10,599 linesWeek of 2026-07-12: -1,253 linesWeek of 2026-07-19: +12,635 linesWeek of 2026-07-19: -3,245 linesWeek of 2026-07-26: +7,357 linesWeek of 2026-07-26: -1,091 linesWeek of 2026-08-02: +5,507 linesWeek of 2026-08-02: -1,550 linesWeek of 2026-08-09: +9,166 linesWeek of 2026-08-09: -422 linesWeek of 2026-08-16: +79 linesWeek of 2026-08-16: -10 linesWeek of 2026-08-23: +5,752 linesWeek of 2026-08-23: -2,418 linesWeek of 2026-08-30: +6,760 linesWeek of 2026-08-30: -796 linesWeek of 2026-09-06: +7,109 linesWeek of 2026-09-06: -935 linesWeek of 2026-09-13: +6,648 linesWeek of 2026-09-13: -496 linesWeek of 2026-09-20: +13,408 linesWeek of 2026-09-20: -1,439 linesWeek of 2026-09-27: +6,709 linesWeek of 2026-09-27: -397 linesWeek of 2026-10-04: +0 linesWeek of 2026-10-04: -0 linesJan 4, 2026Oct 4, 2026
+367.3K lines added, -96.5K removed over the last year.

Commits per week

last 52 weeks
740Week of 2025-10-11: 0 commitsWeek of 2025-10-18: 0 commitsWeek of 2025-10-25: 0 commitsWeek of 2025-11-01: 0 commitsWeek of 2025-11-09: 0 commitsWeek of 2025-11-16: 0 commitsWeek of 2025-11-23: 0 commitsWeek of 2025-11-30: 0 commitsWeek of 2025-12-07: 0 commitsWeek of 2025-12-14: 0 commitsWeek of 2025-12-21: 0 commitsWeek of 2025-12-28: 0 commitsWeek of 2026-01-04: 23 commitsWeek of 2026-01-11: 18 commitsWeek of 2026-01-18: 10 commitsWeek of 2026-01-25: 12 commitsWeek of 2026-02-01: 42 commitsWeek of 2026-02-08: 31 commitsWeek of 2026-02-15: 32 commitsWeek of 2026-02-22: 26 commitsWeek of 2026-03-01: 36 commitsWeek of 2026-03-08: 49 commitsWeek of 2026-03-15: 15 commitsWeek of 2026-03-22: 41 commitsWeek of 2026-03-29: 21 commitsWeek of 2026-04-05: 28 commitsWeek of 2026-04-12: 40 commitsWeek of 2026-04-19: 33 commitsWeek of 2026-04-26: 25 commitsWeek of 2026-05-03: 49 commitsWeek of 2026-05-10: 66 commitsWeek of 2026-05-17: 38 commitsWeek of 2026-05-24: 16 commitsWeek of 2026-05-31: 28 commitsWeek of 2026-06-07: 60 commitsWeek of 2026-06-14: 48 commitsWeek of 2026-06-21: 51 commitsWeek of 2026-06-28: 35 commitsWeek of 2026-07-05: 48 commitsWeek of 2026-07-12: 26 commitsWeek of 2026-07-19: 25 commitsWeek of 2026-07-26: 24 commitsWeek of 2026-08-02: 40 commitsWeek of 2026-08-09: 31 commitsWeek of 2026-08-16: 1 commitsWeek of 2026-08-23: 40 commitsWeek of 2026-08-30: 50 commitsWeek of 2026-09-06: 29 commitsWeek of 2026-09-13: 35 commitsWeek of 2026-09-20: 74 commitsWeek of 2026-09-27: 41 commitsWeek of 2026-10-04: 0 commitsOct 11, 2025Oct 4, 2026
1.3K commits in the last 52 weeks.

When work happens

weekday and hour
SunMonTueWedThuFriSat036912151821Sun 0:00 — 4 commitsSun 1:00 — 3 commitsSun 2:00 — 1 commitsSun 3:00 — 0 commitsSun 4:00 — 4 commitsSun 5:00 — 1 commitsSun 6:00 — 0 commitsSun 7:00 — 1 commitsSun 8:00 — 1 commitsSun 9:00 — 1 commitsSun 10:00 — 5 commitsSun 11:00 — 4 commitsSun 12:00 — 3 commitsSun 13:00 — 6 commitsSun 14:00 — 4 commitsSun 15:00 — 10 commitsSun 16:00 — 19 commitsSun 17:00 — 13 commitsSun 18:00 — 9 commitsSun 19:00 — 5 commitsSun 20:00 — 0 commitsSun 21:00 — 6 commitsSun 22:00 — 9 commitsSun 23:00 — 0 commitsMon 0:00 — 1 commitsMon 1:00 — 1 commitsMon 2:00 — 2 commitsMon 3:00 — 0 commitsMon 4:00 — 1 commitsMon 5:00 — 2 commitsMon 6:00 — 3 commitsMon 7:00 — 1 commitsMon 8:00 — 0 commitsMon 9:00 — 18 commitsMon 10:00 — 19 commitsMon 11:00 — 18 commitsMon 12:00 — 22 commitsMon 13:00 — 5 commitsMon 14:00 — 4 commitsMon 15:00 — 35 commitsMon 16:00 — 17 commitsMon 17:00 — 12 commitsMon 18:00 — 6 commitsMon 19:00 — 12 commitsMon 20:00 — 5 commitsMon 21:00 — 2 commitsMon 22:00 — 4 commitsMon 23:00 — 3 commitsTue 0:00 — 1 commitsTue 1:00 — 0 commitsTue 2:00 — 0 commitsTue 3:00 — 0 commitsTue 4:00 — 1 commitsTue 5:00 — 1 commitsTue 6:00 — 1 commitsTue 7:00 — 1 commitsTue 8:00 — 0 commitsTue 9:00 — 13 commitsTue 10:00 — 17 commitsTue 11:00 — 18 commitsTue 12:00 — 6 commitsTue 13:00 — 3 commitsTue 14:00 — 12 commitsTue 15:00 — 11 commitsTue 16:00 — 13 commitsTue 17:00 — 7 commitsTue 18:00 — 2 commitsTue 19:00 — 3 commitsTue 20:00 — 4 commitsTue 21:00 — 2 commitsTue 22:00 — 3 commitsTue 23:00 — 7 commitsWed 0:00 — 2 commitsWed 1:00 — 1 commitsWed 2:00 — 0 commitsWed 3:00 — 0 commitsWed 4:00 — 1 commitsWed 5:00 — 3 commitsWed 6:00 — 4 commitsWed 7:00 — 1 commitsWed 8:00 — 2 commitsWed 9:00 — 15 commitsWed 10:00 — 30 commitsWed 11:00 — 35 commitsWed 12:00 — 6 commitsWed 13:00 — 11 commitsWed 14:00 — 7 commitsWed 15:00 — 16 commitsWed 16:00 — 18 commitsWed 17:00 — 6 commitsWed 18:00 — 10 commitsWed 19:00 — 9 commitsWed 20:00 — 16 commitsWed 21:00 — 4 commitsWed 22:00 — 2 commitsWed 23:00 — 0 commitsThu 0:00 — 2 commitsThu 1:00 — 0 commitsThu 2:00 — 0 commitsThu 3:00 — 1 commitsThu 4:00 — 1 commitsThu 5:00 — 0 commitsThu 6:00 — 0 commitsThu 7:00 — 1 commitsThu 8:00 — 2 commitsThu 9:00 — 30 commitsThu 10:00 — 24 commitsThu 11:00 — 27 commitsThu 12:00 — 5 commitsThu 13:00 — 2 commitsThu 14:00 — 8 commitsThu 15:00 — 28 commitsThu 16:00 — 16 commitsThu 17:00 — 14 commitsThu 18:00 — 28 commitsThu 19:00 — 0 commitsThu 20:00 — 5 commitsThu 21:00 — 5 commitsThu 22:00 — 14 commitsThu 23:00 — 22 commitsFri 0:00 — 4 commitsFri 1:00 — 9 commitsFri 2:00 — 1 commitsFri 3:00 — 0 commitsFri 4:00 — 0 commitsFri 5:00 — 0 commitsFri 6:00 — 4 commitsFri 7:00 — 1 commitsFri 8:00 — 0 commitsFri 9:00 — 14 commitsFri 10:00 — 42 commitsFri 11:00 — 41 commitsFri 12:00 — 26 commitsFri 13:00 — 5 commitsFri 14:00 — 4 commitsFri 15:00 — 22 commitsFri 16:00 — 36 commitsFri 17:00 — 30 commitsFri 18:00 — 11 commitsFri 19:00 — 0 commitsFri 20:00 — 6 commitsFri 21:00 — 8 commitsFri 22:00 — 2 commitsFri 23:00 — 2 commitsSat 0:00 — 1 commitsSat 1:00 — 0 commitsSat 2:00 — 1 commitsSat 3:00 — 0 commitsSat 4:00 — 0 commitsSat 5:00 — 3 commitsSat 6:00 — 1 commitsSat 7:00 — 1 commitsSat 8:00 — 2 commitsSat 9:00 — 10 commitsSat 10:00 — 28 commitsSat 11:00 — 33 commitsSat 12:00 — 13 commitsSat 13:00 — 1 commitsSat 14:00 — 5 commitsSat 15:00 — 11 commitsSat 16:00 — 25 commitsSat 17:00 — 25 commitsSat 18:00 — 4 commitsSat 19:00 — 1 commitsSat 20:00 — 5 commitsSat 21:00 — 23 commitsSat 22:00 — 10 commitsSat 23:00 — 7 commits
Commit volume by weekday and hour (UTC). Larger dots mean more commits.

Who is committing

last 52 weeks
Maintainer commits146 (11%)
Community commits1,218 (89%)

1,364 commits in total over the last year.

DateListRankStars gained
Sep 18, 2026weekly#17+902
Sep 17, 2026weekly#17+902
Sep 16, 2026weekly#17+1,361
Sep 15, 2026weekly#17+1,361
May 12, 2026daily#15+53
May 11, 2026daily#17+112
May 10, 2026daily#20+157
May 9, 2026daily#15+139
May 8, 2026daily#11+91
  • affaan-m/ECC

    The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.

    272.8K stars · JavaScript

  • NousResearch/hermes-agent

    The agent that grows with you

    251.2K stars · Python

  • react/react

    The library for web and native user interfaces.

    250.9K stars · JavaScript

  • firecrawl/firecrawl

    Supercharge your AI agents with data from the web and beyond. Building the library for superintelligence. 🔥

    188.6K stars · TypeScript

  • microsoft/markitdown

    Python tool for converting files and office documents to Markdown.

    188.4K stars · Python

  • Significant-Gravitas/AutoGPT

    AutoGPT is the vision of accessible AI for everyone, to use and to build on. Our mission is to provide the tools, so that you can focus on what matters.

    187.7K stars · Python