9gouter

module
v0.8.6 Latest Latest
Warning

This package is not in the latest version of its module.

Go to latest
Published: Aug 6, 2026 License: MIT

README

9Gouter Dashboard

9Gouter - FREE AI Router & Token Saver

Never stop coding. Save 20-40% tokens with RTK + auto-fallback to FREE & cheap AI models.

Connect All AI Code Tools (Claude Code, Cursor, Antigravity, Copilot, Codex, Gemini, OpenCode, Cline, OpenClaw...) to 40+ AI Providers & 100+ Models.

npm Downloads Docker Pulls GHCR License

🚀 Quick Start💡 Features📖 Setup🌐 Website

🇻🇳 Tiếng Việt🇨🇳 中文🇯🇵 日本語🇷🇺 Русский🇹🇭 ไทย🇮🇷 فارسی🇮🇩 Indonesia


🤔 Why 9Gouter?

Stop wasting money, tokens and hitting limits:

  • ❌ Subscription quota expires unused every month
  • ❌ Rate limits stop you mid-coding
  • ❌ Tool outputs (git diff, grep, ls...) burn tokens fast
  • ❌ Expensive APIs ($20-50/month per provider)
  • ❌ Manual switching between providers

9Gouter solves this:

  • RTK Token Saver - Auto-compress tool_result content, save 20-40% tokens per request
  • Maximize subscriptions - Track quota, use every bit before reset
  • Auto fallback - Subscription → Cheap → Free, zero downtime
  • Multi-account - Round-robin between accounts per provider
  • Universal - Works with Claude Code, Codex, Cursor, Cline, any CLI tool

🔄 How It Works

┌─────────────┐
│  Your CLI   │  (Claude Code, Codex, OpenClaw, Cursor, Cline...)
│   Tool      │
└──────┬──────┘
       │ http://localhost:20127/v1
       ↓
┌─────────────────────────────────────────────┐
│           9Gouter (Smart Router)            │
│  • RTK Token Saver (cut tool_result tokens) │
│  • Format translation (OpenAI ↔ Claude)     │
│  • Quota tracking                           │
│  • Auto token refresh                       │
└──────┬──────────────────────────────────────┘
       │
       ├─→ [Tier 1: SUBSCRIPTION] Claude Code, Codex, GitHub Copilot
       │   ↓ quota exhausted
       ├─→ [Tier 2: CHEAP] GLM ($0.6/1M), MiniMax ($0.2/1M)
       │   ↓ budget limit
       └─→ [Tier 3: FREE] Kiro, OpenCode Free, Vertex ($300 credits)

Result: Never stop coding, minimal cost + 20-40% token savings via RTK

⚡ Quick Start

1. Install globally:

curl -fsSL https://github.com/Artiffusion-Inc/9gouter/releases/latest/download/9gouter-linux-amd64 -o /usr/local/bin/9gouter \bun install -g 9router& chmod +x /usr/local/bin/9gouter
9router

🎉 Dashboard opens at http://localhost:20127

2. Connect a FREE provider (no signup needed):

Dashboard → Providers → Connect Kiro AI (~50 credits/month free: Claude 4.5 + GLM-5 + MiniMax) or OpenCode Free (no auth) → Done!

3. Use in your CLI tool:

Claude Code/Codex/OpenClaw/Cursor/Cline Settings:
  Endpoint: http://localhost:20127/v1
  API Key: [copy from dashboard]
  Model: kr/claude-sonnet-4.5

That's it! Start coding with FREE AI models.

Alternative: run from source (this repository):

This repository package is private (9gouter-dashboard), so source/Docker execution is the expected local development path.

cp .env.example .env
bun install
PORT=20127 NEXT_PUBLIC_BASE_URL=http://localhost:20127 bun run dev

Production mode:

bun run build
PORT=20127 HOSTNAME=0.0.0.0 NEXT_PUBLIC_BASE_URL=http://localhost:20127 bun run start

Default URLs:

  • Dashboard: http://localhost:20127/dashboard
  • OpenAI-compatible API: http://localhost:20127/v1

Video Guides

Tiết kiệm chi phí LLM với 9Gouter
🇻🇳 Tiếng Việt
Tiết kiệm chi phí LLM cho OpenClaw với 9Gouter
by Mì AI

🇵🇰 اردو / हिन्दी
9Gouter + Claude Code FREE Unlimited Setup
by Build AI With Hamid
9Gouter Setup Tutorial
🇺🇸 English
9Gouter + Claude Code FREE Setup
by Build AI With Hamid
9Gouter Setup Tutorial
🇺🇸 English
9Gouter + Claude Code FREE Setup
by Build AI With Hamid
Claude Code FREE Forever
🇺🇸 English
Claude Code FREE Forever — Unlimited Models
by Build AI With Hamid
Claude CLI Free Setup
🇺🇸 English
Claude CLI Free Setup with 9Gouter 🚀
by CodeVerse Soban
Cài đặt OpenClaw Free A-Z
🇻🇳 Tiếng Việt
Cài Đặt OpenClaw Free Từ A-Z + 9Gouter
by Mai Gia
FREE OpenClaw with Claude Opus
🇺🇸 English
FREE OpenClaw + Claude Opus 4.6
by Build AI With Hamid
Claude CLI Free Setup
🇮🇩 Indonesia
Koding 24 Jam Anti Rate Limit! Hemat Token AI 65% | Tutorial Quick Setup 9Gouter 🚀
by Krisswuh

🇮🇩 Indonesia
Cara Deploy 9Gouter di Hugging Face GRATIS Non-Stop! | Alternatif VPS RAM 16GB
by Krisswuh
این شکلی از هر API ای استفاده کن برای هوش مصنوعی
🇮🇷 Persian-فارسی
این شکلی از هر API ای استفاده کن برای هوش مصنوعی
by Matin SenPai

🇻🇳 Tiếng Việt
Hướng Dẫn Setup OpenClaw + 9Gouter: Tạo Bot Zalo AI Tự Động Từ A-Z
by tuanminhhole

🎬 Made a video about 9Gouter? Submit a Pull Request adding your video to this section — we'll merge it!


🛠️ Supported CLI Tools

9Gouter works seamlessly with all major AI coding tools:

Claude Code
Claude-Code
OpenClaw
OpenClaw
Codex
Codex
OpenCode
OpenCode
Cursor
Cursor
Antigravity
Antigravity
Cline
Cline
Continue
Continue
Droid
Droid
Roo
Roo
Copilot
Copilot
Kilo Code
Kilo Code

🌐 Supported Providers

🔐 OAuth Providers
Claude Code
Claude-Code
Antigravity
Antigravity
Codex
Codex
GitHub
GitHub
Cursor
Cursor
Kimchi
Kimchi
🆓 Free Providers
Kiro
Kiro AI
Claude 4.5 + GLM-5 + MiniMax
50 credits/month free
OpenCode Free
OpenCode Free
No auth • Auto-fetch models
Free (model list varies)
Vertex AI
Vertex AI
Gemini 3 Pro + GLM-5 + DeepSeek
$300 credits free

Note: iFlow, Qwen Code and Gemini CLI free tiers were discontinued in 2026. Use Kiro / OpenCode Free / Vertex instead.

Kiro AI moved to a paid model in Sep 2025 — the free tier is now capped at 50 credits/month (plus 500 trial credits for new accounts in the first 30 days). Paid tiers: Pro $20/mo (1,000 credits), Pro+ $40/mo (2,000), Pro Max $100/mo (5,000), Power $200/mo (10,000). OpenCode Free model list fluctuates over time (some models free only for limited promos) — subject to change without notice. Vertex AI: the $300 free credit for new GCP accounts is still valid, but since Mar 2026 the Gemini API endpoint no longer consumes these credits — call the Vertex AI Studio endpoint instead.

🔑 API Key Providers (40+)
OpenRouter
OpenRouter
GLM
GLM
Kimi
Kimi
MiniMax
MiniMax
OpenAI
OpenAI
Anthropic
Anthropic
Gemini
Gemini
DeepSeek
DeepSeek
Groq
Groq
xAI
xAI
Mistral
Mistral
Perplexity
Perplexity
Together
Together AI
Fireworks
Fireworks
Cerebras
Cerebras
Cohere
Cohere
NVIDIA
NVIDIA
SiliconFlow
SiliconFlow

...and 20+ more providers including Nebius, Chutes, Hyperbolic, and custom OpenAI/Anthropic compatible endpoints


💡 Key Features

Feature What It Does Why It Matters
🚀 RTK Token Saver (RTK ⭐40K) Compress tool outputs (git diff, grep, ls, tree...) before sending to LLM Save 20-40% input tokens per request
🧠 Headroom Token Saver (Headroom) Optional external /v1/compress proxy before provider routing Save more context tokens without changing clients
🪨 Caveman Mode (Caveman ⭐52K) Inject caveman-speak prompt → LLM replies terse, technical substance preserved Save up to 65% output tokens
🐴 Ponytail (Ponytail) Inject "lazy senior dev" prompt → LLM writes minimal, YAGNI-first code (Lite/Full/Ultra) Fewer output tokens, less refactoring
🎯 Smart 3-Tier Fallback Auto-route: Subscription → Cheap → Free Never stop coding, zero downtime
📊 Real-Time Quota Tracking Live token count + reset countdown Maximize subscription value
🔄 Format Translation OpenAI ↔ Claude ↔ Gemini ↔ Cursor ↔ Kiro ↔ Vertex Works with any CLI tool
👥 Multi-Account Support Multiple accounts per provider Load balancing + redundancy
🔄 Auto Token Refresh OAuth tokens refresh automatically No manual re-login needed
🎨 Custom Combos Create unlimited model combinations Tailor fallback to your needs
📝 Request Logging Debug mode with full request/response logs Troubleshoot issues easily
💾 Cloud Sync Sync config across devices Same setup everywhere
📊 Usage Analytics Track tokens, cost, trends over time Optimize spending
🌐 Deploy Anywhere Localhost, VPS, Docker, Cloudflare Workers Flexible deployment options

Set X-9Gouter-Token-Saver: off to bypass all token savers for one chat request.

📖 Feature Details
🚀 RTK Token Saver

Tool outputs (git diff, grep, find, ls, tree, log dumps...) often eat 30-50% of your prompt budget. RTK detects them and applies smart, lossless compression before the request hits the LLM:

  • Filters: git-diff, git-status, grep, find, ls, tree, dedup-log, smart-truncate, read-numbered, search-list
  • Auto-detect: No config needed — RTK peeks the first 1KB of each tool_result and picks the right filter.
  • Safe by design: If a filter fails, throws, or makes output bigger, RTK silently keeps the original text. Errors never break your request.
  • Universal: Works across all formats (OpenAI, Claude, Gemini, Cursor, Kiro, OpenAI Responses) because it runs before any format translation.
  • Default ON: Toggle anytime in Dashboard → Endpoint settings.
Without RTK: 47K tokens sent to LLM
With RTK:    28K tokens sent to LLM   (40% saved · same context · same answer)
🧠 Headroom Token Saver

Headroom is optional and runs separately. 9Gouter calls Headroom's local /v1/compress endpoint, then keeps normal routing, fallback, auth, and usage tracking:

Client → 9Gouter → Headroom /v1/compress → 9Gouter → provider

Local setup:

pip install "headroom-ai[proxy]"
headroom proxy --port 8787

Enable in Dashboard → Endpoint → Token Saver → Headroom. Default URL: http://localhost:8787.

Docker examples:

# Headroom service in same Docker network
http://headroom:8787

# Headroom running on host machine
http://host.docker.internal:8787

If Headroom is down or returns an error, 9Gouter fails open and sends the original request.

🐴 Ponytail (Lazy Senior Dev)

Ponytail injects a "lazy senior dev" system prompt into every request, biasing the LLM toward minimal, YAGNI-first code — deletion over addition, stdlib over new deps, one-liners over abstractions. Adapted from DietrichGebert/ponytail.

  • Lite — Build what's asked, name the lazier alternative.
  • Full — YAGNI ladder enforced: stdlib → native → existing deps → one-liner → minimal code.
  • Ultra — YAGNI extremist: deletion first, ship the one-liner, challenge the rest of the requirement in the same response.
Without Ponytail: verbose code, extra abstractions, "just in case" scaffolding
With Ponytail:    shortest working diff, no unrequested abstractions, fewer tokens

Never trades away: input validation, error handling that prevents data loss, security, accessibility, or anything explicitly requested. Enable in Dashboard → Endpoint → Ponytail. Stacks with Caveman (output terseness) and RTK (input compression).

🎯 Smart 3-Tier Fallback

Create combos with automatic fallback:

Combo: "my-coding-stack"
  1. cc/claude-opus-4-6        (your subscription)
  2. glm/glm-4.7               (cheap backup, $0.6/1M)
  3. if/kimi-k2-thinking       (free fallback)

→ Auto switches when quota runs out or errors occur
📊 Real-Time Quota Tracking
  • Token consumption per provider
  • Reset countdown (5-hour, daily, weekly)
  • Cost estimation for paid tiers
  • Monthly spending reports
🔄 Format Translation

Seamless translation between formats:

  • OpenAIClaudeGeminiCursorKiroVertexAntigravityOllamaOpenAI Responses
  • Your CLI tool sends OpenAI format → 9Gouter translates → Provider receives native format
  • Works with any tool that supports custom OpenAI endpoints
👥 Multi-Account Support
  • Add multiple accounts per provider
  • Auto round-robin or priority-based routing
  • Fallback to next account when one hits quota
🔄 Auto Token Refresh
  • OAuth tokens automatically refresh before expiration
  • No manual re-authentication needed
  • Seamless experience across all providers
🎨 Custom Combos
  • Create unlimited model combinations
  • Mix subscription, cheap, and free tiers
  • Name your combos for easy access
  • Share combos across devices with Cloud Sync
📝 Request Logging
  • Enable debug mode for full request/response logs
  • Track API calls, headers, and payloads
  • Troubleshoot integration issues
  • Export logs for analysis
💾 Cloud Sync
  • Sync providers, combos, and settings across devices
  • Automatic background sync
  • Secure encrypted storage
  • Access your setup from anywhere
Cloud Runtime Notes
  • Prefer server-side cloud variables in production:
    • BASE_URL (internal callback URL used by sync scheduler)
    • CLOUD_URL (cloud sync endpoint base)
  • NEXT_PUBLIC_BASE_URL and NEXT_PUBLIC_CLOUD_URL are still supported for compatibility/UI, but server runtime now prioritizes BASE_URL/CLOUD_URL.
  • Cloud sync requests now use timeout + fail-fast behavior to avoid UI hanging when cloud DNS/network is unavailable.
📊 Usage Analytics
  • Track token usage per provider and model
  • Cost estimation and spending trends
  • Monthly reports and insights
  • Optimize your AI spending

💡 IMPORTANT - Understanding Dashboard Costs:

The "cost" displayed in Usage Analytics is for tracking and comparison purposes only. 9Gouter itself never charges you anything. You only pay providers directly (if using paid services).

Example: If your dashboard shows "$290 total cost" while using Kiro free models, this represents what you would have paid using paid APIs directly. Your actual cost = $0 (Kiro free tier: ~50 credits/mo).

Think of it as a "savings tracker" showing how much you're saving by using free models or routing through 9Gouter!

🌐 Deploy Anywhere
  • 💻 Localhost - Default, works offline
  • ☁️ VPS/Cloud - Share across devices
  • 🐳 Docker - One-command deployment
  • 🚀 Cloudflare Workers - Global edge network

💰 Pricing at a Glance

Tier Provider Cost Quota Reset Best For
🚀 TOKEN SAVER RTK (built-in) FREE Always on Save 20-40% tokens on EVERY request
💳 SUBSCRIPTION Claude Code (Pro/Max) $20-200/mo 5h + weekly Already subscribed
Codex (Plus/Pro) $20-200/mo 5h + weekly OpenAI users
GitHub Copilot $10-19/mo Monthly GitHub users
Cursor IDE $20/mo Monthly Cursor users
💰 CHEAP GLM-5.1 / GLM-4.7 $0.6/1M Daily 10AM Budget backup
MiniMax M2.7 $0.2/1M 5-hour rolling Cheapest option
Kimi K2.5 $9/mo flat 10M tokens/mo Predictable cost
🆓 FREE Kiro AI $0 50 credits/mo Claude 4.5 + GLM-5 + MiniMax free (paid tiers above)
OpenCode Free $0 Varies* No auth, auto-fetch models (list changes over time)
Vertex AI $300 credits New GCP accounts Gemini 3 Pro + DeepSeek + GLM-5 (use Vertex AI Studio endpoint for free credits)

💡 Pro Tip: RTK + Kiro AI + OpenCode Free combo = $0 cost + 20-40% token savings!


📊 Understanding 9Gouter Costs & Billing

9Gouter Billing Reality:

9Gouter software = FREE forever (open source, never charges)
Dashboard "costs" = Display/tracking only (not actual bills)
You pay providers directly (subscriptions or API fees)
FREE providers stay FREE (Kiro ~50 credits/mo, OpenCode Free, Vertex $300 credits = $0 within free-tier limits) — note iFlow/Qwen/Gemini CLI free tiers were discontinued in 2026 ❌ 9Gouter never sends invoices or charges your card

How Cost Display Works:

The dashboard shows estimated costs as if you were using paid APIs directly. This is not billing - it's a comparison tool to show your savings.

Example Scenario:

Dashboard Display:
• Total Requests: 1,662
• Total Tokens: 47M
• Display Cost: $290

Reality Check:
• Provider: Kiro (free tier: ~50 credits/mo)
• Actual Payment: $0.00
• What $290 Means: Amount you SAVED by using free models!

Payment Rules:

  • Subscription providers (Claude Code, Codex): Pay them directly via their websites
  • Cheap providers (GLM, MiniMax): Pay them directly, 9Gouter just routes
  • FREE providers (iFlow, Kiro, Qwen): Genuinely free forever, no hidden charges
  • 9Gouter: Never charges anything, ever

🎯 Use Cases

Case 1: "I have Claude Pro subscription"

Problem: Quota expires unused, rate limits during heavy coding

Solution:

Combo: "maximize-claude"
  1. cc/claude-opus-4-7        (use subscription fully)
  2. glm/glm-5.1               (cheap backup when quota out)
  3. kr/claude-sonnet-4.5      (free emergency fallback)

Monthly cost: $20 (subscription) + ~$5 (backup) = $25 total
vs. $20 + hitting limits = frustration
Case 2: "I want zero cost"

Problem: Can't afford subscriptions, need reliable AI coding

Solution:

Combo: "free-forever"
  1. kr/claude-sonnet-4.5      (Claude 4.5 free via Kiro, ~50 credits/mo)
  2. kr/glm-5                  (GLM-5 free via Kiro)
  3. oc/<auto>                 (OpenCode Free, no auth)

Monthly cost: $0
Quality: Production-ready models + RTK saves 20-40% tokens
Case 3: "I need 24/7 coding, no interruptions"

Problem: Deadlines, can't afford downtime

Solution:

Combo: "always-on"
  1. cc/claude-opus-4-7        (best quality)
  2. cx/gpt-5.5                (second subscription)
  3. glm/glm-5.1               (cheap, resets daily)
  4. minimax/MiniMax-M2.7      (cheapest, 5h reset)
  5. kr/claude-sonnet-4.5      (free via Kiro, ~50 credits/mo)

Result: 5 layers of fallback = zero downtime
Monthly cost: $20-200 (subscriptions) + $10-20 (backup)
Case 4: "I want FREE AI in OpenClaw"

Problem: Need AI assistant in messaging apps (WhatsApp, Telegram, Slack...), completely free

Solution:

Combo: "openclaw-free"
  1. kr/claude-sonnet-4.5      (Claude 4.5 free)
  2. kr/glm-5                  (GLM-5 free)
  3. kr/MiniMax-M2.5           (MiniMax free)

Monthly cost: $0
Access via: WhatsApp, Telegram, Slack, Discord, iMessage, Signal...

❓ Frequently Asked Questions

📊 Why does my dashboard show high costs?

The dashboard tracks your token usage and displays estimated costs as if you were using paid APIs directly. This is not actual billing - it's a reference to show how much you're saving by using free models or existing subscriptions through 9Gouter.

Example:

  • Dashboard shows: "$290 total cost"
  • Reality: You're using Kiro free models (~50 credits/mo)
  • Your actual cost: $0.00
  • What $290 means: Amount you saved by using free models instead of paid APIs!

The cost display is a "savings tracker" to help you understand your usage patterns and optimization opportunities.

💳 Will I be charged by 9Gouter?

No. 9Gouter is free, open-source software that runs on your own computer. It never charges you anything.

You only pay:

  • Subscription providers (Claude Code $20/mo, Codex $20-200/mo) → Pay them directly on their websites
  • Cheap providers (GLM, MiniMax) → Pay them directly, 9Gouter just routes your requests
  • 9Gouter itselfNever charges anything, ever

9Gouter is a local proxy/router. It doesn't have your credit card, can't send invoices, and has no billing system. It's completely free software.

🆓 Are FREE providers really unlimited?

Mostly! The current FREE providers (Kiro, OpenCode Free, Vertex) are genuinely free, but free tiers have limits:

These are free services offered by those respective companies:

  • Kiro AI: ~50 credits/month free (plus 500 trial credits for new accounts in the first 30 days) via AWS Builder ID / Google / GitHub OAuth. Paid tiers available above that.
  • OpenCode Free: No-auth passthrough proxy, models auto-fetched from opencode.ai/zen/v1/models. The free model list fluctuates over time (some models free only for limited promos) — subject to change without notice.
  • Vertex AI: $300 free credits for new Google Cloud accounts (90 days). Since Mar 2026 the Gemini API endpoint no longer consumes these credits — use the Vertex AI Studio endpoint instead.

9Gouter just routes your requests to them - there's no "catch" or future billing from 9Gouter itself. They're truly free services, and 9Gouter makes them easy to use with fallback support.

Discontinued free tiers (no longer recommended):

  • iFlow: Was free unlimited, now changed to paid (2026)
  • Qwen Code: Free OAuth tier fully discontinued by Alibaba on 2026-04-15
  • Gemini CLI: Service fully shut down by Google on 2026-06-18 (replaced by the closed-source Antigravity CLI). Discontinued — do not use.
💰 How do I minimize my actual AI costs?

Free-First Strategy:

  1. Start with 100% free combo:

    1. kr/glm-5 (GLM-5 free via Kiro, ~50 credits/mo)
    2. OpenCode Free models (no auth, auto-fetched)
    3. Vertex AI Gemini 3 Pro (using the Vertex AI Studio endpoint with $300 credits)
    

    Cost: $0/month (within Kiro's free credit cap; OpenCode/Vertex subject to their free-tier limits)

  2. Add cheap backup only if you need it:

    4. glm/glm-4.7 ($0.6/1M tokens)
    

    Additional cost: Only pay for what you actually use

  3. Use subscription providers last:

    • Only if you already have them
    • 9Gouter helps maximize their value through quota tracking

Result: Most users can operate at $0/month using only free tiers!

📈 What if my usage suddenly spikes?

9Gouter's smart fallback prevents surprise charges:

Scenario: You're on a coding sprint and blow through your quotas

Without 9Gouter:

  • ❌ Hit rate limit → Work stops → Frustration
  • ❌ Or: Accidentally rack up huge API bills

With 9Gouter:

  • ✅ Subscription hits limit → Auto-fallback to cheap tier
  • ✅ Cheap tier gets expensive → Auto-fallback to free tier
  • ✅ Never stop coding → Predictable costs

You're in control: Set spending limits per provider in dashboard, and 9Gouter respects them.


📖 Setup Guide

🔐 Subscription Providers (Maximize Value)
Claude Code (Pro/Max)
Dashboard → Providers → Connect Claude Code
→ OAuth login → Auto token refresh
→ 5-hour + weekly quota tracking

Models:
  cc/claude-opus-4-7
  cc/claude-opus-4-6
  cc/claude-sonnet-4-6
  cc/claude-haiku-4-5-20251001

Pro Tip: Use Opus for complex tasks, Sonnet for speed. 9Gouter tracks quota per model!

OpenAI Codex (Plus/Pro)
Dashboard → Providers → Connect Codex
→ OAuth login (port 1455)
→ 5-hour + weekly reset

Models:
  cx/gpt-5.5
  cx/gpt-5.4
  cx/gpt-5.3-codex
  cx/gpt-5.2-codex
GitHub Copilot
Dashboard → Providers → Connect GitHub
→ OAuth via GitHub
→ Monthly reset (1st of month)

Models:
  gh/gpt-5.4
  gh/claude-opus-4.7
  gh/claude-sonnet-4.6
  gh/gemini-3.1-pro-preview
  gh/grok-code-fast-1
Cursor IDE
Dashboard → Providers → Connect Cursor
→ OAuth login
→ Monthly subscription

Models:
  cu/claude-4.6-opus-max
  cu/claude-4.5-sonnet-thinking
  cu/gpt-5.3-codex
💰 Cheap Providers (Backup)
GLM-5.1 / GLM-4.7 (Daily reset, $0.6/1M)
  1. Sign up: Zhipu AI
  2. Get API key from Coding Plan
  3. Dashboard → Add API Key:
    • Provider: glm
    • API Key: your-key

Use: glm/glm-5.1, glm/glm-5, glm/glm-4.7

Pro Tip: Coding Plan offers 3× quota at 1/7 cost! Reset daily 10:00 AM.

MiniMax M2.7 (5h reset, $0.20/1M)
  1. Sign up: MiniMax
  2. Get API key
  3. Dashboard → Add API Key

Use: minimax/MiniMax-M2.7, minimax/MiniMax-M2.5

Pro Tip: Cheapest option for long context (1M tokens)!

Kimi K2.5 ($9/month flat)
  1. Subscribe: Moonshot AI
  2. Get API key
  3. Dashboard → Add API Key

Use: kimi/kimi-k2.5, kimi/kimi-k2.5-thinking

Pro Tip: Fixed $9/month for 10M tokens = $0.90/1M effective cost!

🆓 FREE Providers (Recommended)
Kiro AI (Claude 4.5 + GLM-5 + MiniMax FREE)
Dashboard → Connect Kiro
→ AWS Builder ID, AWS IAM Identity Center, Google, or GitHub
→ Unlimited usage

Models:
  kr/claude-sonnet-4.5
  kr/claude-haiku-4.5
  kr/glm-5
  kr/MiniMax-M2.5
  kr/qwen3-coder-next
  kr/deepseek-3.2

Pro Tip: Best free option for Claude. No API key, no payment, fully unlimited.

OpenCode Free (No auth, auto-fetch models)
Dashboard → Connect OpenCode Free
→ No login required (passthrough proxy)
→ Models auto-fetched from opencode.ai/zen/v1/models

Pro Tip: Fastest setup. Just connect and start coding.

Vertex AI ($300 free credits for new GCP accounts)
Dashboard → Connect Vertex AI
→ Upload Google Cloud Service Account JSON
→ Enable Vertex AI API in your GCP project

Models:
  vertex/gemini-3.1-pro-preview
  vertex/gemini-3-flash-preview
  vertex/gemini-2.5-flash

Vertex Partner (Anthropic / DeepSeek / GLM / Qwen via Vertex):
  vertex-partner/glm-5-maas
  vertex-partner/deepseek-v3.2-maas
  vertex-partner/qwen3-next-80b-a3b-thinking-maas

Pro Tip: New Google Cloud accounts get $300 credits free for 90 days. Plenty for daily coding.

🎨 Create Combos
Example 1: Maximize Subscription → Cheap Backup
Dashboard → Combos → Create New

Name: premium-coding
Models:
  1. cc/claude-opus-4-7 (Subscription primary)
  2. glm/glm-5.1 (Cheap backup, $0.6/1M)
  3. minimax/MiniMax-M2.7 (Cheapest fallback, $0.20/1M)

Use in CLI: premium-coding

Monthly cost example (100M tokens):
  80M via Claude (subscription): $0 extra
  15M via GLM: $9
  5M via MiniMax: $1
  Total: $10 + your subscription
Example 2: Free-Only (Zero Cost)
Name: free-combo
Models:
  1. kr/claude-sonnet-4.5 (Claude 4.5 free via Kiro, ~50 credits/mo)
  2. kr/glm-5 (GLM-5 free via Kiro)
  3. vertex/gemini-3.1-pro-preview ($300 free credits)

Cost: $0 forever (+ 20-40% token savings via RTK)!
🔧 CLI Integration
Cursor IDE
Settings → Models → Advanced:
  OpenAI API Base URL: http://localhost:20127/v1
  OpenAI API Key: [from 9router dashboard]
  Model: cc/claude-opus-4-7

Or use combo: premium-coding

Claude Code

Edit ~/.claude/config.json:

{
  "anthropic_api_base": "http://localhost:20127/v1",
  "anthropic_api_key": "your-9router-api-key"
}
Codex CLI
export OPENAI_BASE_URL="http://localhost:20127"
export OPENAI_API_KEY="your-9router-api-key"

codex "your prompt"
OpenClaw

Option 1 — Dashboard (recommended):

Dashboard → CLI Tools → OpenClaw → Select Model → Apply

Option 2 — Manual: Edit ~/.openclaw/openclaw.json:

{
  "agents": {
    "defaults": {
      "model": {
        "primary": "9router/kr/claude-sonnet-4.5"
      }
    }
  },
  "models": {
    "providers": {
      "9router": {
        "baseUrl": "http://127.0.0.1:20127/v1",
        "apiKey": "sk_9router",
        "api": "openai-completions",
        "models": [
          {
            "id": "kr/claude-sonnet-4.5",
            "name": "Claude Sonnet 4.5 (Kiro Free)"
          }
        ]
      }
    }
  }
}

Note: OpenClaw only works with local 9Gouter. Use 127.0.0.1 instead of localhost to avoid IPv6 resolution issues.

Cline / Continue / RooCode
Provider: OpenAI Compatible
Base URL: http://localhost:20127/v1
API Key: [from dashboard]
Model: cc/claude-opus-4-7
🚀 Deployment
VPS Deployment
# Clone and install
git clone https://github.com/Artiffusion-Inc/9gouter.git
cd 9router
bun install
bun run build

# Configure
export JWT_SECRET="your-secure-secret-change-this"
export INITIAL_PASSWORD="your-password"
export DATA_DIR="/var/lib/9router"
export PORT="20127"
export HOSTNAME="0.0.0.0"
export NODE_ENV="production"
export NEXT_PUBLIC_BASE_URL="http://localhost:20127"
export NEXT_PUBLIC_CLOUD_URL="https://9gouter.dev"
export API_KEY_SECRET="endpoint-proxy-api-key-secret"
export MACHINE_ID_SALT="endpoint-proxy-salt"

# Start
bun run start

# Or use PM2
bun install -g pm2
pm2 start npm --name 9router -- start
pm2 save
pm2 startup
Docker

Published images (multi-platform linux/amd64 + linux/arm64):

Quick start (use published image):

docker run -d \
  --name 9router \
  -p 20127:20127 \
  -v "$HOME/.9router:/app/data" \
  -e DATA_DIR=/app/data \
  Artiffusion-Inc/9gouter:latest

→ Open http://localhost:20127

Build from source (dev):

git clone https://github.com/Artiffusion-Inc/9gouter.git
cd 9router/app
docker build -t 9router .
docker run -d --name 9router -p 20127:20127 \
  -v "$HOME/.9router:/app/data" -e DATA_DIR=/app/data 9router

Container defaults:

  • PORT=20127
  • HOSTNAME=0.0.0.0

Useful commands:

docker logs -f 9router
docker restart 9router
docker stop 9router && docker rm 9router
docker pull Artiffusion-Inc/9gouter:latest   # update to latest

Data persistence: $HOME/.9router/db/data.sqlite on host ↔ /app/data/db/data.sqlite in container.

Environment Variables
Variable Default Description
JWT_SECRET Auto-generated (~/.9router/jwt-secret) JWT signing secret for dashboard auth cookie (override to share across instances)
INITIAL_PASSWORD 123456 First login password when no saved hash exists
DATA_DIR ~/.9router Main app data location (SQLite at $DATA_DIR/db/data.sqlite)
PORT framework default Service port (20127 in examples)
HOSTNAME framework default Bind host (Docker defaults to 0.0.0.0)
NODE_ENV runtime default Set production for deploy
BASE_URL http://localhost:20127 Server-side internal base URL used by cloud sync jobs
CLOUD_URL https://9gouter.dev Server-side cloud sync endpoint base URL
NEXT_PUBLIC_BASE_URL http://localhost:3000 Backward-compatible/public base URL (prefer BASE_URL for server runtime)
NEXT_PUBLIC_CLOUD_URL https://9gouter.dev Backward-compatible/public cloud URL (prefer CLOUD_URL for server runtime)
API_KEY_SECRET endpoint-proxy-api-key-secret HMAC secret for generated API keys
MACHINE_ID_SALT endpoint-proxy-salt Salt for stable machine ID hashing
ENABLE_REQUEST_LOGS false Enables request/response logs under logs/
AUTH_COOKIE_SECURE false Force Secure auth cookie (set true behind HTTPS reverse proxy)
REQUIRE_API_KEY false Enforce Bearer API key on /v1/* routes (recommended for internet-exposed deploys)
HTTP_PROXY, HTTPS_PROXY, ALL_PROXY, NO_PROXY empty Optional outbound proxy for upstream provider calls
SEARXNG_URL http://localhost:8888/search Endpoint for the built-in unauthenticated SearXNG web-search provider

Notes:

  • Lowercase proxy variables are also supported: http_proxy, https_proxy, all_proxy, no_proxy.
  • .env is not baked into Docker image (.dockerignore); inject runtime config with --env-file or -e.
  • On Windows, APPDATA can be used for local storage path resolution.
  • INSTANCE_NAME appears in older docs/env templates, but is currently not used at runtime.
Runtime Files and Storage
  • Main app state: ${DATA_DIR}/db/data.sqlite (SQLite — providers, combos, aliases, keys, settings, usage history)
  • Auto backups: ${DATA_DIR}/db/backups/
  • Optional request/translator logs: <repo>/logs/... when ENABLE_REQUEST_LOGS=true
  • Both ${DATA_DIR} and ~/.9router resolve to the same location in a Docker container — the symlink /root/.9router -> /app/data is created at build time.

📊 Available Models

View all available models

Claude Code (cc/) - Pro/Max:

  • cc/claude-opus-4-7
  • cc/claude-opus-4-6
  • cc/claude-sonnet-4-6
  • cc/claude-sonnet-4-5-20250929
  • cc/claude-haiku-4-5-20251001

Codex (cx/) - Plus/Pro:

  • cx/gpt-5.5
  • cx/gpt-5.4
  • cx/gpt-5.3-codex
  • cx/gpt-5.2-codex
  • cx/gpt-5.1-codex-max

GitHub Copilot (gh/):

  • gh/gpt-5.4
  • gh/claude-opus-4.7
  • gh/claude-sonnet-4.6
  • gh/gemini-3.1-pro-preview
  • gh/grok-code-fast-1

Cursor (cu/) - Subscription:

  • cu/claude-4.6-opus-max
  • cu/claude-4.5-sonnet-thinking
  • cu/gpt-5.3-codex
  • cu/kimi-k2.5

GLM (glm/) - $0.6/1M:

  • glm/glm-5.1
  • glm/glm-5
  • glm/glm-4.7

MiniMax (minimax/) - $0.2/1M:

  • minimax/MiniMax-M2.7
  • minimax/MiniMax-M2.5

Kimi (kimi/) - $9/mo flat:

  • kimi/kimi-k2.5
  • kimi/kimi-k2.5-thinking

Kiro (kr/) - Free (~50 credits/month, paid tiers above):

  • kr/claude-sonnet-4.5
  • kr/claude-haiku-4.5
  • kr/glm-5
  • kr/MiniMax-M2.5
  • kr/qwen3-coder-next
  • kr/deepseek-3.2

OpenCode Free (oc/) - FREE no-auth:

  • Auto-fetched from opencode.ai/zen/v1/models

Vertex AI (vertex/) - $300 free credits:

  • vertex/gemini-3.1-pro-preview
  • vertex/gemini-3-flash-preview
  • vertex/gemini-2.5-flash
  • vertex-partner/glm-5-maas
  • vertex-partner/deepseek-v3.2-maas

🐛 Troubleshooting

"Language model did not provide messages"

  • Provider quota exhausted → Check dashboard quota tracker
  • Solution: Use combo fallback or switch to cheaper tier

Rate limiting

  • Subscription quota out → Fallback to GLM/MiniMax
  • Add combo: cc/claude-opus-4-7 → glm/glm-5.1 → kr/claude-sonnet-4.5

OAuth token expired

  • Auto-refreshed by 9Gouter
  • If issues persist: Dashboard → Provider → Reconnect

High costs

  • Enable RTK in Dashboard → Endpoint settings (default ON, saves 20-40% tokens)
  • Check usage stats in Dashboard
  • Switch primary model to GLM/MiniMax
  • Use free tier (Kiro, OpenCode Free, Vertex) for non-critical tasks

Dashboard opens on wrong port

  • Set PORT=20127 and NEXT_PUBLIC_BASE_URL=http://localhost:20127

First login not working

  • Check INITIAL_PASSWORD in .env
  • If unset, fallback password is 123456

No request logs under logs/

  • Set ENABLE_REQUEST_LOGS=true

🛠️ Tech Stack

  • Runtime: Node.js 20+
  • Framework: Next.js 16
  • UI: React 19 + Tailwind CSS 4
  • Database: SQLite (better-sqlite3 / node:sqlite / sql.js fallback)
  • Streaming: Server-Sent Events (SSE)
  • Auth: OAuth 2.0 (PKCE) + JWT + API Keys

📝 API Reference

Chat Completions
POST http://localhost:20127/v1/chat/completions
Authorization: Bearer your-api-key
Content-Type: application/json

{
  "model": "cc/claude-opus-4-6",
  "messages": [
    {"role": "user", "content": "Write a function to..."}
  ],
  "stream": true
}
List Models
GET http://localhost:20127/v1/models
Authorization: Bearer your-api-key

→ Returns all models + combos in OpenAI format

📧 Support


👥 Contributors

Thanks to all contributors who helped make 9Gouter better!

Contributors


📊 Star Chart

Star Chart

🔀 Forks

OmniRoute — A full-featured TypeScript fork of 9Gouter. Adds 36+ providers, 4-tier auto-fallback, multi-modal APIs (images, embeddings, audio, TTS), circuit breaker, semantic cache, LLM evaluations, and a polished dashboard. 368+ unit tests. Available via npm and Docker.


🙏 Acknowledgments

Built on the shoulders of giants:

  • CLIProxyAPI — original Go implementation that inspired this JavaScript port.
  • RTK Stars — Rust token-saver. 9Gouter ports its compression pipeline to JS → −20-40% input tokens on every request.
  • Caveman Stars by @JuliusBrussee — viral "why use many token when few token do trick". 9Gouter adapts its prompt → −65% output tokens.
  • Ponytail Stars by @DietrichGebert"lazy senior dev" skill. 9Gouter injects its YAGNI-first ladder → fewer tokens, less code, shorter diffs.

Huge thanks to these authors — without their work, 9Gouter's token-saving features wouldn't exist. ⭐ them on GitHub!


📄 License

MIT License - see LICENSE for details.


Built with ❤️ for developers who code 24/7

Directories

Path Synopsis
cmd
9gouter command
internal
adapter/auth
Package auth ports dashboardSession.js concepts into pure Go adapters.
Package auth ports dashboardSession.js concepts into pure Go adapters.
adapter/capabilities
Package capabilities ports open-sse/providers/capabilities.js into Go: the model-capability fallback chain used to resolve vision/reasoning/search/ tools, thinking wire format, and context/output token limits per model.
Package capabilities ports open-sse/providers/capabilities.js into Go: the model-capability fallback chain used to resolve vision/reasoning/search/ tools, thinking wire format, and context/output token limits per model.
adapter/mitm
Package mitm implements the MITM TLS interception proxy for the Go rewrite.
Package mitm implements the MITM TLS interception proxy for the Go rewrite.
adapter/oauth
Package oauth ports the static OAuth provider configurations from open-sse/providers/registry/*.js (the `oauth:` blocks) and src/lib/oauth/constants/oauth.js.
Package oauth ports the static OAuth provider configurations from open-sse/providers/registry/*.js (the `oauth:` blocks) and src/lib/oauth/constants/oauth.js.
adapter/paramsupport
Package paramsupport ports open-sse/translator/concerns/paramSupport.js into Go: a config-driven table of per-provider/model request-param rules applied before dispatch so unsupported params do not reach the upstream and trigger HTTP 400.
Package paramsupport ports open-sse/translator/concerns/paramSupport.js into Go: a config-driven table of per-provider/model request-param rules applied before dispatch so unsupported params do not reach the upstream and trigger HTTP 400.
adapter/pricing
Package pricing ports the legacy JS pricing model (open-sse/providers/pricing.js) into Go: canonical per-model rates (MODEL_PRICING), provider-specific overrides (PROVIDER_PRICING), glob-pattern fallbacks (PATTERN_PRICING), and the calculateCostFromTokens formula.
Package pricing ports the legacy JS pricing model (open-sse/providers/pricing.js) into Go: canonical per-model rates (MODEL_PRICING), provider-specific overrides (PROVIDER_PRICING), glob-pattern fallbacks (PATTERN_PRICING), and the calculateCostFromTokens formula.
adapter/provider
Package provider is the provider adapter: registry, base executor wiring, and provider-specific config.
Package provider is the provider adapter: registry, base executor wiring, and provider-specific config.
adapter/provider/antigravity
Package antigravityexec ports the Antigravity executor.
Package antigravityexec ports the Antigravity executor.
adapter/provider/azure
Package azureexec ports the Azure executor.
Package azureexec ports the Azure executor.
adapter/provider/base
Package base ports the BaseExecutor from open-sse/executors/base.js.
Package base ports the BaseExecutor from open-sse/executors/base.js.
adapter/provider/codebuddy
Package codebuddyexec ports the CodeBuddy-CN executor.
Package codebuddyexec ports the CodeBuddy-CN executor.
adapter/provider/codex
Package codexec ports the OpenAI Codex executor.
Package codexec ports the OpenAI Codex executor.
adapter/provider/commandcode
Package commandcodeexec ports the CommandCode executor.
Package commandcodeexec ports the CommandCode executor.
adapter/provider/cursor
agent.go ports the AgentService protobuf codec from open-sse/executors/cursor.js (upstream v0.5.40, commit 6994cd1f "executeAgent" path).
agent.go ports the AgentService protobuf codec from open-sse/executors/cursor.js (upstream v0.5.40, commit 6994cd1f "executeAgent" path).
adapter/provider/default
Package defaultexec ports the DefaultExecutor from open-sse/executors/default.js.
Package defaultexec ports the DefaultExecutor from open-sse/executors/default.js.
adapter/provider/devincli
Package devincli implements the Devin CLI executor: routes completions through the official Devin CLI binary via the Agent Client Protocol (ACP) JSON-RPC 2.0 over stdio.
Package devincli implements the Devin CLI executor: routes completions through the official Devin CLI binary via the Agent Client Protocol (ACP) JSON-RPC 2.0 over stdio.
adapter/provider/embedding
Package embedding ports the per-provider embeddings adapters from open-sse/handlers/embeddingProviders/*.js.
Package embedding ports the per-provider embeddings adapters from open-sse/handlers/embeddingProviders/*.js.
adapter/provider/gemini-cli
Package geminicliexec ports the Gemini CLI executor.
Package geminicliexec ports the Gemini CLI executor.
adapter/provider/github
Package githubexec ports the GitHub Copilot executor.
Package githubexec ports the GitHub Copilot executor.
adapter/provider/grok-cli
Package grokcliexec ports the Grok CLI / Grok Build executor.
Package grokcliexec ports the Grok CLI / Grok Build executor.
adapter/provider/grok-web
Package grokwebexec ports the Grok Web executor.
Package grokwebexec ports the Grok Web executor.
adapter/provider/iflow
Package iflowexec ports the iFlow executor.
Package iflowexec ports the iFlow executor.
adapter/provider/image
Package image ports the static image-generation provider registry from open-sse/handlers/imageProviders/index.js + the per-provider adapters' static transport config (baseUrl, authType, authHeader, format, bodyFields whitelist).
Package image ports the static image-generation provider registry from open-sse/handlers/imageProviders/index.js + the per-provider adapters' static transport config (baseUrl, authType, authHeader, format, bodyFields whitelist).
adapter/provider/kimchi
Package kimchiexec ports the Kimchi executor.
Package kimchiexec ports the Kimchi executor.
adapter/provider/kiro
eventstream.go ports the binary AWS EventStream codec half of open-sse/executors/kiro.js (upstream v0.5.40, commit 7c7fae39).
eventstream.go ports the binary AWS EventStream codec half of open-sse/executors/kiro.js (upstream v0.5.40, commit 7c7fae39).
adapter/provider/mimo-free
Package mimofreeexec ports the MiMo Free executor.
Package mimofreeexec ports the MiMo Free executor.
adapter/provider/ollama-local
Package ollamalocalexec ports the Ollama Local executor.
Package ollamalocalexec ports the Ollama Local executor.
adapter/provider/opencode
Package opencodeexec ports the OpenCode executor.
Package opencodeexec ports the OpenCode executor.
adapter/provider/opencode-go
Package opencodegoexec ports the OpenCode Go executor.
Package opencodegoexec ports the OpenCode Go executor.
adapter/provider/perplexity-web
Package perplexitywebexec ports the Perplexity Web executor.
Package perplexitywebexec ports the Perplexity Web executor.
adapter/provider/projectid
Package projectid ports open-sse/services/projectId.js: it fetches and caches the real Google Cloud Code project ID bound to an authenticated Antigravity / Gemini CLI account, so requests to those providers carry the user's actual project rather than a random one (which Google's anti-abuse system flags).
Package projectid ports open-sse/services/projectId.js: it fetches and caches the real Google Cloud Code project ID bound to an authenticated Antigravity / Gemini CLI account, so requests to those providers carry the user's actual project rather than a random one (which Google's anti-abuse system flags).
adapter/provider/qoder
Package qoderexec ports the Qoder executor.
Package qoderexec ports the Qoder executor.
adapter/provider/qwen
Package qwenexec ports the Qwen Code executor.
Package qwenexec ports the Qwen Code executor.
adapter/provider/resolver
Package resolver ports the per-provider live-model resolvers from open-sse/services/*Models.js.
Package resolver ports the per-provider live-model resolvers from open-sse/services/*Models.js.
adapter/provider/resolver/tokenrefresh
Package tokenrefresh ports the per-provider OAuth/token-refresh functions from open-sse/services/tokenRefresh/providers.js.
Package tokenrefresh ports the per-provider OAuth/token-refresh functions from open-sse/services/tokenRefresh/providers.js.
adapter/provider/search
Package search ports the static web-search provider registry from the legacy JS build:
Package search ports the static web-search provider registry from the legacy JS build:
adapter/provider/stt
Package stt ports the static STT provider registry from open-sse/providers/registry/{openai,groq,deepgram,gemini,assemblyai}.js.
Package stt ports the static STT provider registry from open-sse/providers/registry/{openai,groq,deepgram,gemini,assemblyai}.js.
adapter/provider/tts
Package tts ports the static TTS provider registry from open-sse/providers/registry/*.js ttsConfig blocks.
Package tts ports the static TTS provider registry from open-sse/providers/registry/*.js ttsConfig blocks.
adapter/provider/vertex
Package vertexexec ports the Vertex executor.
Package vertexexec ports the Vertex executor.
adapter/provider/webfetch
Package webfetch ports the per-provider web-fetch adapters from open-sse/handlers/fetch/index.js (runFirecrawl/runJina/runTavily/runExa).
Package webfetch ports the per-provider web-fetch adapters from open-sse/handlers/fetch/index.js (runFirecrawl/runJina/runTavily/runExa).
adapter/provider/xiaomi-tokenplan
Package xiaomitokenplanexec ports the Xiaomi Tokenplan executor.
Package xiaomitokenplanexec ports the Xiaomi Tokenplan executor.
adapter/pxpipe
Package pxpipe implements a Node subprocess bridge for the pxpipe-proxy library.
Package pxpipe implements a Node subprocess bridge for the pxpipe-proxy library.
adapter/thinking
Package thinking ports open-sse/translator/concerns/thinking.js (maps) and the gemini-level branch of thinkingUnified.js's applyFormat (c4f80d30).
Package thinking ports open-sse/translator/concerns/thinking.js (maps) and the gemini-level branch of thinkingUnified.js's applyFormat (c4f80d30).
adapter/translator
Package jsontosse synthesizes an OpenAI chat.completion.chunk SSE stream from a single, complete OpenAI chat-completion JSON body.
Package jsontosse synthesizes an OpenAI chat.completion.chunk SSE stream from a single, complete OpenAI chat-completion JSON body.
adapter/translator/claude
Package claude implements the Claude-to-OpenAI response translator and the OpenAI-to-Claude response translator.
Package claude implements the Claude-to-OpenAI response translator and the OpenAI-to-Claude response translator.
adapter/translator/codex
Package codex implements the codex format translator.
Package codex implements the codex format translator.
adapter/translator/commandcode
Package commandcode implements the CommandCode-to-OpenAI response translator.
Package commandcode implements the CommandCode-to-OpenAI response translator.
adapter/translator/cursor
Package cursor implements the cursor format translator.
Package cursor implements the cursor format translator.
adapter/translator/gemini
Package gemini implements the Gemini-to-OpenAI response translator and the OpenAI-to-Gemini/Gemini-CLI request translators.
Package gemini implements the Gemini-to-OpenAI response translator and the OpenAI-to-Gemini/Gemini-CLI request translators.
adapter/translator/kiro
Package kiro — conversation.go ports open-sse/translator/concerns/ kiroConversation.js (upstream 16cb40fd).
Package kiro — conversation.go ports open-sse/translator/concerns/ kiroConversation.js (upstream 16cb40fd).
adapter/translator/ollama
Package ollama implements the Ollama-to-OpenAI response translator.
Package ollama implements the Ollama-to-OpenAI response translator.
adapter/translator/openai
Package openai implements the OpenAI base format helpers and the OpenAI→Claude request translator.
Package openai implements the OpenAI base format helpers and the OpenAI→Claude request translator.
adapter/translator/register
Package register side-effect imports every translator subpackage so their init() registrations (translator.RegisterRequest / RegisterResponse) run in the final binary.
Package register side-effect imports every translator subpackage so their init() registrations (translator.RegisterRequest / RegisterResponse) run in the final binary.
adapter/translator/shared
Package shared contains helpers used by multiple translator packages.
Package shared contains helpers used by multiple translator packages.
adapter/transport/http
Package http implements the SSE stream pipe used by the /v1 chat pipeline.
Package http implements the SSE stream pipe used by the /v1 chat pipeline.
adapter/transport/http/accountfallback
Package accountfallback ports the JS account-selection error classification and per-model lock logic from open-sse/config/errorConfig.js + open-sse/services/accountFallback.js + src/sse/services/auth.js (markAccountUnavailable / clearAccountError).
Package accountfallback ports the JS account-selection error classification and per-model lock logic from open-sse/config/errorConfig.js + open-sse/services/accountFallback.js + src/sse/services/auth.js (markAccountUnavailable / clearAccountError).
adapter/transport/http/api
Package api implements the dashboard /api routes for the Go rewrite.
Package api implements the dashboard /api routes for the Go rewrite.
adapter/tunnel
Package tunnel implements the Cloudflare quick-tunnel and Tailscale funnel orchestration for the Go rewrite.
Package tunnel implements the Cloudflare quick-tunnel and Tailscale funnel orchestration for the Go rewrite.
app
Package app is the composition root for the 9Gouter Go rewrite.
Package app is the composition root for the 9Gouter Go rewrite.
domain/auth
Package auth defines the session, principal, and OIDC ports used for dashboard authentication.
Package auth defines the session, principal, and OIDC ports used for dashboard authentication.
domain/chat
Package chat defines core chat request/response entities and pass-through types used by the proxy use case.
Package chat defines core chat request/response entities and pass-through types used by the proxy use case.
domain/format
Package format defines the LLM request/response format enum and endpoint-based format detection ported from open-sse/translator/formats.js.
Package format defines the LLM request/response format enum and endpoint-based format detection ported from open-sse/translator/formats.js.
domain/provider
Package provider defines the Executor and Provider ports used to talk to upstream LLM services.
Package provider defines the Executor and Provider ports used to talk to upstream LLM services.
domain/settings
Package settings defines the configuration entities stored in SQLite.
Package settings defines the configuration entities stored in SQLite.
domain/usage
Package usage defines usage recording entities and the repository port.
Package usage defines usage recording entities and the repository port.
usecase/auth
Package auth provides the dashboard authentication usecase.
Package auth provides the dashboard authentication usecase.
usecase/imageproxy
Package imageproxy implements the /v1/images/generations pipeline for the Go rewrite.
Package imageproxy implements the /v1/images/generations pipeline for the Go rewrite.
usecase/managedashboard
Package managedashboard exposes usecase operations for the dashboard /api routes.
Package managedashboard exposes usecase operations for the dashboard /api routes.
usecase/proxychat
Package proxychat — observability gate backed by the settings repo.
Package proxychat — observability gate backed by the settings repo.
usecase/proxyembeddings
Package proxyembeddings implements the /v1/embeddings upstream pipeline for the Go rewrite.
Package proxyembeddings implements the /v1/embeddings upstream pipeline for the Go rewrite.
usecase/proxyfetch
Package proxyfetch implements the /v1/web/fetch upstream pipeline for the Go rewrite.
Package proxyfetch implements the /v1/web/fetch upstream pipeline for the Go rewrite.
usecase/quotafetch
Package quotafetch ports the legacy JS per-provider usage fetchers (open-sse/services/usage/*.js) into Go.
Package quotafetch ports the legacy JS per-provider usage fetchers (open-sse/services/usage/*.js) into Go.
usecase/searchproxy
Package searchproxy implements the /v1/search pipeline for the Go rewrite.
Package searchproxy implements the /v1/search pipeline for the Go rewrite.
usecase/sttproxy
Package sttproxy implements the /v1/audio/transcriptions pipeline for the Go rewrite.
Package sttproxy implements the /v1/audio/transcriptions pipeline for the Go rewrite.
usecase/ttsproxy
Package ttsproxy implements the /v1/audio/speech pipeline for the Go rewrite.
Package ttsproxy implements the /v1/audio/speech pipeline for the Go rewrite.
usecase/videoproxy
Package videoproxy implements the /v1/videos/* surface for the Go rewrite.
Package videoproxy implements the /v1/videos/* surface for the Go rewrite.
tools
import-backup command
shadowdiff command
shadowdiff mirrors traffic to a primary (legacy JS) backend and a shadow (Go) backend, then diffs status, headers, SSE event sequence, and usage rows.
shadowdiff mirrors traffic to a primary (legacy JS) backend and a shadow (Go) backend, then diffs status, headers, SSE event sequence, and usage rows.

Jump to

Keyboard shortcuts

? : This menu
/ : Search site
f or F : Jump to
y or Y : Canonical URL