Knowledge Base Sections ▾
Navigation
▸ Start here By rolesCategories
- Cursor + Gonka AI - cheap LLM for coding
- Claude Code + Gonka AI - LLM for the terminal
- OpenClaw + Gonka AI - affordable AI agents
- OpenCode: your own model in the terminal
- Continue.dev + Gonka AI - AI for VS Code/JetBrains
- Cline + Gonka AI - AI agent in VS Code
- Aider + Gonka AI - pair programming with AI
- LangChain + Gonka AI - AI applications for pennies
- n8n + Gonka AI - automation with cheap AI
- Open WebUI + Gonka AI - your own ChatGPT
- LibreChat + Gonka AI — open-source ChatGPT
- Hermes Agent + DeepSeek on the Gonka network — an autonomous agent for pennies
- Kilo Code + Gonka AI — AI-Agent in VS Code
- Roo Code + Gonka AI — Autonomous AI Agent in VS Code
- LlamaIndex + Gonka AI — RAG applications for pennies
- PydanticAI + Gonka — typed AI agents for pennies
- Vercel AI SDK + Gonka AI — AI applications in TypeScript for pennies
- TanStack AI + Gonka — AI applications in TypeScript for pennies
- API quick start — curl, Python, TypeScript
- JoinGonka Gateway — a full overview
- Management Keys — SaaS on Gonka
- Cheapest AI API: Provider Comparison 2026
- How to buy AI tokens and an API key: 3 methods in 2026
- Cursor Pro request limit reached — breakdown and cheaper alternative
- Claude Code is cheaper — bill breakdown and switching
- Cline is burning money — why the agent spends so much
- OpenClaw is expensive — why the agent burns through tokens and how to save
- OpenRouter: Cheap Alternative — Comparison with JoinGonka Gateway
- Best AI model for coding in 2026: comparison and prices
- Cheap alternative to GitHub Copilot without limits
- A cheap Windsurf alternative without credits or limits
- The cheapest API for AI agents in 2026
- ZCode: Cheap GLM inference instead of GLM Coding Plan
- JetBrains IDE + JoinGonka Gateway — your own endpoint instead of credits
- GitHub Copilot BYOK — own models instead of quotas
- Zed + JoinGonka Gateway — cheap inference in your editor
- Pi + JoinGonka Gateway — terminal agent on cheap inference
- Codex CLI: your own key instead of a subscription
- DeepSeek Harness: Your Own Provider via JoinGonka Gateway
- MiniMax Code: MiniMax agent with your own key via Gonka
- Warp + JoinGonka Gateway — terminal agent on your own endpoint
- Trae + JoinGonka Gateway — Gonka network models in AI-IDE
- Cherry Studio + JoinGonka Gateway — desktop AI client
- omp (Oh My Pi) + JoinGonka Gateway: an agent with model roles
- OpenHands + JoinGonka Gateway: agent on your own endpoint
- Qwen Code after the closure of qwen-oauth: working via JoinGonka Gateway
- Goose + JoinGonka Gateway: your own provider and key in the keyring
- Crush + JoinGonka Gateway: Charm agent on Gonka network models
- Zoo Code + JoinGonka Gateway: Migrating from Roo Code to Gonka models
- Kimi Code CLI: Moonshot AI agent on your key via Gonka
- Factory Droid + JoinGonka Gateway: BYOK on Gonka network models
- MiMo Code + JoinGonka Gateway: Xiaomi agent on Gonka network models
Tools
GitHub Copilot BYOK — own models instead of quotas
BYOK (bring your own key) is an official GitHub Copilot mechanism that allows you to connect your own model provider and run agentic sessions on it, rather than on the built-in GPT and Claude models. Since January 2026, any OpenAI-compatible endpoint has been supported; as of June 23, BYOK is available in the Copilot app itself in public preview on all paid plans, and as of July 14, in the JetBrains version as well. This means you can keep Copilot as the shell and route inference through JoinGonka Gateway—at the prices of the decentralized Gonka network without consuming premium subscription requests. In VS Code, BYOK does not require a Copilot plan or a GitHub account login.
What BYOK offers and when you need it
A Copilot subscription restricts premium requests with monthly limits; agentic sessions, refactoring, and multi-step tasks consume this quota fastest. BYOK shifts these calls to your provider: consumption is billed at your provider's rates and does not deduct from the Copilot premium quota. Keys are stored in the system keychain and cannot be read back by the interface.
BYOK is justified if you are hitting limits, want models with open weights, or want to pay significantly less for inference: the Gonka network offers the same open-source models hundreds of times cheaper than vendor retail prices — current rates are always available in the live gateway price list.
Timeline: what was included and when
| When | What was released |
|---|---|
| November 2025 | BYOK in public preview: Anthropic, OpenAI, xAI, Microsoft Foundry; Agent, Plan, Ask, and Edit modes in VS Code, JetBrains, Eclipse, Xcode |
| January 2026 | Support for any OpenAI-compatible provider + AWS Bedrock and Google AI Studio; models with Responses API |
| June 23, 2026 | BYOK in the GitHub Copilot app — public preview for all paid plans |
| June 18, 2026 | VS Code: BYOK works without GitHub login and without a Copilot plan; usage is billed by the provider and does not consume the Copilot request quota |
| July 14, 2026 | Custom OpenAI-compatible endpoints in Copilot for JetBrains — on all subscription tiers |
Connecting JoinGonka Gateway: step-by-step
You'll need a jg-… key — it's issued when you register on the gateway, along with 3M free starter tokens, no card required. The key and model are set separately in each Copilot client — here are three ways.
VS Code — Custom Endpoint provider. This is the current path: the former "OpenAI Compatible" provider and the github.copilot.chat.customOAIModels setting are now deprecated.
| Step | Action |
|---|---|
| 1. Model editor | Command Palette → Chat: Manage Language Models (or the gear icon in the chat model list) |
| 2. Provider | Add Models → Custom Endpoint; group name — for example, JoinGonka |
| 3. Key | Paste your jg-… — VS Code stores it in its own secure storage, not in a file |
| 4. API type | Chat Completions |
| 5. Models | VS Code will open chatLanguageModels.json — replace the new group's "models" array with the block below and save the file |
| 6. Selection | Pick the model in the chat model list; if it doesn't show up, restart VS Code |
"models": [
{
"id": "deepseek-ai/DeepSeek-V4-Flash-0731",
"name": "DeepSeek V4 Flash (Gonka)",
"url": "https://gate.joingonka.ai/v1/chat/completions",
"toolCalling": true,
"vision": false,
"maxInputTokens": 347232,
"maxOutputTokens": 32768
},
{
"id": "MiniMaxAI/MiniMax-M2.7",
"name": "MiniMax M2.7 (Gonka)",
"url": "https://gate.joingonka.ai/v1/chat/completions",
"toolCalling": true,
"vision": false,
"maxInputTokens": 191808,
"maxOutputTokens": 8192
},
{
"id": "zai-org/GLM-5.3-Flash",
"name": "GLM 5.3 Flash (Gonka)",
"url": "https://gate.joingonka.ai/v1/chat/completions",
"toolCalling": true,
"vision": false,
"maxInputTokens": 381808,
"maxOutputTokens": 8192
}
]maxInputTokens is the model's context window minus the response ceiling: together with maxOutputTokens it defines the window. Without toolCalling: true, the model won't appear in the list when working with agents.
Copilot CLI — environment variables only. This provider has no settings file; values are set in the shell from which you launch copilot:
export COPILOT_PROVIDER_BASE_URL=https://gate.joingonka.ai/v1
export COPILOT_PROVIDER_TYPE=openai
export COPILOT_PROVIDER_API_KEY=jg-your-key
export COPILOT_MODEL=deepseek-ai/DeepSeek-V4-Flash-0731
export COPILOT_PROVIDER_MAX_PROMPT_TOKENS=347232
export COPILOT_PROVIDER_MAX_OUTPUT_TOKENS=32768
copilotOnly Copilot CLI reads the COPILOT_PROVIDER_* names, so this setup won't interfere with your other tools. The CLI requires tool calling and streaming from the model — every model in the network supports both.
JetBrains — GitHub Copilot plugin. Open Copilot Chat → model list → Manage Models, choose a provider with an OpenAI-compatible endpoint, click Add Models and enter the Base URL https://gate.joingonka.ai/v1, your jg-… key and the model ID. Click Save, check the model under your key and save again — it will now appear in the list alongside the built-in ones.
Tip. The npx @joingonka/setup --tool copilot-byok command prints all three sections with your key and the actual limits of the selected model already filled in — including a ready-to-paste "models" array for VS Code — and verifies with a live request that the gateway accepts the key and the model. It writes no files: each Copilot client keeps the key in its own storage.
Which network model to choose
All three network models support streaming and tool calling. deepseek-ai/DeepSeek-V4-Flash-0731 has one of the longest contexts in the network (380K tokens) and the largest output, up to 32768 tokens: use it for large repositories and long agent sessions. MiniMaxAI/MiniMax-M2.7 is universal, with an output up to 8192 tokens. zai-org/GLM-5.3-Flash is Z.ai's reasoning model with the longest context in the network (390K tokens): it reasons before answering, with an output up to 8192 tokens; part of the response budget goes to reasoning, so don't set the output limit too low in the settings.
Public preview limitations — honestly
BYOK covers agent sessions and chat; inline suggestions and semantic project search remain on Copilot models and require a GitHub account — this is a limitation of Copilot itself, not the provider. In Copilot Business and Enterprise, the local BYOK is managed by administrator policy ("Bring Your Own Language Model Key"): if the option is not in the interface, it is disabled at the organization level. Vision tasks do not work through the gateway: network models are text-only. If a tool sends requests in the Responses API format, the gateway will accept them: POST /v1/responses, streaming, and function tools work. We do not store dialogues, so store is ignored, and previous_response_id will return an error — pass the entire history in input.