Knowledge Base Sections ▾
Navigation
▸ Start here By rolesCategories
- Cursor + Gonka AI - cheap LLM for coding
- Claude Code + Gonka AI - LLM for the terminal
- OpenClaw + Gonka AI - affordable AI agents
- OpenCode: your own model in the terminal
- Continue.dev + Gonka AI - AI for VS Code/JetBrains
- Cline + Gonka AI - AI agent in VS Code
- Aider + Gonka AI - pair programming with AI
- LangChain + Gonka AI - AI applications for pennies
- n8n + Gonka AI - automation with cheap AI
- Open WebUI + Gonka AI - your own ChatGPT
- LibreChat + Gonka AI — open-source ChatGPT
- Hermes Agent + DeepSeek on the Gonka network — an autonomous agent for pennies
- Kilo Code + Gonka AI — AI-Agent in VS Code
- Roo Code + Gonka AI — Autonomous AI Agent in VS Code
- LlamaIndex + Gonka AI — RAG applications for pennies
- PydanticAI + Gonka — typed AI agents for pennies
- Vercel AI SDK + Gonka AI — AI applications in TypeScript for pennies
- TanStack AI + Gonka — AI applications in TypeScript for pennies
- API quick start — curl, Python, TypeScript
- JoinGonka Gateway — a full overview
- Management Keys — SaaS on Gonka
- Cheapest AI API: Provider Comparison 2026
- How to buy AI tokens and an API key: 3 methods in 2026
- Cursor Pro request limit reached — breakdown and cheaper alternative
- Claude Code is cheaper — bill breakdown and switching
- Cline is burning money — why the agent spends so much
- OpenClaw is expensive — why the agent burns through tokens and how to save
- OpenRouter: Cheap Alternative — Comparison with JoinGonka Gateway
- Best AI model for coding in 2026: comparison and prices
- Cheap alternative to GitHub Copilot without limits
- A cheap Windsurf alternative without credits or limits
- The cheapest API for AI agents in 2026
- ZCode: Cheap GLM inference instead of GLM Coding Plan
- JetBrains IDE + JoinGonka Gateway — your own endpoint instead of credits
- GitHub Copilot BYOK — own models instead of quotas
- Zed + JoinGonka Gateway — cheap inference in your editor
- Pi + JoinGonka Gateway — terminal agent on cheap inference
- Codex CLI: your own key instead of a subscription
- DeepSeek Harness: Your Own Provider via JoinGonka Gateway
- MiniMax Code: MiniMax agent with your own key via Gonka
- Warp + JoinGonka Gateway — terminal agent on your own endpoint
- Trae + JoinGonka Gateway — Gonka network models in AI-IDE
- Cherry Studio + JoinGonka Gateway — desktop AI client
- omp (Oh My Pi) + JoinGonka Gateway: an agent with model roles
- OpenHands + JoinGonka Gateway: agent on your own endpoint
- Qwen Code after the closure of qwen-oauth: working via JoinGonka Gateway
- Goose + JoinGonka Gateway: your own provider and key in the keyring
- Crush + JoinGonka Gateway: Charm agent on Gonka network models
- Zoo Code + JoinGonka Gateway: Migrating from Roo Code to Gonka models
- Kimi Code CLI: Moonshot AI agent on your key via Gonka
- Factory Droid + JoinGonka Gateway: BYOK on Gonka network models
- MiMo Code + JoinGonka Gateway: Xiaomi agent on Gonka network models
Tools
Zoo Code + JoinGonka Gateway: Migrating from Roo Code to Gonka models
Zoo Code is a VS Code extension where an AI agent reads and edits project files, runs commands, and completes multi-step tasks in Code, Architect, Ask, Debug, and Orchestrator modes. This is the successor to Roo Code: when the Roo team ceased work on the extension — it was officially closed on May 15, 2026 — development was picked up by a group of former Roo contributors, and the closed project's README refers to Zoo Code as a fork started by the Roo Code community. The code is open under the Apache 2.0 license, with weekly version releases: as of September 23, 2026, the stable version is 3.82.2, and the extension has about 94,000 installations in the VS Code Marketplace.
Among Zoo Code's providers is OpenAI Compatible — any endpoint that speaks OpenAI Chat Completions. JoinGonka Gateway is exactly that, so the agent connects to the models of the decentralized Gonka network — DeepSeek V4 Flash, GLM-5.3 Flash, and MiniMax M2.7 — in the standard way and operates at a flat rate of $0.0069 per million input tokens. If you still have Roo Code installed, its configuration is detailed in a separate guide; here you will find the installation of Zoo Code, migration from Roo, profiles for modes, and error troubleshooting.
The form fields and extension behavior have been verified against the Zoo Code 3.82.2 source code, and queries of the same form sent by the extension were processed through the gateway on September 23, 2026. Field names are provided based on the English interface; in the localized interface, the labels will be translated, but the order remains the same. After address confirmation, 3M free tokens will be credited to your account — enough to test the setup on your project.
Quick start: installation and one command
Step 1: install the extension. In VS Code, open the extensions panel (Ctrl+Shift+X, on macOS Cmd+Shift+X), search for "Zoo Code" and install the extension from the publisher Zoo Code Organization. The same in one command:
code --install-extension ZooCodeOrganization.zoo-codeYou need VS Code 1.100 or newer — as stated in the extension manifest; for VSCodium, Windsurf and other editors without the Marketplace, the extension is published on Open VSX. Besides the stable channel (3.82.2 from September 18, 2026), there's a preview 3.83.x — nightly builds of the main branch; for work, use the stable one. After installation, a Zoo Code icon will appear in the activity bar.
Step 2: get a key. Register at gate.joingonka.ai/register, confirm your address and create a key with the prefix jg- in the "API keys" section. It's convenient to create a separate key for Zoo Code: its usage will show up as a separate line in the statistics.
Step 3: run the installer.
npx @joingonka/setup --tool zoo --model deepseekThe installer will ask for the key — it isn't passed in the command-line arguments so it doesn't end up in the shell history. It has nothing to write: provider profiles along with Zoo Code keys are kept in VS Code's secure secrets storage, not in a file. That's why the installer prints ready-to-use values for the form and the real limits of the selected model:
Zoo Code is configured in the VS Code extension UI (no file is written).
Open the Zoo Code settings (gear icon) → API Provider, then enter:
API Provider: OpenAI Compatible
Base URL: https://gate.joingonka.ai/v1
API Key: jg-your-key
Model: deepseek-ai/DeepSeek-V4-Flash-0731
Then set the real limits of this model under Model Configuration
(the defaults do not fit any Gonka model):
Context Window Size: 380000
Max Output Tokens: 32768
Image Support: off
Zoo Code relies on native tool calling only — all our Gonka models support it.Next comes the result of a live request to the gateway — on success, the line ✓ Verified: the gateway accepted the key, base URL and model. The key, address and model are verified immediately, not on the first task. The key in the printed block is shown in full so it's easy to copy — don't publish this output.
Without the --model flag, the installer will substitute MiniMax M2.7: where the model is chosen manually, it suggests the model with the highest capacity in the network. For Zoo Code, we recommend starting with DeepSeek V4 Flash: in our test, it completed the turn after the tool result on the first request — details in the verification section; the flag also accepts the abbreviations glm (GLM-5.3 Flash) and minimax. Without the --tool flag, the installer will show a menu of 25 tools.
One form detail: there's no separate block titled Model Configuration in Zoo Code 3.82.2. The limit fields are in the same form below, under the explanation "Configure the capabilities and pricing for your custom OpenAI-compatible model", and Max Output Tokens comes before Context Window Size.
Manual configuration: OpenAI Compatible provider
Upon the first launch, Zoo Code displays a greeting with a Choose your provider button: it leads to a form where OpenRouter is pre-selected — change the provider to OpenAI Compatible and click Finish at the end. Later, the same form can be opened via the gear icon in the Zoo Code panel header, under the Providers section; the Save button applies changes.
| Field | Value | What is important |
|---|---|---|
| API Provider | OpenAI Compatible | The provider list supports search |
| Base URL | https://gate.joingonka.ai/v1 | Must include /v1: Zoo Code appends the /chat/completions path itself and requests the model list from the same address |
| API Key | your jg-… key | Stored in the secure VS Code secret storage |
| Model | identifier from the table below | The form pre-fills gpt-4o — replace it. Once the address and key are entered, the gateway models will appear in the list; you can enter the ID manually via the Use custom option |
| Max Output Tokens | per table below | Default is -1, "determined by server". This is the space Zoo Code reserves for the response to decide whether to compress the history |
| Context Window Size | per table below | Default is 128000: with this, DeepSeek V4 Flash history will start compressing when less than a third of the actual window is occupied |
| Image Support | disabled | Enabled by default — this allows Zoo Code to permit image uploads, but the network models are text-only |
| Prompt Caching | disabled, as by default | The checkmark adds cache_control tags to requests for Anthropic-style caching |
| Input Price, Output Price | $0.0069 and $0.021 | Price per million tokens, as a number without the dollar sign. Optional: only needed to estimate task cost |
Network model limits:
| Model | Model | Context Window Size | Max Output Tokens |
|---|---|---|---|
| DeepSeek V4 Flash | deepseek-ai/DeepSeek-V4-Flash-0731 | 380000 | 32768 |
| GLM-5.3 Flash | zai-org/GLM-5.3-Flash | 390000 | 8192 |
| MiniMax M2.7 | MiniMaxAI/MiniMax-M2.7 | 200000 | 8192 |
You don't need to touch anything else: leave Enable streaming and Include max output tokens enabled, keep Enable R1 model parameters and Use Azure disabled (these are flags for the R1 family and Azure OpenAI), and leave the Custom Headers and Extra Body fields empty.
Limits are stored in the profile, not tied to the model: if you change the Model in the same profile, the window and output limit will remain the same. Therefore, it is more convenient to create a separate profile for each model — this is covered in the next section.
Migrating from Roo Code, modes and profiles for models
Zoo Code evolved from the Roo Code codebase, so the migration takes a few minutes and project files remain untouched. The official migration guide covers exporting and importing settings; everything else is transferred via a separate button or read from the old paths:
| What | How to migrate | What to consider |
|---|---|---|
| Provider profiles, settings, custom modes | In Roo Code: settings → Export. In Zoo Code: Import Settings on the welcome screen or Settings → About Zoo Code → Import | Importing merges settings: profiles with the same name are updated, others remain. API keys are stored in the file as plain text — delete it after import |
| Task history | Settings → About Zoo Code → Import history from Roo Code | Copies Roo Code tasks from the same VS Code storage; already migrated tasks are not duplicated |
.roomodes, .roo/rules/, .roo/mcp.json | Nothing needed | Zoo Code reads these project files from the same paths |
| Roo Code Router profiles | Migrated without the provider | Roo router has been removed — configure such a profile again, for example on OpenAI Compatible |
A Roo Code profile that was already pointed at JoinGonka Gateway will migrate with the URL, key, model, and limits — just verify that the model is currently available in the network: the list is provided by GET https://gate.joingonka.ai/v1/models.
Modes and Profiles. There are five built-in modes: Code, Debug, Architect (edits only Markdown files), Ask (answers without edits), and Orchestrator (assigns subtasks to other modes). They can be switched via the list next to the input field, by typing a command with the mode name at the beginning of the message (e.g., /architect), or by using the shortcut Ctrl+. (Cmd+. on macOS). Each mode remembers the profile you last used with it, and Zoo Code automatically switches profiles when you change the mode.
This creates a convenient scheme: one profile per model, tied to specific modes. In Settings → Providers, click "+" next to the Configuration Profile field, name the profile (e.g., Gonka DeepSeek), and fill out the form; repeat for GLM-5.3 Flash and MiniMax M2.7 — the URL and key are the same for all. Then, select the profile in the profile list next to the input field in each mode, and that mode will remember it. Conversely, the lock icon in this list pins one profile for all modes in the workspace folder.
| Mode | Profile | Why |
|---|---|---|
| Code, Debug | DeepSeek V4 Flash | 380K context window and 32768 output limit — headroom for long sessions and large edits; see model overview for details |
| Architect, Ask | GLM-5.3 Flash | Reasons before answering, and in planning or explanation, the chain of thought is more important than volume. Reasoning consumes the same 8192 output limit |
| Orchestrator | DeepSeek V4 Flash | Maintains a long plan with subtask reports; subtasks themselves run in other modes and inherit those modes' profiles |
| Fallback profile | MiniMax M2.7 | Largest capacity in the network — the installer suggests it when no model is set via flag. After a tool result in our test, it responded with text and called the completion tool only after a Zoo Code reminder — the task finished a request later. The primary choice for agentic tasks in Zoo Code is DeepSeek V4 Flash |
Reasoning Effort. The Enable Reasoning Effort checkbox in the form opens a list: Low, Medium, High, Extra High, and Max; while it is disabled, the level is not included in the request. For GLM-5.3 Flash, the toggle is essentially binary: Low turns off reasoning, any other level leaves it fully enabled — it is convenient to keep a separate GLM profile with Low for quick questions. For details, see the GLM-5.3 Flash overview. For DeepSeek V4 Flash and MiniMax M2.7, the gateway does not rewrite the scale: how the level works depends on the model itself.
Verification: what should happen
Open a folder with a small file containing an obvious error, for example calc.py with an addition function that subtracts, and in Code mode, type: "Read calc.py and say in one sentence if it has an error". Zoo Code will ask for permission to read the file — "Zoo wants to read this file", with Approve and Deny buttons if reading was not allowed in advance via Auto-Approve. Then the model will call a completion tool, and the task will end with a Task Completed block with the actual answer; reasoning, if any, is shown in a separate Thinking block. The context usage percentage in the collapsed task header is calculated based on the Context Window Size — another reason to set it correctly.
We reproduced this workflow by issuing a request using the Zoo Code pattern — flow, tools in a strict schema, result of read_file in history. DeepSeek V4 Flash responded with a native attempt_completion call with the correct conclusion: the function subtracts instead of adds. The second half of the check is the gateway dashboard: in the "Usage" section, the request will appear broken down by model, and the key's last used time will be updated.
Without Auto-Approve, an error appears as a API Request Failed block with Retry and Start New Task buttons, accompanied by a Provider Error header with a code and a countdown to retry. The gateway's response text is visible in the message or its details:
| What is visible | What it means | What to do |
|---|---|---|
401 Invalid API key | Gateway did not accept the key | Paste the key again, in full and without spaces; check that it is active in the "API keys" section |
400 Model "gpt-4o" not found. Available: … | The Model field still contains the default gpt-4o or an ID with a typo | Select a model from the list; the gateway lists available identifiers directly in the error |
405 and an HTML page instead of a response, list of models is empty | The /v1 suffix is missing from the Base URL | The address must end with /v1 |
404 Invalid URL (POST /v1/v1/chat/completions) | The Base URL has an extra tail | Keep exactly https://gate.joingonka.ai/v1 |
402 Insufficient balance | Funds have run out on the balance | Top up your account in the "Billing" section; the key remains functional |
429, text contains currently overloaded in the Gonka network | The model ran out of free capacity during peak hours | With Auto-Approve, Zoo Code will repeat the request itself, increasing the delay; without it, click Retry. If it persists, switch the profile to a different model; check network status on the status page |
The model answered with text instead of a tool call, and Zoo Code sent another request by itself; after two such responses in a row — The model failed to use any tools in its response | Zoo Code expects a tool call in every turn and, having not received it, reminds the model about it. This happens after a tool result: the model answers the question, and calls attempt_completion after the reminder | A single repeat is normal, the task will finish by itself. If it happens often, switch the profile to DeepSeek V4 Flash |
How much does it cost
An agent spends tokens differently than a chat. In every request, Zoo Code includes a system prompt with mode rules, tool descriptions, and information about the working folder — that's thousands of tokens even before your question — and a task usually takes several turns. Therefore, the price per token is critical here.
Through JoinGonka Gateway, tokens cost $0.0069 per million for input and $0.021 per million for output — the price is the same for all models in the network and is pulled onto this page from a live source. Order of magnitude by prices as of September 2026:
| Scenario | Consumption | Via Gateway |
|---|---|---|
| One-time task: read a file, find an error | tens of thousands of tokens | hundredths of a cent |
| A day of active work with an agent | 3-7M tokens | a few cents |
| A month of daily development | ~150M tokens | around a dollar |
For comparison — how you can pay for models in Zoo Code in general:
| Method | How it is paid | What to consider |
|---|---|---|
| Vendor key: Anthropic, OpenAI, Google Gemini, etc. | Per tokens at vendor price | Price depends on the model, bill grows with session length |
| Aggregator, e.g. OpenRouter — it is the first on the welcome screen | Per tokens at aggregator price | Wide selection of models under one key |
| JoinGonka Gateway | Per tokens from a prepaid balance, one price for all network models | No subscriptions or monthly quotas; consumption visible in the dashboard by day, key, and model |
With Input Price and Output Price filled in, the cost of the task is visible in its header, but only down to the cent — short tasks on network models will show $0.00 there. The exact consumption and balance are in the dashboard, in the "Usage" and "Billing" sections.
What to consider when working
Approvals and autonomy. By default, Auto-Approve is disabled, and Zoo Code asks for permission for reading, writing, and executing commands; for long tasks, enable the necessary categories in the Auto-approve menu near the input field. For commands, there is a Destructive Command Guard (Settings → Auto-Approve): the extension downloads its executable file, after which the commands it approves are executed automatically, while blocked ones wait for confirmation. With Auto-Approve, Zoo Code also retries failed requests automatically: the pause starts at 5 seconds and doubles with each attempt, but does not exceed 10 minutes.
Streaming response. Do not turn off Enable streaming: in stream mode, the gateway delivers text as it is generated and limits the response only by the model's ceiling, whereas a request without streaming and without an explicit limit is truncated at 1500 response tokens (for GLM-5.3 Flash — at 3000).
Code Indexing. Semantic search through the project (Codebase Indexing) is configured separately from the model, via an embedding provider; there is a local option for it, Semble - Local.
Images. The network models are text-based. For screenshots and mockups, create a separate profile with a model that works with images and switch to it for the duration of such a task.
Privacy. Zoo Code sends error and usage data to the developers, linked to an installation ID — without code and prompts; this can be disabled via the Allow error and usage reporting toggle in Settings → About Zoo Code. For its part, the gateway does not store the content of prompts and responses: only request and token counters are visible in the dashboard.
.roomodes and .roo/ files, settings and history import from Roo. It connects to JoinGonka Gateway using the OpenAI Compatible provider: Base URL https://gate.joingonka.ai/v1, key jg-…, network model instead of the default gpt-4o, and honest Context Window Size and Max Output Tokens (DeepSeek V4 Flash — 380000 and 32768, GLM-5.3 Flash — 390000 and 8192, MiniMax M2.7 — 200000 and 8192), Image Support disabled; values are printed by npx @joingonka/setup --tool zoo --model deepseek. Then proceed profile by profile per model: DeepSeek V4 Flash for Code, Debug, and Orchestrator; GLM-5.3 Flash for Architect and Ask.Want to learn more?
Explore other sections or start earning GNK right now.
Get key and free tokens →