Knowledge Base Sections ▾
Navigation
▸ Start here By rolesCategories
- Cursor + Gonka AI - cheap LLM for coding
- Claude Code + Gonka AI - LLM for the terminal
- OpenClaw + Gonka AI - affordable AI agents
- OpenCode: your own model in the terminal
- Continue.dev + Gonka AI - AI for VS Code/JetBrains
- Cline + Gonka AI - AI agent in VS Code
- Aider + Gonka AI - pair programming with AI
- LangChain + Gonka AI - AI applications for pennies
- n8n + Gonka AI - automation with cheap AI
- Open WebUI + Gonka AI - your own ChatGPT
- LibreChat + Gonka AI — open-source ChatGPT
- Hermes Agent + DeepSeek on the Gonka network — an autonomous agent for pennies
- Kilo Code + Gonka AI — AI-Agent in VS Code
- Roo Code + Gonka AI — Autonomous AI Agent in VS Code
- LlamaIndex + Gonka AI — RAG applications for pennies
- PydanticAI + Gonka — typed AI agents for pennies
- Vercel AI SDK + Gonka AI — AI applications in TypeScript for pennies
- TanStack AI + Gonka — AI applications in TypeScript for pennies
- API quick start — curl, Python, TypeScript
- JoinGonka Gateway — a full overview
- Management Keys — SaaS on Gonka
- Cheapest AI API: Provider Comparison 2026
- How to buy AI tokens and an API key: 3 methods in 2026
- Cursor Pro request limit reached — breakdown and cheaper alternative
- Claude Code is cheaper — bill breakdown and switching
- Cline is burning money — why the agent spends so much
- OpenClaw is expensive — why the agent burns through tokens and how to save
- OpenRouter: Cheap Alternative — Comparison with JoinGonka Gateway
- Best AI model for coding in 2026: comparison and prices
- Cheap alternative to GitHub Copilot without limits
- A cheap Windsurf alternative without credits or limits
- The cheapest API for AI agents in 2026
- ZCode: Cheap GLM inference instead of GLM Coding Plan
- JetBrains IDE + JoinGonka Gateway — your own endpoint instead of credits
- GitHub Copilot BYOK — own models instead of quotas
- Zed + JoinGonka Gateway — cheap inference in your editor
- Pi + JoinGonka Gateway — terminal agent on cheap inference
- Codex CLI: your own key instead of a subscription
- DeepSeek Harness: Your Own Provider via JoinGonka Gateway
- MiniMax Code: MiniMax agent with your own key via Gonka
- Warp + JoinGonka Gateway — terminal agent on your own endpoint
- Trae + JoinGonka Gateway — Gonka network models in AI-IDE
- Cherry Studio + JoinGonka Gateway — desktop AI client
- omp (Oh My Pi) + JoinGonka Gateway: an agent with model roles
- OpenHands + JoinGonka Gateway: agent on your own endpoint
- Qwen Code after the closure of qwen-oauth: working via JoinGonka Gateway
- Goose + JoinGonka Gateway: your own provider and key in the keyring
- Crush + JoinGonka Gateway: Charm agent on Gonka network models
- Zoo Code + JoinGonka Gateway: Migrating from Roo Code to Gonka models
- Kimi Code CLI: Moonshot AI agent on your key via Gonka
- Factory Droid + JoinGonka Gateway: BYOK on Gonka network models
- MiMo Code + JoinGonka Gateway: Xiaomi agent on Gonka network models
Tools
Cherry Studio + JoinGonka Gateway — desktop AI client
Cherry Studio is an open-source desktop AI client for Windows, macOS, and Linux: chat with assistants, compare model responses, agents, knowledge base, translation, and MCP server integration. The project is distributed under the AGPL-3.0 license and has gathered over 50,000 stars on GitHub (September 2026). The client's main principle is BYOK (Bring Your Own Key): you decide which providers perform the inference, and the application provides a unified interface on top of them.
In addition to the default provider list, Cherry Studio features a Custom Provider — a provider with a custom address. Through this, you can connect JoinGonka Gateway: an OpenAI-compatible API to models in the decentralized Gonka network with pay-per-token billing based on live network prices. After registration and address confirmation, you will receive 3M free tokens to your account — feel free to test the client before your first deposit.
This guide provides step-by-step setup for the current V2 lineup with notes for V1, an explanation of the API Host field and its # character rule, model selection, verification, and common errors. Interface labels are based on the English localization of the application.
Installation, version, and key
Installation. Builds for every system are on the project's releases page: an installer and a portable build for Windows, a dmg image for macOS, AppImage, deb, and rpm for Linux — for both x64 and ARM. Chat history and settings are stored locally, on your machine.
V1 or V2. In August 2026 the V2 line launched with a new data structure and reworked settings; the last build of the previous line is 1.9.13. A custom provider exists in both, but the way you add it and some of the labels differ — those spots are flagged below. When moving from V1 to V2, provider and model settings carry over automatically, the original V1 data is preserved, but new data doesn't sync back, and backups from the two lines aren't compatible. Before upgrading, read the migration section in the Cherry Studio documentation.
JoinGonka key. Register on the gateway, open the "API Keys" section in your dashboard, and click "Create key". The key starts with jg- and is shown only once — save it right away. It's convenient to create a separate key for Cherry Studio: in the "Usage" section its spend will show up as its own line.
Connection: custom provider and API Host field
- Click the gear icon to open the settings. Select the Model Provider section (referred to as Model Services in the English documentation).
- Below the provider list, click Add Provider.
- In V2, the Add Custom Provider dialog opens. Enter a Provider Name — for example,
JoinGonka. In the Endpoint settings block, fill in the OpenAI line with the addresshttps://gate.joingonka.ai/v1: a Request path line will appear below the field showing the final path — it must end with/v1/chat/completions. In the API Key field, pastejg-your-keyand click Add. - In V1, the dialog is shorter: provider name and Provider Type — select
OpenAI. The key and address are entered on the provider page itself, in the API Key and API Host fields. - Make sure the provider is enabled: the toggle is in the top-right corner of its page. In V2, new providers are added disabled, and models from a disabled provider do not appear in selection lists.
How Cherry Studio assembles the address. The API Host field takes the base address, and the app appends the path itself, showing the result in the Preview line below the field. The rules are the same in both lines:
| What you enter | Where the request goes | Comment |
|---|---|---|
https://gate.joingonka.ai/v1 | https://gate.joingonka.ai/v1/chat/completions | Recommended option: the version is already in the address, only the path is appended |
https://gate.joingonka.ai | The same address | No version in the address — Cherry Studio will add /v1 itself |
https://gate.joingonka.ai/v1/ | The same address | The trailing slash is simply dropped |
https://gate.joingonka.ai/v1/chat/completions# | The address as-is, without # | The trailing # disables auto-substitution of the version and path. For the gateway this trick is unnecessary — it is needed for services with a non-standard path |
https://gate.joingonka.ai# | https://gate.joingonka.ai/chat/completions | The version is not added, the request goes past the API — the gateway will respond with 405 |
Older instructions mention another rule: a slash at the end of the address supposedly disables /v1. It applied in early versions, up to and including the 1.6 line, and was removed in 1.7: in current builds, only # disables auto-substitution.
Models: list, types, and default model
The provider is added, but it is not visible in the chat yet: Cherry Studio selection lists only include models you have added explicitly. There are two ways to do this.
List from server. On the provider page, click Sync models (in V1 — Manage, then Fetch model list). Cherry Studio will request GET /v1/models from the gateway and display the active network lineup; add the required models using the + button. Manually. The Add Model button opens a form where one field is required — Model ID: the identifier must be entered character for character.
The prices for the network models are the same, so add them all — you will have to choose between them based on behavior rather than budget:
| Model | Model ID | Context and Response | When to use in Cherry Studio |
|---|---|---|---|
| DeepSeek V4 Flash | deepseek-ai/DeepSeek-V4-Flash-0731 | 380K, response up to 32768 tokens | Main assistant model: long documents, large code snippets, detailed answers and translations — the highest response limit in the network |
| MiniMax M2.7 | MiniMaxAI/MiniMax-M2.7 | 200K, response up to 8192 tokens | Fast daily tasks, draft emails, short text edits; suitable as a utility model |
| GLM-5.3 Flash | zai-org/GLM-5.3-Flash | 390K, response up to 8192 tokens | Reasoning model: thinks before answering. Solving complex problems, planning, logic checking; the answer arrives later, and part of the response limit is consumed by the reasoning process |
Model types. Each model in the list has a settings button where you can mark its capabilities: Vision, Reasoning, Tool, and others. Cherry Studio sets some tags automatically, but you should verify them for gateway models. Vision should be unchecked: the network models are text-only. Tool is needed if you are connecting MCP tools — all models in the table support native tool calls. Reasoning is appropriate for GLM-5.3 Flash, as it is a network reasoning model.
Default model. In the Default Model section of the settings, you specify models for utility roles. Set DeepSeek V4 Flash as the main assistant model. For the fast model that invents dialogue titles and prepares search queries, Cherry Studio documentation recommends a light and fast model, not a reasoning one — out of the network models, MiniMax M2.7 fits this role best; do not use GLM-5.3 Flash here due to its mandatory reasoning. Any of the first two models will work for translation.
Verification and diagnostics
On the provider page in the key block, there is a Model Check button (in V1, it is Check next to the key field). Select a model — Cherry Studio will send it a short test request and show the result along with the response latency. If the check passes, go to the chat, select the gateway model in the model switcher, and ask any question.
To ensure that requests are indeed going through the gateway, open its dashboard, go to the "Usage" section: in the "By Keys" block, the Cherry Studio key will show requests and the time of the last access; in the "By Models" block, the selected model will appear. V2 also has its own client statistics in the Usage settings section: tokens and requests by providers, models, and keys. The cost there is for reference only — Cherry Studio estimates it based on the public prices of the models, not on the Gonka network price, so please check the gateway dashboard for exact amounts.
If something goes wrong, the reason is almost always clear from the response:
| What you see | What it means | What to do |
|---|---|---|
401, Invalid API key | The gateway did not recognize the key | Check that the key starts with jg-, is pasted without spaces, and is active in the "API Keys" section |
404, Invalid URL (POST …) | The path contains an extra segment. A typical case is that the full path without # is entered in the API Host, and Cherry Studio appended /chat/completions a second time | Leave the base address https://gate.joingonka.ai/v1 in the field and check against the Preview string |
405 Not Allowed or HTML instead of a response | The request missed the API: the address ends with #, and /v1 is missing | Remove the # or add /v1 |
429 | The model is overloaded: at peak hours, the capacity of a specific model in the network may be fully utilized | Try again in a few seconds or switch to a neighboring model; network status is visible on the status page |
| Check passes, but the model is missing in the chat | The provider is disabled or the model was not added to its list | Enable the provider and add the model via Sync models or Add Model |
| Key is entered, but the chat still doesn't respond | The assistant is running on the default model of another, unconfigured provider | Set the gateway model in the Default Model section |
| For GLM-5.3 Flash, the response is empty or cuts off | The response limit in the assistant settings is too low: reasoning is included in the same limit | Remove the Max tokens restriction or set it with a buffer — from 2000 tokens |
| Model replies that it cannot see the image | Network models are text-only: the gateway replaces the image with a text note | Uncheck Vision for the model; delegate image tasks to a provider with a vision-enabled model — in Cherry Studio, providers coexist seamlessly |
How much it costs
The client itself is free to use; you only pay for the inference. Through the JoinGonka Gateway, tokens cost $0.0069 per million for input and $0.021 per million for output — the price is the same for all models in the network and is pulled onto this page from a live source. Chat consumes tokens significantly more modestly than agent tools, so the approximate scale of costs as of September 2026 is as follows:
| Scenario | Consumption | Via Gateway |
|---|---|---|
| Dialogue with a dozen replies | 10-30 thousand tokens | hundredths of a cent |
| Day of intensive work: questions, translations, document analysis | 0.5-1M tokens | less than a cent |
| Month of daily use | 15-30M tokens | tens of cents |
An important disclaimer about long dialogues: with each reply, the client sends the entire history back to the model, so the hundredth reply costs more than the first. The remedy is a new dialogue for a new task; in Cherry Studio, this is a single button.
| Method | How you pay | What you get |
|---|---|---|
| Chat service subscription | Fixed monthly amount | Models from one vendor in their interface, with plan limits |
| Cherry Studio + vendor key | By tokens at the vendor's price | Your own interface and history on your machine; bill grows with volume |
| Cherry Studio + JoinGonka Gateway | By tokens at the network price, from the gateway balance | Open network models at a single price; usage by days, keys, and models is visible in the dashboard |
What else the combination offers
Comparing models in a single chat. Cherry Studio can ask one question to multiple models at once: select them in the model switcher, and each will respond with a separate request. With a flat rate for all network models, this is the fastest way to understand which tasks are handled well by MiniMax M2.7, where you need a long response from DeepSeek V4 Flash, and where the reasoning of GLM-5.3 Flash pays off — its settings are broken down in the model review.
Second protocol with the same key. The gateway also responds in Anthropic format. In the V2 chat, there is a separate Anthropic line in the Endpoint settings block: enter https://gate.joingonka.ai there, and the Request path line should result in /v1/messages. In V1, the same result is achieved by a separate provider with Provider Type Anthropic and the same address. Under the More options button in V2, there is also an OpenAI Responses line — the gateway also services it, using the same address as the main OpenAI line.
Running terminal agents. On the Code CLI page (referenced as Coding Companion in the documentation), Cherry Studio installs and runs console agents, providing them with the provider and model from its settings; the list of providers is filtered by the protocol required by the specific tool. Therefore, the Anthropic address should be filled in advance if you intend to run Claude Code from here. How to set up the same agent without intermediaries is explained in the Claude Code + Gonka guide.
Knowledge Base. In V2, it works even without an embedding model — using BM25 text search, and with an embedding model, hybrid search is enabled. Gonka network models are generative, so for hybrid mode, use an embedding model from another provider, for example, a local one via Ollama; the gateway model will then answer based on the found fragments.
Need a shared chat for your team in a browser? Cherry Studio is a personal desktop client. For a shared web interface with accounts, check out Open WebUI: it connects to the gateway using the same address and key.
Want to learn more?
Explore other sections or start earning GNK right now.
Get a key and free tokens →