Knowledge Base Sections ▾

Navigation

▸ Start here By roles

Categories

Tools 52
Glossary 12

Tools

Cherry Studio + JoinGonka Gateway — desktop AI client

Cherry Studio is an open-source desktop AI client for Windows, macOS, and Linux: chat with assistants, compare model responses, agents, knowledge base, translation, and MCP server integration. The project is distributed under the AGPL-3.0 license and has gathered over 50,000 stars on GitHub (September 2026). The client's main principle is BYOK (Bring Your Own Key): you decide which providers perform the inference, and the application provides a unified interface on top of them.

In addition to the default provider list, Cherry Studio features a Custom Provider — a provider with a custom address. Through this, you can connect JoinGonka Gateway: an OpenAI-compatible API to models in the decentralized Gonka network with pay-per-token billing based on live network prices. After registration and address confirmation, you will receive 3M free tokens to your account — feel free to test the client before your first deposit.

This guide provides step-by-step setup for the current V2 lineup with notes for V1, an explanation of the API Host field and its # character rule, model selection, verification, and common errors. Interface labels are based on the English localization of the application.

Installation, version, and key

Installation. Builds for every system are on the project's releases page: an installer and a portable build for Windows, a dmg image for macOS, AppImage, deb, and rpm for Linux — for both x64 and ARM. Chat history and settings are stored locally, on your machine.

V1 or V2. In August 2026 the V2 line launched with a new data structure and reworked settings; the last build of the previous line is 1.9.13. A custom provider exists in both, but the way you add it and some of the labels differ — those spots are flagged below. When moving from V1 to V2, provider and model settings carry over automatically, the original V1 data is preserved, but new data doesn't sync back, and backups from the two lines aren't compatible. Before upgrading, read the migration section in the Cherry Studio documentation.

JoinGonka key. Register on the gateway, open the "API Keys" section in your dashboard, and click "Create key". The key starts with jg- and is shown only once — save it right away. It's convenient to create a separate key for Cherry Studio: in the "Usage" section its spend will show up as its own line.

Connection: custom provider and API Host field

  1. Click the gear icon to open the settings. Select the Model Provider section (referred to as Model Services in the English documentation).
  2. Below the provider list, click Add Provider.
  3. In V2, the Add Custom Provider dialog opens. Enter a Provider Name — for example, JoinGonka. In the Endpoint settings block, fill in the OpenAI line with the address https://gate.joingonka.ai/v1: a Request path line will appear below the field showing the final path — it must end with /v1/chat/completions. In the API Key field, paste jg-your-key and click Add.
  4. In V1, the dialog is shorter: provider name and Provider Type — select OpenAI. The key and address are entered on the provider page itself, in the API Key and API Host fields.
  5. Make sure the provider is enabled: the toggle is in the top-right corner of its page. In V2, new providers are added disabled, and models from a disabled provider do not appear in selection lists.

How Cherry Studio assembles the address. The API Host field takes the base address, and the app appends the path itself, showing the result in the Preview line below the field. The rules are the same in both lines:

What you enterWhere the request goesComment
https://gate.joingonka.ai/v1https://gate.joingonka.ai/v1/chat/completionsRecommended option: the version is already in the address, only the path is appended
https://gate.joingonka.aiThe same addressNo version in the address — Cherry Studio will add /v1 itself
https://gate.joingonka.ai/v1/The same addressThe trailing slash is simply dropped
https://gate.joingonka.ai/v1/chat/completions#The address as-is, without #The trailing # disables auto-substitution of the version and path. For the gateway this trick is unnecessary — it is needed for services with a non-standard path
https://gate.joingonka.ai#https://gate.joingonka.ai/chat/completionsThe version is not added, the request goes past the API — the gateway will respond with 405

Older instructions mention another rule: a slash at the end of the address supposedly disables /v1. It applied in early versions, up to and including the 1.6 line, and was removed in 1.7: in current builds, only # disables auto-substitution.

Models: list, types, and default model

The provider is added, but it is not visible in the chat yet: Cherry Studio selection lists only include models you have added explicitly. There are two ways to do this.

List from server. On the provider page, click Sync models (in V1 — Manage, then Fetch model list). Cherry Studio will request GET /v1/models from the gateway and display the active network lineup; add the required models using the + button. Manually. The Add Model button opens a form where one field is required — Model ID: the identifier must be entered character for character.

The prices for the network models are the same, so add them all — you will have to choose between them based on behavior rather than budget:

ModelModel IDContext and ResponseWhen to use in Cherry Studio
DeepSeek V4 Flashdeepseek-ai/DeepSeek-V4-Flash-0731380K, response up to 32768 tokensMain assistant model: long documents, large code snippets, detailed answers and translations — the highest response limit in the network
MiniMax M2.7MiniMaxAI/MiniMax-M2.7200K, response up to 8192 tokensFast daily tasks, draft emails, short text edits; suitable as a utility model
GLM-5.3 Flashzai-org/GLM-5.3-Flash390K, response up to 8192 tokensReasoning model: thinks before answering. Solving complex problems, planning, logic checking; the answer arrives later, and part of the response limit is consumed by the reasoning process

Model types. Each model in the list has a settings button where you can mark its capabilities: Vision, Reasoning, Tool, and others. Cherry Studio sets some tags automatically, but you should verify them for gateway models. Vision should be unchecked: the network models are text-only. Tool is needed if you are connecting MCP tools — all models in the table support native tool calls. Reasoning is appropriate for GLM-5.3 Flash, as it is a network reasoning model.

Default model. In the Default Model section of the settings, you specify models for utility roles. Set DeepSeek V4 Flash as the main assistant model. For the fast model that invents dialogue titles and prepares search queries, Cherry Studio documentation recommends a light and fast model, not a reasoning one — out of the network models, MiniMax M2.7 fits this role best; do not use GLM-5.3 Flash here due to its mandatory reasoning. Any of the first two models will work for translation.

Verification and diagnostics

On the provider page in the key block, there is a Model Check button (in V1, it is Check next to the key field). Select a model — Cherry Studio will send it a short test request and show the result along with the response latency. If the check passes, go to the chat, select the gateway model in the model switcher, and ask any question.

To ensure that requests are indeed going through the gateway, open its dashboard, go to the "Usage" section: in the "By Keys" block, the Cherry Studio key will show requests and the time of the last access; in the "By Models" block, the selected model will appear. V2 also has its own client statistics in the Usage settings section: tokens and requests by providers, models, and keys. The cost there is for reference only — Cherry Studio estimates it based on the public prices of the models, not on the Gonka network price, so please check the gateway dashboard for exact amounts.

If something goes wrong, the reason is almost always clear from the response:

What you seeWhat it meansWhat to do
401, Invalid API keyThe gateway did not recognize the keyCheck that the key starts with jg-, is pasted without spaces, and is active in the "API Keys" section
404, Invalid URL (POST …)The path contains an extra segment. A typical case is that the full path without # is entered in the API Host, and Cherry Studio appended /chat/completions a second timeLeave the base address https://gate.joingonka.ai/v1 in the field and check against the Preview string
405 Not Allowed or HTML instead of a responseThe request missed the API: the address ends with #, and /v1 is missingRemove the # or add /v1
429The model is overloaded: at peak hours, the capacity of a specific model in the network may be fully utilizedTry again in a few seconds or switch to a neighboring model; network status is visible on the status page
Check passes, but the model is missing in the chatThe provider is disabled or the model was not added to its listEnable the provider and add the model via Sync models or Add Model
Key is entered, but the chat still doesn't respondThe assistant is running on the default model of another, unconfigured providerSet the gateway model in the Default Model section
For GLM-5.3 Flash, the response is empty or cuts offThe response limit in the assistant settings is too low: reasoning is included in the same limitRemove the Max tokens restriction or set it with a buffer — from 2000 tokens
Model replies that it cannot see the imageNetwork models are text-only: the gateway replaces the image with a text noteUncheck Vision for the model; delegate image tasks to a provider with a vision-enabled model — in Cherry Studio, providers coexist seamlessly

How much it costs

The client itself is free to use; you only pay for the inference. Through the JoinGonka Gateway, tokens cost $0.0069 per million for input and $0.021 per million for output — the price is the same for all models in the network and is pulled onto this page from a live source. Chat consumes tokens significantly more modestly than agent tools, so the approximate scale of costs as of September 2026 is as follows:

ScenarioConsumptionVia Gateway
Dialogue with a dozen replies10-30 thousand tokenshundredths of a cent
Day of intensive work: questions, translations, document analysis0.5-1M tokensless than a cent
Month of daily use15-30M tokenstens of cents

An important disclaimer about long dialogues: with each reply, the client sends the entire history back to the model, so the hundredth reply costs more than the first. The remedy is a new dialogue for a new task; in Cherry Studio, this is a single button.

MethodHow you payWhat you get
Chat service subscriptionFixed monthly amountModels from one vendor in their interface, with plan limits
Cherry Studio + vendor keyBy tokens at the vendor's priceYour own interface and history on your machine; bill grows with volume
Cherry Studio + JoinGonka GatewayBy tokens at the network price, from the gateway balanceOpen network models at a single price; usage by days, keys, and models is visible in the dashboard

What else the combination offers

Comparing models in a single chat. Cherry Studio can ask one question to multiple models at once: select them in the model switcher, and each will respond with a separate request. With a flat rate for all network models, this is the fastest way to understand which tasks are handled well by MiniMax M2.7, where you need a long response from DeepSeek V4 Flash, and where the reasoning of GLM-5.3 Flash pays off — its settings are broken down in the model review.

Second protocol with the same key. The gateway also responds in Anthropic format. In the V2 chat, there is a separate Anthropic line in the Endpoint settings block: enter https://gate.joingonka.ai there, and the Request path line should result in /v1/messages. In V1, the same result is achieved by a separate provider with Provider Type Anthropic and the same address. Under the More options button in V2, there is also an OpenAI Responses line — the gateway also services it, using the same address as the main OpenAI line.

Running terminal agents. On the Code CLI page (referenced as Coding Companion in the documentation), Cherry Studio installs and runs console agents, providing them with the provider and model from its settings; the list of providers is filtered by the protocol required by the specific tool. Therefore, the Anthropic address should be filled in advance if you intend to run Claude Code from here. How to set up the same agent without intermediaries is explained in the Claude Code + Gonka guide.

Knowledge Base. In V2, it works even without an embedding model — using BM25 text search, and with an embedding model, hybrid search is enabled. Gonka network models are generative, so for hybrid mode, use an embedding model from another provider, for example, a local one via Ollama; the gateway model will then answer based on the found fragments.

Need a shared chat for your team in a browser? Cherry Studio is a personal desktop client. For a shared web interface with accounts, check out Open WebUI: it connects to the gateway using the same address and key.

Cherry Studio connects to the JoinGonka Gateway as its own provider: Settings → Model Provider → Add Provider, address https://gate.joingonka.ai/v1 and key jg-. The client appends the /chat/completions path itself and shows the result in the Preview line; the # symbol at the end of the address disables auto-completion and is not needed for the gateway. Next, enable the provider, add models via Sync models, uncheck Vision for them, and verify the connection with the Model Check button. Use DeepSeek V4 Flash as the main model, MiniMax M2.7 as the utility model, and GLM-5.3 Flash for logic tasks; usage can be tracked in the gateway dashboard, under the "Usage" section.

Want to learn more?

Explore other sections or start earning GNK right now.

Get a key and free tokens →