go.mdx 12 KB

123456789101112131415161718192021222324252627282930313233343536373839404142434445464748495051525354555657585960616263646566676869707172737475767778798081828384858687888990919293949596979899100101102103104105106107108109110111112113114115116117118119120121122123124125126127128129130131132133134135136137138139140141142143144145146147148149150151152153154155156157158159160161162163164165166167168169170171172173174175176177178179180181182183184185186187188189190191192193194195196197198199200201202203204205206207208209210211212213214215216217218219220221222223224225226227228229230231232233234235236237238239240241242243244
  1. ---
  2. title: Go
  3. description: Low cost subscription for open coding models.
  4. ---
  5. import config from "../../../config.mjs"
  6. export const console = config.console
  7. export const email = `mailto:${config.email}`
  8. OpenCode Go is a low cost subscription — **$5 for your first month**, then **$10/month** — that gives you reliable access to popular open coding models.
  9. Go works like any other provider in OpenCode. You subscribe to OpenCode Go and
  10. get your API key. It's **completely optional** and you don't need to use it to
  11. use OpenCode.
  12. It is designed primarily for international users, with models hosted in the US, EU, and Singapore for stable global access.
  13. ---
  14. ## Background
  15. Open models have gotten really good. They now reach performance close to
  16. proprietary models for coding tasks. And because many providers can serve them
  17. competitively, they are usually far cheaper.
  18. However, getting reliable, low latency access to them can be difficult. Providers
  19. vary in quality and availability.
  20. :::tip
  21. We tested a select group of models and providers that work well with OpenCode.
  22. :::
  23. To fix this, we did a couple of things:
  24. 1. We tested a select group of open models and talked to their teams about how to
  25. best run them.
  26. 2. We then worked with a few providers to make sure these were being served
  27. correctly.
  28. 3. Finally, we benchmarked the combination of the model/provider and came up
  29. with a list that we feel good recommending.
  30. OpenCode Go gives you access to these models for **$5 for your first month**, then **$10/month**.
  31. ---
  32. ## How it works
  33. OpenCode Go works like any other provider in OpenCode.
  34. 1. You sign in to **<a href={console}>OpenCode Zen</a>**, subscribe to Go, and
  35. copy your API key.
  36. 2. You run the `/connect` command in the TUI, select `OpenCode Go`, and paste
  37. your API key.
  38. 3. Run `/models` in the TUI to see the list of models available through Go.
  39. :::note
  40. Only one member per workspace can subscribe to OpenCode Go.
  41. :::
  42. The current list of models includes:
  43. - **Grok 4.5**
  44. - **GLM-5.2**
  45. - **GLM-5.1**
  46. - **GPT 5.6 Luna**
  47. - **Kimi K3**
  48. - **Kimi K2.7 Code**
  49. - **Kimi K2.6**
  50. - **MiMo-V2.5**
  51. - **MiMo-V2.5-Pro**
  52. - **MiniMax M3**
  53. - **MiniMax M2.7**
  54. - **Qwen3.7 Max**
  55. - **Qwen3.7 Plus**
  56. - **Qwen3.6 Plus**
  57. - **DeepSeek V4 Pro**
  58. - **DeepSeek V4 Flash**
  59. - **Hy3**
  60. The list of models may change as we test and add new ones.
  61. ---
  62. ## Usage limits
  63. OpenCode Go includes the following limits:
  64. - **5 hour limit** — $12 of usage
  65. - **Weekly limit** — $30 of usage
  66. - **Monthly limit** — $60 of usage
  67. Limits are defined in dollar value. This means your actual request count depends on the model you use. Cheaper models like DeepSeek V4 Flash allow for more requests, while higher-cost models like GLM-5.2 allow for fewer.
  68. The table below provides an estimated request count based on typical Go usage patterns:
  69. | Model | requests per 5 hour | requests per week | requests per month |
  70. | ----------------- | ------------------- | ----------------- | ------------------ |
  71. | Grok 4.5 | 120 | 300 | 600 |
  72. | GLM-5.2 | 880 | 2,150 | 4,300 |
  73. | GLM-5.1 | 880 | 2,150 | 4,300 |
  74. | GPT 5.6 Luna | 2,050 | 5,100 | 10,250 |
  75. | Kimi K3 | 110 | 250 | 490 |
  76. | Kimi K2.7 Code | 1,350 | 3,380 | 6,750 |
  77. | Kimi K2.6 | 1,150 | 2,880 | 5,750 |
  78. | MiMo-V2.5 | 30,100 | 75,200 | 150,400 |
  79. | MiMo-V2.5-Pro | 3,250 | 8,150 | 16,300 |
  80. | MiniMax M3 | 3,200 | 8,000 | 16,000 |
  81. | MiniMax M2.7 | 3,400 | 8,500 | 17,000 |
  82. | Qwen3.7 Max | 950 | 2,390 | 4,770 |
  83. | Qwen3.7 Plus | 4,300 | 10,800 | 21,600 |
  84. | Qwen3.6 Plus | 3,300 | 8,200 | 16,300 |
  85. | DeepSeek V4 Pro | 3,450 | 8,550 | 17,150 |
  86. | DeepSeek V4 Flash | 31,650 | 79,050 | 158,150 |
  87. | Hy3 | 4,300 | 10,750 | 21,500 |
  88. The estimates are based on observed request patterns:
  89. - Grok 4.5 — 1,100 input, 71,500 cached, 220 output tokens per request
  90. - GLM-5.2/5.1 — 700 input, 52,000 cached, 150 output tokens per request
  91. - GPT 5.6 Luna — 1,000 input, 50,000 cached, 220 output tokens per request
  92. - Kimi K3 — 1,050 input, 76,500 cached, 300 output tokens per request
  93. - Kimi K2.7/K2.6 — 870 input, 55,000 cached, 200 output tokens per request
  94. - DeepSeek V4 Pro — 750 input, 82,000 cached, 290 output tokens per request
  95. - DeepSeek V4 Flash — 790 input, 68,000 cached, 280 output tokens per request
  96. - MiniMax M3 — 510 input, 56,000 cached, 190 output tokens per request
  97. - MiniMax M2.7 — 300 input, 55,000 cached, 125 output tokens per request
  98. - MiMo-V2.5 — 830 input, 71,500 cached, 295 output tokens per request
  99. - MiMo-V2.5-Pro — 790 input, 86,000 cached, 305 output tokens per request
  100. - Qwen3.7 Max — 420 input, 66,000 cached, 200 output tokens per request
  101. - Qwen3.7 Plus — 500 input, 57,000 cached, 190 output tokens per request
  102. - Qwen3.6 Plus — 500 input, 57,000 cached, 190 output tokens per request
  103. - Hy3 — 830 input, 71,500 cached, 295 output tokens per request
  104. The estimates are also based on the following prices per 1M tokens and the monthly usage included with each model:
  105. | Model | Input | Output | Cached Read | Cached Write | Usage |
  106. | ---------------------------- | ------ | ------ | ----------- | ------------ | ----- |
  107. | Grok 4.5 | $2.00 | $6.00 | $0.30 | - | $15 |
  108. | GLM-5.2 | $1.40 | $4.40 | $0.26 | - | $60 |
  109. | GLM-5.1 | $1.40 | $4.40 | $0.26 | - | $60 |
  110. | GPT 5.6 Luna (≤ 272K tokens) | $0.20 | $1.20 | $0.02 | $0.25 | $15 |
  111. | GPT 5.6 Luna (> 272K tokens) | $0.40 | $1.80 | $0.04 | $0.50 | $15 |
  112. | Kimi K3 | $3.00 | $15.00 | $0.30 | - | $15 |
  113. | Kimi K2.7 Code | $0.95 | $4.00 | $0.19 | - | $60 |
  114. | Kimi K2.6 | $0.95 | $4.00 | $0.16 | - | $60 |
  115. | MiMo V2.5 | $0.14 | $0.28 | $0.0028 | - | $60 |
  116. | MiMo V2.5 Pro | $0.435 | $0.87 | $0.003625 | - | $15 |
  117. | MiniMax M3 | $0.30 | $1.20 | $0.06 | - | $60 |
  118. | MiniMax M2.7 | $0.30 | $1.20 | $0.06 | $0.375 | $60 |
  119. | MiniMax M2.5 | $0.30 | $1.20 | $0.06 | $0.375 | $60 |
  120. | Qwen3.7 Max | $2.50 | $7.50 | $0.50 | $3.125 | $60 |
  121. | Qwen3.7 Plus (≤ 256K tokens) | $0.40 | $1.60 | $0.04 | $0.50 | $60 |
  122. | Qwen3.7 Plus (> 256K tokens) | $1.20 | $4.80 | $0.12 | $1.50 | $60 |
  123. | Qwen3.6 Plus (≤ 256K tokens) | $0.50 | $3.00 | $0.05 | $0.625 | $60 |
  124. | Qwen3.6 Plus (> 256K tokens) | $2.00 | $6.00 | $0.20 | $2.50 | $60 |
  125. | DeepSeek V4 Pro | $0.435 | $0.87 | $0.003625 | - | $15 |
  126. | DeepSeek V4 Flash | $0.14 | $0.28 | $0.0028 | - | $60 |
  127. | Hy3 | $0.14 | $0.58 | $0.035 | - | $60 |
  128. You can track your current usage in the **<a href={console}>console</a>**.
  129. :::tip
  130. If you reach the usage limit, you can continue using the free models.
  131. :::
  132. Usage limits may change as we learn from early usage and feedback.
  133. ---
  134. ### Usage beyond limits
  135. If you also have credits on your Zen balance, you can enable the **Use balance**
  136. option in the console. When enabled, Go will fall back to your Zen balance
  137. after you've reached your usage limits instead of blocking requests.
  138. ---
  139. ### Why some models have lower usage
  140. With Go, you pay $10/month and we aim to give you 6x that in usage.
  141. For most models, we make this work through bulk discounts and reserved GPU capacity. We then pass those savings on to you through the 6x multiplier.
  142. For some models, we haven't had the opportunity to negotiate a discount or host them at a lower cost, either because the model is new or because their public pricing is already discounted.
  143. For these models, you still get a little more than if you paid the model providers directly; this is why their usage mulitplier is lower in the table above.
  144. ---
  145. ## Endpoints
  146. You can also access Go models through the following API endpoints.
  147. | Model | Model ID | Endpoint | AI SDK Package |
  148. | ----------------- | ----------------- | ------------------------------------------------ | --------------------------- |
  149. | Grok 4.5 | grok-4.5 | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  150. | GLM-5.2 | glm-5.2 | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  151. | GLM-5.1 | glm-5.1 | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  152. | GPT 5.6 Luna | gpt-5.6-luna | `https://opencode.ai/zen/go/v1/responses` | `@ai-sdk/openai` |
  153. | Kimi K3 | kimi-k3 | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  154. | Kimi K2.7 Code | kimi-k2.7-code | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  155. | Kimi K2.6 | kimi-k2.6 | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  156. | DeepSeek V4 Pro | deepseek-v4-pro | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  157. | DeepSeek V4 Flash | deepseek-v4-flash | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  158. | MiMo-V2.5 | mimo-v2.5 | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  159. | MiMo-V2.5-Pro | mimo-v2.5-pro | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  160. | MiniMax M3 | minimax-m3 | `https://opencode.ai/zen/go/v1/messages` | `@ai-sdk/anthropic` |
  161. | MiniMax M2.7 | minimax-m2.7 | `https://opencode.ai/zen/go/v1/messages` | `@ai-sdk/anthropic` |
  162. | MiniMax M2.5 | minimax-m2.5 | `https://opencode.ai/zen/go/v1/messages` | `@ai-sdk/anthropic` |
  163. | Qwen3.7 Max | qwen3.7-max | `https://opencode.ai/zen/go/v1/messages` | `@ai-sdk/anthropic` |
  164. | Qwen3.7 Plus | qwen3.7-plus | `https://opencode.ai/zen/go/v1/messages` | `@ai-sdk/anthropic` |
  165. | Qwen3.6 Plus | qwen3.6-plus | `https://opencode.ai/zen/go/v1/messages` | `@ai-sdk/anthropic` |
  166. | Hy3 | hy3 | `https://opencode.ai/zen/go/v1/chat/completions` | `@ai-sdk/openai-compatible` |
  167. The [model id](/docs/config/#models) in your OpenCode config
  168. uses the format `opencode-go/<model-id>`. For example, for Kimi K3, you would
  169. use `opencode-go/kimi-k3` in your config.
  170. ---
  171. ### Models
  172. You can fetch the full list of available models and their metadata from:
  173. ```
  174. https://opencode.ai/zen/go/v1/models
  175. ```
  176. ---
  177. ## Privacy
  178. The plan is designed primarily for international users, with models hosted in the US, EU, and Singapore for stable global access. Our providers follow a zero-retention policy and do not use your data for model training.
  179. ---
  180. ## Goals
  181. We created OpenCode Go to:
  182. 1. Make AI coding **accessible** to more people with a low cost subscription.
  183. 2. Provide **reliable** access to the best open coding models.
  184. 3. Curate models that are **tested and benchmarked** for coding agent use.
  185. 4. Have **no lock-in** by allowing you to use any other provider with OpenCode as well.