OpenCode Go Bundles DeepSeek V4 Flash, Qwen3.8 Max, and GLM-5.2 for $10/Month
DeepSeek is about to raise API prices significantly, and OpenCode Go arrives as a fixed-cost alternative that bundles multiple frontier Chinese models under one subscription. For developers already using coding agents, it replaces per-token billing anxiety with a predictable $10/month ceiling while keeping access to the same models.
OpenCode Go is a monthly model subscription built for coding agents. For $5 the first month and $10 thereafter, one API key unlocks DeepSeek V4 Flash, Qwen3.8 Max, GLM-5.2, Kimi K3, MiniMax M3, and Kimi K2.7 Code. The service is separate from OpenCode itself, which remains free and open-source, and the key can also be used with other agents that support custom providers.
DeepSeek V4 Flash's cache-hit pricing of $0.0028 per million tokens makes it particularly cheap inside agent loops that carry large, repeated contexts. A shared bill showed a single Codex request with 156k–190k input tokens costing only $0.0006–$0.0014. The plan does carry soft caps: $12 per 5 hours, $30 per week, and $60 per month, with per-model quotas that vary. Qwen3.8 Max, for instance, is capped at $15/month equivalent and is better reserved for complex planning tasks.
Setup takes minutes. Install OpenCode via curl or npm, run `/connect` and paste the Go API key, then use `/models` to pick or switch models. The key works across terminal, desktop, and IDE environments, and model switching requires no reconfiguration.
OpenCode Go is effectively a model arbitrage play: it buys API access at wholesale rates and resells it as a flat subscription, betting that aggregate usage stays within its caps.
The timing is opportunistic. DeepSeek's announced price hike creates immediate demand for a fixed-cost alternative, and OpenCode Go positions itself as the hedge.
Cache-hit economics make or break the value here. DeepSeek V4 Flash's absurdly low cache-read price is what lets a $10 subscription feel unlimited for coding-agent workloads that recycle context heavily.
The per-model quota system means the subscription is really a DeepSeek V4 Flash plan with occasional access to heavier models, not a uniform all-you-can-eat buffet.
The discussion centers on OpenCode's degraded service quality and the timing of the article. One comment claims the platform's model responses have been diluted to the point of unusability, while the other notes the article arrived too late, as the service had already been altered by the time of posting.
opencode is already hated by everyone, the dilution is so bad it's unusable
Too late to the party. If only you had posted this before the 17th, today's the 19th, it's already been changed...