Keep getting rate limited on Zen #9170

Open
opened 2026-02-16 18:11:48 -05:00 by yindo · 6 comments
Owner

Originally created by @thefrana on GitHub (Feb 12, 2026).

Originally assigned to: @fwang on GitHub.

Description

I am using OpenCode Zen with Kimi-K2.5 but I am getting rate limited even after a few queries. I am not even using the free version but it seems paying for the model does not help with anything. I am reaching up to 8th or more attempt. Any ideas what I should do?

Plugins

None

OpenCode version

1.1.59

Steps to reproduce

  1. Open OpenCode
  2. Start using Zen
  3. Rate limits blocks you

Screenshot and/or share link

No response

Operating System

macOS

Terminal

Ghostty

Originally created by @thefrana on GitHub (Feb 12, 2026). Originally assigned to: @fwang on GitHub. ### Description I am using OpenCode Zen with Kimi-K2.5 but I am getting rate limited even after a few queries. I am not even using the free version but it seems paying for the model does not help with anything. I am reaching up to 8th or more attempt. Any ideas what I should do? ### Plugins None ### OpenCode version 1.1.59 ### Steps to reproduce 1. Open OpenCode 2. Start using Zen 3. Rate limits blocks you ### Screenshot and/or share link _No response_ ### Operating System macOS ### Terminal Ghostty
yindo added the bugzen labels 2026-02-16 18:11:48 -05:00
Author
Owner

@github-actions[bot] commented on GitHub (Feb 12, 2026):

This issue might be a duplicate of existing issues. Please check:

  • #11104: API Rate Limit Not Properly Handled - Model Stops Responding After Few Queries
  • #10404: Big Pickle model loops with "too many requests" error in high mode
  • #11242: Rate limit exceeded error messages with Zen models
  • #1712: Implement exponential back-off when hitting rate limits (related feature request)

These issues describe similar rate limiting problems with OpenCode Zen and models. Reviewing these might help you find existing workarounds or follow-up on the core issue.

@github-actions[bot] commented on GitHub (Feb 12, 2026): This issue might be a duplicate of existing issues. Please check: - #11104: API Rate Limit Not Properly Handled - Model Stops Responding After Few Queries - #10404: Big Pickle model loops with "too many requests" error in high mode - #11242: Rate limit exceeded error messages with Zen models - #1712: Implement exponential back-off when hitting rate limits (related feature request) These issues describe similar rate limiting problems with OpenCode Zen and models. Reviewing these might help you find existing workarounds or follow-up on the core issue.
Author
Owner

@vitobotta commented on GitHub (Feb 12, 2026):

I am also having this super frustrating issue. I am using paid Kimi K2.5, yet it happens quite often even when working on a single task at a time.

Is there any support at all for this service?

@vitobotta commented on GitHub (Feb 12, 2026): I am also having this super frustrating issue. I am using *paid* Kimi K2.5, yet it happens quite often even when working on a single task at a time. Is there any support at all for this service?
Author
Owner

@thefrana commented on GitHub (Feb 12, 2026):

I think the problem is Fireworks (IFAIK used in Zen). The metrics on OpenRouter says they are one of the worst providers. Reliability was 50% for Kimi K2.5. I hope Zen will eventually switch to something more stable because this is unbearable.

@thefrana commented on GitHub (Feb 12, 2026): I think the problem is Fireworks (IFAIK used in Zen). The metrics on OpenRouter says they are one of the worst providers. Reliability was 50% for Kimi K2.5. I hope Zen will eventually switch to something more stable because this is unbearable.
Author
Owner

@vitobotta commented on GitHub (Feb 12, 2026):

I switched to OpenRouter, selected the top 5 performing providers and it's working great. I can work on multiple tasks at the same time without any issues and the speed is basically the same, or at least I can't see any difference so far.

@vitobotta commented on GitHub (Feb 12, 2026): I switched to OpenRouter, selected the top 5 performing providers and it's working great. I can work on multiple tasks at the same time without any issues and the speed is basically the same, or at least I can't see any difference so far.
Author
Owner

@thefrana commented on GitHub (Feb 12, 2026):

Thanks. I also tried OpenRouter and it highly varies which provider it selects.

BTW, I found it very interesting that Fireworks does not even bother to show Kimi K2.5 on their status page: https://status.fireworks.ai/

Apparently, it looks like they resign on updating the status website with newer models or the models have such bad uptime they do not even want to show it.

@thefrana commented on GitHub (Feb 12, 2026): Thanks. I also tried OpenRouter and it highly varies which provider it selects. BTW, I found it very interesting that Fireworks does not even bother to show Kimi K2.5 on their status page: https://status.fireworks.ai/ Apparently, it looks like they resign on updating the status website with newer models or the models have such bad uptime they do not even want to show it.
Author
Owner

@vitobotta commented on GitHub (Feb 12, 2026):

I recommend you only enable the best performing providers in the settings, instead of letting OpenRouter pick any provider available, since some of them are slow.

@vitobotta commented on GitHub (Feb 12, 2026): I recommend you only enable the best performing providers in the settings, instead of letting OpenRouter pick any provider available, since some of them are slow.
Sign in to join this conversation.
1 Participants
Notifications
Due Date
No due date set.
Dependencies

No dependencies set.

Reference: anomalyco/opencode#9170