windwhiterain/dsh-llm-quota-retry ↗★ 1

dsh-llm-quota-retry

Keep a DeepSeek Harness agent step alive through an exhausted account quota: after the base provider retry policy gives up on a QUOTA failure, retry the same request once per hour without an attempt limit, and publish the entering, leaving, and current state of that long-term retry to other plugins. 适合需要保证长任务不因临时额度耗尽而中断、需自动重试的无人值守场景。

パッケージ
dsh-llm-quota-retry
互換性
未検証
バージョン
0.2.0
ライセンス
MIT
最終更新
2026/09/30

インストール

$npx -p @deepseek-ai/dsh dsh plugin --profile web add github:windwhiterain/dsh-llm-quota-retry

ドキュメント

README 全文を読む ↗

Configuration

KeyDefaultMeaning
longTermDelayMs3600000Interval between long-term attempts, in milliseconds. Must be a positive finite number no greater than 2147483647.
defaultEnabledtrueThe switch a session starts from before it owns a setting of its own. false makes the whole allowance response opt-in.
poolsnoneNamed, ordered route lists a session may be moved between. See below.
balancesnoneOne balance script per provider, reporting what it has left, plus the apiKeyEnv credential it runs with. See below.
failOpentrueWhether a route nothing reported on counts as usable. false makes an unread provider unusable, so only a route a script vouches for is ever chosen.
exhaustedCooldownMs3600000How long a provider stays marked after a request failed with an exhausted allowance. Must be a positive finite number no greater than 2147483647.

Unknown keys and invalid values throw at activation, so a misconfigured row fails loudly instead of silently retrying on a wrong schedule.

A row that turns the default off makes the allowance response opt-in for the whole deployment; the row below also shortens the interval.

- id: llm-quota-retry
  name: 'dsh-llm-quota-retry'
  config:
    longTermDelayMs: 1800000
    defaultEnabled: false