dsh-provider-rate-limit
by — · jyao-SUSE-power-group/dsh-provider-rate-limit
★ 0 Stars
⑂ 0 Forks
Original language: English
About
Rate-limits LLM requests per provider and per model using a reservation-based token bucket, with queue or reject modes, plus gateway identity rules that rewrite user-agent and inject custom headers.
Similar plugins
Route OpenRouter requests through a configured provider list and quantization cap, injected as provider.only/order with allow_fallbacks, plus provider.quantizations.
★ 0
dsh plugin --profile web add dsh-openrouter-providersLLM request retry plugin for DeepSeek Harness: when a model request fails with a configured HTTP status code, machine code, or provider field=value match, it sleeps for the per-rule delay and re-issue
★ 0
dsh plugin --profile web add dsh-llm-error-retrySub-agent matrix swarm that routes heterogeneous tasks to the most suitable model from an OpenRouter-like gateway plus cfgpu.com/llm/square, dispatches each via in-process subagents or direct LLM call
★ 0
dsh plugin --profile web add dsh-swarm-routerLLM provider adapter that routes model calls through any ACP (Agent Client Protocol) server — Claude Code, Codex, Gemini CLI, Devin and more — with a registry browser, per-server settings UI, lazy aut
★ 0
dsh plugin --profile web add @deepseek-ai/dsh-llm-acpRule-based multi-provider model routing for DeepSeek Harness with pre-first-token failover, cooldown circuit-breaking, usage accounting, and a status API.
★ 0
dsh plugin --profile web add @botton/dsh-model-routerby dylan121322
Adaptive model routing: per-request complexity classification with automatic provider routing.
★ 3
MIT
JavaScript
Aug 14, 2026
dsh plugin --profile web add llm-adaptive