by continuum-ai-corp · v0.0.1
OrcaRouter is an OpenAI-compatible LLM gateway that routes requests across 40+ upstream providers (OpenAI, Anthropic, Google, DeepSeek, Qwen, Grok, and more) to the cheapest or fastest path, with adaptive routing, fallback chains, and pay-below-list-price billing.
This community listing does not yet include every recommended support, privacy, pricing, and permission disclosure. Review the available package permissions before installing.
Available inside your emploidai workspace after installation.
Available inside your emploidai workspace after installation.
OrcaRouter is an OpenAI-compatible LLM gateway that routes requests across 40+ upstream providers (OpenAI, Anthropic, Google, DeepSeek, Qwen, Grok, and more) to the cheapest or fastest path that can serve the model you asked for. Users pay below list price, and the dashboard tracks how much they save in real time.
After installation, sign up at orcarouter.ai, grab an API key from the console, and set it up under Settings → Model Provider in Dify.
orcarouter/autoThe orcarouter/auto model is a virtual router that picks the best upstream per request. Configure its strategy from the routing console. Available strategies:
| Strategy | Behavior |
|---|---|
cheapest | Lowest-priced upstream that can serve the request (default) |
balanced | Trades off price vs latency vs quality |
quality | Highest-quality upstream |
adaptive | Linear contextual bandit picks among candidates based on per-request features (prompt length, code/math/JSON density, declared max_tokens budget tier, MinHash-LSH similarity to recent traffic) |
gated_adaptive | Layers a task-difficulty score on top of adaptive — mundane prompts restricted to a "weak" model pool, hard prompts to a "strong" pool |
Why this matters
orcarouter/auto endpoint serves cheap summarization and premium code-refactor requests with no client-side dispatch logicextra_body)OrcaRouter supports an OpenAI-compatible extension to specify per-request fallback models. Use the Fallback models and Routing mode parameters in the model node:
| Parameter | Example | Effect |
|---|---|---|
Fallback models (string, JSON array) | ["openai/gpt-4o-mini", "openai/gpt-4o"] | If the primary upstream fails, try these in order |
Routing mode | fallback | Activate the fallback list above |
These are translated to the request body's extra_body: {models, route} key — see the API reference.
Some models (OpenAI o1/o3/gpt-5, Anthropic claude-opus-4.7, DeepSeek deepseek-reasoner/deepseek-r1) expose reasoning controls:
reasoning_effort (high / medium / low / minimal), verbosity, exclude_reasoning_tokensenable_thinking, reasoning_budget (token budget), exclude_reasoning_tokensexclude_reasoning_tokens (the model reasons by default)These map onto the upstream provider's native reasoning protocol — OrcaRouter handles the translation.
Reasoning models do not accept temperature. The Dify UI will show temperature only for non-reasoning models.
All prices are configured at the per-model level in this plugin and reflect the effective price you pay through OrcaRouter (which may be below the upstream's list price). See the live models page or GET https://www.orcarouter.ai/api/pricing for the authoritative source.
This plugin is a thin client for the OrcaRouter API. It only transmits:
Authorization: Bearer header on every requestThe plugin itself does not collect, store, log, or transmit any personal data beyond the request payloads required for the OrcaRouter API call. No telemetry, no analytics, no third-party data sharing happens in plugin code.
OrcaRouter's own privacy practices (request logging, billing data retention, etc.) are governed by the OrcaRouter privacy policy: https://www.orcarouter.ai/privacy.html
Plugin source code: https://github.com/Continuum-AI-Corp/dify-plugin-orcarouter