mirror of
https://github.com/diegosouzapw/OmniRoute.git
synced 2026-07-26 09:52:11 +03:00
* chore(release): open v3.8.13 development cycle Bump 3.8.12 → 3.8.13 across package.json, lockfile, electron/, open-sse/, and docs/reference/openapi.yaml; add the [3.8.13] cycle placeholder to the root CHANGELOG and the 41 i18n mirrors. Integration branch for the v3.8.13 cycle — fixes/features land here via per-issue PRs and it merges to main at release time. * fix(ci): skip auto-deploy when VPS host is unreachable from the runner (#3299) Integrated into release/v3.8.13 * fix(dev): auto-rebuild better-sqlite3 on Node ABI mismatch at dev startup (#3301) Integrated into release/v3.8.13 * feat(api): accept path-scoped API keys on client API routes (#3300) Integrated into release/v3.8.13 * fix(sse): harden against empty responses causing Copilot Chat failures (#3297) Integrated into release/v3.8.13 * fix(api): remove Completions.me rickroll provider (discussion #3293) (#3302) Integrated into release/v3.8.13 * fix(opencode-provider): extract contextLength from live model catalog (#3298) Integrated into release/v3.8.13 * feat(web-cookie): self-service login infrastructure + auto-refresh daemon (#3292) Integrated into release/v3.8.13 * docs(changelog): record the v3.8.13 PRs merged this round (#3292/#3300/#3297/#3298/#3301/#3302/#3299) * fix(auth): harden URL token extraction — drop query-string fallback, gate to client routes (security follow-up to #3300) (#3309) Security follow-up to #3300 — integrated into release/v3.8.13 * docs: rename resolve-issues → review-issues skill references * fix(dashboard): keep no-auth providers visible under 'Show configured only' (#3290) (#3312) no-auth providers (opencode, duckduckgo-web, theoldllm, veoaifree-web) never create a DB connection row so stats.total stays 0, which the configured-only filter treated as 'unconfigured' and hid them — even though they are always usable and appear unconditionally in /v1/models. filterConfiguredProviderEntries now treats displayAuthType === 'no-auth' as configured. Co-authored-by: uniQta <uniQta@users.noreply.github.com> * fix(cli): resolve update paths relative to script + recursive backup (#3295) (#3313) omniroute update always failed on a global install: - getCurrentVersion() read package.json from process.cwd(), which on a global npm/brew install is the user's working dir, not the package root → null → 'Could not determine current version'. - createBackup() resolved bin/ from cwd too, and passed the 'cli' directory to copyFileSync → EISDIR, swallowed by the catch → 'Failed to create backup'. Both now resolve package.json/bin relative to the script via import.meta.url, and the backup uses cpSync({recursive:true}) so the cli/ directory is copied. Co-authored-by: uniQta <uniQta@users.noreply.github.com> * fix(theoldllm): read upstream body once to avoid [502] body-already-read (#3296) (#3314) On the cached-token path the executor never enters the refresh branch, so the same upstream Response was read with .text() twice (token-rejection check + final body). A Response body is single-use, so the second read threw 'Body is unusable: Body has already been read', caught and surfaced as [502]. Read the body once into finalBody and only re-read after a token-rejection refetch. Co-authored-by: onizukashonan14-png <onizukashonan14-png@users.noreply.github.com> * fix(sse): strip leaked internal tool envelopes from streaming output (#3311) Integrated into release/v3.8.13 * fix(sse): expose Claude + Gemini budget tiers in the antigravity catalog (#3184) (#3303) Integrated into release/v3.8.13 (#3184) * fix(catalog): compute combo context_length from known targets only (#3304) Integrated into release/v3.8.13 — live contextLength + known-targets combo context (#3298 follow-up) * chore(i18n): add message keys for proxy UI + vscode/ollama endpoint (#3307) Integrated into release/v3.8.13 — i18n message keys for proxy UI + vscode/ollama * feat(dashboard): i18n the proxy settings UI (#3310) Integrated into release/v3.8.13 — i18n the proxy settings UI * feat(api): model catalog enrichment + MCP model-catalog tools (#3306) Integrated into release/v3.8.13 — model catalog enrichment + MCP model-catalog tools, reconciled with #3309 URL-token hardening * test(catalog): align Antigravity preview-alias test with #3303 budget tiers #3303 added the Gemini `-high`/`-low` budget tiers to ANTIGRAVITY_PUBLIC_MODELS (user-callable on the Antigravity OAuth backend, verified via #3184), but did not update the catalog-route test that asserted `antigravity/gemini-3.1-pro-high` must NOT be exposed. The assertion now reflects the intended behavior — the client-visible budget alias IS surfaced — while keeping the legacy `gemini-claude-*` alias keys unexposed. Caught running the full catalog suite on the merged release HEAD (the #3303 round only ran the antigravity-aliases and usage-hardening files). * docs(changelog): record the 6 PRs merged this review round into v3.8.13 #3306/#3307/#3310 (New Features — VS Code split: catalog+MCP, i18n keys, proxy UI i18n), #3311/#3303/#3304 (Bug Fixes — SSE envelope sanitizer, antigravity budget tiers, combo known-targets context_length). * chore(release): finalize v3.8.13 changelog and cleanup Finalize the v3.8.13 changelog with release date, maintenance notes, and contributor credits. Update MCP docs to reference the correct tool inventory diagram, exclude nested .claude worktrees from ESLint scans, and tighten a response sanitizer type guard. * fix(dashboard): refresh connections after provider auth import (#3320) Integrated into release/v3.8.13 — refresh connections after provider auth import * fix(codex): strip client-only params on native /responses passthrough (#3317) (#3325) A /v1/responses request against the built-in codex/ provider does an openai-responses -> openai-responses passthrough (CodexExecutor.transformRequest returns the body early for _nativeCodexPassthrough). It forwarded client-only fields verbatim and the Codex upstream rejected them with 400 Unsupported parameter: prompt_cache_retention / safety_identifier / user — breaking Factory Droid (which injects all three). The chat-completions path already strips these (base.ts #1884, openai-responses translator #2770) but the passthrough skips translation. Strip the three fields in the shared block before the passthrough return; user is removed unconditionally since Codex /responses always rejects it. Co-authored-by: tycronk20 <tycronk20@users.noreply.github.com> * fix(dashboard): normalize agent-bridge /state response to stop page crash (#3318) (#3326) The Agent Bridge page seeded a well-shaped initialData default then replaced it wholesale with the raw /api/tools/agent-bridge/state response. The route returns { server, agents } but the UI reads { serverState, agentStates, bypassPatterns, mappings }, so serverState became undefined and AgentBridgeServerCard crashed on serverState.running — surfaced as the full-page 'Internal Server Error' boundary (client render error, not a real 5xx). Add a shared normalizeAgentBridgeState() that maps the route shape into the page contract (server.running/certExists -> serverState) and always returns safe defaults (never undefined serverState). Wired into both the SSR loader (page.tsx) and the polling hook. The legacy 'agents' entry shape differs from AgentStateEntry so it is not coerced; full route<->page contract reconciliation (port, upstreamCa, bypassPatterns, mappings, agentStates) is a follow-up. Co-authored-by: tycronk20 <tycronk20@users.noreply.github.com> * docs: VS Code/Ollama endpoints + env & i18n tooling (#3319) Integrated into release/v3.8.13 — VS Code/Ollama docs + env & i18n tooling * feat(provider): test-all endpoint, rate-limit overrides, visibility f… (#3267) Integrated into release/v3.8.13 — provider test-all endpoint, rate-limit overrides, model visibility * feat: auto-combo optimization, playground model dropdown, only-configured toggle (#3322) Integrated into release/v3.8.13 — auto-combo candidate expansion + playground dropdown + only-configured toggle * feat(api): VS Code Copilot Ollama-compatible BYOK endpoint (#3316) Integrated into release/v3.8.13 — VS Code Copilot Ollama-compatible BYOK endpoint (reconciled with #3306/#3309 auth hardening) * chore(release): document #3320 in the v3.8.13 changelog + contributor credits --------- Co-authored-by: Felipe Almeman <4226997+zhiru@users.noreply.github.com> Co-authored-by: Wilson <pedbookmed@gmail.com> Co-authored-by: Hernan Javier Ardila Sanchez <hjasgr@gmail.com> Co-authored-by: Paijo <14921983+oyi77@users.noreply.github.com> Co-authored-by: uniQta <uniQta@users.noreply.github.com> Co-authored-by: onizukashonan14-png <onizukashonan14-png@users.noreply.github.com> Co-authored-by: tycronk20 <tycronk20@users.noreply.github.com> Co-authored-by: Vinayrnani <vinayrnani@gmail.com>
36 KiB
36 KiB
title, version, lastUpdated
| title | version | lastUpdated |
|---|---|---|
| Provider Reference | 3.8.12 | 2026-06-06 |
Provider Reference
Auto-generated from
src/shared/constants/providers.ts— do not edit by hand. Regenerate with:npm run gen:provider-referenceLast generated: 2026-06-06
Total providers: 223. See category breakdown below.
Categories
- Free — free tier with API key (configured via dashboard)
- OAuth — sign-in flow handled by OmniRoute, no API key needed
- Web cookie — wraps the provider's web app via cookie auth
- API key — paid provider configured via API key (free credits may apply)
- Local — runs on the user's machine (Ollama, LM Studio, vLLM, etc.)
- Search — web search providers
- Audio — audio-only providers (TTS/STT)
- Upstream proxy — providers that proxy to other providers
- Cloud agent — long-running coding agents (Codex Cloud, Devin, Jules)
- System — OmniRoute-internal providers (loopback, etc.)
Additional tags: image, video, aggregator, enterprise, embed/rerank, self-hosted.
Use the dashboard at /dashboard/providers to enable, configure, and test each provider.
OAuth Providers (19)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
agy |
agy |
Antigravity CLI | OAuth | link | Import your Antigravity CLI (agy) login (paste/upload its token file), auto-detect a local CLI login, or sign in with Google. Shares the Antigravity backend (incl. Claude models). |
amazon-q |
aq |
Amazon Q | OAuth | link | Uses the same AWS Builder ID or imported refresh-token flow as Kiro, but keeps Amazon Q connections separate. |
antigravity |
— | Antigravity | OAuth | — | — |
claude |
cc |
Claude Code | OAuth | — | — |
cline |
cl |
Cline | OAuth | — | — |
codex |
cx |
OpenAI Codex | OAuth | — | — |
cursor |
cu |
Cursor IDE | OAuth | — | — |
devin-cli |
dv |
Devin CLI (Official) | OAuth | link | Requires the Devin CLI binary. Run devin auth login to authenticate, or provide your WINDSURF_API_KEY. Install: https://cli.devin.ai |
gemini-cli |
gemini-cli |
Gemini CLI | OAuth | — | Uses Gemini CLI OAuth / Cloud Code credentials. Pro models require an eligible Google account or paid plan. |
github |
gh |
GitHub Copilot | OAuth | — | — |
gitlab-duo |
gitlab-duo |
GitLab Duo | OAuth | link | OAuth application with ai_features + read_user scopes. Configure GITLAB_DUO_OAUTH_CLIENT_ID and optionally GITLAB_DUO_OAUTH_CLIENT_SECRET on this OmniRoute instance. |
kilocode |
kc |
Kilo Code | OAuth | — | — |
kimi-coding |
kmc |
Kimi Coding | OAuth | — | — |
kiro |
kr |
Kiro AI | OAuth | — | Free tier: 50 credits/month (~25K–100K tokens). ⚠️ Kiro ToS prohibits third-party proxy/harness use. |
qoder |
if |
Qoder AI | OAuth | — | — |
qwen |
qw |
Qwen Code | OAuth | — | ⚠️ DEPRECATED. Qwen OAuth free tier was discontinued on 2026-04-15. Use 'bailian-coding-plan', 'alibaba', 'alibaba-cn', or 'openrouter' provider with API key instead. |
trae |
tr |
Trae | OAuth | link | Trae is an AI-native IDE by ByteDance (SOLO remote agent). Authorize via trae.ai in the popup, or sign in at solo.trae.ai and paste the Cloud-IDE-JWT (sent as 'Authorization: Cloud-IDE-JWT ', ~14-day lifetime) as the access token; web_id/biz_user_id/user_unique_id/scope/tenant/region propagate via providerSpecificData. No headless refresh for pasted tokens — re-paste on expiry. |
windsurf |
ws |
Windsurf (Devin CLI) | OAuth | link | Sign in at windsurf.com to get your token. Visit windsurf.com/show-auth-token after logging in and paste it here, or use the device-code login flow. |
zed |
zd |
Zed IDE | OAuth | link | Zed stores LLM provider credentials (OpenAI, Anthropic, Google, Mistral, xAI) in the OS keychain. Use the Import button below to discover and import them automatically. |
Web Cookie Providers (20)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
adapta-web |
adp-web |
Adapta.org (Adapta One Web) | Web cookie | link | Paste your __client cookie value from .clerk.agent.adapta.one (DevTools → Application → Cookies) |
blackbox-web |
bb-web |
Blackbox Web (Subscription) | Web cookie | link | Paste your __Secure-authjs.session-token value or full cookie header from app.blackbox.ai |
chatgpt-web |
cgpt-web |
ChatGPT Web (Plus/Pro) | Web cookie | link | Paste your __Secure-next-auth.session-token cookie value from chatgpt.com |
claude-web |
cw |
Claude Web | Web cookie | link | Paste your session cookie from claude.ai |
copilot-web |
copilot |
Microsoft Copilot Web | Web cookie | link | Paste your access_token from copilot.microsoft.com (or export a .har file from DevTools while logged in) |
deepseek-web |
ds-web |
DeepSeek Web | Web cookie | link | Paste your userToken from chat.deepseek.com — DevTools → Application → Local Storage → userToken |
doubao-web |
db |
Doubao Web (ByteDance) | Web cookie | link | Paste your session cookie from doubao.com (DevTools → Application → Cookies) |
gemini-web |
gweb |
Gemini Web (Free) | Web cookie | link | Paste your __Secure-1PSID cookie value from gemini.google.com. Optionally add __Secure-1PSIDTS separated by semicolon. |
grok-web |
gw |
Grok Web (Subscription) | Web cookie | link | Paste the full grok.com cookie line from DevTools → Application → Cookies. Include both sso and sso-rw (e.g. sso=...; sso-rw=...) — Grok's anti-bot rejects sso on its own. |
huggingchat |
huggingchat |
HuggingChat (Free) | Web cookie | link | Paste your hf-chat cookie value from huggingface.co/chat (DevTools → Application → Cookies → hf-chat). Optional — works without auth for basic use. |
inner-ai |
in-ai |
Inner.ai (Subscription) | Web cookie | link | Paste your token cookie and email separated by a space: open DevTools → Application → Cookies → .innerai.com, copy the token value, then append a space and your Inner.ai login email. Example: eyJhbG... user@example.com |
kimi-web |
kimi-web |
Kimi Web (Moonshot AI) | Web cookie | link | Paste your session cookie from kimi.moonshot.cn (DevTools → Application → Cookies) |
muse-spark-web |
ms-web |
Muse Spark Web (Meta AI) | Web cookie | link | Paste your abra_sess value or full cookie header from meta.ai |
perplexity-web |
pplx-web |
Perplexity Web (Pro/Max) | Web cookie | link | Paste your __Secure-next-auth.session-token cookie value from perplexity.ai |
phind |
ph |
Phind (Free) | Web cookie | link | Paste your session cookie from phind.com (DevTools → Application → Cookies). Optional — works with free tier. |
poe-web |
poe |
Poe Web (Subscription) | Web cookie | link | Paste your p-b cookie value from poe.com (DevTools → Application → Cookies → p-b) |
qwen-web |
qwen-web |
Qwen Web (Free) | Web cookie | link | Open chat.qwen.ai, log in, then open DevTools → Application → Local Storage → copy the "token" value (or use tongyi_sso_ticket cookie as Bearer token). |
t3-web |
t3chat |
t3.chat (Pro/Free) | Web cookie | link | Open t3.chat in your browser, log in, then open DevTools → Application → Local Storage → https://t3.chat. Copy the value of 'convex-session-id'. Also open DevTools → Network, copy the Cookie header from any request. Paste both values here. See provider setup docs for a step-by-step guide. |
v0-vercel-web |
v0 |
v0 Vercel Web (Code Gen) | Web cookie | link | Paste your session cookie from v0.dev (DevTools → Application → Cookies) |
venice-web |
ven |
Venice Web (Privacy) | Web cookie | link | Paste your session cookie from venice.ai (DevTools → Application → Cookies) |
API Key Providers (paid / paid-with-free-credits) (151)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
360ai |
360ai |
360 AI | API key | link | Get API key at ai.360.cn |
agentrouter |
agentrouter |
AgentRouter | API key, aggregator | link | $200 free credits on signup - multi-model routing gateway |
ai21 |
ai21 |
AI21 Labs | API key | link | $10 trial credits on signup (valid 3 months), no credit card required |
aimlapi |
aiml |
AI/ML API | API key, aggregator | link | $0.025/day free credits — 200+ models (GPT-4o, Claude, Gemini, Llama) via single endpoint |
alibaba |
ali |
Alibaba | API key | link | — |
alibaba-cn |
ali-cn |
Alibaba (China) | API key | link | — |
anthropic |
anthropic |
Anthropic | API key | link | — |
api-airforce |
af |
Api.airforce | API key | link | 55 free tier models including Grok-3, Claude 3.7, Qwen3, Kimi-K2, Gemini 2.5 Flash, DeepSeek-V3 |
arcee-ai |
arcee |
Arcee AI | API key | link | Get API key at arcee.ai |
azure-ai |
azure-ai |
Azure AI Foundry | API key, enterprise | link | Use your Azure AI Foundry key. Base URL can be https://.services.ai.azure.com/openai/v1/ or https://.openai.azure.com/openai/v1/. |
azure-openai |
azure |
Azure OpenAI | API key, enterprise | link | Use your Azure OpenAI API key. Base URL should be your resource endpoint, for example https://my-resource.openai.azure.com. |
baichuan |
baichuan |
Baichuan | API key | link | Get API key at platform.baichuan-ai.com |
baidu |
baidu |
Baidu (ERNIE) | API key | link | Get API key at console.bce.baidu.com |
bailian-coding-plan |
bcp |
Alibaba Coding Plan | API key | link | — |
baseten |
baseten |
Baseten | API key | link | $30 free trial credits for GPU inference |
bazaarlink |
bzl |
BazaarLink | API key | link | Free tier with auto:free routing — zero-cost inference, no credit card required |
bedrock |
bedrock |
Amazon Bedrock | API key, enterprise | link | Use your Amazon Bedrock API key and configure the AWS region where your models are enabled (for example eu-west-2). OmniRoute calls Bedrock's native Converse API directly. |
black-forest-labs |
bfl |
Black Forest Labs | API key, image | link | — |
blackbox |
bb |
Blackbox AI | API key | link | Free tier: unlimited basic chat plus Minimax-M2.5, no credit card required |
bluesminds |
bm |
BluesMinds | API key | link | Free daily pi credits — supports 200+ models including GPT-4o, GPT-4.1, Claude Sonnet 4.5, Gemini 2.0 Flash, DeepSeek V4, Qwen, Kimi K2 |
byteplus |
bpm |
BytePlus ModelArk | API key | link | — |
bytez |
bytez |
Bytez | API key | link | $1 free credits, refreshes every 4 weeks |
cablyai |
cablyai |
CablyAI | API key, aggregator | link | Bearer API key for the CablyAI OpenAI-compatible gateway. |
cerebras |
cerebras |
Cerebras | API key | link | Free Trial: 1M tokens/day, 30K TPM, 5 RPM — no credit card. |
chutes |
chutes |
Chutes.ai | API key, aggregator | link | Bearer API key for the Chutes OpenAI-compatible gateway. |
clarifai |
clarifai |
Clarifai | API key, enterprise | link | Use your Clarifai PAT or app-specific API key. OmniRoute targets the OpenAI-compatible endpoint at https://api.clarifai.com/v2/ext/openai/v1 and authenticates with Authorization: Key . |
cloudflare-ai |
cf |
Cloudflare Workers AI | API key | link | Requires API Token AND Account ID (found at dash.cloudflare.com) |
codestral |
codestral |
Codestral | API key | link | — |
cohere |
cohere |
Cohere | API key | link | Free Trial: 1,000 API calls/month for testing, no credit card required |
command-code |
cmd |
Command Code | API key | link | Use a Command Code API key. Requests are sent to Command Code's /alpha/generate endpoint. |
coze |
coze |
Coze | API key | link | Get API key at coze.com/open/api |
crof |
crof |
CrofAI | API key | link | — |
databricks |
databricks |
Databricks | API key, enterprise | link | — |
datarobot |
datarobot |
DataRobot | API key, enterprise | link | Use your DataRobot API token. Optional Base URL can be the account root (for LLM Gateway) or a deployment URL under /api/v2/deployments/. |
deepinfra |
deepinfra |
DeepInfra | API key | link | Free signup credits for API testing and model exploration |
deepseek |
ds |
DeepSeek | API key | link | 5M free tokens on signup - no credit card required |
dify |
dify |
Dify | API key | link | Get API key from your Dify instance. |
doubao |
doubao |
Doubao | API key | link | Get API key at console.volcengine.com |
empower |
empower |
Empower | API key, aggregator | link | Bearer API key for the Empower OpenAI-compatible endpoint. |
fal-ai |
fal |
Fal.ai | API key, image | link | — |
featherless-ai |
featherless |
Featherless AI | API key | link | Free tier available — no credit card required |
fenayai |
fenayai |
FenayAI | API key, aggregator | link | Bearer API key for the FenayAI OpenAI-compatible gateway. |
firecrawl |
fc |
Firecrawl | API key | link | — |
fireworks |
fireworks |
Fireworks AI | API key | link | $1 free starter credits on signup for API testing |
freeaiapikey |
faik |
FreeAIAPIKey | API key | link | — |
freemodel-dev |
fmd |
FreeModel.dev | API key | link | $300 free credits on signup — no credit card required. Access GPT-5.4 and GPT-5.5 (OpenAI's latest flagship models) through an OpenAI-compatible API. |
friendliai |
friendli |
FriendliAI | API key | link | Free tier for serverless inference — no credit card required |
galadriel |
galadriel |
Galadriel | API key | link | — |
gemini |
gemini |
Gemini (Google AI Studio) | API key | link | Free forever: 1,500 req/day for Gemini 2.5 Flash — no credit card, get key at aistudio.google.com |
getgoapi |
ggo |
GoAPI | API key, aggregator | link | — |
gigachat |
gigachat |
GigaChat (Sber) | API key | link | — |
github-models |
ghm |
GitHub Models | API key | link | Create a GitHub PAT with 'models: read' scope at github.com/settings/tokens |
gitlab |
gitlab |
GitLab Duo PAT | API key | link | GitLab personal access token for the public Code Suggestions API. Configure a self-hosted base URL when not using gitlab.com. |
gitlawb |
glb |
Gitlawb Opengateway (MiMo) | API key | link | Free tier available — no credit card required |
gitlawb-gmi |
glb-gmi |
Gitlawb Opengateway (GMI Cloud) | API key | link | Free tier available — no credit card required |
glhf |
glhf |
GLHF Chat | API key, aggregator | link | Bearer API key for the GLHF OpenAI-compatible gateway. |
glm |
glm |
GLM Coding | API key | link | — |
glm-cn |
glmcn |
GLM Coding (China) | API key | link | — |
glmt |
glmt |
GLM Thinking | API key | link | — |
groq |
groq |
Groq | API key | link | Free tier: 30 RPM / 14.4K RPD — no credit card |
hackclub |
hc |
Hackclub AI | API key, aggregator | link | Sign in with your Hack Club account at ai.hackclub.com. |
haiper |
hp |
Haiper | API key, video | link | Get API key at haiper.ai/haiper-api |
heroku |
heroku |
Heroku AI | API key, enterprise | link | — |
huggingchat |
huggingchat |
HuggingChat | API key | link | No API key required for basic access. |
huggingface |
hf |
HuggingFace | API key | link | Free Inference API for thousands of models (Whisper, VITS, SDXL…) |
hyperbolic |
hyp |
Hyperbolic | API key | link | $1-5 trial credits on signup for serverless inference |
ideogram |
ideo |
Ideogram | API key | link | Get API key at ideogram.ai/docs/api |
iflytek |
iflytek |
iFlytek Spark | API key | link | Get API key at console.xfyun.cn |
inclusionai |
inclusion |
InclusionAI | API key | link | Get API key at inclusionai.com |
inference-net |
inet |
Inference.net | API key | link | $25 free credits on signup plus research grants available |
jina-ai |
jina |
Jina AI | API key, embed/rerank | link | Bearer API key for the Jina AI rerank API. |
jina-reader |
jr |
Jina Reader | API key | link | — |
kie |
kie |
KIE.AI | API key | link | — |
kilo-gateway |
kg |
Kilo Gateway | API key, aggregator | link | — |
kimi |
kimi |
Kimi | API key | link | — |
kimi-coding-apikey |
kmca |
Kimi Coding (API Key) | API key | link | — |
kluster |
kluster |
Kluster AI | API key | link | $5 free credits on signup - DeepSeek R1, Llama 4 Maverick/Scout, Qwen3 235B |
lambda-ai |
lambda |
Lambda AI | API key | link | — |
laozhang |
lz |
LaoZhang AI | API key, aggregator | link | — |
leonardo |
leo |
Leonardo AI | API key, video | link | Get API key at leonardo.ai/developer |
liquid |
liquid |
Liquid AI | API key | link | Get API key at liquid.ai |
llamagate |
llamagate |
LlamaGate | API key | link | — |
llm7 |
llm7 |
LLM7.io | API key | link | No signup required - 2 req/s, 20 RPM, 100 req/hr free tier |
longcat |
lc |
LongCat AI | API key | link | Free: 5M tokens/day on LongCat-2.0-Preview (Flash models retired 2026-05-29); up to 120M/day via feedback. |
maritalk |
maritalk |
Maritalk | API key | link | — |
meta-llama |
meta |
Meta Llama API | API key | link | — |
minimax |
minimax |
Minimax Coding | API key, video | link | — |
minimax-cn |
minimax-cn |
Minimax (China) | API key | link | — |
mistral |
mistral |
Mistral | API key | link | Free Experiment tier: rate-limited access to all models, no credit card required |
modal |
mdl |
Modal | API key, enterprise | link | Use the bearer token that protects your Modal deployment, if enabled. Base URL should point to your OpenAI-compatible Modal app, for example https://--.modal.run/v1. |
monsterapi |
monster |
MonsterAPI | API key | link | Get API key at monsterapi.ai |
moonshot |
moonshot |
Moonshot AI | API key | link | — |
morph |
morph |
Morph | API key | link | Free tier: 250K credits/month, $0 |
nanogpt |
nanogpt |
NanoGPT | API key | link | — |
nebius |
nebius |
Nebius AI | API key | link | ~$1 trial credits on signup for API testing |
nlpcloud |
nlpc |
NLP Cloud | API key | link | Use your NLP Cloud API key in Authorization: Token . OmniRoute targets the chatbot endpoint on https://api.nlpcloud.io/v1/gpu//chatbot by default. |
nomic |
nomic |
Nomic | API key | link | Get API key at atlas.nomic.ai |
nous-research |
nous |
Nous Research | API key | link | Use your Nous Portal API key. OmniRoute targets the official OpenAI-compatible inference endpoint at https://inference-api.nousresearch.com/v1. |
novita |
novita |
Novita AI | API key, aggregator | link | $0.50 trial credits on signup (valid about 1 year) |
nscale |
nscale |
nScale | API key | link | $5 free credits on signup for inference testing |
nvidia |
nvidia |
NVIDIA NIM | API key | link | Free dev access: ~40 RPM, 70+ models (Kimi K2.5, GLM 4.7, DeepSeek V3.2...) |
oci |
oci |
OCI Generative AI | API key, enterprise | link | Use your OCI Generative AI API key or IAM bearer token. Base URL can be https://inference.generativeai..oci.oraclecloud.com/openai/v1/. |
ollama-cloud |
ollamacloud |
Ollama Cloud | API key | link | — |
openai |
openai |
OpenAI | API key | link | — |
opencode-go |
opencode-go |
OpenCode Go | API key | link | — |
opencode-zen |
opencode-zen |
OpenCode Zen | API key | link | — |
openrouter |
openrouter |
OpenRouter | API key, aggregator | link | Free models at $0/token with :free suffix - 20 RPM / 200 RPD |
ovhcloud |
ovh |
OVHcloud AI | API key | link | — |
perplexity |
pplx |
Perplexity | API key | link | — |
phind |
phind |
Phind | API key | link | Get API key at phind.com |
piapi |
pi |
PiAPI | API key, aggregator | link | — |
poe |
poe |
Poe | API key, aggregator | link | Bearer API key for the Poe OpenAI-compatible API. |
pollinations |
pol |
Pollinations AI | API key, video | link | No API key required for free public endpoint. Optional Spore tier: ~0.01 pollen/hour. |
predibase |
predibase |
Predibase | API key | link | $25 free trial credits (30-day validity) |
publicai |
publicai |
PublicAI | API key | link | Free community inference tier for testing |
puter |
pu |
Puter AI | API key | link | Get token at puter.com/dashboard → Copy Auth Token |
qianfan |
qianfan |
Baidu Qianfan | API key | link | — |
recraft |
recraft |
Recraft | API key, image | link | — |
reka |
reka |
Reka | API key | link | Use your Reka API key. OmniRoute supports the OpenAI-compatible base URL https://api.reka.ai/v1 and sends both Authorization and X-Api-Key headers for compatibility. |
runwayml |
runway |
Runway | API key, video | link | Use your Runway API key in Authorization: Bearer . OmniRoute targets the current Runway API at https://api.dev.runwayml.com/v1 and sends the required X-Runway-Version header automatically. |
sambanova |
samba |
SambaNova | API key | link | $5 free credits on signup (30-day validity), no credit card required |
sap |
sap |
SAP Generative AI Hub | API key, enterprise | link | Use your SAP AI Core bearer token. Base URL can be your AI_API_URL root or a deploymentUrl from Generative AI Hub. |
scaleway |
scw |
Scaleway AI | API key | link | 1M free tokens for new accounts — EU/GDPR compliant (Paris), Qwen3 235B & Llama 70B |
sensenova |
sensenova |
SenseNova | API key | link | Get API key at platform.sensenova.cn |
siliconflow |
siliconflow |
SiliconFlow | API key | link | $1 free credits plus permanently free models after identity verification |
snowflake |
snowflake |
Snowflake Cortex | API key, enterprise | link | — |
sparkdesk |
sparkdesk |
SparkDesk | API key | link | Get API key at console.xfyun.cn |
stability-ai |
stability |
Stability AI | API key, image | link | — |
stepfun |
stepfun |
StepFun | API key | link | Get API key at platform.stepfun.com |
suno |
suno |
Suno | API key | link | Paste session cookie from suno.ai (Clerk auth) |
synthetic |
synthetic |
Synthetic | API key, aggregator | link | — |
tencent |
tencent |
Tencent Hunyuan | API key | link | Get API key at console.cloud.tencent.com |
thebai |
thebai |
TheB.AI | API key, aggregator | link | Bearer API key for the TheB.AI OpenAI-compatible gateway. |
together |
together |
Together AI | API key, video | link | $25 signup credits + 3 permanently free models: Llama 3.3 70B, Vision, DeepSeek-R1 distill |
topaz |
topaz |
Topaz | API key, image | link | — |
udio |
udio |
Udio | API key | link | Paste session cookie from udio.com (Supabase auth) |
uncloseai |
unc |
UncloseAI | API key | link | No auth required. API accepts any non-empty string as key for identification. |
upstage |
upstage |
Upstage | API key | link | — |
v0-vercel |
v0 |
v0 (Vercel) | API key | link | — |
venice |
venice |
Venice.ai | API key | link | — |
vercel-ai-gateway |
vag |
Vercel AI Gateway | API key, aggregator | link | — |
vertex |
vertex |
Vertex AI | API key, enterprise | link | Provide Service Account JSON or OAuth access_token |
vertex-partner |
vp |
Vertex AI Partners | API key, enterprise | link | Provide the same Service Account JSON used for Vertex AI partner models. |
volcengine |
volcengine |
Volcengine | API key | link | — |
voyage-ai |
voyage |
Voyage AI | API key, embed/rerank | link | Bearer API key for Voyage AI embeddings and rerank APIs. |
wandb |
wandb |
Weights & Biases Inference | API key | link | — |
watsonx |
watsonx |
IBM watsonx.ai Gateway | API key, enterprise | link | Use your watsonx bearer token. Base URL can be https://.ml.cloud.ibm.com/ml/gateway/v1/ or a self-managed /ml/gateway/v1 endpoint. |
xai |
xai |
xAI (Grok) | API key | link | — |
xiaomi-mimo |
mimo |
Xiaomi MiMo | API key | link | — |
yi |
yi |
Yi (01.AI) | API key | link | Get API key at platform.lingyiwanwu.com |
zai |
zai |
Z.AI | API key | link | — |
Local Providers (11)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
comfyui |
comfyui |
ComfyUI | Local | link | No API key required. Configure the local ComfyUI base URL (default: http://localhost:8188). |
docker-model-runner |
dmr |
Docker Model Runner | Local, self-hosted | link | API key optional. Configure the local Docker Model Runner OpenAI-compatible base URL (default: http://localhost:12434/v1). |
lemonade |
lemonade |
Lemonade Server | Local, self-hosted | link | API key optional. Configure the local Lemonade OpenAI-compatible base URL (default: http://localhost:13305/api/v1). |
llama-cpp |
llamacpp |
llama.cpp | Local, self-hosted | link | API key optional (use any value, e.g. sk-no-key-required). Configure the llama-server OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1). Note: if Llamafile is also installed, both default to port 8080 — run only one at a time or override the port. |
llamafile |
llamafile |
Llamafile | Local, self-hosted | link | API key optional. Configure the local Llamafile OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1). |
lm-studio |
lmstudio |
LM Studio | Local, self-hosted | link | API key optional. Configure the local LM Studio OpenAI-compatible base URL (default: http://localhost:1234/v1). |
oobabooga |
ooba |
oobabooga | Local, self-hosted | link | API key optional. Configure the local oobabooga OpenAI-compatible base URL (default: http://localhost:5000/v1). |
sdwebui |
sdwebui |
SD WebUI | Local | link | No API key required. Configure the local WebUI base URL (default: http://localhost:7860). |
triton |
triton |
NVIDIA Triton | Local, self-hosted | link | API key optional. Configure the Triton OpenAI-compatible base URL (default: http://localhost:8000/v1). |
vllm |
vllm |
vLLM | Local, self-hosted | link | API key optional. Configure the local vLLM OpenAI-compatible base URL (default: http://localhost:8000/v1). |
xinference |
xinference |
XInference | Local, self-hosted | link | API key optional. Configure the local XInference OpenAI-compatible base URL (default: http://localhost:9997/v1). |
Search Providers (11)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
brave-search |
brave-search |
Brave Search | Search | link | Subscription token from Brave Search API dashboard |
exa-search |
exa-search |
Exa Search | Search | link | API key from dashboard.exa.ai |
google-pse-search |
google-pse |
Google Programmable Search | Search | link | Requires a Google API key and your Programmable Search Engine ID (cx) |
linkup-search |
linkup |
Linkup Search | Search | link | Bearer API key from the Linkup dashboard |
ollama-search |
ollama-search |
Ollama Search | Search | link | Same API key as Ollama Cloud (from ollama.com/settings/api-keys) |
perplexity-search |
pplx-search |
Perplexity Search | Search | link | Same API key as Perplexity (pplx-...) |
searchapi-search |
searchapi |
SearchAPI | Search | link | API key from SearchAPI (query param or Bearer auth) |
searxng-search |
searxng |
SearXNG Search | Search | link | API key is optional. Set your SearXNG base URL. Some instances may require a bearer token for access. |
serper-search |
serper-search |
Serper Search | Search | link | API key from serper.dev dashboard |
tavily-search |
tavily-search |
Tavily Search | Search | link | API key from app.tavily.com (format: tvly-...) |
youcom-search |
youcom-search |
You.com Search | Search | link | X-API-Key from the You.com platform dashboard |
Audio-only Providers (7)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
assemblyai |
aai |
AssemblyAI | Audio | link | — |
aws-polly |
polly |
AWS Polly | Audio | link | Use AWS Secret Access Key as API key; set providerSpecificData.accessKeyId and optional region. |
cartesia |
cartesia |
Cartesia | Audio | link | — |
deepgram |
dg |
Deepgram | Audio | link | — |
elevenlabs |
el |
ElevenLabs | Audio | link | — |
inworld |
inworld |
Inworld | Audio | link | — |
playht |
playht |
PlayHT | Audio | link | — |
Upstream Proxy Providers (2)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
9router |
nr |
9router | Upstream proxy | link | — |
cliproxyapi |
cpa |
CLIProxyAPI | Upstream proxy | link | — |
Cloud Agent Providers (3)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
codex-cloud |
codex-cloud |
Codex Cloud | Cloud agent | link | OpenAI API key with Codex Cloud task access. |
devin |
devin |
Devin | Cloud agent | link | Devin API key for cloud agent sessions. |
jules |
jules |
Google Jules | Cloud agent | link | Jules API key for creating and managing cloud coding tasks. |
System Providers (1)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
auto |
auto |
Auto (Zero-Config) | System | — | — |
Sources of truth
- Catalog:
src/shared/constants/providers.ts - Registry (per-model details):
open-sse/config/providerRegistry.ts - Executors:
open-sse/executors/(31 files) - Translators:
open-sse/translator/
See Also
- FREE_TIERS.md — curated free-tier guide
- USER_GUIDE.md — provider setup walkthrough
- ARCHITECTURE.md — overall architecture