Files
OmniRoute/docs/reference/PROVIDER_REFERENCE.md
Diego Rodrigues de Sa e Souza 7b139fdb5e Release v3.8.38 (#5078)
* chore(release): open v3.8.38 development cycle

* fix(executors): strip client_metadata for cerebras and mistral (#4727)

Integrated into release/v3.8.38 (leva 5)

* fix(codebuddy): only send reasoning params when client requests reasoning (#5019)

Integrated into release/v3.8.38 (leva 5)

* fix(sse): keep streaming for forceStream providers when client requests JSON (#5021)

Integrated into release/v3.8.38 (leva 5)

* fix(sse): guard non-JSON SSE lines and duplicate [DONE] (#4937)

Integrated into release/v3.8.38 (leva 5)

* feat(blackbox): refresh provider model catalog (#4935)

Integrated into release/v3.8.38 (leva 5)

* fix(sse): dedupe case-variant Anthropic version/beta headers (#4846)

Integrated into release/v3.8.38 (leva 5)

* feat(sse): Kiro inline <thinking> stream splitter (#4911)

Integrated into release/v3.8.38 (leva 5)

* feat(cursor): parse Composer DeepSeek-style inline tool calls (#4912)

Integrated into release/v3.8.38 (leva 5)

* feat(proxy): auth-less host:port batch import (#4938)

Integrated into release/v3.8.38 (leva 5)

* fix(oauth): support Kiro IDC (organization) token import (#4944)

Integrated into release/v3.8.38 (leva 5)

* fix(translator): preserve cache_control for DashScope OpenAI-compat providers (port from 9router#2069) (#5013)

Integrated into release/v3.8.38 (leva 5)

* fix(tts): resolve Gemini TTS models from catalog (#4934)

Integrated into release/v3.8.38 (leva 5)

* fix(sse): don't cool down the connection on a self-inflicted upstream timeout (504) (#5064)

Integrated into release/v3.8.38 (leva 5)

* fix(sse): robust Anthropic /v1/messages streaming — real ping keepalive + client-disconnect guard (#5063)

Integrated into release/v3.8.38 (leva 5)

* feat(video): add Alibaba DashScope (wan2.7-t2v) provider (#5051)

Integrated into release/v3.8.38 (leva 5)

* fix: preserve model hidden flags (isHidden) across model sync (#5086)

Integrated into release/v3.8.38 (leva 5)

* fix(models): derive model discovery config from registry modelsUrl (#5087)

Integrated into release/v3.8.38 (leva 5)

* fix(compression): replace fileURLToPath(import.meta.url) with runtime anchors for standalone bundle (#5089)

Integrated into release/v3.8.38 (leva 5)

* feat(cc): add summarized thinking display toggle (#5055)

Integrated into release/v3.8.38 (leva 5)

* Harden selected API error responses (#5032)

Integrated into release/v3.8.38 (leva 5)

* chore(quality): rebaseline file-size for leva 5 PR batch drift

6 frozen files grew from merged leva-5 PRs (cursor #4912, kiro #4911,
videoGeneration #5051, default #4727, base #4846, chat #5064); all covered
by per-PR tests. See _rebaseline_2026_06_26_leva5 in the baseline.

* feat(compression): compression playground (Play + Compare tabs) in the studio (#5080)

Integrated into release/v3.8.38

* fix(combo): fail over on empty-content 502 instead of exhausting the provider (#5085) (#5104)

* fix(dashboard): surface detailed credential-validation error in add-connection modal (#5088) (#5106)

* feat(providers): allow local/private provider URLs by default with scoped metadata-safe guard (#5066) (#5107)

* fix(diagnostics): treat non-streaming Claude messages shape as valid output (#5108) (#5116)

* fix(db): translate pt-BR SQLite driver-fallback log lines to English (#5103) (#5115)

* fix(sse): repair release base-reds — malformed-response false positives + header casing + stale tests (#5117)

Repairs the release/v3.8.38 base-reds; unblocks #5078.

* chore(quality): rebaseline file-size for responseSanitizer (#5117) + AddApiKeyModal drift

* fix(translator): forward image tool_result blocks as image_url (#5100)

Base-reds fixed (#5117); image tool_result→image_url. Integrated into release/v3.8.38.

* fix(responses): default text.format for openai-compatible responses providers (#5101)

Base-reds fixed (#5117); default text.format + file-size rebaseline. Integrated into release/v3.8.38.

* feat(dashboard): expose Fusion judgeModel + fusionTuning in the combo editor (#5074)

Base-reds fixed (#5117); Fusion editor + file-size rebaseline. Integrated into release/v3.8.38.

* feat(quota): add opt-in Codex/Claude auto-ping keepalive (#5102)

Base-reds fixed (#5117); auto-ping keepalive + file-size rebaseline. Integrated into release/v3.8.38.

* test(release): relocate 2 orphan test files into the collected flat tests/unit dir (#5120)

Unblocks Lint (test-discovery) on #5078. Integrated into release/v3.8.38.

* fix(translator): preserve reasoning-replay reasoning_content + repair 3 release-green test reds (#5122)

Repairs 3 release-green test reds + test-masking; unblocks #5078.

* test(golden): redact live Node version from provider translate-path snapshot (#5125)

Final golden unblock for #5078.

* test(golden): redact OmniRoute app version from translate-path snapshot (#5126)

Coverage shard golden unblock for #5078.

* Ignore disconnect races during in-band stream error handling (#5007)

Integrated into release/v3.8.38

* Track final connection IDs in failover logs (#5016)

Integrated into release/v3.8.38

* fix(sse): convert Gemini body to OpenAI format in antigravity MITM handler (#4845)

Integrated into release/v3.8.38 (rebased on tip, CHANGELOG re-injected)

* feat(providers): add ZenMux Free session-cookie provider (#5105)

Integrated into release/v3.8.38 (rebased on tip, CHANGELOG re-injected)

* feat(dashboard): click-to-edit model alias in provider page (#5119)

Integrated into release/v3.8.38 (rebased on tip, i18n scope verified, CHANGELOG re-injected)

* feat(mcp): web-session robustness — cookie dedup (PR6) + browser-pool observability (PR7) (#3368) (#5121)

Integrated into release/v3.8.38 (rebased on tip; cookie-dedup branch extracted to findExistingCookieConnection helper → complexity-neutral; CHANGELOG added)

* fix(usage): dedupe request-usage logging and debounce stats (#4940)

Integrated into release/v3.8.38 (rebased on tip; DB-handle hang was stale-base artifact — resetDbInstance already closes the handle, test green 5/5; file-size drift consolidated at release; CHANGELOG re-injected)

* fix(dashboard): key model visibility toggle on canonical providerId (#5091)

Integrated into release/v3.8.38 (retargeted main→release; .tsx visibility-key test green 2/2)

* chore(deps): bump actions/cache from 5.0.5 to 6.0.0 (#5112)

Integrated into release/v3.8.38 (retargeted main→release; workflow-only actions/cache bump — unit failures were stale main base-reds)

* fix(streaming): harden long OpenAI-compatible SSE streams (#5124)

Integrated into release/v3.8.38 (rebased on tip; streamHandler conflict with #5007 disconnect-guard resolved — both coexist, stream-handler 22/22 green)

* feat: Add Grok Build (xAI) provider with OAuth import-token flow (#5020)

Integrated into release/v3.8.38 (rebased on tip; Hard Rule #11 fix — Grok public client_id now via resolvePublicCred(grok_id), 3 literals removed; grok-oauth 7/7 + check:public-creds green)

* feat(providers): add Factory (factory.ai) as a subscription gateway provider (#5065)

Integrated into release/v3.8.38 (rebased on tip; added factory registry test for PR Test Policy + fixed check:env-doc-sync phantom FACTORY_API_KEY; factory loads in PROVIDERS, no Zod issue — that flag was a false positive)

* chore(test): reconcile golden snapshot + apikey count for new providers

#5020 (grok-cli), #5065 (factory), #5105 (zenmux-free) added providers but did
not regenerate tests/snapshots/provider/translate-path.json (now +3 entries) nor
bump the APIKEY_PROVIDERS count (159->160 for the factory gateway). Test-only
reconciliation; no production change.

* fix(resilience): harden quota and model lockout edge cases (#5093)

Integrated into release/v3.8.38 (rebased on tip). TRUST-BUT-VERIFY: dropped the PR's 0dd7df641 'fix unit gates' commit which reverted #5122 reasoning-replay (preserveReasoningContent) + re-introduced #4849 O(n^2) growth, and restored 5 tests it had realigned. Kept only the 3 declared resilience fixes (quota cutoff guard, gemini MIME, model-lockout maxCooldownMs); 23/23 green.

* Hydrate quota cache and scope auto combo candidates (#5015)

Integrated into release/v3.8.38 (rebased on tip). Kept core quota-cache hydration + auto-combo candidate scoping + combos UI; dropped out-of-scope toolCloaking refactor (conflicted with #4813 stripEnumDescriptions — took tip) and the unrelated sse-auth test split. Added quota-cache-hydrate-5015 regression test (Rule #18); combo-account-allowlist 8/8 + hydration 2/2 green.

* chore(quality): reconcile complexity + file-size baselines for v3.8.38 owner-PR batch

complexity 1972->1978 (+6) and file-size providers.ts 1093->1107 / usageHistory.ts
934->983 — drift from the /review-prs merge batch (#4845/#5105/#5020/#4940/#5093/
#5015 + #5121 cookie-dedup helper extraction). check:complexity/check:file-size do
not run on the PR->release fast-path, so the branch accrued unmeasured; all legit
feature/fix growth, not regression. See per-key justifications in each baseline.

* fix(security): exact-host Anthropic baseUrl check (CodeQL js/incomplete-url-substring-sanitization #674) (#5130)

The anthropic-compatible Bearer-fallback gate decided whether a configured baseUrl
targeted the official api.anthropic.com host via a substring `.includes("api.anthropic.com")`.
A look-alike upstream such as `https://api.anthropic.com.evil.test` or
`https://evil.test/?x=api.anthropic.com` matched the substring and was wrongly treated as
official, suppressing the Bearer fallback meant for third-party gateways
(CodeQL #674, js/incomplete-url-substring-sanitization, high).

Replace the substring test with an exported `isOfficialAnthropicBaseUrl()` helper that
parses the URL and compares the hostname for exact equality. Empty baseUrl stays official;
scheme-less hosts are parsed with an assumed https://; an unparseable baseUrl falls back to
third-party (Bearer emitted) as the safer default. Behavior for legitimate official/third-party
baseUrls is unchanged.

Adds tests/unit/anthropic-official-baseurl-host.test.ts covering official, look-alike,
scheme-less, and unparseable inputs plus a static guard that the substring pattern is gone.

* fix(proxy): repair one-click Deno & Cloudflare relay deployments (#5128) (#5132)

* fix(services): embed WS proxy honours LIVE_WS_HOST; reject empty messages early (#5110) (#5133)

* fix(api): resolve /v1/models/{id} case-insensitively (#5082) (#5135)

* fix(providers): add MiniMax M3 & Nemotron 3 Ultra to Cline catalog (#3321) (#5136)

* fix(proxy): make SOCKS5 handshake timeout tunable via SOCKS_HANDSHAKE_TIMEOUT_MS (#5109) (#5137)

* feat(sidebar): add support for colored menu icons (#3812)

Integrated into release/v3.8.38 (recreated on tip — fork had unrelated history; added getSidebarIconAccent regression test, Rule #18). Clean 2-file UI feature.

* fix(providers): complete grok-cli OAuth wiring + zenmux-free web-session metadata

Base-red repair for #5020 (grok-cli) and #5105 (zenmux-free), surfaced by the
full CI on the release PR (#5078) — the PR->release fast-path does not run the
oauth-providers-config / web-session-credentials / provider-consistency gates.

- grok-cli: register in OAUTH_PROVIDERS (providers.ts canonical list, fixes
  check:provider-consistency), add OAUTH_PROVIDER_IDS.GROK_CLI + GROK_CLI_CONFIG
  in oauth constants (provider config now sourced there, not a local literal),
  align oauth-providers-config.test.ts (EXPECTED_PROVIDER_KEYS + config map).
- zenmux-free: declare its web-session credential requirement (full Cookie header)
  in WEB_SESSION_CREDENTIAL_REQUIREMENTS.

Local: oauth-providers-config 27/27, web-session-credentials 4/4, grok-cli-oauth
7/7, check:provider-consistency OK, +115 OAUTH_PROVIDERS tests green.

* Fix resilience settings page response mapping (#5139)

Integrated into release/v3.8.38. Thanks @rdself for the fix and the regression test.

* fix(kiro): retire claude-sonnet-4.5 from catalog + pin 400 model-unavailable test (#5140)

Extracted the real change from #5140 (the bot PR regenerated the entire
freeModelCatalog.data.ts + touched package-lock.json; only the targeted
edits are kept here):
- remove claude-sonnet-4.5 from the Kiro registry entry
- remove the matching kiro free-model catalog row
- pin Kiro's verbatim 400 "Invalid model..." to isModelUnavailableError

Closes #4484

* fix(sidebar): drop orphan `settings` accent color (typecheck:core red) (#5142)

SIDEBAR_ICON_ACCENTS is typed Partial<Record<HideableSidebarItemId, string>>,
but `settings` is not a hideable item id (only `settings-general`,
`settings-appearance`, … and `context-settings` exist; there is no item with
`id: "settings"`), so the accent was unreachable. It broke `typecheck:core`
on the release tip ("'settings' does not exist in type …", introduced by
#3812 colored menu icons). Removing the orphan key restores a clean
typecheck:core (rc=0).

* feat: salvage batch 2 — diagnostics null-guard (#5096) + observed quota reset windows (#5025) (#5141)

* fix(diagnostics): null-guard content blocks in detectMalformedNonStream

A null (or non-object) entry in a Claude-native `content` array made the
non-stream classifier throw `TypeError: Cannot read properties of null
(reading 'type')`, crashing the malformed-response detection path. Guard
before type-asserting each block: a null/non-object block is simply skipped.

Two regression tests added (null block among valid blocks → null; only-null
blocks → empty_choices).

Salvaged from closed PR #5096 (base-stale; only the defensive guard — the
Claude-shape recognition it also carried already landed via #5108).

Co-authored-by: herjarsa <herjarsa@users.noreply.github.com>

* feat(quota): persist observed provider quota reset windows

Adds `provider_quota_reset_events` (migration 108) + `db/quotaResetEvents.ts`
to record real upstream weekly-quota window transitions whenever a quota
refresh shows the reset rolling to a new cycle (different day, later resetAt).
`apiKeyUsageLimits` now prefers the observed window start over the inferred
`resetAt − 7d`, falling back to snapshot inference when no event is recorded
yet. `quotaCache.setQuotaCache` records the transition opportunistically.

`recordProviderQuotaResetEventIfChanged` only fires for the primary weekly
window (not daily/sonnet), is idempotent (INSERT OR IGNORE on the unique
window key), and no-ops when the reset didn't actually roll. 4 unit tests
(tests/unit/lib/quota-reset-events.test.ts).

Salvaged from closed PR #5025 (which bundled this with two unrelated
features + a colliding migration 104). Renumbered to 108; module re-exported
from localDb (Rule #2).

Co-authored-by: Witroch4 <175152067+Witroch4@users.noreply.github.com>

---------

Co-authored-by: herjarsa <herjarsa@users.noreply.github.com>
Co-authored-by: Witroch4 <175152067+Witroch4@users.noreply.github.com>

* docs(i18n): sync 3.8.38 CHANGELOG section to 41 mirrors (unblock docs-accuracy) (#5144)

The root CHANGELOG [3.8.38] section grew with this cycle's merged PRs, but the
docs/i18n/<lang>/CHANGELOG.md mirrors were not re-synced — drifting >25% in body
size and failing check:docs-sync (the "Docs accuracy" fast-gate step) for every
open PR against the release.

Ran scripts/release/sync-changelog-i18n.mjs 3.8.38 3.8.37 to copy the root
[3.8.38] section into all 41 mirrors. check:docs-all now passes (exit 0).

Sections are copied verbatim; the per-language translation pass runs at release
time via i18n:run — this only restores the size-sync the gate enforces.

* feat(compression): pure per-step fidelity checker (4 invariants, fail-open)

* feat(compression): fidelityGate config + rejected breakdown fields

* feat(compression): wire per-step fidelity gate into stacked pipeline (opt-in)

* feat(compression): preview route accepts fidelityGate flag (playground)

* feat(compression): playground fidelity-gate toggle + lane rejection display

* docs(compression): note fidelityGate advanced thresholds are intentionally API-omitted

* refactor(compression): extract fidelity-gate step helpers to shrink strategySelector (file-size gate)

bodyToText and gateAdvance moved to fidelityGateStep.ts; StackAccumulator exported.
strategySelector: 889->854 (-35). Residual +6 vs pre-Milestone-B frozen 848 is the
irreducible StackOptions.fidelityGate field + two stacked-loop dispatch reads + import.
Baseline updated to 854 with justification. No cycle introduced (import type only).
940 compression tests pass; typecheck clean.

* test(usage): wire usageHistoryDedup under unit runner brace-list (#5145)

Integrated into release/v3.8.38.

* feat: salvage batch from closed stale PRs (#5038, #5057, #5076) (#5138)

Integrated into release/v3.8.38.

* test(combo): deterministic routing-decision matrix for all 17 strategies (#5146)

Integrated into release/v3.8.38.

* feat(compression): fuzzy near-duplicate dedup (session-dedup 2nd pass + playground toggle) (#5143)

Integrated into release/v3.8.38.

* chore(quality): rebaseline file-size for sidebarVisibility.ts + chat.ts drift (#5147)

Mid-cycle drift on release/v3.8.38 from already-merged PRs that the fast-path
(PR->release skips check:file-size) let accumulate without a bump:

- src/shared/constants/sidebarVisibility.ts 1100->1198 (#3812 colored menu
  icons, per-item accent map; #5142 dropped one orphan, net still above frozen)
- src/sse/handlers/chat.ts 1560->1575 (#5064 self-inflicted-timeout cooldown
  skip + #5124 long OpenAI-compatible SSE hardening + #5110 embed-WS
  LIVE_WS_HOST honour / early empty-message reject)

Each covered by its own PR tests; structural shrink of chat.ts tracked in #3501.
Unblocks the Fast Quality Gates for PRs targeting release/v3.8.38.

* chore(release): finalize v3.8.38 CHANGELOG + cycle reconciliation

- Reconcile [3.8.38]: +18 bullets (compression fidelity-gate/fuzzy-dedup #5143,
  quota keepalive #5102, web-session robustness #5121, MiniMax/Nemotron #5136,
  model-visibility #5091, failover logs #5016, disconnect races #5007, sidebar
  orphan #5142, SRE playbooks salvage #5138, new Security #5130 + Maintenance roll-up)
- Credit salvaged-PR authors (@JxnLexn / @KooshaPari / @herjarsa / @Witroch4)
- Remove phantom bullet for CLOSED-not-merged #5092 (setup aggregator never landed)
- Fix isHidden bullet PR citation #4389 -> #5086 (@herjarsa)
- Back-fill forgotten v3.8.36 bullet: #5026 crypto.randomUUID ID-gen (@hamsa0x7)
- Sync 41 i18n CHANGELOG mirrors; README What's New -> v3.8.38
- Rebaseline cycle drift: eslint 3987->4002, cognitive 833->841, dead-exports
  345->346, cyclomatic 1978->1980 (file-size handled by #5147)

* fix(i18n): add missing English UI labels (#5153)

Integrated into release/v3.8.38

* Preserve non-stream reasoning fields for compatible clients (#5155)

Integrated into release/v3.8.38

* feat(compression): ionizer engine — lossy JSON-array sampling reversible via CCR (#5148)

Integrated into release/v3.8.38

* test(combo): gated live smoke for combo strategies (in-process + VPS HTTP) (#5151)

Integrated into release/v3.8.38

* test: refresh release expectations to match current code (#5150)

Integrated into release/v3.8.38 (test-only base-red alignment extracted from #5150)

---------

Co-authored-by: Éder Costa <eder.almeida.costa@gmail.com>
Co-authored-by: José Victor Ferreira <root@josevictor.me>
Co-authored-by: Hernan Javier Ardila Sanchez <hjasgr@gmail.com>
Co-authored-by: fulorgnas <46461624+fulorgnas@users.noreply.github.com>
Co-authored-by: Randi <55005611+rdself@users.noreply.github.com>
Co-authored-by: Jan Leon <Jan.gaschler@gmail.com>
Co-authored-by: R. Beltran <rbeltran8000@gmail.com>
Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>
Co-authored-by: KooshaPari <42529354+KooshaPari@users.noreply.github.com>
Co-authored-by: Ramel Tecnologia - Rafa Martins <146174365+rafacpti23@users.noreply.github.com>
Co-authored-by: herjarsa <herjarsa@users.noreply.github.com>
Co-authored-by: Witroch4 <175152067+Witroch4@users.noreply.github.com>
2026-06-27 09:07:12 -03:00

107 KiB
Raw Blame History

title, version, lastUpdated
title version lastUpdated
Provider Reference 3.8.31 2026-06-20

Provider Reference

Auto-generated from src/shared/constants/providers.ts — do not edit by hand. Regenerate with: npm run gen:provider-reference Last generated: 2026-06-20

Total providers: 231. See category breakdown below.

Categories

  • Free — free tier with API key (configured via dashboard)
  • OAuth — sign-in flow handled by OmniRoute, no API key needed
  • Web cookie — wraps the provider's web app via cookie auth
  • API key — paid provider configured via API key (free credits may apply)
  • Local — runs on the user's machine (Ollama, LM Studio, vLLM, etc.)
  • Search — web search providers
  • Audio — audio-only providers (TTS/STT)
  • Upstream proxy — providers that proxy to other providers
  • Cloud agent — long-running coding agents (Codex Cloud, Devin, Jules)
  • System — OmniRoute-internal providers (loopback, etc.)

Additional tags: image, video, aggregator, enterprise, embed/rerank, self-hosted.

Use the dashboard at /dashboard/providers to enable, configure, and test each provider.


OAuth Providers (19)

ID Alias Name Tags Website Notes
agy agy Antigravity CLI OAuth link Import your Antigravity CLI (agy) login (paste/upload its token file), auto-detect a local CLI login, or sign in with Google. Shares the Antigravity backend (incl. Claude models).
amazon-q aq Amazon Q OAuth link Uses the same AWS Builder ID or imported refresh-token flow as Kiro, but keeps Amazon Q connections separate.
antigravity Antigravity OAuth
claude cc Claude Code OAuth
cline cl Cline OAuth
codex cx OpenAI Codex OAuth
cursor cu Cursor IDE OAuth
devin-cli dv Devin CLI (Official) OAuth link Requires the Devin CLI binary. Run devin auth login to authenticate, or provide your WINDSURF_API_KEY. Install: https://cli.devin.ai
gemini-cli gemini-cli Gemini CLI OAuth Uses Gemini CLI OAuth / Cloud Code credentials. Pro models require an eligible Google account or paid plan.
github gh GitHub Copilot OAuth
gitlab-duo gitlab-duo GitLab Duo OAuth link OAuth application with ai_features + read_user scopes. Configure GITLAB_DUO_OAUTH_CLIENT_ID and optionally GITLAB_DUO_OAUTH_CLIENT_SECRET on this OmniRoute instance.
kilocode kc Kilo Code OAuth
kimi-coding kmc Kimi Coding OAuth
kiro kr Kiro AI OAuth Free tier: 50 credits/month (~25K100K tokens). ⚠️ Kiro ToS prohibits third-party proxy/harness use.
qoder if Qoder AI OAuth
qwen qw Qwen Code OAuth ⚠️ DEPRECATED. Qwen OAuth free tier was discontinued on 2026-04-15. Use 'bailian-coding-plan', 'alibaba', 'alibaba-cn', or 'openrouter' provider with API key instead.
trae tr Trae OAuth link Trae is an AI-native IDE by ByteDance (SOLO remote agent). Authorize via trae.ai in the popup, or sign in at solo.trae.ai and paste the Cloud-IDE-JWT (sent as 'Authorization: Cloud-IDE-JWT ', ~14-day lifetime) as the access token; web_id/biz_user_id/user_unique_id/scope/tenant/region propagate via providerSpecificData. No headless refresh for pasted tokens — re-paste on expiry.
windsurf ws Windsurf (Devin CLI) OAuth link In the Windsurf / VS Code IDE, open the command palette and run Windsurf: Provide Auth Token (or click the Jupyter "Get Windsurf Authentication Token" button), then copy the shown token and paste it here. Note: opening windsurf.com/show-auth-token directly only renders a "Redirecting" page — the IDE must initiate the flow (it adds a ?state=... param) for the token to appear.
zed zd Zed IDE OAuth link Zed stores LLM provider credentials (OpenAI, Anthropic, Google, Mistral, xAI) in the OS keychain. Use the Import button below to discover and import them automatically.
ID Alias Name Tags Website Notes
adapta-web adp-web Adapta.org (Adapta One Web) Web cookie link Paste your __client cookie value from .clerk.agent.adapta.one (DevTools → Application → Cookies)
blackbox-web bb-web Blackbox Web (Subscription) Web cookie link Paste your __Secure-authjs.session-token value or full cookie header from app.blackbox.ai
chatgpt-web cgpt-web ChatGPT Web (Plus/Pro) Web cookie link Paste your __Secure-next-auth.session-token cookie value from chatgpt.com
claude-web cw Claude Web Web cookie link Paste your session cookie from claude.ai
copilot-web copilot Microsoft Copilot Web Web cookie link Paste your access_token from copilot.microsoft.com (or export a .har file from DevTools while logged in)
deepseek-web ds-web DeepSeek Web Web cookie link Paste your userToken from chat.deepseek.com — DevTools → Application → Local Storage → userToken
doubao-web db Doubao Web (ByteDance) Web cookie link Paste your session cookie from doubao.com (DevTools → Application → Cookies)
gemini-business gembiz Gemini Business (Enterprise) Web cookie link From your enterprise account: open business.gemini.google/home/cid/{your-cid}, then copy **Secure-1PSID and **Secure-1PSIDTS cookies from DevTools → Application → Cookies. Paste as a cookie header below.
gemini-web gweb Gemini Web (Free) Web cookie link Paste your **Secure-1PSID cookie value from gemini.google.com. Optionally add **Secure-1PSIDTS separated by semicolon.
grok-web gw Grok Web (Subscription) Web cookie link Paste the full grok.com cookie line from DevTools → Application → Cookies. Include both sso and sso-rw (e.g. sso=...; sso-rw=...) — Grok's anti-bot rejects sso on its own.
huggingchat huggingchat HuggingChat (Free) Web cookie link Paste your hf-chat cookie value from huggingface.co/chat (DevTools → Application → Cookies → hf-chat). Optional — works without auth for basic use.
inner-ai in-ai Inner.ai (Subscription) Web cookie link Paste your token cookie and email separated by a space: open DevTools → Application → Cookies → .innerai.com, copy the token value, then append a space and your Inner.ai login email. Example: eyJhbG... user@example.com
kimi-web kimi-web Kimi Web (Moonshot AI) Web cookie link Paste your session cookie from kimi.moonshot.cn (DevTools → Application → Cookies)
lmarena lma LMArena (Free) Web cookie link Paste the full Cookie header from lmarena.ai (DevTools → Network → request → Cookie). The session is now split across arena-auth-prod-v1.0, .1, … — copy the whole header. Optional — works with free tier for basic comparisons.
muse-spark-web ms-web Muse Spark Web (Meta AI) Web cookie link Paste your abra_sess value or full cookie header from meta.ai
perplexity-web pplx-web Perplexity Web (Pro/Max) Web cookie link Paste your __Secure-next-auth.session-token cookie value from perplexity.ai
phind ph Phind (Free) Web cookie link ⚠️ DEPRECATED. Phind shut down its API (2026-01); the /api/chat endpoint no longer serves (sweep 2026-06-19).
poe-web poe Poe Web (Subscription) Web cookie link Paste your p-b cookie value from poe.com (DevTools → Application → Cookies → p-b)
qwen-web qwen-web Qwen Web (Free) Web cookie link Open chat.qwen.ai, log in, then open DevTools → Application → Local Storage → copy the "token" value (or use tongyi_sso_ticket cookie as Bearer token).
t3-web t3chat t3.chat (Pro/Free) Web cookie link Open t3.chat in your browser, log in, then open DevTools → Application → Local Storage → https://t3.chat. Copy the value of 'convex-session-id'. Also open DevTools → Network, copy the Cookie header from any request. Paste both values here. See provider setup docs for a step-by-step guide.
v0-vercel-web v0 v0 Vercel Web (Code Gen) Web cookie link Paste your session cookie from v0.dev (DevTools → Application → Cookies)
venice-web ven Venice Web (Privacy) Web cookie link Paste your session cookie from venice.ai (DevTools → Application → Cookies)

API Key Providers (paid / paid-with-free-credits) (157)

ID Alias Name Tags Website Notes
360ai 360ai 360 AI API key link Get API key at ai.360.cn
agentrouter agentrouter AgentRouter API key, aggregator link $200 free credits on signup - multi-model routing gateway
ai21 ai21 AI21 Labs API key link $10 trial credits on signup (valid 3 months), no credit card required
aimlapi aiml AI/ML API API key, aggregator link Free tier paused (2026) — AI/ML API is now pay-as-you-go only (min $20 top-up); no recurring free credits.
alibaba ali Alibaba API key link
alibaba-cn ali-cn Alibaba (China) API key link
anthropic anthropic Anthropic API key link
api-airforce af Api.airforce API key link 55 free tier models including Grok-3, Claude 3.7, Qwen3, Kimi-K2, Gemini 2.5 Flash, DeepSeek-V3
arcee-ai arcee Arcee AI API key link Get API key at arcee.ai
azure-ai azure-ai Azure AI Foundry API key, enterprise link Use your Azure AI Foundry key. Base URL can be https://.services.ai.azure.com/openai/v1/ or https://.openai.azure.com/openai/v1/.
azure-openai azure Azure OpenAI API key, enterprise link Use your Azure OpenAI API key. Base URL should be your resource endpoint, for example https://my-resource.openai.azure.com.
baichuan baichuan Baichuan API key link Get API key at platform.baichuan-ai.com
baidu baidu Baidu (ERNIE) API key link Get API key at console.bce.baidu.com
bailian-coding-plan bcp Alibaba Coding Plan API key link
baseten baseten Baseten API key link $30 free trial credits for GPU inference
bazaarlink bzl BazaarLink API key link Free tier with auto:free routing — zero-cost inference, no credit card required
bedrock bedrock Amazon Bedrock API key, enterprise link Use your Amazon Bedrock API key and configure the AWS region where your models are enabled (for example eu-west-2). OmniRoute calls Bedrock's native Converse API directly.
black-forest-labs bfl Black Forest Labs API key, image link
blackbox bb Blackbox AI API key link Free tier: unlimited basic chat plus Minimax-M2.5, no credit card required
bluesminds bm BluesMinds API key link Free daily pi credits — supports 200+ models including GPT-4o, GPT-4.1, Claude Sonnet 4.5, Gemini 2.0 Flash, DeepSeek V4, Qwen, Kimi K2
byteplus bpm BytePlus ModelArk API key link
bytez bytez Bytez API key link $1 free credits, refreshes every 4 weeks
cablyai cablyai CablyAI API key, aggregator link Bearer API key for the CablyAI OpenAI-compatible gateway.
cerebras cerebras Cerebras API key link Free Trial: 1M tokens/day, 30K TPM, 5 RPM — no credit card.
chutes chutes Chutes.ai API key, aggregator link Bearer API key for the Chutes OpenAI-compatible gateway.
clarifai clarifai Clarifai API key, enterprise link Use your Clarifai PAT or app-specific API key. OmniRoute targets the OpenAI-compatible endpoint at https://api.clarifai.com/v2/ext/openai/v1 and authenticates with Authorization: Key .
cloudflare-ai cf Cloudflare Workers AI API key link Requires API Token AND Account ID (found at dash.cloudflare.com)
codestral codestral Codestral API key link
cohere cohere Cohere API key link Free Trial: 1,000 API calls/month for testing, no credit card required
command-code cmd Command Code API key link Use a Command Code API key. Requests are sent to Command Code's /alpha/generate endpoint.
coze coze Coze API key link Get API key at coze.com/open/api
crof crof CrofAI API key link
databricks databricks Databricks API key, enterprise link
datarobot datarobot DataRobot API key, enterprise link Use your DataRobot API token. Optional Base URL can be the account root (for LLM Gateway) or a deployment URL under /api/v2/deployments/.
deepinfra deepinfra DeepInfra API key link Free signup credits for API testing and model exploration
deepseek ds DeepSeek API key link 5M free tokens on signup - no credit card required
dify dify Dify API key link Get API key from your Dify instance.
dit dai DIT.ai API key link Use your dit.ai API key in Authorization: Bearer . Fully OpenAI-compatible — a drop-in replacement, just change the base URL to https://api.dit.ai/v1.
doubao doubao Doubao API key link Get API key at console.volcengine.com
empower empower Empower API key, aggregator link Bearer API key for the Empower OpenAI-compatible endpoint.
fal-ai fal Fal.ai API key, image link
factory factory Factory API key, aggregator link Bearer API key for the Factory (Factory Droids) OpenAI-compatible gateway. Same backend the droid CLI uses; subscription tier via app.factory.ai. OAuth follow-up tracked in issue #5060.
featherless-ai featherless Featherless AI API key link Free tier available — no credit card required
fenayai fenayai FenayAI API key, aggregator link Bearer API key for the FenayAI OpenAI-compatible gateway.
firecrawl fc Firecrawl API key link
fireworks fireworks Fireworks AI API key link $1 free starter credits on signup for API testing
freeaiapikey faik FreeAIAPIKey API key link
freemodel-dev fmd FreeModel.dev API key link $300 free credits on signup — no credit card required. Access GPT-5.4 and GPT-5.5 (OpenAI's latest flagship models) through an OpenAI-compatible API.
friendliai friendli FriendliAI API key link Free tier for serverless inference — no credit card required
galadriel galadriel Galadriel API key link ⚠️ DEPRECATED. api.galadriel.ai no longer resolves (sweep 2026-06-19); the inference API appears discontinued.
gemini gemini Gemini (Google AI Studio) API key link Free forever: 1,500 req/day for Gemini 2.5 Flash — no credit card, get key at aistudio.google.com
getgoapi ggo GoAPI API key, aggregator link
gigachat gigachat GigaChat (Sber) API key link
github-models ghm GitHub Models API key link Create a GitHub PAT with 'models: read' scope at github.com/settings/tokens
gitlab gitlab GitLab Duo PAT API key link GitLab personal access token for the public Code Suggestions API. Configure a self-hosted base URL when not using gitlab.com.
gitlawb glb Gitlawb Opengateway (MiMo) API key link Free MiMo (xiaomi/mimo-v2.5) revoked 2026-05 — Opengateway is now a pay-as-you-go credit gateway; no recurring free model.
gitlawb-gmi glb-gmi Gitlawb Opengateway (GMI Cloud) API key link Free Nemotron promo ended 2026-06 — the GMI Cloud route is now pay-as-you-go credit only.
glhf glhf GLHF Chat API key, aggregator link ⚠️ DEPRECATED. glhf.chat shut down (2026); its api.laf.run gateway no longer serves the catalog (sweep 2026-06-19).
glm glm GLM Coding API key link
glm-cn glmcn GLM Coding (China) API key link
glmt glmt GLM Thinking API key link
groq groq Groq API key link Free tier: 30 RPM / 14.4K RPD — no credit card
hackclub hc Hackclub AI API key, aggregator link Sign in with your Hack Club account at ai.hackclub.com.
haiper hp Haiper API key, video link Get API key at haiper.ai/haiper-api
heroku heroku Heroku AI API key, enterprise link
huggingchat huggingchat HuggingChat API key link No API key required for basic access.
huggingface hf HuggingFace API key link Free Inference API for thousands of models (Whisper, VITS, SDXL…)
hyperbolic hyp Hyperbolic API key link $1-5 trial credits on signup for serverless inference
ideogram ideo Ideogram API key link Get API key at ideogram.ai/docs/api
iflytek iflytek iFlytek Spark API key link Get API key at console.xfyun.cn
inclusionai inclusion InclusionAI API key link ⚠️ DEPRECATED. api.inclusionai.tech no longer resolves (sweep 2026-06-19); the inference API appears discontinued.
inference-net inet Inference.net API key link $25 free credits on signup plus research grants available
jina-ai jina Jina AI API key, embed/rerank link Bearer API key for the Jina AI rerank API.
jina-reader jr Jina Reader API key link
kie kie KIE.AI API key link
kilo-gateway kg Kilo Gateway API key, aggregator link
kimi kimi Kimi API key link
kimi-coding-apikey kmca Kimi Coding (API Key) API key link
kluster kluster Kluster AI API key link ⚠️ DEPRECATED. kluster.ai shut down (2026-06-09); api.kluster.ai no longer resolves (sweep 2026-06-19). Use another OpenAI-compatible provider.
lambda-ai lambda Lambda AI API key link
laozhang lz LaoZhang AI API key, aggregator link
leonardo leo Leonardo AI API key, video link Get API key at leonardo.ai/developer
liquid liquid Liquid AI API key link Get API key at liquid.ai
llamagate llamagate LlamaGate API key link
llm7 llm7 LLM7.io API key link No signup required - 2 req/s, 20 RPM, 100 req/hr free tier
longcat lc LongCat AI API key link Free: 5M tokens/day on LongCat-2.0-Preview (Flash models retired 2026-05-29); up to 120M/day via feedback.
maritalk maritalk Maritalk API key link
meta-llama meta Meta Llama API API key link
minimax minimax Minimax Coding API key, video link
minimax-cn minimax-cn Minimax (China) API key link
mistral mistral Mistral API key link Free Experiment tier: rate-limited access to all models, no credit card required
modal mdl Modal API key, enterprise link Use the bearer token that protects your Modal deployment, if enabled. Base URL should point to your OpenAI-compatible Modal app, for example https://--.modal.run/v1.
monsterapi monster MonsterAPI API key link Get API key at monsterapi.ai
moonshot moonshot Moonshot AI API key link
morph morph Morph API key link Free tier: 250K credits/month, $0
nanogpt nanogpt NanoGPT API key link
nebius nebius Nebius AI API key link ~$1 trial credits on signup for API testing
nlpcloud nlpc NLP Cloud API key link Use your NLP Cloud API key in Authorization: Token . OmniRoute targets the chatbot endpoint on https://api.nlpcloud.io/v1/gpu//chatbot by default.
nomic nomic Nomic API key link Get API key at atlas.nomic.ai
nous-research nous Nous Research API key link Use your Nous Portal API key. OmniRoute targets the official OpenAI-compatible inference endpoint at https://inference-api.nousresearch.com/v1.
novita novita Novita AI API key, aggregator link $0.50 trial credits on signup (valid about 1 year)
nscale nscale nScale API key link $5 free credits on signup for inference testing
nvidia nvidia NVIDIA NIM API key link Free dev access: ~40 RPM, 70+ models (Kimi K2.5, GLM 4.7, DeepSeek V3.2...)
oci oci OCI Generative AI API key, enterprise link Use your OCI Generative AI API key or IAM bearer token. Base URL can be https://inference.generativeai..oci.oraclecloud.com/openai/v1/.
ollama-cloud ollamacloud Ollama Cloud API key link
openadapter oad OpenAdapter API key link Use your OpenAdapter API key in Authorization: Bearer sk-cv-. Fully OpenAI-compatible. API base URL: https://api.openadapter.in/v1.
openai openai OpenAI API key link
opencode-go opencode-go OpenCode Go API key link
opencode-zen opencode-zen OpenCode Zen API key link
openrouter openrouter OpenRouter API key, aggregator link Free models at $0/token with :free suffix - 20 RPM / 200 RPD
orcarouter orcarouter OrcaRouter API key link
ovhcloud ovh OVHcloud AI API key link
perplexity pplx Perplexity API key link
phind phind Phind API key link Get API key at phind.com
piapi pi PiAPI API key, aggregator link
poe poe Poe API key, aggregator link Bearer API key for the Poe OpenAI-compatible API.
pollinations pol Pollinations AI API key, video link Free keyless tier: openai, openai-fast, openai-large, qwen-coder, mistral, deepseek, grok, gemini-flash-lite-3.1, perplexity-fast, perplexity-reasoning. Premium models (claude, gemini, midijourney) require a Pollinations API key from enter.pollinations.ai.
predibase predibase Predibase API key link ⚠️ DEPRECATED. serving.app.predibase.com no longer resolves (sweep 2026-06-19); the managed serving API appears discontinued.
publicai publicai PublicAI API key link Requires an API key — one-time signup credit, then paid
puter pu Puter AI API key link Get token at puter.com/dashboard → Copy Auth Token
qianfan qianfan Baidu Qianfan API key link
recraft recraft Recraft API key, image link
reka reka Reka API key link Use your Reka API key. OmniRoute supports the OpenAI-compatible base URL https://api.reka.ai/v1 and sends both Authorization and X-Api-Key headers for compatibility.
runwayml runway Runway API key, video link Use your Runway API key in Authorization: Bearer . OmniRoute targets the current Runway API at https://api.dev.runwayml.com/v1 and sends the required X-Runway-Version header automatically.
sambanova samba SambaNova API key link $5 free credits on signup (30-day validity), no credit card required
sap sap SAP Generative AI Hub API key, enterprise link Use your SAP AI Core bearer token. Base URL can be your AI_API_URL root or a deploymentUrl from Generative AI Hub.
scaleway scw Scaleway AI API key link 1M free tokens for new accounts — EU/GDPR compliant (Paris), Qwen3 235B & Llama 70B
sensenova sensenova SenseNova API key link Get API key at platform.sensenova.cn
siliconflow siliconflow SiliconFlow API key link $1 free credits plus permanently free models after identity verification
snowflake snowflake Snowflake Cortex API key, enterprise link
sparkdesk sparkdesk SparkDesk API key link Get API key at console.xfyun.cn
stability-ai stability Stability AI API key, image link
stepfun stepfun StepFun API key link Get API key at platform.stepfun.com
suno suno Suno API key link Paste session cookie from suno.ai (Clerk auth)
synthetic synthetic Synthetic API key, aggregator link
tencent tencent Tencent Hunyuan API key link Get API key at console.cloud.tencent.com
thebai thebai TheB.AI API key, aggregator link Bearer API key for the TheB.AI OpenAI-compatible gateway.
together together Together AI API key, video link $25 signup credits + 3 permanently free models: Llama 3.3 70B, Vision, DeepSeek-R1 distill
tokenrouter trk TokenRouter API key link Use your TokenRouter API key in Authorization: Bearer . Fully OpenAI-compatible. API base URL: https://api.tokenrouter.com/v1.
topaz topaz Topaz API key, image link
udio udio Udio API key link Paste session cookie from udio.com (Supabase auth)
uncloseai unc UncloseAI API key link No auth required. API accepts any non-empty string as key for identification.
upstage upstage Upstage API key link
v0-vercel v0 v0 (Vercel) API key link
venice venice Venice.ai API key link
vercel-ai-gateway vag Vercel AI Gateway API key, aggregator link
vertex vertex Vertex AI API key, enterprise link Provide Service Account JSON or OAuth access_token
vertex-partner vp Vertex AI Partners API key, enterprise link Provide the same Service Account JSON used for Vertex AI partner models.
volcengine volcengine Volcengine API key link
voyage-ai voyage Voyage AI API key, embed/rerank link Bearer API key for Voyage AI embeddings and rerank APIs.
wafer wafer Wafer AI API key link
wandb wandb Weights & Biases Inference API key link
watsonx watsonx IBM watsonx.ai Gateway API key, enterprise link Use your watsonx bearer token. Base URL can be https://.ml.cloud.ibm.com/ml/gateway/v1/ or a self-managed /ml/gateway/v1 endpoint.
xai xai xAI (Grok) API key link
xiaomi-mimo mimo Xiaomi MiMo API key link
yi yi Yi (01.AI) API key link Get API key at platform.lingyiwanwu.com
zai zai Z.AI API key link
zenmux zm ZenMux API key link Use your ZenMux API key in Authorization: Bearer . ZenMux is fully OpenAI-compatible. Base URL: https://zenmux.ai/api/v1.

Local Providers (11)

ID Alias Name Tags Website Notes
comfyui comfyui ComfyUI Local link No API key required. Configure the local ComfyUI base URL (default: http://localhost:8188).
docker-model-runner dmr Docker Model Runner Local, self-hosted link API key optional. Configure the local Docker Model Runner OpenAI-compatible base URL (default: http://localhost:12434/v1).
lemonade lemonade Lemonade Server Local, self-hosted link API key optional. Configure the local Lemonade OpenAI-compatible base URL (default: http://localhost:13305/api/v1).
llama-cpp llamacpp llama.cpp Local, self-hosted link API key optional (use any value, e.g. sk-no-key-required). Configure the llama-server OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1). Note: if Llamafile is also installed, both default to port 8080 — run only one at a time or override the port.
llamafile llamafile Llamafile Local, self-hosted link API key optional. Configure the local Llamafile OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1).
lm-studio lmstudio LM Studio Local, self-hosted link API key optional. Configure the local LM Studio OpenAI-compatible base URL (default: http://localhost:1234/v1).
oobabooga ooba oobabooga Local, self-hosted link API key optional. Configure the local oobabooga OpenAI-compatible base URL (default: http://localhost:5000/v1).
sdwebui sdwebui SD WebUI Local link No API key required. Configure the local WebUI base URL (default: http://localhost:7860).
triton triton NVIDIA Triton Local, self-hosted link API key optional. Configure the Triton OpenAI-compatible base URL (default: http://localhost:8000/v1).
vllm vllm vLLM Local, self-hosted link API key optional. Configure the local vLLM OpenAI-compatible base URL (default: http://localhost:8000/v1).
xinference xinference XInference Local, self-hosted link API key optional. Configure the local XInference OpenAI-compatible base URL (default: http://localhost:9997/v1).

Search Providers (11)

ID Alias Name Tags Website Notes
brave-search brave-search Brave Search Search link Subscription token from Brave Search API dashboard
exa-search exa-search Exa Search Search link API key from dashboard.exa.ai
google-pse-search google-pse Google Programmable Search Search link Requires a Google API key and your Programmable Search Engine ID (cx)
linkup-search linkup Linkup Search Search link Bearer API key from the Linkup dashboard
ollama-search ollama-search Ollama Search Search link Same API key as Ollama Cloud (from ollama.com/settings/api-keys)
perplexity-search pplx-search Perplexity Search Search link Same API key as Perplexity (pplx-...)
searchapi-search searchapi SearchAPI Search link API key from SearchAPI (query param or Bearer auth)
searxng-search searxng SearXNG Search Search link API key is optional. Set your SearXNG base URL. Some instances may require a bearer token for access.
serper-search serper-search Serper Search Search link API key from serper.dev dashboard
tavily-search tavily-search Tavily Search Search link API key from app.tavily.com (format: tvly-...)
youcom-search youcom-search You.com Search Search link X-API-Key from the You.com platform dashboard

Audio-only Providers (7)

ID Alias Name Tags Website Notes
assemblyai aai AssemblyAI Audio link
aws-polly polly AWS Polly Audio link Use AWS Secret Access Key as API key; set providerSpecificData.accessKeyId and optional region.
cartesia cartesia Cartesia Audio link
deepgram dg Deepgram Audio link
elevenlabs el ElevenLabs Audio link
inworld inworld Inworld Audio link
playht playht PlayHT Audio link

Upstream Proxy Providers (2)

ID Alias Name Tags Website Notes
9router nr 9router Upstream proxy link
cliproxyapi cpa CLIProxyAPI Upstream proxy link

Cloud Agent Providers (3)

ID Alias Name Tags Website Notes
codex-cloud codex-cloud Codex Cloud Cloud agent link OpenAI API key with Codex Cloud task access.
devin devin Devin Cloud agent link Devin API key for cloud agent sessions.
jules jules Google Jules Cloud agent link Jules API key for creating and managing cloud coding tasks.

System Providers (1)

ID Alias Name Tags Website Notes
auto auto Auto (Zero-Config) System

Sources of truth

See Also