mirror of
https://github.com/diegosouzapw/OmniRoute.git
synced 2026-07-26 09:52:11 +03:00
* chore(release): open v3.8.38 development cycle
* fix(executors): strip client_metadata for cerebras and mistral (#4727)
Integrated into release/v3.8.38 (leva 5)
* fix(codebuddy): only send reasoning params when client requests reasoning (#5019)
Integrated into release/v3.8.38 (leva 5)
* fix(sse): keep streaming for forceStream providers when client requests JSON (#5021)
Integrated into release/v3.8.38 (leva 5)
* fix(sse): guard non-JSON SSE lines and duplicate [DONE] (#4937)
Integrated into release/v3.8.38 (leva 5)
* feat(blackbox): refresh provider model catalog (#4935)
Integrated into release/v3.8.38 (leva 5)
* fix(sse): dedupe case-variant Anthropic version/beta headers (#4846)
Integrated into release/v3.8.38 (leva 5)
* feat(sse): Kiro inline <thinking> stream splitter (#4911)
Integrated into release/v3.8.38 (leva 5)
* feat(cursor): parse Composer DeepSeek-style inline tool calls (#4912)
Integrated into release/v3.8.38 (leva 5)
* feat(proxy): auth-less host:port batch import (#4938)
Integrated into release/v3.8.38 (leva 5)
* fix(oauth): support Kiro IDC (organization) token import (#4944)
Integrated into release/v3.8.38 (leva 5)
* fix(translator): preserve cache_control for DashScope OpenAI-compat providers (port from 9router#2069) (#5013)
Integrated into release/v3.8.38 (leva 5)
* fix(tts): resolve Gemini TTS models from catalog (#4934)
Integrated into release/v3.8.38 (leva 5)
* fix(sse): don't cool down the connection on a self-inflicted upstream timeout (504) (#5064)
Integrated into release/v3.8.38 (leva 5)
* fix(sse): robust Anthropic /v1/messages streaming — real ping keepalive + client-disconnect guard (#5063)
Integrated into release/v3.8.38 (leva 5)
* feat(video): add Alibaba DashScope (wan2.7-t2v) provider (#5051)
Integrated into release/v3.8.38 (leva 5)
* fix: preserve model hidden flags (isHidden) across model sync (#5086)
Integrated into release/v3.8.38 (leva 5)
* fix(models): derive model discovery config from registry modelsUrl (#5087)
Integrated into release/v3.8.38 (leva 5)
* fix(compression): replace fileURLToPath(import.meta.url) with runtime anchors for standalone bundle (#5089)
Integrated into release/v3.8.38 (leva 5)
* feat(cc): add summarized thinking display toggle (#5055)
Integrated into release/v3.8.38 (leva 5)
* Harden selected API error responses (#5032)
Integrated into release/v3.8.38 (leva 5)
* chore(quality): rebaseline file-size for leva 5 PR batch drift
6 frozen files grew from merged leva-5 PRs (cursor #4912, kiro #4911,
videoGeneration #5051, default #4727, base #4846, chat #5064); all covered
by per-PR tests. See _rebaseline_2026_06_26_leva5 in the baseline.
* feat(compression): compression playground (Play + Compare tabs) in the studio (#5080)
Integrated into release/v3.8.38
* fix(combo): fail over on empty-content 502 instead of exhausting the provider (#5085) (#5104)
* fix(dashboard): surface detailed credential-validation error in add-connection modal (#5088) (#5106)
* feat(providers): allow local/private provider URLs by default with scoped metadata-safe guard (#5066) (#5107)
* fix(diagnostics): treat non-streaming Claude messages shape as valid output (#5108) (#5116)
* fix(db): translate pt-BR SQLite driver-fallback log lines to English (#5103) (#5115)
* fix(sse): repair release base-reds — malformed-response false positives + header casing + stale tests (#5117)
Repairs the release/v3.8.38 base-reds; unblocks #5078.
* chore(quality): rebaseline file-size for responseSanitizer (#5117) + AddApiKeyModal drift
* fix(translator): forward image tool_result blocks as image_url (#5100)
Base-reds fixed (#5117); image tool_result→image_url. Integrated into release/v3.8.38.
* fix(responses): default text.format for openai-compatible responses providers (#5101)
Base-reds fixed (#5117); default text.format + file-size rebaseline. Integrated into release/v3.8.38.
* feat(dashboard): expose Fusion judgeModel + fusionTuning in the combo editor (#5074)
Base-reds fixed (#5117); Fusion editor + file-size rebaseline. Integrated into release/v3.8.38.
* feat(quota): add opt-in Codex/Claude auto-ping keepalive (#5102)
Base-reds fixed (#5117); auto-ping keepalive + file-size rebaseline. Integrated into release/v3.8.38.
* test(release): relocate 2 orphan test files into the collected flat tests/unit dir (#5120)
Unblocks Lint (test-discovery) on #5078. Integrated into release/v3.8.38.
* fix(translator): preserve reasoning-replay reasoning_content + repair 3 release-green test reds (#5122)
Repairs 3 release-green test reds + test-masking; unblocks #5078.
* test(golden): redact live Node version from provider translate-path snapshot (#5125)
Final golden unblock for #5078.
* test(golden): redact OmniRoute app version from translate-path snapshot (#5126)
Coverage shard golden unblock for #5078.
* Ignore disconnect races during in-band stream error handling (#5007)
Integrated into release/v3.8.38
* Track final connection IDs in failover logs (#5016)
Integrated into release/v3.8.38
* fix(sse): convert Gemini body to OpenAI format in antigravity MITM handler (#4845)
Integrated into release/v3.8.38 (rebased on tip, CHANGELOG re-injected)
* feat(providers): add ZenMux Free session-cookie provider (#5105)
Integrated into release/v3.8.38 (rebased on tip, CHANGELOG re-injected)
* feat(dashboard): click-to-edit model alias in provider page (#5119)
Integrated into release/v3.8.38 (rebased on tip, i18n scope verified, CHANGELOG re-injected)
* feat(mcp): web-session robustness — cookie dedup (PR6) + browser-pool observability (PR7) (#3368) (#5121)
Integrated into release/v3.8.38 (rebased on tip; cookie-dedup branch extracted to findExistingCookieConnection helper → complexity-neutral; CHANGELOG added)
* fix(usage): dedupe request-usage logging and debounce stats (#4940)
Integrated into release/v3.8.38 (rebased on tip; DB-handle hang was stale-base artifact — resetDbInstance already closes the handle, test green 5/5; file-size drift consolidated at release; CHANGELOG re-injected)
* fix(dashboard): key model visibility toggle on canonical providerId (#5091)
Integrated into release/v3.8.38 (retargeted main→release; .tsx visibility-key test green 2/2)
* chore(deps): bump actions/cache from 5.0.5 to 6.0.0 (#5112)
Integrated into release/v3.8.38 (retargeted main→release; workflow-only actions/cache bump — unit failures were stale main base-reds)
* fix(streaming): harden long OpenAI-compatible SSE streams (#5124)
Integrated into release/v3.8.38 (rebased on tip; streamHandler conflict with #5007 disconnect-guard resolved — both coexist, stream-handler 22/22 green)
* feat: Add Grok Build (xAI) provider with OAuth import-token flow (#5020)
Integrated into release/v3.8.38 (rebased on tip; Hard Rule #11 fix — Grok public client_id now via resolvePublicCred(grok_id), 3 literals removed; grok-oauth 7/7 + check:public-creds green)
* feat(providers): add Factory (factory.ai) as a subscription gateway provider (#5065)
Integrated into release/v3.8.38 (rebased on tip; added factory registry test for PR Test Policy + fixed check:env-doc-sync phantom FACTORY_API_KEY; factory loads in PROVIDERS, no Zod issue — that flag was a false positive)
* chore(test): reconcile golden snapshot + apikey count for new providers
#5020 (grok-cli), #5065 (factory), #5105 (zenmux-free) added providers but did
not regenerate tests/snapshots/provider/translate-path.json (now +3 entries) nor
bump the APIKEY_PROVIDERS count (159->160 for the factory gateway). Test-only
reconciliation; no production change.
* fix(resilience): harden quota and model lockout edge cases (#5093)
Integrated into release/v3.8.38 (rebased on tip). TRUST-BUT-VERIFY: dropped the PR's 0dd7df641 'fix unit gates' commit which reverted #5122 reasoning-replay (preserveReasoningContent) + re-introduced #4849 O(n^2) growth, and restored 5 tests it had realigned. Kept only the 3 declared resilience fixes (quota cutoff guard, gemini MIME, model-lockout maxCooldownMs); 23/23 green.
* Hydrate quota cache and scope auto combo candidates (#5015)
Integrated into release/v3.8.38 (rebased on tip). Kept core quota-cache hydration + auto-combo candidate scoping + combos UI; dropped out-of-scope toolCloaking refactor (conflicted with #4813 stripEnumDescriptions — took tip) and the unrelated sse-auth test split. Added quota-cache-hydrate-5015 regression test (Rule #18); combo-account-allowlist 8/8 + hydration 2/2 green.
* chore(quality): reconcile complexity + file-size baselines for v3.8.38 owner-PR batch
complexity 1972->1978 (+6) and file-size providers.ts 1093->1107 / usageHistory.ts
934->983 — drift from the /review-prs merge batch (#4845/#5105/#5020/#4940/#5093/
#5015 + #5121 cookie-dedup helper extraction). check:complexity/check:file-size do
not run on the PR->release fast-path, so the branch accrued unmeasured; all legit
feature/fix growth, not regression. See per-key justifications in each baseline.
* fix(security): exact-host Anthropic baseUrl check (CodeQL js/incomplete-url-substring-sanitization #674) (#5130)
The anthropic-compatible Bearer-fallback gate decided whether a configured baseUrl
targeted the official api.anthropic.com host via a substring `.includes("api.anthropic.com")`.
A look-alike upstream such as `https://api.anthropic.com.evil.test` or
`https://evil.test/?x=api.anthropic.com` matched the substring and was wrongly treated as
official, suppressing the Bearer fallback meant for third-party gateways
(CodeQL #674, js/incomplete-url-substring-sanitization, high).
Replace the substring test with an exported `isOfficialAnthropicBaseUrl()` helper that
parses the URL and compares the hostname for exact equality. Empty baseUrl stays official;
scheme-less hosts are parsed with an assumed https://; an unparseable baseUrl falls back to
third-party (Bearer emitted) as the safer default. Behavior for legitimate official/third-party
baseUrls is unchanged.
Adds tests/unit/anthropic-official-baseurl-host.test.ts covering official, look-alike,
scheme-less, and unparseable inputs plus a static guard that the substring pattern is gone.
* fix(proxy): repair one-click Deno & Cloudflare relay deployments (#5128) (#5132)
* fix(services): embed WS proxy honours LIVE_WS_HOST; reject empty messages early (#5110) (#5133)
* fix(api): resolve /v1/models/{id} case-insensitively (#5082) (#5135)
* fix(providers): add MiniMax M3 & Nemotron 3 Ultra to Cline catalog (#3321) (#5136)
* fix(proxy): make SOCKS5 handshake timeout tunable via SOCKS_HANDSHAKE_TIMEOUT_MS (#5109) (#5137)
* feat(sidebar): add support for colored menu icons (#3812)
Integrated into release/v3.8.38 (recreated on tip — fork had unrelated history; added getSidebarIconAccent regression test, Rule #18). Clean 2-file UI feature.
* fix(providers): complete grok-cli OAuth wiring + zenmux-free web-session metadata
Base-red repair for #5020 (grok-cli) and #5105 (zenmux-free), surfaced by the
full CI on the release PR (#5078) — the PR->release fast-path does not run the
oauth-providers-config / web-session-credentials / provider-consistency gates.
- grok-cli: register in OAUTH_PROVIDERS (providers.ts canonical list, fixes
check:provider-consistency), add OAUTH_PROVIDER_IDS.GROK_CLI + GROK_CLI_CONFIG
in oauth constants (provider config now sourced there, not a local literal),
align oauth-providers-config.test.ts (EXPECTED_PROVIDER_KEYS + config map).
- zenmux-free: declare its web-session credential requirement (full Cookie header)
in WEB_SESSION_CREDENTIAL_REQUIREMENTS.
Local: oauth-providers-config 27/27, web-session-credentials 4/4, grok-cli-oauth
7/7, check:provider-consistency OK, +115 OAUTH_PROVIDERS tests green.
* Fix resilience settings page response mapping (#5139)
Integrated into release/v3.8.38. Thanks @rdself for the fix and the regression test.
* fix(kiro): retire claude-sonnet-4.5 from catalog + pin 400 model-unavailable test (#5140)
Extracted the real change from #5140 (the bot PR regenerated the entire
freeModelCatalog.data.ts + touched package-lock.json; only the targeted
edits are kept here):
- remove claude-sonnet-4.5 from the Kiro registry entry
- remove the matching kiro free-model catalog row
- pin Kiro's verbatim 400 "Invalid model..." to isModelUnavailableError
Closes #4484
* fix(sidebar): drop orphan `settings` accent color (typecheck:core red) (#5142)
SIDEBAR_ICON_ACCENTS is typed Partial<Record<HideableSidebarItemId, string>>,
but `settings` is not a hideable item id (only `settings-general`,
`settings-appearance`, … and `context-settings` exist; there is no item with
`id: "settings"`), so the accent was unreachable. It broke `typecheck:core`
on the release tip ("'settings' does not exist in type …", introduced by
#3812 colored menu icons). Removing the orphan key restores a clean
typecheck:core (rc=0).
* feat: salvage batch 2 — diagnostics null-guard (#5096) + observed quota reset windows (#5025) (#5141)
* fix(diagnostics): null-guard content blocks in detectMalformedNonStream
A null (or non-object) entry in a Claude-native `content` array made the
non-stream classifier throw `TypeError: Cannot read properties of null
(reading 'type')`, crashing the malformed-response detection path. Guard
before type-asserting each block: a null/non-object block is simply skipped.
Two regression tests added (null block among valid blocks → null; only-null
blocks → empty_choices).
Salvaged from closed PR #5096 (base-stale; only the defensive guard — the
Claude-shape recognition it also carried already landed via #5108).
Co-authored-by: herjarsa <herjarsa@users.noreply.github.com>
* feat(quota): persist observed provider quota reset windows
Adds `provider_quota_reset_events` (migration 108) + `db/quotaResetEvents.ts`
to record real upstream weekly-quota window transitions whenever a quota
refresh shows the reset rolling to a new cycle (different day, later resetAt).
`apiKeyUsageLimits` now prefers the observed window start over the inferred
`resetAt − 7d`, falling back to snapshot inference when no event is recorded
yet. `quotaCache.setQuotaCache` records the transition opportunistically.
`recordProviderQuotaResetEventIfChanged` only fires for the primary weekly
window (not daily/sonnet), is idempotent (INSERT OR IGNORE on the unique
window key), and no-ops when the reset didn't actually roll. 4 unit tests
(tests/unit/lib/quota-reset-events.test.ts).
Salvaged from closed PR #5025 (which bundled this with two unrelated
features + a colliding migration 104). Renumbered to 108; module re-exported
from localDb (Rule #2).
Co-authored-by: Witroch4 <175152067+Witroch4@users.noreply.github.com>
---------
Co-authored-by: herjarsa <herjarsa@users.noreply.github.com>
Co-authored-by: Witroch4 <175152067+Witroch4@users.noreply.github.com>
* docs(i18n): sync 3.8.38 CHANGELOG section to 41 mirrors (unblock docs-accuracy) (#5144)
The root CHANGELOG [3.8.38] section grew with this cycle's merged PRs, but the
docs/i18n/<lang>/CHANGELOG.md mirrors were not re-synced — drifting >25% in body
size and failing check:docs-sync (the "Docs accuracy" fast-gate step) for every
open PR against the release.
Ran scripts/release/sync-changelog-i18n.mjs 3.8.38 3.8.37 to copy the root
[3.8.38] section into all 41 mirrors. check:docs-all now passes (exit 0).
Sections are copied verbatim; the per-language translation pass runs at release
time via i18n:run — this only restores the size-sync the gate enforces.
* feat(compression): pure per-step fidelity checker (4 invariants, fail-open)
* feat(compression): fidelityGate config + rejected breakdown fields
* feat(compression): wire per-step fidelity gate into stacked pipeline (opt-in)
* feat(compression): preview route accepts fidelityGate flag (playground)
* feat(compression): playground fidelity-gate toggle + lane rejection display
* docs(compression): note fidelityGate advanced thresholds are intentionally API-omitted
* refactor(compression): extract fidelity-gate step helpers to shrink strategySelector (file-size gate)
bodyToText and gateAdvance moved to fidelityGateStep.ts; StackAccumulator exported.
strategySelector: 889->854 (-35). Residual +6 vs pre-Milestone-B frozen 848 is the
irreducible StackOptions.fidelityGate field + two stacked-loop dispatch reads + import.
Baseline updated to 854 with justification. No cycle introduced (import type only).
940 compression tests pass; typecheck clean.
* test(usage): wire usageHistoryDedup under unit runner brace-list (#5145)
Integrated into release/v3.8.38.
* feat: salvage batch from closed stale PRs (#5038, #5057, #5076) (#5138)
Integrated into release/v3.8.38.
* test(combo): deterministic routing-decision matrix for all 17 strategies (#5146)
Integrated into release/v3.8.38.
* feat(compression): fuzzy near-duplicate dedup (session-dedup 2nd pass + playground toggle) (#5143)
Integrated into release/v3.8.38.
* chore(quality): rebaseline file-size for sidebarVisibility.ts + chat.ts drift (#5147)
Mid-cycle drift on release/v3.8.38 from already-merged PRs that the fast-path
(PR->release skips check:file-size) let accumulate without a bump:
- src/shared/constants/sidebarVisibility.ts 1100->1198 (#3812 colored menu
icons, per-item accent map; #5142 dropped one orphan, net still above frozen)
- src/sse/handlers/chat.ts 1560->1575 (#5064 self-inflicted-timeout cooldown
skip + #5124 long OpenAI-compatible SSE hardening + #5110 embed-WS
LIVE_WS_HOST honour / early empty-message reject)
Each covered by its own PR tests; structural shrink of chat.ts tracked in #3501.
Unblocks the Fast Quality Gates for PRs targeting release/v3.8.38.
* chore(release): finalize v3.8.38 CHANGELOG + cycle reconciliation
- Reconcile [3.8.38]: +18 bullets (compression fidelity-gate/fuzzy-dedup #5143,
quota keepalive #5102, web-session robustness #5121, MiniMax/Nemotron #5136,
model-visibility #5091, failover logs #5016, disconnect races #5007, sidebar
orphan #5142, SRE playbooks salvage #5138, new Security #5130 + Maintenance roll-up)
- Credit salvaged-PR authors (@JxnLexn / @KooshaPari / @herjarsa / @Witroch4)
- Remove phantom bullet for CLOSED-not-merged #5092 (setup aggregator never landed)
- Fix isHidden bullet PR citation #4389 -> #5086 (@herjarsa)
- Back-fill forgotten v3.8.36 bullet: #5026 crypto.randomUUID ID-gen (@hamsa0x7)
- Sync 41 i18n CHANGELOG mirrors; README What's New -> v3.8.38
- Rebaseline cycle drift: eslint 3987->4002, cognitive 833->841, dead-exports
345->346, cyclomatic 1978->1980 (file-size handled by #5147)
* fix(i18n): add missing English UI labels (#5153)
Integrated into release/v3.8.38
* Preserve non-stream reasoning fields for compatible clients (#5155)
Integrated into release/v3.8.38
* feat(compression): ionizer engine — lossy JSON-array sampling reversible via CCR (#5148)
Integrated into release/v3.8.38
* test(combo): gated live smoke for combo strategies (in-process + VPS HTTP) (#5151)
Integrated into release/v3.8.38
* test: refresh release expectations to match current code (#5150)
Integrated into release/v3.8.38 (test-only base-red alignment extracted from #5150)
---------
Co-authored-by: Éder Costa <eder.almeida.costa@gmail.com>
Co-authored-by: José Victor Ferreira <root@josevictor.me>
Co-authored-by: Hernan Javier Ardila Sanchez <hjasgr@gmail.com>
Co-authored-by: fulorgnas <46461624+fulorgnas@users.noreply.github.com>
Co-authored-by: Randi <55005611+rdself@users.noreply.github.com>
Co-authored-by: Jan Leon <Jan.gaschler@gmail.com>
Co-authored-by: R. Beltran <rbeltran8000@gmail.com>
Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com>
Co-authored-by: KooshaPari <42529354+KooshaPari@users.noreply.github.com>
Co-authored-by: Ramel Tecnologia - Rafa Martins <146174365+rafacpti23@users.noreply.github.com>
Co-authored-by: herjarsa <herjarsa@users.noreply.github.com>
Co-authored-by: Witroch4 <175152067+Witroch4@users.noreply.github.com>
107 KiB
107 KiB
title, version, lastUpdated
| title | version | lastUpdated |
|---|---|---|
| Provider Reference | 3.8.31 | 2026-06-20 |
Provider Reference
Auto-generated from
src/shared/constants/providers.ts— do not edit by hand. Regenerate with:npm run gen:provider-referenceLast generated: 2026-06-20
Total providers: 231. See category breakdown below.
Categories
- Free — free tier with API key (configured via dashboard)
- OAuth — sign-in flow handled by OmniRoute, no API key needed
- Web cookie — wraps the provider's web app via cookie auth
- API key — paid provider configured via API key (free credits may apply)
- Local — runs on the user's machine (Ollama, LM Studio, vLLM, etc.)
- Search — web search providers
- Audio — audio-only providers (TTS/STT)
- Upstream proxy — providers that proxy to other providers
- Cloud agent — long-running coding agents (Codex Cloud, Devin, Jules)
- System — OmniRoute-internal providers (loopback, etc.)
Additional tags: image, video, aggregator, enterprise, embed/rerank, self-hosted.
Use the dashboard at /dashboard/providers to enable, configure, and test each provider.
OAuth Providers (19)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
agy |
agy |
Antigravity CLI | OAuth | link | Import your Antigravity CLI (agy) login (paste/upload its token file), auto-detect a local CLI login, or sign in with Google. Shares the Antigravity backend (incl. Claude models). |
amazon-q |
aq |
Amazon Q | OAuth | link | Uses the same AWS Builder ID or imported refresh-token flow as Kiro, but keeps Amazon Q connections separate. |
antigravity |
— | Antigravity | OAuth | — | — |
claude |
cc |
Claude Code | OAuth | — | — |
cline |
cl |
Cline | OAuth | — | — |
codex |
cx |
OpenAI Codex | OAuth | — | — |
cursor |
cu |
Cursor IDE | OAuth | — | — |
devin-cli |
dv |
Devin CLI (Official) | OAuth | link | Requires the Devin CLI binary. Run devin auth login to authenticate, or provide your WINDSURF_API_KEY. Install: https://cli.devin.ai |
gemini-cli |
gemini-cli |
Gemini CLI | OAuth | — | Uses Gemini CLI OAuth / Cloud Code credentials. Pro models require an eligible Google account or paid plan. |
github |
gh |
GitHub Copilot | OAuth | — | — |
gitlab-duo |
gitlab-duo |
GitLab Duo | OAuth | link | OAuth application with ai_features + read_user scopes. Configure GITLAB_DUO_OAUTH_CLIENT_ID and optionally GITLAB_DUO_OAUTH_CLIENT_SECRET on this OmniRoute instance. |
kilocode |
kc |
Kilo Code | OAuth | — | — |
kimi-coding |
kmc |
Kimi Coding | OAuth | — | — |
kiro |
kr |
Kiro AI | OAuth | — | Free tier: 50 credits/month (~25K–100K tokens). ⚠️ Kiro ToS prohibits third-party proxy/harness use. |
qoder |
if |
Qoder AI | OAuth | — | — |
qwen |
qw |
Qwen Code | OAuth | — | ⚠️ DEPRECATED. Qwen OAuth free tier was discontinued on 2026-04-15. Use 'bailian-coding-plan', 'alibaba', 'alibaba-cn', or 'openrouter' provider with API key instead. |
trae |
tr |
Trae | OAuth | link | Trae is an AI-native IDE by ByteDance (SOLO remote agent). Authorize via trae.ai in the popup, or sign in at solo.trae.ai and paste the Cloud-IDE-JWT (sent as 'Authorization: Cloud-IDE-JWT ', ~14-day lifetime) as the access token; web_id/biz_user_id/user_unique_id/scope/tenant/region propagate via providerSpecificData. No headless refresh for pasted tokens — re-paste on expiry. |
windsurf |
ws |
Windsurf (Devin CLI) | OAuth | link | In the Windsurf / VS Code IDE, open the command palette and run Windsurf: Provide Auth Token (or click the Jupyter "Get Windsurf Authentication Token" button), then copy the shown token and paste it here. Note: opening windsurf.com/show-auth-token directly only renders a "Redirecting" page — the IDE must initiate the flow (it adds a ?state=... param) for the token to appear. |
zed |
zd |
Zed IDE | OAuth | link | Zed stores LLM provider credentials (OpenAI, Anthropic, Google, Mistral, xAI) in the OS keychain. Use the Import button below to discover and import them automatically. |
Web Cookie Providers (22)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
adapta-web |
adp-web |
Adapta.org (Adapta One Web) | Web cookie | link | Paste your __client cookie value from .clerk.agent.adapta.one (DevTools → Application → Cookies) |
blackbox-web |
bb-web |
Blackbox Web (Subscription) | Web cookie | link | Paste your __Secure-authjs.session-token value or full cookie header from app.blackbox.ai |
chatgpt-web |
cgpt-web |
ChatGPT Web (Plus/Pro) | Web cookie | link | Paste your __Secure-next-auth.session-token cookie value from chatgpt.com |
claude-web |
cw |
Claude Web | Web cookie | link | Paste your session cookie from claude.ai |
copilot-web |
copilot |
Microsoft Copilot Web | Web cookie | link | Paste your access_token from copilot.microsoft.com (or export a .har file from DevTools while logged in) |
deepseek-web |
ds-web |
DeepSeek Web | Web cookie | link | Paste your userToken from chat.deepseek.com — DevTools → Application → Local Storage → userToken |
doubao-web |
db |
Doubao Web (ByteDance) | Web cookie | link | Paste your session cookie from doubao.com (DevTools → Application → Cookies) |
gemini-business |
gembiz |
Gemini Business (Enterprise) | Web cookie | link | From your enterprise account: open business.gemini.google/home/cid/{your-cid}, then copy **Secure-1PSID and **Secure-1PSIDTS cookies from DevTools → Application → Cookies. Paste as a cookie header below. |
gemini-web |
gweb |
Gemini Web (Free) | Web cookie | link | Paste your **Secure-1PSID cookie value from gemini.google.com. Optionally add **Secure-1PSIDTS separated by semicolon. |
grok-web |
gw |
Grok Web (Subscription) | Web cookie | link | Paste the full grok.com cookie line from DevTools → Application → Cookies. Include both sso and sso-rw (e.g. sso=...; sso-rw=...) — Grok's anti-bot rejects sso on its own. |
huggingchat |
huggingchat |
HuggingChat (Free) | Web cookie | link | Paste your hf-chat cookie value from huggingface.co/chat (DevTools → Application → Cookies → hf-chat). Optional — works without auth for basic use. |
inner-ai |
in-ai |
Inner.ai (Subscription) | Web cookie | link | Paste your token cookie and email separated by a space: open DevTools → Application → Cookies → .innerai.com, copy the token value, then append a space and your Inner.ai login email. Example: eyJhbG... user@example.com |
kimi-web |
kimi-web |
Kimi Web (Moonshot AI) | Web cookie | link | Paste your session cookie from kimi.moonshot.cn (DevTools → Application → Cookies) |
lmarena |
lma |
LMArena (Free) | Web cookie | link | Paste the full Cookie header from lmarena.ai (DevTools → Network → request → Cookie). The session is now split across arena-auth-prod-v1.0, .1, … — copy the whole header. Optional — works with free tier for basic comparisons. |
muse-spark-web |
ms-web |
Muse Spark Web (Meta AI) | Web cookie | link | Paste your abra_sess value or full cookie header from meta.ai |
perplexity-web |
pplx-web |
Perplexity Web (Pro/Max) | Web cookie | link | Paste your __Secure-next-auth.session-token cookie value from perplexity.ai |
phind |
ph |
Phind (Free) | Web cookie | link | ⚠️ DEPRECATED. Phind shut down its API (2026-01); the /api/chat endpoint no longer serves (sweep 2026-06-19). |
poe-web |
poe |
Poe Web (Subscription) | Web cookie | link | Paste your p-b cookie value from poe.com (DevTools → Application → Cookies → p-b) |
qwen-web |
qwen-web |
Qwen Web (Free) | Web cookie | link | Open chat.qwen.ai, log in, then open DevTools → Application → Local Storage → copy the "token" value (or use tongyi_sso_ticket cookie as Bearer token). |
t3-web |
t3chat |
t3.chat (Pro/Free) | Web cookie | link | Open t3.chat in your browser, log in, then open DevTools → Application → Local Storage → https://t3.chat. Copy the value of 'convex-session-id'. Also open DevTools → Network, copy the Cookie header from any request. Paste both values here. See provider setup docs for a step-by-step guide. |
v0-vercel-web |
v0 |
v0 Vercel Web (Code Gen) | Web cookie | link | Paste your session cookie from v0.dev (DevTools → Application → Cookies) |
venice-web |
ven |
Venice Web (Privacy) | Web cookie | link | Paste your session cookie from venice.ai (DevTools → Application → Cookies) |
API Key Providers (paid / paid-with-free-credits) (157)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
360ai |
360ai |
360 AI | API key | link | Get API key at ai.360.cn |
agentrouter |
agentrouter |
AgentRouter | API key, aggregator | link | $200 free credits on signup - multi-model routing gateway |
ai21 |
ai21 |
AI21 Labs | API key | link | $10 trial credits on signup (valid 3 months), no credit card required |
aimlapi |
aiml |
AI/ML API | API key, aggregator | link | Free tier paused (2026) — AI/ML API is now pay-as-you-go only (min $20 top-up); no recurring free credits. |
alibaba |
ali |
Alibaba | API key | link | — |
alibaba-cn |
ali-cn |
Alibaba (China) | API key | link | — |
anthropic |
anthropic |
Anthropic | API key | link | — |
api-airforce |
af |
Api.airforce | API key | link | 55 free tier models including Grok-3, Claude 3.7, Qwen3, Kimi-K2, Gemini 2.5 Flash, DeepSeek-V3 |
arcee-ai |
arcee |
Arcee AI | API key | link | Get API key at arcee.ai |
azure-ai |
azure-ai |
Azure AI Foundry | API key, enterprise | link | Use your Azure AI Foundry key. Base URL can be https://.services.ai.azure.com/openai/v1/ or https://.openai.azure.com/openai/v1/. |
azure-openai |
azure |
Azure OpenAI | API key, enterprise | link | Use your Azure OpenAI API key. Base URL should be your resource endpoint, for example https://my-resource.openai.azure.com. |
baichuan |
baichuan |
Baichuan | API key | link | Get API key at platform.baichuan-ai.com |
baidu |
baidu |
Baidu (ERNIE) | API key | link | Get API key at console.bce.baidu.com |
bailian-coding-plan |
bcp |
Alibaba Coding Plan | API key | link | — |
baseten |
baseten |
Baseten | API key | link | $30 free trial credits for GPU inference |
bazaarlink |
bzl |
BazaarLink | API key | link | Free tier with auto:free routing — zero-cost inference, no credit card required |
bedrock |
bedrock |
Amazon Bedrock | API key, enterprise | link | Use your Amazon Bedrock API key and configure the AWS region where your models are enabled (for example eu-west-2). OmniRoute calls Bedrock's native Converse API directly. |
black-forest-labs |
bfl |
Black Forest Labs | API key, image | link | — |
blackbox |
bb |
Blackbox AI | API key | link | Free tier: unlimited basic chat plus Minimax-M2.5, no credit card required |
bluesminds |
bm |
BluesMinds | API key | link | Free daily pi credits — supports 200+ models including GPT-4o, GPT-4.1, Claude Sonnet 4.5, Gemini 2.0 Flash, DeepSeek V4, Qwen, Kimi K2 |
byteplus |
bpm |
BytePlus ModelArk | API key | link | — |
bytez |
bytez |
Bytez | API key | link | $1 free credits, refreshes every 4 weeks |
cablyai |
cablyai |
CablyAI | API key, aggregator | link | Bearer API key for the CablyAI OpenAI-compatible gateway. |
cerebras |
cerebras |
Cerebras | API key | link | Free Trial: 1M tokens/day, 30K TPM, 5 RPM — no credit card. |
chutes |
chutes |
Chutes.ai | API key, aggregator | link | Bearer API key for the Chutes OpenAI-compatible gateway. |
clarifai |
clarifai |
Clarifai | API key, enterprise | link | Use your Clarifai PAT or app-specific API key. OmniRoute targets the OpenAI-compatible endpoint at https://api.clarifai.com/v2/ext/openai/v1 and authenticates with Authorization: Key . |
cloudflare-ai |
cf |
Cloudflare Workers AI | API key | link | Requires API Token AND Account ID (found at dash.cloudflare.com) |
codestral |
codestral |
Codestral | API key | link | — |
cohere |
cohere |
Cohere | API key | link | Free Trial: 1,000 API calls/month for testing, no credit card required |
command-code |
cmd |
Command Code | API key | link | Use a Command Code API key. Requests are sent to Command Code's /alpha/generate endpoint. |
coze |
coze |
Coze | API key | link | Get API key at coze.com/open/api |
crof |
crof |
CrofAI | API key | link | — |
databricks |
databricks |
Databricks | API key, enterprise | link | — |
datarobot |
datarobot |
DataRobot | API key, enterprise | link | Use your DataRobot API token. Optional Base URL can be the account root (for LLM Gateway) or a deployment URL under /api/v2/deployments/. |
deepinfra |
deepinfra |
DeepInfra | API key | link | Free signup credits for API testing and model exploration |
deepseek |
ds |
DeepSeek | API key | link | 5M free tokens on signup - no credit card required |
dify |
dify |
Dify | API key | link | Get API key from your Dify instance. |
dit |
dai |
DIT.ai | API key | link | Use your dit.ai API key in Authorization: Bearer . Fully OpenAI-compatible — a drop-in replacement, just change the base URL to https://api.dit.ai/v1. |
doubao |
doubao |
Doubao | API key | link | Get API key at console.volcengine.com |
empower |
empower |
Empower | API key, aggregator | link | Bearer API key for the Empower OpenAI-compatible endpoint. |
fal-ai |
fal |
Fal.ai | API key, image | link | — |
factory |
factory |
Factory | API key, aggregator | link | Bearer API key for the Factory (Factory Droids) OpenAI-compatible gateway. Same backend the droid CLI uses; subscription tier via app.factory.ai. OAuth follow-up tracked in issue #5060. |
featherless-ai |
featherless |
Featherless AI | API key | link | Free tier available — no credit card required |
fenayai |
fenayai |
FenayAI | API key, aggregator | link | Bearer API key for the FenayAI OpenAI-compatible gateway. |
firecrawl |
fc |
Firecrawl | API key | link | — |
fireworks |
fireworks |
Fireworks AI | API key | link | $1 free starter credits on signup for API testing |
freeaiapikey |
faik |
FreeAIAPIKey | API key | link | — |
freemodel-dev |
fmd |
FreeModel.dev | API key | link | $300 free credits on signup — no credit card required. Access GPT-5.4 and GPT-5.5 (OpenAI's latest flagship models) through an OpenAI-compatible API. |
friendliai |
friendli |
FriendliAI | API key | link | Free tier for serverless inference — no credit card required |
galadriel |
galadriel |
Galadriel | API key | link | ⚠️ DEPRECATED. api.galadriel.ai no longer resolves (sweep 2026-06-19); the inference API appears discontinued. |
gemini |
gemini |
Gemini (Google AI Studio) | API key | link | Free forever: 1,500 req/day for Gemini 2.5 Flash — no credit card, get key at aistudio.google.com |
getgoapi |
ggo |
GoAPI | API key, aggregator | link | — |
gigachat |
gigachat |
GigaChat (Sber) | API key | link | — |
github-models |
ghm |
GitHub Models | API key | link | Create a GitHub PAT with 'models: read' scope at github.com/settings/tokens |
gitlab |
gitlab |
GitLab Duo PAT | API key | link | GitLab personal access token for the public Code Suggestions API. Configure a self-hosted base URL when not using gitlab.com. |
gitlawb |
glb |
Gitlawb Opengateway (MiMo) | API key | link | Free MiMo (xiaomi/mimo-v2.5) revoked 2026-05 — Opengateway is now a pay-as-you-go credit gateway; no recurring free model. |
gitlawb-gmi |
glb-gmi |
Gitlawb Opengateway (GMI Cloud) | API key | link | Free Nemotron promo ended 2026-06 — the GMI Cloud route is now pay-as-you-go credit only. |
glhf |
glhf |
GLHF Chat | API key, aggregator | link | ⚠️ DEPRECATED. glhf.chat shut down (2026); its api.laf.run gateway no longer serves the catalog (sweep 2026-06-19). |
glm |
glm |
GLM Coding | API key | link | — |
glm-cn |
glmcn |
GLM Coding (China) | API key | link | — |
glmt |
glmt |
GLM Thinking | API key | link | — |
groq |
groq |
Groq | API key | link | Free tier: 30 RPM / 14.4K RPD — no credit card |
hackclub |
hc |
Hackclub AI | API key, aggregator | link | Sign in with your Hack Club account at ai.hackclub.com. |
haiper |
hp |
Haiper | API key, video | link | Get API key at haiper.ai/haiper-api |
heroku |
heroku |
Heroku AI | API key, enterprise | link | — |
huggingchat |
huggingchat |
HuggingChat | API key | link | No API key required for basic access. |
huggingface |
hf |
HuggingFace | API key | link | Free Inference API for thousands of models (Whisper, VITS, SDXL…) |
hyperbolic |
hyp |
Hyperbolic | API key | link | $1-5 trial credits on signup for serverless inference |
ideogram |
ideo |
Ideogram | API key | link | Get API key at ideogram.ai/docs/api |
iflytek |
iflytek |
iFlytek Spark | API key | link | Get API key at console.xfyun.cn |
inclusionai |
inclusion |
InclusionAI | API key | link | ⚠️ DEPRECATED. api.inclusionai.tech no longer resolves (sweep 2026-06-19); the inference API appears discontinued. |
inference-net |
inet |
Inference.net | API key | link | $25 free credits on signup plus research grants available |
jina-ai |
jina |
Jina AI | API key, embed/rerank | link | Bearer API key for the Jina AI rerank API. |
jina-reader |
jr |
Jina Reader | API key | link | — |
kie |
kie |
KIE.AI | API key | link | — |
kilo-gateway |
kg |
Kilo Gateway | API key, aggregator | link | — |
kimi |
kimi |
Kimi | API key | link | — |
kimi-coding-apikey |
kmca |
Kimi Coding (API Key) | API key | link | — |
kluster |
kluster |
Kluster AI | API key | link | ⚠️ DEPRECATED. kluster.ai shut down (2026-06-09); api.kluster.ai no longer resolves (sweep 2026-06-19). Use another OpenAI-compatible provider. |
lambda-ai |
lambda |
Lambda AI | API key | link | — |
laozhang |
lz |
LaoZhang AI | API key, aggregator | link | — |
leonardo |
leo |
Leonardo AI | API key, video | link | Get API key at leonardo.ai/developer |
liquid |
liquid |
Liquid AI | API key | link | Get API key at liquid.ai |
llamagate |
llamagate |
LlamaGate | API key | link | — |
llm7 |
llm7 |
LLM7.io | API key | link | No signup required - 2 req/s, 20 RPM, 100 req/hr free tier |
longcat |
lc |
LongCat AI | API key | link | Free: 5M tokens/day on LongCat-2.0-Preview (Flash models retired 2026-05-29); up to 120M/day via feedback. |
maritalk |
maritalk |
Maritalk | API key | link | — |
meta-llama |
meta |
Meta Llama API | API key | link | — |
minimax |
minimax |
Minimax Coding | API key, video | link | — |
minimax-cn |
minimax-cn |
Minimax (China) | API key | link | — |
mistral |
mistral |
Mistral | API key | link | Free Experiment tier: rate-limited access to all models, no credit card required |
modal |
mdl |
Modal | API key, enterprise | link | Use the bearer token that protects your Modal deployment, if enabled. Base URL should point to your OpenAI-compatible Modal app, for example https://--.modal.run/v1. |
monsterapi |
monster |
MonsterAPI | API key | link | Get API key at monsterapi.ai |
moonshot |
moonshot |
Moonshot AI | API key | link | — |
morph |
morph |
Morph | API key | link | Free tier: 250K credits/month, $0 |
nanogpt |
nanogpt |
NanoGPT | API key | link | — |
nebius |
nebius |
Nebius AI | API key | link | ~$1 trial credits on signup for API testing |
nlpcloud |
nlpc |
NLP Cloud | API key | link | Use your NLP Cloud API key in Authorization: Token . OmniRoute targets the chatbot endpoint on https://api.nlpcloud.io/v1/gpu//chatbot by default. |
nomic |
nomic |
Nomic | API key | link | Get API key at atlas.nomic.ai |
nous-research |
nous |
Nous Research | API key | link | Use your Nous Portal API key. OmniRoute targets the official OpenAI-compatible inference endpoint at https://inference-api.nousresearch.com/v1. |
novita |
novita |
Novita AI | API key, aggregator | link | $0.50 trial credits on signup (valid about 1 year) |
nscale |
nscale |
nScale | API key | link | $5 free credits on signup for inference testing |
nvidia |
nvidia |
NVIDIA NIM | API key | link | Free dev access: ~40 RPM, 70+ models (Kimi K2.5, GLM 4.7, DeepSeek V3.2...) |
oci |
oci |
OCI Generative AI | API key, enterprise | link | Use your OCI Generative AI API key or IAM bearer token. Base URL can be https://inference.generativeai..oci.oraclecloud.com/openai/v1/. |
ollama-cloud |
ollamacloud |
Ollama Cloud | API key | link | — |
openadapter |
oad |
OpenAdapter | API key | link | Use your OpenAdapter API key in Authorization: Bearer sk-cv-. Fully OpenAI-compatible. API base URL: https://api.openadapter.in/v1. |
openai |
openai |
OpenAI | API key | link | — |
opencode-go |
opencode-go |
OpenCode Go | API key | link | — |
opencode-zen |
opencode-zen |
OpenCode Zen | API key | link | — |
openrouter |
openrouter |
OpenRouter | API key, aggregator | link | Free models at $0/token with :free suffix - 20 RPM / 200 RPD |
orcarouter |
orcarouter |
OrcaRouter | API key | link | — |
ovhcloud |
ovh |
OVHcloud AI | API key | link | — |
perplexity |
pplx |
Perplexity | API key | link | — |
phind |
phind |
Phind | API key | link | Get API key at phind.com |
piapi |
pi |
PiAPI | API key, aggregator | link | — |
poe |
poe |
Poe | API key, aggregator | link | Bearer API key for the Poe OpenAI-compatible API. |
pollinations |
pol |
Pollinations AI | API key, video | link | Free keyless tier: openai, openai-fast, openai-large, qwen-coder, mistral, deepseek, grok, gemini-flash-lite-3.1, perplexity-fast, perplexity-reasoning. Premium models (claude, gemini, midijourney) require a Pollinations API key from enter.pollinations.ai. |
predibase |
predibase |
Predibase | API key | link | ⚠️ DEPRECATED. serving.app.predibase.com no longer resolves (sweep 2026-06-19); the managed serving API appears discontinued. |
publicai |
publicai |
PublicAI | API key | link | Requires an API key — one-time signup credit, then paid |
puter |
pu |
Puter AI | API key | link | Get token at puter.com/dashboard → Copy Auth Token |
qianfan |
qianfan |
Baidu Qianfan | API key | link | — |
recraft |
recraft |
Recraft | API key, image | link | — |
reka |
reka |
Reka | API key | link | Use your Reka API key. OmniRoute supports the OpenAI-compatible base URL https://api.reka.ai/v1 and sends both Authorization and X-Api-Key headers for compatibility. |
runwayml |
runway |
Runway | API key, video | link | Use your Runway API key in Authorization: Bearer . OmniRoute targets the current Runway API at https://api.dev.runwayml.com/v1 and sends the required X-Runway-Version header automatically. |
sambanova |
samba |
SambaNova | API key | link | $5 free credits on signup (30-day validity), no credit card required |
sap |
sap |
SAP Generative AI Hub | API key, enterprise | link | Use your SAP AI Core bearer token. Base URL can be your AI_API_URL root or a deploymentUrl from Generative AI Hub. |
scaleway |
scw |
Scaleway AI | API key | link | 1M free tokens for new accounts — EU/GDPR compliant (Paris), Qwen3 235B & Llama 70B |
sensenova |
sensenova |
SenseNova | API key | link | Get API key at platform.sensenova.cn |
siliconflow |
siliconflow |
SiliconFlow | API key | link | $1 free credits plus permanently free models after identity verification |
snowflake |
snowflake |
Snowflake Cortex | API key, enterprise | link | — |
sparkdesk |
sparkdesk |
SparkDesk | API key | link | Get API key at console.xfyun.cn |
stability-ai |
stability |
Stability AI | API key, image | link | — |
stepfun |
stepfun |
StepFun | API key | link | Get API key at platform.stepfun.com |
suno |
suno |
Suno | API key | link | Paste session cookie from suno.ai (Clerk auth) |
synthetic |
synthetic |
Synthetic | API key, aggregator | link | — |
tencent |
tencent |
Tencent Hunyuan | API key | link | Get API key at console.cloud.tencent.com |
thebai |
thebai |
TheB.AI | API key, aggregator | link | Bearer API key for the TheB.AI OpenAI-compatible gateway. |
together |
together |
Together AI | API key, video | link | $25 signup credits + 3 permanently free models: Llama 3.3 70B, Vision, DeepSeek-R1 distill |
tokenrouter |
trk |
TokenRouter | API key | link | Use your TokenRouter API key in Authorization: Bearer . Fully OpenAI-compatible. API base URL: https://api.tokenrouter.com/v1. |
topaz |
topaz |
Topaz | API key, image | link | — |
udio |
udio |
Udio | API key | link | Paste session cookie from udio.com (Supabase auth) |
uncloseai |
unc |
UncloseAI | API key | link | No auth required. API accepts any non-empty string as key for identification. |
upstage |
upstage |
Upstage | API key | link | — |
v0-vercel |
v0 |
v0 (Vercel) | API key | link | — |
venice |
venice |
Venice.ai | API key | link | — |
vercel-ai-gateway |
vag |
Vercel AI Gateway | API key, aggregator | link | — |
vertex |
vertex |
Vertex AI | API key, enterprise | link | Provide Service Account JSON or OAuth access_token |
vertex-partner |
vp |
Vertex AI Partners | API key, enterprise | link | Provide the same Service Account JSON used for Vertex AI partner models. |
volcengine |
volcengine |
Volcengine | API key | link | — |
voyage-ai |
voyage |
Voyage AI | API key, embed/rerank | link | Bearer API key for Voyage AI embeddings and rerank APIs. |
wafer |
wafer |
Wafer AI | API key | link | — |
wandb |
wandb |
Weights & Biases Inference | API key | link | — |
watsonx |
watsonx |
IBM watsonx.ai Gateway | API key, enterprise | link | Use your watsonx bearer token. Base URL can be https://.ml.cloud.ibm.com/ml/gateway/v1/ or a self-managed /ml/gateway/v1 endpoint. |
xai |
xai |
xAI (Grok) | API key | link | — |
xiaomi-mimo |
mimo |
Xiaomi MiMo | API key | link | — |
yi |
yi |
Yi (01.AI) | API key | link | Get API key at platform.lingyiwanwu.com |
zai |
zai |
Z.AI | API key | link | — |
zenmux |
zm |
ZenMux | API key | link | Use your ZenMux API key in Authorization: Bearer . ZenMux is fully OpenAI-compatible. Base URL: https://zenmux.ai/api/v1. |
Local Providers (11)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
comfyui |
comfyui |
ComfyUI | Local | link | No API key required. Configure the local ComfyUI base URL (default: http://localhost:8188). |
docker-model-runner |
dmr |
Docker Model Runner | Local, self-hosted | link | API key optional. Configure the local Docker Model Runner OpenAI-compatible base URL (default: http://localhost:12434/v1). |
lemonade |
lemonade |
Lemonade Server | Local, self-hosted | link | API key optional. Configure the local Lemonade OpenAI-compatible base URL (default: http://localhost:13305/api/v1). |
llama-cpp |
llamacpp |
llama.cpp | Local, self-hosted | link | API key optional (use any value, e.g. sk-no-key-required). Configure the llama-server OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1). Note: if Llamafile is also installed, both default to port 8080 — run only one at a time or override the port. |
llamafile |
llamafile |
Llamafile | Local, self-hosted | link | API key optional. Configure the local Llamafile OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1). |
lm-studio |
lmstudio |
LM Studio | Local, self-hosted | link | API key optional. Configure the local LM Studio OpenAI-compatible base URL (default: http://localhost:1234/v1). |
oobabooga |
ooba |
oobabooga | Local, self-hosted | link | API key optional. Configure the local oobabooga OpenAI-compatible base URL (default: http://localhost:5000/v1). |
sdwebui |
sdwebui |
SD WebUI | Local | link | No API key required. Configure the local WebUI base URL (default: http://localhost:7860). |
triton |
triton |
NVIDIA Triton | Local, self-hosted | link | API key optional. Configure the Triton OpenAI-compatible base URL (default: http://localhost:8000/v1). |
vllm |
vllm |
vLLM | Local, self-hosted | link | API key optional. Configure the local vLLM OpenAI-compatible base URL (default: http://localhost:8000/v1). |
xinference |
xinference |
XInference | Local, self-hosted | link | API key optional. Configure the local XInference OpenAI-compatible base URL (default: http://localhost:9997/v1). |
Search Providers (11)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
brave-search |
brave-search |
Brave Search | Search | link | Subscription token from Brave Search API dashboard |
exa-search |
exa-search |
Exa Search | Search | link | API key from dashboard.exa.ai |
google-pse-search |
google-pse |
Google Programmable Search | Search | link | Requires a Google API key and your Programmable Search Engine ID (cx) |
linkup-search |
linkup |
Linkup Search | Search | link | Bearer API key from the Linkup dashboard |
ollama-search |
ollama-search |
Ollama Search | Search | link | Same API key as Ollama Cloud (from ollama.com/settings/api-keys) |
perplexity-search |
pplx-search |
Perplexity Search | Search | link | Same API key as Perplexity (pplx-...) |
searchapi-search |
searchapi |
SearchAPI | Search | link | API key from SearchAPI (query param or Bearer auth) |
searxng-search |
searxng |
SearXNG Search | Search | link | API key is optional. Set your SearXNG base URL. Some instances may require a bearer token for access. |
serper-search |
serper-search |
Serper Search | Search | link | API key from serper.dev dashboard |
tavily-search |
tavily-search |
Tavily Search | Search | link | API key from app.tavily.com (format: tvly-...) |
youcom-search |
youcom-search |
You.com Search | Search | link | X-API-Key from the You.com platform dashboard |
Audio-only Providers (7)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
assemblyai |
aai |
AssemblyAI | Audio | link | — |
aws-polly |
polly |
AWS Polly | Audio | link | Use AWS Secret Access Key as API key; set providerSpecificData.accessKeyId and optional region. |
cartesia |
cartesia |
Cartesia | Audio | link | — |
deepgram |
dg |
Deepgram | Audio | link | — |
elevenlabs |
el |
ElevenLabs | Audio | link | — |
inworld |
inworld |
Inworld | Audio | link | — |
playht |
playht |
PlayHT | Audio | link | — |
Upstream Proxy Providers (2)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
9router |
nr |
9router | Upstream proxy | link | — |
cliproxyapi |
cpa |
CLIProxyAPI | Upstream proxy | link | — |
Cloud Agent Providers (3)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
codex-cloud |
codex-cloud |
Codex Cloud | Cloud agent | link | OpenAI API key with Codex Cloud task access. |
devin |
devin |
Devin | Cloud agent | link | Devin API key for cloud agent sessions. |
jules |
jules |
Google Jules | Cloud agent | link | Jules API key for creating and managing cloud coding tasks. |
System Providers (1)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
auto |
auto |
Auto (Zero-Config) | System | — | — |
Sources of truth
- Catalog:
src/shared/constants/providers.ts - Registry (per-model details):
open-sse/config/providerRegistry.ts - Executors:
open-sse/executors/(31 files) - Translators:
open-sse/translator/
See Also
- FREE_TIERS.md — curated free-tier guide
- USER_GUIDE.md — provider setup walkthrough
- ARCHITECTURE.md — overall architecture