mirror of
https://github.com/diegosouzapw/OmniRoute.git
synced 2026-08-13 18:52:18 +03:00
* fix(ci): clear base-reds on release/v3.8.50 (round 3) - CHANGELOG.md: restore the top [Unreleased] section dropped by the #10189 reconcile (docs-sync gate: first section must be Unreleased) - env-doc-sync: document CONDUCTOR_ORCHESTRATOR_TOKEN + CONDUCTOR_SPOKESPERSON_URL in .env.example/ENVIRONMENT.md; allowlist the CI-only GITHUB_STEP_SUMMARY and TS7_BASE_REF (ts7 ratchet signals); drop a stray merge artifact line - providers: restore the audited chatanywhere metadata entry that base-reds round 2 dropped together with its duplicate — the provider was half-wired (registry+endpoint without APIKEY metadata), which is what the wave3 test catches; re-pin providers-constants-split at the measured 228 - docs counts: 338 -> 339 (today's +2 void-ai/helixmind, -1 Puter) via gen:provider-reference + README/AGENTS/llm.txt/package.json/diagrams/i18n mirrors - file-size ratchet: annotated rebaseline for the two pre-existing drifts (ModelSelectModal 1138, gateways 1250) following the 2026-08-11 precedent Refs #9985 * fix(ci): base-reds round 3b — stale sibling tests + mode-pack weight contract - check-docs-counts-sync.test.ts: drop the imports/subtests of the four helpers #10196 removed from the gate script (readMcpFactsFromSource, listLocalizedDocs, makeRequiredCountsValidator, checkFreeTierInventory) — the new-API tests that #10196 added stay; the file now loads again under the node runner - quota-connection-recovery.test.ts: convert from vitest APIs to node:test — the file lives in tests/unit/*.test.ts (node-runner glob) and the vitest runtime crashes when imported outside vitest, killing the whole shard entry - modePacks.ts: re-normalize all six mode packs to sum 1.0 — #8940 added sessionAvailability: 0.05 to every pack without rebalancing (1.05 total); ratios preserved exactly (÷1.05), so post-normalizeScoringWeights behavior is unchanged; restores the declared sum-to-1.0 contract the 4235 test pins Refs #9985 * fix(ci): base-reds round 3c — vitest siblings, weights default, secrets FP, mutation tap - DistributeProxiesButton.test.tsx: wrap renders in NextIntlClientProvider — #9245 localized the component (useTranslations) and left the test without the intl context, failing all 14 cases - scoring.ts: re-normalize DEFAULT_WEIGHTS to sum 1.0 (same #8940 class as the mode packs — sessionAvailability added without rebalancing; ratios preserved) - .gitleaks.toml: generalize the kimi sponsor-banner localStorage-key allowlist to -v\d+ — #10200 bumped v1→v2 and the stale regex regressed the secrets ratchet with a false positive - stryker.conf.json: register 6 covering unit tests in tap.testFiles (4 modules) so their mutant kills count — unblocks check:mutation-test-coverage --strict Refs #9985 * fix(ci): base-reds round 3d — inspector factor gap, stale registry/gap tests, i18n key sync - comboScoringInspector: add cacheAffinity/sessionAvailability/connectionDensity to FACTOR_KEYS + the factor-key type — calculateScore() weighs them but the breakdown omitted them, so the explained contributions never summed to the reported score (inspector bug, red on the pure tip) - combo-scoring-inspector.test: make the explicit-weights override sum-neutral (±0.05 shift) so it stays valid for any DEFAULT_WEIGHTS values — the hardcoded override only summed to 1.0 against the pre-#8940 defaults, which is also why explicit weights silently fell back to 'default' on the tip - unorouter-registry.test: align to the canonical .com host (api.unorouter.ai 301-redirects there, verified live) and to wave4's live model discovery (passthrough, no static seed) — the .ai/auto-model expectations were stale - check-migration-numbering.test: 147 left KNOWN_GAPS when 147_api_keys_model_access_mode.sql landed — assert absent (same as 143) - i18n: sync-ui pass — 35,914 missing UI keys stamped as __MISSING__ placeholders across 42 locales (mechanical; greens the pt-BR key-presence integrity test; coverage pct unchanged by design — translation is a separate workstream) Refs #9985 * fix(ci): base-reds round 3e — 2 real defects + 14 stale sibling tests (waves A-E) Real defects fixed: - src/lib/db/apiKeys.ts: #9313's empty-allowlist early return bypassed the group permission check, silently disabling group deny rules (#8817) for every key without a per-key allowlist; fall-through restored, restricted+[] deny-all kept - open-sse/utils/proxyFetch.ts: #10032 re-appended the raw transport error to the propagated message, reintroducing the proxy user:password leak #9837 closed; new redactProxyDetailsInMessage() keeps the reason, redacts URL/credentials - .github/workflows/quality.yml: #10134 added the TS7 ratchet as a separate blocking step AFTER the aggregated gates — the exact #8542 masking mechanism; folded into the non-fail-fast loop (still blocking, still PR-only) ⚠️ CI edit, gate-strengthening — explicit owner sign-off requested on the PR - src/i18n/messages/ko.json: 3 machine-mistranslation regressions caught by the #8244 glossary checker (장애인→비활성화됨, 양말5://→socks5://, 비클로드→Claude가 아닌) Stale sibling tests aligned to deliberately-moved contracts (each cites its mover): request-log-detail-layout + -stream (#9245 intl provider), repro-8542 pin update, quality-rail-gate-membership (#10134 shape), agentSkills-routes 45→46 (#9058), cloudflare-ai-catalog-8717 (#8804 supersedes #8808), executor-xai (#9994), vision-bridge-claude-wire (#9463 minimax→openai), sse-auth forced-pin (#8893), tls-proxy-context (strengthened leak guards), rate-limit-local-error-classification (#9164/#9342), minimax-thinking-signature (#9463), codebuddy-cn (#9723 +1 test), github-copilot-custom-model (#9050), providers-g4f-batch3 (#9584), synced-capability-warmup (#9199, stricter), sidebar-tools-group (#8221), oauth-modal-grok-cli-paste (#9245); agentSkills/catalog.ts comment 45→46; file-size rebaseline for proxyFetch (+19, annotated) Refs #9985 * fix(ci): base-reds round 3f — waves F-J: 9 more real defects + stale sibling sweep Real production defects fixed (all red on the pure tip, each with its origin): - routeGuard.ts: #8949 accidentally DELETED the /api/providers/[id]/login local-only pattern — the route spawns a browser, so the loopback gate for a process-spawning route was gone (Hard Rules #15/#17); restored (314 guard tests green) - agentSkills generator: #9058's category dispatch gave the config category an empty body, wiping skills/config-codex-cli/SKILL.md at the #10131 sync; fixed + SKILL.md regenerated via the official generator - imageRegistry: #9982 broke same-provider bare aliasing (antigravity preview id sent upstream unresolved); new resolveSameProviderBareAlias() keeps the fal cross-provider fix intact - imageRegistry: #9982's prefix strip handed the bare nano-banana ids to fal-ai, violating the pinned 2026-07-31 operator decision (adobe-firefly owns them); fal entries made prefix-only (dispatch already re-prefixes) - mediaGeneration/fal.ts: the missing-credential 401 guard was lost when #10198 deleted the superseded falHandler — tests were hitting the live network - bottleneckPatch/rateLimitManager: #9041's merge clobbered #9604, resurrecting the Bottleneck v2.19.5 heartbeat bug (reservoir never refills); patched the library defect at the root and re-aligned chat-rate-limit-body-lock to the working reservoir contract - processSupervisor.mjs: #9761 regressed the Node spawn to bare "node" (the #9156 launchd bug) and dropped #9209's ipv4first args; both restored - openai-responses/pureHelpers: #9423's Agent null-sentinel was unreachable on the schemaless JSON-string path; gate extended - i18n en.json: #8222's regen reverted the #9976 unclosed-tag fix and #8559's combo-cooldown copy; #9038 shipped 40 t() calls with no messages (runtime MISSING_MESSAGE); all restored/added + official sync-ui stamps, and vi's zero-marker policy re-established via the sanctioned translation backend Stale sibling tests aligned (movers cited inline): chat-helpers (#9447), executor-antigravity (#9351), video-fal-grok (#9982), visionBridge (#9759), web-session-credentials (#8974), production-build-module-integrity (positive anchor added), agentSkills-generator/skillManifestsLint/skills-injection/ agentSkillTools-mcp/listCapabilities-a2a (#9058), memory-settings (#10010), model-catalog-policy-invalidation (#8906), model-alias-seed (#9485), reactive-context-compaction (#8949), combo-provider-wildcard (broken upsert helper), oauth-google-loopback (43-locale resurrected-key removal) Validation: 501/501 across the 47 touched test files; typecheck:core, lint, file-size, docs-sync all green. Refs #9985 * fix(ci): base-reds round 3g — wave K/L: 4 more real defects + stale alignments Real defects: - base/reasoningEffort.ts: the stale duplicate cherry-pick #9612 re-added the codex minimal→low rewrite that #9883 had deliberately removed (OMP minimal passthrough); block removed again - cursorImages.ts: #9840 wired prepareCursorImageForWire (sharp re-encode, fail-closed) into the SHARED resolveCursorImages, breaking zai-web and conol-web image uploads (HTTP 400 'undecodable'); new prepareForWire opt-out, Cursor default path unchanged (8 cursor suites green) - modelCapabilities/snapshot: catalog prepare still issued 323 per-model reads of model_context_overrides + max_input_tokens overrides, violating #9199's bulk-load contract; both now resolve from the snapshot single pass - v1-models-discovery-conformance: re-pinned to the bounded 30s SWR window (#9199/#10198) — the old 'stale-first regardless of age' contract is gone Stale tests aligned (movers cited inline): codex-tools-strict-default (#9828 redundant-oneOf strip), devin-providers (#9245 i18n), db-migrationrunner- constants-split (147→151 renumber #8228), gitlab-duo-oauth-setup (#9245), chatcore-extracted-modules (#9161 outbound-protocol keying) compression-api CI failures were cascade artifacts of codex-tools-strict-default failing in the same force-exit shard process — no own defect (171/171 local). Refs #9985 * fix(test): compression-api — register both describes before the runner starts The DATA_DIR setup + route/db top-level awaits sat BETWEEN the two describes; under --test-force-exit (the CI unit-runner flag) the process exits once the already-registered tests finish, so on slow CI machines the whole second describe died as 'Promise resolution is still pending' — the recurring CI-only shard-2 failure that never reproduced locally without the flag. Moved to the top of the file; 10/10 under --test-force-exit locally. Refs #9985 * fix(quality): freeze modelCapabilities.ts at 1006 (annotated) — snapshot routing growth Refs #9985 * fix(quality): move the modelCapabilities freeze into the frozen map (nested schema) Refs #9985 * fix(i18n): translate all 39,718 pending UI keys across 42 locales (owner-approved) Mass-translated every __MISSING__ placeholder via the official i18n:sync-ui --translate-markers pipeline (operator backend), restoring i18nUiCoverage to the 100 baseline (was 89.9 after the merge-storm UI landings + the 42 keys #9038 never shipped). Post-pass repairs, all caught by the existing gates: - glossary: retired renderings the machine reintroduced normalized again (提供商→提供者 zh-CN/zh-TW, 鏈接→連結, 文檔→文件, 調用→呼叫, 供應商→提供者, 響應→回應, 不活躍→未啟用 zh-TW; 클로드→Claude, 옴니루트→OmniRoute ko); DATA_DIR forbidden rendering avoided via 数据文件夹 rephrase - ICU integrity: 120 values with renamed/dropped {params} repaired (39 positional renames, 81 reset to the en source — functional over fluent) Validation: glossary/pt-BR/vi/deno-relay/settings-keys/value-drift/google- loopback suites 76/76; placeholder diff en×42 locales = 0; worst-locale coverage = 100.0%. Refs #9985 --------- Co-authored-by: backryun <bakryun0718@proton.me>
62 KiB
62 KiB
title, version, lastUpdated
| title | version | lastUpdated |
|---|---|---|
| Provider Reference | 3.8.50 | 2026-08-12 |
Provider Reference
Auto-generated from
src/shared/constants/providers.ts— do not edit by hand. Regenerate with:npm run gen:provider-referenceLast generated: 2026-08-12
Total providers: 339. See category breakdown below.
Categories
- Free — free tier with API key (configured via dashboard)
- No-auth — public endpoints that require no key or sign-in at all
- OAuth — sign-in flow handled by OmniRoute, no API key needed
- Web cookie — wraps the provider's web app via cookie auth
- API key — paid provider configured via API key (free credits may apply)
- Local — runs on the user's machine (Ollama, LM Studio, vLLM, etc.)
- Search — web search providers
- Audio — audio-only providers (TTS/STT)
- Upstream proxy — providers that proxy to other providers
- Cloud agent — long-running coding agents (Codex Cloud, Devin, Jules)
- System — OmniRoute-internal providers (loopback, etc.)
Additional tags: image, video, aggregator, enterprise, embed/rerank, self-hosted.
Tool calling (where shown): native — real function-calling API; emulated — the tools array is prompt-emulated via webTools.ts (regex-parsed <tool>{...}</tool> blocks); none — tools is currently silently dropped. See #7286.
Use the dashboard at /dashboard/providers to enable, configure, and test each provider.
No-auth Providers (no key required) (10)
| ID | Alias | Name | Tags | Website | Notes | Tool calling |
|---|---|---|---|---|---|---|
aihorde |
horde |
AI Horde | No-auth | link | No API key required — uses AI Horde's documented anonymous key. Adding a free aihorde.net key is optional and only buys higher queue priority (kudos). | — |
auggie |
aug |
Augment (Auggie CLI) | No-auth | link | No API key stored by OmniRoute. Install the Auggie CLI and run auggie login on this machine, then OmniRoute spawns it locally for each request. |
— |
chipotle |
pepper |
Chipotle Pepper AI (Free) | No-auth | link | No credentials required. Uses Chipotle's public support chatbot via reverse-engineered SockJS/STOMP protocol. | — |
devin-cli-agentic |
dva |
Devin CLI Agentic Bridge | No-auth | link | Authentication is owned by the official Devin CLI in its isolated bridge volume. | emulated |
duckduckgo-web |
ddgw |
DuckDuckGo AI Chat | No-auth | link | No credentials required — DuckDuckGo AI Chat is anonymous and free. | emulated |
felo-web |
felo |
Felo | No-auth | link | No credentials required — Felo is a free, no-signup chat/search aggregator. | — |
mimocode |
mcode |
MiMoCode (Free) | No-auth | link | No API key required. The executor auto-generates JWT tokens via device fingerprint bootstrap. | — |
opencode |
oc |
OpenCode Free | No-auth | link | No API key required — uses OpenCode's public free endpoint. | — |
theoldllm |
tllm |
The Old LLM (Free) | No-auth | link | No credentials required. The executor auto-generates access tokens via an embedded Playwright browser instance. | — |
veoaifree-web |
veo-free |
Veo AI Free | No-auth, video | link | No auth required. Rate limited to 6 requests/hour per IP. | — |
OAuth Providers (25)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
agy |
agy |
Antigravity CLI | OAuth | link | Import your Antigravity CLI (agy) login (paste/upload its token file), auto-detect a local CLI login, or sign in with Google. Shares the Antigravity backend (incl. Claude models). |
amazon-q |
aq |
Amazon Q | OAuth | link | Uses the same AWS Builder ID or imported refresh-token flow as Kiro, but keeps Amazon Q connections separate. |
antigravity |
— | Antigravity | OAuth | — | — |
claude |
cc |
Claude Code | OAuth | — | — |
cline |
cl |
Cline | OAuth | — | — |
clinepass |
cp |
ClinePass | OAuth | link | ClinePass is Cline's $9.99/mo subscription bundling 10 open coding models. Sign in with your Cline account (same login as the Cline CLI/IDE), or paste a direct ClinePass API key (app.cline.bot → Settings → API Keys). A ClinePass subscription unlocks the cline-pass/* models. Reuses the Cline WorkOS OAuth flow. |
codebuddy-cn |
cbcn |
CodeBuddy CN | OAuth | link | Tencent CodeBuddy CN (copilot.tencent.com). Sign in via the official CLI device-code flow, or paste a direct API key (sent as Authorization: Bearer). Catalog: GLM / Kimi / MiniMax / DeepSeek / Hunyuan. |
codex |
cx |
OpenAI Codex | OAuth | — | — |
cursor |
cu |
Cursor IDE | OAuth | — | — |
devin-cli |
dv |
Devin CLI | OAuth | link | Requires the Devin CLI binary. Run devin auth login to authenticate, or provide your WINDSURF_API_KEY. Install: https://cli.devin.ai |
devin-desktop |
— | Devin Desktop | OAuth | link | Paste an existing Devin API key from an authenticated Devin session. Key export availability and steps vary by Devin version and account. |
ghe-copilot |
ghe-copilot |
GitHub Enterprise Copilot | OAuth | — | Enter your GHE instance URL (e.g., https://ghe.company.com) in provider settings, then authenticate via device flow. |
github |
gh |
GitHub Copilot | OAuth | — | — |
gitlab-duo |
gitlab-duo |
GitLab Duo | OAuth | link | GitLab Duo OAuth is not configured. Register an OAuth application at https://gitlab.com/-/profile/applications with redirect URI http://localhost:20128/callback and scopes "ai_features read_user", then set GITLAB_DUO_OAUTH_CLIENT_ID (and optionally GITLAB_DUO_OAUTH_CLIENT_SECRET) and restart. |
grok-cli |
gc |
Grok Build | OAuth | — | Sign in with your browser, or paste your ~/.grok/auth.json (or the JWT access token) from the Grok Build CLI; refresh_token is rotated automatically either way. |
kilocode |
kc |
Kilo Code | OAuth | — | — |
kimi-coding |
kmc |
Kimi Code CLI | OAuth | link | Sign in with the same Kimi account used by Kimi Code CLI. OmniRoute uses the CLI OAuth flow and Kimi Coding Plan endpoints. |
kiro |
kr |
Kiro AI | OAuth | — | Free tier: 50 credits/month (~25K–100K tokens). ⚠️ Kiro ToS prohibits third-party proxy/harness use. |
openference |
of |
Openference | OAuth | link | Sign in with your Openference account to route requests through api.openference.com. An active plan is required for inference — OAuth may authenticate but return 402 without one. |
qoder |
if |
Qoder | OAuth | — | — |
raycast |
rc |
Raycast Pro AI | OAuth | link | Unofficial integration — uses your Raycast Pro subscription via credentials from the macOS app (Auto-Import or manual capture). May break on Raycast updates. Not for redistribution; personal use only. |
trae |
tr |
Trae | OAuth | link | Trae is an AI-native IDE by ByteDance (SOLO remote agent). Authorize via trae.ai in the popup, or sign in at solo.trae.ai and paste the Cloud-IDE-JWT (sent as 'Authorization: Cloud-IDE-JWT ', ~14-day lifetime) as the access token; web_id/biz_user_id/user_unique_id/scope/tenant/region propagate via providerSpecificData. No headless refresh for pasted tokens — re-paste on expiry. |
xai-oauth |
xao |
xAI OAuth (Grok) | OAuth | link | Sign in with xAI to use api.x.ai models such as Grok 4.5. This is separate from Grok Build JWT sessions, which use cli-chat-proxy.grok.com and grok-build model aliases. |
zed |
zd |
Zed IDE | OAuth | link | Zed stores LLM provider credentials (OpenAI, Anthropic, Google, Mistral, xAI) in the OS keychain. Use the Import button below to discover and import them automatically. |
zed-hosted |
— | Zed Hosted Models | OAuth | link | Sign in with your Zed account (native-app sign-in). OmniRoute generates a one-time RSA keypair and opens zed.dev to authorize it — on a remote/headless install, copy the resulting 127.0.0.1 callback URL from your browser's address bar and paste it back here. Distinct from the 'Zed IDE' credential-import entry above: this proxies chat completions through Zed's own hosted model aggregator (cloud.zed.dev), fronting Anthropic/OpenAI/Google/xAI models under your Zed plan. |
Web Cookie Providers (34)
| ID | Alias | Name | Tags | Website | Notes | Tool calling |
|---|---|---|---|---|---|---|
adapta-web |
adp-web |
Adapta.org (Adapta One Web) | Web cookie | link | Paste your __client cookie value from .clerk.agent.adapta.one (DevTools → Application → Cookies) | emulated |
adobe-firefly |
firefly |
Adobe Firefly (Image/Video) | Web cookie | link | RECOMMENDED: firefly.adobe.com signed-in → F12 → Network → click firefly-3p.ff.adobe.io (generate-async or models/discovery) → Request Headers → Authorization → copy the token AFTER 'Bearer ' (starts with eyJ…). Cookie-only from firefly.adobe.com mints a GUEST token → 401/403; only multi-domain IMS cookies (adobelogin.com) or that Bearer JWT work. Unofficial/experimental media + Limits. | — |
blackbox-web |
bb-web |
Blackbox Web (Subscription) | Web cookie | link | Paste your __Secure-authjs.session-token value or full cookie header from app.blackbox.ai | emulated |
chatgpt-web |
cgpt-web |
ChatGPT Web (Plus/Pro) | Web cookie | link | Paste your __Secure-next-auth.session-token cookie value from chatgpt.com | emulated |
chatgpt-web-codex |
cgpt-codex |
ChatGPT Web (Codex) | Web cookie | link | Paste the full ChatGPT Cookie header. OmniRoute verifies it in an isolated headless browser profile. | native |
claude-web |
cw |
Claude Web | Web cookie | link | Paste your session cookie from claude.ai | none |
conol-web |
cnl |
Conol (Unofficial/Experimental) | Web cookie | link | Use browser sign-in, or paste the full Cookie header from conol.ai. The __Secure-better-auth.session_token cookie is required. | — |
copilot-m365-web |
m365copilot |
Microsoft 365 Copilot (BizChat) | Web cookie | link | Sign in at m365.cloud.microsoft/chat, then open DevTools → Network → filter 'WS' → click the Chathub WebSocket connection. Copy both the access_token query parameter AND the account-specific Chathub path segment from its request URL (wss://…/Chathub/?…&access_token=…). It is NOT an Authorization: Bearer header on an XHR/Fetch request. The token is short-lived; this is an unofficial integration. | — |
copilot-web |
copilot |
Microsoft Copilot Web | Web cookie | link | Paste the access_token from an authenticated copilot.microsoft.com request (DevTools → Network → Authorization), or export a HAR while logged in | — |
deepseek-web |
ds-web |
DeepSeek Web | Web cookie | link | Paste your userToken from chat.deepseek.com — DevTools → Application → Local Storage → userToken | emulated |
doubao-web |
db |
Dola Web (ByteDance) | Web cookie | link | Paste the full Cookie header from www.dola.com. It should include sessionid, ttwid, and s_v_web_id. If s_v_web_id is unavailable, fp=verify_... from a chat/completion request URL can be used as a fallback. | — |
gemini-business |
gembiz |
Gemini Business (Enterprise) | Web cookie | link | From your enterprise account: open business.gemini.google/home/cid/{your-cid}, then copy __Secure-1PSID and __Secure-1PSIDTS cookies from DevTools → Application → Cookies. Paste as a cookie header below. | — |
gemini-web |
gweb |
Gemini Web (Free) | Web cookie | link | Paste your __Secure-1PSID cookie value from gemini.google.com. Optionally add __Secure-1PSIDTS separated by semicolon. | emulated |
grok-web |
gw |
Grok Web (Subscription) | Web cookie | link | Paste the full grok.com cookie line from DevTools → Application → Cookies. Include both sso and sso-rw (e.g. sso=...; sso-rw=...) — Grok's anti-bot rejects sso on its own. |
— |
hailuo-web |
hailuo-web |
Hailuo Web (MiniMax) | Web cookie | link | Open hailuo.ai, log in, then open DevTools → Application → Local Storage → copy the "_token" value. device_id/uuid fingerprint fields are derived automatically; if requests fail, re-capture _token (sessions can expire). | — |
huggingchat |
huggingchat |
HuggingChat (Free) | Web cookie | link | Paste the full Cookie header from huggingface.co/chat (DevTools → Network → /chat/conversation → Request Headers → Cookie). It should include hf-chat and may also include token / aws-waf-token. | — |
hyperagent |
ha |
HyperAgent (Unofficial/Experimental) | Web cookie | link | Paste the full Cookie header from hyperagent.com (DevTools → Network → any request → Request Headers → Cookie). Session cookies power chat + billing usage. | — |
inner-ai |
in-ai |
Inner.ai (Subscription) | Web cookie | link | Paste your token cookie and email separated by a space: open DevTools → Application → Cookies → .innerai.com, copy the token value, then append a space and your Inner.ai login email. Example: eyJhbG... user@example.com | emulated |
kimi-web |
kimi-web |
Kimi Web | Web cookie | link | Paste access_token from www.kimi.com DevTools → Application → Local Storage. A legacy kimi-auth cookie is also accepted. | — |
lmarena |
lma |
Arena (Free) | Web cookie | link | Paste the full Cookie header from arena.ai (DevTools → Network → request → Cookie). Include arena-auth-prod-v1.0/.1… and cf_clearance/__cf_bm when present. OmniRoute uses Chrome TLS impersonation; if Arena still 403s, set providerSpecificData.recaptchaV3Token from a live browser session. | — |
microsoft-designer-web |
msdesigner |
Microsoft Designer (Image Generation) | Web cookie | link | Sign in at designer.microsoft.com, then open DevTools → Network, generate an image, and find the request to DallE.ashx?action=GetDallEImagesCogSci. Copy the value of its Authorization: Bearer header (the access_token — no 'Bearer ' prefix). The token is short-lived; this is an unofficial, reverse-engineered integration. | — |
muse-spark-web |
ms-web |
Muse Spark Web (Meta AI) | Web cookie | link | Paste your ecto_1_sess cookie AND the ecto1:... WS auth token from meta.ai. Capture the ecto1: token in DevTools → Network → WS → the clippy request's Authorization query param. Example: ecto_1_sess=4240a308...NVDg0; ecto1:ABCD... | emulated |
notion-web |
nw |
Notion AI Web (Unofficial/Experimental) | Web cookie | link | Paste only the token_v2 cookie VALUE from app.notion.com (DevTools → Application → Cookies → token_v2). Do not paste token_v2= or the full Cookie header. Workspace is auto-detected; space_id / notion_user_id are optional. | — |
perplexity-web |
pplx-web |
Perplexity Web (Pro/Max) | Web cookie | link | Paste your __Secure-next-auth.session-token cookie value from perplexity.ai | emulated |
poe-web |
poe |
Poe Web (Subscription) | Web cookie | link | Paste your p-b cookie value from poe.com (DevTools → Application → Cookies → p-b) | — |
promptql |
pql |
PromptQL (Unofficial/Experimental) | Web cookie | link | Paste the Bearer JWT from prompt.ql.app DevTools → Network → graphql → Authorization (token only). Optional projectId + session Cookie for refresh. | — |
qwen-web |
qwen-web |
Qwen Web (Free) | Web cookie | link | Open chat.qwen.ai, log in, then open DevTools → Application → Local Storage → copy the "token" value (or use tongyi_sso_ticket cookie as Bearer token). | emulated |
t3-web |
t3chat |
t3.chat (Pro/Free) | Web cookie | link | Open t3.chat in your browser, log in, then open DevTools → Application → Local Storage → https://t3.chat. Copy the value of 'convex-session-id'. Also open DevTools → Network, copy the Cookie header from any request. Paste both values here. See provider setup docs for a step-by-step guide. | emulated |
tinycms-web |
tcw |
TinyCMS Web (Free/Sub) | Web cookie | link | Go to site.tinycms.xyz, open DevTools → Application → Local Storage, copy the value of 'app-config-uuid' (starts with 'R'), and paste it here. | — |
v0-vercel-web |
v0-vercel-web |
v0 Vercel Web (Code Gen) | Web cookie | link | Paste your session cookie from v0.dev (DevTools → Application → Cookies) | — |
venice-web |
ven |
Venice Web (Privacy) | Web cookie | link | Paste your session cookie from venice.ai (DevTools → Application → Cookies) | — |
yuanbao-web |
ybw |
Tencent Yuanbao (Free) | Web cookie | link | Log in to yuanbao.tencent.com, then paste the full Cookie header (DevTools → Network → any /api request → Request Headers → Cookie). It must contain hy_user and hy_token. | — |
zai-web |
zw |
Z.ai Web | Web cookie | link | Copy the "token" value from chat.z.ai → DevTools → Application → Local Storage. Do not copy cookies; OmniRoute handles the per-request CAPTCHA through its browser transport. | — |
zenmux-free |
zmf |
ZenMux Free (Web) | Web cookie | link | Login at zenmux.ai, then export all cookies using EditThisCookie or Cookie-Editor and paste the full Cookie header string here. Refresh every ~30 days. | — |
API Key Providers (paid / paid-with-free-credits) (228)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
360ai |
360ai |
360 AI | API key | link | Get API key at ai.360.cn |
agentrouter |
agentrouter |
AgentRouter | API key, aggregator | link | $200 free credits on signup - multi-model routing gateway |
agnes |
agnes |
Agnes AI | API key, video | link | Get API key at agnes-ai.com |
ai21 |
ai21 |
AI21 Labs | API key | link | $10 trial credits on signup (valid 3 months), no credit card required |
aimlapi |
aiml |
AI/ML API | API key, aggregator | link | Free tier paused (2026) — AI/ML API is now pay-as-you-go only (min $20 top-up); no recurring free credits. |
ainative |
ainative |
AINative Studio | API key | link | Create a free API key at ainative.studio (no card), then paste it here as a Bearer token. |
aion |
aion |
Aion Labs | API key | link | Create a free API key at aionlabs.ai (no card), then paste it here as a Bearer token. |
alibaba |
ali |
Alibaba Cloud Model Studio | API key | link | — |
alibaba-cn |
ali-cn |
Alibaba (China) | API key | link | — |
ant-ling |
ling |
Ant Ling / Ring (inclusionAI) | API key | link | Register and create an API key at the Ant Ling API console (https://chat.ant-ling.com/open), then paste it here. OmniRoute routes chat traffic to https://api.ant-ling.com/v1/chat/completions; the provider is OpenAI-compatible and also exposes an Anthropic-compatible surface. |
anthropic |
anthropic |
Anthropic | API key | link | — |
anyapi |
anyapi |
AnyAPI AI | API key, aggregator | link | Free plan: 100,000 ANY Tokens/day and 100 RPM for eligible Free/Basic models; no credit card required. |
api-airforce |
af |
Api.airforce | API key | link | 55 free tier models including Grok-3, Claude 3.7, Qwen3, Kimi-K2, Gemini 2.5 Flash, DeepSeek-V3 |
arcee-ai |
arcee |
Arcee AI | API key | link | Get API key at arcee.ai |
auriko |
auriko |
Auriko | API key, aggregator | link | Free plan publishes 1,000 Platform RPM and 10,000 BYOK RPM. Platform inference still passes through provider cost; this is not a free-token pool or unlimited free inference. |
azure-ai |
azure-ai |
Azure AI Foundry | API key, enterprise | link | Use your Azure AI Foundry key. Base URL can be https://.services.ai.azure.com/openai/v1/ or https://.openai.azure.com/openai/v1/. |
azure-openai |
azure |
Azure OpenAI | API key, enterprise | link | Use your Azure OpenAI API key. Base URL should be your resource endpoint, for example https://my-resource.openai.azure.com. |
bai |
bai |
b.ai | API key | link | Bearer API key for the b.ai OpenAI-compatible LLM gateway (distinct from TheB.AI). Create a key at https://docs.b.ai, then use https://api.b.ai/v1 as the OpenAI-compatible base URL. |
baichuan |
baichuan |
Baichuan | API key | link | Get API key at platform.baichuan-ai.com |
baidu |
baidu |
Baidu (ERNIE) | API key | link | Get API key at console.bce.baidu.com |
bailian-coding-plan |
bcp |
Alibaba Token Plan | API key | link | — |
baseten |
baseten |
Baseten | API key | link | $30 free trial credits for GPU inference |
bazaarlink |
bzl |
BazaarLink | API key | link | Use your BazaarLink API key (starts with sk-bl-) in Authorization: Bearer . OpenAI SDK works with base URL https://bazaarlink.ai/api/v1. Models use provider/model-name format. |
bedrock |
bedrock |
Amazon Bedrock | API key, enterprise | link | Use your Amazon Bedrock API key and configure the AWS region where your models are enabled (for example eu-west-2). OmniRoute calls Bedrock's native Converse API directly. |
black-forest-labs |
bfl |
Black Forest Labs | API key, image | link | — |
blackbox |
bb |
Blackbox AI | API key | link | Limited free access is available through Blackbox; model availability and account limits apply |
bluesminds |
bm |
BluesMinds | API key | link | Free daily pi credits — supports 200+ models including GPT-4o, GPT-4.1, Claude Sonnet 4.5, Gemini 2.0 Flash, DeepSeek V4, Qwen, Kimi K2 |
byteplus |
bpm |
BytePlus ModelArk | API key | link | — |
bytez |
bytez |
Bytez | API key | link | $1 free credits, refreshes every 4 weeks |
cerebras |
cerebras |
Cerebras | API key | link | Free Trial: 1M tokens/day, 30K TPM, 5 RPM — no credit card. |
charm-hyper |
charm-hyper |
Charm Hyper | API key | link | 100 free monthly Hypercredits on signup |
chat-oripe |
chat-oripe |
Chat Oripe | API key, aggregator | link | Official metadata advertises 2M tokens/month, but the public site and documentation were blocked during audit; treat the quota and brand mapping as unconfirmed. |
chatanywhere |
chatanywhere |
ChatAnywhere | API key, aggregator | link | Personal, educational or research use only: public documentation cites 10,000 points/day and 200 requests/day per IP/key; do not use for commercial traffic. |
cheaperinference |
cinf |
Cheaper Inference | API key | link | — |
chenzk |
chenzk |
Chenzk API | API key | link | — |
chutes |
chutes |
Chutes.ai | API key, aggregator | link | Bearer API key for the Chutes OpenAI-compatible gateway. |
clarifai |
clarifai |
Clarifai | API key, enterprise | link | Use your Clarifai PAT or app-specific API key. OmniRoute targets the OpenAI-compatible endpoint at https://api.clarifai.com/v2/ext/openai/v1 and authenticates with Authorization: Key . |
cloudcode-one |
cloudcode-one |
CloudCode.ONE | API key, aggregator | link | Published free models include glm-4.7-flash and glm-4.6v-flash; no numeric quota is published, and key creation may require credit or a coupon. |
cloudflare-ai |
cf |
Cloudflare Workers AI | API key | link | Requires API Token AND Account ID (found at dash.cloudflare.com) |
clova-studio |
clova |
Naver CLOVA Studio | API key | link | — |
codestral |
codestral |
Codestral | API key | link | — |
cohere |
cohere |
Cohere | API key | link | Free Trial: 1,000 API calls/month for testing, no credit card required |
command-code |
cmd |
Command Code | API key | link | Use a Command Code API key. Requests are sent to Command Code's /alpha/generate endpoint. |
coze |
coze |
Coze | API key | link | Get API key at coze.com/open/api |
crof |
crof |
CrofAI | API key | link | — |
dahl |
dahl |
Dahl | API key | link | Click 'Add Account' to auto-generate a token, or add a manual API key. |
databricks |
databricks |
Databricks | API key, enterprise | link | — |
datarobot |
datarobot |
DataRobot | API key, enterprise | link | Use your DataRobot API token. Optional Base URL can be the account root (for LLM Gateway) or a deployment URL under /api/v2/deployments/. |
deepai |
deepai |
DeepAI | API key, image | link | Use your DeepAI API key. Get one at deepai.org — requires a Pro subscription ($9.99/mo). |
deepinfra |
deepinfra |
DeepInfra | API key | link | Free signup credits for API testing and model exploration |
deepseek |
ds |
DeepSeek | API key | link | 5M free tokens on signup - no credit card required |
dgrid |
dgrid |
DGrid | API key | link | DGrid Free Models Router: 10 requests/minute and 100 requests/day. A $5 lifetime top-up unlocks up to 20 requests/minute and 1,000 requests/day. |
dify |
dify |
Dify | API key | link | Get API key from your Dify instance. |
digitalocean |
digitalocean |
DigitalOcean | API key | link | — |
dit |
dai |
DIT.ai | API key | link | Use your dit.ai API key in Authorization: Bearer . Fully OpenAI-compatible — a drop-in replacement, just change the base URL to https://api.dit.ai/v1. |
doubao |
doubao |
Doubao | API key | link | Get API key at console.volcengine.com |
dxnt |
dxnt |
DXNT / DX Token | API key, aggregator | link | Free accounts are documented at 100 calls/day; the quota may increase through invitations and can vary by account. |
electronhub |
electronhub |
Electron Hub | API key, aggregator | link | Free plan: 5 RPM, $0.25 weekly credits and 10 Neutrinos/day for :free models; family budgets also apply. |
empower |
empower |
Empower | API key, aggregator | link | Bearer API key for the Empower OpenAI-compatible endpoint. |
factory |
factory |
Factory | API key | link | Bearer API key for the Factory OpenAI-compatible gateway. |
fal-ai |
fal |
Fal.ai | API key, image | link | — |
fastrouter |
fastrouter |
FastRouter | API key, aggregator | link | Models with the :free suffix allow 10 requests/day per organization and model; availability may change. |
featherless-ai |
featherless |
Featherless AI | API key | link | Free tier available — no credit card required |
fenayai |
fenayai |
FenayAI | API key, aggregator | link | Bearer API key for the FenayAI OpenAI-compatible gateway. |
fireworks |
fireworks |
Fireworks AI | API key | link | $1 free starter credits on signup for API testing |
free-ai |
free-ai |
Free.ai | API key, aggregator | link | 30,000 tokens/day cover self-hosted models after email verification. Usage beyond the pool can bill at raw cost, and premium external models are paid. |
freeaiapikey |
faik |
FreeAIAPIKey | API key | link | — |
freeinference |
freeinference |
FreeInference | API key, aggregator | link | Free research access without a card; non-Harvard applicants require manual approval and no numeric quota is publicly guaranteed. |
freemodel-dev |
fmd |
FreeModel.dev | API key | link | $300 free credits on signup — no credit card required. Access GPT-5.4 and GPT-5.5 (OpenAI's latest flagship models) through an OpenAI-compatible API. |
freepik |
fpk |
Freepik (Mystic) | API key, image | link | Get API key at freepik.com/developers (Mystic image endpoint) |
freetheai |
fta |
FreeTheAi | API key, aggregator | link | Join the FreeTheAi Discord to get your free API key. |
friendliai |
friendli |
FriendliAI | API key | link | Free tier for serverless inference — no credit card required |
g4f-gemini |
g4fgem |
g4f.space — Gemini | API key, aggregator | link | No auth required. Free tier is limited to 5 requests/minute — sign up at g4f.dev/members.html for higher limits. |
g4f-groq |
g4fgroq |
g4f.space — Groq | API key, aggregator | link | No auth required. Free tier is limited to 5 requests/minute — sign up at g4f.dev/members.html for higher limits. |
g4f-nvidia |
g4fnv |
g4f.space — NVIDIA | API key, aggregator | link | No auth required. Free tier is limited to 5 requests/minute — sign up at g4f.dev/members.html for higher limits. |
g4f-ollama |
g4foll |
g4f.space — Ollama | API key, aggregator | link | No auth required. Free tier is limited to 5 requests/minute — sign up at g4f.dev/members.html for higher limits. |
g4f-pollinations |
g4fpol |
g4f.space — Pollinations | API key, aggregator | link | No auth required. Free tier is limited to 5 requests/minute — sign up at g4f.dev/members.html for higher limits. |
galadriel |
galadriel |
Galadriel | API key | link | ⚠️ DEPRECATED. api.galadriel.ai no longer resolves (sweep 2026-06-19); the inference API appears discontinued. |
gemini |
gemini |
Gemini (Google AI Studio) | API key | link | Free tier available through Google AI Studio; current per-model quotas and regional limits apply |
getgoapi |
ggo |
GoAPI | API key, aggregator | link | — |
gigachat |
gigachat |
GigaChat (Sber) | API key | link | — |
gitlab |
gitlab |
GitLab Duo PAT | API key | link | GitLab personal access token for the public Code Suggestions API. Configure a self-hosted base URL when not using gitlab.com. |
gitlawb |
glb |
Gitlawb Opengateway (MiMo) | API key | link | Free MiMo (xiaomi/mimo-v2.5) revoked 2026-05 — Opengateway is now a pay-as-you-go credit gateway; no recurring free model. |
gitlawb-gmi |
glb-gmi |
Gitlawb Opengateway (GMI Cloud) | API key | link | Free Nemotron promo ended 2026-06 — the GMI Cloud route is now pay-as-you-go credit only. |
glm |
glm |
GLM Coding | API key | link | — |
glm-cn |
glmcn |
GLM Coding (China) | API key | link | — |
glmt |
glmt |
GLM Thinking | API key | link | — |
groq |
groq |
Groq | API key | link | Free tier: 30 RPM / 14.4K RPD — no credit card |
hackclub |
hc |
Hackclub AI | API key, aggregator | link | Sign in with your Hack Club account at ai.hackclub.com. |
haiper |
hp |
Haiper | API key, video | link | Get API key at haiper.ai/haiper-api |
hcnsec |
hcnsec |
Huancheng Public API | API key | link | Get API key at api.hcnsec.cn |
helixmind |
helixmind |
HelixMind | API key, aggregator | link | Previously circulated 3 RPM/50 RPD and no-card claims were not confirmed during the 2026-08-02 audit; current quota and billing require account verification. |
helyxai |
helyxai |
Helyx AI | API key, aggregator | link | Operational Free plan documents 100,000 tokens/day; the site's separate 2M+ marketing claim conflicts and is not treated as a quota guarantee. |
heroku |
heroku |
Heroku AI | API key, enterprise | link | — |
huggingface |
hf |
HuggingFace | API key | link | Free Inference API for thousands of models (Whisper, VITS, SDXL…) |
hyperbolic |
hyp |
Hyperbolic | API key | link | $1-5 trial credits on signup for serverless inference |
ideogram |
ideo |
Ideogram | API key | link | Get API key at ideogram.ai/docs/api |
iflytek |
iflytek |
iFlytek Spark | API key | link | Get API key at console.xfyun.cn |
inception |
inception |
Inception | API key | link | 10M free tokens on signup, no credit card required. |
inference-net |
inet |
Inference.net | API key | link | $25 free credits on signup plus research grants available |
internlm |
internlm |
InternLM (Intern-S1) | API key | link | Free monthly quota ~1M input / 3M output tokens (~10 RPM) |
jina-ai |
jina |
Jina AI | API key, embed/rerank | link | Bearer API key for the Jina AI rerank API. |
jina-reader |
jr |
Jina Reader | API key | link | — |
kenari |
kenari |
Kenari | API key | link | Use your Kenari API key (kn-...) in Authorization: Bearer . Fully OpenAI-compatible. API base URL: https://kenari.id/v1. |
kie |
kie |
KIE.AI | API key | link | — |
kilo-gateway |
kg |
Kilo Gateway | API key, aggregator | link | — |
kimi |
kimi |
Kimi (Legacy Moonshot API) | API key | link | — |
kimi-coding-apikey |
kmca |
Kimi Code API Key | API key | link | — |
lambda-ai |
lambda |
Lambda AI | API key | link | — |
laozhang |
lz |
LaoZhang AI | API key, aggregator | link | — |
leonardo |
leo |
Leonardo AI | API key, video | link | Get API key at leonardo.ai/developer |
liquid |
liquid |
Liquid AI | API key | link | Get API key at liquid.ai |
literouter |
literouter |
LiteRouter | API key, aggregator | link | Free model variants use the :free suffix; daily credit limits vary by model and free input is capped at 5,000 tokens. |
llamagate |
llamagate |
LlamaGate | API key | link | — |
llm-kiwi |
llmkiwi |
LLM.Kiwi | API key, aggregator | link | Free plan exposes auto and hrLLM; the published 40 requests/hour limit applies to hrLLM. |
llm7 |
llm7 |
LLM7.io | API key | link | Use any non-empty key (for example 'unused'). If older built-in models return model_unavailable, use Available Models → Import from /models or Auto-Sync; verified live model: gemini-3.1-flash-lite. |
llmgateway |
llmgateway |
LLM Gateway | API key, aggregator | link | Hosted Free plan: free-priced models are limited to 5 requests per 10 minutes when the account has no credits. |
longcat |
lc |
LongCat AI | API key | link | Free: one-time 10M-token grant after account signup + KYC verification (LongCat-2.0). One-time only — not a recurring daily/monthly allowance. |
maritalk |
maritalk |
Maritalk | API key | link | — |
meganova-ai |
meganova-ai |
MegaNova AI | API key, aggregator | link | Free signup without a card. Published Tier 1 per-model quotas total 550 requests/day; they are not a shared global pool, and paid overage can apply if enabled. |
meta-llama |
meta |
Meta Llama API | API key | link | — |
minimax |
minimax |
Minimax Coding | API key, video | link | — |
minimax-cn |
minimax-cn |
Minimax (China) | API key | link | — |
mistral |
mistral |
Mistral | API key | link | Free Experiment tier: rate-limited access to all models, no credit card required |
mixedbread |
mxbai |
Mixedbread AI | API key | link | Bearer API key for the Mixedbread embeddings API. |
mixlayer |
mixlayer |
Mixlayer | API key, aggregator | link | The qwen/qwen3.5-4b-free model is free for prototyping and rate-limited; no fixed public RPM or daily quota is confirmed. |
mnn-ai |
mnn-ai |
MNN AI | API key, aggregator | link | Free plan: $1 monthly credits, 10 RPM and access only to models marked Free. |
modal |
mdl |
Modal | API key, enterprise | link | Use the bearer token that protects your Modal deployment, if enabled. Base URL should point to your OpenAI-compatible Modal app, for example https://--.modal.run/v1. |
modelscope |
ms |
ModelScope | API key | link | Free tier via ModelScope API-Inference — Alibaba account required. |
monsterapi |
monster |
MonsterAPI | API key | link | Get API key at monsterapi.ai |
moonshot |
moonshot |
Kimi | API key | link | — |
morph |
morph |
Morph | API key | link | Free tier: 250K credits/month, $0 |
muse-code |
mc |
Muse Code (Meta) | API key | link | Use your META_API_KEY env var as a Bearer token. Muse Code CLI uses the OpenAI Responses API wire format (POST /responses). |
naga-ac |
naga |
Naga.ac | API key, aggregator | link | Get API key at naga.ac — Google/GitHub/Discord signup available. |
naga-ai |
naga-ai |
Naga AI | API key, aggregator | link | Models marked :free are publicly listed, but no numeric quota is confirmed. Naga's policy warns that free-tier prompts and outputs may be collected or used for training. |
nanogpt |
nanogpt |
NanoGPT | API key | link | — |
nara |
nara |
NaraRouter | API key | link | Get a free API key via NaraRouter's Telegram channel, then paste it here as a Bearer token. |
navy |
navy |
NavyAI | API key | link | Create a free API key from the NavyAI dashboard, then paste it here as a Bearer token. |
nebius |
nebius |
Nebius AI | API key | link | ~$1 trial credits on signup for API testing |
nlpcloud |
nlpc |
NLP Cloud | API key | link | Use your NLP Cloud API key in Authorization: Token . OmniRoute targets the chatbot endpoint on https://api.nlpcloud.io/v1/gpu//chatbot by default. |
nomic |
nomic |
Nomic | API key | link | Get API key at atlas.nomic.ai |
nous-research |
nous |
Nous Research | API key | link | Use your Nous Portal API key. OmniRoute targets the official OpenAI-compatible inference endpoint at https://inference-api.nousresearch.com/v1. |
novita |
novita |
Novita AI | API key, video, aggregator | link | $0.50 trial credits on signup (valid about 1 year) |
nscale |
nscale |
nScale | API key | link | $5 free credits on signup for inference testing |
nube |
nube |
Nube.sh | API key | link | — |
nvidia |
nvidia |
NVIDIA NIM | API key | link | Free dev access: ~40 RPM, 70+ models (Kimi K2.5, GLM 4.7, DeepSeek V3.2...) |
oci |
oci |
OCI Generative AI | API key, enterprise | link | Use your OCI Generative AI API key or IAM bearer token. Base URL can be https://inference.generativeai..oci.oraclecloud.com/openai/v1/. |
ofoxai |
ofoxai |
OfoxAI | API key, aggregator | link | The current catalog advertises 10+ free models without a public numeric quota; review upstream provenance, retention and training terms before production use. |
ollama-cloud |
ollamacloud |
Ollama Cloud | API key | link | — |
openadapter |
oad |
OpenAdapter | API key | link | Use your OpenAdapter API key in Authorization: Bearer sk-cv-. Fully OpenAI-compatible. API base URL: https://api.openadapter.in/v1. |
openai |
openai |
OpenAI | API key | link | — |
opencode-go |
opencode-go |
OpenCode Go | API key | link | — |
opencode-zen |
opencode-zen |
OpenCode Zen | API key | link | — |
openference-api |
ofa |
Openference API | API key | link | Free plan: 3-day trial with open-source models — no credit card required |
openrouter |
openrouter |
OpenRouter | API key, aggregator | link | Free models at $0/token with :free suffix - 20 RPM / 200 RPD |
openvecta |
openvecta |
OpenVecta | API key | link | Free credits on signup for OpenAI-compatible inference across LLMs, embeddings, and reasoning models |
orcarouter |
orcarouter |
OrcaRouter | API key | link | — |
ovhcloud |
ovh |
OVHcloud AI | API key | link | — |
perplexity |
pplx |
Perplexity | API key | link | — |
piapi |
pi |
PiAPI | API key, aggregator | link | — |
pioneer |
pn |
Pioneer AI | API key | link | $75 free usage credits — no credit card required |
plamo |
plamo |
PLaMo | API key | link | — |
poe |
poe |
Poe | API key, aggregator | link | Bearer API key for the Poe OpenAI-compatible API. |
poixe-ai |
poixe-ai |
Poixe AI | API key, aggregator | link | Current public free limits are small and model-group specific: 2 RPM/5 RPD for large-cup models and 20 RPM/50 RPD for small-cup models. |
pollinations |
pol |
Pollinations AI | API key, video | link | Anonymous/keyless access to the documented free models is best-effort. Local v3.8.50 verification (2026-07-31) returned 401 via OmniRoute and Cloudflare 1010 on direct upstream probes from the same network. Premium models still require a Pollinations API key from enter.pollinations.ai. |
poolside |
poolside |
Poolside | API key | link | Laguna S 2.1 and XS 2.1 are free during Preview; no public numeric quota is published. |
predibase |
predibase |
Predibase | API key | link | ⚠️ DEPRECATED. serving.app.predibase.com no longer resolves (sweep 2026-06-19); the managed serving API appears discontinued. |
publicai |
publicai |
PublicAI | API key | link | Requires an API key — one-time signup credit, then paid |
qianfan |
qianfan |
Baidu Qianfan | API key | link | — |
qiniu |
qiniu |
Qiniu | API key | link | — |
qwen-cloud |
qwc |
Qwen Cloud | API key | link | — |
qwen-cloud-token-plan |
qct |
Qwen Cloud Token Plan | API key | link | — |
recraft |
recraft |
Recraft | API key, image | link | — |
regolo |
regolo |
Regolo AI | API key | link | Get your Regolo API key from regolo.ai, then paste it here as a Bearer token. |
reka |
reka |
Reka | API key | link | Use your Reka API key. OmniRoute supports the OpenAI-compatible base URL https://api.reka.ai/v1 and sends both Authorization and X-Api-Key headers for compatibility. |
requesty |
requesty |
Requesty | API key | link | Free tier ~200 requests/day - multi-model routing gateway (300+ models) |
routeway |
routeway |
Routeway | API key | link | Create a free API key at routeway.ai, then paste it here as a Bearer token. |
runwayml |
runway |
Runway | API key, video | link | Use your Runway API key in Authorization: Bearer . OmniRoute targets the current Runway API at https://api.dev.runwayml.com/v1 and sends the required X-Runway-Version header automatically. |
sambanova |
samba |
SambaNova | API key | link | $5 free credits on signup (30-day validity), no credit card required |
sap |
sap |
SAP Generative AI Hub | API key, enterprise | link | Use your SAP AI Core bearer token. Base URL can be your AI_API_URL root or a deploymentUrl from Generative AI Hub. |
sarvam |
sarvam |
Sarvam AI | API key | link | ₹1,000 in free signup credits — never expire |
scaleway |
scw |
Scaleway AI | API key | link | 1M free tokens for new accounts — EU/GDPR compliant (Paris), Qwen3 235B & Llama 70B |
sealion |
sealion |
SEA-LION | API key | link | Sign in at sea-lion.ai with Google (no card, no region wall), create an API key, then paste it here. |
segmind |
segmind |
Segmind | API key, image, video | link | Use your Segmind API key in the x-api-key header. OmniRoute targets https://api.segmind.com/v1/ and returns the generated image/video bytes directly. |
sensenova |
sensenova |
SenseNova | API key | link | Get API key at platform.sensenova.cn |
siliconflow |
siliconflow |
SiliconFlow | API key | link | $1 free credits plus currently listed $0 models after identity verification; availability and limits may change |
snowflake |
snowflake |
Snowflake Cortex | API key, enterprise | link | — |
sparkdesk |
sparkdesk |
SparkDesk | API key | link | Get API key at console.xfyun.cn |
speka |
speka |
Speka AI | API key, aggregator | link | Free plan: $1 monthly usage, 10 RPM, one API key and access to open models and the playground; no card required. |
stability-ai |
stability |
Stability AI | API key, image | link | — |
stepfun |
stepfun |
StepFun | API key | link | Get API key at platform.stepfun.com |
sumopod |
sumopod |
SumoPod | API key | link | Use your SumoPod API key (sk-...) in Authorization: Bearer . Fully OpenAI-compatible. API base URL: https://ai.sumopod.com/v1. |
suno |
suno |
Suno | API key | link | Paste session cookie from suno.ai (Clerk auth) |
synthetic |
synthetic |
Synthetic | API key, aggregator | link | — |
tencent |
tencent |
Tencent Hunyuan | API key | link | Get API key at console.cloud.tencent.com |
thebai |
thebai |
TheB.AI | API key, aggregator | link | Bearer API key for the TheB.AI OpenAI-compatible gateway. |
tinyfish |
tf |
TinyFish Fetch | API key | link | X-API-Key from agent.tinyfish.ai/api-keys |
together |
together |
Together AI | API key, video | link | — |
tokenreply |
tokenreply |
TokenReply | API key, aggregator | link | Free-tagged models have model- and campaign-specific daily limits; no fixed global free quota is published. |
tokenrouter |
trk |
TokenRouter | API key | link | Use your TokenRouter API key in Authorization: Bearer . Fully OpenAI-compatible. API base URL: https://api.tokenrouter.com/v1. |
topaz |
topaz |
Topaz | API key, image | link | — |
typhoon |
typhoon |
Typhoon | API key | link | Free API key with a 5 req/s and 200 req/m rate limit. |
udio |
udio |
Udio | API key | link | Paste session cookie from udio.com (Supabase auth) |
uncloseai |
unc |
UncloseAI | API key | link | No auth required. API accepts any non-empty string as key for identification. If older built-in models return 404, use Available Models → Import from /models or Auto-Sync; verified live model: solidrust/Hermes-3-Llama-3.1-8B-AWQ. |
unorouter |
unorouter |
UnoRouter | API key, aggregator | link | Models with the :free suffix do not debit balance; limit is 1 request/minute per free model per user. |
upstage |
upstage |
Upstage | API key | link | — |
v0-vercel |
v0 |
v0 (Vercel) | API key | link | — |
venice |
venice |
Venice.ai | API key | link | — |
vercel-ai-gateway |
vag |
Vercel AI Gateway | API key, aggregator | link | — |
vertex |
vertex |
Vertex AI | API key, enterprise | link | Provide Service Account JSON or OAuth access_token |
vertex-partner |
vp |
Vertex AI Partners | API key, enterprise | link | Provide the same Service Account JSON used for Vertex AI partner models. |
void-ai |
void-ai |
Void AI | API key, aggregator | link | The public model catalog marks some models with a free plan requirement, but access is conditional and no numeric quota is confirmed. |
volcengine |
volcengine |
Volcengine | API key | link | — |
voyage-ai |
voyage |
Voyage AI | API key, embed/rerank | link | Bearer API key for Voyage AI embeddings and rerank APIs. |
wafer |
wafer |
Wafer AI | API key | link | — |
wandb |
wandb |
Weights & Biases Inference | API key | link | — |
watsonx |
watsonx |
IBM watsonx.ai Gateway | API key, enterprise | link | Use your watsonx bearer token. Base URL can be https://.ml.cloud.ibm.com/ml/gateway/v1/ or a self-managed /ml/gateway/v1 endpoint. |
writer |
writer |
Writer | API key | link | — |
x5lab |
x5lab |
X5Lab | API key | link | Use your X5Lab API key (x5-...) in Authorization: Bearer . Fully OpenAI-compatible. API base URL: https://api.x5lab.dev/v1. |
xai |
xai |
xAI (Grok) | API key | link | — |
xiaomi-mimo |
mimo |
Xiaomi MiMo | API key | link | — |
xiaomi-mimo-token-plan |
mimotp |
Xiaomi MiMo Token Plan | API key | link | — |
yi |
yi |
Yi (01.AI) | API key | link | Get API key at platform.lingyiwanwu.com |
yolo-auto |
yolo-auto |
Yolo-Auto | API key, aggregator | link | Free API access is request-limited and intended for testing; no numeric daily quota is published and free access is not promised indefinitely. |
zai |
zai |
Z.AI | API key | link | — |
zenmux |
zm |
ZenMux | API key | link | Use your ZenMux API key in Authorization: Bearer . ZenMux is fully OpenAI-compatible. Base URL: https://zenmux.ai/api/v1. |
zerolimitai |
zerolimitai |
ZeroLimitAI | API key, aggregator | link | Temporary free trial is advertised, but official pages conflict between 3 and 7 days; a 100-calls/day claim is not treated as permanent. |
zylo-api |
zylo |
Zylo API | API key, aggregator | link | Basic plan: 10 RPM, 7,200 requests/day and 200,000 tokens/day; limited to Basic text models. |
Local Providers (12)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
comfyui |
comfyui |
ComfyUI | Local | link | No API key required. Configure the local ComfyUI base URL (default: http://localhost:8188). |
docker-model-runner |
dmr |
Docker Model Runner | Local, self-hosted | link | API key optional. Configure the local Docker Model Runner OpenAI-compatible base URL (default: http://localhost:12434/v1). |
lemonade |
lemonade |
Lemonade Server | Local, self-hosted | link | API key optional. Configure the local Lemonade OpenAI-compatible base URL (default: http://localhost:13305/api/v1). |
llama-cpp |
llamacpp |
llama.cpp | Local, self-hosted | link | API key optional (use any value, e.g. sk-no-key-required). Configure the llama-server OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1). Note: if Llamafile is also installed, both default to port 8080 — run only one at a time or override the port. |
llamafile |
llamafile |
Llamafile | Local, self-hosted | link | API key optional. Configure the local Llamafile OpenAI-compatible base URL (default: http://127.0.0.1:8080/v1). |
lm-studio |
lmstudio |
LM Studio | Local, self-hosted | link | API key optional. Configure the local LM Studio OpenAI-compatible base URL (default: http://localhost:1234/v1). |
ollama-local |
ollama |
Ollama | Local, self-hosted | link | No API key required. Ollama runs locally — configure its OpenAI-compatible base URL (default: http://localhost:11434/v1) and make sure Ollama is running before connecting. |
oobabooga |
ooba |
oobabooga | Local, self-hosted | link | API key optional. Configure the local oobabooga OpenAI-compatible base URL (default: http://localhost:5000/v1). |
sdwebui |
sdwebui |
SD WebUI | Local | link | No API key required. Configure the local WebUI base URL (default: http://localhost:7860). |
triton |
triton |
NVIDIA Triton | Local, self-hosted | link | API key optional. Configure the Triton OpenAI-compatible base URL (default: http://localhost:8000/v1). |
vllm |
vllm |
vLLM | Local, self-hosted | link | API key optional. Configure the local vLLM OpenAI-compatible base URL (default: http://localhost:8000/v1). |
xinference |
xinference |
XInference | Local, self-hosted | link | API key optional. Configure the local XInference OpenAI-compatible base URL (default: http://localhost:9997/v1). |
Search Providers (12)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
brave-search |
brave-search |
Brave Search | Search | link | Subscription token from Brave Search API dashboard |
exa-search |
exa-search |
Exa Search | Search | link | API key from dashboard.exa.ai |
firecrawl |
fc |
Firecrawl | Search | link | API key from firecrawl.dev/app/api-keys (or set your self-hosted Firecrawl base URL) |
google-pse-search |
google-pse |
Google Programmable Search | Search | link | Requires a Google API key and your Programmable Search Engine ID (cx) |
linkup-search |
linkup |
Linkup Search | Search | link | Bearer API key from the Linkup dashboard |
ollama-search |
ollama-search |
Ollama Search | Search | link | Same API key as Ollama Cloud (from ollama.com/settings/keys) |
perplexity-search |
pplx-search |
Perplexity Search | Search | link | Same API key as Perplexity (pplx-...) |
searchapi-search |
searchapi |
SearchAPI | Search | link | API key from SearchAPI (query param or Bearer auth) |
searxng-search |
searxng |
SearXNG Search | Search | link | API key is optional. Set your SearXNG base URL. Some instances may require a bearer token for access. |
serper-search |
serper-search |
Serper Search | Search | link | API key from serper.dev dashboard |
tavily-search |
tavily-search |
Tavily Search | Search | link | API key from app.tavily.com (format: tvly-...) |
youcom-search |
youcom-search |
You.com Search | Search | link | X-API-Key from the You.com platform dashboard |
Audio-only Providers (12)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
assemblyai |
aai |
AssemblyAI | Audio | link | — |
aws-polly |
polly |
AWS Polly | Audio | link | Use AWS Secret Access Key as API key; set providerSpecificData.accessKeyId and optional region. |
cartesia |
cartesia |
Cartesia | Audio | link | — |
deepgram |
dg |
Deepgram | Audio | link | — |
elevenlabs |
el |
ElevenLabs | Audio | link | — |
fishaudio |
fishaudio |
Fish Audio | Audio | link | — |
gladia |
gladia |
Gladia | Audio | link | — |
inworld |
inworld |
Inworld | Audio | link | — |
playht |
playht |
PlayHT | Audio | link | — |
rev-ai |
revai |
Rev AI | Audio | link | — |
soniox |
sx |
Soniox | Audio | link | — |
speechmatics |
sm |
Speechmatics | Audio | link | Free tier — 8 hours/month, no credit card required. Batch (async) mode only. |
Upstream Proxy Providers (2)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
9router |
nr |
9router | Upstream proxy | link | — |
cliproxyapi |
cpa |
CLIProxyAPI | Upstream proxy | link | — |
Cloud Agent Providers (3)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
codex-cloud |
codex-cloud |
Codex Cloud | Cloud agent | link | OpenAI API key with Codex Cloud task access. |
devin |
devin |
Devin | Cloud agent | link | Devin API key for cloud agent sessions. |
jules |
jules |
Google Jules | Cloud agent | link | Jules API key for creating and managing cloud coding tasks. |
System Providers (1)
| ID | Alias | Name | Tags | Website | Notes |
|---|---|---|---|---|---|
auto |
auto |
Auto (Zero-Config) | System | — | — |
Sources of truth
- Catalog:
src/shared/constants/providers.ts - Registry (per-model details):
open-sse/config/providerRegistry.ts - Executors:
open-sse/executors/(100 implementations) - Translators:
open-sse/translator/
See Also
- FREE_TIERS.md — curated free-tier guide
- USER_GUIDE.md — provider setup walkthrough
- ARCHITECTURE.md — overall architecture