mirror of
https://github.com/diegosouzapw/OmniRoute.git
synced 2026-08-04 22:32:12 +03:00
* chore(release): open v3.8.10 development cycle Bump 3.8.9 → 3.8.10 across package.json, lockfile, electron, open-sse, and docs/reference/openapi.yaml; add the [3.8.10] CHANGELOG section (root + 41 i18n mirrors) as the integration target for the cycle. Entries land here as work merges into release/v3.8.10; finalized by the release flow. * fix(providers): resolve web provider alias collisions Assign unique aliases to HuggingChat, Kimi Web, and Qwen Web so they no longer shadow primary providers or trigger startup warnings. Add a unit test to enforce provider alias uniqueness and prevent future collisions. Also expand local ignore and VS Code exclude rules for agent, build, and worktree artifacts. * fix(responses): normalize image_url parts across input paths (#3150) Normalize image_url parts across all Responses input paths. Integrated into release/v3.8.10. * fix(api-manager): preserve API key expiration local time (#3146) Preserve API key expiration local time + clear button. Integrated into release/v3.8.10. * Strip previous_response_id for stateless Responses upstreams (#3143) Strip previous_response_id for stateless Responses upstreams (auto/strip/preserve). Integrated into release/v3.8.10. * fix(opencode-plugin): map thinking cap to interleaved in model+combo (#3138) Map caps.thinking to ModelV2.capabilities.interleaved for opencode-plugin. Integrated into release/v3.8.10. * fix(providers): use synced models as fallback for all providers (#3148) Use synced models as authoritative local catalog for all providers (+regression test). Integrated into release/v3.8.10. * fix(qoder): bifurcate validation by token type — PAT→Cosy, regular API key→dashscope (#3149) Bifurcate Qoder validation by token type (PAT→Cosy, regular→dashscope) +regression test. Integrated into release/v3.8.10. * fix(antigravity): dynamic model resolution via MITM alias table (#3144) Dynamic antigravity MITM model resolution in the executor (+bug fix +regression test; DB import dropped from client-reachable config). Integrated into release/v3.8.10. * Feature/batch allow big (#3128) Podman deployment options + larger upload body-size limits (+CONTAINER_HOST docs). Integrated into release/v3.8.10. * fix(fireworks): preserve fully-qualified router/model IDs (#3133) (#3160) Fireworks router IDs (accounts/fireworks/routers/...) were double-prefixed with accounts/fireworks/models/ → upstream 404. Add optional acceptedModelIdPrefixes to the registry entry and skip the prepend when the model already starts with an accepted prefix. Co-authored-by: KooshaPari <KooshaPari@users.noreply.github.com> * fix(llama-cpp): route to configured local baseUrl instead of OpenAI (#3136) (#3161) llama-cpp was missing from the local-provider group in buildUrl(), so it fell through to the OpenAI baseUrl and returned an OpenAI 401. Add the case to resolve the connection's providerSpecificData.baseUrl. Co-authored-by: tjengbudi <tjengbudi@users.noreply.github.com> * fix(t3-chat-web): parse cookies + convexSessionId from stored credential (#3007) (#3162) The executor read credentials.cookies/convexSessionId, but the pipeline only stores the pasted string under apiKey → t3.chat always 400'd. Parse both values from apiKey (fallback accessToken), mirroring validation.ts. Co-authored-by: minhtran162 <minhtran162@users.noreply.github.com> * fix(minimax): stop capping MiniMax-M3 / M2.7 max_tokens at 8192 (#3141) (#3163) MiniMax-M3 had no MODEL_SPECS entry and capitalized MiniMax-M2.7 missed its lowercase spec (case-sensitive lookup) → both fell to the 8192 default cap. Add the M3 spec (512K output), alias the capitalized ids, and make getModelSpec lookups case-insensitive. Co-authored-by: totaltube <totaltube@users.noreply.github.com> * fix(github-copilot): discover model catalog live from api.githubcopilot.com (#3120, #3121) (#3164) The github (Copilot) provider had a static hardcoded catalog with no discovery source, so Import Models never refreshed (#3120) and advertised non-entitled models that 400 on use (#3121). Add a live /models fetch with fallback to the static list. Co-authored-by: gabrielmoreira <gabrielmoreira@users.noreply.github.com> * fix(combo): invalidate nested-combo cache on edits + log DATA_DIR (#3147) (#3165) Editing a combo did not invalidate the 10s nested-combo expansion caches (chat.ts getCombosCachedForChat + chatCore.ts getCombosCached; the exported clearCombosCache was dead code), so a removed nested target/model could be served as a phantom for up to 10s. Wire a shared monotonic combos-cache version in readCache (bumped by invalidateDbCache("combos") on every combo write); both cache layers treat a version mismatch as a miss. Also log the resolved DATA_DIR/SQLITE_FILE absolute path at DB init so the reporter's 'persists across restart + volume wipe' symptom (a multi-replica Docker volume/DATA_DIR mismatch, not a routing bug) is diagnosable from logs. Includes consolidated CHANGELOG entries for #3133/#3136/#3007/#3141/#3120/#3121. Co-authored-by: ViFigueiredo <ViFigueiredo@users.noreply.github.com> * fix(web-tools): parse bare JSON tool calls (#3157) Parse bare JSON tool calls for deepseek-web (#2820) + fuzzy tool-name matching. Integrated into release/v3.8.10. * fix(misc): minor fixes across reasoning cache, account fallback, binary manager (#3177) Misc: ProviderProfile export, DeepSeek reasoning regex, binary guard. Integrated into release/v3.8.10. * fix(kiro): minor OAuth social exchange tweaks (#3176) Kiro social OAuth: optional targetProvider passthrough. Integrated into release/v3.8.10. * deps: bump hono from 4.12.18 to 4.12.23 (#3179) Bump hono to 4.12.23. Integrated into release/v3.8.10. * fix(providerRegistry): update kilocode format and executor (#3166) kilocode: openai format + default executor (matches kilo-gateway) + registry test. Integrated into release/v3.8.10. * feat(metrics): cross-request TTFT and gap latency after tool calls (#3173) Cross-request TTFT + gap-after-tool latency metrics (+test). Integrated into release/v3.8.10. * feat(dashboard): provider stats API endpoint and dashboard page (#3175) Provider stats dashboard + API (SQL moved to db module per Hard Rule #5, +test). Integrated into release/v3.8.10. * fix(usage): sequential+spaced OAuth quota sync, reactive force-refresh, actionable 401 (#3156) Sequential+spaced OAuth quota sync, reactive force-refresh on 401, actionable 401 in UI. Integrated into release/v3.8.10. * fix(healthcheck): per-provider proactive-refresh skip list (rescue short-TTL OAuth) (#3159) Per-provider proactive-refresh skip list (OMNIROUTE_HEALTHCHECK_SKIP_PROVIDERS) to rescue short-TTL OAuth. Integrated into release/v3.8.10. * feat(quota): show OAuth token expiry on provider cards (small, blue, informative) (#3178) Show OAuth token expiry on provider cards (small, blue, informative). Integrated into release/v3.8.10. * fix(providers): empty refresh must not resurface just-cleared synced models (#3181) Empty refresh must not resurface just-cleared synced models (fixes the release-blocking provider-models-route test). Integrated into release/v3.8.10. * chore(release): v3.8.10 — 2026-06-04 (finalize CHANGELOG) --------- Co-authored-by: Wilson <pedbookmed@gmail.com> Co-authored-by: Xiangzhe <32761048+xz-dev@users.noreply.github.com> Co-authored-by: Jan Leon <Jan.gaschler@gmail.com> Co-authored-by: M.M <mr.maatoug@gmail.com> Co-authored-by: Hernan Javier Ardila Sanchez <hjasgr@gmail.com> Co-authored-by: Markus Hartung <mail@hartmark.se> Co-authored-by: KooshaPari <KooshaPari@users.noreply.github.com> Co-authored-by: tjengbudi <tjengbudi@users.noreply.github.com> Co-authored-by: minhtran162 <minhtran162@users.noreply.github.com> Co-authored-by: totaltube <totaltube@users.noreply.github.com> Co-authored-by: gabrielmoreira <gabrielmoreira@users.noreply.github.com> Co-authored-by: ViFigueiredo <ViFigueiredo@users.noreply.github.com> Co-authored-by: PizzaV <103120356+pizzav-xyz@users.noreply.github.com> Co-authored-by: dependabot[bot] <49699333+dependabot[bot]@users.noreply.github.com> Co-authored-by: Nicolas Lorin <androw95220@gmail.com>
1669 lines
51 KiB
TypeScript
1669 lines
51 KiB
TypeScript
import test from "node:test";
|
|
import assert from "node:assert/strict";
|
|
import fs from "node:fs";
|
|
import os from "node:os";
|
|
import path from "node:path";
|
|
|
|
const TEST_DATA_DIR = fs.mkdtempSync(path.join(os.tmpdir(), "omniroute-chat-pipeline-"));
|
|
process.env.DATA_DIR = TEST_DATA_DIR;
|
|
process.env.REQUIRE_API_KEY = "false";
|
|
process.env.API_KEY_SECRET = process.env.API_KEY_SECRET || "test-chat-pipeline-secret";
|
|
|
|
const core = await import("../../src/lib/db/core.ts");
|
|
const providersDb = await import("../../src/lib/db/providers.ts");
|
|
const combosDb = await import("../../src/lib/db/combos.ts");
|
|
const settingsDb = await import("../../src/lib/db/settings.ts");
|
|
const apiKeysDb = await import("../../src/lib/db/apiKeys.ts");
|
|
const callLogsDb = await import("../../src/lib/usage/callLogs.ts");
|
|
const readCacheDb = await import("../../src/lib/db/readCache.ts");
|
|
const { invalidateMemorySettingsCache } = await import("../../src/lib/memory/settings.ts");
|
|
const { skillRegistry } = await import("../../src/lib/skills/registry.ts");
|
|
const { skillExecutor } = await import("../../src/lib/skills/executor.ts");
|
|
const { handleChat } = await import("../../src/sse/handlers/chat.ts");
|
|
const { initTranslators } = await import("../../open-sse/translator/index.ts");
|
|
const { clearInflight } = await import("../../open-sse/services/requestDedup.ts");
|
|
const { setCliCompatProviders } = await import("../../open-sse/config/cliFingerprints.ts");
|
|
const { BaseExecutor } = await import("../../open-sse/executors/base.ts");
|
|
const { getCodexClientVersion } = await import("../../open-sse/config/codexClient.ts");
|
|
const { GEMINI_CLI_VERSION, GEMINI_CLI_GOOGLE_API_NODE_CLIENT_VERSION } =
|
|
await import("../../open-sse/services/geminiCliHeaders.ts");
|
|
const { getCircuitBreaker, resetAllCircuitBreakers } =
|
|
await import("../../src/shared/utils/circuitBreaker.ts");
|
|
const { clearProviderFailure } = await import("../../open-sse/services/accountFallback.ts");
|
|
|
|
const originalFetch = globalThis.fetch;
|
|
const originalRetryDelayMs = BaseExecutor.RETRY_CONFIG.delayMs;
|
|
|
|
type SeedConnectionOverrides = {
|
|
name?: string;
|
|
authType?: string;
|
|
apiKey?: string;
|
|
accessToken?: string;
|
|
refreshToken?: string;
|
|
tokenType?: string;
|
|
expiresAt?: string;
|
|
tokenExpiresAt?: string;
|
|
isActive?: boolean;
|
|
testStatus?: string;
|
|
priority?: number;
|
|
rateLimitedUntil?: string | number | null;
|
|
providerSpecificData?: Record<string, unknown>;
|
|
};
|
|
|
|
type FetchCall = {
|
|
url: string;
|
|
method?: string;
|
|
headers: Record<string, string>;
|
|
body: Record<string, any> | null;
|
|
};
|
|
|
|
type SeedApiKeyOptions = {
|
|
name?: string;
|
|
noLog?: boolean;
|
|
allowedConnections?: string[];
|
|
allowedModels?: string[];
|
|
};
|
|
|
|
function toPlainHeaders(headers: HeadersInit | undefined | null) {
|
|
if (!headers) return {};
|
|
if (headers instanceof Headers) return Object.fromEntries(headers.entries());
|
|
if (Array.isArray(headers)) return Object.fromEntries(headers);
|
|
return Object.fromEntries(
|
|
Object.entries(headers).map(([key, value]) => [key, value == null ? "" : String(value)])
|
|
);
|
|
}
|
|
|
|
function buildRequest({
|
|
url = "http://localhost/v1/chat/completions",
|
|
body,
|
|
authKey = null,
|
|
headers = {},
|
|
}: {
|
|
url?: string;
|
|
body?: unknown;
|
|
authKey?: string | null;
|
|
headers?: Record<string, string>;
|
|
} = {}) {
|
|
const requestHeaders: Record<string, string> = {
|
|
"Content-Type": "application/json",
|
|
...headers,
|
|
};
|
|
|
|
if (authKey) {
|
|
requestHeaders.Authorization = `Bearer ${authKey}`;
|
|
}
|
|
|
|
return new Request(url, {
|
|
method: "POST",
|
|
headers: requestHeaders,
|
|
body: typeof body === "string" ? body : JSON.stringify(body),
|
|
});
|
|
}
|
|
|
|
function buildOpenAIResponse(text = "ok", model = "gpt-4o-mini", usage = null) {
|
|
return new Response(
|
|
JSON.stringify({
|
|
id: "chatcmpl_json",
|
|
object: "chat.completion",
|
|
model,
|
|
choices: [
|
|
{
|
|
index: 0,
|
|
message: { role: "assistant", content: text },
|
|
finish_reason: "stop",
|
|
},
|
|
],
|
|
usage: usage || {
|
|
prompt_tokens: 4,
|
|
completion_tokens: 2,
|
|
total_tokens: 6,
|
|
},
|
|
}),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "application/json" },
|
|
}
|
|
);
|
|
}
|
|
|
|
function buildOpenAIToolCallResponse({
|
|
model = "gpt-4o-mini",
|
|
toolName = "lookupWeather@1.0.0",
|
|
toolCallId = "call_weather",
|
|
argumentsObject = { location: "Sao Paulo" },
|
|
} = {}) {
|
|
return new Response(
|
|
JSON.stringify({
|
|
id: "chatcmpl_tool",
|
|
object: "chat.completion",
|
|
model,
|
|
choices: [
|
|
{
|
|
index: 0,
|
|
message: {
|
|
role: "assistant",
|
|
content: "",
|
|
tool_calls: [
|
|
{
|
|
id: toolCallId,
|
|
type: "function",
|
|
function: {
|
|
name: toolName,
|
|
arguments: JSON.stringify(argumentsObject),
|
|
},
|
|
},
|
|
],
|
|
},
|
|
finish_reason: "tool_calls",
|
|
},
|
|
],
|
|
usage: {
|
|
prompt_tokens: 6,
|
|
completion_tokens: 4,
|
|
total_tokens: 10,
|
|
},
|
|
}),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "application/json" },
|
|
}
|
|
);
|
|
}
|
|
|
|
function buildClaudeResponse(text = "ok", model = "claude-3-5-sonnet-20241022") {
|
|
return new Response(
|
|
JSON.stringify({
|
|
id: "msg_json",
|
|
type: "message",
|
|
role: "assistant",
|
|
model,
|
|
content: [{ type: "text", text }],
|
|
stop_reason: "end_turn",
|
|
usage: {
|
|
input_tokens: 10,
|
|
output_tokens: 4,
|
|
},
|
|
}),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "application/json" },
|
|
}
|
|
);
|
|
}
|
|
|
|
function buildClaudeStreamResponse(text = "streamed from claude", model = "claude-sonnet-4-6") {
|
|
return new Response(
|
|
[
|
|
"event: message_start",
|
|
`data: ${JSON.stringify({
|
|
type: "message_start",
|
|
message: {
|
|
id: "msg_stream",
|
|
type: "message",
|
|
role: "assistant",
|
|
model,
|
|
usage: { input_tokens: 12, output_tokens: 0 },
|
|
},
|
|
})}`,
|
|
"",
|
|
"event: content_block_start",
|
|
`data: ${JSON.stringify({
|
|
type: "content_block_start",
|
|
index: 0,
|
|
content_block: { type: "text", text: "" },
|
|
})}`,
|
|
"",
|
|
"event: content_block_delta",
|
|
`data: ${JSON.stringify({
|
|
type: "content_block_delta",
|
|
index: 0,
|
|
delta: { type: "text_delta", text },
|
|
})}`,
|
|
"",
|
|
"event: message_delta",
|
|
`data: ${JSON.stringify({
|
|
type: "message_delta",
|
|
delta: { stop_reason: "end_turn" },
|
|
usage: { output_tokens: 3 },
|
|
})}`,
|
|
"",
|
|
"event: message_stop",
|
|
`data: ${JSON.stringify({ type: "message_stop" })}`,
|
|
"",
|
|
].join("\n"),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "text/event-stream" },
|
|
}
|
|
);
|
|
}
|
|
|
|
function buildGeminiResponse(text = "ok", model = "gemini-2.5-flash") {
|
|
return new Response(
|
|
JSON.stringify({
|
|
responseId: "resp_gemini",
|
|
modelVersion: model,
|
|
createTime: "2026-04-05T12:00:00.000Z",
|
|
candidates: [
|
|
{
|
|
content: {
|
|
parts: [{ text }],
|
|
},
|
|
finishReason: "STOP",
|
|
},
|
|
],
|
|
usageMetadata: {
|
|
promptTokenCount: 5,
|
|
candidatesTokenCount: 7,
|
|
totalTokenCount: 12,
|
|
},
|
|
}),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "application/json" },
|
|
}
|
|
);
|
|
}
|
|
|
|
function buildOpenAIStreamResponse(text = "streamed from openai") {
|
|
return new Response(
|
|
[
|
|
`data: ${JSON.stringify({
|
|
id: "chatcmpl_stream",
|
|
object: "chat.completion.chunk",
|
|
choices: [{ index: 0, delta: { role: "assistant", content: text } }],
|
|
})}`,
|
|
"",
|
|
"data: [DONE]",
|
|
"",
|
|
].join("\n"),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "text/event-stream" },
|
|
}
|
|
);
|
|
}
|
|
|
|
function buildOpenAIResponsesSSE({
|
|
text = "responses streamed from codex",
|
|
model = "gpt-5.1-codex",
|
|
usage = null,
|
|
} = {}) {
|
|
return new Response(
|
|
[
|
|
`data: ${JSON.stringify({
|
|
type: "response.completed",
|
|
response: {
|
|
id: "resp_stream",
|
|
object: "response",
|
|
status: "completed",
|
|
model,
|
|
output: [
|
|
{
|
|
id: "msg_stream",
|
|
type: "message",
|
|
role: "assistant",
|
|
content: [{ type: "output_text", text, annotations: [] }],
|
|
},
|
|
],
|
|
usage: usage || {
|
|
input_tokens: 120,
|
|
output_tokens: 30,
|
|
prompt_tokens_details: {
|
|
cached_tokens: 40,
|
|
},
|
|
cache_creation_input_tokens: 11,
|
|
completion_tokens_details: {
|
|
reasoning_tokens: 13,
|
|
},
|
|
},
|
|
},
|
|
})}`,
|
|
"",
|
|
"data: [DONE]",
|
|
"",
|
|
].join("\n"),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "text/event-stream" },
|
|
}
|
|
);
|
|
}
|
|
|
|
function buildOpenAIResponsesJson({
|
|
text = "responses compacted from codex",
|
|
model = "gpt-5.5",
|
|
usage = null,
|
|
} = {}) {
|
|
return new Response(
|
|
JSON.stringify({
|
|
id: "resp_compact",
|
|
object: "response",
|
|
status: "completed",
|
|
model,
|
|
output: [
|
|
{
|
|
id: "msg_compact",
|
|
type: "message",
|
|
role: "assistant",
|
|
content: [{ type: "output_text", text, annotations: [] }],
|
|
},
|
|
],
|
|
output_text: text,
|
|
usage: usage || {
|
|
input_tokens: 90,
|
|
output_tokens: 15,
|
|
total_tokens: 105,
|
|
},
|
|
}),
|
|
{
|
|
status: 200,
|
|
headers: { "Content-Type": "application/json" },
|
|
}
|
|
);
|
|
}
|
|
|
|
async function resetStorage() {
|
|
globalThis.fetch = originalFetch;
|
|
process.env.REQUIRE_API_KEY = "false";
|
|
clearInflight();
|
|
resetAllCircuitBreakers();
|
|
apiKeysDb.resetApiKeyState();
|
|
readCacheDb.invalidateDbCache();
|
|
invalidateMemorySettingsCache();
|
|
await new Promise((resolve) => setTimeout(resolve, 20));
|
|
core.resetDbInstance();
|
|
fs.rmSync(TEST_DATA_DIR, { recursive: true, force: true });
|
|
fs.mkdirSync(TEST_DATA_DIR, { recursive: true });
|
|
initTranslators();
|
|
}
|
|
|
|
async function seedConnection(provider, overrides: SeedConnectionOverrides = {}) {
|
|
return providersDb.createProviderConnection({
|
|
provider,
|
|
authType: overrides.authType || "apikey",
|
|
name: overrides.name || `${provider}-primary`,
|
|
email: overrides.email,
|
|
apiKey: overrides.apiKey || `sk-${provider}-${Math.random().toString(16).slice(2, 10)}`,
|
|
accessToken: overrides.accessToken,
|
|
refreshToken: overrides.refreshToken,
|
|
tokenType: overrides.tokenType,
|
|
expiresAt: overrides.expiresAt,
|
|
tokenExpiresAt: overrides.tokenExpiresAt,
|
|
isActive: overrides.isActive ?? true,
|
|
testStatus: overrides.testStatus || "active",
|
|
priority: overrides.priority,
|
|
rateLimitedUntil: overrides.rateLimitedUntil,
|
|
providerSpecificData: overrides.providerSpecificData || {},
|
|
});
|
|
}
|
|
|
|
async function seedApiKey({
|
|
name = "chat-pipeline-key",
|
|
noLog = false,
|
|
allowedConnections,
|
|
allowedModels,
|
|
}: SeedApiKeyOptions = {}) {
|
|
const key = await apiKeysDb.createApiKey(name, "machine-test");
|
|
const updates: Record<string, unknown> = {};
|
|
if (noLog) updates.noLog = true;
|
|
if (allowedConnections) updates.allowedConnections = allowedConnections;
|
|
if (allowedModels) updates.allowedModels = allowedModels;
|
|
if (Object.keys(updates).length > 0) {
|
|
await apiKeysDb.updateApiKeyPermissions(key.id, updates);
|
|
}
|
|
return key;
|
|
}
|
|
|
|
function ensureLegacyMemoryTable() {
|
|
const db = core.getDbInstance();
|
|
db.exec(`
|
|
CREATE TABLE IF NOT EXISTS memory (
|
|
id TEXT PRIMARY KEY,
|
|
apiKeyId TEXT NOT NULL,
|
|
sessionId TEXT,
|
|
type TEXT NOT NULL,
|
|
key TEXT,
|
|
content TEXT NOT NULL,
|
|
metadata TEXT,
|
|
createdAt TEXT NOT NULL,
|
|
updatedAt TEXT NOT NULL,
|
|
expiresAt TEXT
|
|
)
|
|
`);
|
|
}
|
|
|
|
function insertLegacyMemory(apiKeyId, content) {
|
|
const db = core.getDbInstance();
|
|
const now = new Date().toISOString();
|
|
const hasModernTable = Boolean(
|
|
db.prepare("SELECT name FROM sqlite_master WHERE type = 'table' AND name = 'memories'").get()
|
|
);
|
|
|
|
if (hasModernTable) {
|
|
db.prepare(
|
|
`
|
|
INSERT INTO memories (
|
|
id, api_key_id, session_id, type, key, content, metadata, created_at, updated_at, expires_at
|
|
) VALUES (?, ?, ?, ?, ?, ?, ?, ?, ?, ?)
|
|
`
|
|
).run(
|
|
`mem_${Math.random().toString(16).slice(2, 10)}`,
|
|
apiKeyId,
|
|
"",
|
|
"factual",
|
|
"pref",
|
|
content,
|
|
"{}",
|
|
now,
|
|
now,
|
|
null
|
|
);
|
|
return;
|
|
}
|
|
|
|
ensureLegacyMemoryTable();
|
|
db.prepare(
|
|
`
|
|
INSERT INTO memory (
|
|
id, apiKeyId, sessionId, type, key, content, metadata, createdAt, updatedAt, expiresAt
|
|
) VALUES (?, ?, ?, ?, ?, ?, ?, ?, ?, ?)
|
|
`
|
|
).run(
|
|
`mem_${Math.random().toString(16).slice(2, 10)}`,
|
|
apiKeyId,
|
|
"",
|
|
"factual",
|
|
"pref",
|
|
content,
|
|
"{}",
|
|
now,
|
|
now,
|
|
null
|
|
);
|
|
}
|
|
|
|
async function waitFor(fn, timeoutMs = 1500) {
|
|
const startedAt = Date.now();
|
|
while (Date.now() - startedAt < timeoutMs) {
|
|
const result = await fn();
|
|
if (result) return result;
|
|
await new Promise((resolve) => setTimeout(resolve, 25));
|
|
}
|
|
return null;
|
|
}
|
|
|
|
async function getLatestCallLog() {
|
|
const rows = await callLogsDb.getCallLogs({ limit: 5 });
|
|
if (!Array.isArray(rows) || rows.length === 0) return null;
|
|
return callLogsDb.getCallLogById(rows[0].id);
|
|
}
|
|
|
|
async function getResponsesCallLogs() {
|
|
const rows = await callLogsDb.getCallLogs({ limit: 200 });
|
|
if (!Array.isArray(rows) || rows.length === 0) return [];
|
|
return rows.filter((row) => row.path === "/v1/responses");
|
|
}
|
|
|
|
test.beforeEach(async () => {
|
|
BaseExecutor.RETRY_CONFIG.delayMs = 0;
|
|
await resetStorage();
|
|
});
|
|
|
|
test.afterEach(async () => {
|
|
BaseExecutor.RETRY_CONFIG.delayMs = originalRetryDelayMs;
|
|
setCliCompatProviders([]);
|
|
await resetStorage();
|
|
});
|
|
|
|
test.after(async () => {
|
|
BaseExecutor.RETRY_CONFIG.delayMs = originalRetryDelayMs;
|
|
globalThis.fetch = originalFetch;
|
|
clearInflight();
|
|
resetAllCircuitBreakers();
|
|
core.resetDbInstance();
|
|
fs.rmSync(TEST_DATA_DIR, { recursive: true, force: true });
|
|
});
|
|
|
|
test("chat pipeline handles OpenAI passthrough with valid API key auth", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-primary" });
|
|
const apiKey = await seedApiKey();
|
|
const fetchCalls: FetchCall[] = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
method: init.method || "GET",
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponse("OpenAI passthrough");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
authKey: apiKey.key,
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Hello OpenAI" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /\/chat\/completions$/);
|
|
assert.equal(fetchCalls[0].headers.Authorization, "Bearer sk-openai-primary");
|
|
assert.equal(fetchCalls[0].body.messages[0].content, "Hello OpenAI");
|
|
assert.equal(json.choices[0].message.content, "OpenAI passthrough");
|
|
});
|
|
|
|
test("chat pipeline persists Codex responses cache and reasoning tokens to call logs", async () => {
|
|
await seedConnection("codex", { apiKey: "sk-codex-primary" });
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponsesSSE();
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses",
|
|
body: {
|
|
model: "codex/gpt-5.1-codex",
|
|
stream: false,
|
|
input: "Persist cache + reasoning usage",
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
const callLog = await waitFor(() => getLatestCallLog());
|
|
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /\/responses$/);
|
|
assert.equal(fetchCalls[0].headers.Authorization, "Bearer sk-codex-primary");
|
|
assert.equal(json.object, "response");
|
|
assert.equal(json.output[0].type, "message");
|
|
assert.equal(json.output[0].content[0].text, "responses streamed from codex");
|
|
assert.equal(json.output_text, "responses streamed from codex");
|
|
assert.equal(json.usage.input_tokens_details.cached_tokens, 40);
|
|
assert.equal(json.usage.output_tokens_details.reasoning_tokens, 13);
|
|
|
|
assert.ok(callLog, "expected a call log row to be created");
|
|
assert.equal(callLog.provider, "codex");
|
|
assert.equal(callLog.path, "/v1/responses");
|
|
assert.equal(callLog.tokens.cacheRead, 40);
|
|
assert.equal(callLog.tokens.cacheWrite, 11);
|
|
assert.equal(callLog.tokens.reasoning, 13);
|
|
});
|
|
|
|
test("chat pipeline applies global Codex priority service tier inside combos", async () => {
|
|
await seedConnection("codex", { apiKey: "sk-codex-combo-priority" });
|
|
await settingsDb.updateSettings({
|
|
codexServiceTier: { enabled: true, tier: "priority" },
|
|
});
|
|
await combosDb.createCombo({
|
|
name: "codex-priority-combo",
|
|
strategy: "priority",
|
|
config: { maxRetries: 0, retryDelayMs: 0 },
|
|
models: ["codex/gpt-5.5"],
|
|
});
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponsesSSE({ text: "combo priority ok", model: "gpt-5.5" });
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "codex-priority-combo",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Use Codex combo priority" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /\/responses$/);
|
|
assert.equal(fetchCalls[0].headers.Authorization, "Bearer sk-codex-combo-priority");
|
|
assert.equal(fetchCalls[0].body.service_tier, "priority");
|
|
assert.equal(json.choices[0].message.content, "combo priority ok");
|
|
});
|
|
|
|
test("chat pipeline applies Codex CLI fingerprint to OAuth responses requests", async () => {
|
|
setCliCompatProviders(["codex"]);
|
|
await seedConnection("codex", {
|
|
apiKey: "unused-for-oauth",
|
|
authType: "oauth",
|
|
accessToken: "codex-oauth-token",
|
|
providerSpecificData: {
|
|
openaiStoreEnabled: false,
|
|
requestDefaults: { reasoningEffort: "high" },
|
|
codexInstallationId: "11111111-1111-4111-a111-111111111111",
|
|
},
|
|
});
|
|
|
|
const fetchCalls = [];
|
|
globalThis.fetch = async (url, init = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
bodyString: String(init.body || ""),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponsesSSE({ text: "fingerprint ok" });
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses",
|
|
body: {
|
|
model: "codex/gpt-5.5-low",
|
|
stream: false,
|
|
conversation_id: "conv_codex_fingerprint",
|
|
input: [
|
|
{
|
|
type: "message",
|
|
role: "user",
|
|
content: [{ type: "input_text", text: "Reply with fingerprint ok" }],
|
|
},
|
|
],
|
|
},
|
|
})
|
|
);
|
|
|
|
await response.json();
|
|
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
const call = fetchCalls[0];
|
|
assert.match(call.url, /chatgpt\.com\/backend-api\/codex\/responses$/);
|
|
assert.equal(call.headers.Authorization, "Bearer codex-oauth-token");
|
|
assert.equal(call.headers.Accept, "text/event-stream");
|
|
assert.equal(call.headers.Version, getCodexClientVersion());
|
|
assert.equal(call.headers["Openai-Beta"], "responses=experimental");
|
|
assert.equal(call.headers["X-Codex-Beta-Features"], "responses_websockets");
|
|
assert.equal(call.headers["User-Agent"], "codex-cli/0.132.0 (Windows 10.0.26200; x64)");
|
|
assert.equal(call.headers["x-codex-window-id"], "conv_codex_fingerprint:0");
|
|
assert.ok(call.headers["x-client-request-id"], "expected Codex request id header");
|
|
assert.ok(call.headers["x-codex-turn-metadata"], "expected Codex turn metadata header");
|
|
|
|
const headerOrder = Object.keys(call.headers);
|
|
assert.ok(headerOrder.indexOf("Content-Type") < headerOrder.indexOf("Authorization"));
|
|
assert.ok(headerOrder.indexOf("Authorization") < headerOrder.indexOf("Accept"));
|
|
assert.ok(headerOrder.indexOf("Accept") < headerOrder.indexOf("User-Agent"));
|
|
|
|
const bodyOrder = Object.keys(JSON.parse(call.bodyString));
|
|
assert.deepEqual(bodyOrder.slice(0, 7), [
|
|
"model",
|
|
"stream",
|
|
"input",
|
|
"instructions",
|
|
"store",
|
|
"reasoning",
|
|
"prompt_cache_key",
|
|
]);
|
|
assert.equal(call.body.model, "gpt-5.5");
|
|
assert.equal(call.body.store, false);
|
|
assert.equal(
|
|
call.body.client_metadata["x-codex-installation-id"],
|
|
"11111111-1111-4111-a111-111111111111"
|
|
);
|
|
});
|
|
|
|
test("chat pipeline strips previous_response_id from stateless Codex responses by default", async () => {
|
|
await seedConnection("codex", {
|
|
apiKey: "sk-codex-stateless-responses",
|
|
providerSpecificData: { openaiStoreEnabled: false },
|
|
});
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponsesSSE({ text: "stateless responses ok", model: "gpt-5.5" });
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses",
|
|
body: {
|
|
model: "codex/gpt-5.5",
|
|
stream: false,
|
|
previous_response_id: "resp_vs_code_prev",
|
|
input: [
|
|
{
|
|
type: "message",
|
|
role: "user",
|
|
content: [{ type: "input_text", text: "Second VS Code turn" }],
|
|
},
|
|
],
|
|
},
|
|
})
|
|
);
|
|
|
|
await response.json();
|
|
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /\/responses$/);
|
|
assert.equal(fetchCalls[0].body.previous_response_id, undefined);
|
|
assert.equal(fetchCalls[0].body.store, false);
|
|
});
|
|
|
|
test("chat pipeline preserve mode forwards previous_response_id for responses requests", async () => {
|
|
await settingsDb.updateSettings({ responsesPreviousResponseIdMode: "preserve" });
|
|
await seedConnection("codex", {
|
|
apiKey: "sk-codex-preserve-responses",
|
|
providerSpecificData: { openaiStoreEnabled: false },
|
|
});
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponsesSSE({ text: "preserve responses ok", model: "gpt-5.5" });
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses",
|
|
body: {
|
|
model: "codex/gpt-5.5",
|
|
stream: false,
|
|
previous_response_id: "resp_preserved_prev",
|
|
input: "Second stateful turn",
|
|
},
|
|
})
|
|
);
|
|
|
|
await response.json();
|
|
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.equal(fetchCalls[0].body.previous_response_id, "resp_preserved_prev");
|
|
});
|
|
|
|
test("chat pipeline treats Codex /responses/compact as non-streaming JSON", async () => {
|
|
await seedConnection("codex", { apiKey: "sk-codex-compact" });
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponsesJson();
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses/compact",
|
|
headers: { Accept: "text/event-stream" },
|
|
body: {
|
|
model: "codex/gpt-5.5",
|
|
input: "Compact this session",
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as { object?: string; output_text?: string };
|
|
const callLog = await waitFor(() => getLatestCallLog());
|
|
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /\/responses\/compact$/);
|
|
assert.equal(fetchCalls[0].headers.Accept, "application/json");
|
|
assert.equal(fetchCalls[0].body.stream, undefined);
|
|
assert.equal(fetchCalls[0].body.store, undefined);
|
|
assert.equal(json.object, "response");
|
|
assert.equal(json.output_text, "responses compacted from codex");
|
|
|
|
assert.ok(callLog, "expected a compact call log row to be created");
|
|
assert.equal(callLog.provider, "codex");
|
|
assert.equal(callLog.path, "/v1/responses/compact");
|
|
assert.equal(callLog.status, 200);
|
|
});
|
|
|
|
test("chat pipeline serves repeated /v1/responses requests as MISS then HIT and logs cache hits separately", async () => {
|
|
await seedConnection("codex", { apiKey: "sk-codex-cache-seq" });
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponsesSSE({
|
|
text: "cached semantic response",
|
|
usage: {
|
|
input_tokens: 21,
|
|
output_tokens: 7,
|
|
prompt_tokens_details: {
|
|
cached_tokens: 5,
|
|
},
|
|
cache_creation_input_tokens: 2,
|
|
completion_tokens_details: {
|
|
reasoning_tokens: 3,
|
|
},
|
|
},
|
|
});
|
|
};
|
|
|
|
const uniquePrompt = `semantic-cache-seq-${Math.random().toString(16).slice(2)}`;
|
|
const requestBody = {
|
|
model: "codex/gpt-5.3-codex",
|
|
stream: false,
|
|
temperature: 0,
|
|
input: [{ role: "user", content: [{ type: "input_text", text: uniquePrompt }] }],
|
|
};
|
|
|
|
const beforeCount = (await getResponsesCallLogs()).length;
|
|
|
|
const firstResponse = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses",
|
|
body: requestBody,
|
|
})
|
|
);
|
|
|
|
const secondResponse = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses",
|
|
body: requestBody,
|
|
})
|
|
);
|
|
|
|
const thirdResponse = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/responses",
|
|
body: requestBody,
|
|
})
|
|
);
|
|
|
|
await firstResponse.json();
|
|
await secondResponse.json();
|
|
await thirdResponse.json();
|
|
|
|
assert.equal(firstResponse.status, 200);
|
|
assert.equal(secondResponse.status, 200);
|
|
assert.equal(thirdResponse.status, 200);
|
|
|
|
assert.equal(firstResponse.headers.get("X-OmniRoute-Cache"), "MISS");
|
|
assert.equal(secondResponse.headers.get("X-OmniRoute-Cache"), "HIT");
|
|
assert.equal(thirdResponse.headers.get("X-OmniRoute-Cache"), "HIT");
|
|
|
|
assert.equal(fetchCalls.length, 1, "expected upstream to be called only once for MISS");
|
|
assert.match(fetchCalls[0].url, /\/responses$/);
|
|
|
|
const callLogs = await waitFor(async () => {
|
|
const rows = await getResponsesCallLogs();
|
|
return rows.length === beforeCount + 3 ? rows : null;
|
|
}, 2000);
|
|
|
|
assert.ok(callLogs, "expected /v1/responses call logs to be recorded");
|
|
assert.equal(callLogs.length, beforeCount + 3, "expected MISS plus two HIT call logs");
|
|
|
|
const newLogs = callLogs.slice(0, 3);
|
|
assert.equal(newLogs.filter((row) => row.cacheSource === "upstream").length, 1);
|
|
assert.equal(newLogs.filter((row) => row.cacheSource === "semantic").length, 2);
|
|
|
|
const callLog = await waitFor(() => getLatestCallLog());
|
|
assert.ok(callLog, "expected a call log row to exist");
|
|
assert.equal(callLog.path, "/v1/responses");
|
|
assert.equal(callLog.status, 200);
|
|
});
|
|
|
|
test("chat pipeline translates OpenAI requests to Claude and returns OpenAI-shaped responses", async () => {
|
|
await seedConnection("claude", { apiKey: "sk-claude-primary" });
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildClaudeResponse("Claude translated reply");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "claude/claude-3-5-sonnet-20241022",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Hello Claude" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /\?beta=true$/);
|
|
assert.equal(fetchCalls[0].headers["x-api-key"], "sk-claude-primary");
|
|
assert.equal(fetchCalls[0].body.messages[0].role, "user");
|
|
assert.equal(fetchCalls[0].body.messages[0].content[0].text, "Hello Claude");
|
|
assert.equal(json.object, "chat.completion");
|
|
assert.equal(json.choices[0].message.content, "Claude translated reply");
|
|
});
|
|
|
|
test("chat pipeline translates OpenAI requests to Gemini and returns OpenAI-shaped responses", async () => {
|
|
await seedConnection("gemini", { apiKey: "sk-gemini-primary" });
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildGeminiResponse("Gemini translated reply");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "gemini/gemini-2.5-flash",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Hello Gemini" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /generateContent$/);
|
|
assert.equal(fetchCalls[0].headers["x-goog-api-key"], "sk-gemini-primary");
|
|
assert.equal(fetchCalls[0].body.contents[0].role, "user");
|
|
assert.equal(fetchCalls[0].body.contents[0].parts[0].text, "Hello Gemini");
|
|
assert.equal(json.object, "chat.completion");
|
|
assert.equal(json.choices[0].message.content, "Gemini translated reply");
|
|
});
|
|
|
|
test("chat pipeline sends Gemini CLI OAuth requests with native Cloud Code transport", async () => {
|
|
setCliCompatProviders(["gemini-cli"]);
|
|
await seedConnection("gemini-cli", {
|
|
authType: "oauth",
|
|
apiKey: "unused-for-oauth",
|
|
accessToken: "gemini-cli-oauth-token",
|
|
providerSpecificData: { projectId: "stored-project" },
|
|
});
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
|
|
if (String(url).endsWith("loadCodeAssist")) {
|
|
return new Response(JSON.stringify({ cloudaicompanionProject: "fresh-project" }), {
|
|
status: 200,
|
|
headers: { "Content-Type": "application/json" },
|
|
});
|
|
}
|
|
|
|
return buildGeminiResponse("Gemini CLI translated reply", "gemini-3-flash-preview");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "gemini-cli/gemini-3-flash-preview",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Hello Gemini CLI" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 2);
|
|
|
|
const loadCodeAssistCall = fetchCalls[0];
|
|
assert.match(loadCodeAssistCall.url, /loadCodeAssist$/);
|
|
assert.equal(loadCodeAssistCall.headers.Authorization, "Bearer gemini-cli-oauth-token");
|
|
assert.equal(loadCodeAssistCall.body.metadata.ideType, "IDE_UNSPECIFIED");
|
|
|
|
const generateCall = fetchCalls[1];
|
|
assert.match(generateCall.url, /generateContent$/);
|
|
assert.equal(generateCall.headers.Authorization, "Bearer gemini-cli-oauth-token");
|
|
assert.equal(generateCall.headers.Accept, "application/json");
|
|
assert.match(
|
|
generateCall.headers["User-Agent"],
|
|
new RegExp(
|
|
`^GeminiCLI/${GEMINI_CLI_VERSION.replaceAll(".", "\\.")}/gemini-3-flash-preview .* google-api-nodejs-client/${GEMINI_CLI_GOOGLE_API_NODE_CLIENT_VERSION.replaceAll(".", "\\.")}$`
|
|
)
|
|
);
|
|
assert.match(generateCall.headers["X-Goog-Api-Client"], /^gl-node\/\d+\.\d+\.\d+$/);
|
|
assert.equal(generateCall.body.project, "fresh-project");
|
|
assert.equal(generateCall.body.model, "gemini-3-flash-preview");
|
|
assert.equal(generateCall.body.userAgent, undefined);
|
|
assert.equal(generateCall.body.requestId, undefined);
|
|
assert.equal(generateCall.body.user_prompt_id, generateCall.body.request.session_id);
|
|
const keys = Object.keys(generateCall.body).slice(0, 4);
|
|
assert.deepEqual(keys.sort(), ["model", "project", "request", "user_prompt_id"]);
|
|
assert.equal(generateCall.body.request.sessionId, undefined);
|
|
assert.match(generateCall.body.request.session_id, /^[0-9a-f-]{36}$/i);
|
|
assert.equal(generateCall.body.request.contents.at(-1).parts[0].text, "Hello Gemini CLI");
|
|
assert.equal(json.object, "chat.completion");
|
|
assert.equal(json.choices[0].message.content, "Gemini CLI translated reply");
|
|
});
|
|
|
|
test("chat pipeline translates Claude-format requests into OpenAI upstream and back to Claude", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-claude-route" });
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponse("OpenAI answered Claude client");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
url: "http://localhost/v1/messages",
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
max_tokens: 128,
|
|
system: [{ text: "Be brief" }],
|
|
messages: [{ role: "user", content: [{ type: "text", text: "Hello from Claude client" }] }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.match(fetchCalls[0].url, /\/chat\/completions$/);
|
|
assert.equal(fetchCalls[0].body.messages[0].role, "system");
|
|
assert.equal(fetchCalls[0].body.messages[0].content, "Be brief");
|
|
assert.equal(fetchCalls[0].body.messages[1].content, "Hello from Claude client");
|
|
assert.equal(json.type, "message");
|
|
assert.equal(json.role, "assistant");
|
|
assert.equal(json.content[0].text, "OpenAI answered Claude client");
|
|
});
|
|
|
|
test("chat pipeline converts Claude SSE streams into OpenAI SSE output", async () => {
|
|
await seedConnection("claude", { apiKey: "sk-claude-stream" });
|
|
|
|
globalThis.fetch = async () => buildClaudeStreamResponse("Streamed Claude chunk");
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "claude/claude-sonnet-4-6",
|
|
stream: true,
|
|
messages: [{ role: "user", content: "Stream this" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const raw = await response.text();
|
|
assert.equal(response.status, 200);
|
|
assert.equal(response.headers.get("Content-Type"), "text/event-stream");
|
|
assert.match(raw, /chat\.completion\.chunk/);
|
|
assert.match(raw, /Streamed Claude chunk/);
|
|
assert.match(raw, /\[DONE\]/);
|
|
});
|
|
|
|
test("chat pipeline rejects invalid API keys and malformed JSON bodies", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-invalid-key-path" });
|
|
|
|
const invalidKeyResponse = await handleChat(
|
|
buildRequest({
|
|
authKey: "does-not-exist",
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
messages: [{ role: "user", content: "Hello" }],
|
|
},
|
|
})
|
|
);
|
|
const invalidKeyJson = (await invalidKeyResponse.json()) as any;
|
|
|
|
const invalidJsonResponse = await handleChat(
|
|
new Request("http://localhost/v1/chat/completions", {
|
|
method: "POST",
|
|
headers: { "Content-Type": "application/json" },
|
|
body: "{bad-json",
|
|
})
|
|
);
|
|
const invalidJson = (await invalidJsonResponse.json()) as any;
|
|
|
|
assert.equal(invalidKeyResponse.status, 401);
|
|
assert.match(invalidKeyJson.error.message, /Invalid API key|Incorrect API key/i);
|
|
assert.equal(invalidJsonResponse.status, 400);
|
|
assert.match(invalidJson.error.message, /Invalid JSON body/i);
|
|
});
|
|
|
|
test("chat pipeline allows unauthenticated requests through to provider resolution when called directly (authz pipeline enforces REQUIRE_API_KEY at route level)", async () => {
|
|
process.env.REQUIRE_API_KEY = "true";
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Missing auth" }],
|
|
},
|
|
})
|
|
);
|
|
const json = (await response.json()) as any;
|
|
|
|
// handleChat does not enforce REQUIRE_API_KEY — that's the authz pipeline's job.
|
|
// Without provider credentials seeded, the request falls through to the "no credentials" path.
|
|
assert.equal(response.status, 400);
|
|
assert.match(json.error.message, /No credentials for provider/i);
|
|
});
|
|
|
|
test("chat pipeline returns 400 when the model field is omitted", async () => {
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
stream: false,
|
|
messages: [{ role: "user", content: "No model selected" }],
|
|
},
|
|
})
|
|
);
|
|
const json = (await response.json()) as any;
|
|
|
|
assert.equal(response.status, 400);
|
|
assert.match(json.error.message, /Missing model/i);
|
|
});
|
|
|
|
test("chat pipeline treats Accept text/event-stream as streaming mode and returns a session header", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-accept-stream" });
|
|
|
|
globalThis.fetch = async () => buildOpenAIStreamResponse("Accept header stream");
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
headers: { Accept: "application/json, text/event-stream" },
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
messages: [{ role: "user", content: "Stream via Accept" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const raw = await response.text();
|
|
assert.equal(response.status, 200);
|
|
assert.equal(response.headers.get("Content-Type"), "text/event-stream");
|
|
assert.ok(response.headers.get("X-OmniRoute-Session-Id"));
|
|
assert.match(raw, /Accept header stream/);
|
|
assert.match(raw, /\[DONE\]/);
|
|
});
|
|
|
|
test("chat pipeline supports local mode without Authorization on explicit combos", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-local-combo" });
|
|
await combosDb.createCombo({
|
|
name: "local-router",
|
|
strategy: "priority",
|
|
models: ["openai/gpt-4o-mini"],
|
|
});
|
|
const fetchCalls = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
});
|
|
return buildOpenAIResponse("Local combo route");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "local-router",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "No auth header here" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.equal(json.choices[0].message.content, "Local combo route");
|
|
});
|
|
|
|
test("chat pipeline honors noLog by redacting persisted call log payloads", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-no-log" });
|
|
const apiKey = await seedApiKey({ noLog: true });
|
|
|
|
globalThis.fetch = async () => buildOpenAIResponse("No-log reply");
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
authKey: apiKey.key,
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Do not persist payloads" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
assert.equal(response.status, 200);
|
|
|
|
const callLog = await waitFor(() => getLatestCallLog());
|
|
assert.ok(callLog, "expected a call log row to be created");
|
|
assert.equal(callLog.apiKeyId, apiKey.id);
|
|
assert.equal(callLog.requestBody, null);
|
|
assert.equal(callLog.responseBody, null);
|
|
assert.equal(callLog.artifactRelPath, null);
|
|
});
|
|
|
|
test("chat pipeline returns current no-credentials contract when no provider connection exists", async () => {
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Hello" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 400);
|
|
assert.match(json.error.message, /No credentials for provider: openai/);
|
|
});
|
|
|
|
test("chat pipeline surfaces upstream 500 responses as structured errors", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-500" });
|
|
|
|
globalThis.fetch = async () =>
|
|
new Response(JSON.stringify({ error: { message: "provider exploded" } }), {
|
|
status: 500,
|
|
headers: { "Content-Type": "application/json" },
|
|
});
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Trigger 500" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 500);
|
|
assert.match(json.error.message, /\[500\]: provider exploded/);
|
|
});
|
|
|
|
test("chat pipeline returns 429 with Retry-After when the upstream rate-limits the only account", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-429" });
|
|
await settingsDb.updateSettings({
|
|
requestRetry: 0,
|
|
maxRetryIntervalSec: 0,
|
|
});
|
|
let attempts = 0;
|
|
|
|
globalThis.fetch = async () => {
|
|
attempts += 1;
|
|
return new Response(
|
|
JSON.stringify({
|
|
error: {
|
|
message: "Rate limit exceeded. Your quota will reset after 30s.",
|
|
},
|
|
}),
|
|
{
|
|
status: 429,
|
|
headers: { "Content-Type": "application/json" },
|
|
}
|
|
);
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Trigger 429" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 429);
|
|
assert.ok(attempts >= 1, "expected at least one upstream attempt");
|
|
assert.ok(Number(response.headers.get("Retry-After")) >= 1);
|
|
assert.match(json.error.message, /\[openai\/gpt-4o-mini\]/);
|
|
});
|
|
|
|
test("chat pipeline keeps provider breaker closed for repeated connection-scoped 429s", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-429-breaker" });
|
|
await settingsDb.updateSettings({
|
|
requestRetry: 0,
|
|
maxRetryIntervalSec: 0,
|
|
});
|
|
|
|
globalThis.fetch = async () =>
|
|
new Response(
|
|
JSON.stringify({
|
|
error: {
|
|
message: "Rate limit exceeded. Your quota will reset after 30s.",
|
|
},
|
|
}),
|
|
{
|
|
status: 429,
|
|
headers: { "Content-Type": "application/json" },
|
|
}
|
|
);
|
|
|
|
for (let i = 0; i < 3; i += 1) {
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: `Trigger 429 #${i + 1}` }],
|
|
},
|
|
})
|
|
);
|
|
assert.equal(response.status, 429);
|
|
}
|
|
|
|
const breaker = getCircuitBreaker("openai");
|
|
const status = breaker.getStatus();
|
|
|
|
assert.equal(status.state, "CLOSED");
|
|
assert.equal(status.failureCount, 0);
|
|
});
|
|
|
|
test("chat pipeline maps upstream timeouts to 504 responses", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-timeout" });
|
|
|
|
globalThis.fetch = async () => {
|
|
const error = new Error("upstream timed out");
|
|
error.name = "TimeoutError";
|
|
throw error;
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Trigger timeout" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 504);
|
|
assert.match(json.error.message, /\[504\]: upstream timed out/);
|
|
});
|
|
|
|
test("chat pipeline injects memory context before sending the upstream request", async () => {
|
|
// Reset provider failure state to avoid circuit breaker interference
|
|
clearProviderFailure("openai");
|
|
await seedConnection("openai", { apiKey: "sk-openai-memory" });
|
|
const apiKey = await seedApiKey();
|
|
await settingsDb.updateSettings({
|
|
memoryEnabled: true,
|
|
memoryMaxTokens: 400,
|
|
memoryRetentionDays: 30,
|
|
memoryStrategy: "recent",
|
|
});
|
|
invalidateMemorySettingsCache();
|
|
insertLegacyMemory(apiKey.id, "User prefers concise answers.");
|
|
|
|
const fetchCalls = [];
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIResponse("Memory-aware reply");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
authKey: apiKey.key,
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Summarize my preference" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.equal(fetchCalls[0].body.messages[0].role, "system");
|
|
assert.match(fetchCalls[0].body.messages[0].content, /User prefers concise answers/);
|
|
assert.equal(json.choices[0].message.content, "Memory-aware reply");
|
|
});
|
|
|
|
test("chat pipeline injects skills into tools and intercepts tool calls with skill output", async () => {
|
|
// Reset provider failure state to avoid circuit breaker interference
|
|
clearProviderFailure("openai");
|
|
await seedConnection("openai", { apiKey: "sk-openai-skills" });
|
|
const apiKey = await seedApiKey();
|
|
await settingsDb.updateSettings({ skillsEnabled: true });
|
|
invalidateMemorySettingsCache();
|
|
|
|
const handlerName = `weather-handler-${Date.now()}`;
|
|
skillExecutor.registerHandler(handlerName, async (input) => ({
|
|
forecast: `Sunny in ${input.location}`,
|
|
}));
|
|
|
|
await skillRegistry.register({
|
|
apiKeyId: apiKey.id,
|
|
name: "lookupWeather",
|
|
version: "1.0.0",
|
|
description: "Return a canned forecast",
|
|
schema: {
|
|
input: {
|
|
type: "object",
|
|
properties: {
|
|
location: { type: "string" },
|
|
},
|
|
},
|
|
output: {
|
|
type: "object",
|
|
},
|
|
},
|
|
handler: handlerName,
|
|
enabled: true,
|
|
});
|
|
|
|
const fetchCalls = [];
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
fetchCalls.push({
|
|
url: String(url),
|
|
body: init.body ? JSON.parse(String(init.body)) : null,
|
|
});
|
|
return buildOpenAIToolCallResponse();
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
authKey: apiKey.key,
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Check the weather" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(fetchCalls.length, 1);
|
|
assert.ok(Array.isArray(fetchCalls[0].body.tools));
|
|
assert.equal(fetchCalls[0].body.tools[0].function.name, "lookupWeather@1.0.0");
|
|
assert.equal(json.choices[0].finish_reason, "tool_calls");
|
|
assert.equal(json.tool_results[0].tool_call_id, "call_weather");
|
|
assert.equal(JSON.parse(json.tool_results[0].output).forecast, "Sunny in Sao Paulo");
|
|
});
|
|
|
|
test("chat pipeline falls back to the next account after a provider failure", async () => {
|
|
// Reset provider failure state to avoid circuit breaker interference
|
|
clearProviderFailure("openai");
|
|
await seedConnection("openai", {
|
|
name: "openai-primary",
|
|
apiKey: "sk-openai-primary-fallback",
|
|
priority: 1,
|
|
});
|
|
await seedConnection("openai", {
|
|
name: "openai-secondary",
|
|
apiKey: "sk-openai-secondary-fallback",
|
|
priority: 2,
|
|
});
|
|
const seenAuthHeaders = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
const headers = toPlainHeaders(init.headers);
|
|
seenAuthHeaders.push(headers.Authorization);
|
|
if (seenAuthHeaders.length === 1) {
|
|
return new Response(JSON.stringify({ error: { message: "first account failed" } }), {
|
|
status: 500,
|
|
headers: { "Content-Type": "application/json" },
|
|
});
|
|
}
|
|
return buildOpenAIResponse("Second account succeeded");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Use account fallback" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.deepEqual(seenAuthHeaders, [
|
|
"Bearer sk-openai-primary-fallback",
|
|
"Bearer sk-openai-secondary-fallback",
|
|
]);
|
|
assert.equal(json.choices[0].message.content, "Second account succeeded");
|
|
});
|
|
|
|
test("chat pipeline falls back across combo models when the first provider fails", async () => {
|
|
// Reset provider failure state to avoid circuit breaker interference
|
|
clearProviderFailure("openai");
|
|
clearProviderFailure("claude");
|
|
await seedConnection("openai", { apiKey: "sk-openai-combo-fail" });
|
|
await seedConnection("claude", { apiKey: "sk-claude-combo-fail" });
|
|
await combosDb.createCombo({
|
|
name: "combo-fallback",
|
|
strategy: "priority",
|
|
config: { maxRetries: 0, retryDelayMs: 0 },
|
|
models: ["openai/gpt-4o-mini", "claude/claude-3-5-sonnet-20241022"],
|
|
});
|
|
const attempts = [];
|
|
|
|
globalThis.fetch = async (url, init: RequestInit = {}) => {
|
|
const call = {
|
|
url: String(url),
|
|
headers: toPlainHeaders(init.headers),
|
|
};
|
|
attempts.push(call);
|
|
if (attempts.length === 1) {
|
|
return new Response(JSON.stringify({ error: { message: "openai combo miss" } }), {
|
|
status: 503,
|
|
headers: { "Content-Type": "application/json" },
|
|
});
|
|
}
|
|
return buildClaudeResponse("Claude combo fallback");
|
|
};
|
|
|
|
const response = await handleChat(
|
|
buildRequest({
|
|
body: {
|
|
model: "combo-fallback",
|
|
stream: false,
|
|
messages: [{ role: "user", content: "Use combo fallback" }],
|
|
},
|
|
})
|
|
);
|
|
|
|
const json = (await response.json()) as any;
|
|
assert.equal(response.status, 200);
|
|
assert.equal(attempts.length, 2);
|
|
assert.match(attempts[0].url, /\/chat\/completions$/);
|
|
assert.match(attempts[1].url, /\?beta=true$/);
|
|
assert.equal(json.choices[0].message.content, "Claude combo fallback");
|
|
});
|
|
|
|
test("chat pipeline deduplicates concurrent identical non-stream requests", async () => {
|
|
// Reset provider failure state to avoid circuit breaker interference
|
|
clearProviderFailure("openai");
|
|
await seedConnection("openai", { apiKey: "sk-openai-dedup" });
|
|
let fetchCount = 0;
|
|
|
|
globalThis.fetch = async () => {
|
|
fetchCount += 1;
|
|
await new Promise((resolve) => setTimeout(resolve, 25));
|
|
return buildOpenAIResponse("Deduplicated response");
|
|
};
|
|
|
|
const requestA = buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
temperature: 0,
|
|
messages: [{ role: "user", content: "Deduplicate this request" }],
|
|
},
|
|
});
|
|
const requestB = buildRequest({
|
|
body: {
|
|
model: "openai/gpt-4o-mini",
|
|
stream: false,
|
|
temperature: 0,
|
|
messages: [{ role: "user", content: "Deduplicate this request" }],
|
|
},
|
|
});
|
|
|
|
const [responseA, responseB] = await Promise.all([handleChat(requestA), handleChat(requestB)]);
|
|
const [jsonA, jsonB] = await Promise.all([responseA.json(), responseB.json()]);
|
|
|
|
assert.equal(responseA.status, 200);
|
|
assert.equal(responseB.status, 200);
|
|
assert.equal(fetchCount, 1);
|
|
assert.equal(jsonA.choices[0].message.content, "Deduplicated response");
|
|
assert.equal(jsonB.choices[0].message.content, "Deduplicated response");
|
|
});
|