mirror of
https://github.com/diegosouzapw/OmniRoute.git
synced 2026-08-04 22:32:12 +03:00
* chore(release): open v3.8.39 development cycle * docs(changelog): backfill 5 v3.8.38 bullets merged after release finalize These PRs squash-merged into release/v3.8.38 between the CHANGELOG finalize (ff57be32f) and the merge-to-main (ae6e2342d), so they shipped in the v3.8.38 tag but had no bullet: - feat(compression): Ionizer engine (lossy JSON-array sampling + CCR) (#5148) - fix(sse): preserve non-stream reasoning fields (#5155, @rdself) - fix(i18n): add missing English UI labels (#5153, @rdself) - test(combo): gated live smoke (#5151) + release-expectations refresh (#5150, @KooshaPari) (#5129 exact-host Anthropic baseUrl is already covered by the #5130 bullet — same CodeQL #674.) Synced 41 i18n CHANGELOG mirrors. * feat(compression): TOON best-of-N candidate encoder + encoder A/B table (#5163) Integrated into release/v3.8.39. TOON best-of-N candidate encoder (GCF default, fail-open). 17/17 unit tests pass on merge result; CI reds were base-stale + Quality Ratchet DRIFT. * fix(zenmux): normalize vendor-prefixed GLM system roles (#5158) Integrated into release/v3.8.39. ZenMux vendor-prefixed GLM system-role normalization; 12/12 role-normalizer tests pass on merge result. CI reds base-stale. * [codex] fix xAI OAuth test and reasoning effort (#5157) Integrated into release/v3.8.39. xAI reasoning-effort normalization (max/xhigh→high) + OAuth test config; 46/46 xai-translator tests pass on merge result. CI reds base-stale. * docs(i18n): add Traditional Chinese (zh-TW) README and update zh-CN to latest (#5162) Integrated into release/v3.8.39. Traditional Chinese (zh-TW) README + zh-CN refresh; docs-only. * test(security): guard PII redaction stays opt-in (default off) + Hard Rule #20 (#5159) Integrated into release/v3.8.39. PII opt-in regression guard + Hard Rule #20; rebased to strip base-drift (+81/-1). 5/5 guard tests pass; flip-proof verified. * test(combo): deterministic context-relay universal-handoff coverage (closes phase-2 TODO) (#5168) Integrated into release/v3.8.39. Deterministic context-relay universal-handoff coverage (3 tests); 3/3 pass on merge result. * docs(i18n): full sync zh-TW and zh-CN README with canonical English v3.8.39 (#5171) Integrated into release/v3.8.39. Full zh-TW docs tree + zh-CN sync with canonical English v3.8.39; docs-only. * fix(serve): honour HOSTNAME from .env instead of hardcoding 0.0.0.0 (#5134) (#5170) Integrated into release/v3.8.39. HOSTNAME env override in serve (#5134) + regression test (4/4, TDD flip-proof verified). * fix(sse): resolve nameless deepseek-web tool blocks via parameter-schema match (#5154) (#5173) Integrated into release/v3.8.39. Schema-based nameless deepseek-web tool-block resolution (#5154); 6/6 tests pass on merge result (incl. ambiguous/no-match negatives + named-tag no-regression). * fix(sse): normalize array user content for Command Code to avoid upstream 400 (#5166) (#5174) Integrated into release/v3.8.39. Normalize array user content for Command Code (#5166, user-array/400 symptom); 4/4 tests pass on merge result. * fix(sse): defer </think> close so it never leaks before tool_calls (#5123) (#5175) Integrated into release/v3.8.39. Defer </think> close so it never leaks before tool_calls (#5123); 4/4 tests pass (incl. #4633 no-regression). CHANGELOG synced to keep all 3 v3.8.39 fixes. * fix(dashboard): use amber for home update-step warning icon (#5176) Integrated into release/v3.8.39. Amber for home update-step warning icon; 1/1 UI test. * fix(api): LAN/Tailscale dashboard — host-aware CSP + GET-exempt version route + combo field errors (#5083) (#5177) Integrated into release/v3.8.39. Host-aware CSP (ReDoS/injection-safe host validation) + GET-exempt /api/system/version (POST/spawn stays LOCAL_ONLY, exact-match safe-methods-only) + COMBO_002 firstField. 44/44 tests + route-guard membership gate green. CHANGELOG synced to keep all 4 v3.8.39 fixes. * fix(api): replace #5083 global middleware CSP with declarative ws: scheme (#5083) Follow-up to PR #5177 (merged): that version implemented the LAN-CSP fix (Bug 1) with a new global `src/middleware.ts` + `src/server/csp.ts`, which contradicts the project's documented architecture — 'No global Next.js middleware — interception is route-specific' (CLAUDE.md / AGENTS.md) — and was merged unverified (middleware vs next.config header precedence was never confirmed in a real build). This replaces that approach with the minimal, declarative equivalent: • next.config.mjs: connect-src now permits the bare `ws:` scheme (symmetric with the bare `wss:` already allowed) so the dashboard can reach its own Live WS server from a LAN/Tailscale host. No middleware. • Removes src/middleware.ts, src/server/csp.ts, and tests/unit/csp-host-aware.test.ts. • Adds tests/unit/csp-lan-ws-5083.test.ts (incl. a guard asserting src/middleware.ts does NOT exist, so the global-middleware approach cannot silently return). Bugs 2 (GET-exempt /api/system/version) and 3 (COMBO_002 field surfacing) from #5177 are unaffected and remain in place. Co-authored-by: KooshaPari <KooshaPari@users.noreply.github.com> * test(combo): end-to-end quota-share DRR routing-decision coverage (matrix parity) (#5179) Integrated into release/v3.8.39. Quota-share DRR routing-decision coverage (matrix parity); 2/2 pass on merge result. * feat(agent-bridge): graceful cert-install fallback with manual guide for containers (#4546) (#5178) Integrated into release/v3.8.39. Agent-bridge graceful cert-install fallback + manual guide (#4546); 6/6 tests pass on merge result. * fix(antigravity): family-scoped quota lockout (gemini/claude buckets) (#5180) Integrated into release/v3.8.39 — family-scoped antigravity quota lockout. Rebased from v3.8.37 + validated (vitest 5/5, typecheck clean, full combo-matrix green, model-lockout 99/0). Same-model cross-account retry (chat.ts) deferred pending live antigravity VPS validation. * fix(cli): force NODE_ENV to match dev/start run mode in custom Next server (#5189) Integrated into release/v3.8.39. Force NODE_ENV to match dev/start run mode in custom Next server; 2/2 source-scan+ordering tests pass on merge result. * feat(compression): CCR ranged/grep/stats retrieval (ReDoS-safe, backward-compat) (#5187) Integrated into release/v3.8.39. CCR ranged/grep/stats retrieval (safe-regex ReDoS guard + length/match caps); 17/17 tests pass on merge result. * docs(combo): sync all combo/routing-strategy docs to current state + document test coverage (#5185) Integrated into release/v3.8.39. Combo/routing-strategy docs sync; docs-only. * fix(mcp): return 404 (not 400) for unknown Streamable HTTP session id (#5169) (#5191) * fix(api): respect blocked Auto (Zero-Config) provider in /v1/models catalog (#5192) (#5194) * test(combo): deterministic context-relay codex quota-handoff coverage (closes last gap) (#5195) * test(ci): wire antigravity-quota-family under test:vitest (fix test-discovery orphan) (#5196) * fix(oauth): antigravity login no longer hangs — fire-and-forget onboarding + bounded post-exchange (#5193) Antigravity OAuth hang fix (no-PKCE/no-openid + bounded post-exchange + exchange-500 fix). Includes #5200 (Koosha) revert + owner rebaseline to keep documented comments. Integrated into release/v3.8.39. * feat(oauth): remote Antigravity login via local helper + paste-credentials (#5203) Remote Antigravity login: local helper (omniroute login antigravity) + paste-credentials. Integrated into release/v3.8.39. * fix(translator): accept Claude Messages shape in non-stream malformed-200 guard (#5156) Integrated into release/v3.8.39 * fix(cli): default dev bundler to Turbopack (16.2.x panic no longer reproduces) (#5206) Integrated into release/v3.8.39 * fix(cli): auto-calibrate server V8 heap from physical RAM (#5172) (#5213) The server was spawned with a fixed --max-old-space-size=512 (omniroute serve) or no heap flag at all (Electron), so RAM-rich boxes still OOM-crashed under load (Ineffective mark-compacts near heap limit ~500MB) with many providers/ accounts and large model catalogs. New calibrateHeapFallbackMb(os.totalmem()) defaults the heap to ~35% of RAM clamped [512,4096], wired into serve.mjs and electron/main.js. Explicit OMNIROUTE_MEMORY_MB still wins (#2939 unchanged). Also addresses #5160 (same OOM root); #5152 (docker) benefits via the same knob. Closes #5172 * fix(proxy): coalesce fast-fail health probes (#5208) Integrated into release/v3.8.39 * fix(proxy): close dispatchers when clearing cache (#5202) Integrated into release/v3.8.39 * fix(cli): raise dev server Node heap limit to 8GB to prevent OOM (#5198) Integrated into release/v3.8.39 * fix(auth): allow synthetic no-auth fallback for mimocode (#5205) Integrated into release/v3.8.39 * fix(oauth): preserve Antigravity refresh_token on empty/omitted upstream response (#3850) (#5214) Google's OAuth refresh tokens are non-rotating: the refresh response usually omits refresh_token and occasionally returns it as an empty string. The Antigravity executor used `typeof tokens.refresh_token === "string" ? ... ` which accepts "" (typeof "" === "string") and overwrote the stored token with empty, nulling it on first refresh. Now treats non-string OR empty as absent and preserves credentials.refreshToken, matching refreshGoogleToken semantics. Closes #3850 * fix(responses): normalize non-array input (#5204) Integrated into release/v3.8.39 * fix(stream): normalize safety finish reasons via shared helper (#5197) Integrated into release/v3.8.39 * fix(request-logger): never render negative '(-100%)' compression badge (#5201) Integrated into release/v3.8.39 * fix(combo): reject empty responses api output (#5207) Integrated into release/v3.8.39 — combo failover now rejects empty Responses API output (validateQuality). Baseline rebaseline dropped (main-measured drift; maintainer rebaselines at release). * fix(pwa): prefer cached navigation before offline page (#5209) Integrated into release/v3.8.39 — PWA service worker prefers cached navigation before offline page (#5165). * chore(release): v3.8.39 — 2026-06-28 * chore(release): rebaseline openapi+i18n coverage ratchet drift for v3.8.39 --------- Co-authored-by: Arthur Bodera <abodera@gmail.com> Co-authored-by: Nguyen Minh <lop123thcs@gmail.com> Co-authored-by: lunkerchen <labanchen@gmail.com> Co-authored-by: Ankit <177378174+anki1kr@users.noreply.github.com> Co-authored-by: KooshaPari <KooshaPari@users.noreply.github.com> Co-authored-by: Ardem2025 <ardemb22@gmail.com> Co-authored-by: backryun <bakryun0718@proton.me> Co-authored-by: Anton <39598727+NomenAK@users.noreply.github.com> Co-authored-by: KooshaPari <42529354+KooshaPari@users.noreply.github.com> Co-authored-by: Wilson <pedbookmed@gmail.com> Co-authored-by: Randi <55005611+rdself@users.noreply.github.com>
283 lines
11 KiB
TypeScript
283 lines
11 KiB
TypeScript
// tests/integration/combo-matrix/context-relay-handoff.test.ts
|
|
//
|
|
// Deterministic in-process tests for context-relay UNIVERSAL HANDOFF behavior:
|
|
// the session-context transfer that fires when a model switch is detected,
|
|
// regardless of provider (provider-agnostic). This is the distinguishing
|
|
// behavior of context-relay that was deferred as TODO(phase-2) in the
|
|
// original combo-matrix coverage.
|
|
//
|
|
// What is tested here:
|
|
// 1. Universal handoff fires (extra summary dispatch + DB record) when:
|
|
// - universalHandoffConfig.enabled = true (default)
|
|
// - request carries x-omniroute-session-id header
|
|
// - session_model_history records a DIFFERENT prior model than the combo target
|
|
// 2. Control (no model switch): prevModel === currModel — handoff must NOT fire.
|
|
// 3. Control (no session ID): no header passed — handoff must NOT fire.
|
|
//
|
|
// Codex-specific block (lines 2143-2183 in combo.ts):
|
|
// Requires strategy === "context-relay" AND provider === "codex" AND a live
|
|
// codex session connection + quota fetcher. The seams (getSessionConnection /
|
|
// fetchCodexQuota) are NOT exported from contextHandoff.ts as testable hooks;
|
|
// there is no in-process mechanism equivalent to registerQuotaFetcher for the
|
|
// codex path. This block requires a real codex provider connection (VPS) and
|
|
// is NOT covered here. The universal-handoff path above is the user-visible,
|
|
// provider-agnostic behavior and is now fully covered.
|
|
|
|
import test from "node:test";
|
|
import assert from "node:assert/strict";
|
|
import { createComboRoutingHarness, providerFromUrl } from "../_comboRoutingHarness.ts";
|
|
|
|
const h = await createComboRoutingHarness("combo-relay-handoff");
|
|
const {
|
|
BaseExecutor,
|
|
combosDb,
|
|
handleChat,
|
|
buildRequest,
|
|
seedConnection,
|
|
resetStorage,
|
|
buildOpenAIResponse,
|
|
waitFor,
|
|
toPlainHeaders,
|
|
} = h;
|
|
|
|
// Import DB helpers AFTER harness so they share the same DB instance (DATA_DIR
|
|
// is set by the harness before any import triggers DB init).
|
|
const { recordSessionModelUsage, getHandoff } = await import(
|
|
"../../../src/lib/db/contextHandoffs.ts"
|
|
);
|
|
|
|
// A minimal but valid handoff-JSON blob that parseHandoffJSON will accept.
|
|
// Must have at minimum a non-empty "summary" field.
|
|
const SCRIPTED_SUMMARY_JSON = JSON.stringify({
|
|
summary: "User initiated a context-relay handoff test and requested a model switch.",
|
|
keyDecisions: ["switched from gemini to openai"],
|
|
taskProgress: "in-progress",
|
|
activeEntities: ["context-relay-test"],
|
|
});
|
|
|
|
const COMBO_NAME = "m-relay-handoff";
|
|
// The session ID must use the external-session format: extractExternalSessionId in chat.ts
|
|
// reads "x-session-id" (not "x-omniroute-session-id") and prefixes the result with "ext:".
|
|
// We must seed session_model_history with this exact "ext:"-prefixed ID so that
|
|
// getLastSessionModel(relayOptions.sessionId, comboName) returns the prior model.
|
|
const SESSION_HEADER_VALUE = "relay-handoff-session-001";
|
|
const SESSION_ID = `ext:${SESSION_HEADER_VALUE}`; // matches what extractExternalSessionId produces
|
|
const PREV_MODEL = "gemini/gemini-2.5-flash"; // the "old" model — must differ from combo target
|
|
const CURR_MODEL = "openai/gpt-4o-mini"; // the combo target model (provider/model format)
|
|
|
|
// Build a Request with the session-id header.
|
|
function relayRequest(withSessionId = true) {
|
|
const headers: Record<string, string> = withSessionId
|
|
? { "x-session-id": SESSION_HEADER_VALUE } // x-session-id → relayOptions.sessionId = "ext:..."
|
|
: {};
|
|
return buildRequest({
|
|
headers,
|
|
body: {
|
|
model: COMBO_NAME,
|
|
stream: false,
|
|
messages: [
|
|
{ role: "user", content: "Prior turn — building something important in the test session." },
|
|
],
|
|
},
|
|
});
|
|
}
|
|
|
|
// Install a recording fetch that:
|
|
// • returns a valid handoff JSON (wrapped in an OpenAI completion) for the
|
|
// internal summary request (identified by _omnirouteInternalRequest flag in body)
|
|
// • returns a normal OpenAI response for every other call
|
|
function installHandoffAwareFetch() {
|
|
h.calls.length = 0;
|
|
|
|
globalThis.fetch = async (url: any, init: any = {}) => {
|
|
const u = String(url);
|
|
const provider = providerFromUrl(u);
|
|
const headers = toPlainHeaders(init?.headers);
|
|
|
|
// Parse request body to detect the internal summary request
|
|
let bodyObj: Record<string, unknown> = {};
|
|
try {
|
|
bodyObj =
|
|
typeof init?.body === "string"
|
|
? (JSON.parse(init.body) as Record<string, unknown>)
|
|
: ((init?.body ?? {}) as Record<string, unknown>);
|
|
} catch {
|
|
bodyObj = {};
|
|
}
|
|
|
|
const call = {
|
|
index: h.calls.length,
|
|
provider,
|
|
url: u,
|
|
authorization: headers.authorization,
|
|
model: typeof bodyObj.model === "string" ? bodyObj.model : undefined,
|
|
};
|
|
h.calls.push(call);
|
|
|
|
// Return valid handoff JSON for the internal summary generation request
|
|
if (bodyObj._omnirouteInternalRequest === "universal-handoff") {
|
|
return buildOpenAIResponse(SCRIPTED_SUMMARY_JSON);
|
|
}
|
|
|
|
// Normal success response for the real request
|
|
return buildOpenAIResponse("assistant reply ok");
|
|
};
|
|
}
|
|
|
|
test.beforeEach(async () => {
|
|
BaseExecutor.RETRY_CONFIG.delayMs = 0;
|
|
await resetStorage();
|
|
});
|
|
test.afterEach(async () => {
|
|
BaseExecutor.RETRY_CONFIG.delayMs = h.originalRetryDelayMs;
|
|
await resetStorage();
|
|
});
|
|
test.after(async () => {
|
|
await h.cleanup();
|
|
});
|
|
|
|
// ── Test 1: universal handoff fires on model switch ────────────────────────────
|
|
//
|
|
// Seed session history so getLastSessionModel returns PREV_MODEL (gemini).
|
|
// Combo target is CURR_MODEL (openai). The switch is detected and
|
|
// maybeGenerateUniversalHandoff fires via setImmediate → calls handleSingleModel
|
|
// (extra upstream dispatch) → parses response → upsertHandoff writes to DB.
|
|
//
|
|
// Observable 1: h.calls has ≥2 entries (main request + summary dispatch).
|
|
// Observable 2: getHandoff(sessionId, comboName) returns a record with a
|
|
// non-empty .summary (proves the FULL path executed, not just the
|
|
// dispatch). The assertion FAILS if the handoff did not fire.
|
|
test("context-relay universal handoff: fires and writes handoff record on model switch", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-handoff" });
|
|
await seedConnection("gemini", { apiKey: "sk-gemini-handoff" });
|
|
|
|
await combosDb.createCombo({
|
|
name: COMBO_NAME,
|
|
strategy: "context-relay",
|
|
config: { maxRetries: 0, retryDelayMs: 0, stickyRoundRobinLimit: 1 },
|
|
models: [
|
|
// Single target — openai/gpt-4o-mini — so modelStr = CURR_MODEL
|
|
{ id: "rh-openai", kind: "model", providerId: "openai", model: "gpt-4o-mini" },
|
|
],
|
|
});
|
|
|
|
// Seed the prior model usage so getLastSessionModel returns PREV_MODEL.
|
|
// universalHandoffConfig.enabled = true by DEFAULT_UNIVERSAL_HANDOFF_CONFIG.
|
|
recordSessionModelUsage(SESSION_ID, COMBO_NAME, PREV_MODEL, "gemini", undefined);
|
|
|
|
installHandoffAwareFetch();
|
|
|
|
const r = await handleChat(relayRequest(/* withSessionId */ true));
|
|
assert.equal(r.status, 200, "main request must succeed");
|
|
|
|
// Wait for the setImmediate + generateUniversalHandoffAsync to complete and
|
|
// write the DB record. Poll for up to 2 s — typically resolves in <100 ms.
|
|
const handoff = await waitFor(
|
|
() => getHandoff(SESSION_ID, COMBO_NAME),
|
|
2000
|
|
);
|
|
|
|
assert.ok(
|
|
handoff !== null,
|
|
"universal handoff record must be written to DB when a model switch is detected"
|
|
);
|
|
assert.ok(
|
|
typeof handoff!.summary === "string" && handoff!.summary.length > 0,
|
|
`handoff.summary must be non-empty; got: ${JSON.stringify(handoff!.summary)}`
|
|
);
|
|
assert.equal(
|
|
handoff!.comboName,
|
|
COMBO_NAME,
|
|
"handoff must be keyed to the correct combo"
|
|
);
|
|
assert.equal(
|
|
handoff!.sessionId,
|
|
SESSION_ID,
|
|
"handoff must be keyed to the correct session"
|
|
);
|
|
|
|
// Extra dispatch observable: main (index 0) + summary (index ≥ 1).
|
|
assert.ok(
|
|
h.calls.length >= 2,
|
|
`expected ≥2 upstream dispatches (main + summary); got ${h.calls.length}: ${JSON.stringify(h.calls.map((c) => ({ i: c.index, p: c.provider, m: c.model })))}`
|
|
);
|
|
});
|
|
|
|
// ── Test 2: control — no model switch, handoff must NOT fire ──────────────────
|
|
//
|
|
// DO NOT seed a prior model. getLastSessionModel returns null → prevModel is
|
|
// null → the `if (prevModel && prevModel !== modelStr)` branch is skipped →
|
|
// no handoff. The assertion FAILS if the code incorrectly fires a handoff.
|
|
test("context-relay universal handoff: does NOT fire when no prior model is recorded (no switch)", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-noswitch" });
|
|
|
|
await combosDb.createCombo({
|
|
name: COMBO_NAME,
|
|
strategy: "context-relay",
|
|
config: { maxRetries: 0, retryDelayMs: 0, stickyRoundRobinLimit: 1 },
|
|
models: [{ id: "ns-openai", kind: "model", providerId: "openai", model: "gpt-4o-mini" }],
|
|
});
|
|
|
|
// No recordSessionModelUsage call → getLastSessionModel returns null → no switch.
|
|
installHandoffAwareFetch();
|
|
|
|
const r = await handleChat(relayRequest(true));
|
|
assert.equal(r.status, 200);
|
|
|
|
// Give setImmediate time to fire if the bug were present.
|
|
await new Promise((res) => setTimeout(res, 250));
|
|
|
|
const handoff = getHandoff(SESSION_ID, COMBO_NAME);
|
|
assert.equal(
|
|
handoff,
|
|
null,
|
|
"handoff must NOT be written when no prior model exists (prevModel is null)"
|
|
);
|
|
|
|
// Only the main request — no extra summary dispatch.
|
|
assert.equal(
|
|
h.calls.length,
|
|
1,
|
|
`expected exactly 1 upstream dispatch (main only); got ${h.calls.length}`
|
|
);
|
|
});
|
|
|
|
// ── Test 3: control — no session ID, handoff must NOT fire ────────────────────
|
|
//
|
|
// Session ID gate: `relayOptions?.sessionId` must be truthy.
|
|
// Without the x-omniroute-session-id header, sessionId = null → block is skipped.
|
|
test("context-relay universal handoff: does NOT fire when x-omniroute-session-id header is absent", async () => {
|
|
await seedConnection("openai", { apiKey: "sk-openai-nosid" });
|
|
|
|
await combosDb.createCombo({
|
|
name: COMBO_NAME,
|
|
strategy: "context-relay",
|
|
config: { maxRetries: 0, retryDelayMs: 0, stickyRoundRobinLimit: 1 },
|
|
models: [{ id: "sid-openai", kind: "model", providerId: "openai", model: "gpt-4o-mini" }],
|
|
});
|
|
|
|
// Seed prior model — but without the header the session block won't fire.
|
|
recordSessionModelUsage(SESSION_ID, COMBO_NAME, PREV_MODEL, "gemini", undefined);
|
|
|
|
installHandoffAwareFetch();
|
|
|
|
// Send WITHOUT the session header.
|
|
const r = await handleChat(relayRequest(/* withSessionId */ false));
|
|
assert.equal(r.status, 200);
|
|
|
|
await new Promise((res) => setTimeout(res, 250));
|
|
|
|
const handoff = getHandoff(SESSION_ID, COMBO_NAME);
|
|
assert.equal(
|
|
handoff,
|
|
null,
|
|
"handoff must NOT be written when x-omniroute-session-id header is absent (sessionId gate)"
|
|
);
|
|
|
|
assert.equal(
|
|
h.calls.length,
|
|
1,
|
|
`expected exactly 1 upstream dispatch (main only, no summary); got ${h.calls.length}`
|
|
);
|
|
});
|