mirror of
https://github.com/diegosouzapw/OmniRoute.git
synced 2026-08-14 11:12:17 +03:00
* fix(ci): clear base-reds on release/v3.8.50 (round 3) - CHANGELOG.md: restore the top [Unreleased] section dropped by the #10189 reconcile (docs-sync gate: first section must be Unreleased) - env-doc-sync: document CONDUCTOR_ORCHESTRATOR_TOKEN + CONDUCTOR_SPOKESPERSON_URL in .env.example/ENVIRONMENT.md; allowlist the CI-only GITHUB_STEP_SUMMARY and TS7_BASE_REF (ts7 ratchet signals); drop a stray merge artifact line - providers: restore the audited chatanywhere metadata entry that base-reds round 2 dropped together with its duplicate — the provider was half-wired (registry+endpoint without APIKEY metadata), which is what the wave3 test catches; re-pin providers-constants-split at the measured 228 - docs counts: 338 -> 339 (today's +2 void-ai/helixmind, -1 Puter) via gen:provider-reference + README/AGENTS/llm.txt/package.json/diagrams/i18n mirrors - file-size ratchet: annotated rebaseline for the two pre-existing drifts (ModelSelectModal 1138, gateways 1250) following the 2026-08-11 precedent Refs #9985 * fix(ci): base-reds round 3b — stale sibling tests + mode-pack weight contract - check-docs-counts-sync.test.ts: drop the imports/subtests of the four helpers #10196 removed from the gate script (readMcpFactsFromSource, listLocalizedDocs, makeRequiredCountsValidator, checkFreeTierInventory) — the new-API tests that #10196 added stay; the file now loads again under the node runner - quota-connection-recovery.test.ts: convert from vitest APIs to node:test — the file lives in tests/unit/*.test.ts (node-runner glob) and the vitest runtime crashes when imported outside vitest, killing the whole shard entry - modePacks.ts: re-normalize all six mode packs to sum 1.0 — #8940 added sessionAvailability: 0.05 to every pack without rebalancing (1.05 total); ratios preserved exactly (÷1.05), so post-normalizeScoringWeights behavior is unchanged; restores the declared sum-to-1.0 contract the 4235 test pins Refs #9985 * fix(ci): base-reds round 3c — vitest siblings, weights default, secrets FP, mutation tap - DistributeProxiesButton.test.tsx: wrap renders in NextIntlClientProvider — #9245 localized the component (useTranslations) and left the test without the intl context, failing all 14 cases - scoring.ts: re-normalize DEFAULT_WEIGHTS to sum 1.0 (same #8940 class as the mode packs — sessionAvailability added without rebalancing; ratios preserved) - .gitleaks.toml: generalize the kimi sponsor-banner localStorage-key allowlist to -v\d+ — #10200 bumped v1→v2 and the stale regex regressed the secrets ratchet with a false positive - stryker.conf.json: register 6 covering unit tests in tap.testFiles (4 modules) so their mutant kills count — unblocks check:mutation-test-coverage --strict Refs #9985 * fix(ci): base-reds round 3d — inspector factor gap, stale registry/gap tests, i18n key sync - comboScoringInspector: add cacheAffinity/sessionAvailability/connectionDensity to FACTOR_KEYS + the factor-key type — calculateScore() weighs them but the breakdown omitted them, so the explained contributions never summed to the reported score (inspector bug, red on the pure tip) - combo-scoring-inspector.test: make the explicit-weights override sum-neutral (±0.05 shift) so it stays valid for any DEFAULT_WEIGHTS values — the hardcoded override only summed to 1.0 against the pre-#8940 defaults, which is also why explicit weights silently fell back to 'default' on the tip - unorouter-registry.test: align to the canonical .com host (api.unorouter.ai 301-redirects there, verified live) and to wave4's live model discovery (passthrough, no static seed) — the .ai/auto-model expectations were stale - check-migration-numbering.test: 147 left KNOWN_GAPS when 147_api_keys_model_access_mode.sql landed — assert absent (same as 143) - i18n: sync-ui pass — 35,914 missing UI keys stamped as __MISSING__ placeholders across 42 locales (mechanical; greens the pt-BR key-presence integrity test; coverage pct unchanged by design — translation is a separate workstream) Refs #9985 * fix(ci): base-reds round 3e — 2 real defects + 14 stale sibling tests (waves A-E) Real defects fixed: - src/lib/db/apiKeys.ts: #9313's empty-allowlist early return bypassed the group permission check, silently disabling group deny rules (#8817) for every key without a per-key allowlist; fall-through restored, restricted+[] deny-all kept - open-sse/utils/proxyFetch.ts: #10032 re-appended the raw transport error to the propagated message, reintroducing the proxy user:password leak #9837 closed; new redactProxyDetailsInMessage() keeps the reason, redacts URL/credentials - .github/workflows/quality.yml: #10134 added the TS7 ratchet as a separate blocking step AFTER the aggregated gates — the exact #8542 masking mechanism; folded into the non-fail-fast loop (still blocking, still PR-only) ⚠️ CI edit, gate-strengthening — explicit owner sign-off requested on the PR - src/i18n/messages/ko.json: 3 machine-mistranslation regressions caught by the #8244 glossary checker (장애인→비활성화됨, 양말5://→socks5://, 비클로드→Claude가 아닌) Stale sibling tests aligned to deliberately-moved contracts (each cites its mover): request-log-detail-layout + -stream (#9245 intl provider), repro-8542 pin update, quality-rail-gate-membership (#10134 shape), agentSkills-routes 45→46 (#9058), cloudflare-ai-catalog-8717 (#8804 supersedes #8808), executor-xai (#9994), vision-bridge-claude-wire (#9463 minimax→openai), sse-auth forced-pin (#8893), tls-proxy-context (strengthened leak guards), rate-limit-local-error-classification (#9164/#9342), minimax-thinking-signature (#9463), codebuddy-cn (#9723 +1 test), github-copilot-custom-model (#9050), providers-g4f-batch3 (#9584), synced-capability-warmup (#9199, stricter), sidebar-tools-group (#8221), oauth-modal-grok-cli-paste (#9245); agentSkills/catalog.ts comment 45→46; file-size rebaseline for proxyFetch (+19, annotated) Refs #9985 * fix(ci): base-reds round 3f — waves F-J: 9 more real defects + stale sibling sweep Real production defects fixed (all red on the pure tip, each with its origin): - routeGuard.ts: #8949 accidentally DELETED the /api/providers/[id]/login local-only pattern — the route spawns a browser, so the loopback gate for a process-spawning route was gone (Hard Rules #15/#17); restored (314 guard tests green) - agentSkills generator: #9058's category dispatch gave the config category an empty body, wiping skills/config-codex-cli/SKILL.md at the #10131 sync; fixed + SKILL.md regenerated via the official generator - imageRegistry: #9982 broke same-provider bare aliasing (antigravity preview id sent upstream unresolved); new resolveSameProviderBareAlias() keeps the fal cross-provider fix intact - imageRegistry: #9982's prefix strip handed the bare nano-banana ids to fal-ai, violating the pinned 2026-07-31 operator decision (adobe-firefly owns them); fal entries made prefix-only (dispatch already re-prefixes) - mediaGeneration/fal.ts: the missing-credential 401 guard was lost when #10198 deleted the superseded falHandler — tests were hitting the live network - bottleneckPatch/rateLimitManager: #9041's merge clobbered #9604, resurrecting the Bottleneck v2.19.5 heartbeat bug (reservoir never refills); patched the library defect at the root and re-aligned chat-rate-limit-body-lock to the working reservoir contract - processSupervisor.mjs: #9761 regressed the Node spawn to bare "node" (the #9156 launchd bug) and dropped #9209's ipv4first args; both restored - openai-responses/pureHelpers: #9423's Agent null-sentinel was unreachable on the schemaless JSON-string path; gate extended - i18n en.json: #8222's regen reverted the #9976 unclosed-tag fix and #8559's combo-cooldown copy; #9038 shipped 40 t() calls with no messages (runtime MISSING_MESSAGE); all restored/added + official sync-ui stamps, and vi's zero-marker policy re-established via the sanctioned translation backend Stale sibling tests aligned (movers cited inline): chat-helpers (#9447), executor-antigravity (#9351), video-fal-grok (#9982), visionBridge (#9759), web-session-credentials (#8974), production-build-module-integrity (positive anchor added), agentSkills-generator/skillManifestsLint/skills-injection/ agentSkillTools-mcp/listCapabilities-a2a (#9058), memory-settings (#10010), model-catalog-policy-invalidation (#8906), model-alias-seed (#9485), reactive-context-compaction (#8949), combo-provider-wildcard (broken upsert helper), oauth-google-loopback (43-locale resurrected-key removal) Validation: 501/501 across the 47 touched test files; typecheck:core, lint, file-size, docs-sync all green. Refs #9985 * fix(ci): base-reds round 3g — wave K/L: 4 more real defects + stale alignments Real defects: - base/reasoningEffort.ts: the stale duplicate cherry-pick #9612 re-added the codex minimal→low rewrite that #9883 had deliberately removed (OMP minimal passthrough); block removed again - cursorImages.ts: #9840 wired prepareCursorImageForWire (sharp re-encode, fail-closed) into the SHARED resolveCursorImages, breaking zai-web and conol-web image uploads (HTTP 400 'undecodable'); new prepareForWire opt-out, Cursor default path unchanged (8 cursor suites green) - modelCapabilities/snapshot: catalog prepare still issued 323 per-model reads of model_context_overrides + max_input_tokens overrides, violating #9199's bulk-load contract; both now resolve from the snapshot single pass - v1-models-discovery-conformance: re-pinned to the bounded 30s SWR window (#9199/#10198) — the old 'stale-first regardless of age' contract is gone Stale tests aligned (movers cited inline): codex-tools-strict-default (#9828 redundant-oneOf strip), devin-providers (#9245 i18n), db-migrationrunner- constants-split (147→151 renumber #8228), gitlab-duo-oauth-setup (#9245), chatcore-extracted-modules (#9161 outbound-protocol keying) compression-api CI failures were cascade artifacts of codex-tools-strict-default failing in the same force-exit shard process — no own defect (171/171 local). Refs #9985 * fix(test): compression-api — register both describes before the runner starts The DATA_DIR setup + route/db top-level awaits sat BETWEEN the two describes; under --test-force-exit (the CI unit-runner flag) the process exits once the already-registered tests finish, so on slow CI machines the whole second describe died as 'Promise resolution is still pending' — the recurring CI-only shard-2 failure that never reproduced locally without the flag. Moved to the top of the file; 10/10 under --test-force-exit locally. Refs #9985 * fix(quality): freeze modelCapabilities.ts at 1006 (annotated) — snapshot routing growth Refs #9985 * fix(quality): move the modelCapabilities freeze into the frozen map (nested schema) Refs #9985 * fix(i18n): translate all 39,718 pending UI keys across 42 locales (owner-approved) Mass-translated every __MISSING__ placeholder via the official i18n:sync-ui --translate-markers pipeline (operator backend), restoring i18nUiCoverage to the 100 baseline (was 89.9 after the merge-storm UI landings + the 42 keys #9038 never shipped). Post-pass repairs, all caught by the existing gates: - glossary: retired renderings the machine reintroduced normalized again (提供商→提供者 zh-CN/zh-TW, 鏈接→連結, 文檔→文件, 調用→呼叫, 供應商→提供者, 響應→回應, 不活躍→未啟用 zh-TW; 클로드→Claude, 옴니루트→OmniRoute ko); DATA_DIR forbidden rendering avoided via 数据文件夹 rephrase - ICU integrity: 120 values with renamed/dropped {params} repaired (39 positional renames, 81 reset to the en source — functional over fluent) Validation: glossary/pt-BR/vi/deno-relay/settings-keys/value-drift/google- loopback suites 76/76; placeholder diff en×42 locales = 0; worst-locale coverage = 100.0%. Refs #9985 --------- Co-authored-by: backryun <bakryun0718@proton.me>
372 lines
12 KiB
TypeScript
372 lines
12 KiB
TypeScript
import test from "node:test";
|
|
import assert from "node:assert/strict";
|
|
import fs from "node:fs";
|
|
import os from "node:os";
|
|
import path from "node:path";
|
|
|
|
const TEST_DATA_DIR = fs.mkdtempSync(path.join(os.tmpdir(), "omniroute-rl-local-errors-"));
|
|
process.env.DATA_DIR = TEST_DATA_DIR;
|
|
process.env.API_KEY_SECRET = "test-rate-limit-local-error-secret";
|
|
|
|
// Dynamic imports are required because DATA_DIR must be set before DB modules evaluate.
|
|
const core = await import("../../src/lib/db/core.ts");
|
|
const providersDb = await import("../../src/lib/db/providers.ts");
|
|
const { handleComboChat } = await import("../../open-sse/services/combo.ts");
|
|
const {
|
|
isComboRequestScopedFailure,
|
|
isRequestScopedUpstreamFailure,
|
|
shouldRecordProviderBreakerFailure,
|
|
shouldSkipConnDisable,
|
|
} = await import("../../open-sse/services/combo/comboPredicates.ts");
|
|
const {
|
|
LEGACY_RATE_LIMIT_QUEUE_TIMEOUT_CODE,
|
|
RATE_LIMIT_EXECUTION_TIMEOUT_CODE,
|
|
RATE_LIMIT_QUEUE_WEDGED_CODE,
|
|
getTrustedLocalRateLimitError,
|
|
getTrustedLocalRateLimitResponse,
|
|
inheritTrustedLocalRateLimitResponse,
|
|
markLocalRateLimitError,
|
|
markTrustedLocalRateLimitResponse,
|
|
} = await import("../../open-sse/services/rateLimitManager/errors.ts");
|
|
const accountFallback = await import("../../open-sse/services/accountFallback.ts");
|
|
const providerCooldown = await import("../../open-sse/services/providerCooldownTracker.ts");
|
|
const rateLimitSemaphore = await import("../../open-sse/services/rateLimitSemaphore.ts");
|
|
const { createStreamingErrorResult } =
|
|
await import("../../open-sse/handlers/chatCore/streamErrorResult.ts");
|
|
const { shouldTripProviderBreakerForResult } =
|
|
await import("../../src/sse/handlers/chatPredicates.ts");
|
|
|
|
const LOCAL_ERROR_MESSAGE = "OmniRoute repaired a local limiter queue";
|
|
|
|
function createLocalLimiterSseResponse(connectionId: string, code = RATE_LIMIT_QUEUE_WEDGED_CODE) {
|
|
const error = markLocalRateLimitError(new Error(LOCAL_ERROR_MESSAGE), code);
|
|
const { response } = createStreamingErrorResult(
|
|
getTrustedLocalRateLimitError(error)?.status ?? 503,
|
|
LOCAL_ERROR_MESSAGE,
|
|
code,
|
|
"rate_limit_queue_wedged"
|
|
);
|
|
response.headers.set("X-OmniRoute-Selected-Connection-Id", connectionId);
|
|
return markTrustedLocalRateLimitResponse(response, error);
|
|
}
|
|
|
|
function createUpstreamCollisionResponse(connectionId: string) {
|
|
return new Response(
|
|
JSON.stringify({
|
|
error: {
|
|
message: "Provider emitted a colliding code",
|
|
code: RATE_LIMIT_QUEUE_WEDGED_CODE,
|
|
type: "rate_limit_queue_wedged",
|
|
},
|
|
}),
|
|
{
|
|
status: 503,
|
|
headers: {
|
|
"content-type": "application/json",
|
|
"X-OmniRoute-Selected-Connection-Id": connectionId,
|
|
},
|
|
}
|
|
);
|
|
}
|
|
|
|
function createSuccessResponse(connectionId: string) {
|
|
return new Response(JSON.stringify({ choices: [{ message: { content: "fallback ok" } }] }), {
|
|
status: 200,
|
|
headers: {
|
|
"content-type": "application/json",
|
|
"X-OmniRoute-Selected-Connection-Id": connectionId,
|
|
},
|
|
});
|
|
}
|
|
|
|
const log = { info() {}, warn() {}, error() {}, debug() {} };
|
|
const settings = {
|
|
modelLockout: {
|
|
enabled: true,
|
|
errorCodes: [503],
|
|
baseCooldownMs: 3_000,
|
|
maxCooldownMs: 5_000,
|
|
maxBackoffSteps: 10,
|
|
useExponentialBackoff: true,
|
|
},
|
|
};
|
|
|
|
test.afterEach(() => {
|
|
accountFallback.clearAllModelLockouts();
|
|
accountFallback.clearProviderFailure("openai");
|
|
providerCooldown.clearCooldownState();
|
|
rateLimitSemaphore.resetAll();
|
|
core.resetDbInstance();
|
|
fs.rmSync(TEST_DATA_DIR, { recursive: true, force: true });
|
|
fs.mkdirSync(TEST_DATA_DIR, { recursive: true });
|
|
});
|
|
|
|
test.after(() => {
|
|
accountFallback.clearAllModelLockouts();
|
|
accountFallback.clearProviderFailure("openai");
|
|
providerCooldown.clearCooldownState();
|
|
rateLimitSemaphore.resetAll();
|
|
core.resetDbInstance();
|
|
fs.rmSync(TEST_DATA_DIR, { recursive: true, force: true });
|
|
});
|
|
|
|
test("execution-timeout classification requires trusted provenance; queue codes classify by string (#9164/#9342)", () => {
|
|
const executionError = markLocalRateLimitError(
|
|
new Error("local execution expiration"),
|
|
RATE_LIMIT_EXECUTION_TIMEOUT_CODE
|
|
);
|
|
const localResponse = markTrustedLocalRateLimitResponse(
|
|
new Response("local", { status: 504 }),
|
|
executionError
|
|
);
|
|
const collisionResponse = createUpstreamCollisionResponse("collision-conn");
|
|
|
|
assert.deepEqual(getTrustedLocalRateLimitError(executionError), {
|
|
code: RATE_LIMIT_EXECUTION_TIMEOUT_CODE,
|
|
status: 504,
|
|
});
|
|
assert.equal(getTrustedLocalRateLimitResponse(localResponse)?.status, 504);
|
|
const wrappedResponse = inheritTrustedLocalRateLimitResponse(
|
|
localResponse,
|
|
new Response("wrapped local", { status: 504 })
|
|
);
|
|
assert.equal(getTrustedLocalRateLimitResponse(wrappedResponse)?.status, 504);
|
|
assert.equal(
|
|
isRequestScopedUpstreamFailure({ code: RATE_LIMIT_EXECUTION_TIMEOUT_CODE }),
|
|
false,
|
|
"an upstream-controlled code string must not establish local provenance"
|
|
);
|
|
assert.equal(
|
|
isComboRequestScopedFailure(localResponse, "local execution expiration", {
|
|
code: RATE_LIMIT_EXECUTION_TIMEOUT_CODE,
|
|
}),
|
|
true
|
|
);
|
|
// #9164 (3898305df0) deliberately widened the contract: the rate_limit_queue_*
|
|
// code strings are OmniRoute-owned backpressure codes and classify as
|
|
// request-scoped even without WeakMap provenance (an upstream collision is
|
|
// accepted as fail-safe: worst case a colliding provider 503 skips health
|
|
// penalties, it never amplifies into fallback storms).
|
|
assert.equal(
|
|
isComboRequestScopedFailure(collisionResponse, "provider collision", {
|
|
code: RATE_LIMIT_QUEUE_WEDGED_CODE,
|
|
}),
|
|
true
|
|
);
|
|
assert.equal(
|
|
shouldTripProviderBreakerForResult(
|
|
{
|
|
status: 504,
|
|
response: localResponse,
|
|
errorCode: RATE_LIMIT_EXECUTION_TIMEOUT_CODE,
|
|
},
|
|
false,
|
|
false
|
|
),
|
|
false
|
|
);
|
|
assert.equal(
|
|
shouldTripProviderBreakerForResult(
|
|
{
|
|
status: 503,
|
|
response: collisionResponse,
|
|
errorCode: RATE_LIMIT_QUEUE_WEDGED_CODE,
|
|
},
|
|
false,
|
|
false
|
|
),
|
|
false,
|
|
"#9342 (47c819df66): RATE_LIMIT_QUEUE_* codes are OmniRoute backpressure and never trip the provider breaker, provenance or not"
|
|
);
|
|
assert.equal(
|
|
shouldSkipConnDisable(
|
|
{
|
|
status: 504,
|
|
response: localResponse,
|
|
errorCode: RATE_LIMIT_EXECUTION_TIMEOUT_CODE,
|
|
},
|
|
false,
|
|
false,
|
|
"openai"
|
|
),
|
|
true
|
|
);
|
|
assert.equal(
|
|
shouldSkipConnDisable(
|
|
{
|
|
status: 503,
|
|
response: collisionResponse,
|
|
errorCode: RATE_LIMIT_QUEUE_WEDGED_CODE,
|
|
},
|
|
false,
|
|
false,
|
|
"openai"
|
|
),
|
|
true,
|
|
"#9164: the queue-code string alone marks the failure request-scoped, so the connection is not disabled"
|
|
);
|
|
assert.equal(
|
|
shouldRecordProviderBreakerFailure({
|
|
isStreamReadinessFailure: false,
|
|
status: 504,
|
|
sameProviderNext: false,
|
|
skipProviderBreaker: false,
|
|
requestScopedFailure: true,
|
|
error: executionError,
|
|
isProxyUnreachable: false,
|
|
}),
|
|
false
|
|
);
|
|
});
|
|
|
|
test("legacy queue-timeout code classifies as request-scoped with or without provenance (#9164)", () => {
|
|
const untrusted = new Response("legacy collision", { status: 503 });
|
|
const legacyError = markLocalRateLimitError(
|
|
new Error("legacy local timeout"),
|
|
LEGACY_RATE_LIMIT_QUEUE_TIMEOUT_CODE
|
|
);
|
|
const trusted = markTrustedLocalRateLimitResponse(
|
|
new Response("legacy local timeout", { status: 503 }),
|
|
legacyError
|
|
);
|
|
|
|
// #9164 added rate_limit_queue_timeout to REQUEST_SCOPED_UPSTREAM_ERROR_CODES,
|
|
// so the code string is sufficient — trusted provenance is no longer required
|
|
// for this classification (it still works, next assertion).
|
|
assert.equal(
|
|
isComboRequestScopedFailure(untrusted, "legacy collision", {
|
|
code: LEGACY_RATE_LIMIT_QUEUE_TIMEOUT_CODE,
|
|
}),
|
|
true
|
|
);
|
|
assert.equal(
|
|
isComboRequestScopedFailure(trusted, "legacy local timeout", {
|
|
code: LEGACY_RATE_LIMIT_QUEUE_TIMEOUT_CODE,
|
|
}),
|
|
true
|
|
);
|
|
});
|
|
|
|
for (const strategy of ["priority", "round-robin"] as const) {
|
|
test(`${strategy} fallback preserves all health state for a trusted local SSE failure`, async () => {
|
|
const connection = await providersDb.createProviderConnection({
|
|
provider: "openai",
|
|
authType: "apikey",
|
|
name: `local-wedge-${strategy}`,
|
|
apiKey: `sk-local-wedge-${strategy}`,
|
|
isActive: true,
|
|
testStatus: "active",
|
|
rateLimitedUntil: null,
|
|
backoffLevel: 0,
|
|
providerSpecificData: {},
|
|
});
|
|
const models = [
|
|
{
|
|
kind: "model",
|
|
model: "openai/gpt-local-first",
|
|
connectionId: connection.id,
|
|
},
|
|
{
|
|
kind: "model",
|
|
model: "openai/gpt-local-second",
|
|
connectionId: connection.id,
|
|
},
|
|
];
|
|
const calls: string[] = [];
|
|
|
|
const result = await handleComboChat({
|
|
body: {},
|
|
combo: {
|
|
name: `local-wedge-${strategy}-combo`,
|
|
strategy,
|
|
models,
|
|
config: {
|
|
maxRetries: 1,
|
|
retryDelayMs: 0,
|
|
fallbackDelayMs: 0,
|
|
maxConcurrency: 1,
|
|
},
|
|
},
|
|
handleSingleModel: async (_body, modelStr) => {
|
|
calls.push(modelStr);
|
|
return calls.length === 1
|
|
? createLocalLimiterSseResponse(connection.id)
|
|
: createSuccessResponse(connection.id);
|
|
},
|
|
isModelAvailable: async () => true,
|
|
log,
|
|
settings,
|
|
allCombos: null,
|
|
});
|
|
|
|
assert.equal(result.status, 200, `attempted targets: ${calls.join(", ")}`);
|
|
assert.deepEqual(calls, ["openai/gpt-local-first", "openai/gpt-local-second"]);
|
|
assert.equal(accountFallback.isModelLocked("openai", connection.id, "gpt-local-first"), false);
|
|
assert.equal(
|
|
accountFallback.getProviderBreakerState("openai")?.failureCount ?? 0,
|
|
0,
|
|
"local failure must not increment the provider breaker"
|
|
);
|
|
assert.equal(
|
|
providerCooldown.isProviderInCooldown("openai", connection.id),
|
|
false,
|
|
"local failure must not enter provider cooldown"
|
|
);
|
|
const semaphoreStates = Object.values(rateLimitSemaphore.getStats());
|
|
assert.equal(
|
|
semaphoreStates.some((state) => state.rateLimitedUntil !== null),
|
|
false,
|
|
"local failure must not cool a round-robin semaphore"
|
|
);
|
|
const storedConnection = await providersDb.getProviderConnectionById(connection.id);
|
|
assert.equal(storedConnection?.testStatus, "active");
|
|
assert.equal(storedConnection?.rateLimitedUntil ?? null, null);
|
|
});
|
|
}
|
|
|
|
test("an upstream body colliding with local queue codes is treated as local backpressure (#9164)", async () => {
|
|
const connection = await providersDb.createProviderConnection({
|
|
provider: "openai",
|
|
authType: "apikey",
|
|
name: "upstream-local-code-collision",
|
|
apiKey: "sk-upstream-local-code-collision",
|
|
isActive: true,
|
|
testStatus: "active",
|
|
providerSpecificData: {},
|
|
});
|
|
|
|
const result = await handleComboChat({
|
|
body: {},
|
|
combo: {
|
|
name: "upstream-local-code-collision-combo",
|
|
strategy: "priority",
|
|
models: [
|
|
{
|
|
kind: "model",
|
|
model: "openai/gpt-collision",
|
|
connectionId: connection.id,
|
|
},
|
|
],
|
|
config: { maxRetries: 1, retryDelayMs: 0, fallbackDelayMs: 0 },
|
|
},
|
|
handleSingleModel: async () => createUpstreamCollisionResponse(connection.id),
|
|
isModelAvailable: async () => true,
|
|
log,
|
|
settings,
|
|
allCombos: null,
|
|
});
|
|
|
|
assert.equal(result.status, 503);
|
|
// #9164 (3898305df0): isLocalQueueCapacityErrorBody matches the queue-code
|
|
// string in the body, so the combo returns the 503 without upstream fallback
|
|
// and without counting it toward provider health — the collision is accepted
|
|
// as fail-safe (no breaker/cooldown penalties, but also no retry amplification).
|
|
assert.equal(
|
|
accountFallback.getProviderBreakerState("openai")?.failureCount ?? 0,
|
|
0,
|
|
"a queue-code collision body is classified local backpressure and skips breaker accounting"
|
|
);
|
|
const storedConnection = await providersDb.getProviderConnectionById(connection.id);
|
|
assert.equal(storedConnection?.testStatus, "active", "the connection must not be disabled");
|
|
});
|