mirror of
https://github.com/diegosouzapw/OmniRoute.git
synced 2026-09-13 18:32:12 +03:00
Landed with the design call resolved per the owner's pick — **option 1**: the synced store is now endpoint-agnostic (persistDiscoveredModels and managedModelImport no longer drop non-chat models at write time), and chat selectability moved to read time (auto-pool expansion in autoStrategy applies filterChatSelectableModels; the models-route projection already had its chatOnly filter). Your discovery test now passes end-to-end (3/3): /api/show capabilities persist per connection and image/embedding requests route through the advertising host. Reconciliation notes: conflicted areas merged onto the current tip (adobe discovery import, requestedModel preflight signature, resolvedProvider fast-path coexists with the synced-route override — explicit resolution wins); carried base-red drains (#10055 memoization, #11071 test variants) dropped as already-landed; the managed-model-import exclusion test was propagated to the new contract (image/video models persist; the read filter still hides them from chat pickers — pinned by a new assertion). Full battery: 205/206 focused (the one red is a confirmed periodic-timer timing flake on the loaded devbox — 20/20 isolated), autoCombo vitest 30/30, combo suites 46/46, gates + typecheck clean. Thank you @yourspraveen — the capability probe + routing design was right; it just needed the store contract opened up. Fixes #11087.
44 lines
1.5 KiB
TypeScript
44 lines
1.5 KiB
TypeScript
import test from "node:test";
|
|
import assert from "node:assert/strict";
|
|
|
|
const rlm = await import("../../open-sse/services/rateLimitManager.ts");
|
|
const {
|
|
enableRateLimitProtection,
|
|
withRateLimit,
|
|
updateFromHeaders,
|
|
updateFromResponseBody,
|
|
__resetRateLimitManagerForTests,
|
|
} = rlm;
|
|
|
|
test.beforeEach(async () => {
|
|
await __resetRateLimitManagerForTests();
|
|
});
|
|
|
|
test("updateFromResponseBody overwrites updateFromHeaders retry-after", async () => {
|
|
enableRateLimitProtection("test-seq-1");
|
|
await withRateLimit("openai", "test-seq-1", "gpt-4", async () => "ok");
|
|
const headers = new Headers({ "retry-after": "5" });
|
|
updateFromHeaders("openai", "test-seq-1", headers, 429, "gpt-4");
|
|
updateFromResponseBody("openai", "test-seq-1", JSON.stringify({ retry_after: 10 }), 429, "gpt-4");
|
|
});
|
|
|
|
test("no retry-after in either source leaves limiter state unchanged", async () => {
|
|
enableRateLimitProtection("test-seq-2");
|
|
await withRateLimit("openai", "test-seq-2", "gpt-4", async () => "ok");
|
|
const headers = new Headers({});
|
|
updateFromHeaders("openai", "test-seq-2", headers, 200, "gpt-4");
|
|
updateFromResponseBody("openai", "test-seq-2", "{}", 200, "gpt-4");
|
|
});
|
|
|
|
test("response body retry-after is parsed correctly", async () => {
|
|
enableRateLimitProtection("test-seq-3");
|
|
await withRateLimit("openai", "test-seq-3", "gpt-4", async () => "ok");
|
|
updateFromResponseBody(
|
|
"openai",
|
|
"test-seq-3",
|
|
JSON.stringify({ data: { retry_after: 30 } }),
|
|
429,
|
|
"gpt-4"
|
|
);
|
|
});
|