Files
OmniRoute/tests/unit/nvidia-eol-catalog.test.ts
Praveen K Palaniswamy 65e81158ab fix(ollama): route models by advertised capability (#11088)
Landed with the design call resolved per the owner's pick — **option 1**: the synced store is now endpoint-agnostic (persistDiscoveredModels and managedModelImport no longer drop non-chat models at write time), and chat selectability moved to read time (auto-pool expansion in autoStrategy applies filterChatSelectableModels; the models-route projection already had its chatOnly filter). Your discovery test now passes end-to-end (3/3): /api/show capabilities persist per connection and image/embedding requests route through the advertising host.

Reconciliation notes: conflicted areas merged onto the current tip (adobe discovery import, requestedModel preflight signature, resolvedProvider fast-path coexists with the synced-route override — explicit resolution wins); carried base-red drains (#10055 memoization, #11071 test variants) dropped as already-landed; the managed-model-import exclusion test was propagated to the new contract (image/video models persist; the read filter still hides them from chat pickers — pinned by a new assertion). Full battery: 205/206 focused (the one red is a confirmed periodic-timer timing flake on the loaded devbox — 20/20 isolated), autoCombo vitest 30/30, combo suites 46/46, gates + typecheck clean.

Thank you @yourspraveen — the capability probe + routing design was right; it just needed the store contract opened up. Fixes #11087.
2026-08-23 11:45:01 -03:00

55 lines
1.8 KiB
TypeScript

import test from "node:test";
import assert from "node:assert/strict";
import { FREE_MODEL_BUDGETS } from "../../open-sse/config/freeModelCatalog.data.ts";
import reviewedLiveIds from "../../open-sse/config/nvidiaHostedModels.snapshot.json" with { type: "json" };
import { nvidiaProvider } from "../../open-sse/config/providers/registry/nvidia/index.ts";
const registryIds = new Set(nvidiaProvider.models.map((model) => model.id));
const documentedFreeIds = new Set(
FREE_MODEL_BUDGETS.filter((model) => model.provider === "nvidia").map((model) => model.modelId)
);
const reviewedIds = new Set(reviewedLiveIds);
test("NVIDIA registry excludes retired DeepSeek V4 models", () => {
assert.ok(
!registryIds.has("deepseek-ai/deepseek-v4-pro"),
"retired deepseek-ai/deepseek-v4-pro must not be advertised"
);
assert.ok(
!registryIds.has("deepseek-ai/deepseek-v4-flash"),
"retired deepseek-ai/deepseek-v4-flash must not be advertised"
);
});
test("NVIDIA static lifecycle metadata excludes known EOL models", () => {
for (const modelId of ["z-ai/glm-5.1", "deepseek-ai/deepseek-v4-pro"]) {
assert.ok(
!reviewedIds.has(modelId),
`${modelId} must not remain in the reviewed NVIDIA hosted-model snapshot`
);
assert.ok(
!documentedFreeIds.has(modelId),
`${modelId} must not remain in the NVIDIA free-model catalog`
);
}
});
test("NVIDIA cleanup preserves the healthy GLM replacement", () => {
assert.ok(registryIds.has("z-ai/glm-5.2"), "z-ai/glm-5.2 must remain in the NVIDIA registry");
assert.ok(
reviewedIds.has("z-ai/glm-5.2"),
"z-ai/glm-5.2 must remain in the reviewed NVIDIA hosted-model snapshot"
);
assert.ok(
documentedFreeIds.has("z-ai/glm-5.2"),
"z-ai/glm-5.2 must remain in the NVIDIA free-model catalog"
);
});