mirror of
https://github.com/Routstr/routstr-core.git
synced 2026-08-09 19:04:47 +00:00
A generic (OpenAI-compatible) upstream whose /models response omits pricing was silently defaulting to $0.001/M tokens and a 4096 context window. For a provider like DeepSeek that reports no price, this undercharged real usage by ~280x — a direct money leak — while presenting a plausible-looking price. Resolve each model through trust-ordered sources instead: the provider's native schema (Venice's model_spec) first, then litellm's bundled cost map (curated list prices), then the OpenRouter feed. Capture the richer metadata those sources carry (cache rates, modalities, max output tokens, context) rather than only price and context. When no source knows the model, import it disabled with a warning rather than invent a number, so an operator can price it before it serves traffic. Context has no trustworthy source of last resort, but it is not a billing input, so a model priced without a reported context window falls back to an id-based estimate. The whole source-incomplete fallback (price chain + context estimate) lives in one pricing_resolver module so it can later be hoisted into the base provider unchanged. Co-Authored-By: Claude Opus 4.8 <noreply@anthropic.com>