LiteLLM fork with a corgi per-request routing brain: filters candidate models by capability, then optimizes cost or latency within the capability frontier, steered by X-Router-* headers.
python cost-optimization llm-proxy litellm ai-gateway llm-router llm-gateway openai-compatible model-routing routing-strategy per-request-routing
-
Updated
Jul 28, 2026 - Python