Files
alfadb f06bf181d2 feat(openai): support Fast mode service_tier across responses/chat/WS paths
- Accept fast|priority (canonical priority), flex|auto|default|scale on
  /v1/responses and /v1/chat/completions; reject unknown/empty/non-string
  with HTTP 400; omitted and null stay compatible.
- Propagate service_tier through JSON/SSE, Responses<->Chat conversions,
  fallback paths and HTTP->upstream WebSocket bridge.
- Billing prefers the upstream terminal tier; the outbound (policy-
  transformed) tier is used only when upstream omits the field.
  Explicit upstream default bills Standard even when Fast was requested.
- Pricing: Fast premium 2x Standard for gpt-5.6-sol/terra/luna and
  gpt-5.4; 2.5x for gpt-5.5; channel FastMultiplier stays authoritative.
- Live verification (official Codex 0.149.0 + gateway, HTTP & WS):
  upstream ChatGPT backend may return terminal default even when the
  account catalog advertises priority; billing follows the actual tier.
2026-08-24 11:52:48 +08:00
..
2025-12-18 13:50:39 +08:00

Model Pricing Data

This directory contains a local copy of the mirrored model pricing data as a fallback mechanism.

Source

The original file is maintained by the LiteLLM project and mirrored into the price-mirror branch of this repository via GitHub Actions:

Purpose

This local copy serves as a fallback when the remote file cannot be downloaded due to:

  • Network restrictions
  • Firewall rules
  • DNS resolution issues
  • GitHub being blocked in certain regions
  • Docker container network limitations

Update Process

The pricingService will:

  1. First attempt to download the latest version from GitHub
  2. If download fails, use this local copy as fallback
  3. Log a warning when using the fallback file

Manual Update

To manually update this file with the latest pricing data (if automation is unavailable):

curl -s https://raw.githubusercontent.com/BerriAI/litellm/main/model_prices_and_context_window.json -o model_prices_and_context_window.json

File Format

The file contains JSON data with model pricing information including:

  • Model names and identifiers
  • Input/output token costs
  • Context window sizes
  • Model capabilities

Last updated: 2025-08-10