Skip to main content
Use this table to read or compare Valar’s public list prices without parsing the visual pricing page. All amounts are USD per 1M tokens. Model IDs are case-sensitive. Retired IDs keep working as aliases of their replacement and are billed at the replacement’s rates: requests for deepseek-ai/DeepSeek-V4-Flash are served as deepseek-ai/DeepSeek-V4.1-Flash. This table contains the same models and rates as Pricing, covering Now, Priority, Standard, and Flex. Hidden and restricted models are excluded. Both tables are generated from the same public catalog and update when the docs are published; this is not a live quote or a feed of organization-specific rates.

Read or export

No API key is required. Open this page as Markdown, or connect to the docs MCP server and ask:
The Markdown table is the published export. Any JSON or CSV you ask an agent to produce is a conversion of that table, not a separate API response.

Public rates

Each row identifies a model and a completion window. Input is uncached input; cached input is input served from the prompt cache; output includes generated reasoning tokens. See Completion windows for scheduling behavior. Claude and OpenAI models also bill input tokens written into the prompt cache at a cache-write rate: 1.25x the input rate, or 2x for a Claude write cached with the 1-hour TTL. Other models have no cache-write charge. See Cache writes. For your organization’s available models and negotiated rates, use the separate authenticated GET /v1/models API. The public docs MCP does not call that API.