Open-weight coding models
These are the cheaper targets you route traffic to. They are regular catalog models, and ValarCode prices each request at the model’s Now tier rate, which you can see on the Pricing page.Frontier Claude tiers
These stand in for the Opus, Sonnet, Haiku, and Fable classes your harness asks for, and they are the baseline that savings are measured against. They are also the fallback when a key’s routing cannot be read.
Claude Opus 4.8 and Sonnet 4.6 are also selectable in the routing picker, though nothing falls back to them.
The frontier Claude tiers are specific to ValarCode routing. They do not appear on the public Models page and are not callable as standalone models from the Responses or Chat Completions APIs. They exist so a cohort can hold a Claude baseline and so savings compare against real Claude list prices.
How each harness is served
ValarCode meets each tool on its own API:
Every endpoint requires a coding key. Requests are forwarded to the resolved target model, and the response is rewritten so the model your harness sees is the one it asked for. That keeps the harness working across restarts, whatever target actually served the request.
Whatever the target, Valar serves it on throughput-optimized inference tuned for agentic workloads. That highly efficient serving is what makes the open-weight models a practical everyday default rather than something you reach for only on cheap background work.
Next steps
Model routing
Point each alias at one of these models per cohort.
Analytics & savings
See the served-model mix and per-model savings.