/v1/chat/completions). All of them are reasoning models — they think before answering and return the chain-of-thought in reasoning_content. Context window: 256k tokens.
Available models
Reasoning tokens are billed as completion tokens — budget
max_tokens generously. Prompt-cache hits are billed at a reduced input rate automatically.Quick example
Parameters
Response notes
API Reference
View the interactive API playground.