Model Library

Explore Models

Production-ready models optimized for performance and cost. Deploy in seconds via our unified API.

GPT-OSS-120B is OpenAI's higher-capability open-weight model for complex reasoning, coding agents, function calling and structured outputs. Use it for workloads that require multi-step problem solving, planning, tool use or more difficult coding tasks, especially when you want an open model that can be customized or deployed outside OpenAI's hosted API.

openai/gpt-oss-120b
$0.039per 1M tokens

GPT-OSS-20B is OpenAI's lower-latency open-weight reasoning model for frequent, well-defined developer and agent tasks. Use it for extraction, classification, routing, structured outputs, tool calls, lightweight coding and high-volume agent loops where responsiveness matters more than maximum reasoning capability. It is the natural GPT-OSS choice when 120B would be unnecessarily heavy for the task.

openai/gpt-oss-20b
$0.030per 1M tokens