GPT-OSS-120B is OpenAI's open-source lightweight model with 120 billion parameters, offering fast inference and a 128K token context window for developers.
| Provider | Input / 1M | Output / 1M | Cached / 1M |
|---|---|---|---|
| $0.037 | $0.170 | - | |
| $0.050 | $0.250 | - | |
| $0.060 | $0.390 | - | |
| $0.070 | $0.300 | - | |
| $0.100 | $0.400 | - | |
| $0.150 | $0.600 | - | |
| $0.150 | $0.600 | - | |
| $0.150 | $0.600 | - | |
| $0.150 | $0.600 | - | |
| $0.150 | $0.600 | - | |
| $0.150 | $0.600 | - | |
| $0.174 | $0.697 | - | |
| $0.188 | $0.700 | $0.094 | |
| $0.350 | $0.750 | - |
Prices updated daily. Last check: Sep 6, 2026
Input, output, and batch rates, plus alternatives, for one provider at a time.
Benchmarks measured Sep 2026. Scores are independent evaluations, not vendor-reported.
GPT-OSS-120B is well-suited for applications requiring fast, reliable text generation without the complexity of tool use or multimodal capabilities. Its open-source nature makes it ideal for organizations needing local deployment for data privacy, custom fine-tuning for domain-specific tasks, or integration into products without API dependencies. Common use cases include content generation, document summarization, customer service chatbots, code commenting, and batch text processing workflows where the 128K context window enables handling of long documents. The model's lightweight design and fast inference make it particularly valuable for high-throughput applications or resource-constrained environments where deploying larger frontier models would be impractical.
GPT-OSS-120B pricing varies by provider and deployment method. Since it's open-source, you can also run it locally without per-token costs. Check the pricing table above for current rates across API providers.
GPT-OSS-120B excels at text generation tasks requiring fast inference and long context handling, such as content creation, document summarization, and chatbot applications. Its open-source nature makes it ideal for organizations needing local deployment, custom fine-tuning, or applications with data privacy requirements.
GPT-OSS-120B trades some advanced capabilities for accessibility and control. Unlike proprietary GPT models, it lacks tool calling and multimodal features but offers open-source weights for local deployment and customization. It's designed for use cases where fast, reliable text generation is needed without the full feature set of frontier models.