gpt-oss-120B
Advanced open reasoning model with enterprise-grade capabilities.
About model
Enterprise-Ready Open Reasoning:
gpt-oss-120B delivers sophisticated chain-of-thought reasoning capabilities in a fully open model. Built with community feedback and released under Apache 2.0, this 120B parameter model provides transparency, customization, and deployment flexibility for organizations requiring complete data security & privacy control.
Model | FrontierMath Tier 4 | GPQA Diamond | HLE | SciCode | GDPval-AA | Terminal-Bench 2.1 | Agent Arena | FrontierCode | DeepSWE |
|---|---|---|---|---|---|---|---|---|---|
gpt-oss-120B | 78.2% | 18% | 39% | 15% | 26% | Related open-source models | Competitor closed-source models | ||
87.8% | 92.6% | 53% | 60% | 62% | 85% | 53.5% | 70% | ||
73.2% | 93.2% | 53% | 56% | 68% | 89% | +16.4pp | 53.4% | 74% | |
82.9% | 94.1% | 47% | 56% | 61% | 88% | 47.5% | 73% | ||
24.4% | 93.1% | 40% | 54% | 51% | 82% | +4.2pp | 42.4% | 54% | |
61.0% | 91.1% | 37% | 53% | 54% | 81% | 39.8% | 67% |
API usage
Endpoint:
Model card
Architecture Overview:
• Mixture-of-Experts (MoE) architecture with SwiGLU activations
• Alternating attention layers between full context and sliding 128-token window
• Learned attention sink per-head for enhanced performance
Training Methodology:
• Comprehensive safety training and evaluation protocols
• Community feedback integration from global listening sessions
• Rigorous testing under Preparedness Framework
• Standard GPT-4o tokenizer with additional Harmony format tokens
Performance Characteristics:
• Native FP4 quantization for efficient inference
• 128K context window with RoPE positional encoding
• Chain-of-thought reasoning with adjustable effort levelsApplications & use cases
Enterprise Applications:
• Complex reasoning and analysis tasks
• Research and development support
• Technical documentation generation
• Strategic planning and decision support
Developer Use Cases:
• Code generation and review
• API development and integration
• System architecture design
• Technical troubleshooting and debugging
Industry Solutions:
• Healthcare: Clinical decision support and medical research
• Finance: Risk analysis and regulatory compliance
• Legal: Contract analysis and legal research
• Education: Curriculum development and tutoring systems
Deployment Scenarios:
• On-premises infrastructure for data sovereignty
• Private cloud deployments for security compliance
• Custom fine-tuning for domain-specific applications
• Multi-modal integration with existing systems
- TypeReasoningChat
- Main use casesChatSmall & FastMedium General Purpose
- FeaturesJSON Mode
- SpeedHigh
- IntelligenceHigh
- DeploymentServerlessDedicated
- Endpoint
- Parameters120B
- Context length128K
- Input price
$0.15 / 1M tokens
- Output price
$0.60 / 1M tokens
- Input modalitiesText
- Output modalitiesText
- ReleasedAugust 4, 2025
- Last updatedAugust 18, 2025
- External link
- CategoryChat
