Models/gpt-oss-120b

gpt-oss-120bActive · Cloud

WandB Inference · Released Aug 2025 · Apache-2.0

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

Context
131K
Max output
131K
Input
$15000.00
Output
$60000.00
Cached in
-
Latency
-

Capabilities

Core
Streaming
Function calling
JSON mode
Structured outputs
Multimodal
Vision
Image input
Audio
Image generation
Video
Advanced
Reasoning
Tool calling
MCP
Prompt caching
Batch API
Embeddings

Benchmarks

Aider Polyglot41.8

Details

ReleasedAug 2025
LicenseApache-2.0
InputText
OutputText
Throughput-
Availability100%

Related models