Models/gpt-oss-120b

gpt-oss-120bActive · Inference

Replicate · Released Aug 2025 · Apache-2.0

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

Context
0
Max output
0
Input
$0.18
Output
$0.72
Cached in
-
Latency
-

Capabilities

Core
Streaming
Function calling
JSON mode
Structured outputs
Multimodal
Vision
Image input
Audio
Image generation
Video
Advanced
Reasoning
Tool calling
MCP
Prompt caching
Batch API
Embeddings

Benchmarks

Aider Polyglot41.8

Details

ProviderReplicate
ReleasedAug 2025
LicenseApache-2.0
InputText
OutputText
Throughput-
Availability100%

Related models