Models/Qwen3.8-2.4T-A95B

Qwen3.8-2.4T-A95BActive · Inference

Together AI · Released Aug 2026 · Apache-2.0

Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is...

Context
1.0M
Max output
0
Input
$2.00
Output
$6.00
Cached in
$0.25
Latency
-

Capabilities

Core
Streaming
Function calling
JSON mode
Structured outputs
Multimodal
Vision
Image input
Audio
Image generation
Video
Advanced
Reasoning
Tool calling
MCP
Prompt caching
Batch API
Embeddings

Details

ProviderTogether AI
ReleasedAug 2026
LicenseApache-2.0
InputText
OutputText
Throughput-
Availability100%

Related models