Models/mistral-nemo@latest

mistral-nemo@latestActive · Cloud

Google - Vertex AI · Released Jul 2024 · Apache-2.0

A 12B parameter model with a 128k token context length built by Mistral in collaboration with NVIDIA. The model is multilingual, supporting English, French, German, Spanish, Italian, Portuguese, Chinese, Japanese,...

Context
128K
Max output
128K
Input
$0.15
Output
$0.15
Cached in
-
Latency
-

Capabilities

Core
Streaming
Function calling
JSON mode
Structured outputs
Multimodal
Vision
Image input
Audio
Image generation
Video
Advanced
Reasoning
Tool calling
MCP
Prompt caching
Batch API
Embeddings

Benchmarks

BBH29.7
GPQA5.4
IFEval63.8
MATH Lvl 512.7
MMLU-Pro28
MUSR8.5
Open LLM Average24.7

Details

ReleasedJul 2024
LicenseApache-2.0
InputText
OutputText
Throughput-
Availability99.8%

Related models