Pixtral 12B

Mistral vision-language model for image understanding and multimodal chat

BENCHMARKS

Not in the Epoch Capabilities Index. Epoch benchmarks what it can run, so this is a gap in coverage, not a poor result.

RELEASE DATESep 1, 2024
WEIGHTSOpen weights
CONTEXT WINDOW128K
MAX OUTPUT128K
REASONINGNo
TOOL CALLINGYes
INPUTStext, image
OUTPUTStext

Specifications from Models.dev. Availability and limits may differ by provider. Benchmark scores come from Epoch AI under CC BY 4.0.