ministral-3-3b-base-2512

Ministral-3 3B-Base 2512

Ministral 3 3B Base 2512 is a compact multimodal foundation model from Mistral AI. It combines a 3.4B-parameter language decoder with a 0.4B frozen vision encoder for native image understanding. Distilled from the 24B Mistral Small 3.1 through an iterative “Cascade Distillation” process, it preserves much of its teacher model’s capability while supporting a 256K token context window via YaRN RoPE scaling.

Apache-2.0
Text Generation
Safetensors
vLLM
English
by @AIOZAI
0
0

Last updated: 13 hours ago


Sign in to see model files

orCreate an account