AI Briefing
KO

Mistral Small 3.1

·2025.03.17 09:00

Key point

Mistral released Mistral Small 3.1, an open-source model with multimodal capabilities and a 128k context.

Details

Mistral Small 3.1 is a model that builds on the existing Mistral Small 3, extending text performance, multimodal understanding, and a context window of up to 128k tokens. It shows better performance than similar models such as Gemma 3 and GPT-4o Mini, and provides fast inference speed of 150 tokens per second.

This model is an open-source model distributed under the Apache 2.0 license, offering the following key features:

  • Lightweight: Can be run on a single RTX 4090 or a Mac with 32GB RAM, making it suitable for on-device use.
  • Fast response: Optimized for real-time conversational services such as virtual assistants.
  • Low-latency function calling: Enables rapid execution in automated workflows and agentic environments.
  • Domain-specific fine-tuning: Can be customized into expert models for specific fields such as legal, medical, and technical support.

Currently, Base and Instruct models can be downloaded via Hugging Face, and are available through the Mistral AI API and Google Cloud Vertex AI. Support on NVIDIA NIM and Microsoft Azure AI Foundry is planned for the future.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.