AI Briefing
KO

Mistral Small 3

·2025.01.30 09:00

Key point

Mistral AI has released Mistral Small 3, a 24B-parameter model optimized for low-latency performance.

1 / 2

Details

Mistral AI has released Mistral Small 3, a 24B-parameter model optimized for low-latency performance, under the Apache 2.0 license. This model can compete with larger models such as Llama 3.3 70B or Qwen 32B, and is a strong open-source alternative that can replace GPT-4o-mini.

Mistral Small 3 is designed to handle 80% of generative AI tasks that require strong language understanding and instruction-following ability along with very low latency. In particular, by reducing the number of layers compared to competing models, it shortens forward pass time, delivering more than 3x the speed of Llama 3.3 70B instruct on the same hardware.

Key performance metrics are as follows:

  • Achieves over 81% accuracy on MMLU
  • Latency performance of over 150 tokens/s
  • Model size optimized for local deployment

This model was trained without using RL (reinforcement learning) or synthetic data, making it different in nature from models like Deepseek R1. Instead, it is well suited for use as a strong base model for building reasoning capabilities.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.