AI Briefing
KO

Qwen2.5-Math: A World-Class Open Source Math-Specialized LLM

·2024.09.19 01:00

Key point

Qwen has open-sourced the Qwen2.5-Math series, which supports CoT and TIR to enhance mathematical problem-solving ability.

1 / 2

Details

Qwen has released the Qwen2.5-Math series, an upgrade of the existing Qwen2-Math. This series consists of Base models (1.5B/7B/72B), Instruct models, and a math-specific Reward Model (72B).

Unlike the previous model, which only used CoT (Chain-of-Thought) for English math problems, Qwen2.5-Math supports both CoT and TIR (Tool-integrated Reasoning) in both English and Chinese. In particular, TIR greatly improves precise calculation and symbolic manipulation ability for complex computation or algorithmic reasoning.

To improve performance, the following technical improvements were applied.

  • Initialized from the Qwen2.5 base model to strengthen language understanding and code generation ability
  • Built the Qwen Math Corpus v2 (utilizing over 1T tokens)
  • Training optimization via Rejection Sampling and GRPO (Group Relative Policy Optimization)

As a result, it achieved dramatic performance improvements over the previous model on major math benchmarks such as GSM8K, MATH, and Gaokao.

This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.

Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.