Introducing Qwen1.5
Key point
Qwen1.5, the next-generation model in the Qwen series, has been released with various sizes and improved performance.
Details
Qwen1.5, the next-generation version of the Qwen series, has been released. This update focuses on both improving model performance and optimizing the developer experience.
It offers Base and Chat models in various sizes ranging from 0.5B to 110B, and also includes MoE (Mixture-of-Experts) models. All models support a context length of up to 32,768 tokens.
Key features include the following:
- Integrated into Hugging Face transformers, allowing easy use without needing to set
trust_remote_code. - Significantly enhanced Alignment with human preferences and multilingual capabilities.
- Support for various quantized models including Int4, Int8 GPTQ, AWQ, GGUF.
In terms of performance, Qwen1.5-72B outperforms Llama2-70B across all benchmarks, demonstrating outstanding capabilities in language understanding, reasoning, and mathematics. In addition, the smaller models under 7B also show strong performance that can compete with existing excellent small models.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.