Grok 4.1
Key point
xAI has released Grok 4.1, which significantly enhances emotional intelligence and creative capabilities.
Details
Grok 4.1 has been released to all users via grok.com, X, and the iOS and Android apps. This model maintains the existing sharp intelligence and reliability while significantly improving creative, emotional, and collaborative interaction capabilities.
xAI utilized Reinforcement Learning infrastructure to optimize the model's style, personality, usefulness, and alignment. In particular, to optimize hard-to-verify reward signals, they introduced a new approach that uses state-of-the-art agentic reasoning models as reward models to autonomously evaluate and iterate on responses.
Key performance metrics are as follows:
- LMArena Text Leaderboard: The Grok 4.1 Thinking (
quasarflux) model took the overall #1 spot with 1483 Elo, while the non-reasoningtensormodel also ranked #2 with 1465 Elo. - EQ-Bench: A benchmark measuring emotional intelligence (EQ) demonstrated improvements in empathy and interpersonal skills.
- Creative Writing v3: The model also recorded high performance in creative writing.
Grok 4.1 showed an overwhelming performance improvement, achieving a high preference rate of 64.78% over the previous model in user preference surveys.
This summary was generated automatically by AI. Check the original for the author's claims and context. Copyright belongs to the original author.
Our guide explains how the AI works. Report summary errors, attribution issues, or removal requests via Contact.