xAI launches Grok 4.7 at bargain prices, but benchmarks reveal a wide gap to Claude and GPT-6
Source: The Decoder (opens in a new tab) · Matthias Bastian
Intel Summary
xAI has released Grok 4.7, positioning the model as a lower-cost option. According to benchmark scores from the Artificial Analysis Intelligence Index, Grok 4.7 scored 46 points, placing it mid-pack behind leading models Claude Fable 5.1 and GPT-6, which each scored 53 points, with a wider performance gap observed in agentic coding tasks.
Why It Matters
The release highlights a trade-off for software teams and enterprise buyers between API pricing and frontier model performance, particularly for complex software engineering and agentic workflows where higher-tier models maintain a measured lead.
Part of an ongoing development
Developing storyIndependent reportingXAI released Grok 4.7
xAI has released Grok 4.7, positioning the model as a lower-cost option. According to benchmark scores from the Artificial Analysis Intelligence Index, Grok 4.7 scored 46 points, placing it mid-pack behind leading models Claude Fable 5.1 and GPT-6, which each scored 53 points, with a wider performance gap observed in agentic coding tasks. Claims are as reported; this summary makes no determination about accuracy or significance.
- Confidence
- Moderate confidence
- Corroboration
- Corroborated
More coverage of this development
- SpaceX launches Grok 4.7 with long-horizon processing, safety upgradesSiliconANGLEIndependent reporting
Organizations & Entities
Related Intelligence
- ReportSame development
SpaceX launches Grok 4.7 with long-horizon processing, safety upgrades
SpaceX has introduced Grok 4.7, its latest large language model, featuring long-horizon processing capabilities and updated safety controls. The launch marks the continued development of the Grok model family following the integration of Elon Musk's AI startup xAI into SpaceX.
SiliconANGLE - DevelopmentDevelopingAlso involving Artificial Analysis
Google DeepMind releases Gemini 3.8 Live audio models
Google DeepMind released Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two new speech-to-speech audio models for developers. The models reportedly lead the Artificial Analysis speech-to-speech leaderboard and are offered at $1.38 per hour of voice conversation to undercut OpenAI's GPT-Live-1. Claims are as reported; this summary makes no determination about accuracy or significance.
2 independent sources - DevelopmentNewAlso involving xAI
Amazon Web Services makes xAI Grok 4.6 available in Amazon Bedrock
Amazon Web Services has made xAI's Grok 4.6 model available in Amazon Bedrock. According to AWS, the model features a 500K token context window and four reasoning effort levels for coding, knowledge work, and long-running agents. Claims are as reported; this summary makes no determination about accuracy or significance.
- DevelopmentNewAlso involving Claude Fable 5.1
Anthropic launched Claude Opus 5.5
Anthropic has launched Claude Opus 5.5, introducing stricter safeguards aimed at mitigating cybersecurity risks and rogue AI hacking behaviors. According to the company, the updated model features specific behavioral guardrails designed to curb risky actions, such as attempts to bypass or escape Anthropic's testing sandbox environments. Claims are as reported; this summary makes no determination about accuracy or significance.