
When you purchase through links on our site, we may earn an affiliate commission. Here’s how it works .
Mistral has released Mistral Large 4, scoring 38 on the Artificial Analysis Intelligence Index; France is back to having the most intelligent model from outside the US and China
@MistralAI has released Mistral Large 4 in Research Public Preview, with plans to release the weights of the 1T parameter (49B active) model at the end of October. It achieves 38 on the Artificial Analysis Intelligence Index, comparable to GPT-6 Luna (max, 38) and DeepSeek V4.1 Flash (max, 39). It also achieves 50% on the Artificial Analysis Cyber Index, level with GLM-5.3-Flash and ahead of models such as Kimi K3 and DeepSeek V4.1 Flash (max).
Key benchmarking results for Mistral Large 4 Preview:
➤ Most intelligent model from outside the US and China: Mistral Large 4 Preview scores 38 on the Intelligence Index, comparable to DeepSeek V4.1 Flash (max, 39) and GPT-6 Luna (max, 38). This makes it the most intelligent model from outside the US and China, ahead of countries such as South Korea and the United Arab Emirates
➤ Level with GLM-5.3-Flash on cyber defense capability: Mistral Large 4 Preview scores 50 on the Artificial Analysis Cyber Index, level with GLM-5.3-Flash (50) and behind MiMo-V2.6-Pro (56). Once its weights are released, it will rank among the top three open weights models on the Cyber Index. Its strongest result is on CyberGym-E2E-AA, where it scores 82%, ahead of MiMo-V2.6-Pro (79%) and GPT-6 Luna (max, 78%)
➤ Over 4x the Cost per Task of similar-intelligence open weights models: Mistral Large 4 Preview costs $1.13 per Intelligence Index task with standard pricing of $1.36/$4.18 per 1M input/output tokens, with $0.14 per 1M cached input tokens. For the first two weeks, Mistral Large 4 Preview will be served at a 50% launch discount, bringing its Cost per Task down to $0.57. This is still more costly than GLM-5.3-Flash ($0.25) and DeepSeek V4.1 Flash (max, $0.27)
➤ Strong document and image reasoning: Mistral Large 4 Preview scores 19% on GDP.pdf, on par with MiMo-V2.6-Pro (19%) and behind Kimi K3 (22%). This is an 18-point improvement from Mistral Large 3, partly driven by improvements in their API, which now accepts 100 images per request, up from 8 for previous Mistral models
Key model details:
➤ Context Window: 512k tokens
➤ Multimodality: Text and image input, with text output
➤ Pricing: $1.36/$4.18 per 1M input/output tokens ($0.14 per 1M cached input tokens), with 50% off for the first two weeks ($0.68/$2.09)
➤ Availability: Research Public Preview on Mistral's API, with open weights planned for the end of October October 6, 2026
ML4 scored 38 on the Artificial Analysis Intelligence Index v4.3.2, which AA says makes it the “most intelligent model from outside the US and China." At least five Chinese open models scored higher on the index: DeepSeek V4.1 Flash (max) at 39, GLM-5.3-Flash at 42, Kimi K3 (max) at 44, GLM-5.3 (max) at 45, and MiMo-V2.6-Pro at 46. While MiMo-V2.6-Pro is comparable in total size to ML4, GLM-5.3-Flash is about one-third the size at 320B parameters total and 18B versus ML4’s 49B–52B active. Unlike the rest, ML4 isn’t open yet: Mistral says its weights are due “by the end of the month.”
Mistral Large 4 vs. Chinese open models on Artificial Analysis' Intelligence Index (Oct. 6) Model
Cost is a different story . ML4 comes in at $1.13 per AA task at list price, or $0.57 at the 50% discounted launch price. This compares to $0.13 for MiMo-V2.6-Pro and $0.25 for GLM-5.3-Flash. Although comparatively expensive, ML4 scored what AA calls its “strongest result” and Mistral calls “the highest of any model” with an 81.7% score on CyberGym-E2E-AA, compared with GLM-5.3-Flash’s 74%, from a model whose post-training run Mistral claims is “still in flight.”
While the score is impressive, it is only one of three tests in the Cyber Index; ML4 only scores 16% in DeepsecBench-AA and 51% in CWE-Bench-AA. However, AA’s projection is that, once the weights ship, it “will rank among the top three open weights models on the Cyber Index.” This leaves the door open for ML4 to apply to narrower tasks.
Mistral Large 4 on Artificial Analysis' Cyber Index and its three tests (Oct. 6) Model
OpenAI GPT-6 Astra (max; declined some tasks)
Key considerations
- Investor positioning can change fast
- Volatility remains possible near catalysts
- Macro rates and liquidity can dominate flows
Reference reading
- https://www.tomshardware.com/tech-industry/artificial-intelligence/SPONSORED_LINK_URL
- https://www.tomshardware.com/tech-industry/artificial-intelligence/independent-tests-rank-mistrals-new-trillion-parameter-large-4-the-best-ai-model-outside-the-u-s-and-china-but-chinese-open-weights-still-overcome-europes-best-efforts#main
- https://www.tomshardware.com/my-account
- AMD's EPYC Verano AI host CPU will reportedly use a special SB1 socket
- Elevate your Steam Deck and ROG Ally's SSD storage for just $0.13 per GB
- Upgrade to Wi-Fi 6E or Wi-Fi 7 with these stellar savings
- Musk rejects rumors of TSMC takeover of Terafab — Intel reaffirms 14A node deal as Musk floats cleanroom 'sublease'
- Contain the Chaos: ‘CONTROL Resonant’ Launches on GeForce NOW
Informational only. No financial advice. Do your own research.