No. 80
2026-10-07
yapay zeka nabzı — ham değil, demlenmişthe AI pulse — brewed, not raw
 AI KRİTİKAI CRITIQUE
Görsel, büyük ve endüstriyel bir odada uzun bir sıra halinde dizilmiş koyu renkli sunucu dolaplarını, gri ve mavi tonlarında stilize bir çubuk grafik ile geometrik şekillerin üzerine bindirilmiş halini göstermektedir.

Mistral Large 4, Çin dışı en iyi açık model iddiasıyla test edildiMistral Large 4 tested as the best open model outside China

Habere gitRead the article    Türkçe (otomatik çeviri) English (automatic translation)

ÖzetSummary

Mistral Large 4, Avrupa'da 3.800 Grace Blackwell GPU'su üzerinde eğitilen 1.050 milyar parametreli açık ağırlıklı bir modeldir. 6 Ekim 2026'da önizleme olarak yayımlanan model, 'Çin dışındaki en yetenekli açık model' unvanını iddia ediyor. Bağımsız kıyaslamalarda model, açık ağırlıklı modeller arasında 8. sırada yer alıyor. Altı Çinli modelin gerisinde kalmasına rağmen, en iyi ABD'li açık modelden 12 puan önde. Siber güvenlik puanları ise nispeten daha güçlü.Mistral Large 4 is a 1.05-trillion-parameter open-weight model trained on 3,800 Grace Blackwell GPUs in Europe. Released as a preview on October 6, 2026, the model claims the title of 'the most capable open model outside of China.' In independent benchmarks, the model ranks 8th among open-weight models. While it trails six Chinese models, it leads the best US open model by 12 points. Cybersecurity scores are relatively stronger.

Neden ÖnemliWhy it matters

Modelin konumu, Avrupa yapay zekasının Çinli açık ağırlıklı liderlerle arasındaki farkı onlara bağımlı olmadan kapatıp kapatamayacağını test etmesi açısından önemli. Mistral, kendi platformunda Z.ai'nin GLM-5.3 modelini barındırıyor. Bu durum, şirketin pazarlama stratejisiyle gizlemeye çalıştığı bir bağımlılığı gözler önüne seriyor. Fiyatların yarıya indirilmesi ve lisansın hâlâ açıklanmamış olması, ticari ve hukuki tablonun henüz netleşmediğine işaret ediyor. Bu nedenle 'açık' iddiası en iyi ihtimalle geçici nitelikte.The model's positioning is significant as it tests whether European AI can close the gap with Chinese open-weight leaders without relying on them. Mistral hosts Z.ai's GLM-5.3 model on its own platform, exposing a dependency the company's marketing strategy attempts to obscure. The halving of prices and the fact that the license remains unannounced indicate that the commercial and legal landscape is not yet settled. Therefore, the 'open' claim is, at best, temporary.

Öne ÇıkanlarHighlights

  • Artificial Analysis, Large 4'ü 38 puanla değerlendirirken MiMo-V2.6-Pro 46, GLM-5.3 ise 45 puan alıyor. Altı Çinli açık model bu modelin önünde.Artificial Analysis rates Large 4 at 38 points, while MiMo-V2.6-Pro scores 46 and GLM-5.3 scores 45. Six Chinese open models rank ahead of this model.
  • Mistral'in kendi DeepSWE grafiğinde Kimi K3 (%68) ve DeepSeek V4.1 Flash (%74,2) yer almıyor. Her iki model de Large 4'ün %61,7'sini geride bırakıyor.Mistral's own DeepSWE chart omits Kimi K3 (68%) and DeepSeek V4.1 Flash (74.2%). Both models outperform Large 4's 61.7%.
  • Siber Endeks: 49,5 puan. 18 model arasında 5. sırada. Mistral'in pazarlama stratejisinin bu modele en çok lehine olduğu alan.Cyber Index: 49.5 points. Ranked 5th among 18 models. This is the area where Mistral's marketing strategy benefits the model the most.

EleştiriCritical take

Mistral'in 'Çin dışındaki en yetenekli açık model' iddiası, Güney Kore'nin Motif 3 modelini (34 puan, sadece 4 puan geride) hariç tuttuğunuzda ve ağırlıkları, lisansı ile nihai RL ayarlı sonuçları hâlâ beklenmede olan bir önizlemeyi kabul ettiğinizde teknik olarak savunulabilir. Kendi grafiklerinde hangi Çinli modellerin yer alacağını seçerek yapılan bu seçici kıyaslama, 'ciddi bir fark' anlatısını önemli ölçüde zayıflatıyor.Mistral's claim of being 'the most capable open model outside of China' is technically defensible only if you exclude South Korea's Motif 3 model (34 points, just 4 points behind) and accept a preview whose weights, license, and final RL-tuned results are still pending. This selective benchmarking, achieved by choosing which Chinese models appear in its own charts, significantly weakens the narrative of a 'serious gap.'

Bu kritik, yerel yapay zeka modeli Qwen3.8-27B tarafından yazılmıştır. Haberi ben seçtim; kritik metnini yayımlanmadan önce tek tek denetlemiyorum.This critique was written by the local AI model Qwen3.8-27B. I chose the story; I do not review each critique before it is published.
Bültene dönBack to the issue