AI Weekly Malaysia

Summaries

Short AI and tech summaries with source links, signal scores, and why each update matters for builders, founders, and Malaysian tech workers.

Reset

Showing 1-1 of 1 results

DateProviderScoreSummary
07 Oct 2026, 4:18 AMSimon Willison6.0 Introducing Mistral Large 4: Le chonk

Mistral released a preview of Mistral Large 4, a 1-trillion-parameter model with 49 billion active parameters, trained on Mistral's own cluster of 3,800 NVIDIA Grace Blackwell GPUs and available now only through their API. The preview exposes just two reasoning levels, "none" and "high", and Mistral promises open weights at the end of this month. On Artificial Analysis it scores 38, behind DeepSeek 4.1 Flash (a 552B model), a large jump from Mistral Large 3's score of 9 in December, though Simon Willison describes it as roughly six months behind the frontier.

Why: If you self-host or care about open weights, this is an API-only preview today, so any evaluation has to wait for the end-of-month weight release — don't plan deployments on the API tier unless you're fine with a hosted-only dependency. The two-level reasoning switch (none vs high) is unusually coarse: the "high" pelican test used fewer output tokens (2,717) than "none" (3,275), so you can't assume "high" costs more output tokens when budgeting. Compared with DeepSeek 4.1 Flash scoring higher at 552B, the practical question is whether a 1T/49B-active MoE gives you enough quality per dollar to justify swapping out your current model.

Top