Skip to contentSkip to stories

Updated

#Industry news

Oct 8

TodayOct 8Thu4 items

Oct 7

Oct 7Wed
  1. Ars Technica · AIAI score63

    Mistral releases Le Chonk, a 1 trillion-parameter open-weight model

    AIMistral has released Mistral Large 4, nicknamed Le Chonk, a 1 trillion-parameter model it says can be used and customized by anyone. It is in preview, with a final version due by the end of the month, and is optimized for coding and cyberdefense as well as manufacturing, finance, and electrical engineering tasks. Mistral claims it is the most capable open-weight model developed outside China and says it was trained from scratch rather than through distillation.

Oct 6

Oct 6Tue
  1. Artificial Analysis ArticlesAI score54

    Mistral Large 4 Preview scores 38 on Artificial Analysis Intelligence Index

    AIMistral has released Mistral Large 4 in Research Public Preview, with open weights for the 1T parameter (49B active) model planned for the end of October. It scores 38 on the Artificial Analysis Intelligence Index, comparable to GPT-6 Luna (max, 38) and DeepSeek V4.1 Flash (max, 39), and 50 on the Cyber Index. The source calls it the most intelligent model from outside the US and China, and notes costs of $1.13 per Intelligence Index task at standard pricing.

Oct 5

Oct 5Mon
  1. meng shaoAI score47

    Reflection previews Beam, a 501B-parameter open agentic model

    AIReflection AI previewed Beam, an MoE open model with 501B total and 23B active parameters, claiming 3–4x better inference efficiency than GLM 5.2. The model was pretrained from scratch on 23.8T tokens in four weeks, and its RL run used 10,500 GB300 GPUs over four weeks, which the post describes as possibly the largest publicly recorded. Reflection positions Beam as a workhorse open model for enterprises, governments, and developers, with full weights due this month.

Oct 3

Oct 3Sat

Sep 30

Sep 30Wed
  1. Artificial Analysis ArticlesAI score39

    Upstage Releases Solar Mini 4 Reasoning Model, Scoring 24 on Intelligence Index

    AIKorean AI lab Upstage has released Solar Mini 4, a proprietary reasoning model that scores 24 on the Artificial Analysis Intelligence Index with 35B total and 3B active parameters. It is priced at $0.10/$0.40 per 1M input/output tokens and has a 1M-token context window, but averages 7.1 minutes per task due to heavy output token use. Its weights are not released, and its size cannot be independently verified.

Sep 28

Sep 28Mon
  1. Mike KriegerAI score67

    Anthropic releases Claude Sonnet 5.5, 30% faster and up to 30% cheaper than Sonnet 5

    AIAnthropic has released Claude Sonnet 5.5, the second model in the Claude 5.5 family. The company says it is more than 30% faster than Sonnet 5 and costs up to 30% less for most work.

    Why it matters: The post gives concrete speed and price changes against Sonnet 5, which helps readers judge whether the upgrade fits their workloads and budgets.

Sep 10

Sep 10Thu
  1. Mustafa SuleymanAI score38

    Microsoft's MAI models top rankings across image, transcription, and cyber benchmarks

    AIMustafa Suleyman says Microsoft launched eight new models in two months, with five debuting at #1 on their respective leaderboards. MAI-Transcribe-2 is described as the fastest, most accurate, and cheapest transcription model, while MAI-Image 2.5, 2.6, and 2.6 Flash debuted at #1 on the Artificial Analysis leaderboard. MAI-Code-1.1-Flash, launched in GitHub Copilot four weeks ago, now accounts for a third of small-model traffic there.

Sep 4

Sep 4Fri

Sep 1

Sep 1Tue

Aug 25

Aug 25Tue

Aug 12

Aug 12Wed
  1. Liquid AI NewsletterAI score46

    Liquid AI releases LFM2.5-2.6B model for on-device agentic workloads

    AILiquid AI has released LFM2.5-2.6B, a model optimized to run agentic workflows entirely on-device without cloud escalation. The company said it is designed for high-volume agentic tasks and chained workflows while staying on-device. Separately, Liquid AI and MacPaw announced a long-term partnership to co-develop local AI technology for Mac, with LFMs running on Apple silicon through MacPaw's Elix inference engine.

  2. Michael TruellAI score62

    Grok 4.6 is released with gains on agentic and knowledge-work benchmarks

    AIGrok 4.6 is released as a significant improvement over Grok 4.5 at the same price, according to the announcement. The author says it is significantly better at difficult tasks and knowledge work, combining Opus-class intelligence and polish with low cost and high speed. A comparison table shows Grok 4.6 High scoring 61 on the AA Intelligence Index, versus 56 for Grok 4.5 High, and 1753 on GDPval-AA v2, versus 1526.

Jul 30

Jul 30Thu

Jul 8

Jul 8Wed
  1. Aman SangerAI score40

    Very excited for this model.

    AIIt’s an enormous improvement over composer 2.5 and trained entirely from scratch. It’s been a pleasure working with the SpaceXAI team on it. Even more excited by the slope of the effort and the models to come.

  2. Michael TruellAI score57

    Cursor and SpaceXAI release Grok 4.5, a coding-focused model

    AICursor co-founder Michael Truell announced Grok 4.5, a model trained with SpaceXAI that the post calls Opus-class, fast, and low cost. He says it is a significant step up over Composer 2.5 and has become the daily driver for many on the Cursor team. A benchmark table shows Grok 4.5 at 83.3% on Terminal-Bench 2.1 and 78.0% on SWE-Bench Multilingual, with the post saying more releases will follow.

May 18

May 18Mon

Nov 18, 2025

Nov 18, 2025Tue