Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. Vaibhav (VB) SrivastavXAI score34

    Codex Auto-review now free for ChatGPT sign-in users

    AIOpenAI's Codex "Approve for me" mode uses a separate Auto-review agent to check actions needing approval, such as running commands outside the sandbox or accessing extra files and network resources. It reduces approval prompts during long tasks while keeping sandbox protections, and it is now free with ChatGPT sign-in without drawing from plan usage.

    Image from @reach_vb's post
  2. OpenRouter · New modelsBlogAI score43

    Google's Nano Banana 2.1 image model improves product recontextualization and editing

    AIGoogle has released Nano Banana 2.1, an image generation and editing model on its Flash tier that succeeds Nano Banana 2 and Nano Banana Pro. The source says it improves product recontextualization, though the excerpt provides no further details on benchmarks, pricing, or availability.

  3. Google ResearchOfficialAI score13

    Google sponsors COLM 2026 with 28 papers and 16 workshops

    AIGoogle, as a Diamond Sponsor of COLM 2026, will have researchers from Google Research and Google DeepMind present 28 papers and participate in 16 workshops at the Hilton Union Square in San Francisco from October 6–9. Attendees can visit Google booth #107 to explore work on agentic systems, reasoning, and multimodal AI.

    Image from @GoogleResearch's post
  4. 👩‍💻 Paige BaileyXAI score20

    Google's Gemini apps rank high in new a16z consumer AI rankings

    AIGoogle has five products in a16z's top 50 consumer AI web apps, with the Gemini app at #2, NotebookLM at #9, Google AI Studio at #11, Google Labs at #18, and Antigravity at #44. Paige Bailey highlights Google AI Studio, a developer workbench, reaching the overall consumer top ten. Steven B. Johnson notes Google is the only company with two products in the top ten and four in the top twenty.

  5. ARC PrizeOfficialAI score46

    DeepSeek V4.1 Flash scores 72.9% on ARC-AGI-2 at $0.13/task

    AIDeepSeek V4.1 Flash reaches 72.9% on ARC-AGI-2 at $0.13 per task and 94.5% on ARC-AGI-1 at $0.07 per task, according to ARC Prize verification. Compared with V4 Flash's best scores, it gains 11.5 points on ARC-AGI-2 and 5.5 points on ARC-AGI-1, but costs about 250% more per task.

    Image from @arcprize's post
  6. Aravind SrinivasXAI score13

    Perplexity Decider is the best decision model, per Tetris test

    AIPerplexity's Decider model was ranked best among eight decision models in a Tetris benchmark, according to a quoted post from @alokbishoyi97. The post says Decider consistently placed at the top of the tests, which were run through the Tetris Royale playground.

  7. GranolaOfficialAI score4

    Granola lists four San Francisco events for October with partners

    AIGranola announces four San Francisco events in October, including a GTM Series talk with Together AI's CRO Kai Mak on October 15. Other dates include a pickup basketball game on October 19, a GTM Hackathon with Anthropic on October 20, and a Craft at Speed demo night with Linear on October 22.

    Image from @meetgranola's post
  8. Amazon ScienceOfficialAI score8

    Amazon Science Hosts Live Talks at COLM 2026 Booth

    AIAmazon is presenting live talks at the COLM 2026 conference booth this week, covering super weights in LLMs, multi-agent orchestration, and responsible AI. Meet-the-scientist sessions are also scheduled throughout the week, with the full schedule available via the linked page.

    Image from @AmazonScience's post
  9. Vercel DevelopersOfficialAI score59

    Mistral Large 4 is now available on Vercel AI Gateway

    AIVercel says Mistral Large 4 is live on its AI Gateway, describing it as an open-weight, natively multimodal model that reasons across text and images. Mistral's quoted post says the model has 1T total parameters with 49B active, and that it is available via API today, with open weights due at the end of October.

  10. Google ResearchOfficialAI score31

    Google Earth AI's population model tackles public health data gaps globally

    AIGoogle Research shares results from five partner-driven case studies showing how Google Earth AI's Population Dynamics Foundation Model addresses public health data gaps. The post frames the work as empowering global health research, with details available in the linked blog.

    Image from @GoogleResearch's post
  11. Google ResearchOfficialAI score51

    Google's PDFM location embeddings improve five global public health tasks

    AIGoogle Research reports that Population Dynamics Foundation Model (PDFM) embeddings, built from search trends, mobility, built environment, and weather signals, were tested by partners across five public health tasks. The embeddings improved results in cross-border MMR vaccination coverage, dengue forecasting, postpartum depression screening, and cholera outbreak prediction, and matched census inputs for cardiovascular mortality nowcasting.

  12. WaymoOfficialAI score23

    Waymo unveils silver Ojai with leatherette seats and wireless charging

    AIWaymo introduced a silver version of its Ojai vehicle, featuring leatherette seats, wireless charging, and refreshed interior details. The vehicle will serve riders soon in San Francisco, Los Angeles, and Las Vegas, with more cities to follow over time.

    Image from @Waymo's post
  13. The Next PlatformNewsAI score20

    When One Datacenter Is No Longer Enough: Cisco on Scale-Across AI Networking

    AICisco SVP Rakesh Chopra discusses the "Scale-Across" approach to networking AI training workloads spread across multiple data centers. He describes how Silicon One architecture and Intelligent Collective Networking aim to manage synchronized GPU traffic over long-distance fiber links. The interview covers power efficiency and hardware-accelerated MACsec and IPsec security.

  14. DatabricksOfficialAI score32

    Databricks' Lakebase rearchitects databases to spin up at AI speeds

    AIDatabricks describes Lakebase as a ground-up rearchitecture of databases that can spin them up and shut them down at AI speeds. The company says this suits the many small, fast workloads generated by AI-driven development.

    Image from @databricks's post
  15. merveXAI score72

    Mistral Large 4 will open its weights at the end of October

    AIMistral announced Mistral Large 4, which it describes as a natively multimodal model with 1T parameters and 49B active. Mistral says it is available via API now, with open weights to follow at the end of October, and a Hugging Face page is listed for the release.

    Why it matters: The quoted Mistral announcement gives specific size, activation, and API details, and the open-weights timing matters for teams weighing open model options.

    Image from @mervenoyann's post
  16. Simon WillisonXAI score36

    Mistral's Pelican SVG Test Passes, Tied to Mistral Large 4 Context

    AISimon Willison reports that Mistral can now generate his pelican SVG test, shared via a Markdown SVG renderer. The post links to a rendered result but gives no benchmark or scoring details. Background from Mistral's own announcement describes Mistral Large 4 as a 1T-parameter, natively multimodal model with 49B active parameters, available via API today and with open weights planned for end of October.

    Image from @simonw's post
  17. Thomas WolfXAI score62

    Mistral Large 4 open weights are set for release at end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active, now available via API. Open weights are scheduled for release at the end of October, with a countdown page on Hugging Face showing October 31, 2026.

    Why it matters: The post pairs a Mistral Large 4 announcement with a dated open-weights release, giving a concrete timeline for readers tracking European open models.

    Image from @Thom_Wolf's post
  18. Aravind SrinivasXAI score42

    Perplexity Computer plays real-time StarCraft against itself, Blue wins 2-5

    AIPerplexity's Computer ran two agents playing StarCraft against each other in real time, with the game never paused while each agent thought. Blue, playing with 41 Dragoons, lost the final match 2-5 to Red, which used High Templar and Psionic Storm after Blue failed to scout Red's build. Each agent received only its own fog-of-war-limited game state, and video input was not provided.

    Video from @AravSrinivas's post
  19. Mustafa SuleymanXAI score42

    Daron Acemoglu predicts AI will replace only 5% of human work in 10 years

    AINobel laureate Daron Acemoglu argues in the first issue of The Humanist Review, published by MAI, that AI will replace only about 5% of what humans do over the next decade. He says AI is not yet visible in productivity statistics and projects roughly 1.5% added to GDP over 10 years, and he urges building pro-worker tools that make people better at their jobs.

  20. Liquid AIOfficialAI score14

    Liquid AI partners with Arm on hardware-aware efficient AI models

    AILiquid AI announced a partnership with Arm to bring its efficient AI models, optimized for Arm platforms, into Arm Total Design for Physical AI. The collaboration aims to help accelerate full-stack solutions for real-world deployment, with Liquid AI COO Jeffrey Li featured in the accompanying video.

  21. Nathan LambertXAI score62

    Mistral Large 4 announced as a 1T-parameter multimodal open-weights model

    AIMistral has announced Mistral Large 4, a natively multimodal model with 1T total parameters and 49B active parameters. The company says it is the best open-weights model from the US or Europe on aggregated benchmarks, and that it is available via API today, with open weights due at the end of October.

  22. Mistral AIOfficialAI score30

    Mistral AI points readers to Mistral Large 4 announcement

    AIMistral AI's post links to a news page about Mistral Large 4, but the post text itself gives no details on the model's capabilities, specifications, or pricing. The linked announcement is the only source of further information.

  23. ElevenLabsOfficialAI score49

    ElevenLabs' Eleven v4 and v4 Turbo top Artificial Analysis TTS leaderboard

    AIElevenLabs' Eleven v4 and Eleven v4 Turbo rank #1 and #2 on the Artificial Analysis text-to-speech leaderboard. Per Artificial Analysis, Eleven v4 Turbo leads the Provider Voice arena at an Elo of 1,334, costs $40 per 1M characters, and generates 96 characters per second.

  24. clem 🤗XAI score35

    Mistral Large 4.0 model page appears on Hugging Face

    AIClément Delangue of Hugging Face shared a link to a Hugging Face model page for mistralai/Mistral-Large-4.0-1T05-A52B. The post itself gives no further details about the model's capabilities, release terms, or benchmarks.

    Image from @ClementDelangue's post
  25. LM StudioOfficialAI score44

    LM Studio posts "We're so back" amid Mistral Large 4 news

    AILM Studio posted "We're so back" with no further details in the main post. Quoted context from Mistral AI says Mistral Large 4 has 1T parameters, 49B active, is natively multimodal, and is available via API today, with open weights planned for end of October.

  26. Sophia YangXAI score26

    Reinforcement learning infrastructure scales to tens of thousands of parallel rollouts

    AIThe post describes a reinforcement learning system that autoscales an actor fleet to run tens of thousands of rollouts in parallel with asynchronous training, designed for trajectories of millions of tokens with multiple compactions and low staleness. New methods at both stages reduce off-policy drift, and the setup runs on 3k GPUs producing about 33B tokens per day, with roughly 16B trainable after filtering and masking. Rewards rise across representative environments as the policy learns harder tasks.

    Image from @sophiamyang's post
  27. Yuchen JinXAI score72

    Mistral Large 4 launches as a 1T-parameter multimodal model with open weights due end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active, available via API today. Mistral claims it is the best open weights model from the US or Europe on aggregated benchmarks, with open weights set for release at the end of October. The author quotes this claim and comments that it appears to beat GLM-5.3.

    Why it matters: The quoted announcement gives specific size, multimodal, and deployment claims, with open weights promised later, useful for comparing it against other open models.

  28. OpenRouter · New modelsBlogAI score62

    Mistral Large 4 is listed on OpenRouter with a 1M-token context window

    AIMistral AI's Mistral Large 4 is listed on OpenRouter as a frontier multimodal model accepting text and image input. The listing says it is built for reasoning, coding, and agentic workloads and offers a 1M-token context window. The feed excerpt is truncated, so further details such as pricing or availability are not confirmed here.

  29. Sophia YangXAI score45

    Mistral Large 4 tops benchmarks across cybersecurity, legal, and agentic tasks

    AIMistral Large 4 is a 1T-parameter natively multimodal model with 49B active parameters, which the Mistral account says leads open-weights models from the US or Europe on aggregated benchmarks. The post claims it beats closed frontier models on visual grounding and posts strong results across cybersecurity, legal, and agentic behavior. It is available via API now, with open weights due at the end of October.

    Image from @sophiamyang's post
  30. Julien ChaumondXAI score70

    Mistral Large 4 announced with open weights due end of October

    AIJulien Chaumond reposted Mistral's announcement of Mistral Large 4, a 1T-parameter natively multimodal model with 49B active parameters. Mistral says it is available via API today, with open weights scheduled for release at the end of October, and is working privately with cybersecurity partners.

    Why it matters: The post lays out Mistral Large 4's scale, multimodal design, and availability timeline, which helps readers gauge the open-weights landscape outside China.

  31. Arthur MenschXAI score48

    Mistral Large 4 trained on own compute, RL shows no saturation

    AIArthur Mensch says Mistral trained its model on its own compute, and reinforcement learning shows no sign of saturating. The post accompanies Mistral's announcement of Mistral Large 4, a 1T-parameter natively multimodal model with 49B active parameters, available via API today and with open weights planned for end of October.

  32. Georgi GerganovXAI score29

    Upgrade Qwen3.8-27B to DFlash for extra llama.cpp speed

    AIGeorgi Gerganov says users of Qwen3.8-27B with MTP can get extra speed by switching to DFlash speculative decoding in llama.cpp. The command uses --spec-type draft-dflash with --spec-draft-n-max 7, and it requires the latest llama.cpp v0.6.0.

  33. Guillaume Lample @ NeurIPS 2024XAI score62

    Mistral Large 4 (ML4) is released, with more coming and hiring expanding

    AIGuillaume Lample announced that Mistral's Science team has shipped ML4, which the post links to the Mistral Large 4 news page. He said more is coming soon and that the team is scaling alongside its compute, with hiring open in Europe, the US, and Montreal for frontier open-weight models and large-scale RL systems.

    Why it matters: The post links the ML4 release to a stated hiring push for frontier open-weight models and large-scale RL, which shows where the team is investing next.

  34. Guillaume Lample @ NeurIPS 2024XAI score26

    Mistral model beats GLM 5.3 on STEM, CAD, and finance tasks

    AIOn human evaluation, the model outperforms GLM 5.3 on STEM, CAD, and finance tasks and performs on par on agentic coding. The post is part 5 of a thread, so the model's name and other details come from earlier posts not included here.

    Image from @GuillaumeLample's post