Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri
  1. Mistral AI · new models on Hugging FaceOfficialAI score47

    Mistral releases Voxtral Mini 4B Realtime Arabic speech-to-text model

    AIMistral releases Voxtral Mini 4B Realtime Arabic, a streaming speech-to-text model for Arabic dialects and Modern Standard Arabic under the Apache 2.0 License. The model has about 4.4 billion parameters, is fine-tuned from Voxtral-Mini-4B-Realtime-2602, and reaches an average 8.82% Character Error Rate across seven Arabic benchmarks at a 480 ms transcription delay. It can be run with vLLM or Transformers 5.2.0 or later.

  2. Mistral AI · new models on Hugging FaceOfficialAI score32

    Mistral releases LIDstral-Arabic, a language and dialect classifier for Arabic-script text

    AIMistral AI has released LIDstral-Arabic, a fast classifier that identifies Modern Standard Arabic, Arabic dialects, and non-Arabic languages written in Arabic script across 51 classes. On Moroccan Darija, it scores 88.67% F1, versus 72.65% for LahjatBERT ALDi CL and 71.28% for GlotLID v3, across 84,870 evaluation examples. The model runs on CPU and is available from a private Hugging Face repository under Apache 2.0.

  3. Andrew CurranXAI score14

    Andrew Curran says OpenAI's unnamed Aeon model solved most recent math results

    AIAndrew Curran says OpenAI's unnamed internal model, which he calls Aeon, reached most of the recent math results in one shot from a single prompt. He says the model did not exist before the end of August, is still training, and that each result used about three hours of thinking across 4,000 problems OpenAI gave it. He argues models can now devise novel self-improvement methods, which he says is what labs are now reacting to internally.

    Image from @AndrewCurran_'s post
  4. Codex · GitHub ReleasesOfficialAI score13

    Codex 0.162.1 fixes TUI crash and startup failures

    AIOpenAI releases Codex 0.162.1, a bug-fix update for the Codex CLI. It fixes a TUI crash when asynchronous questions contain multiple lines, preserving line breaks and complete hyperlink destinations. It also fixes startup failures caused by mismatches between a running background server's feature settings and CLI defaults.

  5. Replit ⠕OfficialAI score4

    Replit says it backs partnerships that go beyond ads

    AIReplit says it believes in partnerships that create content rather than just ads, describing a creator economy where both partners benefit. The post gives no specific partners, deals, or figures.

  6. elvisXAI score62

    StepFun's Step 5 Preview targets long coding agent runs

    AIElvis Saravia says he has tested StepFun's Step 5 Preview as a coding agent since early access and found that it checks its own work and stops when tasks are done. The post says the model is built for engineering tasks such as bug fixing, multi-file features, and refactoring, plus frontend generation and financial report output.

    Image from @omarsar0's post
  7. Bloomberg · TechnologyNewsAI score18

    Deutsche Bank says bearish bond narrative has gone too far

    AIDeutsche Bank strategist George Saravelos says negative sentiment on fixed-income assets this year has overshot. He argues the market underprices the risk of an artificial-intelligence blowup that could trigger a rush into bonds.

  8. Microsoft CopilotOfficialAI score30

    Microsoft unveils new Copilot Home combining Chat and Cowork

    AIMicrosoft says the new Copilot Home brings Chat and Cowork into one experience, so users can move from thinking to doing without losing context. Microsoft Copilot EVP Jacob Andreou explains why Home is the new starting point for work.

    Video from @MSFTCopilot's post
  9. LangChain BlogOfficialAI score40

    LangChain adds emoji reactions to Managed Deep Agents Slack channels

    AILangChain's Managed Deep Agents v0.9 adds a reactions attribute for Slack channels that accepts either an emoji string or a callable returning one. The article shows a function that returns a bug emoji when a message contains "broken" and eyes otherwise. It also shows a TypeSafe Classifier that picks from a seven-emoji vocabulary and falls back to eyes below 25% confidence.

  10. a16z NewsBlogAI score40

    a16z leads investment in TypeSafe AI, maker of Jev System One model

    AIa16z says it is leading an investment in TypeSafe AI, whose Jev model hands decisions to code as typed values and reached 1 trillion tokens generated three days after launch. The company says Jev costs roughly 1/100 to 1/500 of frontier models and runs 100x faster on classification tasks at comparable accuracy. TypeSafe says 25% of the Fortune 500 have integrated Jev.

  11. a16z NewsBlogAI score33

    Prediction markets show no partisan bias in election pricing, NBER study finds

    AIA preliminary NBER working paper by Prof. Zitzewitz, covering over 100 years of prediction markets, finds no statistically significant bias by political affiliation, gender, race, or age. The only exception is non-US elections, where markets appear to overrate right-leaning candidates, but that result is not statistically significant. Separately, prediction markets had Flávio Bolsonaro's Brazilian presidential rise about three weeks before his first-round win.

  12. 🚨 AI News | TestingCatalogXAI score62

    Anthropic moves dynamic workflows in Claude Managed Agents into public beta

    AIAnthropic has expanded dynamic workflows in Claude Managed Agents into a public beta, according to Testing Catalog. Users can configure their agents for multiagent orchestration, with Claude planning and operating a fleet of agents to achieve a goal. The post also links a video from Anthropic's ClaudeDevs account, which the author describes as a new SWE norm.

    Video from @testingcatalog's post
  13. Vaibhav (VB) SrivastavXAI score4

    OpenAI rolls out invites for DevDay Exchange Berlin, Paris, and London

    AIOpenAI says invites for its DevDay Exchange events in Berlin, Paris, and London are rolling out now. Recipients are asked to register as soon as possible to secure a spot. People still waiting can reply with their city, what they are building or exploring, and why they want to attend.

    Image from @reach_vb's post
  14. merveXAI score28

    Hugging Face lets agents train Qwen3.8-27B on Nebius GPUs

    AIHugging Face launches an arena where users bring their own agent, which gets Nebius GPUs to build RL environments that improve Qwen3.8-27B across eight domains. The arena runs on PostTrainArena from BenchFlow, with compute from Nebius. Setup requires only a few steps through the linked OpenEnv Arena space.

    Video from @mervenoyann's post
  15. ClaudeDevsOfficialAI score60

    Claude Managed Agents adds dynamic workflows in public beta

    AIAnthropic's ClaudeDevs account announces that dynamic workflows for Claude Managed Agents are now available in public beta. The feature is a new type of multiagent orchestration in which a lead agent writes a plan that runs across many agents in phases, then combines their results at the end.

    Why it matters: The post describes how a lead agent plans work across many agents in phases and merges their results, a structure useful for understanding complex agent orchestration.

    Video from @ClaudeDevs's post
  16. Perplexity DevelopersOfficialAI score34

    Perplexity releases cookbook for a browser agent using the Decisions API

    AIPerplexity Developers says its new cookbook builds a browser agent that sends a screenshot and questions to pplx-decider-v1.1-27b through the Decisions API, which accepts text and image inputs. The developer's code converts the returned probabilities into clicks, scrolls, and stops.

  17. LangChainOfficialAI score16

    LangChain opens Interrupt session recordings on-demand

    AILangChain says its Interrupt archive is available, with every session watchable on-demand at Context from the quoted post: LangChain simplified agent authentication, memory, and channels, and added web search as a pre-built tool for Managed Deep Agents.