Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. Arena.aiOfficialAI score30

    Arena raises Series B led by Lightspeed and Khosla Ventures

    AIArena has announced a Series B round co-led by Lightspeed and Khosla Ventures, with participation from Salesforce Ventures, 01 Advisors, Dell Technologies Capital, and Endeavor Catalyst. Existing investors Andreessen Horowitz, Felicis, a16z's AMP, QuantumLight, and The House Fund also supported the round.

  2. Arena.aiOfficialAI score55

    Arena raises $200M Series B and launches Alignment Index for AI agents

    AIArena announced a $200M Series B at a $3.1B valuation and released its Alignment Index, a benchmark measuring agent safety and alignment. The index is built from 90K+ real-world agent sessions across 27 models and tracks Unauthorized Action, False Attribution, and Deceptive Completion. OpenAI's GPT-6.1-Sol leads with a score of 87.9, ahead of Claude-Opus-5.5 at 83.2 and Grok-4.7 at 82.7.

    Video from @arena's post
  3. Arena.aiOfficialAI score22

    Arena reports GPT-6 model variants' false attribution rates

    AIArena found that some models misquote users while others credit users with others' work in false attribution cases. GPT-6 Luna and Astra rarely misquoted users, at 15.6% and 28.6%, but often misattributed statements, at 53.1% and 48.2%. Sibling model GPT-6 Sol had the highest rate of misstating the user's history, at 23.5%.

    Image from @arena's post
  4. OpenAI NewsOfficialAI score26

    How Oracle turns days of work into minutes with ChatGPT and Codex

    AIOracle is using ChatGPT Work and Codex to turn specialist knowledge into fast, repeatable workflows across recruiting, engineering, and operations. The source does not provide figures, timelines, or specific results beyond the headline's claim that days of work can take minutes.

  5. The Robot ReportNewsAI score38

    Helm.ai reports $70M in signed commercial contracts for its physical AI foundation models

    AIHelm.ai said it signed $70 million in commercial contracts for its foundation models for physical AI over 12 months, spanning global automotive OEMs, Tier 1 suppliers, and industrial automation companies. The Redwood City, Calif.-based company said it has projects bound for production in autonomous vehicles, mining, and construction, and that it is on a path to break even. CEO Vladislav Voroninski said its models are trained on unsupervised "deep teaching" and are environment-agnostic.

  6. SantiagoXAI score40

    Seedance 2.5 tops evaluation of world models for physical consistency

    AISantiago says physical consistency is the most important and hardest feature of a world model, and that many generated videos show objects defying gravity. He reports that Seedance 2.5 is currently the best among the evaluated world models. The post links to a physics evaluation benchmark in which eight video world models reached a top score of 57.76/100.

  7. SiliconANGLE · AINewsAI score22

    Willow picks CoreWeave for AI model training and forward-deployed support

    AIWillow Care Inc., maker of the AI dictation app Willow Voice, chose CoreWeave for its forward-deployed support rather than compute alone, according to co-founder and CTO Lawrence Liu. Liu said CoreWeave's reinforcement learning infrastructure lets Willow focus on eval alignment, while Willow fine-tunes its own speech recognition model and pairs it with a compact post-processing LLM. He said inference demand is growing faster than training as dictation use climbs.

  8. GoogleOfficialAI score40

    Google AI estimates gestational age within four days in clinical study

    AIIn a prospective clinical study, Google's models pinpointed gestational age within 4 days of accuracy. The company says that precision could meaningfully affect clinical care, and that extending such tools to low-resource settings could help reduce maternal deaths and close care gaps worldwide.

    Image from @Google's post
  9. Boris PowerXAI score34

    Boris Power says progress on a result has been remarkable

    AIBoris Power, who owns OpenAI's account, praised the pace of progress on a result he called remarkable. The context from @0xdoug reports a PR merged and a tightened bound from κ = 2⁻¹⁸² to κ = 2⁻¹⁵, a community effort across contributors.

  10. Leandro von WerraXAI score70

    Carbon-A open model and database predict 566 million gene candidates across 22,617 species

    AICarbon-A is an open model that predicts gene locations directly from DNA, and it has been used to annotate genomes from over 22,000 species. The release includes a database of 566 million gene candidates, about 16 times the gene annotations in the RefSeq dataset. Wet-lab RNA experiments supported 239 candidates missing from RefSeq across cats, Syrian hamsters, chickens, and Arabidopsis.

    Why it matters: The source ties an open gene-annotation model to specific wet-lab checks and gene counts, helping readers judge how far its predictions extend beyond well-studied genomes.

  11. Don't Worry About the Vase (Zvi Mowshowitz)BlogAI score62

    AI #189: New Math covers OpenAI's math results and Claude's new pricing

    AIOpenAI reportedly posted solutions to 90 of the top 500 open math problems, using an average of three hours of Pro-level compute per question. Anthropic released Claude Haiku 5.5 at $0.10 input and $0.50 output per million tokens, and the author says Jay Clayton was named AI Czar to head a new taskforce.

  12. Thomas WolfXAI score67

    Carbon-A open model and database find 566 million candidate genes across 22,617 species

    AIThomas Wolf says Carbon-A, an open model that finds genes directly in DNA, has been released with a database of 566.34 million candidate genes across 22,617 species. The team reports wet-lab validation of several new genes in cats, chickens and arabidopsis, and RNA evidence for 239 genes missing from reference annotations of common species.

    This story has a top pick“Carbon-A open model and database predict 566 million gene candidates across 22,617 species”

  13. elvisXAI score32

    Drama 3 voice model offers fine-grained tone and emotion control

    AIFish Audio's Drama 3 voice model lets users direct tone, emotion, and pacing in plain language, and can shift emotion mid-sentence. The poster, who found it remarkably effective in testing, says the control over delivery is unlike anything previously seen. A preview is available through the API as drama-3-preview.

  14. clem 🤗XAI score22

    Hugging Face shares a Space to make your own Reachy Mini dance

    AIClément Delangue of Hugging Face points users to a Hugging Face Space from Pollen Robotics where they can create a dance for Reachy Mini. The post gives no further details about how the dance tool works.

  15. ZyphraOfficialAI score3

    Zyphra, an open superintelligence company in San Francisco, is hiring

    AIZyphra describes itself as an open superintelligence research and product company based in San Francisco, aiming to build human-aligned AI that helps individuals and organizations reach their fullest potential. The post invites applicants to join the company through its job listings.

  16. ZyphraOfficialAI score34

    Zyphra's lossless method cuts communication for MoE expert routing

    AIZyphra says its approach is lossless: the same tokens still reach the same experts, with unchanged architecture, routing decisions, and training objective. By reorganizing where experts and tokens live, it reduces the communication needed to perform the same computation.

  17. ZyphraOfficialAI score38

    Zyphra reports up to 2.63x faster MoE token exchange in Megatron-LM

    AIZyphra reports that its MoE training optimizations speed up token exchange by 1.16x to 2.63x and full training steps by up to 1.41x in Megatron-LM on 8 to 64 GPUs. The gains are largest when each token uses more experts and those experts span several nodes.

  18. ZyphraOfficialAI score18

    Token shuffling routes tokens to predicted experts without extra network traffic

    AIZyphra reports that expert routing in mixture-of-experts models is predictable across layers, since the experts a token uses in one layer indicate which it will need next. Its token shuffling method moves each token to the GPU holding those experts within a transfer that already runs after attention, adding no network traffic.

    Image from @ZyphraAI's post
  19. ZyphraOfficialAI score32

    MoE training spends 45-60% of step time on cross-node token exchange

    AIIn Zyphra's runs, MoE token exchange between experts consumed 13-24% of step time on one node and 45-60% across four nodes. Because experts are spread across GPUs and nodes, tokens must be sent to their experts and returned, making this communication a major training cost as models scale.

    Image from @ZyphraAI's post
  20. Andrew CurranXAI score9

    Claude joke: planet named after Opus 5.5 and Fable 5.1 usage

    AIAndrew Curran jokes that a newly discovered planet should be named Claude, since he primarily used Opus 5.5 and Fable 5.1. The post is light commentary tied to a quoted thread describing a planet hidden in NASA telescope data for seven years.

  21. siddharthXAI score22

    Databuddy, a YC F26 analytics platform with an AI analyst, tried free

    AIA developer says they tried Databuddy, an analytics tool from the new YC batch, for free on a small game they built. The tool tracks visitor data through conversions and includes an AI analyst that explains user behavior. The post links to databuddy.cc, and the background post says Databuddy is part of YC F26.

    Video from @buildwithsid's post
  22. Allie K. MillerXAI score18

    Meta adds a new icon to Instagram profiles, expects millions more users

    AIMeta is reportedly placing a new icon on Instagram's profile page, following the same approach it used to promote Reels and Threads. The post predicts millions of additional users within the next week, though it provides no figures or details to support that estimate.

    Image from @alliekmiller's post
  23. Nathan LambertXAI score10

    AI news grows bizarre, raising uncertainty about what comes next

    AINathan Lambert says the strange state of AI news, including kids vibe hacking with Claude Code and governments vibe governing, greatly increases uncertainty about what is coming. He calls this meme-heavy AI timeline one of the scariest developments.

  24. The Verge · AINewsAI score52

    Google's experimental AI Edge Foresight transcribes meetings fully offline on Mac

    AIGoogle has released AI Edge Foresight, a free experimental note-taking app that transcribes meetings and audio files entirely offline on macOS. It runs on the on-device EmbeddingGemma 2 model and turns shorthand notes into polished notes based on the transcript. Google says files, meeting audio, and notes never leave the computer, and the app is currently optimized only for Macs with Apple Silicon.