Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

Oct 9Fri
  1. ARC PrizeOfficialAI score42

    ARC Prize 2026 ARC-AGI-2 high score reaches 88.06%

    AITufa Labs posted an 88.06% score on ARC-AGI-2, a new high for the ARC Prize 2026 leaderboard. ARC Prize says a $150K bonus prize, on top of guaranteed prizes, will be split among all teams scoring over 85%.

    Image from @arcprize's post
  2. Ben TossellXAI score10

    Ben Tossell lists AI meeting tools that alert and take notes

    AIBen Tossell lists seven AI meeting tools, including dot, Grok bot, Poke, Instinct, OpenAI's meetings plugin, Granola, and Dia, that alert users to meetings, join Zoom calls, or take notes. He says the list means he should no longer miss meetings, but he says he was late because of the post itself.

  3. Andrew CurranXAI score62

    OpenAI responds to three fired employees' letter on safety and trust

    AIOpenAI's research leaders say they parted ways with Jasmine, Mikita, and Tomek after an investigation found they violated policies on handling sensitive information. The company says the decision was not about raising safety concerns and that it is finalizing contracts with third-party safety assessors, with details to follow in the coming weeks.

  4. Hamel HusainXAI score22

    How to make the team case for investing in evaluations

    AIHamel Husain advises reviewing real user interactions and showing the team the problems found. He says to then fix those problems and show what improved as the case for investing in evaluations.

    Image from @HamelHusain's post
  5. Rohan PaulXAI score46

    Sabi raises $50M to build AI brain-reading baseball cap

    AISabi has raised a $50M seed round led by Vinod Khosla and Accel to build a baseball cap that turns brain signals into AI prompts. The cap can hold up to 100,000 sensors of 1 to 5 millimeters each, feeding a brain foundation model trained on 100,000 hours of labeled neural recordings. Sabi says the cap can predict a user's next 3-4 keystrokes from brain signals before they type them.

    Image from @rohanpaul_ai's post
  6. 🚨 AI News | TestingCatalogXAI score62

    Sabi raises $50M for a wearable brain-reading cap

    AISabi has raised $50M from Khosla Ventures, Accel, Initialized, DST, and Collab Fund, with Kevin Weil also participating. The company is building the Sabi Cap, a non-invasive wearable that aims to read brain signals and turn thoughts into text. Sabi says it designs its own custom chips and neuroimaging sensors, including a non-contact EEG chip fabricated by TSMC.

    Image from @testingcatalog's post
  7. SantiagoXAI score43

    Sabi raises $50 million for a wearable brain-to-text cap

    AISabi has raised $50 million from Khosla Ventures, Accel, Initialized, Kevin Weil, and DST Global to build a wearable brain-computer interface. The company says its TSMC-fabricated chip reads brain signals without touching the scalp, with custom sensors collecting data and its own AI model decoding it into text. The device is a cap rather than an implant.

  8. PixVerseOfficialAI score14

    PixVerse hosts sessions demoing ChatGPT plugin video creation

    AIPixVerse says each session includes a platform walkthrough, a live OpenAI demo showing ChatGPT generating a creative brief and finished video via the PixVerse Plugin, a creator sharing their workflow, and live Q&A. The post presents these as recurring sessions rather than a new product launch.

  9. OpenAI DevelopersOfficialAI score46

    Codex on Windows gets new MXC-based sandbox mode

    AIOpenAI says Codex on Windows now has a new sandbox mode built on Microsoft's Execution Containers (MXC), offering faster setup, stronger network enforcement, and granular file access controls. The mode requires a compatible Windows 11 device. Background from Microsoft's announcement says MXC is now generally available on Windows 11, keeping agents within boundaries the operating system enforces.

  10. LangChainOfficialAI score29

    LangSmith data shows Claude Sonnet 5 and GPT-5.6 Luna gaining ground

    AILangChain reports that over the last month Claude Sonnet 5 rose from #9 to #2 in model adoption, with 51% more organizations using it. GPT-5.6 Luna climbed from #3 to #1 in call footprints, up 65% in calls, while smaller, faster models dominate call footprints overall. Two open-weights models entered the adoption top 10 but do not lead in call volume.

    Image from @LangChain's post
  11. TechRadar · AINewsAI score38

    Small businesses using AI expect faster growth and hiring than non-adopters

    AIA New York Federal Reserve Bank study found that US small businesses using AI are more optimistic about hiring and revenue than non-adopters. AI adopters expected a 33-point net increase in employment over the next year, versus 15 points for non-adopters. Only 31% of businesses have actually reported higher sales from AI so far.

  12. CNBC · TechnologyNewsAI score26

    AI shopping agents like Muse could reshape retail stocks

    AIThe original article discusses AI agents such as Muse that can shop on a user's behalf and what that could mean for retail stocks. The supplied text contains only site navigation and footer material, so no specific product details, figures, or market impacts can be confirmed.

  13. Baseten BlogOfficialAI score61

    How to choose which layers to run at NVFP4 quantization precision

    AIBaseten explains how to decide which layers of a model can run in 4-bit NVFP4 without losing needed information. The post compares architecture-based heuristics, isolated-layer sensitivity scoring, and SaturationQuant, which accounts for other quantized layers. It also covers calibration with representative data and block-level scales of 16 values.

    Why it matters: The post explains how to choose which layers run at NVFP4 precision using heuristics, sensitivity scoring, and saturation-aware scoring, with clear calibration steps.

  14. The Verge · AINewsAI score47

    Instinct AI agent holds its own against Muse and Dots in testing

    AIInstinct, a startup AI agent that works through text messages, handled online tasks such as swim lesson searches, an eye doctor appointment email, and an Ikea return about as well as Muse and Dots, according to The Verge's testing. The startup was valued at $10 billion in late September, and it has no app or subscription fee for now, with access by invitation or waitlist. Its founder, Noah Shinn, says its focus on a personal assistant sets it apart from OpenAI and Meta.

  15. The Verge · AINewsAI score38

    Trump orders federal use of "super intelligence" instead of "artificial intelligence"

    AIPresident Trump has ordered federal agencies to avoid the term "AI" and adopt "super intelligence," renaming the Center for AI Standards and Innovation to CAISSI and calling anyone who uses "Artificial Intelligence" THE ENEMY. Elon Musk, Jeff Bezos, Jensen Huang and Sam Altman have praised the term, though OpenAI says it won't change its name. The article argues the rename clashes with existing laws such as the Take It Down Act and with the earlier, different meaning of superintelligence used by Sen. Bernie Sanders.

  16. GuizangXAI score22

    Guizang releases a one-click Grok bot for daily AI news videos

    AIGuizang says he turned his workflow into a Grok bot that users can install with one click. The bot runs on Grok's cloud virtual machine to collect content, write code, and render a daily morning AI news video without using a local computer.

  17. elvisXAI score38

    Sabi raises $50M to build a wearable brain-AI interface cap

    AISabi has raised $50M from Khosla Ventures, Accel, Initialized, Kevin Weil, and DST Global to build a wearable brain-computer interface cap. The cap carries 100,000 sensors and aims to let people talk to AI agents by thinking. The post's author says improved hardware could unlock new research and personal agentic applications.

  18. TransformerBlogAI score46

    Democrats prepare competing AI regulation bills ahead of possible congressional takeover

    AIDemocrats are drafting competing AI regulation proposals as they prepare for a possible congressional takeover next month. Their bills include a federal standard-setting agency and an emergency shutdown switch for models, with Sen. Maria Cantwell's framework calling for constant government and independent oversight of frontier models. Party members oppose the FRONTIER Act's preemption provisions, and none of the proposals is expected to pass this year.

  19. Arena.aiOfficialAI score24

    Arena weekly update: Nano Banana 2.1, Mistral Large 4, Claude Haiku 5.5 rankings

    AIArena's weekly update says Nano Banana 2.1 ranked in the top six across three Image Arena modes, with #4 in Multi-Image Edit at 1431 points. Mistral Large 4 placed #43 overall in Agent Arena, 11 spots above Mistral Medium 3.5, and Claude Haiku 5.5 (High) landed #30 in Code Arena WebDev at 1587 points, priced at $0.10/$0.50 per 1M input/output tokens. The post also introduces Arena's Alignment Index and announces a $200M Series B at a $3.1B valuation.

  20. Boris PowerXAI score28

    Boris Power calls OpenAI integer multiplication progress "Wow!"

    AIBoris Power, who owns the OpenAI account, posted only the word "Wow!" with no details. Background from a separate post says the integer multiplication problem #109 witness value κ rose to 2⁻¹⁰·⁵⁴⁷ (about 6.6857 × 10⁻⁴), past the 2⁻¹¹ threshold. The author notes gains are now fractional and a major breakthrough is still needed.

  21. dexXAI score8

    Dex Horthy posts "yep" endorsing Geoffrey Huntley's AI-replacement warning

    AIDex Horthy replied "yep" to Geoffrey Huntley's post arguing that teams too busy with their normal jobs to experiment with AI are being prepared for replacement. Huntley's linked post says AI use is no longer optional for employment and offers hiring advice for 2026 AI-first candidates.

  22. DatabricksOfficialAI score25

    Databricks pairs Temporal and Lakebase for durable cloud agents

    AIDatabricks has published a reference implementation pairing Temporal with Lakebase Postgres so cloud agents can survive worker, container, or deployment replacement. The design keeps recorded work and evidence and review state queryable, and lets human decisions arrive days later. Unity Catalog remains the governed policy source through synced tables.

    Image from @databricks's post
  23. Ben TossellXAI score5

    Ben Tossell teases an idea with no details yet

    AIBen Tossell (@bentossell) posts only "i have an idea..." without describing the concept. He quotes Adel Wu (@adelwu_), who says AI startups struggle to hire people who combine online culture fluency, taste, AI and technical understanding, and execution skill.