Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 9

TodayOct 9Fri
  1. dexXAI score43

    Dex Horthy posts a one-word teaser, "he cook"

    AIDex Horthy (@dexhorthy) posted only the words "he cook" on X, with no further detail. The post is a short reaction and does not describe a product, release, or result on its own. Background from the quoted post by @0xblacklight describes a serverless background agent that created a GitHub pull request from an issue.

  2. Sierra BlogOfficialAI score62

    Sierra publishes draft Personal Agent Protocol, called Poppy, with 35 new design partners

    AISierra has published a draft of the Personal Agent Protocol, known as Poppy, and named 35 additional design partners, including Adyen, Bank of America, Mastercard, OpenAI, PayPal, and Visa. Under the protocol, companies publish a /.well-known/poppy.json discovery file, and personal agents start sessions, identify themselves, and sign in through OAuth with session tokens limited to approved access. The company says the draft will be followed by design workshops and a reference implementation over the next month.

    Why it matters: The draft specifies how personal agents identify themselves, obtain customer-approved access, and work with company websites, APIs, or agents, which helps readers assess its practical effect on agent-driven transactions.

  3. MarkTechPostNewsAI score67

    OpenAI launches Decisions API in public beta with typed answers

    AIOpenAI has released the Decisions API in public beta, returning typed probabilities, choices, and scores instead of prose from text and images. OpenAI says it runs about 10x faster than the Responses API and costs $0.10 per 1M input tokens with no output charges. The article notes that OpenAI has not published accuracy data and that TypeSafe Jev offers cheaper input pricing at $0.042 per 1M tokens.

  4. MarkTechPostNewsAI score62

    Alibaba Qwen releases Qwen-Image-2.1-Turbo, an 8-step 7B image model

    AIAlibaba's Qwen team released Qwen-Image-2.1-Turbo, an accelerated checkpoint of Qwen-Image-2.1 that generates and edits images in 8 denoising steps instead of 40. The model keeps the same 7B architecture and offers a hosted API at CNY 0.1 per image, while its weights are under a Qwen Research License that requires separate permission for commercial self-hosting.

  5. Google GemmaOfficialAI score46

    Google AI Pro and Ultra plans add A100 and H100 GPUs to Colab

    AIGoogle AI Pro plans now include Colab access to A100 GPUs with 80GB VRAM, and Ultra subscribers can use H100 GPUs. With that memory, users can run Gemma 4 31B in bf16, fully fine-tune Gemma 4 E4B in bf16, LoRA-tune Gemma 4 26B A4B in bf16, and QLoRA-tune Gemma 4 31B in bf16.

    Image from @googlegemma's post
  6. 🚨 AI News | TestingCatalogXAI score36

    Microsoft releases Microsoft-Decision-1, a 9B model for fast decisions

    AIMicrosoft has made Microsoft-Decision-1 available on Microsoft Foundry, a model post-trained on Qwen3.5-9B for fast, single-pass decision scoring. Microsoft says it achieved the highest accuracy across a 36-benchmark comparison of nearly 150,000 questions, and runs 4.5 times faster than Quyet-1.0-Large and 35 times faster than GPT-6 Sol. Microsoft plans to rebase it on other models, including MAI and OpenAI models.

    Video from @testingcatalog's post
  7. Prime IntellectOfficialAI score44

    Prime Intellect extends RL training to multi-agent swarms

    AIPrime Intellect says swarms have costs, since messages consume tokens, lose information, and agents must coordinate to avoid duplicated work. The company is extending its RL training infrastructure from individual agents to multi-agent systems, letting developers express arbitrary agent interactions and train them.

  8. ZDNet · AINewsAI score46

    Amazon launches Alexa Tablets with Alexa+ and Google Play access starting at $230

    AIAmazon announces three Alexa Tablets with Alexa+ built into the interface, starting at $230 for the Tablet 8, $330 for the Tablet 11, and $500 for the Tablet 12 Pro. The tablets are the first of Amazon's newer models to support Google Play alongside Amazon's app store, and they ship October 14 after pre-orders open. Amazon also launches two Kids Tablets, the Kids Tablet 8 at $230 and the Kids Tablet 11 at $330, which run Android instead of FireOS.

  9. Claude Code · GitHub ReleasesOfficialAI score33

    Claude Code v2.1.296 adds gateway policy controls and fixes hook and permission bugs

    AIAnthropic releases Claude Code v2.1.296, which adds a code key to the Claude apps gateway's managed.policies[] and an allow_large option to the Read tool for reading large text files in one call. The release also adds autoCompactWindow for subagents and CLAUDE_CODE_WORKFLOW_SUBAGENT_MODEL, and fixes many bugs in hooks, MCP servers, permission checks and self-hosted runners.

  10. GitHub Copilot ChangelogOfficialAI score36

    Copilot code review adds organization billing and review request controls

    AIGitHub adds two Copilot code review admin controls. Organization owners can bill code reviews from members with a Copilot license to the owning organization instead of member quotas, which requires AI Credits paid usage and allows an optional budget. Owners and repository admins can also restrict review requests to users whose Copilot license comes from their organization or enterprise.

  11. OpenAI · YouTubeOfficialAI score31

    Set up your dot in the ChatGPT mobile app

    AIOpenAI says users can now create a dot in the ChatGPT mobile app. Users can edit its name, customize its avatar, and set it as the first conversation they see when opening the app.

  12. RadixArkOfficialAI score22

    RadixArk praises Proximal for training coding agents with Miles

    AIRadixArk says Proximal is using Miles to train coding agents and calls it a flexible, scalable foundation for teams running their own training workloads. Proximal says its training framework is built on Miles, with runs on Modal's on-demand GPU clusters and serverless GPUs for inference. Its sandboxing infrastructure runs on Kubernetes and can handle millions of concurrent rollouts.

  13. Microsoft CopilotOfficialAI score40

    Microsoft brings full Office apps into Copilot for Frontier users

    AIMicrosoft says the full Word, Excel, and PowerPoint apps are now rolling out to the Copilot app in Frontier. Users can create, edit, and collaborate on documents in one connected workspace, with branded templates and version history. The workspace also offers a library of frontier models, with an Auto router that picks a model for each task.

  14. ChatGPTOfficialAI score22

    ChatGPT app lets users create dots from phones

    AIOpenAI says users can now create their dot directly from their phone in the ChatGPT app on iOS and Android. The post presents this as the first of several fresh updates for dots, and does not give further details.

    Video from @ChatGPT's post
  15. ChatGPTOfficialAI score28

    ChatGPT's dot can now delegate work to Codex threads

    AIOpenAI says its dot can start work in Codex and follow up on existing threads, drawing on ChatGPT conversations, Codex threads, and automations. The dot also decides whether to continue a thread or start a fresh one, and can review and edit Scheduled Tasks in ChatGPT Work.

    Video from @ChatGPT's post
  16. dexXAI score38

    HumanLayer releases teleport and orchestrate commands with a minimalist UI

    AIHumanLayer announces a new release with /hl:teleport, which moves a local session to any remote host the user owns or launches without losing context. The release also adds /hl:orchestrate, which lets HumanLayer drive its own tasks, including splitting work, forking workflows, and moving artifacts, and it ships a minimalist UI with rounded corners and less visual noise.

    Image from @dexhorthy's post
  17. The DecoderNewsAI score62

    Anthropic adds dynamic workflows letting Claude orchestrate up to 1,000 parallel agents

    AIAnthropic is adding dynamic workflows to Claude Managed Agents, letting a lead agent plan tasks, distribute them to up to 1,000 parallel sub-agents, and merge their results. In Anthropic's test, a 116,000-line codebase with 70 hidden bugs saw a single agent catch 14 to 27 per run, while the dynamic workflow consistently caught 66. To activate it, users select the "multiagent_20261001" agent type, and Anthropic recommends starting small because the workflows can use a lot of tokens.

  18. Cloudflare Blog · AIOfficialAI score55

    Cloudflare releases Clef-omni with audio and video input and cuts Clef-flash price

    AICloudflare releases Clef-omni, an open-weight decision model that accepts audio, video, image, and text input in a single API call. Clef-flash's price falls from $0.09 to $0.038 per M input tokens, while its hosted context window drops from 64k to 24k. Cloudflare also reports median latency reductions of 1.7 to 2.0 times for the Clef model on Workers AI.

  19. 🚨 AI News | TestingCatalogXAI score34

    Grok Bot users can have their bot claim an email address

    AISpaceXAI says Grok Bot users can ask their bot to claim its own email address for contacting others and signing up for newsletters. Testing Catalog reports the bot subscribed it to its daily AI Brief newsletter without issue. Admins must enable the feature for their team, and it is rolling out to users starting today.

    Image from @testingcatalog's post
  20. ClineOfficialAI score39

    Cline offers free access to Upstage's Solar Mini 4 model

    AICline is offering Solar Mini 4 free, a new 35B mixture-of-experts model from Korean lab Upstage with 3B active parameters. It has a 524K context window and runs at 208 tokens per second. Cline says it scores 24 on the AAII, the highest of any model at 3B active and within one point of Nemotron 3 Ultra, which uses 55B active.

  21. Satya NadellaXAI score38

    Microsoft unveils Microsoft-Decision-1, a model for fast decision-making

    AIMicrosoft introduces Microsoft-Decision-1, a new model for fast decision-making that it says outperforms both LLMs and other decision models on structured decision tasks in latency and quality. The company says it is already testing the model across Microsoft for uses including incident response, quality control, and scientific discovery.

    Video from @satyanadella's post
  22. OpenAI DevelopersOfficialAI score33

    Codex adds composer predictions for Pro users in beta

    AIOpenAI says composer predictions in Codex is now in beta for Pro users. The feature suggests a user's next message based on their conversation and how they phrase requests. OpenAI calls it one of the most loved features its team has tested internally.

    Video from @OpenAIDevs's post