Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 6

Oct 6Tue
  1. Philipp SchmidXAI score70

    EmbeddingGemma 2 releases native multimodal embeddings built on Gemma 4

    AIGoogle releases EmbeddingGemma 2, its first native multimodal embedding model, built on Gemma 4 under Apache 2.0. It embeds over 100 languages, code, images, audio, and video into one vector, with an 8,192-token context and four sizes from 270M to 740M parameters. Matryoshka output dimensions of 768, 512, 256, or 128 are supported, and the model is available in Sentence Transformers and LiteRT-LM, with a reported 14% gain on MTEB Code.

    Why it matters: The release extends an embedding model to text, code, images, audio, and video in one vector, a useful option for retrieval systems that mix media types.

  2. Boris PowerXAI score22

    Frontier AI research taste reportedly doubling every three months since December 2025

    AIResearch by pzeroresearch estimates that frontier models' experimental research taste has doubled roughly every three months since December 2025, with Opus 5.5 now exceeding their expert human baseline. The author of the main post, Boris Power, calls the plot very interesting for recursive self-improvement implications, while noting that the details matter for doing useful work at frontier labs.

  3. vLLMOfficialAI score60

    vLLM Adds Day-0 Support for Google's EmbeddingGemma 2 Multimodal Embeddings

    AIvLLM announced day-0 support for EmbeddingGemma 2 from Google DeepMind, a bidirectional omni-modal embedding model that maps text, image, audio, video, and interleaved inputs into one vector space. Users can try it with the latest vLLM nightly build using the command vllm serve google/embeddinggemma-2 --runner pooling. The quoted Google post says the model is built on the Gemma 4 architecture and released under Apache 2.0.

    Image from @vllm_project's post
  4. elvisXAI score41

    Parsewave audit fixes 206 verifier bugs in AutomationBench

    AIParsewave audited all 600 public tasks in Zapier's AutomationBench and human review confirmed 206 real verifier bugs, all of which were fixed in AutomationBench Verified. Replaying 1,235 Kimi K3 runs on the old and fixed verifiers changed 27.9% of grades, with pass rate rising from 18.8% to 43.8% where verifiers were too strict and falling from 60.2% to 49.7% where they were too lenient.

  5. Logan KilpatrickXAI score40

    Nano Banana 2.1 released with higher quality and lower price

    AIGoogle released Nano Banana 2.1, an update to its image generation and editing model, offering higher-quality images and bug fixes from previous versions. The update also comes with a new lower price point. It is available through the API, AI Studio, and the Gemini app.

    Image from @OfficialLoganK's post
  6. ZDNet · AINewsAI score60

    Google limits free Gemini users to Flash-Lite model from October 9

    AIStarting October 9, free Google Gemini users will only have access to the Flash-Lite model, and AI Plus subscribers will lose Pro access. Google is also moving to compute-based limits that refresh every five hours, with higher limits for paid plans. The AI Pro plan at $20 per month will gain access to the Deep Think reasoning mode.

  7. Google · Innovation & AIOfficialAI score42

    Google Study Tests AI-Guided Blind Sweep Ultrasounds for Pregnant Women in Kenya and Chicago

    AIGoogle researchers, working with Northwestern Medicine and Jacaranda Health, trained healthcare workers to perform "blind sweep" ultrasounds analyzed by machine learning models. The models estimated gestational age and fetal presentation as accurately as a trained sonographer in a study of 1,000 mothers each in Nairobi and Chicago. The AI processes results on the device, so it needs no electricity supply or Wi-Fi.

  8. WaymoOfficialAI score54

    Waymo starts fully autonomous driving tests in Detroit

    AIWaymo says its vehicles may begin appearing on Detroit roads this week without a human driver. The company says it is not yet picking up riders and invites users to download the Waymo app to be among the first when service launches.

    Video from @Waymo's post
  9. Arthur MenschXAI score34

    Mistral's Arthur Mensch touts AGI building from a train station

    AIArthur Mensch posted that Mistral is building AGI from a train station. The quoted post adds that ML4 was trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's European datacenters, including its Bruyères-le-Châtel cluster funded by its Series B. Mistral says Series C and D clusters are coming soon to support longer training and faster iteration.

  10. Matt ShumerXAI score34

    AgentID lets AI agents sign in to apps like Google login

    AIAgentMail has launched AgentID, a sign-in system that lets millions of AI agents with AgentMail accounts log into third-party apps within minutes. The post frames it as the agent-era equivalent of Sign in with Google, and offers three months of free AgentMail plans to developers who integrate it and share a screenshot.

  11. laurenXAI score42

    Developer automates releases and QA with Grok Bot agents in Slack

    AIA developer used Grok Bot to build two Slack team bots, sandcastle for release management and poteto for engineering, automating their release and QA process. Sandcastle DMs contributors PR links, kicks off builds, and runs a fuzz swarm of 10+ Grok 4.7 xhigh agents, while poteto triages and fixes issues via Cursor Projects.

    Image from @poteto's post
  12. X.PINXAI score38

    DeepSeek nears 80 billion yuan funding round, Tencent and CATL investing

    AIDeepSeek is close to raising at least 80 billion yuan ($12 billion), up from an original target of about 50 billion yuan, according to Bloomberg citing people familiar with the matter. Tencent and battery maker CATL are among the largest investors in the round, which is expected to close soon. DeepSeek is planning an IPO in early 2027, though details could still change.

  13. Sierra BlogOfficialAI score62

    Sierra and Meta announce Personal Agent Protocol, an open standard for personal agents

    AISierra and Meta are developing Personal Agent Protocol, an open standard defining how personal agents interact with businesses, with industry partners including Genesys, Instinct, Rocket, Shopify, Stripe, and Walmart. The protocol uses OAuth sessions where consumers choose read-only or write access and companies choose whether agents reach them through websites, APIs via MCP and OpenAPI, or their own agents. The authors plan to publish the v0.1 specification later this month along with a reference implementation.

    Why it matters: The post specifies how personal agents would authenticate and reach businesses through websites, APIs, or company agents, which matters for anyone building agent integrations.

  14. ReplicateOfficialAI score29

    Nano Banana 2.1 image model now live on Replicate

    AIReplicate has made Nano Banana 2.1, Google DeepMind's latest image model, available, optimized for a balance of price and performance. The model offers improved visual design, mask-based editing, and subject consistency for more natural-looking images.

    Image from @replicate's post
  15. ClaudeOfficialAI score42

    Claude now edits Google Docs, Sheets, and Slides in beta

    AIAnthropic's Claude can now open Google Docs, Sheets, and Slides beside the chat, either from a pasted Google file link or a request for a new file, so users and Claude can edit together. Access follows existing Google sharing permissions. The feature is in beta on all paid plans.

  16. Nano Banana 2.1OfficialAI score40

    Nano Banana 2.1 released with gains in design, editing, and consistency

    AIGoogle announces Nano Banana 2.1, an upgraded image model that outperforms its previous versions across the board. The company says it brings notable improvements in visual design, mask-based editing, subject consistency, and more natural-looking imagery. It is available now in the Gemini app and Google AI Studio.

    Image from @NanoBanana's post
  17. Yuchen JinXAI score34

    Reflection's Beam and Mistral Large 4 near GLM-5.2 level

    AIYuchen Jin says Reflection's Beam and Mistral Large 4 both reached roughly GLM-5.2 level within the past two days. He suggests the Western versus Chinese open-source model gap may come down to Chinese labs being able to distill Anthropic and OpenAI models, which Western labs cannot.

  18. Design ArenaOfficialAI score40

    Google's Nano Banana 2.1 image model now available on Design Arena

    AIGoogle DeepMind's Nano Banana 2.1, Google's latest image generation and editing model, is now available on Design Arena. The model brings improvements in visual design, mask-based editing, and subject consistency, aiming to produce more natural-looking images with greater control across editing workflows.

    Image from @DesignArena's post