Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. Google GeminiOfficialAI score34

    Gemini Live's Guided Vision now rolls out to Android devices

    AIGuided Vision in Gemini Live is now available on Android devices running Android 9 and above, where Gemini Live is supported. Users need the latest Gemini app and Android system software to get started.

  2. AnthropicOfficialAI score49

    Anthropic expands Cyber Verification Program for verified security professionals

    AIAnthropic is expanding its Cyber Verification Program to give verified security professionals broader access to its most capable models. Through the program, they can use Claude Mythos 5.1, Opus 5.5, and Sonnet 5.5 with safeguards designed for defensive work. New tiers will also allow authorized offensive work such as penetration testing and red-teaming.

  3. Google GeminiOfficialAI score46

    Gemini Live adds Guided Vision for real-time visual assistance

    AIGoogle's Gemini Live now includes Guided Vision, built with blind and low-vision community input, offering conversational real-time visual assistance. Users can share their camera to receive dynamic audio descriptions and verbal cues to help them explore their surroundings.

    Video from @GeminiApp's post
  4. Claude Code · GitHub ReleasesOfficialAI score40

    Claude Code v2.1.292 adds plugin marketplace flag and fixes security issues

    AIClaude Code v2.1.292 adds a --marketplace option to claude plugin install, which adds the marketplace if needed and then installs the plugin from it. The release also adds an effort parameter to the Agent tool and fixes several security issues, including permission prompts bypassed for network (UNC) file reads and a sandboxed read path that could return files outside approved access.

  5. OpenRouterOfficialAI score32

    OpenRouter adds Gemini Nano Banana 2.1 with image and panorama features

    AIOpenRouter now offers google/gemini-nano-banana-2.1, which accepts up to 14 reference images and supports Grounding with Google Search. The model produces cleaner 1:4, 4:1, 1:8 and 8:1 panoramas at 2K and 4K. Pricing is $0.0336 per image at 1K, $0.0504 at 2K and $0.1134 at 4K.

  6. Lydia Hallie ✨XAI score23

    Claude Code cloud sessions one-time bonus credit claimable until Oct 7

    AIAnthropic's Lydia Hallie says users can still claim a one-time bonus credit for Claude Code cloud sessions until October 7 by running /claim-credit. Cloud sessions run Claude Code on a fresh VM per task, letting several run at once even after the laptop is closed.

  7. Philipp SchmidXAI score22

    Embedding Gemma runs in browser via WebGPU demo

    AIPhilipp Schmid shares a Hugging Face Space that runs Gemma embedding models in the browser using WebGPU. The demo, a webml-community project, lets users generate embeddings locally without server-side inference.

    Video from @_philschmid's post
  8. ClaudeDevsOfficialAI score20

    Claude Pro and Max users can claim one-time cloud sessions bonus credit

    AIAnthropic's ClaudeDevs account says Pro or Max plan subscribers on September 23 can still claim a one-time bonus credit for cloud sessions by running /claim-credit in Claude Code by October 7 at 11:59pm PT. Cloud sessions draw on this credit first before counting toward plan limits, and the credit expires November 4.

  9. ClaudeDevsOfficialAI score38

    Claude Code cloud sessions run parallel tasks on fresh VMs

    AIAnthropic's ClaudeDevs says Claude Code cloud sessions run each task on a fresh VM, so users can start several at once. The sessions keep running after the user closes their laptop. A field guide covers seven suitable workflows and how to connect GitHub.

  10. PikaOfficialAI score42

    Nano Banana 2.1 launches on Pika with better visuals and speed

    AIPika has made Nano Banana 2.1 available on its platform, citing improved visual quality and faster generation speeds. The post says the model incorporates the real-world intelligence of Gemini, though it gives no benchmark figures, pricing, or technical specifications.

    Video from @pika_labs's post
  11. Google ResearchOfficialAI score36

    Google Research demos Co-Director for coherent long-form AI video generation

    AIGoogle Research will demo Co-Director, a hierarchical multi-agent framework that optimizes video generation and consistency for long-form storytelling, at the #COLM2026 Google booth #107 today at 1:00 PM PT. The demo showcases interactive cinematic narratives, and the team's blog post details the approach.

    Image from @GoogleResearch's post
  12. Philipp SchmidXAI score70

    EmbeddingGemma 2 releases native multimodal embeddings built on Gemma 4

    AIGoogle releases EmbeddingGemma 2, its first native multimodal embedding model, built on Gemma 4 under Apache 2.0. It embeds over 100 languages, code, images, audio, and video into one vector, with an 8,192-token context and four sizes from 270M to 740M parameters. Matryoshka output dimensions of 768, 512, 256, or 128 are supported, and the model is available in Sentence Transformers and LiteRT-LM, with a reported 14% gain on MTEB Code.

    Why it matters: The release extends an embedding model to text, code, images, audio, and video in one vector, a useful option for retrieval systems that mix media types.

  13. NVIDIA AIOfficialAI score16

    NVIDIA to livestream DGX Spark smart routing for hybrid AI

    AINVIDIA AI announced a live broadcast titled "DGX Spark Live: Smart Routing for Hybrid AI" on X. The post provides no further details on the routing approach, specifications, or results.

  14. vLLMOfficialAI score60

    vLLM Adds Day-0 Support for Google's EmbeddingGemma 2 Multimodal Embeddings

    AIvLLM announced day-0 support for EmbeddingGemma 2 from Google DeepMind, a bidirectional omni-modal embedding model that maps text, image, audio, video, and interleaved inputs into one vector space. Users can try it with the latest vLLM nightly build using the command vllm serve google/embeddinggemma-2 --runner pooling. The quoted Google post says the model is built on the Gemma 4 architecture and released under Apache 2.0.

    Image from @vllm_project's post
  15. Logan KilpatrickXAI score40

    Nano Banana 2.1 released with higher quality and lower price

    AIGoogle released Nano Banana 2.1, an update to its image generation and editing model, offering higher-quality images and bug fixes from previous versions. The update also comes with a new lower price point. It is available through the API, AI Studio, and the Gemini app.

    Image from @OfficialLoganK's post
  16. Mistral AIOfficialAI score47

    Mistral Large 4 solves 18 of 19 CTF challenges in speedrun test

    AIMistral Large 4 solved 18 of 19 challenges in a CTF speedrun, with tool calls and solve times drawn from actual runs. The post frames the model as efficient at reasoning over diverse complex challenges compared with other models.

    Video from @MistralAI's post
  17. ZDNet · AINewsAI score60

    Google limits free Gemini users to Flash-Lite model from October 9

    AIStarting October 9, free Google Gemini users will only have access to the Flash-Lite model, and AI Plus subscribers will lose Pro access. Google is also moving to compute-based limits that refresh every five hours, with higher limits for paid plans. The AI Pro plan at $20 per month will gain access to the Deep Think reasoning mode.

  18. Google · Innovation & AIOfficialAI score42

    Google Study Tests AI-Guided Blind Sweep Ultrasounds for Pregnant Women in Kenya and Chicago

    AIGoogle researchers, working with Northwestern Medicine and Jacaranda Health, trained healthcare workers to perform "blind sweep" ultrasounds analyzed by machine learning models. The models estimated gestational age and fetal presentation as accurately as a trained sonographer in a study of 1,000 mothers each in Nairobi and Chicago. The AI processes results on the device, so it needs no electricity supply or Wi-Fi.

  19. WaymoOfficialAI score54

    Waymo starts fully autonomous driving tests in Detroit

    AIWaymo says its vehicles may begin appearing on Detroit roads this week without a human driver. The company says it is not yet picking up riders and invites users to download the Waymo app to be among the first when service launches.

    Video from @Waymo's post
  20. Arthur MenschXAI score34

    Mistral's Arthur Mensch touts AGI building from a train station

    AIArthur Mensch posted that Mistral is building AGI from a train station. The quoted post adds that ML4 was trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's European datacenters, including its Bruyères-le-Châtel cluster funded by its Series B. Mistral says Series C and D clusters are coming soon to support longer training and faster iteration.

  21. clem 🤗XAI score62

    Mistral Large 4 announced with API access today and open weights due end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active parameters. It is available via API today, with open weights planned for the end of October. Clément Delangue, Hugging Face's CEO, reacted by noting that the model cannot be the best open-weight model until its weights are actually released.

    Why it matters: The quoted announcement gives the parameter scale, active count, and availability path, which help readers compare it with other open-weight releases.

  22. X.PINXAI score38

    DeepSeek nears 80 billion yuan funding round, Tencent and CATL investing

    AIDeepSeek is close to raising at least 80 billion yuan ($12 billion), up from an original target of about 50 billion yuan, according to Bloomberg citing people familiar with the matter. Tencent and battery maker CATL are among the largest investors in the round, which is expected to close soon. DeepSeek is planning an IPO in early 2027, though details could still change.

  23. Sierra BlogOfficialAI score62

    Sierra and Meta announce Personal Agent Protocol, an open standard for personal agents

    AISierra and Meta are developing Personal Agent Protocol, an open standard defining how personal agents interact with businesses, with industry partners including Genesys, Instinct, Rocket, Shopify, Stripe, and Walmart. The protocol uses OAuth sessions where consumers choose read-only or write access and companies choose whether agents reach them through websites, APIs via MCP and OpenAPI, or their own agents. The authors plan to publish the v0.1 specification later this month along with a reference implementation.

    Why it matters: The post specifies how personal agents would authenticate and reach businesses through websites, APIs, or company agents, which matters for anyone building agent integrations.

  24. LumaOfficialAI score13

    Luma argues no single model suits every creative stage

    AILuma Labs says there is no one model that excels at every task, since ideation and final-frame generation require different strengths. The post's point is that creative work should be matched to the best model for each stage rather than a single model for all.

    Video from @LumaLabsAI's post
  25. Google FlowOfficialAI score38

    Google Flow adds Nano Banana 2.1 for improved image generation and editing

    AIGoogle Flow now offers Nano Banana 2.1, which Google says improves visual quality and subject consistency for both image generation and editing. Nano Banana Pro and Nano Banana 2 Lite remain available in Google Flow alongside the new model.

  26. ReplicateOfficialAI score29

    Nano Banana 2.1 image model now live on Replicate

    AIReplicate has made Nano Banana 2.1, Google DeepMind's latest image model, available, optimized for a balance of price and performance. The model offers improved visual design, mask-based editing, and subject consistency for more natural-looking images.

    Image from @replicate's post