Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 6

Oct 6Tue
  1. PikaOfficialAI score42

    Nano Banana 2.1 launches on Pika with better visuals and speed

    AIPika has made Nano Banana 2.1 available on its platform, citing improved visual quality and faster generation speeds. The post says the model incorporates the real-world intelligence of Gemini, though it gives no benchmark figures, pricing, or technical specifications.

    Video from @pika_labs's post
  2. Google ResearchOfficialAI score36

    Google Research demos Co-Director for coherent long-form AI video generation

    AIGoogle Research will demo Co-Director, a hierarchical multi-agent framework that optimizes video generation and consistency for long-form storytelling, at the #COLM2026 Google booth #107 today at 1:00 PM PT. The demo showcases interactive cinematic narratives, and the team's blog post details the approach.

    Image from @GoogleResearch's post
  3. Philipp SchmidXAI score70

    EmbeddingGemma 2 releases native multimodal embeddings built on Gemma 4

    AIGoogle releases EmbeddingGemma 2, its first native multimodal embedding model, built on Gemma 4 under Apache 2.0. It embeds over 100 languages, code, images, audio, and video into one vector, with an 8,192-token context and four sizes from 270M to 740M parameters. Matryoshka output dimensions of 768, 512, 256, or 128 are supported, and the model is available in Sentence Transformers and LiteRT-LM, with a reported 14% gain on MTEB Code.

    Why it matters: The release extends an embedding model to text, code, images, audio, and video in one vector, a useful option for retrieval systems that mix media types.

  4. vLLMOfficialAI score60

    vLLM Adds Day-0 Support for Google's EmbeddingGemma 2 Multimodal Embeddings

    AIvLLM announced day-0 support for EmbeddingGemma 2 from Google DeepMind, a bidirectional omni-modal embedding model that maps text, image, audio, video, and interleaved inputs into one vector space. Users can try it with the latest vLLM nightly build using the command vllm serve google/embeddinggemma-2 --runner pooling. The quoted Google post says the model is built on the Gemma 4 architecture and released under Apache 2.0.

    Why it matters: The post gives a runnable serve command and day-0 vLLM support, showing how to deploy the new multimodal embedding model locally.

    Image from @vllm_project's post
  5. Logan KilpatrickXAI score40

    Nano Banana 2.1 released with higher quality and lower price

    AIGoogle released Nano Banana 2.1, an update to its image generation and editing model, offering higher-quality images and bug fixes from previous versions. The update also comes with a new lower price point. It is available through the API, AI Studio, and the Gemini app.

    Image from @OfficialLoganK's post
  6. Mistral AIOfficialAI score47

    Mistral Large 4 solves 18 of 19 CTF challenges in speedrun test

    AIMistral Large 4 solved 18 of 19 challenges in a CTF speedrun, with tool calls and solve times drawn from actual runs. The post frames the model as efficient at reasoning over diverse complex challenges compared with other models.

    Video from @MistralAI's post
  7. ZDNet · AINewsAI score60

    Google limits free Gemini users to Flash-Lite model from October 9

    AIStarting October 9, free Google Gemini users will only have access to the Flash-Lite model, and AI Plus subscribers will lose Pro access. Google is also moving to compute-based limits that refresh every five hours, with higher limits for paid plans. The AI Pro plan at $20 per month will gain access to the Deep Think reasoning mode.

  8. Google · Innovation & AIOfficialAI score42

    Google Study Tests AI-Guided Blind Sweep Ultrasounds for Pregnant Women in Kenya and Chicago

    AIGoogle researchers, working with Northwestern Medicine and Jacaranda Health, trained healthcare workers to perform "blind sweep" ultrasounds analyzed by machine learning models. The models estimated gestational age and fetal presentation as accurately as a trained sonographer in a study of 1,000 mothers each in Nairobi and Chicago. The AI processes results on the device, so it needs no electricity supply or Wi-Fi.

  9. WaymoOfficialAI score54

    Waymo starts fully autonomous driving tests in Detroit

    AIWaymo says its vehicles may begin appearing on Detroit roads this week without a human driver. The company says it is not yet picking up riders and invites users to download the Waymo app to be among the first when service launches.

    Video from @Waymo's post
  10. Arthur MenschXAI score34

    Mistral's Arthur Mensch touts AGI building from a train station

    AIArthur Mensch posted that Mistral is building AGI from a train station. The quoted post adds that ML4 was trained on 3,800 NVIDIA Grace Blackwell GPUs in Mistral's European datacenters, including its Bruyères-le-Châtel cluster funded by its Series B. Mistral says Series C and D clusters are coming soon to support longer training and faster iteration.

  11. clem 🤗XAI score62

    Mistral Large 4 announced with API access today and open weights due end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active parameters. It is available via API today, with open weights planned for the end of October. Clément Delangue, Hugging Face's CEO, reacted by noting that the model cannot be the best open-weight model until its weights are actually released.

    Why it matters: The quoted announcement gives the parameter scale, active count, and availability path, which help readers compare it with other open-weight releases.

  12. X.PINXAI score38

    DeepSeek nears 80 billion yuan funding round, Tencent and CATL investing

    AIDeepSeek is close to raising at least 80 billion yuan ($12 billion), up from an original target of about 50 billion yuan, according to Bloomberg citing people familiar with the matter. Tencent and battery maker CATL are among the largest investors in the round, which is expected to close soon. DeepSeek is planning an IPO in early 2027, though details could still change.

  13. Sierra BlogOfficialAI score62

    Sierra and Meta announce Personal Agent Protocol, an open standard for personal agents

    AISierra and Meta are developing Personal Agent Protocol, an open standard defining how personal agents interact with businesses, with industry partners including Genesys, Instinct, Rocket, Shopify, Stripe, and Walmart. The protocol uses OAuth sessions where consumers choose read-only or write access and companies choose whether agents reach them through websites, APIs via MCP and OpenAPI, or their own agents. The authors plan to publish the v0.1 specification later this month along with a reference implementation.

    Why it matters: The post specifies how personal agents would authenticate and reach businesses through websites, APIs, or company agents, which matters for anyone building agent integrations.

  14. Google FlowOfficialAI score38

    Google Flow adds Nano Banana 2.1 for improved image generation and editing

    AIGoogle Flow now offers Nano Banana 2.1, which Google says improves visual quality and subject consistency for both image generation and editing. Nano Banana Pro and Nano Banana 2 Lite remain available in Google Flow alongside the new model.

  15. ReplicateOfficialAI score29

    Nano Banana 2.1 image model now live on Replicate

    AIReplicate has made Nano Banana 2.1, Google DeepMind's latest image model, available, optimized for a balance of price and performance. The model offers improved visual design, mask-based editing, and subject consistency for more natural-looking images.

    Image from @replicate's post
  16. ClaudeOfficialAI score42

    Claude now edits Google Docs, Sheets, and Slides in beta

    AIAnthropic's Claude can now open Google Docs, Sheets, and Slides beside the chat, either from a pasted Google file link or a request for a new file, so users and Claude can edit together. Access follows existing Google sharing permissions. The feature is in beta on all paid plans.

  17. ClaudeOfficialAI score62

    Claude now works inside Google Docs, Sheets, and Slides

    AIand files from those apps can also be opened within Claude. In Google Workspace, Claude appears in a sidebar next to the open file, reads its contents, and edits it in place, with each edit available for user approval before it is applied.

    Why it matters: The source describes Claude working within Google Workspace files and approving edits, a direct change to how users in those apps can collaborate with the model.

    Video from @claudeai's post
  18. Nano Banana 2.1OfficialAI score40

    Nano Banana 2.1 released with gains in design, editing, and consistency

    AIGoogle announces Nano Banana 2.1, an upgraded image model that outperforms its previous versions across the board. The company says it brings notable improvements in visual design, mask-based editing, subject consistency, and more natural-looking imagery. It is available now in the Gemini app and Google AI Studio.

    Image from @NanoBanana's post
  19. Design ArenaOfficialAI score40

    Google's Nano Banana 2.1 image model now available on Design Arena

    AIGoogle DeepMind's Nano Banana 2.1, Google's latest image generation and editing model, is now available on Design Arena. The model brings improvements in visual design, mask-based editing, and subject consistency, aiming to produce more natural-looking images with greater control across editing workflows.

    Image from @DesignArena's post
  20. eric zakariassonXAI score32

    Grok turns one product photo into an 8-second vertical video ad

    AIA pipeline built with Grok takes a single product photo and produces an eight-second vertical ad, writing the brief, generating three scenes, selecting the best one, animating it, and adding a voiceover. Each run costs about $1.35, and the code is available in the xai-cookbook repository on GitHub.

    Video from @ericzakariasson's post
  21. eric zakariassonXAI score36

    Grok pipeline turns a one-line premise into a short film

    AIxAI's storyboard-to-film example uses grok-4.7 to plan four shots from a one-line premise, then Grok Imagine draws and animates each one with text-to-speech narration. Every shot is an edit of the first keyframe, which keeps the main character consistent across scenes. The code is available in the xai-cookbook repository on GitHub.

    Video from @ericzakariasson's post
  22. eric zakariassonXAI score29

    Cursor adds five SpaceXAI TypeScript SDK demo apps to cookbook

    AICursor has added five apps to its cookbook, all built on the new SpaceXAI TypeScript SDK. The apps cover premise-to-short-film, picture-to-video ads, screenshot-to-React-component, X posts-to-sentiment dashboard, and link-to-podcast conversion. Demos and code are available in the post.

    Video from @ericzakariasson's post
  23. LangChainOfficialAI score46

    LangChain video shows how to build a model router into an agent harness

    AILangChain's Sydney Runkle presents a four-step method for building a model router into a coding agent harness: understanding tasks, understanding models, building the router, and tracking task outcomes. The background post says the router cut costs by 64% without reducing quality by sending tasks that do not need a frontier model to cheaper models.

  24. 👩‍💻 Paige BaileyXAI score20

    Nano Banana 2.1 Gets Day 0 Support on fal

    AIGoogle's Gemini Nano Banana 2.1 image model is available on fal from day zero, according to Paige Bailey's post thanking fal. The fal background adds that it generates significantly faster than Nano Banana 2 and improves visual design, mask-based editing, and subject consistency.

  25. AMDOfficialAI score24

    OpenAI picks AMD EPYC Turin CPUs to host Jalapeño ASICs

    AIOpenAI selected AMD EPYC "Turin" CPUs to host its Jalapeño AI ASIC deployment, citing platform strength, partner experience, and reducing unnecessary risk. The post, which links to a Tom's Hardware report, says AMD aims to deliver reliable solutions for leaders facing aggressive AI performance goals.