Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 6

Oct 6Tue
  1. OpenAI DevelopersOfficialAI score22

    OpenAI's GPT-6 Luna adds predicate, choice, and score outputs

    AIOpenAI's GPT-6 Luna accepts text and image inputs and supports three output types: predicates that estimate the probability a statement is true, choices that select from predefined options with confidence scores, and scores that evaluate an input against a numeric range.

    Video from @OpenAIDevs's post
  2. Aravind SrinivasXAI score40

    Perplexity halves Decision API input pricing to $0.02 per million tokens

    AIPerplexity cut its Decision API input pricing by half, to $0.02 per million input tokens. The reduction follows the release of pplx-decider-v1.1-27b, an open-weights multimodal decision model that scores highest on Hugging Face's Decision Index 0.3 benchmark. The model costs half as much as v1.

  3. v0OfficialAI score27

    v0 iOS app now supports ChatGPT subscriptions

    AIThe v0 iOS app now lets users connect their ChatGPT subscription to build with their ChatGPT tokens. The feature is available to ChatGPT Plus and Pro subscribers.

    Video from @v0's post
  4. Aravind SrinivasXAI score36

    Perplexity Computer Offers Engineering Design as a Service

    AIPerplexity's Computer can generate an interactive 3D preview of a design and export it as an editable CAD file. It can also modify the design in FreeCAD and record a video of the edits.

  5. ChatGPTOfficialAI score44

    ChatGPT Meetings plugin takes notes and drafts follow-ups in beta

    AIOpenAI's ChatGPT Meetings plugin takes notes during meetings and saves a personalized summary and next steps in ChatGPT Space. Users can keep notes private or share them with their team, then ask ChatGPT to update a project plan or draft a follow-up. It is in beta for Pro and Business users in the ChatGPT desktop app on macOS, with Enterprise coming soon.

    Image from @ChatGPT's post
  6. Boris ChernyXAI score38

    Boris Cherny shares prompts for formally verifying Claude Agent SDK

    AIBoris Cherny says he used Opus 5.5 with Lean to formally verify the Claude Agent SDK, with a couple of short prompts producing 16 PRs fixing bugs and race conditions. He also reports that TLA+ works well, sometimes combined with Lean to find data flow, concurrency, and state management issues. The post links to his actual prompts as another example.

  7. NVIDIA Technical BlogOfficialAI score37

    Scale Bitwise-Deterministic Pretraining with NVIDIA Megatron Core

    AINVIDIA's technical blog describes bitwise determinism for large-scale pretraining with Megatron Core, which makes training runs easier to debug, validate, and resume reproducibly. The source says these benefits matter most for models with trillions of parameters trained across thousands of GPUs, where multiple parallelism dimensions, low-precision computation, and distributed checkpointing complicate failure reproduction and fix validation.

  8. Google DeepMindOfficialAI score67

    Google DeepMind releases EmbeddingGemma 2, an open multimodal embedding model for on-device use

    AIGoogle DeepMind has released EmbeddingGemma 2, an open 740 million parameter model that maps text, images, audio, and video into one embedding space. It is built on the Gemma 4 architecture under an Apache 2.0 license and supports an 8K token context window. The company reports a code benchmark gain from 68.76 to 78.68 on MTEB Code and says the model can run on-device with about 567MB of active RAM for the full multimodal version on a Google Pixel 11 Pro.

    Why it matters: The release shows how a 740M-parameter embedding model can cover text, code, images, audio, and video on local hardware, with memory and storage figures to compare against other on-device options.

  9. Google FlowOfficialAI score20

    Google Flow adds custom text effects to videos with Omni 1.1 Flash

    AIGoogle Flow users can add custom text overlays to existing videos by dragging a video into the prompt box, selecting Omni 1.1 Flash, and describing the scene and text placement. Writing the text in quotes and describing its style, with "preserve all other motion" to keep the original video's look, is the suggested method.

    Video from @FlowbyGoogle's post
  10. falOfficialAI score34

    fal Now Available in ChatGPT and Codex for Media Generation

    AIfal is now available inside ChatGPT and Codex, letting users generate images and videos without leaving the chat. Generated media can be browsed directly in the conversation, and users can access their fal Media Library from ChatGPT. The library can be pinned to the sidebar for quick access to assets in any chat.

    Video from @fal's post
  11. TiboXAI score34

    OpenAI API removes friction for developers building on it

    AIOpenAI says it removed some friction for developers building on its API, while the main post gives no specific changes. Background from @OpenAIDevs says the five paid usage tiers become three—Build, Launch, and Grow—and Grow, the new highest tier, requires $500 in total API payments, down from $1,000 for the previous top tier.

  12. NVIDIA Technical BlogOfficialAI score36

    How DOCA GPUNetIO Unifies GPU-Initiated Networking Across the NVIDIA Software Stack

    AINVIDIA's DOCA GPUNetIO lets GPU applications control networking and data movement directly, rather than routing each transaction through the CPU. The source says host-driven network handling adds latency on the critical path and limits how quickly distributed applications can respond in real time. The provided text is truncated, so details of the unified software stack are not available.

  13. Google GeminiOfficialAI score22

    Gemini's Guided Vision offers visual help for everyday tasks

    AIGoogle Gemini's Guided Vision supports everyday tasks such as reading fine print, finding objects, describing items, and exploring one's immediate surroundings. The feature is also presented as practical visual assistance for older adults, people with low literacy, and anyone reading in low light.

  14. Google GeminiOfficialAI score29

    Gemini Live's Guided Vision coaches users to reframe their camera

    AIWith Guided Vision turned on, Gemini Live verbally prompts users to pan, tilt, or step back when their camera framing is too high, too close, or off to the side. These prompts help Gemini get the clear visual context it needs to answer questions.

  15. Google GeminiOfficialAI score34

    Gemini Live's Guided Vision now rolls out to Android devices

    AIGuided Vision in Gemini Live is now available on Android devices running Android 9 and above, where Gemini Live is supported. Users need the latest Gemini app and Android system software to get started.

  16. AnthropicOfficialAI score49

    Anthropic expands Cyber Verification Program for verified security professionals

    AIAnthropic is expanding its Cyber Verification Program to give verified security professionals broader access to its most capable models. Through the program, they can use Claude Mythos 5.1, Opus 5.5, and Sonnet 5.5 with safeguards designed for defensive work. New tiers will also allow authorized offensive work such as penetration testing and red-teaming.

  17. Google GeminiOfficialAI score46

    Gemini Live adds Guided Vision for real-time visual assistance

    AIGoogle's Gemini Live now includes Guided Vision, built with blind and low-vision community input, offering conversational real-time visual assistance. Users can share their camera to receive dynamic audio descriptions and verbal cues to help them explore their surroundings.

    Video from @GeminiApp's post
  18. Claude Code · GitHub ReleasesOfficialAI score40

    Claude Code v2.1.292 adds plugin marketplace flag and fixes security issues

    AIClaude Code v2.1.292 adds a --marketplace option to claude plugin install, which adds the marketplace if needed and then installs the plugin from it. The release also adds an effort parameter to the Agent tool and fixes several security issues, including permission prompts bypassed for network (UNC) file reads and a sandboxed read path that could return files outside approved access.

  19. OpenRouterOfficialAI score32

    OpenRouter adds Gemini Nano Banana 2.1 with image and panorama features

    AIOpenRouter now offers google/gemini-nano-banana-2.1, which accepts up to 14 reference images and supports Grounding with Google Search. The model produces cleaner 1:4, 4:1, 1:8 and 8:1 panoramas at 2K and 4K. Pricing is $0.0336 per image at 1K, $0.0504 at 2K and $0.1134 at 4K.

  20. Lydia Hallie ✨XAI score23

    Claude Code cloud sessions one-time bonus credit claimable until Oct 7

    AIAnthropic's Lydia Hallie says users can still claim a one-time bonus credit for Claude Code cloud sessions until October 7 by running /claim-credit. Cloud sessions run Claude Code on a fresh VM per task, letting several run at once even after the laptop is closed.

  21. Philipp SchmidXAI score22

    Embedding Gemma runs in browser via WebGPU demo

    AIPhilipp Schmid shares a Hugging Face Space that runs Gemma embedding models in the browser using WebGPU. The demo, a webml-community project, lets users generate embeddings locally without server-side inference.

    Video from @_philschmid's post
  22. ClaudeDevsOfficialAI score20

    Claude Pro and Max users can claim one-time cloud sessions bonus credit

    AIAnthropic's ClaudeDevs account says Pro or Max plan subscribers on September 23 can still claim a one-time bonus credit for cloud sessions by running /claim-credit in Claude Code by October 7 at 11:59pm PT. Cloud sessions draw on this credit first before counting toward plan limits, and the credit expires November 4.

  23. ClaudeDevsOfficialAI score38

    Claude Code cloud sessions run parallel tasks on fresh VMs

    AIAnthropic's ClaudeDevs says Claude Code cloud sessions run each task on a fresh VM, so users can start several at once. The sessions keep running after the user closes their laptop. A field guide covers seven suitable workflows and how to connect GitHub.

  24. PikaOfficialAI score42

    Nano Banana 2.1 launches on Pika with better visuals and speed

    AIPika has made Nano Banana 2.1 available on its platform, citing improved visual quality and faster generation speeds. The post says the model incorporates the real-world intelligence of Gemini, though it gives no benchmark figures, pricing, or technical specifications.

    Video from @pika_labs's post
  25. Google ResearchOfficialAI score36

    Google Research demos Co-Director for coherent long-form AI video generation

    AIGoogle Research will demo Co-Director, a hierarchical multi-agent framework that optimizes video generation and consistency for long-form storytelling, at the #COLM2026 Google booth #107 today at 1:00 PM PT. The demo showcases interactive cinematic narratives, and the team's blog post details the approach.

    Image from @GoogleResearch's post
  26. Philipp SchmidXAI score70

    EmbeddingGemma 2 releases native multimodal embeddings built on Gemma 4

    AIGoogle releases EmbeddingGemma 2, its first native multimodal embedding model, built on Gemma 4 under Apache 2.0. It embeds over 100 languages, code, images, audio, and video into one vector, with an 8,192-token context and four sizes from 270M to 740M parameters. Matryoshka output dimensions of 768, 512, 256, or 128 are supported, and the model is available in Sentence Transformers and LiteRT-LM, with a reported 14% gain on MTEB Code.

    Why it matters: The release extends an embedding model to text, code, images, audio, and video in one vector, a useful option for retrieval systems that mix media types.

  27. vLLMOfficialAI score60

    vLLM Adds Day-0 Support for Google's EmbeddingGemma 2 Multimodal Embeddings

    AIvLLM announced day-0 support for EmbeddingGemma 2 from Google DeepMind, a bidirectional omni-modal embedding model that maps text, image, audio, video, and interleaved inputs into one vector space. Users can try it with the latest vLLM nightly build using the command vllm serve google/embeddinggemma-2 --runner pooling. The quoted Google post says the model is built on the Gemma 4 architecture and released under Apache 2.0.

    Why it matters: The post gives a runnable serve command and day-0 vLLM support, showing how to deploy the new multimodal embedding model locally.

    Image from @vllm_project's post
  28. Logan KilpatrickXAI score40

    Nano Banana 2.1 released with higher quality and lower price

    AIGoogle released Nano Banana 2.1, an update to its image generation and editing model, offering higher-quality images and bug fixes from previous versions. The update also comes with a new lower price point. It is available through the API, AI Studio, and the Gemini app.

    Image from @OfficialLoganK's post
  29. Mistral AIOfficialAI score47

    Mistral Large 4 solves 18 of 19 CTF challenges in speedrun test

    AIMistral Large 4 solved 18 of 19 challenges in a CTF speedrun, with tool calls and solve times drawn from actual runs. The post frames the model as efficient at reasoning over diverse complex challenges compared with other models.

    Video from @MistralAI's post