Skip to contentSkip to stories

Updated

All AI news

Oct 6

Oct 6Tue
  1. Paige BaileyAI score60

    Google releases Nano Banana 2.1 image model at $0.034 per image

    AIGoogle's Nano Banana 2.1, model gemini-nano-banana-2.1, is now available and is said to outperform the previous Pro model at about a quarter of the price, $0.034 per image versus $0.134. The quoted post lists improved instruction following, better in-image text rendering, grounding with Google Image Search, and up to 5 characters of consistency plus 14 reference images. It is available in Google AI Studio, the Gemini API, Google Cloud, the Gemini app, and Flow. The author's own post is a playful reaction praising its design ability and shows a generated vegan basketball food truck poster.

  2. Microsoft ResearchAI score36

    Jennifer Neville on learning from surprising AI failures and evaluation beyond benchmarks

    AIMicrosoft Research podcast host Chad Atalla interviews Jennifer Neville, a partner research manager at Microsoft, about her path into AI and her work on how evaluation exposes surprising failures in models tested beyond traditional benchmarks. The conversation also covers practical guidance for working with current AI systems and why examining underlying data matters when results defy expectations.

  3. Philipp SchmidAI score62

    Google releases Nano Banana 2.1 image model at $0.034 per image

    AIGoogle's Nano Banana 2.1 (gemini-nano-banana-2.1) is now available and outperforms the previous Pro model at $0.034 per image, versus $0.134 before. It adds improved instruction following, better in-image text rendering, grounding with Google Image Search, and consistency for up to 5 characters with 14 reference images. It is available in Google AI Studio, the Gemini API, Google Cloud, the Gemini app, and Flow by Google.

  4. Google for DevelopersAI score40

    Google's multimodal embedding toolkit runs fully offline on device

    AIGoogle's new multimodal embedding setup processes image, audio, and video entirely offline with zero server calls. Its modular design lets developers drop unused vision and audio components to save memory, and flexible dimension sizes cut local database storage by up to 6x. It can also pair with Gemma 4 to build RAG pipelines with minimal memory and processing requirements.

  5. Google DeepMindAI score58

    Google DeepMind releases EmbeddingGemma 2 with 740M parameters under Apache 2.0

    AIGoogle DeepMind released EmbeddingGemma 2, a 740M-parameter embedding model, under an Apache 2.0 license. The post says it is competitive across benchmarks and outperforms some specialist models more than twice its size, and that developers can use it for multimodal search or pair it with Gemma 4 for on-device RAG. Weights are available on Hugging Face and Kaggle.

  6. Sundar PichaiAI score62

    Google releases EmbeddingGemma 2, an open multimodal embedding model for on-device use

    AIGoogle introduces EmbeddingGemma 2, its first open, natively multimodal embedding model, covering text, code, image, video, and audio tasks. It has a 740M parameter form factor, is positioned for offline, privacy-first RAG when paired with Gemma 4, and the post claims it outperforms some specialist models more than twice its size. Weights are available now on Hugging Face.

  7. Sierra BlogAI score27

    Sierra launches partner ecosystem to extend its AI customer agents across systems and markets

    AISierra introduced a partner ecosystem of technology platforms, marketplaces, and service partners to bring its AI agents to more businesses. Sierra agents connect securely to systems including contact center platforms, payment providers, electronic health records, property management software, and billing systems. The platform is also available through leading cloud and frontier lab marketplaces, and consulting firms and systems integrators help build agents.

  8. Google LabsAI score57

    Google Flow Music Spaces can now export custom tools as VST3/AU plugins

    AIGoogle Flow Music lets creators build custom instruments or effects from natural language, and Spaces can now be exported as VST3/AU plugins. These plugins run inside producers' Digital Audio Workstations, so tools can fit existing production workflows. The source gives producer Khris Riddick-Tynes's "No Chaser" plugin as an example for checking instrumentals and vocals.

  9. Google DeepMind · The KeywordAI score72

    Google releases EmbeddingGemma 2, an open multimodal embedding model for on-device use

    AIGoogle DeepMind has released EmbeddingGemma 2, a 740-million-parameter embedding model that maps text, images, audio, and video into a shared space and runs on local hardware under an Apache 2.0 license. Matryoshka Representation Learning lets developers truncate output vectors from 768 dimensions to 512, 256, or 128, and the model supports an 8K-token context window. The model weights are available on Hugging Face and Kaggle, with Gemini Enterprise Agent Platform availability coming soon.

    Why it matters: The release shows how a 740M-parameter multimodal embedder runs locally with a 768-to-128 dimension truncation option, useful for judging on-device retrieval designs.

  10. Sebastian RaschkaAI score14

    Schmidhuber and Newfield discuss AI and humanity's future in morning chat

    AISebastian Raschka posted about a spontaneous morning conversation with Juergen Schmidhuber and Jake Newfield on AI and its philosophical implications. He said his microphone disconnected during the chat but that Schmidhuber made the more interesting points. The discussion was part of a broader conversation on humanity's future with AI, hosted by HermetiqAI.