Skip to contentSkip to stories

Updated

Open source

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. ollamaOfficialAI score55

    Google DeepMind's EmbeddingGemma 2 is now available on Ollama

    AIOllama announced that Google DeepMind's EmbeddingGemma 2 is now available on Ollama. The author describes it as made for consumer devices and multimodal, and gives the command ollama pull embeddinggemma-2 to download it. The quoted DeepMind post says the model is a natively multimodal open model for on-device embeddings that unifies code, images, audio, and video in a shared space.

  2. GammaOfficialAI score42

    Gamma 5 launches with rebuilt design, editing, and visual storytelling tools

    AIGamma announces Gamma 5, a rebuilt version of its presentation platform that it says overhauls how the product thinks, designs, and edits. The company says the release addresses concerns that AI-generated output looked too similar across tools, and it revamps the agent, design tools, editing, import, export, and connectors. Gamma says teams can build presentations, docs, social assets, and graphics that follow their brand or a new aesthetic, using every frontier and image model under the hood.

  3. SGLangOfficialAI score62

    SGLang adds support for Kandinsky 6.0 Video audio-visual generation

    AISGLang now supports Kandinsky 6.0 Video, which generates video and synchronized audio together from text or an image. The model comes in Lite (3B) and Pro (29B) sizes, with built-in super-resolution up to 1920×1080. A sample sglang serve command for the Pro distilled model is included.

    Why it matters: The post shows the exact serve command and model size options, which lets engineers judge whether this open video model fits their hardware and pipeline.

    Image from @sgl_project's post
  4. AMDOfficialAI score18

    Zyphra trains ZAYA1-8B reasoning model on full AMD stack

    AIZyphra trained its ZAYA1-8B reasoning model from scratch on a full-stack AMD platform, according to AMD's post. VP of AI Engineering Quentin Anthony credits access to open software libraries and direct collaboration with AMD for enabling bigger model training and efficient compute use.

    Video from @AMD's post
  5. Gemini CLI · GitHub ReleasesOfficialAI score14

    Gemini CLI v0.63.0 released with retry indicator and auth loop fixes

    AIGemini CLI v0.63.0 adds a retry progress indicator during connection recovery and fixes an infinite authentication loop caused by file contention, headless keyring issues, and supervisor state drops. The release also bounds tool output size and cleans up temporary directories when background shell execution exits, alongside fixes for MCP enablement config handling and stdin restoration after capability detection.

  6. NVIDIA Technical BlogOfficialAI score37

    Scale Bitwise-Deterministic Pretraining with NVIDIA Megatron Core

    AINVIDIA's technical blog describes bitwise determinism for large-scale pretraining with Megatron Core, which makes training runs easier to debug, validate, and resume reproducibly. The source says these benefits matter most for models with trillions of parameters trained across thousands of GPUs, where multiple parallelism dimensions, low-precision computation, and distributed checkpointing complicate failure reproduction and fix validation.

  7. Google DeepMindOfficialAI score67

    Google DeepMind releases EmbeddingGemma 2, an open multimodal embedding model for on-device use

    AIGoogle DeepMind has released EmbeddingGemma 2, an open 740 million parameter model that maps text, images, audio, and video into one embedding space. It is built on the Gemma 4 architecture under an Apache 2.0 license and supports an 8K token context window. The company reports a code benchmark gain from 68.76 to 78.68 on MTEB Code and says the model can run on-device with about 567MB of active RAM for the full multimodal version on a Google Pixel 11 Pro.

    Why it matters: The release shows how a 740M-parameter embedding model can cover text, code, images, audio, and video on local hardware, with memory and storage figures to compare against other on-device options.

  8. MiniMax (official)OfficialAI score12

    MiniMax hosts AI events at SF Tech Week with partner companies

    AIMiniMax is taking part in SF Tech Week with a series of events on October 6, 7, and 8, featuring partners including Friendli.AI, Anaconda, Kilo Code, Novita AI, Artificial Analysis, Nous Research, RadixArk, Vercel, Fireworks AI, DigitalOcean, Modular, and Evermind. The programming covers frontier models, high-speed inference, agents, and open-source AI stacks, plus a Magnific-hosted talk on growing creative AI products.

    Image from @MiniMax_AI's post
  9. Claude Code · GitHub ReleasesOfficialAI score40

    Claude Code v2.1.292 adds plugin marketplace flag and fixes security issues

    AIClaude Code v2.1.292 adds a --marketplace option to claude plugin install, which adds the marketplace if needed and then installs the plugin from it. The release also adds an effort parameter to the Agent tool and fixes several security issues, including permission prompts bypassed for network (UNC) file reads and a sandboxed read path that could return files outside approved access.

  10. Philipp SchmidXAI score22

    Embedding Gemma runs in browser via WebGPU demo

    AIPhilipp Schmid shares a Hugging Face Space that runs Gemma embedding models in the browser using WebGPU. The demo, a webml-community project, lets users generate embeddings locally without server-side inference.

    Video from @_philschmid's post
  11. Philipp SchmidXAI score70

    EmbeddingGemma 2 releases native multimodal embeddings built on Gemma 4

    AIGoogle releases EmbeddingGemma 2, its first native multimodal embedding model, built on Gemma 4 under Apache 2.0. It embeds over 100 languages, code, images, audio, and video into one vector, with an 8,192-token context and four sizes from 270M to 740M parameters. Matryoshka output dimensions of 768, 512, 256, or 128 are supported, and the model is available in Sentence Transformers and LiteRT-LM, with a reported 14% gain on MTEB Code.

    Why it matters: The release extends an embedding model to text, code, images, audio, and video in one vector, a useful option for retrieval systems that mix media types.

  12. vLLMOfficialAI score60

    vLLM Adds Day-0 Support for Google's EmbeddingGemma 2 Multimodal Embeddings

    AIvLLM announced day-0 support for EmbeddingGemma 2 from Google DeepMind, a bidirectional omni-modal embedding model that maps text, image, audio, video, and interleaved inputs into one vector space. Users can try it with the latest vLLM nightly build using the command vllm serve google/embeddinggemma-2 --runner pooling. The quoted Google post says the model is built on the Gemma 4 architecture and released under Apache 2.0.

    Why it matters: The post gives a runnable serve command and day-0 vLLM support, showing how to deploy the new multimodal embedding model locally.

    Image from @vllm_project's post
  13. clem 🤗XAI score62

    Mistral Large 4 announced with API access today and open weights due end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active parameters. It is available via API today, with open weights planned for the end of October. Clément Delangue, Hugging Face's CEO, reacted by noting that the model cannot be the best open-weight model until its weights are actually released.

    Why it matters: The quoted announcement gives the parameter scale, active count, and availability path, which help readers compare it with other open-weight releases.

  14. Yuchen JinXAI score34

    Reflection's Beam and Mistral Large 4 near GLM-5.2 level

    AIYuchen Jin says Reflection's Beam and Mistral Large 4 both reached roughly GLM-5.2 level within the past two days. He suggests the Western versus Chinese open-source model gap may come down to Chinese labs being able to distill Anthropic and OpenAI models, which Western labs cannot.

  15. eric zakariassonXAI score29

    Cursor adds five SpaceXAI TypeScript SDK demo apps to cookbook

    AICursor has added five apps to its cookbook, all built on the new SpaceXAI TypeScript SDK. The apps cover premise-to-short-film, picture-to-video ads, screenshot-to-React-component, X posts-to-sentiment dashboard, and link-to-podcast conversion. Demos and code are available in the post.

    Video from @ericzakariasson's post
  16. 👩‍💻 Paige BaileyXAI score54

    EmbeddingGemma 2 launches as an Apache 2.0 multimodal embeddings model

    AIGoogle's EmbeddingGemma 2 is an open embeddings model for on-device use that covers code, image, video, audio, and text. It comes in modular sizes from 270M text/code to 740M full multimodal, supports Matryoshka truncation down to 128 dimensions, and reports a 14% gain on MTEB Code over v1 under an Apache 2.0 license. The author's post highlights the release and a Hugging Face demo, while the benchmark table compares it with several models.

    Video from @DynamicWebPaige's post
  17. Dongxi NLPXAI score34

    Mistral AI releases Mistral Large 4, dubbed "Le Chonk"

    AIMistral AI has released Mistral Large 4, a model nicknamed "Le Chonk," according to a post by Dongxi NLP. The post also highlights "sovereign AI" as a keyword, tying the release to the theme of national or independent AI capability. No specifications, benchmarks, or pricing are given in the post itself.

  18. Unsloth AIOfficialAI score62

    Google releases EmbeddingGemma 2, an open multimodal embedding model for on-device use

    AIGoogle released EmbeddingGemma 2, a 740M-parameter open embedding model under Apache 2.0 that combines a 270M text model with vision (170M) and audio (300M) encoders. The 270M text model can run locally with 0.5GB of RAM, and the full multimodal model with 1GB, and Unsloth provides GGUF files and fine-tuning support.

    Why it matters: The post pairs the model's parameter split and local memory footprint with a benchmark table, showing how the multimodal embedding model compares with other embedding models.

    Image from @UnslothAI's post