Skip to contentSkip to stories

Updated

Open source

Showing low-relevance items too. Hide low-relevance items

Jul 27

Jul 27Mon
  1. Air Street PressBlogAI score72

    Black Forest Labs releases FLUX 3, extended to video and robot control

    AIBlack Forest Labs released FLUX 3, a multimodal model trained on images, video, and audio, and mimic built FLUX-mimic on its video backbone to control robots. In a soft-body kitting task, mimic reports a 95% success rate without single-task fine-tuning, compared with 55% for an adapted π0.5 model. FLUX 3 Video is in early access, with action prediction offered to selected partners and an open-weight backbone planned.

  2. Jensen HuangXAI score44

    Nvidia launches Open Secure AI Alliance to strengthen defenders against AI attacks

    AIJensen Huang says defenders need a frontier AI ecosystem combining the best open and closed models, citing how an open-weight frontier model helped contain the Hugging Face intrusion when closed AI blocked essential forensics. Nvidia has created the Open Secure AI Alliance to develop new techniques and tools for safeguarding software and agents by sharing models, tooling, and research in the open.

Jul 26

Jul 26Sun
  1. Fireworks AI BlogOfficialAI score60

    Fireworks AI adds open-weight Kimi K3 with US-only serverless endpoints

    AIFireworks AI made the open-weight Kimi K3 available for inference and training on its platform, with US-only serverless endpoints and Zero Data Retention. In its own head-to-head with Opus 5, the post reports K3 at 92.7% accuracy and $0.52 per task on SWE (480) against Opus 5's 94.8% and $1.05, with the vendor claiming up to 5x better cost efficiency per task.

    Why it matters: The post compares Kimi K3 with Opus 5 on accuracy and cost per task, giving readers concrete figures to judge the open model against closed alternatives for their own workloads.

  2. Jeremy HowardXAI score16

    Jeremy Howard criticizes employees who ignore their employer's interests

    AIJeremy Howard argues that many people with otherwise sound judgment seem unable to think clearly about actions that affect the company paying them. The post is a general observation and cites no specific company, product, or event. Its quoted context about Nvidia's CUDA and GPU driver open source release is not the post's main subject.

Jul 25

Jul 25Sat
  1. Fireworks AI BlogOfficialAI score51

    Fireworks AI Enables LoRA Training on Kimi K3 in Private Preview

    AIFireworks AI has made Kimi K3 available for Multi-LoRA serving and training in private preview through Fireworks Serverless Training. The post explains how small LoRA adapters can be trained on K3 and served with live merge or multi-LoRA deployment, and it reports two example tasks, Countdown and Frozen Lake, with reward curves.

Jul 24

Jul 24Fri
  1. Mira MuratiXAI score16

    Murati says useful AI knowledge must be distributed, backing Jensen Huang's vision

    AIMira Murati argues that the knowledge making AI useful is spread across scientists, engineers, clinicians, and firms, so AI must itself be distributed to benefit from it. She says she agrees with Jensen Huang that this is a future worth building. The post accompanies Huang's shared NVIDIA letter arguing that open models strengthen safety, cybersecurity, innovation, and sovereignty alongside frontier closed models.

  2. Bryan CatanzaroXAI score36

    Bryan Catanzaro argues open AI models should be treated as infrastructure

    AIBryan Catanzaro, NVIDIA's account owner, argues the central US AI leadership question is whether AI models will be treated as infrastructure like the internet or electricity. He says open models will be at the heart of this infrastructure, enabling companies from startups to established industry leaders, and making sovereignty possible. He concludes policymakers seeking to keep American AI at the forefront should recognize open models as the critical infrastructure of the AI age.

Jul 23

Jul 23Thu
  1. Sequoia CapitalBlogAI score58

    Western AI Builders Depend on Chinese Open Models Through Distillation

    AIThe essay argues that Western companies increasingly rely on Chinese open-weight models like Qwen and Kimi for post-training, while Western labs cannot lawfully distill from American frontier models. It says Qwen's share of new open-model fine-tunes rose from 1% in January 2024 to 69% by February 2026, citing ATOM's Report. The authors propose controlled teacher access and tighter enforcement against foreign distillation as a domestic alternative.

  2. Matei ZahariaXAI score36

    Berkeley STAR Lab packages AI research optimizers into one GEPA API

    AIBerkeley's STAR Lab packaged multiple LLM-based "autoresearch" algorithms into a single API within the GEPA package, letting users mix and match them. The optimizers can be applied to tasks including prompt writing, agent design, and code optimization. The quoted thread adds that GEPA, AutoResearch, and Meta-Harness each win on different tasks, and that the new optimize_anything omni meta-optimizer beats every standalone optimizer at a matched budget.

  3. Gemini NotebookOfficialAI score34

    Gemini Notebook rolls out Collections for web users

    AIGoogle's Gemini Notebook has rolled out Collections to 100% of web users, letting people group notebooks like photo albums or playlists. Notebooks can belong to multiple collections or stay only in the main "My Notebooks" tab, with no rigid folder structure. The account asks users which features they want next.

  4. Bryan CatanzaroXAI score33

    NVIDIA says it is now HuggingFace's biggest institutional contributor

    AINVIDIA says it has become the largest institutional contributor on HuggingFace and expects to keep publishing open data, techniques, and models. The company frames the effort as enabling organizations to build and deploy AI their own way, and as serving its own interests because AI growth expands NVIDIA's opportunities.

    Image from @ctnzr's post
  5. BAAI · new models on Hugging FaceOfficialAI score62

    BAAI releases AREX-Base, a 122B deep research agent model

    AIBAAI has released AREX-Base, a 122B-total, 10B-activated Mixture-of-Experts deep research agent built on Qwen3.5-122B-A10B with a 262,144-token context. The model uses an inner research loop and an outer self-improvement loop, and the source reports it scoring 82.5 on BrowseComp and 85.4 on GAIA, under Apache 2.0.

    Why it matters: The release pairs a 122B-parameter deep research agent with benchmark tables against frontier and open models, letting readers compare its search-agent results directly.

  6. Andrew NgXAI score65

    Andrew Ng announces OpenWorker, an open-source agent that delivers finished work

    AIAndrew Ng and Rohit Prasad announced OpenWorker, an open-source agent that produces deliverables such as documents, Slack messages, and calendar updates across files and everyday tools. It checks in before consequential actions, runs on Mac with Windows support coming soon, and works with user-supplied API keys for models including GPT 5.6 Sol, Claude Fable, Gemini 3.6, open-weight models, or local Ollama models. Source code is available on GitHub, and the tool requires the user's own API key.

    Video from @AndrewYNg's post

Jul 21

Jul 21Tue
  1. Soumith ChintalaXAI score45

    Soumith Chintala says Poolside's Laguna S 2.1 suits agentic work on DGX Spark

    AISoumith Chintala praised Poolside's Laguna S 2.1 as looking strong for agentic use and said it fits on a single NVIDIA DGX Spark. The quoted Poolside release describes it as a 118B total-parameter Mixture-of-Experts model with 8B active per token, up to 1M-token context, and thinking and no-thinking modes, with weights openly available under OpenMDW-1.1.

  2. Bryan CatanzaroXAI score57

    Poolside releases open-weight Laguna S 2.1 for agentic coding

    AIPoolside released Laguna S 2.1, an open-weight model with 118B total parameters and 8B active per token. The author says it performs strongly on agentic coding and long-horizon tasks, and it can run on a single NVIDIA DGX Spark. Weights are on Hugging Face under the OpenMDW-1.1 license, with access also available through OpenRouter and Poolside's API.

  3. JetBrains AI BlogOfficialAI score55

    JetBrains Air adds ACP agents, local models, and Java/Kotlin code intelligence

    AIJetBrains Air now connects to ACP-compatible coding agents, including GitHub Copilot CLI, OpenCode, Pi, and Cline, through the Agent Client Protocol. The release also adds Beta Java and Kotlin navigation and diagnostics powered by the IntelliJ IDEA code engine, local model support through Ollama or LM Studio, and Docker-based agent tasks on Windows.

  4. Meta AI BlogOfficialAI score44

    Meta's SAM 3 and DINOv3 Power SYNAPS-I's Genesis Mission Imaging Pipeline

    AISYNAPS-I, a multi-lab Genesis Mission project led by Lawrence Berkeley National Laboratory, uses Meta's open-source SAM 3 and DINOv3 models to segment X-ray and micro-CT scientific imagery. The fine-tuned pipeline, run on 300 A100 GPUs, reduced a grapevine xylem analysis from a month of expert annotation per time step to about 15 minutes. The team can deploy the open models inside secure national lab infrastructure, where research data must remain.

Jul 20

Jul 20Mon

Jul 18

Jul 18Sat

Jul 17

Jul 17Fri

Jul 16

Jul 16Thu
  1. Soumith ChintalaXAI score60

    Kimi K3 launches as a 2.8 trillion parameter open-weight model

    AIMoonshot AI announced Kimi K3, a native multimodal model with 2.8 trillion parameters and a 1 million token context window. The announcement cites up to 6.3x faster decoding in million-token contexts and about 25% higher training efficiency, and says open weights arrive by July 27, 2026. The author, Soumith Chintala, reposted it with a brief note of congratulations.

  2. Mistral AI · new models on Hugging FaceOfficialAI score46

    Mistral releases Shieldstral-1.0-3B, a policy-adaptive multimodal safety classifier

    AIMistral AI released Shieldstral-1.0-3B, a 3B-parameter multimodal safety classifier that judges content against natural-language policies and outputs a continuous safety score. It moderates text, image, and text-plus-image content in a single forward pass and can be retargeted to new policies at inference time without retraining. The Apache 2.0 open-weight model is built on Ministral-3-3B-Base-2512 and trained on sequences up to 32k tokens.

Jul 15

Jul 15Wed
  1. Junyang LinXAI score42

    Junyang Lin Asks Whether Inkling's Small Model Is Open-Sourced

    AIJunyang Lin praised Thinking Machines' new architecture in Inkling, which reasons across text, image, and audio, and asked whether the small model is open-sourced. The quoted announcement says the full weights are available and that Inkling is available for fine-tuning on Tinker.

  2. Soumith ChintalaXAI score38

    Modal trains DFlash speculator, faster than MTP for inference

    AIModal has trained a DFlash speculator that runs much faster than MTP, according to Soumith Chintala. The speculator is backed by Inkling by Thinking Machines, which Modal says delivers 67% higher throughput and interactivity on Modal Auto Endpoints with SGLang.

    Image from @soumithchintala's post
  3. Leandro von WerraXAI score42

    Thinking Machines releases Inkling, a multimodal model with open weights

    AIThinking Machines has introduced Inkling, a model that reasons across text, image, and audio, with full weights made available. It is available for fine-tuning on Tinker and can be tried in the Inkling Playground. Hugging Face's Leandro von Werra praised the release for its grounded writing, interesting details, and strong ecosystem integration.