Skip to contentSkip to stories

Updated

#Open source/Repo

Showing low-relevance items too. Hide low-relevance items

Jun 4

Jun 4Thu
  1. Cohere · new models on Hugging FaceOfficialAI score60

    Cohere releases North Mini Code 1.0, a 30B-A3B open-weights coding model

    AICohere and Cohere Labs released North Mini Code 1.0, an open-weights 30B-A3B mixture-of-experts model for code generation and agentic terminal tasks, under Apache 2.0. The model has 256K context and 64K max output, and is trained for tool use. Its benchmark table lists Terminal-Bench v2 at 36.0, SWE-Bench Verified at 67.6, and LiveCodeBench v6 at 70.3, below Qwen3.6 on several tasks.

    Why it matters: The card lists benchmark results against Qwen3.6, Gemma4, and other models, showing where North Mini Code trails on some coding and agentic tasks.

Jun 3

Jun 3Wed

Jun 2

Jun 2Tue
  1. MiniMax · new models on Hugging FaceOfficialAI score68

    MiniMax releases M3, a native multimodal model with 1M context

    AIMiniMax has released MiniMax-M3, a native multimodal model with a 1M-token context window, roughly 428B total parameters, and about 23B activated parameters. The model introduces MiniMax Sparse Attention, which the source says delivers 9× prefill and 15× decode speedups over M2 at 1M context. M3 supports enabled, adaptive, and disabled reasoning modes through the thinking parameter, and weights are available on Hugging Face.

    Why it matters: The source gives concrete attention-efficiency figures and three reasoning modes, which helps readers judge long-context cost against deployment choices.

  2. ByteDance · new models on Hugging FaceOfficialAI score44

    ByteDance Releases Bernini-R Diffusers Weights for Video Generation and Editing

    AIByteDance has open-sourced the inference code and model weights of the Bernini Renderer (Bernini-R), a DiT-based renderer paired with an MLLM-based semantic planner for video generation and editing. A diffusers-format version, ByteDance/Bernini-R-Diffusers, bundles the Wan2.2 base components with the Bernini-R transformer weights for direct loading, and the framework requires a CUDA GPU with PyTorch 2.5.1+cu124.

Jun 1

Jun 1Mon
  1. PaddlePaddleOfficialAI score36

    PaddleOCR and ERNIE Image now available as official Dify plugins

    AIPaddleOCR and ERNIE Image are now available as official Dify plugins, bringing document parsing and image generation into Dify's agent workflows. PaddleOCR, powered by PP-OCRv5, PP-StructureV3, and PaddleOCR-VL, turns images, scanned PDFs, and multilingual documents into structured data for chunking, vectorization, and RAG, with private or on-prem deployment supported. ERNIE Image offers free generation, a Turbo mode with 8-step inference, and an OpenAI-style API.

    Image from @PaddlePaddle's post

May 28

May 28Thu

May 27

May 27Wed

May 26

May 26Tue

May 20

May 20Wed
  1. Stability AIOfficialAI score62

    Stability AI releases Stable Audio 3.0 model family with open-weight music models

    AIStability AI released Stable Audio 3.0, a family of four audio models trained on fully licensed data. Three of them, Small SFX, Small and Medium, have open weights on Hugging Face, while Large is available through the Stability AI API and enterprise self-hosting. Outputs can be distributed and commercialized under the Stability AI Community License, and organizations with more than $1M in annual revenue can use the Enterprise License.

    Why it matters: The source specifies which models are open-weight, their licensing terms, and clip-length limits, which matters for anyone deciding whether to build on them.

  2. PaddlePaddleOfficialAI score36

    PaddleOCR 3.5 adds Hugging Face Transformers as inference backend

    AIPaddleOCR 3.5 now supports Hugging Face Transformers as an inference backend, letting users run PP-OCRv5 and PaddleOCR-VL 1.5 models directly within the Transformers ecosystem. Users can select it with engine="transformers" while keeping the same PaddleOCR pipeline, which the post says eases integration for RAG and Document AI applications.

May 19

May 19Tue
  1. koray kavukcuogluXAI score72

    Google rolls out Gemini 3.5 Flash globally across consumer, developer, and enterprise platforms

    AIGoogle is rolling out Gemini 3.5 Flash globally for consumers in the Gemini app and Search AI Mode. It is also available to developers through the Gemini API, Google Antigravity, and Google AI Studio, and to businesses on the Gemini Enterprise Agent Platform.

    Why it matters: The post shows where each Gemini 3.5 Flash access path goes, from consumer apps to developer and enterprise platforms, which helps readers pick the right entry point.

May 15

May 15Fri
  1. Fidji SimoXAI score60

    ChatGPT adds a personal finance preview for U.S. Pro users

    AIChatGPT is previewing a personal finance experience for Pro users in the U.S., who can securely connect financial accounts and see where their money is going. Users can ask questions based on the information they choose to connect, and the author says this follows the similar health records connection feature.

May 8

May 8Fri

May 7

May 7Thu
  1. Sam BowmanXAI score38

    Anthropic donates open-source alignment testing tool Petri to Meridian Labs

    AIAnthropic is donating Petri, its open-source interactive behavioral-evals tool for alignment testing, to Meridian Labs so development can continue independently. Working with Meridian, Anthropic has also released a major update improving the adaptability, realism, and depth of Petri's tests. Developers are invited to try the tool and contribute.

Apr 28

Apr 28Tue

Apr 27

Apr 27Mon
  1. Mistral AI · new models on Hugging FaceOfficialAI score36

    Mistral Medium 3.5 EAGLE draft model released for speculative decoding on Hugging Face

    AIMistral AI has released mistralai/Mistral-Medium-3.5-128B-EAGLE, an EAGLE draft model for speculative decoding with the 128B dense Mistral Medium 3.5. The companion model, which the source says replaces Mistral Medium 3.1 and Magistral in Le Chat and Devstral 2 in Vibe, has a 256k context window, handles text and image input with text output, and is served with vLLM or SGLang using three speculative tokens. The model is released under a Modified MIT License that allows commercial use with exceptions for companies with large revenue.

Apr 26

Apr 26Sun
  1. Xiaomi MiMoOfficialAI score87

    Xiaomi releases open-source MiMo-V2.5-Pro for long-horizon agentic coding

    AIXiaomi released and open-sourced MiMo-V2.5-Pro, a 1.02T-parameter Mixture-of-Experts model with 42B active parameters and a 1M-token context window. The company reports gains in agentic tasks, software engineering, and long-horizon work, including a Rust SysY compiler task finished in 4.3 hours across 672 tool calls. Weights and tokenizer are on Hugging Face, and API pricing is unchanged.

    Why it matters: The release pairs a 1.02T-parameter open-weight model with long-horizon agent results and token-efficiency claims, useful for judging its fit in coding and agent workflows.

Apr 23

Apr 23Thu
  1. Apple · new models on Hugging FaceOfficialAI score40

    Apple releases CADD-Base-7B, a masked diffusion model for code generation

    AIApple has released CADD-Base-7B on Hugging Face, a 7B masked diffusion language model for code generation that uses Continuously Augmented Discrete Diffusion (CADD) to guide discrete denoising with a continuous flow-matching signal. The model loads through Transformers with trust_remote_code, and its diffusion_generate method supports CADD sampling modes "weighted" and "argmax" with alg options such as "entropy" and "maskgit_plus". The release builds on DiffuCoder and reuses Dream's modeling architecture and generation utilities.

  2. OpenAI Alignment Research BlogOfficialAI score44

    OpenAI Open-Sources Chain-of-Thought Monitorability Evaluation Datasets and Code

    AIOpenAI is releasing a subset of datasets, reference code, and the g-mean 2 metric for evaluating chain-of-thought monitorability. The release includes most datasets from its monitorability suite, while some evaluations relying on private or restricted data were omitted. The company says it will keep reporting monitorability results in future frontier reasoning model system cards.

Apr 21

Apr 21Tue
  1. NVIDIA AI DeveloperOfficialAI score29

    NVIDIA OpenShell v0.0.34 adds live sandbox policy updates and VM installs

    AINVIDIA's OpenShell v0.0.34 release lets users update sandbox policy without restarting the runtime. The update also adds install-vm, which installs the gateway and VM driver with new --driver-dir support, and sandbox get, which shows the active runtime policy. Supervisor seccomp improvements and HTTP normalization are included as well.

Apr 20

Apr 20Mon

Apr 17

Apr 17Fri
  1. OpenAI · new models on Hugging FaceOfficialAI score41

    OpenAI Releases Privacy Filter, an Open-Weight PII Detection Model on Hugging Face

    AIOpenAI released Privacy Filter, a bidirectional token-classification model that detects and masks personally identifiable information in text under the Apache 2.0 license. The model has 1.5B total parameters with 50M active, supports a 128,000-token context window, and can run in a web browser or on a laptop. Users can fine-tune it and adjust precision/recall tradeoffs through preset operating points.

Apr 13

Apr 13Mon
  1. BAAIOfficialAI score40

    ClawKeeper v1.0 releases open-source security framework for OpenClaw AI agents

    AIBAAI announces ClawKeeper v1.0, an open-source security framework for OpenClaw AI agents, combining Skill-based command policies, Plugin-based runtime monitoring, and a Watcher system-level observer. The independent Watcher is designed to block high-risk operations such as prompt injections, key leaks, rogue commands, and remote code execution, even if the agent is compromised. The paper is available on arXiv and the project code is hosted on GitHub.

Apr 9

Apr 9Thu
  1. Awni HannunXAI score24

    Running YOLO26 Locally on Apple Silicon with MLX

    AIA blog post and release show how to run YOLO26 locally using MLX. The linked background post describes YOLO26-MLX as a native Apple Silicon port without PyTorch or an external GPU, with up to 2.6x faster inference and up to 1.7x faster training.

    Image from @awnihannun's post

Apr 6

Apr 6Mon
  1. Tri DaoXAI score32

    Fast Muon optimizer coming to Blackwell consumer GPUs

    AITri Dao says a fast Muon optimizer is coming to consumer cards, since its symmetric matmul kernels work once Blackwell consumer GPU mainloop support is in place. Background from @jcz42 reports Gram Newton-Schulz symmetric kernels now support RTX 5090, with 2x faster Newton-Schulz and 1.7x faster optimizer time on 15 layers of Gemma-4 E2B.

  2. Black Forest Labs · new models on Hugging FaceOfficialAI score41

    FLUX.2 Small Decoder offers faster, lower-VRAM drop-in replacement for FLUX.2 decoder

    AIBlack Forest Labs released FLUX.2 Small Decoder, a distilled VAE decoder that works as a drop-in replacement for the standard FLUX.2 decoder on Hugging Face. It decodes about 1.4x faster and uses about 1.4x less VRAM at decode time, with ~28M decoder parameters versus ~50M in the full decoder and minimal quality loss. It is available under the Apache 2.0 license and is compatible with FLUX.2-klein-4B, FLUX.2-klein-9B, FLUX.2-klein-9b-kv, and FLUX.2-dev.

Apr 1

Apr 1Wed
  1. Jim FanXAI score62

    CaP-X open-sources agentic robotics toolkit, benchmark, and RL setup

    AIJim Fan announced the open-source release of CaP-X, an agentic robotics framework in which LLM-driven agents control robot arms and humanoids through perception and actuation APIs. The release includes CaP-Gym with 187 manipulation tasks across RoboSuite, LIBERO-PRO, and BEHAVIOR, and CaP-Bench, which evaluates 12 frontier LLMs and VLMs across 8 tiers. The post also reports that a 7B open-source model rose from 20% to 72% success after 50 RL training iterations, with synthesized programs transferring to real robots.

    Video from @DrJimFan's post

Mar 30

Mar 30Mon

Mar 26

Mar 26Thu
  1. Intern Large ModelsOfficialAI score44

    DataChef: RL framework auto-generates data recipes for LLM adaptation

    AIDataChef, an AI4AI framework, uses reinforcement learning to automatically generate optimal data recipes for adapting LLMs. Its DataChef-32B model, using an efficient proxy reward system, matches Gemini-3-Pro in recipe generation, with its recipes surpassing expert-curated ones on AIME'25 and ClimaQA benchmarks.

    Image from @intern_lm's post

Mar 22

Mar 22Sun
  1. FunAudioLLM (Alibaba Tongyi) · new models on Hugging FaceOfficialAI score32

    PrismAudio Adds Reinforcement Learning to Video-to-Audio Generation with Chain-of-Thought Planning

    AIPrismAudio is a framework that integrates reinforcement learning into video-to-audio generation, using a Chain-of-Thought planning mechanism. It builds on ThinkSound by splitting single-step reasoning into four CoT modules for semantic, temporal, aesthetic, and spatial dimensions, each with targeted reward functions. Code, model weights, and datasets are released for research and educational use under the MIT License, and commercial use requires explicit author authorization.

Mar 20

Mar 20Fri
  1. Aman SangerXAI score22

    Cursor's Composer 2 model praised, built on an open-source base

    AIAman Sanger of Cursor says Composer 2 is a really good model and he is excited for more people to try it. The quoted reply from Lee Robinson says Composer 2 started from an open-source base, with only about one-quarter of the final model's compute coming from that base. Cursor plans full pretraining in the future and says it is following the license through its inference partner terms.

Mar 17

Mar 17Tue
  1. Apple · new models on Hugging FaceOfficialAI score44

    Apple releases SimpleSD-30B-instruct, a self-distilled Qwen code model for research

    AIApple has released apple/SimpleSD-30B-instruct, a research checkpoint built on Qwen that uses Simple Self-Distillation to improve code generation without rewards, verifiers, or teacher models. On LiveCodeBench, the model scores 55.3% pass@1 on LCBv6 versus 42.4% for its base, Qwen3-30B-A3B-Instruct-2507. The checkpoints are for reproducibility, not optimized Qwen releases, and are available under the Apple Machine Learning Research Model License.