Skip to contentSkip to stories

Updated

#Open source/Repo

Mar 11

Mar 11Wed
  1. Mistral AI · new models on Hugging FaceAI score62

    Mistral AI releases Leanstral-2603, an open-source Lean 4 proof agent

    AIMistral AI released Leanstral 119B A6B on Hugging Face as an open-source code agent for Lean 4 proof engineering. The model uses 128 experts with 4 active per token, 6.5B activated parameters, a 256k token context window, and accepts text and image input under the Apache 2.0 license. The page also documents vLLM server deployment and Mistral Vibe integration.

    Why it matters: The source specifies Leanstral's 119B MoE architecture, 256k context, Apache 2.0 license, and vLLM setup, showing how the Lean 4 proof agent could be deployed locally.

Mar 9

Mar 9Mon
  1. Black Forest Labs · new models on Hugging FaceAI score39

    Black Forest Labs releases FLUX.2 [klein] 9B-KV with KV-cache for faster multi-reference editing

    AIBlack Forest Labs has released FLUX.2 [klein] 9B-KV, a variant of FLUX.2 [klein] 9B that caches reference-image key-value pairs to speed up multi-reference editing by up to 2.5 times. The 9B flow model, which uses an 8B Qwen3 text embedder and is step-distilled to 4 inference steps, is available for non-commercial use under the FLUX Non-Commercial License and fits in about 29GB VRAM.

Mar 4

Mar 4Wed
  1. Mistral AI · new models on Hugging FaceAI score67

    Mistral Small 4 unifies instruct, reasoning, and coding in one open model

    AIMistral Small 4 combines instruct, reasoning, and Devstral capabilities in one multimodal model with 119B total parameters, 6.5B active per token, and a 256k context window. The source reports a 40% reduction in latency-optimized end-to-end completion time and 3x more requests per second in throughput-optimized setups versus Mistral Small 3. It is released under Apache 2.0 and supports reasoning mode toggling per request.

    Why it matters: The source lists architecture, context length, and mode-switching controls, letting readers compare this release's design with earlier Mistral Small models.

Feb 10

Feb 10Tue
  1. Z.ai (GLM) · new models on Hugging FaceAI score72

    Z.ai releases GLM-5, a 744B-parameter open model for agentic engineering

    AIZ.ai launches GLM-5, scaling from 355B to 744B total parameters with 40B active and pre-training data from 23T to 28.5T tokens. The model integrates DeepSeek Sparse Attention to reduce deployment cost and reports strong results on reasoning, coding, and agentic benchmarks against GLM-4.7, DeepSeek-V3.2, Kimi K2.5, and several frontier models.

    Why it matters: The source gives concrete scale, data, and benchmark comparisons against named frontier models, showing where GLM-5 sits among open-source and proprietary systems.

Jan 29

Jan 29Thu
  1. Z.ai (GLM) · new models on Hugging FaceAI score60

    Z.ai releases open-source GLM-OCR multimodal document model

    AIZ.ai has released GLM-OCR, a 0.9B-parameter multimodal OCR model for complex document understanding, under the MIT License. The model scores 94.62 on OmniDocBench V1.5 and supports deployment through vLLM, SGLang, and Ollama, with an official SDK for document parsing.

    Why it matters: The page gives benchmark scores, a 0.9B parameter size, and supported serving frameworks, which help readers weigh OCR deployment options against heavier alternatives.

Jan 23

Jan 23Fri
  1. Mistral AI · new models on Hugging FaceAI score67

    Mistral Small 4 unifies instruct, reasoning, and coding in one open model

    AIMistral Small 4 is a 119B-parameter MoE model with 6.5B active per token and a 256k context window, combining instruct, reasoning, and Devstral-style coding in one model. It accepts text and image input, lets users set reasoning_effort per request, and is released under Apache 2.0. The model card reports a 40% latency reduction and 3x throughput versus Mistral Small 3 in its tested setups, and its benchmark chart shows reasoning scores on GPQA Diamond, MMLU Pro, AIME-style text tasks, and MMMU-Pro.

    Why it matters: The model card names concrete architecture, context, and licensing details, letting readers compare its reasoning toggle and efficiency claims against other open models.

Jan 21

Jan 21Wed
  1. Mistral AI · new models on Hugging FaceAI score65

    Mistral releases open-weight Voxtral Mini 4B Realtime 2602 speech model

    AIMistral AI released Voxtral Mini 4B Realtime 2602, a multilingual realtime speech-transcription model with 13 supported languages under the Apache 2.0 license. The model has a configurable transcription delay from 240ms to 2.4s, and it matches leading offline open-source models at a 480ms delay. The source says it is optimized for on-device deployment and is currently supported only in vLLM.

    Why it matters: The source specifies the 480ms delay operating point, 4B size, Apache 2.0 license, and vLLM serving path, which matter for teams weighing realtime transcription deployment.

Jan 19

Jan 19Mon
  1. Z.ai (GLM) · new models on Hugging FaceAI score62

    Z.ai releases GLM-4.7-Flash, a 30B-A3B MoE model for lightweight deployment

    AIZ.ai has released GLM-4.7-Flash, a 30B-A3B MoE model that it positions as the strongest model in the 30B class. The model reports SWE-bench Verified 59.2 and τ²-Bench 79.5, and supports local deployment through vLLM and SGLang.

    Why it matters: The source lists benchmark scores against Qwen3-30B-A3B-Thinking-2507 and GPT-OSS-20B, letting readers compare the 30B-class MoE model directly with its named rivals.

Jan 14

Jan 14Wed
  1. Black Forest Labs · new models on Hugging FaceAI score62

    Black Forest Labs releases FLUX.2 [klein] 4B image model under Apache 2.0

    AIBlack Forest Labs released FLUX.2 [klein] 4B, a 4 billion parameter model that unifies text-to-image generation and image editing with multi-reference support. The source says it runs on consumer GPUs such as the RTX 3090 or 4070 with about 13GB VRAM, and its open weights are available under the Apache 2.0 license.

    Why it matters: The source specifies a 4 billion parameter model running on about 13GB VRAM under Apache 2.0, which helps readers judge whether local image generation fits their hardware.

Jan 1

Jan 1Thu
  1. Moonshot AI (Kimi) · new models on Hugging FaceAI score75

    Moonshot AI releases open-source multimodal agent model Kimi K2.5

    AIMoonshot AI released Kimi K2.5, an open-source native multimodal agentic model built by continual pretraining on about 15 trillion mixed visual and text tokens. The model card reports a 1T-parameter Mixture-of-Experts architecture with 32B activated parameters and a 256K context length, and it lists benchmark results against GPT-5.2, Claude 4.5 Opus, Gemini 3 Pro, DeepSeek V3.2, and Qwen3-VL-235B-A22B-Thinking. Weights and code are released under a Modified MIT License, with API access on the Moonshot platform.

    Why it matters: The model card gives a full benchmark table against GPT-5.2, Claude 4.5 Opus, and Gemini 3 Pro, useful for comparing open multimodal agent models.

Dec 20, 2025

Dec 20, 2025Sat
  1. MiniMax · new models on Hugging FaceAI score74

    MiniMax-M2.1 open-sources weights for coding and agent tasks

    AIMiniMax has released MiniMax-M2.1 model weights on Hugging Face, with API access on the MiniMax Open Platform and the MiniMax Agent product. The company reports gains over M2 on coding and agent benchmarks such as SWE-bench Verified (74.0) and VIBE average (88.6), and says it outperforms Claude Sonnet 4.5 on multilingual scenarios.

    Why it matters: The release pairs open weights with a broad benchmark table against Claude and GPT models, letting readers compare coding and agent claims directly.

Dec 16, 2025

Dec 16, 2025Tue
  1. MiniMax · new models on Hugging FaceAI score38

    MiniMax Releases VTP-Large-f16d64 Visual Tokenizer With Technical Report and Pretrained Weights

    AIMiniMax released the technical report and pretrained weights for VTP-Large-f16d64, a visual tokenizer that jointly optimizes contrastive, self-supervised, and reconstruction losses. The model scores 78.2 zero-shot accuracy, 85.7 linear probing, and 0.36 rFID, and its generation performance scales with pretraining compute, parameters, and data. Checkpoint weights were listed as "released very soon" in the source.

  2. Xiaomi MiMoAI score78

    Xiaomi releases open-source MiMo-V2-Flash MoE model for reasoning and coding

    AIXiaomi released and open-sourced MiMo-V2-Flash, a Mixture-of-Experts model with 309B total and 15B active parameters, under the MIT license. The company reports 73.4% on SWE-Bench Verified, the top score among open-source models, and inference at 150 tokens per second for $0.1 per million input tokens and $0.3 per million output tokens. It supports a hybrid thinking mode and a 256k context window.

    Why it matters: The post gives architecture, speculative decoding speedup, and pricing figures, which help readers judge how the efficiency claims are achieved and what they cost.

Dec 12, 2025

Dec 12, 2025Fri
  1. Apple · new models on Hugging FaceAI score46

    Apple's SHARP Turns a Single Photo into a 3D Scene in Under a Second

    AIApple has released SHARP, a model that generates a 3D Gaussian representation of a scene from a single photograph in less than a second on a standard GPU. The output renders in real time as high-resolution photorealistic views of nearby camera positions, with metric absolute scale, and the paper reports reductions of 25–34% in LPIPS and 21–43% in DISTS versus the best prior model.

Dec 11, 2025

Dec 11, 2025Thu
  1. OpenAI · new models on Hugging FaceAI score42

    OpenAI Releases circuit-sparsity Sparse Model Weights on Hugging Face

    AIOpenAI has published weights for a sparse model from Gao et al. 2025, used for qualitative results on bracket counting and variable binding, on Hugging Face under the openai/circuit-sparsity repository. The release includes a standalone Hugging Face implementation that loads the converted model and tokenizer with trust_remote_code and runs sample generation. The project is licensed under Apache License 2.0.

Dec 10, 2025

Dec 10, 2025Wed
  1. FunAudioLLM (Alibaba Tongyi) · new models on Hugging FaceAI score42

    Fun-CosyVoice3-0.5B-2512 Released as Open-Source Multilingual Text-to-Speech Model

    AIAlibaba's FunAudioLLM has released Fun-CosyVoice3-0.5B-2512, a 0.5B-parameter LLM-based text-to-speech model on Hugging Face, with an RL variant also published. The model supports zero-shot voice cloning across 9 languages and 18+ Chinese dialects and accents, with streaming output at latency as low as 150ms. On the source's test-en benchmark, it reports a 2.24% WER and 71.8% speaker similarity, and the RL version reports 1.68% WER.

Dec 2, 2025

Dec 2, 2025Tue
  1. Apple · new models on Hugging FaceAI score36

    Apple releases CLaRa-7B-Instruct for compressed-document retrieval-augmented QA

    AIApple has published CLaRa-7B-Instruct on Hugging Face, an instruction-tuned unified RAG model with built-in semantic document compression at 16× and 128× ratios. The model answers instruction-following questions directly from compressed document representations, and its paper, GitHub repository, and transformers usage example are referenced in the release.

Nov 4, 2025

Nov 4, 2025Tue
  1. Moonshot AI (Kimi) · new models on Hugging FaceAI score82

    Moonshot AI releases open-source Kimi K2 Thinking reasoning agent model

    AIMoonshot AI released Kimi K2 Thinking, an open-source thinking model that interleaves step-by-step reasoning with tool calls across 200 to 300 sequential invocations. The model is a 1T-parameter mixture-of-experts with 32B activated parameters and a 256k context window, and it uses native INT4 quantization for roughly 2x faster generation. The model card reports benchmark results on HLE, BrowseComp, and other tests, and recommends vLLM, SGLang, or KTransformers for deployment.

    Why it matters: The model card gives benchmark tables, quantization details, and deployment settings, letting readers compare Kimi K2 Thinking against GPT-5 and other models on specific tasks.

Oct 30, 2025

Oct 30, 2025Thu
  1. Moonshot AI (Kimi) · new models on Hugging FaceAI score60

    Moonshot AI releases Kimi Linear 48B hybrid linear attention models on Hugging Face

    AIMoonshot AI released Kimi Linear, a hybrid linear attention architecture with 48B total and 3B activated parameters and a 1M-token context length, on Hugging Face. The model card reports up to 6.3x faster TPOT than MLA at 1M tokens and up to 75% lower KV cache needs, and says it outperforms full attention on long-context and RL-style benchmarks.

    Why it matters: The model card gives concrete long-context speed and memory figures for a hybrid attention design, useful for judging whether linear attention can replace full attention in practice.

Jun 22, 2025

Jun 22, 2025Sun
  1. Cognition Blog (Devin, Windsurf)AI score57

    Cognition details blockdiff, an open-source file format for instant VM disk snapshots

    AICognition built and open-sourced blockdiff, a file format that creates block-level diffs of VM disks using only filesystem metadata. The company reports its otterlink hypervisor cut snapshot times from 30 to 60 minutes on EC2 to about 5 to 10 seconds for a 128 GB disk with a 5 GB diff, roughly a 200x speedup. The post also covers the sparse file and copy-on-write concepts behind the approach and why OverlayFS, ZFS, and qcow2 were not chosen.

May 21, 2025

May 21, 2025Wed
  1. Cognition Blog (Devin, Windsurf)AI score62

    Cognition launches official DeepWiki MCP server for indexed GitHub repos

    AICognition launched the official DeepWiki Model Context Protocol server, which is free and requires no login or authentication. It gives programmatic access to ask_question, read_wiki_contents, and read_wiki_structure for GitHub repositories indexed on DeepWiki.com. Private repositories require a Devin account with GitHub connected, and open-source maintainers can apply for $500 in Devin credits.

    Why it matters: The source names the three tools and the access path, showing how indexed GitHub repositories can be queried programmatically without login.

May 18, 2025

May 18, 2025Sun
  1. Cognition Blog (Devin, Windsurf)AI score22

    Cognition Revives Devin Open Source Initiative With $500 Credits for Projects

    AICognition is bringing back its Devin Open Source Initiative, offering $500 in Devin ACU credits to open-source GitHub projects with over 100 forks. Projects below that threshold will still be considered. Eligible projects must have an OSI-approved license and be actively maintained, and maintainers can apply through a linked form.

May 4, 2025

May 4, 2025Sun
  1. Cognition Blog (Devin, Windsurf)AI score42

    DeepWiki launches free AI-generated docs for public GitHub repositories

    AICognition has launched DeepWiki, a free public version of its Devin Wiki and Devin Search tools that helps developers understand codebases. Users can view docs for any repo by replacing github.com with deepwiki.com in the URL, and more than 50,000 top public GitHub repositories are already indexed. Private repositories require a Devin account.

Dec 11, 2024

Dec 11, 2024Wed
  1. Cognition Blog (Devin, Windsurf)AI score38

    Devin Open Source Initiative Gives Maintainers 500 Free ACUs for Repo Work

    AICognition is launching the Devin Open Source Initiative, giving selected open source maintainers 500 free ACUs on a Devin Teams plan as part of Devin's general availability launch. The post shows Devin contributing pull requests to projects including Anthropic's MCP Inspector, Dagger, and nanoGPT, with maintainers still reviewing the results. Devin's GitHub integration forwards PR comments and CI checks to help refine changes, though the company warns a human should still verify final quality.