Skip to contentSkip to stories

Updated

#Open source/Repo

Showing low-relevance items too. Hide low-relevance items

Oct 2

Oct 2Fri
  1. Guillermo RauchXAI score34

    Muse Ships Open-Source ESP32 Firmware and Linux SDK for Gadgets

    AIMuse has released Muse Gadgets, an open-source ESP32 firmware and Linux SDK for building hardware devices that work with Muse. Developers can obtain an API token from gadgets.muse.ai and use a coding agent with the GitHub repo to build peripherals. Guillermo Rauch praised the team's rapid shipping.

  2. Aravind SrinivasXAI score62

    Perplexity open-sources models, an inference engine, and security tools

    AIPerplexity has released several open source projects, including the pplx-decider-v1-27b multimodal decision model, the pplx-embed-v2-context-9b-preview contextual embeddings model, and the Lily local inference engine for Apple silicon. The post also lists the 0.6B on-device PII-Tracer classifier with its PII-TRACE benchmark, the WANDR research agent benchmark, and the Numbat and Bumblebee security tools, and says more open source releases are coming soon.

  3. SGLangOfficialAI score38

    SGLang v0.5.20 adds Intel XPU support and faster RL rollouts

    AISGLang has released v0.5.20, bringing Intel XPU into standard releases alongside RL sampling masks that make rollouts more reliable with up to 52% faster decode. The update also adds Unified Radix Tree SWA branching-point caching, which the project says lifts cache hit rate about 20 points and cuts TTFT by roughly one-third, plus up to 12.5× faster ROCm model loading. New models named in the release include GLM-5.3-Flash, Qwen3.8-Flash-Next, K2 Horizon, Hy4-Preview, FastH3, and VDN-H3.

  4. eric zakariassonXAI score47

    xAI releases experimental TypeScript SDK with Grok models and tools

    AIxAI has released an experimental TypeScript SDK, installable via npm install @xai-official/sdk, that covers text, voice, image, and video in one package. It provides access to the latest Grok models along with server-side tools including real-time X search, web search, code execution, and remote MCP.

    Video from @ericzakariasson's post
  5. DatabricksOfficialAI score44

    Omnigent: open-source meta-harness coordinating Claude Code and Codex agents

    AIDatabricks' new open-source meta-harness, Omnigent, lets multiple coding agents such as Claude Code and Codex share sessions, rules, and security policies in one system. A walkthrough by @leonvz demonstrates forking work across agents, multi-agent review and debate with Debby, and splitting implementation across subagents with Polly.

    Video from @databricks's post
  6. François CholletXAI score28

    Keras community call outlines pluggable backends and KerasHub updates

    AIKeras is moving to a pluggable backend design, with MLX and PaddlePaddle backends upcoming as add-on libraries. The team is reducing the operations needed to ship new backends and streamlining unit testing so a single harness can test all ops, such as casting consistency. KerasHub also gains many new models and is shifting its preprocessing from tf-text to PyGrain.

  7. Ant LingOfficialAI score27

    Ling-3.1-flash now available free on OpenRouter

    AIAnt Ling has made Ling-3.1-flash available on OpenRouter at no cost, inviting users to try it and share feedback. The post provides no details on model size, benchmarks, context length, or pricing beyond the free access.

  8. Hugging Face BlogOfficialAI score70

    Ai2 open-sources AstaBrief 8B, a fast model for generating cited research reports

    AIAi2 released AstaBrief 8B, an open-weights model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. The model runs as Fast mode in Asta, averaging 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The post also describes filtering synthetic training data by citation density and building DPO pairs judged by two models that agreed.

    Why it matters: The post explains how supervised fine-tuning, preference data, and citation-density filtering were used to build a cited-report model, which is useful for teams training their own models.

  9. Liquid AIOfficialAI score64

    Hugging Face guide shows multi-harness RL for coding agents via a capture proxy

    AILiquid AI shared a Hugging Face guide to multi-harness reinforcement learning for coding agents, in which a proxy records the token ids and logprobs vLLM samples so training works without changing the harness. Per the quoted post, LFM2.5-2.6B rose from 42% to 54% after training across four harnesses at once, and imitation fine-tuning on 3,189 rollouts from Qwen3.8-27B plateaued at 47.5%, below both RL runs. The proxy, trainer, tasks, SFT data, training code and seven trained models are described as open.

  10. Google ResearchOfficialAI score60

    Google's TEE-based federated learning system adds verifiable privacy guarantees

    AIGoogle announces a next-generation federated learning system that uses Trusted Execution Environments to provide verifiable, auditable data anonymization. The system publishes access policies to a public transparency log and is deployed in Gboard, which has launched English and Japanese next-word prediction models with stronger privacy guarantees and improved accuracy. Training time has also sped up significantly because computation moved to the server and is parallelized across many machines.

    Why it matters: The post shows how Trusted Execution Environments make federated learning's privacy claims externally verifiable, rather than relying on trust in the server operator.

  11. merveXAI score36

    llama.cpp adds support for decision models on modest hardware

    AIllama.cpp now supports decision models, which route tickets, moderate content, or choose an agent's next step by returning a probability for every option. Five open models from 144M to 27B parameters are supported at launch, and the team says more will follow in the coming days. Because most decision models do not need large GPUs, they are a good fit for llama.cpp, and a Hugging Face blog post explains how to set them up.

  12. Ai2 (Allen Institute for AI)OfficialAI score67

    Ai2 open-sources AstaBrief 8B, a fast open-weights scientific report model

    AIAi2 released AstaBrief 8B, a model that turns a research question and retrieved literature excerpts into a cited report, along with its training data. In Asta's Generate a report feature, Fast mode averages 51.1 seconds per report versus 178.5 seconds for Thinking mode, about 3.5x faster. The model is built on Qwen3-8B with supervised fine-tuning and DPO, and institutions can run its open weights on their own infrastructure.

    Why it matters: The post explains the data filtering and one-pass generation choices behind a fast open-weights report model, showing what worked and what did not.

Oct 1

Oct 1Thu
  1. ComfyUIOfficialAI score34

    YUI, the first animated short film made with Comfy Agent

    AIComfyUI released YUI, described as the first animated short film made with Comfy Agent, a 4:20 film its creator made in three days. Comfy Agent handled shot regeneration, side-by-side model comparisons, and character consistency checks that previously required weeks of manual testing and generation.

  2. Guillermo RauchXAI score38

    Guillermo Rauch says verification engineering is the future of software

    AIGuillermo Rauch argues that the future is verification engineering, spanning proofs, end-to-end tests, benchmarks, and linters. He expects some of these tests to be deterministic and others agentic, and he says the approach looks great. The quoted post introduces e2e, an open-source agentic testing framework that mixes deterministic and agentic APIs and runs locally or in CI.

  3. GoodfireOfficialAI score58

    Goodfire Says AI Biosecurity Risks Are Next After Cybersecurity Risks

    AIGoodfire says AI cybersecurity risks are already here and that biosecurity risks are next, as models improve at biology. The post presents this as both an opportunity for science and medicine and a reason for stronger security. It quotes Demis Hassabis announcing SynthID for biology, a watermarking approach for AI-generated proteins, published in Nature with SynthID Bio tools open sourced.

  4. Josh WoodwardXAI score34

    Google launches Stitch CLI to generate design ideas from terminal

    AIGoogle has introduced the @google/stitch CLI, letting users generate screens and design systems without leaving the terminal. It connects to local coding agents and can send a local dev server snapshot to Stitch. The tool complements the existing Stitch MCP and SDK, and can also be driven through agents such as Antigravity.

  5. merveXAI score46

    Hugging Face clarifies ml-intern options, one trained model for $6

    AIHugging Face says ml-intern is an open-source ML engineering and research harness usable free on local setups, and it is also hosted on Hugging Chat with no-code access. A second hosted option runs on Hugging Face infrastructure, where ml-intern selects the cheapest GPU for a task so models can be trained for a few dollars. MaziyarPanahi reportedly trained a model by prompting alone for $6.60 on an NVIDIA A100 in 16 minutes.

  6. Cloudflare Blog · AIOfficialAI score58

    Cloudflare releases open-source Clef decision models and an RL fine-tuning service

    AICloudflare released Clef and Clef-flash, two decision models hosted on Workers AI and open-sourced on Hugging Face under Apache 2.0, and launched a reinforcement learning fine-tuning service. In Cloudflare's tests, Clef classified a domain in 2.2s versus 4.7s for gpt-oss-120b, and the models are Jev-API compatible. The company is offering fine-tuning first through a forward-deployed engineering team, with a self-serve platform planned later.

Sep 30

Sep 30Wed
  1. Sophia YangXAI score15

    Fireworks AI highlighted as leader in open model inference economy

    AISophia Yang praised Fireworks AI for leading the open model inference economy, citing a plot from Nathan Lambert showing exponential growth in daily tokens processed. The chart compares leading companies using public disclosures for Together, Baseten, and Fireworks, plus OpenRouter API data, with Lambert noting he had underestimated OpenRouter's growth.