Skip to contentSkip to stories

Updated

Coding

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 9

TodayOct 9Fri
  1. SGLangOfficialAI score37

    Miles runs full RL loop on Rubin GPUs with SGLang and Megatron

    AIMiles runs the full reinforcement learning loop on Nvidia Rubin, using SGLang for rollout, Megatron for training, and one container image. On a single 4-GPU tray, Qwen3-30B-A3B's GSM8K reward rises from about 45% to about 95% over 50 rollouts, matching the GB300 curve. The post also reports DeepSeek-V4-Flash end-to-end rollout and training, and Qwen3.5-35B-A3B agentic RL with 64 concurrent mini-SWE-agent sandboxes on SWE-bench Verified, where reward holds near 0.6 and median response length falls about 30%.

  2. RadixArkOfficialAI score22

    RadixArk praises Proximal for training coding agents with Miles

    AIRadixArk says Proximal is using Miles to train coding agents and calls it a flexible, scalable foundation for teams running their own training workloads. Proximal says its training framework is built on Miles, with runs on Modal's on-demand GPU clusters and serverless GPUs for inference. Its sandboxing infrastructure runs on Kubernetes and can handle millions of concurrent rollouts.

Oct 8

Oct 8Thu
  1. Anil Chandra Naidu MatchaXAI score29

    Open-Voyager releases open-source creative-work agent harness

    AIOpen-Voyager is announced as a free, open-source harness for creative work, described as a Codex or Claude Code equivalent. The post says it integrates with 600+ models and links to a GitHub repository for self-hosting. The background post says the original Voyager is built for video, graphics, and games, and can drive tools such as Blender, Resolve, and After Effects.

    Video from @matchaman11's post
  2. meng shaoXAI score77

    Theo open-sources tsc-rs, a Rust port of the TypeScript 7 compiler

    AITheo, creator of the T3 Stack, open-sourced tsc-rs, a line-by-line Rust port of Microsoft's Go-native TypeScript 7 compiler, type checker, and language server under MIT, pinned to typescript-go commit 673a5f17. The author reports tsc-rs is about 1.61× faster than tsc 7 and about 2.95× faster than bun check on six real-app benchmarks on an Apple M4 Pro. The port passes all 181,711 ported Go tests, and CLI output matches the Go version on 120 open-source repos except for known edge cases such as monorepo rootDir and tsc -b incremental output.

    Why it matters: The post reports a benchmarked, test-verified Rust port of the TypeScript 7 compiler, with pinned upstream and stated edge cases useful for judging its compatibility.

    Image from @shao__meng's post
  3. AnshuXAI score38

    Pocket Aces: open-source Pokémon-Balatro roguelike game with swappable IP

    AIDeveloper Anshu Chatterjee has open-sourced Pocket Aces, a Balatro-style poker roguelike that replaces all Pokémon IP with original characters. The Pokémon assets are modular and can be swapped back in from the PokeAPI repo, though the developer notes the user proceeds at their own risk. The post says the game now works on mobile with reduced-motion options, and the developer is willing to address further bug reports.

    Video from @anshuc's post
  4. Zhihao JiaXAI score62

    Lithos AI open-sources lithos-metal for fast local inference on Apple M5 Max

    AILithos AI says it is open-sourcing lithos-metal, which uses megakernels and DSpark speculative decoding. The post claims Qwen3.8-27B reaches a peak of over 200 tokens per second per user on a single Apple M5 Max. It says users can try the tool with any coding agent in one command, and links to the code on GitHub and a technical blog.

    Video from @JiaZhihao's post
  5. Daniel HanXAI score38

    Unsloth adds OS-level sandboxing for Linux, Mac, and Windows

    AIUnsloth now supports OS-level sandboxing on Linux via bwrap, on Mac via seatbelt, and on Windows via Microsoft's MXC. Per-tool-call latency is under 100ms across all three, and its software-style sandboxing with regex AST checks adds about 3ms. The Windows integration was built in collaboration with Microsoft.

  6. Unsloth AIOfficialAI score44

    Unsloth adds Windows OS-level sandboxing via Microsoft's mxc

    AIUnsloth now supports OS-level sandboxing on Windows by integrating Microsoft's open-source mxc repository for sandboxed code execution. The integration adds under 100 ms of overhead, according to the post. A setup guide is available in Unsloth's documentation.

    Image from @UnslothAI's post
  7. Augment Code BlogOfficialAI score62

    Augment Code sells Cosmos, Auggie CLI, and Context Engine assets to Harness

    AIAugment Code is selling select assets, including Cosmos, Auggie CLI, and the Code Context Engine, to Harness, and the product team is moving to Harness. The company says Harness's integrated platform delivers these capabilities to customers more effectively than building them independently. Harness describes itself as building the Autonomous SDLC Platform for shipping AI-written code across enterprises.

    Why it matters: The announcement shows how a coding AI company is folding its products into a larger software delivery platform, a shift that shapes how enterprise teams will buy these tools.

  8. Tessl BlogOfficialAI score44

    Tessl Code Review Uses Repo-Owned Lenses to Make AI Review Context-Driven

    AITessl's Code Review defines review standards as skills in the repository, called lenses, routed to files by a repo-owned profile file. Because these team-visible standards are portable, lessons from review can feed back into code generation and maintenance, not only the next review.

  9. vLLMOfficialAI score62

    vLLM v0.31.0 adds DeepSeek-V4.1-Flash support and new serving features

    AIvLLM v0.31.0 is released with 717 commits from 307 contributors, including 96 first-time contributors. Highlights include DeepSeek-V4.1-Flash support, a vllm preload command that keeps weights in GPU memory across restarts, and Model Runner V2 with draft-model speculative decoding. The release also adds large-scale serving, scheduling, and HiSparse fixes, with full notes linked on GitHub.

    Image from @vllm_project's post

Oct 7

Oct 7Wed
  1. Gizmodo · AINewsAI score51

    Vibe-Coded Artcraft Suite Offers Free Photoshop Alternative on GitHub

    AIDeveloper Brandon Thomas used Claude Opus 5.5 and Rust to build Artcraft, a free open-source suite with Photocraft, Vectorcraft, and other apps that mimic Adobe products. The author tested Photocraft and found its basic editing commands worked where expected, but Free Transform behaved unpredictably. Thomas describes the software as early alpha and invites developers to contribute.

  2. Ars Technica · AINewsAI score46

    Artcraft releases open source clones of Adobe Photoshop, Premiere and other apps built with Claude

    AIDeveloper Brandon Thomas's Artcraft has launched seven open source apps in Rust that aim to replicate the interfaces and tools of Adobe Photoshop, Illustrator, Premiere, Lightroom, After Effects, InDesign, and Acrobat Pro. Thomas said he used Anthropic's Claude Opus 5.5 to build the clean-room replacements, with WebAssembly versions available for browser use. The apps remain in a "super early alpha" state, and commenters have pointed out many current shortcomings.

  3. Vercel DevelopersOfficialAI score23

    Glyph Cluster is available on Vercel AI Gateway in stealth

    AIVercel says Glyph Cluster, listed as stealth/glyph-cluster, is now free on AI Gateway for a limited time while in stealth. Access is restricted to paid users, the model is positioned for coding and agentic knowledge work, and prompts may be used for model improvement.

Oct 6

Oct 6Tue
  1. Kilo (acq. by Anaconda)OfficialAI score29

    Kilo launches Kilo Desktop, a unified app for 500+ AI models

    AIKilo has launched Kilo Desktop, a single app offering access to more than 500 models from major labs, including open-source and local models. It includes agents that plan, code, and debug alongside users, plus built-in notebooks, local model support, and conda environments.

    Image from @kilocode's post
  2. Latent SpaceBlogAI score60

    Reflection launches Beam, a 501B-parameter open-weight coding model

    AIReflection announced Beam, a text-only 501B-total, 23B-active MoE model for coding, agentic, and scientific work, trained from scratch with full weights under Apache 2.0 promised this month. Self-reported results include 80.9 on SWE-bench Verified and 3–4x the inference efficiency of GLM 5.2, while the roundup notes that GLM 5.3, Kimi K3, Qwen 3.8 Max, and DeepSeek V4.1 Flash are generally ahead.

  3. Mastra BlogOfficialAI score67

    Mastra launches Agent Controller GA, a runtime for long-running agent sessions

    AIMastra has released Agent Controller in general availability, a runtime that hosts long-running agent sessions around the agent loop. The team says it was first built for Mastra Code and expanded to support Mastra Factory, which runs many concurrent sessions, and that memory usage in long-running Mastra Code processes dropped from 2–20 GB to 300–750 MB after optimizing UI state snapshots.

    Why it matters: The post explains how the controller evolved from one developer's session to many concurrent sessions, with measured memory and storage changes useful to engineers building multi-user agent apps.

Oct 5

Oct 5Mon
  1. ReflectionOfficialAI score23

    Reflection AI's Beam model pretrained in four weeks on 24T tokens

    AIReflection AI says its Beam model was pretrained in 4 weeks on 24T high-quality tokens, giving it innate coding capabilities. The company credits MoE stability improvements and large-scale data curation and deduplication for a base model it claims outperforms open-source base models of the same class. It presents this strong reasoning foundation as what makes sustained reinforcement learning gains possible.

    Image from @reflection_ai's post
  2. Guillermo RauchXAI score44

    gdp-ts brings compile-time authorization proofs to TypeScript APIs

    AIGuillermo Rauch introduced gdp-ts, a library, linter, and AI skill that uses "proofs" to enforce that sensitive functions are called only after an authorization check. The TypeScript typechecker verifies these proofs at compile time, aiming to stop security bugs from shipping, including those written by AI agents. The README models a Vercel API constraint requiring a role and entitlement proof to change a Project's password.

    Video from @rauchg's post

Oct 2

Oct 2Fri
  1. SGLangOfficialAI score58

    SGLang v0.5.21 adds native decisions API and new model support

    AISGLang has released v0.5.21 with a native Decisions API that turns an LLM or VLM into a low-latency classifier and scorer. The release also lets /v1/score rerank search or RAG results in one call, lets PD instances switch between prefill and decode without restarting, and adds support for models including DeepSeek-V4.1 Flash, Kimi K3, and GLM-5.3-Flash on AMD MI355X. The announcement reports a 22% faster first token on long prompts for DeepSeek-V4.1 Flash and 20.6% higher prefill throughput for Kimi K3 in PD serving.

    Image from @sgl_project's post

Sep 30

Sep 30Wed
  1. ClineOfficialAI score34

    Cline desktop app can run agents on remote Linux servers over SSH

    AICline says its desktop app can keep running on a laptop while the agent executes on any Linux machine reachable via SSH, configured under Settings → Remote. The setup requires no root access, no npm, and no public port, uploading a self-contained helper and tunneling only the authenticated Cline protocol.

Sep 29

Sep 29Tue
  1. Kilo (acq. by Anaconda)OfficialAI score32

    Kilo adds Sign in with ChatGPT across all its surfaces

    AIKilo now supports Sign in with ChatGPT, letting users apply their ChatGPT plan across VS Code, JetBrains, the CLI, Cloud Agents, Code Reviews, and the mobile app. The integration enables running 10 agents in parallel, and Cloud Agents continue working after the user closes their laptop.

    Image from @kilocode's post
  2. ModelScopeOfficialAI score54

    IQuest-Q1 released as 320B MoE model for long-horizon coding agents

    AIModelScope announced IQuest-Q1, a 320B MoE model with 15B active parameters and a 512K context window for agentic coding. The post reports scores of 84.5 on CyberGym, 83.2 on Terminal-Bench 2.1, 64.6 on DeepSWE v1.1, and 63.0 on NL2Repo, and says weights are released under the IQuest-Q1 License.

    Image from @ModelScope2022's post
  3. vLLMOfficialAI score58

    IQuest-Q1 320B MoE coding model gets day-0 support in vLLM

    AIvLLM announced day-0 support for IQuest-Q1, a 320B-parameter MoE model with 15B active per token, 256 experts with 8 active, and a 524,288-token context. The post credits existing vLLM features such as the hybrid KV cache coordinator, sinks attention path, and EAGLE speculative decoding with probabilistic draft sampling. The linked material includes a Docker image and vllm serve commands, with and without recursive MTP.

    Image from @vllm_project's post
  4. SGLangOfficialAI score53

    SGLang adds Day-0 support for IQuest-Q1 with a single-node serve command

    AISGLang says it has Day-0 support for IQuest-Q1, an open-source sparse MoE model with 320B total and 15B active parameters for coding and agentic tasks. The post includes a single-node serving command for H200 GPUs in BF16, using tensor parallelism of 8, EAGLE speculative decoding, and the iquest_q1 reasoning and tool-call parsers. The image marks the command as not verified.

    Image from @sgl_project's post

Sep 28

Sep 28Mon
  1. FireworksOfficialAI score22

    Normal Factory's CAD Arena joins the Specialized Intelligence Index

    AINormal Factory joins the Specialized Intelligence Index with CAD Arena, which tests whether AI agents can turn engineering drawings into accurate, editable CAD parts. The benchmark evaluates agents across five CAD platforms, extending the SII into engineering design.

    Image from @FireworksAI_HQ's post

Sep 24

Sep 24Thu

Sep 17

Sep 17Thu
  1. Hacker News · Launch HN, YC launches (10+ points)BlogAI score58

    Skillsync launches tool to move AI chat sessions across coding agents

    AISkillsync, a Y Combinator W26 company, launched a tool that converts AI chat sessions between coding agents, including messages, reasoning and tool calls. The conversion engine txcript is open source, and the local-first app runs on Mac with a CLI and MCP support. Only sessions shared into team workspaces leave the user's machine.

Sep 16

Sep 16Wed
  1. TinkerOfficialAI score32

    Sundial trains Inkling-Small to fix LaTeX errors in under a second

    AISundial fine-tuned Thinking Machines' Inkling-Small with RLVR on 3,978 verified TeX.StackExchange fixes, using rewards for compilation and PDF match and penalties for removed content. The trained model fixes 83.7% of LaTeX errors in under one second at $0.0013 per fix, according to the post. Sundial says it is rolling out the model in its editor, applying fixes as suggestions and rebuilding the PDF.

  2. Kilo (acq. by Anaconda)OfficialAI score40

    Kilo Mobile lets users run full AI agent loops from their phone

    AIKilo Mobile now lets users spawn Cloud Agents, start sessions on remote machines, and dictate prompts by voice from a phone. Users can also review and comment on pull requests and approve Security Agent remediations without a laptop. On iPhone, Live Activities show session status on the Lock Screen when an agent needs input.

    Image from @kilocode's post

Sep 15

Sep 15Tue
  1. Zed BlogOfficialAI score72

    Zed launches Delta public beta to replace pull requests with agent threads

    AIZed has launched the public beta of Delta, a multiplayer environment for coding with agents and reviewing their work, which replaces pull requests with shared threads. Delta is built on DeltaDB, which records edits and messages between Git commits, and it is free during the beta, with paid plans for individuals and teams to follow.

    Why it matters: The post explains how Delta replaces pull requests with shared agent threads and DeltaDB, showing a concrete alternative to the GitHub review workflow.