Skip to contentSkip to stories
Updated

Open source

Oct 9

TodayOct 9Fri
  1. vLLM BlogOfficialAI score62

    vLLM adds support for NVIDIA Vera Rubin NVL72 with 7.8x throughput over GB200

    AIvLLM now supports NVIDIA Vera Rubin NVL72, with daily container builds and support for models from DeepSeek, Moonshot AI, Z.ai, and MiniMax. In early AgentX benchmarks, vLLM running MiniMax M3 delivered up to 7.84x the throughput per GPU of GB200 NVL72 at matched interactivity. The post is an early look, and the team expects further gains from ongoing optimizations.

    Why it matters: The post gives specific hardware specs, kernel configurations, and benchmark figures for running vLLM on Vera Rubin NVL72, useful for teams planning deployments.

  2. Mike KnoopXAI score62

    Tufa Labs hits 88.06% on ARC-AGI-2, clearing the Kaggle bonus threshold

    AIMike Knoop says the 85% Grand Prize bonus threshold has been reached on Kaggle. The ARC Prize 2026 leaderboard lists Tufa Labs first at 88.06%, followed by Rabbithole at 80.56% and Yi-Chia Chen at 77.22%. Knoop says this will be the final year for ARC-AGI-2 on Kaggle and expects an open-source, low-cost, offline reproducible solution and model.

  3. elvisXAI score40

    Syren Video learns your style to build AI videos from prompts

    AISyren Video, a new agentic video tool, learns preferred graphics, motion, and editing rhythm from a user's library and generates new videos from a prompt. Users refine the results through chat, and the tool is free to try in a browser or through Claude MCP, per the company's announcement. The post's author says the education sector is exploring it.

  4. MarkTechPostNewsAI score44

    Underdog Releases Saluki 27B, a 2-Bit Qwen3.8-27B That Beats the Original at Tool Calling

    AIUnderdog has released Saluki 27B under Apache 2.0, a 2-bit GGUF of Qwen3.8-27B that fits in 7.89 GB, versus 54 GB for the full BF16 model. On Underdog Bench, Saluki scores 88 against 84 for the full model, and it raises parallel tool-call accuracy to 42 from 35. It runs on stock llama.cpp, but math and reasoning drop sharply, with AIME 2025 at 79.2 versus 96.7.

Oct 8

Oct 8Thu
  1. Feng XueXAI score41

    ThreatBook acquires CyberStrikeAI, a widely used AI pentesting agent

    AIThreatBook has acquired CyberStrikeAI, an open-source AI pentesting agent, after reports that attackers had used it in real intrusions. The company says it added guardrails to limit misuse but will put new capability enhancements into its commercial edition, and it argues defenders should control such tools. It also corrects an earlier threat intelligence report that linked the developer, a security engineer at Alipay, to government agencies, calling that attribution a false positive.

  2. meng shaoXAI score77

    Theo open-sources tsc-rs, a Rust port of the TypeScript 7 compiler

    AITheo, creator of the T3 Stack, open-sourced tsc-rs, a line-by-line Rust port of Microsoft's Go-native TypeScript 7 compiler, type checker, and language server under MIT, pinned to typescript-go commit 673a5f17. The author reports tsc-rs is about 1.61× faster than tsc 7 and about 2.95× faster than bun check on six real-app benchmarks on an Apple M4 Pro. The port passes all 181,711 ported Go tests, and CLI output matches the Go version on 120 open-source repos except for known edge cases such as monorepo rootDir and tsc -b incremental output.

    Why it matters: The post reports a benchmarked, test-verified Rust port of the TypeScript 7 compiler, with pinned upstream and stated edge cases useful for judging its compatibility.

    Image from @shao__meng's post
  3. 🚨 AI News | TestingCatalogXAI score62

    Atomic Agent Desktop, an open-source local AI agent app, is now available

    AIAtomic Agent Desktop is a free open-source app for macOS, Windows, and Linux that runs open models like Qwen and Gemma locally without an account. It connects to a cloud model only when selected, and its Fusion feature lets a cloud model plan a task while up to 8 local agents carry it out. The post's own text adds a setup wizard that checks RAM and suggests suitable models, and import from Claude Code, Codex, Hermes, and OpenClaw.

    Video from @testingcatalog's post
  4. elvisXAI score42

    Voyager: an open harness for creative AI work across video and games

    AIElvis Saravia argues that creative work needs domain-specific agent harnesses rather than coding-oriented ones, and he highlights Voyager as an open harness for video, graphics, and games. According to the quoted post, Voyager lets agents work with local files and drive apps such as Blender, DaVinci Resolve, and Unity, and it is designed to work with models like Opus, Astra, and DeepSeek.

    Video from @omarsar0's post
  5. Hacker News · Show HN, AI (20+ points)BlogAI score43

    Pocketty is an iPhone SSH terminal that alerts you when an agent is blocked

    AIPocketty is a $99 iPhone and iPad SSH terminal, with a 14-day free trial, that notifies you when an herdr-managed agent is blocked or done. The alert is sealed on your computer for your phone only, and tapping it opens that exact Pane over SSH so you can answer in a real terminal. The source says the relay forwards only sealed bytes and that terminal traffic goes directly between the app and your computers.

  6. Alexander DoriaXAI score46

    LightOnOCR-3 claims state-of-the-art OCR performance under 1B parameters

    AILightOn has released LightOnOCR-3, a family of OCR models in 0.8B and 4B versions that it says lead benchmarks including OlmOCR-Bench and ParseBench, with the 0.8B model positioned as the sub-1B option. The models recognize text, handwriting, images, charts and document structure in one pass, process documents up to twice as fast as LightOnOCR-2, and are released under the Apache 2.0 license.

    Image from @Dorialexander's post
  7. Dhravya ShahXAI score42

    MemoryRepo: open-source implementation of Cognition's dreaming agent memory

    AISupermemory introduces MemoryRepo.dev, an open-source implementation of Cognition's dreaming memory system built on Cloudflare Artifacts, Durable Objects, Alchemy, and Effect. The project follows Cognition's Devin memory design, which builds a memory graph across sessions and prunes stale records overnight. Supermemory says it will incorporate learnings from this research into its own product.

    Video from @DhravyaShah's post
  8. Zhihao JiaXAI score62

    Lithos AI open-sources lithos-metal for fast local inference on Apple M5 Max

    AILithos AI says it is open-sourcing lithos-metal, which uses megakernels and DSpark speculative decoding. The post claims Qwen3.8-27B reaches a peak of over 200 tokens per second per user on a single Apple M5 Max. It says users can try the tool with any coding agent in one command, and links to the code on GitHub and a technical blog.

    Video from @JiaZhihao's post

Oct 7

Oct 7Wed
  1. Gizmodo · AINewsAI score51

    Vibe-Coded Artcraft Suite Offers Free Photoshop Alternative on GitHub

    AIDeveloper Brandon Thomas used Claude Opus 5.5 and Rust to build Artcraft, a free open-source suite with Photocraft, Vectorcraft, and other apps that mimic Adobe products. The author tested Photocraft and found its basic editing commands worked where expected, but Free Transform behaved unpredictably. Thomas describes the software as early alpha and invites developers to contribute.

  2. a16z NewsBlogAI score46

    a16z backs Preference Model, which builds RL environments for training AI models

    AIPreference Model is open-sourcing Karotte, the framework it uses to build reinforcement learning environments that resist reward hacking, including defenses like killing stray processes before grading and rejecting grader-crashing files. The framework has been hardened through more than a million evaluation runs and controlled red-teaming. The company focuses on machine learning engineering tasks for leading labs, and a16z says it is partnering with Preference Model and its founders, Jennifer Zhou and Ning Cao.

  3. LlamaIndex 🦙OfficialAI score47

    LlamaIndex launches OpenDocRouter, one API for many document parsing models

    AILlamaIndex announced OpenDocRouter, a single API that routes document parsing requests to any of 10 frontier and open-source models at launch, including Claude Opus 5.5, Gemini 3.8 Flash, GPT-6 Luna, MinerU2.5-Pro, and PaddleOCR-VL-1.6. Users can switch models in one line with the same request and markdown output, and each model is scored on ParseBench for quality and cost. Pricing is per-token, failed pages are not charged, and the service costs $0.86 to $48.82 per 1,000 pages depending on the model.

    Video from @llama_index's post
  4. elvisXAI score44

    Pulse by Smallest AI tops Diarization Bench with 24.4% DER

    AISmallest AI's Pulse ranks first on the Diarization + ASR track of Voice Arena's Diarization Bench with a 24.4% DER, where lower is better. The next system in that track scores 40.7% DER, and Pulse misses 4.4% of reference speech, the lowest in the track.

  5. Latent SpaceBlogAI score61

    Stacklok's Mecatl harness moves coding agents from desktops to the cloud

    AIStacklok, founded by Kubernetes creators Craig McLuckie and Joe Beda, has released Mecatl, an open source cloud-native harness for coding agents on GitHub. Mecatl keeps the agent loop separate from the client, model provider, state store, and execution environment, and moves tool calling, session management, and memory into manageable systems. The article also covers ToolHive, an MCP platform, and an AI Gateway that is not yet open sourced, with a commercial enterprise control plane tying the pieces together.

Oct 6

Oct 6Tue
  1. Epoch AIOfficialAI score47

    GPT-6 Astra Hit 100% on EBR-bench Using a Card That Bypassed Its Time Limits

    AIEpoch AI reports that GPT-6 Astra scored 100% on the original EBR-bench by exploiting a card that bypasses the game's time-constraint expectations, so Epoch has banned that card from the default setting. Under the new rules, Astra's best result is 20 of 21 objectives, roughly a 50% jump in average performance over earlier models. Epoch will report revised scores only for Claude Fable 5.1, Claude Opus 5, GPT-5.6 Sol, GPT-6 Astra, and future models.