Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 23

Sep 23Wed
  1. QwenAI score60

    Qwen Intelligence launches three mobile agents and opens its benchmark suite

    AIAlibaba's Qwen launched Qwen Intelligence with three mobile agents: a Mobile Planner Agent, a Mobile-Use Agent, and a Mobile Creative Agent. The post reports benchmark results including MobileWorld 82.1, MobileWorld-Real 92.2, and AndroidDaily 97.2, plus a 90% end-to-end success rate, and says the MobilePA-Bench, MobileWorld, MobileWorld-Real, and MobileWorld-Safety benchmarks are open.

    Image from @Alibaba_Qwen's post
  2. Tencent HyAI score22

    Tencent Hunyuan's Hy Image3.5 preview free for two weeks on OnSolo

    AITencent Hunyuan has made its Hy Image3.5 preview available on the OnSolo platform, free for two weeks. The model targets short drama character sheets, full-motion video game assets, and keyframes, with characters kept consistent across episodes and edits that refine rather than regenerate images. OnSolo's background post says the preview supports 5 references at 2K resolution and is free for Members during the two-week window.

  3. Prime Intellect BlogAI score60

    Prime Intellect makes Prime Sandboxes generally available as microVMs for agentic RL

    AIPrime Intellect has made Prime Sandboxes generally available, offering each sandbox as a full Linux virtual machine with its own kernel and support for Docker Compose. The product is available through its CLI/SDK and RL suite, with accounts starting at 1,024 concurrent sandboxes, and pricing listed at $0.02 per vCPU-hour, $0.0125 per GiB-hour of memory, and $0.0002 per GiB-hour of disk, valid through December 22. The company says GPU microVMs, snapshotting, sandbox forking, and persistent workspaces are planned next.

    Why it matters: The post explains why full VMs rather than gVisor containers matter for agentic RL, since silent environment differences can reward behaviors that fail to transfer.

Sep 22

Sep 22Tue
  1. Tencent HyAI score43

    Tencent Hunyuan previews Hy Image3.5 in ComfyUI

    AITencent Hunyuan's Hy Image3.5 preview is now available in ComfyUI, with a claimed 30% higher win rate in human evaluation than Hy Image3.0. The model handles text-to-image and image-to-image in one model at up to 2K resolution, with multilingual text and small print rendering correctly. It also keeps identity and product features consistent across scene, outfit, and style changes.

  2. Google Developers BlogAI score62

    Antigravity SDK adds local Gemma 4 26B agent support via LiteRT

    AIGoogle announced that the Antigravity SDK supports local agent workflows, with initial support for Gemma 4 26B A4B through Google AI Edge's LiteRT. The post includes Python setup steps and says a recommended machine has more than 24GB VRAM or unified memory. It also describes a hybrid pattern in which a cloud Gemini 3.8 Flash planner hands work to local Gemma 4 26B models, with 97.2% of tokens in one recorded run staying local.

    Why it matters: The source shows how to run an agent with a local Gemma 4 26B model using LiteRT, plus a hybrid cloud-planner pattern that keeps most tokens on-device.

  3. Fireworks AI BlogAI score46

    Fireworks ARCv3 cuts RL weight-update payloads nearly 50% for cross-region training

    AIFireworks released ARCv3, a lossless compressor for BF16 weight-update deltas sent from trainers to RL rollout machines. Across 1,000 production RL deltas, ARCv3 produced payloads nearly 50% smaller than ARCv2, averaging about 0.19% of the BF16 weight size versus 0.36%. ARCv3 is available through the Fireworks Training API as fireworks-delta-compression.

  4. Sierra BlogAI score34

    Sierra Lets Companies See, Edit, and Export Their AI Agents' Logic and Data

    AISierra says its platform makes enterprise AI agents visible and editable, with journeys, policies, and actions viewable in Agent Studio and testable through Simulations and Experiments before rollout. Customers can export agent logic in a portable structured format, access conversation logs and performance data through export APIs, and manage the agent's code in a Git repository. Sierra agents also connect to existing systems through MCP, REST, GraphQL, or custom integrations.

  5. Comfy BlogAI score42

    ComfyUI Speeds Up MiniMax H3 Video VAE Encoding and Decoding

    AIComfyUI's update makes the MiniMax H3 video VAE encode up to about 2.2x faster and decode 1.4-2.7x faster, cutting a 1344x768, 129-frame round trip on an RTX 5090 from 24.3 to 12.7 seconds. The gains come from a fused encoder kernel enabled by default, fp16 accumulation support in a custom convolution, and an int8 decoder, and the source says the changes are visually lossless to the eye. Users need ComfyUI v0.36.0 or above, and the int8 VAE file is a drop-in replacement for the standard one.

  6. Alex AlbertAI score60

    Claude Cowork and chat are merging into a single Claude

    AIAnthropic's Claude Cowork and chat are being merged into one Claude, which can take a question or report and keep working after the user closes their laptop. When something is unclear, Claude asks, and the user keeps the final say. The update rolls out to Pro and Max over the next few weeks, and the author notes it is coming to everyone soon.

  7. StepFunAI score43

    StepFun open-sources onPanda for token-level LLM annotation and inspection

    AIStepFun has open-sourced onPanda, a tool used internally for LLM data annotation and model inspection, letting users correct tokens and let models continue. The company reports a 52% lower median annotation time versus manual post-editing, with SFT and preference data combined in one workflow. It also supports token probability and top-k inspection, token-by-token decoding control, and browser-based testing across SVG generation, web development, and agent tasks.

  8. StepFunAI score27

    StepFun's Step Code tops Terminal-Bench 2.1 and Multi-Frame with fewer tokens

    AIStepFun's Step Code passed 72 of 89 tasks (80.9%) on Terminal-Bench 2.1, tying for the highest pass rate among evaluated harnesses while using fewer tokens than the other tied leaders. On Multi-Frame, it passed 110 of 150 tasks (73.3%) and averaged 5.09M tokens per task, the highest pass rate and lowest token use among six harnesses evaluated.

    Image from @StepFun_ai's post
  9. StepFunAI score52

    StepFun releases Step Code v0.1.0 as an open-source coding CLI

    AIStepFun has released Step Code v0.1.0, an open-source command-line tool under the MIT License that covers reading and editing code, running tests, and shipping from one CLI. The post reports 80.9% on Terminal-Bench 2.1 and 73.3% on Multi-Frame, a 150-task long-horizon benchmark from StepFun. It also includes one-command static site publishing with StepPage and links the GitHub repository.

    Image from @StepFun_ai's post
  10. Daniel HanAI score42

    Qwen-Image-2.1 runs locally in Unsloth Desktop via INT8, FP8, GGUF

    AIDaniel Han says Qwen-Image-2.1 works in Unsloth Desktop through INT8, FP8, and GGUF builds, with Unsloth also releasing dynamic GGUFs for it. Pinned RAM offloading lets INT8 and FP8 fit under 6–8GB of VRAM while remaining relatively fast. The linked Unsloth post says the 7B model runs on 12GB VRAM and performs on par with Nano Banana 2.0.

  11. Tencent HyAI score34

    Tencent Hunyuan's Hy Image3.5 preview launches free on Miora for two weeks

    AITencent Hunyuan has released a preview of Hy Image3.5 on Miora, a design platform, and is offering it free for two weeks. The model keeps existing canvas workflows and remembers users' brand rules while making edits. Miora's background post says the free period runs through October 7 and supports up to 2K output for text-to-image and image-to-image.

  12. OpenBMBAI score59

    VoxWeft runs real-time interpretation locally on Apple Silicon using VoxCPM2

    AIOpenBMB highlights VoxWeft, an open-source simultaneous interpretation system for Apple Silicon built by developer @HenryZ30734018 on an MLX implementation of VoxCPM2. The system turns live speech into translated speech on-device, with first audio streaming in about 170 ms on an M5 MacBook. VoxCPM2 generates speech in 30 languages, supports direct language-pair interpretation, and clones a target voice from about 5 seconds of reference audio.

    Video from @OpenBMB's post
  13. TechNode · AIAI score60

    Alibaba's T-Head unveils Zhenwu V900 AI chip with full-stack system design

    AIT-Head, Alibaba's chip subsidiary, unveiled the Zhenwu V900 AI chip for training and inference at the 2026 Apsara Conference in Hangzhou. The company claims three times the performance of its predecessor, the Zhenwu M890, with 216GB of memory, 1,200GB/s inter-chip bandwidth, and mass production expected in the first quarter of 2027.