Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. vLLMOfficialAI score47

    vLLM-Omni adds day-0 support for Kandinsky 6.0 Video

    AIvLLM-Omni now supports Kandinsky 6.0 Video from launch day, with inference ready at release. Kandinsky 6.0 Video generates 5-second clips with synchronized audio and lip-sync from text or image inputs. Kandinsky's code and checkpoints are released under the MIT license, with Lite (3B) and Pro (29B) variants.

  2. IThome · AINewsAI score41

    Strata engine runs 125B Qwen3.8 model on 12GB GPU at 94 tokens/s

    AIDeveloper Niko1221 has open-sourced Strata, an engine that runs a quantized 125B-parameter Qwen3.8-Flash-Next model on consumer GPUs with at least 12GB of VRAM. Strata loads the MoE model into RAM and keeps only frequently used experts in VRAM, and uses a lightweight model for speculative decoding. On an NVIDIA RTX 5070 with 12GB VRAM, the Q2_0 quantization reaches 94 tokens per second for output.

  3. Latent SpaceBlogAI score60

    Reflection launches Beam, a 501B-parameter open-weight coding model

    AIReflection announced Beam, a text-only 501B-total, 23B-active MoE model for coding, agentic, and scientific work, trained from scratch with full weights under Apache 2.0 promised this month. Self-reported results include 80.9 on SWE-bench Verified and 3–4x the inference efficiency of GLM 5.2, while the roundup notes that GLM 5.3, Kimi K3, Qwen 3.8 Max, and DeepSeek V4.1 Flash are generally ahead.

  4. falOfficialAI score38

    Kandinsky 6.0 launches on fal with Pro and Lite video models

    AIfal has made Kandinsky 6.0 available, offering Pro for cinematic, high-fidelity generation and Lite for fast, low-cost iteration. The release includes native synchronized audio and built-in upscaling, plus standalone VSR and VSR Lite video super-resolution tools.

    Video from @fal's post
  5. EveryBlogAI score36

    Every launches the Every Agent, an agentic coworker in Slack

    AIEvery has launched the Every Agent, an agentic coworker that lives in Slack and helps teams delegate complex work and share AI experiments. It also sends personalized Frontier Alerts when new models or tools ship, and the company says it charges zero percent markup on tokens, so customers pay what Every pays.

  6. Mastra BlogOfficialAI score67

    Mastra launches Agent Controller GA, a runtime for long-running agent sessions

    AIMastra has released Agent Controller in general availability, a runtime that hosts long-running agent sessions around the agent loop. The team says it was first built for Mastra Code and expanded to support Mastra Factory, which runs many concurrent sessions, and that memory usage in long-running Mastra Code processes dropped from 2–20 GB to 300–750 MB after optimizing UI state snapshots.

    Why it matters: The post explains how the controller evolved from one developer's session to many concurrent sessions, with measured memory and storage changes useful to engineers building multi-user agent apps.

Oct 5

Oct 5Mon
  1. Teknium 🪽XAI score23

    Teknium publishes catalog of open hardware for Hermes Agent

    AITeknium has created a catalog of open platform hardware and devices that Hermes Agent, or any agent, can build on, integrate with, or run inside. The post is a brief announcement that links to the catalog at with no further specifications or pricing given.

    Image from @Teknium's post
  2. Matt ShumerXAI score13

    Matt Shumer shares early progress on a live Hogwarts build

    AIMatt Shumer says his Hogwarts project is starting to take shape and invites others to help build it at spawn.co. The post is brief and includes no technical details, specifications, or release information.

    Video from @mattshumer_'s post
  3. meng shaoXAI score47

    Reflection previews Beam, a 501B-parameter open agentic model

    AIReflection AI previewed Beam, an MoE open model with 501B total and 23B active parameters, claiming 3–4x better inference efficiency than GLM 5.2. The model was pretrained from scratch on 23.8T tokens in four weeks, and its RL run used 10,500 GB300 GPUs over four weeks, which the post describes as possibly the largest publicly recorded. Reflection positions Beam as a workhorse open model for enterprises, governments, and developers, with full weights due this month.

    Image from @shao__meng's post
  4. RadixArkOfficialAI score14

    RadixArk joins three SF Tech Week events on open-source AI and RL

    AIRadixArk is speaking at three SF Tech Week events this week on open-source AI, reinforcement learning, and AI infrastructure. Mao Cheng joins an October 7 panel with Novita Labs on the latest Miles post-training release, inference, and agent guardrails. Shi Dong will share research on Miles and recursive self-improvement at an October 8 MiniMax event, and SGLang core contributor Yuwei An will speak at an October 8 AI infra meetup with SkyPilot and H Company.

    Image from @radixark's post
  5. ClineOfficialAI score19

    Cline launches a desktop app alongside its CLI

    AICline says its CLI can be installed globally with npm i -g cline, and it has also released a new Desktop app. The company notes several free model promotions are available to try the Desktop app.

  6. SemiAnalysisBlogAI score52

    Anthropic subscriptions give over 5x the API-equivalent value of OpenAI's

    AISemiAnalysis measured usage meters on Anthropic and OpenAI subscription plans to estimate each plan's API-equivalent value. At mid-tier models, it found Anthropic offers roughly 5x the value of OpenAI, after OpenAI halved its $200 plan limits and introduced a $500 tier. The analysis also argues that subscriptions take a large share of inference compute while providing a small share of revenue, so their limits materially affect lab margins.

  7. ThariqXAI score22

    Thariq shares a Claude Code skill for generating better HTML plans

    AIThariq, who works at Anthropic, is developing a skill for Claude Code that produces HTML plans using simple language, code snippets, surfaced questions, and mockups. Linting is used to reduce common failure cases Claude encounters, and he is seeking feedback before a broader release.

    Video from @trq212's post
  8. PikaOfficialAI score22

    Pika lets users try Ideogram 4.5 on its platform

    AIPika announces that users can try Ideogram 4.5 through its create platform, linking to an Ideogram app page in its image tools section. The post provides no details on features, pricing, or capabilities.

  9. PikaOfficialAI score22

    Pika offers Ideogram 4.5 edits at up to 59% lower cost

    AIPika says users can make repeated edits with Ideogram 4.5 without degrading the original image quality. The post states the editing is up to 59% less expensive on Pika and the Pika API Club.

    Video from @pika_labs's post
  10. ReflectionOfficialAI score42

    Reflection AI previews Beam, a 500B open model under Apache 2.0

    AIReflection AI says its Beam model, with a 500B form factor, combines strong agentic performance and efficient reasoning for enterprises, governments, and developers. Beam is in final red-teaming and will be released this month under an Apache 2.0 license, with quantized FP8 and NVFP4 versions for efficient deployment. Early access sign-ups are open on the company's platform.

  11. ReflectionOfficialAI score14

    Reflection AI says Beam leads in inference efficiency

    AIReflection AI says its model Beam is 3-4x more efficient than GLM 5.2 and more than 4x more efficient than leading Western open models in inference. The company says this means Beam completes tasks faster and cheaper.

    Image from @reflection_ai's post
  12. SunoOfficialAI score12

    Dream Relic releases new album "Lost In a Dream" on Suno

    AIDream Relic has released his new album "Lost In a Dream" on Suno, following the viral success of his earlier tracks "Seven-Eleven Halo" and "Time is a Limited Allowance." The album is available on Suno's platform via a direct album link.

    Video from @suno's post
  13. IdeogramOfficialAI score20

    Ideogram 4.5 turns famous paintings into realistic photos

    AIIdeogram 4.5 can convert famous paintings into realistic photographic scenes using the single prompt "Turn this painting into a realistic scene." The example shown is Edvard Munch's The Scream (1893), reimagined as a photo.

  14. MuseOfficialAI score22

    Muse runs natively on a 15-year-old PSP with voice replies

    AIDeveloper Wob Soriano ported Muse's gadget client to C so it runs natively on a PSP, a handheld kept in his house for over 15 years. Holding R and speaking into the built-in microphone gets an answer from Muse, with OpenAI providing the voice.

  15. Elad GilXAI score40

    Era launches free simulated enterprises for testing AI agents

    AIEra, launched by Ofir Ehrlich's team, generates a complete simulated company spanning Salesforce, Slack, Jira, Zendesk, Gong, and Deel, plus cloud databases and storage. Agents interact with it through live MCP and API interfaces, and because Era generated the company, it knows the exact ground truth for testing and benchmarking. The post says the product is live today and free.

  16. SantiagoXAI score34

    Utah approves Nolla Health's AI app to issue acne prescriptions

    AINolla Health has reportedly become the first U.S. organization to receive regulatory approval for an AI system to issue initial prescriptions, starting with acne treatment in Utah. The app scans a user's face, asks a few questions, creates a personalized plan, prescribes medication when needed, and tracks progress over time. Users also have access to a physician at no extra cost.