Skip to contentSkip to stories

Updated

#Video

Showing low-relevance items too. Hide low-relevance items

Sep 18

Sep 18Fri
  1. WanOfficialAI score16

    Wan3.0 ranks #2 in Overall Video on OpenArt Arena

    AIWan3.0 placed second in the Overall Video category on OpenArt Arena, according to its developers. The post highlights support for 30-second takes, native audio, and any reference input for creators.

    Image from @Alibaba_Wan's post

Sep 17

Sep 17Thu
  1. vLLM BlogOfficialAI score38

    vLLM Adds NVIDIA Hardware Video Decoding to Scale Multi-GPU Video Captioning

    AIvLLM now supports NVIDIA hardware video decoding through PyNvVideoCodec, moving video decoding off the CPU so multi-GPU video captioning can scale to 8 GPUs. In benchmarks on 8xH100 GPUs, GPU-based decoding provides more than double the throughput of the CPU-based decoder for Qwen/Qwen3-VL-8B-Instruct with 8 single-GPU vLLM replicas. The functionality is included in standard CUDA vLLM releases, and PyNvVideoCodec==2.0.4 is required for custom installations.

  2. World LabsOfficialAI score32

    World Labs' Atlas generates real-time flythrough from 32 input images

    AIWorld Labs says its Atlas model turns 32 input images into a real-time flight through NVIDIA's Voyager headquarters. Trained on NVIDIA Blackwell GPUs, Atlas uses the images as 3D spatial context to generate new views with pixel-perfect camera control.

    Video from @theworldlabs's post
  3. WanOfficialAI score44

    Wan3.0 generates single 30-second video shots with director-level control

    AIAlibaba's Wan3.0 video model now produces a single 30-second shot directly, up from 15 seconds and a year ago's 5-second clips. It adds director-level control and omni-reference input accepting up to five videos, letting creators generate long takes instead of stitching short clips. A filmmaker used Wan3.0 in a production workflow to make Soulscape and Johnny Mai.

    Video from @Alibaba_Wan's post

Sep 16

Sep 16Wed
  1. Kling AIOfficialAI score3

    Kling AI shares a link to a full film

    AIKling AI's post links to a full film hosted at watch.fountain0.com. The post itself gives no further details about the film's content, production, or release.

  2. Kling AIOfficialAI score38

    Fountain 0's ODYSSEUS: The Fall, fully generated with Kling 3.0, is out

    AIFountain 0's new feature film ODYSSEUS: The Fall, directed by Ash Koosha, is now available in full with every shot generated by Kling 3.0. The team previously premiered Dreams of Violets at the 2026 Tribeca Festival as the first AI feature film accepted into a major film festival.

    Video from @Kling_ai's post
  3. WanOfficialAI score34

    Wan 3.0 generates a full anime-style fantasy fight sequence

    AIQwen's Wan 3.0 video model produced a complete fantasy fight between a mage and a swordswoman, with strong timing, poses, and complex layouts that the post calls anime-ready. The quoted post notes Wan is stronger than Seedance 2.5 on timing and layouts, while Seedance is better at keeping characters on-model, and that all three support Vid2Vid and Omni Reference.

  4. WanOfficialAI score7

    Wan 3.0 used to generate a magical girl video clip

    AIWan (@Alibaba_Wan) shares a magical girl clip made with Wan 3.0. The post notes it is also a test of effect generation, so more magical girl content may follow. The clip was shared alongside the TapNow tool.

  5. inclusionAI (Ant Ling) · new models on Hugging FaceOfficialAI score55

    inclusionAI releases Realtime-Venus full-duplex audio-visual models on Hugging Face

    AIinclusionAI has published Realtime-Venus on Hugging Face with two 9B checkpoints: Realtime-Venus-Omni for audio-visual interaction and Realtime-Venus-Audio for audio-only conversation. Both are built on MiniCPM-o 4.5 with a Qwen3-8B backbone and support full-duplex dialogue, proactive responses, and training-free long-video memory. The asynchronous Realtime-Venus-Harness runtime is hosted in a separate GitHub repository.

Sep 15

Sep 15Tue
  1. Jazzyear · InsightsNewsAI score67

    HiDream's vivago R1 agent targets five-minute AI video delivery

    AIHiDream.ai launched vivago R1, a content creation agent, globally, with a domestic version upgrade. The company says R1 can output five-minute high-quality videos through agent planning, with a claimed 85% usable-output rate and support for multi-round extensions. It also released HiDream-O1-Video-1.0, a native omni-modal video model supporting single shots of 5 to 20 seconds at 1080p.

  2. OdysseyOfficialAI score38

    Odyssey unveils Odyssey-3, a foundation world model for robotics and more

    AIOdyssey announced Odyssey-3, a foundation world model it describes as a major step forward. The post claims it can control robots, power humanoids, drive cars, train AIs, pilot drones, and play video games, though it gives no benchmarks or technical specifications.

    Video from @odysseyml's post
  3. Kling AIOfficialAI score15

    Kling AI joins TIFF: The Market 2026 with industry panel and showcase

    AIKling AI took part in TIFF: The Market 2026 on September 10 through an industry panel, networking reception, and on-site showcase. The company said the event explored how AI video is becoming part of professional filmmaking, from feature-length animation to real-world production workflows. Kling AI also said it continues to join leading film festivals such as Cannes and Toronto.

    Image from @Kling_ai's post
  4. MiniMax Design (H3)OfficialAI score26

    MiniMax Design canvas runs Astra agent and Blender to produce full scene

    AIMiniMax Design lets a single brief drive a full production workflow on one canvas, with the Astra agent working in Blender through an official connector to build the scene and camera direction. The final video is generated with MiniMax H3 from the same canvas, with outputs syncing directly onto the canvas.

  5. MiniMax Design (H3)OfficialAI score16

    Hailuo AI's MiniMax H3 brings GPT-6 Astra-generated set to life

    AIHailuo AI, owned by MiniMax, promotes its MiniMax H3 model as a way to create an army of characters without casting or catering costs. A creator, @koldo2k, says they used GPT-6 Astra-generated work to build a set and story, then animated it with MiniMax H3.

Sep 14

Sep 14Mon
  1. MiniMax (official)OfficialAI score41

    MiniMax H3 video generation exceeds 2× real-time on 8× B200

    AIMiniMax H3 with SGLang-Diffusion and VDN-H3 generates 14.4 seconds of 768p video in 9.0 seconds end-to-end after warmup on 8× B200 GPUs. Eight-step denoising takes 6.9 seconds, exceeding 2× real-time, with no measured quality regression versus dense 50-step H3 across 103 test prompts.

  2. RadixArkOfficialAI score43

    SGLang-Diffusion runs MiniMax H3 video generation faster than playback

    AIRadixArk's SGLang-Diffusion, paired with VDN-H3, generates 14.4 seconds of 768p video in 9.0 seconds on 8× B200 GPUs. The 8-step denoising alone takes 6.9 seconds, which is over 2× real time, and the team reports no measured quality regression against dense 50-step MiniMax H3 across 103 test prompts.

  3. MiniMax Design (H3)OfficialAI score9

    MiniMax Hailuo AI Bestiary contest closes tonight with $8,000 prize

    AIThe Hailuo AI Bestiary contest closes tonight, September 14 at 23:59 PT, with $8,000 cash, 200,000 Credits, and ten Audience Choice awards. Entrants must submit a 30-second-or-longer short film generated with MiniMax H3 on any myth, era, or world.

  4. MiniMax (official)OfficialAI score36

    MiniMax H3 community projects speed up open-source video generation

    AIMiniMax highlighted open-source community progress on its H3 video generation model, which it built with native stereo audio and multimodal reference control. Recent highlights include FastH3's 4-step distillation running on DGX Spark and Apple Silicon, and NVIDIA's Sol-H3 generating 15 seconds of 768p video with audio in 6.6 seconds on 8×B300 in a warm-inference benchmark. Other releases include VDN's faster-inference attention work with code and weights, and 8-step Acc-LoRAs from Alibaba PAI, with LightX2V offering 4- and 8-step Turbo LoRAs.

    Image from @MiniMax_AI's post
  5. MiniMax Design (H3)OfficialAI score22

    Hailuo AI highlights AI-driven 3D workflow integration with Blender

    AIHailuo AI promotes bringing AI tools into professional 3D production workflows. A quoted post from @akiyoshisan describes connecting MiniMax Design, using GPT-6 Astra, to Blender via MCP for direct 3D creation. The creator says this lets 3D representation in Blender be built into an AI production flow rather than relying on MiniMax Design alone.

Sep 10

Sep 10Thu

Sep 9

Sep 9Wed

Sep 8

Sep 8Tue
  1. Fei-Fei LiXAI score13

    Fei-Fei Li says Atlas runs in real time

    AIFei-Fei Li announced on X that Atlas runs in real time. The post gives no further technical details, benchmarks, or availability information beyond the claim itself.

Sep 4

Sep 4Fri

Sep 3

Sep 3Thu
  1. Gemini NotebookOfficialAI score28

    Gemini Notebook rolls out Short Video Overviews in 70+ new languages

    AIGoogle's Gemini Notebook has fully rolled out its international expansion of Short Video Overviews to all web users, adding support for more than 70 new languages and three new English variants. Mobile availability is coming soon, according to the post.

    Video from @Gemini_Notebook's post

Sep 1

Sep 1Tue
  1. Gemini NotebookOfficialAI score28

    Gemini Notebook adds Short Video Overviews in 70+ languages

    AIGemini Notebook now turns sources into roughly 60-second vertical videos called Short Video Overviews in more than 70 languages, plus three new English variants. The feature is rolling out on web and mobile for Ultra and Pro subscribers.

    Video from @Gemini_Notebook's post
  2. Google AI StudioOfficialAI score75

    Google adds agentic video understanding to Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite

    AIGoogle AI Studio says agentic video understanding is now available across Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite via the Gemini API. The company reports cost reductions of up to 66%, token consumption reductions of up to 88% and accuracy gains of up to 7% on standard video benchmarks. Developers enable it by setting processing to "agentic" in the API configuration, at standard token pricing.

    Why it matters: The source gives concrete cost and token figures and explains how the agentic loop replaces fixed-rate frame ingestion, helping developers weigh it against their current video pipelines.

  3. Google AI DevelopersOfficialAI score44

    Gemini adds agentic video understanding across three Flash models

    AIGemini 3.7 Flash, 3.6 Flash, and 3.5 Flash-Lite now support agentic video understanding. The feature is available today for video uploads and YouTube videos through the Gemini API in Google AI Studio and Gemini Enterprise Agent Platform.

  4. Google AI DevelopersOfficialAI score34

    Gemini 3.7 Flash counts rapid claps using agentic video understanding

    AIGoogle's Gemini 3.7 Flash accurately counts every clap in a video by using a new agentic video understanding capability that automatically adapts its processing speed. Static video processing defaults to 1 FPS, which can miss split-second movements or confuse claps with snaps and clicks.

    Video from @googleaidevs's post
  5. Google AI StudioOfficialAI score62

    Google AI Studio introduces agentic video understanding with Gemini

    AIin which the model decides what to watch, at what speed, and through which modality. It fetches only the moments and signals it needs instead of ingesting media at a fixed frame rate. The post says this cuts costs by up to 66% and token consumption by up to 88% while boosting accuracy, and it is available now via the Gemini API and in AI Studio.

    Video from @GoogleAIStudio's post