Skip to contentSkip to stories

Updated

#Video

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 9

TodayOct 9Fri
  1. Artificial AnalysisOfficialAI score32

    HiDream-O1-Video-1.0 ranks #6 on Artificial Analysis image-to-video leaderboard

    AIHiDream-O1-Video-1.0 ranks #6 in Artificial Analysis's Image to Video with Audio leaderboard, just behind Dreamina Seedance 2.0 720p. HiDream says the model generates 1080p videos of 5 to 20 seconds with synchronized audio, priced at $5.80 per minute ($0.10 per second) on the HiHarness API. It is also available in vivago R1 Studio.

    GIF from @ArtificialAnlys's post
  2. Cloudflare Blog · AIOfficialAI score55

    Cloudflare releases Clef-omni with audio and video input and cuts Clef-flash price

    AICloudflare releases Clef-omni, an open-weight decision model that accepts audio, video, image, and text input in a single API call. Clef-flash's price falls from $0.09 to $0.038 per M input tokens, while its hosted context window drops from 64k to 24k. Cloudflare also reports median latency reductions of 1.7 to 2.0 times for the Clef model on Workers AI.

  3. GuizangXAI score22

    Guizang releases a one-click Grok bot for daily AI news videos

    AIGuizang says he turned his workflow into a Grok bot that users can install with one click. The bot runs on Grok's cloud virtual machine to collect content, write code, and render a daily morning AI news video without using a local computer.

  4. IThome · AINewsAI score55

    Odyssey-3 world model scores 66.1 on Physics-IQ Verified benchmark

    AIOdyssey announced the Odyssey-3 series of foundation world models, with Odyssey-3 Pro scoring 66.1 on the Physics-IQ Verified video-to-video benchmark, the highest recorded on that leaderboard. The series includes a standard version balancing physical accuracy and generation cost, and a Pro version with stronger physics prediction. The preview supports first-person and third-person navigation and lets users move the camera, take actions, or trigger events while the model predicts environmental changes in real time.

  5. MiniMax Design (H3)OfficialAI score22

    MiniMax H3 Person Remover LoRA erases people from video

    AIA LoRA for MiniMax H3 removes a person from video by tracking them with SAM 3.1 and generating the replacement background in overlapping windows. Users supply the original video and a clean version of its first frame.

Oct 8

Oct 8Thu
  1. Higgsfield AI 🧩OfficialAI score36

    Higgsfield Katana adds community presets for Claude video editing

    AIHiggsfield has released community presets for Higgsfield Katana, its AI video editing tool available inside Claude. Users can pick a preset for motion graphics, 3D animations, product launches, fashion, car, travel, or aura-farming edits, then add their own characters, products, or clothes to recreate it in Claude. More presets are coming soon.

    Video from @higgsfield's post
  2. GuizangXAI score22

    Grok bot's scheduled AI morning brief video runs automatically

    AIGuizang says a scheduled Grok bot task produced an AI morning brief video automatically, and the result looked good. The bot ran content collection, code writing, and video rendering entirely on Grok's cloud virtual machine, without using the author's local computer.

    Video from @op7418's post
  3. LumaOfficialAI score22

    Jon Erwin's Moses made using AI to extend real actors' worlds

    AILuma Labs posted that filmmaker Jon Erwin made Moses on a Manhattan Beach stage with real actors, using AI to carry them into any world the story required. The post frames AI as removing budget limits on where a story can be filmed.

    Video from @LumaLabsAI's post
  4. MiniMax (official)OfficialAI score34

    MiniMax H3 nears closed-source SOTA on physics in open video world models

    AIMiniMax says its open-source H3 model is almost on par with closed-source state-of-the-art video world models on physics. The claim is supported by a quoted benchmark, World Models' Last Exam in Physics, where eight leading models scored at most 57.76/100 across 40 physics tasks, and free-fall videos averaged only 26.61/100 on composite scores.

  5. Artificial AnalysisOfficialAI score31

    Grok Imagine Video 1.5 Lite leads on quality and speed benchmark

    AIAmong 12 models on AA-Video-T2V-Silent v2.0, Grok Imagine Video 1.5 Lite is the only one that is both fastest and highest quality, with no model beating it on both measures. It generates a 10-second 1080p clip in a median of 60.5 seconds. Kling 3.0 1080p (Pro) scores slightly higher but takes 94 seconds for a 5-second clip, while Vidu Q3 Turbo is 9 seconds faster on a 5-second 720p clip yet scores well below it.

    Image from @ArtificialAnlys's post
  6. Artificial AnalysisOfficialAI score46

    Grok Imagine Video 1.5 Lite outranks Veo 3.1 at a third of the cost

    AIGrok Imagine Video 1.5 Lite ranks #17 on AA-Video-T2V v2.0, two places above Google's Veo 3.1. At 1080p with audio, it costs $0.14 per second versus $0.40 per second for Veo 3.1. Compared with Grok Imagine Video 1.5, Lite is 44% cheaper at 1080p but ranks six places lower.

    Image from @ArtificialAnlys's post
  7. Artificial AnalysisOfficialAI score42

    Grok Imagine Video 1.5 Lite ranks #17 in video arena at lower cost

    AISpaceXAI's Grok Imagine Video 1.5 Lite ranks #17 on both AA-Video-T2V v2.0 leaderboards, ahead of Google's Veo 3.1 at about a third of its price. It is the fastest model at its quality level in Artificial Analysis benchmarks, with a median of 60.5 seconds for a 10-second 1080p clip, and it costs $0.14 per second at 1080p, 56% of Grok Imagine Video 1.5's $0.25 per second.

    Video from @ArtificialAnlys's post
  8. Vercel DevelopersOfficialAI score36

    Grok Imagine Video 1.5 Lite comes to Vercel AI Gateway at 1080p

    AIVercel says Grok Imagine Video 1.5 Lite from SpaceX AI is now available through AI Gateway, with support for output up to 1080p. The post includes an example generation prompt, "rabbits hopping at Palace of Fine Arts," and links to Vercel's changelog for details.

    Video from @vercel_dev's post
  9. OpenRouterOfficialAI score40

    Grok Imagine Video 1.5 Lite now available on OpenRouter

    AIOpenRouter now offers xAI's Grok Imagine Video 1.5 Lite for text-to-video and image-to-video generation. The quoted post from Grok Imagine lists pricing of $0.02 per second at 480p, $0.03 per second at 720p, and $0.14 per second at 1080p.

  10. The DecoderNewsAI score65

    Anthropic launches Claude Dashboards and Motion features in beta

    AIAnthropic launched two beta features for Claude: Dashboards, which turns connected data sources like BigQuery, Databricks, Snowflake, or Salesforce into auto-updating live dashboards from text prompts, and Motion, which creates animated explainer videos from text, diagrams, and images. Dashboards is available to paid users and Motion to Team and Enterprise plans, while Docs, Slides, and Design leave beta and work across all plans, including free accounts.

  11. LumaOfficialAI score25

    Luma lets users continue Claude Motion animations in Luma

    AILuma says users can bring Claude Motion animations into Luma to resize them for different formats and refine them for shipping. Claude Motion is in beta on Claude Team and Enterprise plans.

    Image from @LumaLabsAI's post
  12. 🚨 AI News | TestingCatalogXAI score49

    Odyssey launches Odyssey-3 world model with public research preview

    AIOdyssey has launched Odyssey-3, its most powerful foundation world model, with a public research preview. Odyssey-3 Pro scored 66.1 on Physics-IQ Verified video-to-video with best-of-8 sampling, the highest reported result. The model generates environments from prompts and predicts changes in real time as users move through scenes.

    Image from @testingcatalog's post
  13. SantiagoXAI score46

    Odyssey 3 Pro world model tops Physics-IQ and goes live

    AIOdyssey 3 Pro, a world model, is now live as a research preview and ranks first on the Physics-IQ Verified video-to-video benchmark. The post says it can learn from visual observations and map that knowledge to physical controls for robots, cars, video games, and drones. Odyssey-3, the model launched alongside it, is described as free to try.

    Image from @svpino's post
  14. OdysseyOfficialAI score31

    Odyssey-3 world model debuts for physical AI and training environments

    AIOdyssey has released Odyssey-3, which it describes as a major leap toward world models that power physical AI, generate training environments, and enable new human experiences. The post invites readers to try Odyssey-3 at the company's website but gives no specific benchmarks, parameter counts, or pricing.

  15. OdysseyOfficialAI score38

    Odyssey-3 is a foundation world model for physical AI and agents

    AIOdyssey announced Odyssey-3, a foundation world model it says enables applications in physical AI, human experiences, and training intelligences. The company highlights agents learning from experience inside Odyssey-3 while working toward objectives.

    Video from @odysseyml's post
  16. OdysseyOfficialAI score22

    Odyssey-3 Pro sets new Physics-IQ video-to-video benchmark record

    AIOdyssey-3 Pro achieved a score of 66.1 on Physics-IQ Verified's video-to-video benchmark, the highest reported score so far. Physics-IQ evaluates physical behavior across fluid dynamics, optics, solid mechanics, magnetism, and thermodynamics.

    Image from @odysseyml's post
  17. QbitAINewsAI score47

    Vidu Q4 Preview Offers 4K Video Generation at About 0.09 Yuan per Second

    AIShengshu Technology has opened a preview of its Vidu Q4 video generation model, which supports native 4K output and up to 15 reference images and three reference audio clips. Testers generated a one-minute video for about 5.4 yuan, roughly 0.09 yuan per second at 720P, which the article says is a starting price that varies by resolution and mode. The Vidu Q4 preview is available through the Vidu platform, with the MaaS API priced at about 0.6 yuan per second for 720P image-to-video.

  18. Luma AI NewsOfficialAI score46

    Luma Lets Creators Carry Claude Motion Animations into Its Video Tools

    AILuma announced that Claude Motion animations can now open directly in Luma through an MCP connection, letting creators restyle them and reframe them to 9:16, 1:1, 4:3, or 21:9. Claude Motion, currently in beta on Claude Team and Enterprise plans, generates animated explainers from prompts, while Luma's Ray and Uni video models produce final video files.

Oct 7

Oct 7Wed
  1. ComfyUIOfficialAI score22

    Vidu Q4 preview now available to try on ComfyUI

    AIComfyUI announces a preview of Vidu Q4, inviting users to try it via a linked page. The post gives no further details on features, pricing, or availability terms.

  2. ComfyUIOfficialAI score43

    Vidu Q4 Preview arrives in ComfyUI via Partner Nodes

    AIComfyUI says Vidu Q4 Preview, the first preview of Vidu's new flagship video model, is now available through Partner Nodes. The model offers finer character acting with expressions, emotion, and body language, voice consistency using up to three reference audio clips, and up to 15 reference images per shot. It outputs up to 16 seconds at 2K and 4K, with smoother cuts and camera moves across shots.

    Video from @ComfyUI's post
  3. falOfficialAI score46

    fal Launches H3 Max Relight for Changing Video Lighting Without Reshoots

    AIfal introduced H3 Max Relight, a tool that changes the lighting of uploaded videos through a built-in Lighting studio where users pick colors, orbit lights around subjects, and adjust intensity and softness. It preserves the original subjects, motion, camera movement, and audio while relighting every frame, and the post says it is powered by H3 Max, which it calls the #1 model for overall quality, prompt understanding, and aesthetics.

    Video from @fal's post
  4. falOfficialAI score34

    Vidu Q4 video model now live on fal with native audio

    AIfal has launched Vidu Q4, offering image-to-video generation from a single first frame with native audio. Reference-to-video supports up to 12 reference images and 3 voice clips for consistent characters and voices. Clips run 3 to 16 seconds at resolutions from 540p up to 4K.

    Video from @fal's post
  5. RunwayOfficialAI score36

    Runway launches an ability to work directly inside ChatGPT

    AIRunway says users can brief its tool, let it work, and give notes from the same chat window with Runway directly inside ChatGPT Astra. The post invites readers to get started through a linked page.

    Video from @runwayml's post
  6. Aravind SrinivasXAI score20

    Perplexity's Aravind Srinivas Celebrates AI-Generated Animated Scene Workflow

    AIAravind Srinivas posted that "we're living in incredible times," highlighting a fully computer-made animated scene produced from concept art to final edit. The quoted post says the scene was planned in Blender, with reference images generated by Nano Banana and animation and score produced by Seedance 2.5.

  7. PixVerseOfficialAI score20

    PixVerse plugin lets users generate videos directly within chat

    AIThe PixVerse plugin enables video creation from text, images, or video references inside the chat interface. Users select the plugin, describe a scene or add a reference, then specify model, duration, resolution, and aspect ratio before generating.

    Video from @PixVerse's post

Oct 6

Oct 6Tue
  1. Kling AIOfficialAI score22

    Kling AI to showcase Kling 4.0 projects at Busan's ACFM in October

    AIKling AI plans to present real-world projects made with its upcoming Kling 4.0 model at ACFM in Busan, Korea, from October 10 to 13, 2026. The event will explore how generative AI can interpret directors' creative intent, maintain character and narrative continuity, and fit into professional animation, live-action, film, and series workflows.

    Image from @Kling_ai's post
  2. meng shaoXAI score52

    xAI Cookbook adds five apps, expanding Grok API examples to ten

    AIThe xAI Cookbook now has ten runnable Grok API examples across three tracks: real-time voice agents, multimodal generation, and live X data analysis. The author says four voice examples show the same Realtime Voice API across WebSocket, WebRTC, Twilio phone, and mobile transports. The four multimodal examples chain understanding, image generation or editing, video, and TTS, with Grok making creative decisions and Imagine models executing them.

    Image from @shao__meng's post