Updated
#Video
Updated
Showing low-relevance items too. Hide low-relevance items
May 15
Yann LeCun@ylecunXAI score9
May 11
Soumith Chintala@soumithchintalaXAI score22Thinky previews real-time interaction models for human-AI collaboration
AISoumith Chintala, a Thinky-linked voice, said the company is at step one of a plan to increase human-AI bandwidth and raise the ceiling of joint intelligence. He shared a preview of interaction models, described as real-time collaborative tools that talk, listen, watch, and think alongside people. A linked Thinking Machines post describes the approach and early results.
Apr 17
NVIDIA AI Developer@NVIDIAAIDevOfficialAI score23NVIDIA's Cosmos Cookoff winners showcase Cosmos Reason 2 projects
AINVIDIA highlighted the developers who won the Cosmos Cookoff and how they used Cosmos Reason 2 to build projects spanning disaster-response drones, explainable visual AI, and intelligent security systems. The post links to a YouTube showcase and a LinkedIn recap of the event.

Mar 22
FunAudioLLM (Alibaba Tongyi) · new models on Hugging FaceOfficialAI score32 PrismAudio Adds Reinforcement Learning to Video-to-Audio Generation with Chain-of-Thought Planning
AIPrismAudio is a framework that integrates reinforcement learning into video-to-audio generation, using a Chain-of-Thought planning mechanism. It builds on ThinkSound by splitting single-step reasoning into four CoT modules for semantic, temporal, aesthetic, and spatial dimensions, each with targeted reward functions. Code, model weights, and datasets are released for research and educational use under the MIT License, and commercial use requires explicit author authorization.
Mar 17
Xiaomi MiMoOfficialPickAI score71 Xiaomi releases MiMo-V2-Omni, an omni-modal model for agentic tasks
AIXiaomi introduces MiMo-V2-Omni, a single model that fuses image, video, and audio encoders into a shared backbone with native tool calling and UI grounding. The company reports benchmark results against Gemini 3 Pro, Claude Opus 4.6, and GPT 5.2, and demonstrates browser-based shopping and video-publishing workflows run through the OpenClaw agent scaffold. It also states the model supports over 10 hours of continuous audio understanding.
Why it matters: The page gives benchmark comparisons, a driving-risk demo, and browser-task walkthroughs, letting readers check how far the omni-modal claims extend into agent use.
Jan 14
Chip Huyen@chiproXAI score14Agentic Hackathon projects tackle long-running tasks, retrieval, and multimodal agents
AIChip Huyen praised projects at last weekend's Agentic Hackathon, which hosted by MongoDB and Cerebral Valley, where she served as a judge. Teams tackled long-running tasks such as memory management, recovery from mid-task failures, and consistency across steps and sub-agents, along with adaptive retrieval across databases, search indices, and websites. Finalist demos are scheduled in San Francisco tomorrow, with talks by Douglas Eck.

Dec 11, 2025
Runway ResearchOfficialPickAI score62 Runway Introduces GWM-1, a Real-Time General World Model Family
AIRunway announced GWM-1, its first general world model family, built on Gen-4.5 and generating frames autoregressively in real time under interactive control. It comes in three variants: GWM Worlds for explorable environments, GWM Avatars for conversational characters, and GWM Robotics for robotic manipulation. Runway also says it is working toward unifying these domains under a single base world model, and GWM Robotics includes a Python SDK.
Why it matters: The post separates three GWM-1 variants and ties each to a concrete use, which clarifies where a general world model would fit compared with a single model.
Nov 1, 2025
Runway ResearchOfficialPickAI score72 Runway releases Gen-4.5, ranked first on the Text-to-Video benchmark
AIRunway announced Gen-4.5, a video generation model that it says holds the top position on the Artificial Analysis Text-to-Video benchmark with 1,247 Elo points. The model is available across all paid Runway plans at comparable pricing, and the post lists limitations including causal reasoning errors, object permanence failures, and success bias.
Why it matters: The post separates Runway's own ranking claim from the listed limitations, such as causal reasoning and object permanence errors, which helps judge where the model is reliable.