Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. Design ArenaOfficialAI score40

    Google's Nano Banana 2.1 image model now available on Design Arena

    AIGoogle DeepMind's Nano Banana 2.1, Google's latest image generation and editing model, is now available on Design Arena. The model brings improvements in visual design, mask-based editing, and subject consistency, aiming to produce more natural-looking images with greater control across editing workflows.

    Image from @DesignArena's post
  2. eric zakariassonXAI score32

    Grok turns one product photo into an 8-second vertical video ad

    AIA pipeline built with Grok takes a single product photo and produces an eight-second vertical ad, writing the brief, generating three scenes, selecting the best one, animating it, and adding a voiceover. Each run costs about $1.35, and the code is available in the xai-cookbook repository on GitHub.

    Video from @ericzakariasson's post
  3. eric zakariassonXAI score36

    Grok pipeline turns a one-line premise into a short film

    AIxAI's storyboard-to-film example uses grok-4.7 to plan four shots from a one-line premise, then Grok Imagine draws and animates each one with text-to-speech narration. Every shot is an edit of the first keyframe, which keeps the main character consistent across scenes. The code is available in the xai-cookbook repository on GitHub.

    Video from @ericzakariasson's post
  4. eric zakariassonXAI score29

    Cursor adds five SpaceXAI TypeScript SDK demo apps to cookbook

    AICursor has added five apps to its cookbook, all built on the new SpaceXAI TypeScript SDK. The apps cover premise-to-short-film, picture-to-video ads, screenshot-to-React-component, X posts-to-sentiment dashboard, and link-to-podcast conversion. Demos and code are available in the post.

    Video from @ericzakariasson's post
  5. LangChainOfficialAI score46

    LangChain video shows how to build a model router into an agent harness

    AILangChain's Sydney Runkle presents a four-step method for building a model router into a coding agent harness: understanding tasks, understanding models, building the router, and tracking task outcomes. The background post says the router cut costs by 64% without reducing quality by sending tasks that do not need a frontier model to cheaper models.

  6. 👩‍💻 Paige BaileyXAI score20

    Nano Banana 2.1 Gets Day 0 Support on fal

    AIGoogle's Gemini Nano Banana 2.1 image model is available on fal from day zero, according to Paige Bailey's post thanking fal. The fal background adds that it generates significantly faster than Nano Banana 2 and improves visual design, mask-based editing, and subject consistency.

  7. AMDOfficialAI score24

    OpenAI picks AMD EPYC Turin CPUs to host Jalapeño ASICs

    AIOpenAI selected AMD EPYC "Turin" CPUs to host its Jalapeño AI ASIC deployment, citing platform strength, partner experience, and reducing unnecessary risk. The post, which links to a Tom's Hardware report, says AMD aims to deliver reliable solutions for leaders facing aggressive AI performance goals.

  8. ClaudeOfficialAI score25

    Claude Startup Stack bundles partner credits worth up to $45,000

    AIAnthropic's Claude Startup Stack combines offers from companies building with Claude, including Linear, Lovable, ElevenLabs, Granola, and Hex. Startups get discounts and credits worth up to $45,000 on tools spanning sales, design, data engineering, and more.

    Video from @claudeai's post
  9. Microsoft ResearchOfficialAI score16

    Jennifer Neville on winding research paths and practical AI evaluations

    AIIn a Microsoft Research Podcast episode, Jennifer Neville discusses her nonlinear route into computer science and her push for more practical evaluations of today's AI systems. The post offers little beyond this framing, so no specific models, benchmarks, or results are mentioned.

    Video from @MSFTResearch's post
  10. ARC PrizeOfficialAI score39

    Grok 4.7 reasoning tokens track its ARC-AGI-2 public scores

    AIARC Prize reports that Grok 4.7's reasoning-token usage correlates with its ARC-AGI-2 public scores. The low setting averaged about 10k tokens per test-pair attempt and scored 25%, while medium through xhigh used 86k to 120k tokens and scored 57.5% to 60%. ARC Prize suggests the lower token usage may help explain the low setting's lower score.

    Image from @arcprize's post
  11. ARC PrizeOfficialAI score28

    Grok 4.7 scores 1.8% on ARC-AGI-3 standard harness

    AIGrok 4.7 scored 1.8% on ARC-AGI-3 in the standard harness, which lets models carry notes between turns, slightly below the 2.1% reported for Grok 4.7 in that setting. In a new provider adapter harness that preserves opaque reasoning and enables auto compaction, the score rose to 10.0%.

  12. ARC PrizeOfficialAI score22

    Grok 4.7 uses more reasoning tokens than Grok 4.6 on ARC-AGI-2

    AIGrok 4.7 used more reasoning tokens on average than Grok 4.6 on ARC-AGI-2 semi-private tasks at medium, high, and xhigh reasoning levels, raising its cost per task. Per test-pair attempt, medium used 136% more tokens, high 125% more, and xhigh 173% more, while low used 27% fewer. A chart compares the two models at xhigh on the 20 public tasks where Grok 4.7 increased token use the most.

    Image from @arcprize's post
  13. ARC PrizeOfficialAI score38

    Grok 4.7 scores 1.8% on ARC-AGI-3, trails Grok 4.6 on ARC-AGI-2

    AISpaceXAI's Grok 4.7 scored 90.2% on ARC-AGI-1 at $0.64 per task, higher than Grok 4.6, according to ARC Prize. It reached 61.4% on ARC-AGI-2 at $2.01 per task and 1.8% on ARC-AGI-3 under the standard harness ($2.7k), or 10.0% with the provider adapter harness ($4.8k), both lower than Grok 4.6 on those two benchmarks.

    Image from @arcprize's post
  14. Ai2OfficialAI score13

    Arman Cohan previews COLM 2026 work on RL and research agents

    AIAi2 faculty research scientist Arman Cohan shared a thread previewing his group's upcoming COLM 2026 presentations. The background post says the work covers reinforcement learning with metacognitive rewards, on-policy self-distillation with rubric rewards, and evolving research agents. The main post itself contains only a call to see the thread, so no results or figures are reported.

  15. Ai2OfficialAI score8

    Maarten Sap's Sapling lab reports 10 papers at COLM 2026

    AIMaarten Sap, senior research scientist and technical AI safety lead, announced that his Sapling lab has 10 papers accepted at COLM 2026. The post is a conference-acceptance announcement and includes a link to the main post, with no details on the papers' topics or findings.

  16. Ai2OfficialAI score4

    Pao Siangliulue to discuss human-agent collaboration at COLM 2026

    AIAi2 research scientist Pao Siangliulue says she will attend COLM 2026 in San Francisco from Tuesday to Friday. She plans to discuss human-agent collaboration, long-horizon agents, AI for science, research idea generation, and opportunities at Ai2. She will be at the Ai2 booth on Tuesday from 10 to 11:30 am and at the Ai2 Open House that evening from 6 to 9 pm.

  17. 👩‍💻 Paige BaileyXAI score54

    EmbeddingGemma 2 launches as an Apache 2.0 multimodal embeddings model

    AIGoogle's EmbeddingGemma 2 is an open embeddings model for on-device use that covers code, image, video, audio, and text. It comes in modular sizes from 270M text/code to 740M full multimodal, supports Matryoshka truncation down to 128 dimensions, and reports a 14% gain on MTEB Code over v1 under an Apache 2.0 license. The author's post highlights the release and a Hugging Face demo, while the benchmark table compares it with several models.

    Video from @DynamicWebPaige's post
  18. KreaOfficialAI score36

    Nano Banana 2.1 launches on Krea with 4K generation support

    AIKrea has made Nano Banana 2.1 available on its platform, promising better prompt understanding, sharper edits, and stronger subject consistency. The model now supports generations up to 4K resolution.

    Video from @krea_ai's post
  19. Alexandr WangXAI score4

    Alexandr Wang mocks Jon Stewart's Daily Show Meta AI segment

    AIMeta AI leader Alexandr Wang reacted to The Daily Show's segment featuring Jon Stewart talking with Meta's Muse AI agent, asking which corporate figure booked the appearance. The show's background description portrays Muse as an AI agent Stewart chats with in a joking, suspicious tone.

  20. Google WorkspaceOfficialAI score8

    Google shows how Workspace Studio and Gemini cut busywork

    AIGoogle's productivity advisor Laura Mae Martin explains how Workspace Studio and Gemini can lighten users' workloads by reducing routine tasks. The post promotes a Google article on the topic but gives no specific features, figures, or availability details.

    Image from @GoogleWorkspace's post
  21. Dongxi NLPXAI score34

    Mistral AI releases Mistral Large 4, dubbed "Le Chonk"

    AIMistral AI has released Mistral Large 4, a model nicknamed "Le Chonk," according to a post by Dongxi NLP. The post also highlights "sovereign AI" as a keyword, tying the release to the theme of national or independent AI capability. No specifications, benchmarks, or pricing are given in the post itself.

  22. Google FlowOfficialAI score6

    Google Flow shares a full documentary on YouTube

    AIGoogle Flow's X account shares a link to watch a full documentary on YouTube. The post provides no further details about the film's subject or content.

  23. Google FlowOfficialAI score20

    Google Flow partners with Divine on "Tequila Dance" music video

    AIGoogle Flow says it partnered with Divine to create the music video for "Tequila Dance," framing the project as a showcase of how technology can amplify human expression. The post provides no further details about the tools or production process used.

    Video from @FlowbyGoogle's post
  24. 👩‍💻 Paige BaileyXAI score60

    Google releases Nano Banana 2.1 image model at $0.034 per image

    AIGoogle's Nano Banana 2.1, model gemini-nano-banana-2.1, is now available and is said to outperform the previous Pro model at about a quarter of the price, $0.034 per image versus $0.134. The quoted post lists improved instruction following, better in-image text rendering, grounding with Google Image Search, and up to 5 characters of consistency plus 14 reference images. It is available in Google AI Studio, the Gemini API, Google Cloud, the Gemini app, and Flow. The author's own post is a playful reaction praising its design ability and shows a generated vegan basketball food truck poster.

    Why it matters: The quoted announcement gives concrete changes and a price drop for image generation, useful for weighing cost against the previous Pro model.

    Video from @DynamicWebPaige's post
  25. Microsoft ResearchOfficialAI score36

    Jennifer Neville on learning from surprising AI failures and evaluation beyond benchmarks

    AIMicrosoft Research podcast host Chad Atalla interviews Jennifer Neville, a partner research manager at Microsoft, about her path into AI and her work on how evaluation exposes surprising failures in models tested beyond traditional benchmarks. The conversation also covers practical guidance for working with current AI systems and why examining underlying data matters when results defy expectations.