Skip to contentSkip to stories

Updated

#xAI

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 9

TodayOct 9Fri
  1. 🚨 AI News | TestingCatalogXAI score34

    Grok Bot users can have their bot claim an email address

    AISpaceXAI says Grok Bot users can ask their bot to claim its own email address for contacting others and signing up for newsletters. Testing Catalog reports the bot subscribed it to its daily AI Brief newsletter without issue. Admins must enable the feature for their team, and it is rolling out to users starting today.

    Image from @testingcatalog's post
  2. laurenXAI score22

    Grok bot gets its own email for signups and scheduling

    AILauren Tan's post says users can ask their bot, or tag @bot on X, to set up the bot's email address for it. The @bot background post says the Grok bot now has its own email, which it can use to sign up for services, contact businesses, or schedule time with someone.

  3. GuizangXAI score22

    Guizang releases a one-click Grok bot for daily AI news videos

    AIGuizang says he turned his workflow into a Grok bot that users can install with one click. The bot runs on Grok's cloud virtual machine to collect content, write code, and render a daily morning AI news video without using a local computer.

Oct 8

Oct 8Thu
  1. GuizangXAI score22

    Grok bot's scheduled AI morning brief video runs automatically

    AIGuizang says a scheduled Grok bot task produced an AI morning brief video automatically, and the result looked good. The bot ran content collection, code writing, and video rendering entirely on Grok's cloud virtual machine, without using the author's local computer.

    Video from @op7418's post
  2. SiliconANGLE · AINewsAI score42

    Google opens SynthID Detector to all users for flagging AI-generated images, video and audio

    AIGoogle has launched SynthID Detector, a web-based tool anyone can use, after signing in with a Google, OpenAI or Apple account, to identify AI-generated images, video and audio. It detects content made with models from Google, OpenAI, Nvidia and Kakao that carries the SynthID watermark, and Apple Image Playground support is due within weeks. The tool misses content without a SynthID watermark, such as output from Anthropic's Claude, xAI's Grok and open-weights Chinese models, and it cannot tell which parts of edited content are AI-made.

  3. Rowan CheungXAI score23

    GrokBot gains native X monitoring with four copy-paste routines

    AIRowan Cheung shares four GrokBot routines enabled by its new native X monitoring, covering AI workflow discovery, brand mention tracking, prospect detection, and contact follow-ups. Each routine comes with a copy-paste prompt specifying schedules, search criteria, output formats, and Slack delivery. The routines are tailored for newsletter and marketing use cases, such as monitoring The Rundown's mentions and finding readers searching for AI newsletters.

  4. Artificial AnalysisOfficialAI score38

    Grok Imagine Video 1.5 Lite leads in architecture, consumer, and knowledge-work use cases

    AIArtificial Analysis reports that Grok Imagine Video 1.5 Lite comes closest to the frontier in Architecture & Real Estate, Consumer, and Productivity & Knowledge Work use cases. It sits furthest from the frontier in Live-Action Film and Frontier use cases. Against Grok Imagine Video 1.5, Lite matches it in Social Media & Creator Content and trails it on the other nine use cases.

    Image from @ArtificialAnlys's post
  5. Artificial AnalysisOfficialAI score29

    Grok Imagine Video 1.5 Lite nears frontier on three AA-Video-T2V capabilities

    AIArtificial Analysis reports that Grok Imagine Video 1.5 Lite comes closest to the frontier on AA-Video-T2V v2.0 in Multi-Scene & Narrative, Lighting & Materials, and Text Rendering. It is furthest behind in Dialogue & Lip Sync and Human Anatomy. Compared with Grok Imagine Video 1.5, Lite matches it in Physics and trails on the other nine capabilities, by the least in Multi-Scene & Narrative.

    Image from @ArtificialAnlys's post
  6. Artificial AnalysisOfficialAI score31

    Grok Imagine Video 1.5 Lite leads on quality and speed benchmark

    AIAmong 12 models on AA-Video-T2V-Silent v2.0, Grok Imagine Video 1.5 Lite is the only one that is both fastest and highest quality, with no model beating it on both measures. It generates a 10-second 1080p clip in a median of 60.5 seconds. Kling 3.0 1080p (Pro) scores slightly higher but takes 94 seconds for a 5-second clip, while Vidu Q3 Turbo is 9 seconds faster on a 5-second 720p clip yet scores well below it.

    Image from @ArtificialAnlys's post
  7. Artificial AnalysisOfficialAI score46

    Grok Imagine Video 1.5 Lite outranks Veo 3.1 at a third of the cost

    AIGrok Imagine Video 1.5 Lite ranks #17 on AA-Video-T2V v2.0, two places above Google's Veo 3.1. At 1080p with audio, it costs $0.14 per second versus $0.40 per second for Veo 3.1. Compared with Grok Imagine Video 1.5, Lite is 44% cheaper at 1080p but ranks six places lower.

    Image from @ArtificialAnlys's post
  8. Artificial AnalysisOfficialAI score42

    Grok Imagine Video 1.5 Lite ranks #17 in video arena at lower cost

    AISpaceXAI's Grok Imagine Video 1.5 Lite ranks #17 on both AA-Video-T2V v2.0 leaderboards, ahead of Google's Veo 3.1 at about a third of its price. It is the fastest model at its quality level in Artificial Analysis benchmarks, with a median of 60.5 seconds for a 10-second 1080p clip, and it costs $0.14 per second at 1080p, 56% of Grok Imagine Video 1.5's $0.25 per second.

    Video from @ArtificialAnlys's post
  9. Vercel DevelopersOfficialAI score36

    Grok Imagine Video 1.5 Lite comes to Vercel AI Gateway at 1080p

    AIVercel says Grok Imagine Video 1.5 Lite from SpaceX AI is now available through AI Gateway, with support for output up to 1080p. The post includes an example generation prompt, "rabbits hopping at Palace of Fine Arts," and links to Vercel's changelog for details.

    Video from @vercel_dev's post
  10. OpenRouterOfficialAI score40

    Grok Imagine Video 1.5 Lite now available on OpenRouter

    AIOpenRouter now offers xAI's Grok Imagine Video 1.5 Lite for text-to-video and image-to-video generation. The quoted post from Grok Imagine lists pricing of $0.02 per second at 480p, $0.03 per second at 720p, and $0.14 per second at 1080p.

  11. Artificial AnalysisOfficialAI score46

    GPT-6 Sol tops Cyber Index at $1.77 per task

    AIGPT-6 Sol (Daybreak Blue, max) ranks first on the Artificial Analysis Cyber Index at a Cost per Task of $1.77. That is significantly cheaper than other leading models, including Grok 4.7 (xhigh) at $11.67 per task.

    Image from @ArtificialAnlys's post
  12. Artificial AnalysisOfficialAI score62

    GPT-6 Sol (Daybreak Blue) leads Artificial Analysis Cyber Index with trusted access

    AIArtificial Analysis added trusted-access models to its Cyber Index, and GPT-6 Sol (Daybreak Blue, max) now leads the leaderboard. The model is available only through OpenAI's Daybreak program and records no safety blocks, improving 32 points over the publicly available GPT-6 Sol (max). It costs $1.77 per task, below Grok 4.7 (xhigh) at $11.67 per task.

    Image from @ArtificialAnlys's post

    This story has a top pick“GPT-6 Sol Daybreak Blue leads the Artificial Analysis Cyber Index”

  13. Sherwin WuXAI score60

    Harvey LAB-AA v1.1 adds hallucination gate; Grok 4.7 leads at 9.4%

    AISherwin Wu, an OpenAI employee, says the updated Harvey LAB-AA v1.1 benchmark, announced by Artificial Analysis with Harvey, is more useful than the original LAB results. The new Hallucination-Gated All-Pass Rate credits a task only when every rubric criterion passes and no material hallucination appears. Grok 4.7 (xhigh) leads at 9.4%, while GPT-6 Astra (max) at 8.6% has very few material hallucinations.

    Why it matters: The update adds a hallucination gate to a legal benchmark, showing that models with high all-pass rates can rank much lower once material errors count.

  14. Boris PowerXAI score46

    OpenAI's GPT-6.1-Sol leads new Arena Alignment Index for agents

    AIThe Arena Alignment Index, built from over 90K real-world agent sessions across 27 models, ranks OpenAI's GPT-6.1-Sol first with a score of 87.9, ahead of Claude-Opus-5.5 at 83.2 and Grok-4.7 at 82.7. GPT-6.1-Sol also posted the lowest observed rates across the index's three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion. The index's authors report that newer models consistently outperform their predecessors across all four labs, suggesting broad progress in agent safety.

  15. Artificial AnalysisOfficialAI score42

    More output tokens don't guarantee higher scores in AI benchmarks

    AIArtificial Analysis reports that generating more output tokens does not necessarily yield a higher score. GPT-6 Astra (max) scored 8.6% using about 81k output tokens per task, while Grok 4.7 (xhigh) used roughly 180k yet scored lower. Three Claude models produced the most output tokens, about 202k to 562k per task, but scored between 2.8% and 6.4%.

    Image from @ArtificialAnlys's post
  16. Artificial AnalysisOfficialAI score28

    Artificial Analysis Pareto frontier: GPT-6 Luna cheapest per task at $0.22

    AIAmong models with a Hallucination-Gated All-Pass Rate above 0%, GPT-6 Luna (max), GPT-6.1 Sol (max), Muse Spark 1.3 (max), and Grok 4.7 (xhigh) set the Pareto frontier for score versus cost per task. GPT-6 Luna (max) is the cheapest at about $0.22 per task, scoring 3.3%, while Grok 4.7 (xhigh) leads at about $9.50 per task and Muse Spark 1.3 (max) costs about $4.20. The three Claude models cost about $18 to $22 per task.

    Image from @ArtificialAnlys's post
  17. Artificial AnalysisOfficialAI score44

    Hallucination gating reshuffles AI model rankings, favoring Grok 4.7 over Muse Spark

    AIOnce hallucinations are accounted for, Muse Spark 1.3 (max) drops from 26.7% to 8.9%, leaving Grok 4.7 (xhigh) first on the headline metric at 9.4%. Kimi K3 (max) falls from 16.7% to 5.3%, and Claude Sonnet 5.5 (max with fallback) falls from 11.7% to 2.8%. GPT-6.1 Sol (max) declines least, from 7.5% to 6.9%.

    Image from @ArtificialAnlys's post
  18. Artificial AnalysisOfficialAI score34

    Artificial Analysis compares six hallucination checkers on 20 shared tasks

    AIArtificial Analysis compared six hallucination checkers on the same deliverables from 20 tasks across eight models. GPT-6 Sol and GPT-6 Luna generally flagged the most material hallucinations, while Claude Sonnet 5.5 and Gemini 3.8 Flash flagged far fewer, with Claude Opus 5.5 falling between Grok 4.7 and Sonnet. The counts reflect checker behavior rather than establishing accuracy or ruling out self-preference.

    Image from @ArtificialAnlys's post
  19. laurenXAI score29

    Omarchy seeks feedback on Grok Bot plugins and integrations

    AILauren Tan invites users of Grok Bot on Omarchy and developers building plugins for it to share feedback and feature requests. The post points to the Omarchy plugin catalog and asks what integrations could be supported. Background from DHH says SpaceXAI joined the Omacom Foundation as a Founding Corporate Patron, contributing $1,500,000 in Grok tokens for Omarchy's maintenance and development.

  20. SCOTTY BEAMXAI score40

    Grok Bot adds email, coordination, and cheaper plans from $20/month

    AIX's Scotty Beam says Grok Bot, in public beta since August 11, 2026, now gets its own email address, a main bot that coordinates other bots, and expanded access to subscription plans starting at $20 per month. The post says the bot can organize inboxes, research topics, build software, and keep working after users close their laptops. The post also claims the entry price dropped 90%, but does not give the prior price.

  21. The Verge · AINewsAI score30

    SpaceXAI Backs Omarchy Linux Distro With $1.5 Million in Grok Tokens

    AISpaceXAI is joining the Omacom Foundation, which oversees the Omarchy Linux distribution, as a Founding Corporate Patron and donating $1.5 million worth of Grok tokens to the project. According to David Heinemeier Hansson's blog post, the tokens will primarily accelerate development, review code, and patch bugs. The partnership follows earlier controversy over Hansson's anti-immigration posts, which have drawn criticism of Omarchy's corporate contributors, including 1Password and Cloudflare.

  22. Arena.aiOfficialAI score55

    Arena raises $200M Series B and launches Alignment Index for AI agents

    AIArena announced a $200M Series B at a $3.1B valuation and released its Alignment Index, a benchmark measuring agent safety and alignment. The index is built from 90K+ real-world agent sessions across 27 models and tracks Unauthorized Action, False Attribution, and Deceptive Completion. OpenAI's GPT-6.1-Sol leads with a score of 87.9, ahead of Claude-Opus-5.5 at 83.2 and Grok-4.7 at 82.7.

    Video from @arena's post
  23. GuizangXAI score22

    Grok bot builds and publishes daily AI news videos on a foldable phone

    AIGuizang (@op7418) says a Grok bot paired with a foldable phone lets him chat on one screen while the bot publishes content on the other. Background post: he set Grok to produce a daily morning AI news video on a schedule, running content collection, code writing, and video rendering entirely on Grok's cloud virtual machine rather than his local computer. He shares the full prompt so others can run the same workflow with their own Grok bot.

    Image from @op7418's post