Skip to contentSkip to stories

Updated

#xAI

Oct 8

Oct 8Thu
  1. SiliconANGLE · AIAI score42

    Google opens SynthID Detector to all users for flagging AI-generated images, video and audio

    AIGoogle has launched SynthID Detector, a web-based tool anyone can use, after signing in with a Google, OpenAI or Apple account, to identify AI-generated images, video and audio. It detects content made with models from Google, OpenAI, Nvidia and Kakao that carries the SynthID watermark, and Apple Image Playground support is due within weeks. The tool misses content without a SynthID watermark, such as output from Anthropic's Claude, xAI's Grok and open-weights Chinese models, and it cannot tell which parts of edited content are AI-made.

  2. SiliconANGLE · AIAI score49

    US venture deal value hits record $515.8B, but exits lag behind

    AIUS venture capital deal value reached a record $515.8 billion in the first nine months of 2026, driven largely by giant AI rounds for OpenAI and Anthropic, according to the PitchBook-NVCA Venture Monitor. Exit activity has not kept pace, with third-quarter exit value heavily dependent on SpaceX's $60 billion all-stock purchase of Anysphere, and IPOs remaining sparse.

  3. Rowan CheungAI score23

    GrokBot gains native X monitoring with four copy-paste routines

    AIRowan Cheung shares four GrokBot routines enabled by its new native X monitoring, covering AI workflow discovery, brand mention tracking, prospect detection, and contact follow-ups. Each routine comes with a copy-paste prompt specifying schedules, search criteria, output formats, and Slack delivery. The routines are tailored for newsletter and marketing use cases, such as monitoring The Rundown's mentions and finding readers searching for AI newsletters.

  4. Artificial AnalysisAI score38

    Grok Imagine Video 1.5 Lite leads in architecture, consumer, and knowledge-work use cases

    AIArtificial Analysis reports that Grok Imagine Video 1.5 Lite comes closest to the frontier in Architecture & Real Estate, Consumer, and Productivity & Knowledge Work use cases. It sits furthest from the frontier in Live-Action Film and Frontier use cases. Against Grok Imagine Video 1.5, Lite matches it in Social Media & Creator Content and trails it on the other nine use cases.

  5. Artificial AnalysisAI score29

    Grok Imagine Video 1.5 Lite nears frontier on three AA-Video-T2V capabilities

    AIArtificial Analysis reports that Grok Imagine Video 1.5 Lite comes closest to the frontier on AA-Video-T2V v2.0 in Multi-Scene & Narrative, Lighting & Materials, and Text Rendering. It is furthest behind in Dialogue & Lip Sync and Human Anatomy. Compared with Grok Imagine Video 1.5, Lite matches it in Physics and trails on the other nine capabilities, by the least in Multi-Scene & Narrative.

  6. Artificial AnalysisAI score31

    Grok Imagine Video 1.5 Lite sits on the quality and speed frontier on AA-Video-T2V-Silent v2.0 Among the 12 models on AA-Video-T2V-Silent…

    AI…v2.0 that we benchmark for generation speed, no model is both faster and higher quality than Grok Imagine Video 1.5 Lite. It generates a 10 second 1080p clip in a median of 60.5 seconds. Kling 3.0 1080p (Pro) scores slightly higher and takes 94 seconds for a 5 second clip. Vidu Q3 Turbo is 9 seconds faster on a 5 second 720p clip, and scores well below it.

  7. Artificial AnalysisAI score42

    Grok Imagine Video 1.5 Lite ranks #17 in video arena at lower cost

    AISpaceXAI's Grok Imagine Video 1.5 Lite ranks #17 on both AA-Video-T2V v2.0 leaderboards, ahead of Google's Veo 3.1 at about a third of its price. It is the fastest model at its quality level in Artificial Analysis benchmarks, with a median of 60.5 seconds for a 10-second 1080p clip, and it costs $0.14 per second at 1080p, 56% of Grok Imagine Video 1.5's $0.25 per second.

  8. Artificial AnalysisAI score62

    GPT-6 Sol Daybreak Blue leads the Artificial Analysis Cyber Index

    AIArtificial Analysis added trusted-access models to its Cyber Index, and GPT-6 Sol (Daybreak Blue, max) now ranks first. The model is available only through OpenAI's Daybreak program and records no safety blocks across the Index. Its overall score is 32 points higher than the publicly available GPT-6 Sol (max), at a cost of $1.77 per task versus $11.67 for Grok 4.7 (xhigh).

  9. Artificial AnalysisAI score28

    Among models with a Hallucination-Gated All-Pass Rate above 0%, four set the Pareto frontier for score vs.

    AICost per Task: GPT-6 Luna (max), GPT-6.1 Sol (max), Muse Spark 1.3 (max) and Grok 4.7 (xhigh). Grok 4.7 (xhigh) leads at ~$9.50 per task and Muse Spark 1.3 (max) comes second at ~$4.20, while the three Claude models cost ~$18 to ~$22 per task. GPT-6 Luna (max) is the cheapest at ~$0.22 per task, scoring 3.3%.

  10. Artificial AnalysisAI score34

    Artificial Analysis compares six hallucination checkers on 20 shared tasks

    AIArtificial Analysis compared six hallucination checkers on the same deliverables from 20 tasks across eight models. GPT-6 Sol and GPT-6 Luna generally flagged the most material hallucinations, while Claude Sonnet 5.5 and Gemini 3.8 Flash flagged far fewer, with Claude Opus 5.5 falling between Grok 4.7 and Sonnet. The counts reflect checker behavior rather than establishing accuracy or ruling out self-preference.

  11. The Verge · AIAI score30

    SpaceXAI Backs Omarchy Linux Distro With $1.5 Million in Grok Tokens

    AISpaceXAI is joining the Omacom Foundation, which oversees the Omarchy Linux distribution, as a Founding Corporate Patron and donating $1.5 million worth of Grok tokens to the project. According to David Heinemeier Hansson's blog post, the tokens will primarily accelerate development, review code, and patch bugs. The partnership follows earlier controversy over Hansson's anti-immigration posts, which have drawn criticism of Omarchy's corporate contributors, including 1Password and Cloudflare.

  12. Gizmodo · AIAI score13

    Trump Declares Anyone Using "Artificial Intelligence" Term "THE ENEMY" on Truth Social

    AIPresident Donald Trump posted on Truth Social that the White House considers anyone who uses the term "Artificial Intelligence," instead of "Super Intelligence," to be "THE ENEMY." The article reports that Nvidia's Jensen Huang, Meta's Mark Zuckerberg, Elon Musk, and Jeff Bezos have adopted the "super intelligence" term, and that ai.gov was changed to read "SI" while si.gov appears dormant.

  13. GuizangAI score26

    Grok bot auto-generates a daily AI news video in the cloud

    AIThe author set up a Grok bot to produce a daily morning AI news video on a schedule, running content collection, code writing, and video rendering entirely on Grok's cloud virtual machine without local computers. The author says the results are quite good and shares the full prompt so others can run the same workflow with their own Grok bot.

  14. Artificial Analysis ArticlesAI score50

    Harvey LAB-AA v1.1 adds hallucination checks to legal AI benchmark

    AIHarvey LAB-AA v1.1 adds hallucination checks that audit every model deliverable against task source documents, with material hallucinations zeroing a task's score. GPT-6 Astra averaged 0.03 material hallucinations per task across 120 tasks, while Gemini 3.8 Flash averaged 13.96. Harvey uses GPT-6 Sol (high) as the hallucination checker, separate from its three-judge rubric panel.

Oct 7

Oct 7Wed
  1. GuizangAI score35

    Grok Bot can now search, read, and monitor X posts

    AIGrok bot 最近真猛啊! 现在可以搜索、阅读甚至监控指定的推特内容 比如你可以让他在 tibo 重置 Codex 额度的时候通知你 这些都是免费的,不通过以前的 xapi 实现 --- Translation: Grok bot has been really impressive lately! Now it can search, read, and even monitor specified tweets on X For example, you can have it notify you when tibo resets the Codex quota All of this is free, and not implemented via the previous xapi

  2. Epoch AIAI score67

    Epoch tests six AI models on real Epoch work and finds they cannot yet fully automate it

    AIEpoch gave six models 11 real work tasks from its own operations, including graphic design, data insights, and research design, and graded outputs against employee standards. Fable 5.1 and GPT-6 Astra led on average task performance, reliably handling well-defined work such as coding and computational analysis. The report finds that all models still fail on open-ended judgment, including matching Epoch's standards, designing informative experiments, and generating diverse ideas, so the authors conclude AI cannot yet replace workers at Epoch.

    Why it matters: The report separates well-defined task reliability from open-ended judgment failures, which benchmark scores on easily verifiable tasks would miss.