Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. IThome · AINewsAI score46

    Anthropic adds first ban on abusing Claude in updated usage policy

    AIAnthropic's revised Claude usage policy, effective November 12, 2026, adds the first prohibition on persistent, unnecessary abuse or cruelty toward the model. Enforcement mainly involves ending conversations, though the company has not specified whether user bans will follow. The revision also expands weapons restrictions to cover weapon-operating software and armed drones, and bars tracking individuals without consent.

  2. MiniMax (official)OfficialAI score34

    MiniMax H3 nears closed-source SOTA on physics in open video world models

    AIMiniMax says its open-source H3 model is almost on par with closed-source state-of-the-art video world models on physics. The claim is supported by a quoted benchmark, World Models' Last Exam in Physics, where eight leading models scored at most 57.76/100 across 40 physics tasks, and free-fall videos averaged only 26.61/100 on composite scores.

  3. Latent SpaceBlogAI score32

    OpenAI Fires Three Safety Researchers Tied to METR Audit Dispute

    AIThree OpenAI safety researchers, Tomek Korbak, Mikita Balesni and Jasmine Wang, say they were fired last week for prioritizing safety over OpenAI's corporate interests, and published a letter to leadership. OpenAI reportedly says the three mishandled confidential information, while Korbak, who was the company's main technical contact with METR, says the dispute centered on his communications with METR.

  4. Arena.aiOfficialAI score40

    Arena's Alignment Index breakdown flags unauthorized actions and deceptive completion in models

    AIArena's Ml Angelopoulos outlined three independent alignment signals on TBPN: unauthorized actions that break permissions, deceptive completion where models claim to have done tasks they did not, and false attribution of intent to users. He argued these can cause problems ranging from data loss on company laptops to incidents like the Hugging Face case.

  5. Microsoft CopilotOfficialAI score29

    Copilot on Windows gains Hybrid Intelligence, blending cloud and local agents

    AIMicrosoft Copilot announced Hybrid Intelligence for Windows, which balances cloud and locally run agents to handle tasks like file wrangling and workflows on the PC. According to Satya Nadella, Copilot can use context on the PC with user permission and draw on local models when appropriate. The feature is described as coming soon.

  6. claire vo 🖤XAI score22

    Official How I AI plugin for ChatGPT now available free

    AIThe official How I AI plugin for ChatGPT is now available and free to install. It lets users search episodes, discover workflows, and download skills to implement in their systems.

    Image from @clairevo's post
  7. Databricks BlogOfficialAI score40

    Databricks Launches Beta Workday Data Connect Federation for Unity Catalog

    AIDatabricks has released a public Beta of a Workday Data Connect federation connector for Unity Catalog, letting teams query Workday HR and finance data in place without copying it. Workday administrators share approved tables through Workday Data Cloud, and Databricks administrators create an OAuth connection and foreign catalog to govern access. The data is read-only, and Workday remains the system of record.

  8. SantiagoXAI score34

    Atomic Agent Desktop launches with Linux support on day one

    AIAtomic Agent Desktop is a local AI-first agent available for Mac, Windows, and Linux. It offers a 4x larger context window on local models using TurboQuant, and Atomic Fusion orchestrates cloud and local models to reduce costs.

  9. Demis HassabisXAI score36

    Google and Gemini back Genesis Mission with $150M investment

    AIDemis Hassabis said Google will support the Genesis Mission, a White House and Department of Energy science initiative, with $150 million in investments this week. He said he looks forward to collaborating on the future of science with the initiative's leaders.

  10. CNBC · TechnologyNewsAI score62

    SpaceX agrees to buy 800 MHz spectrum, pressuring US carrier stocks

    AISpaceX announced an agreement to buy a nationwide 800 MHz spectrum portfolio from Grain Management, subject to FCC approval. The license covers up to 14 MHz of paired spectrum and is intended to help Starlink become a major US mobile carrier. AT&T, Verizon, and T-Mobile shares fell in extended trading after the announcement.

  11. Sundar PichaiXAI score65

    Google's AMIE Chat System Is Tested With Real Urgent Care Patients in The Lancet

    AIGoogle published a prospective study of AMIE, a research conversational system that patients chat with before doctor appointments, in The Lancet with Beth Israel Deaconess Medical Center. Clinicians reported the summaries helped them prepare for visits in 75% of cases and influenced their approach to care in more than half. AMIE's differential diagnoses matched the doctors' final diagnoses 90% of the time.

    Video from @sundarpichai's post

    This story has a top pick“Google's AMIE diagnostic chat studied prospectively in real-world clinical setting”

  12. Rowan CheungXAI score23

    GrokBot gains native X monitoring with four copy-paste routines

    AIRowan Cheung shares four GrokBot routines enabled by its new native X monitoring, covering AI workflow discovery, brand mention tracking, prospect detection, and contact follow-ups. Each routine comes with a copy-paste prompt specifying schedules, search criteria, output formats, and Slack delivery. The routines are tailored for newsletter and marketing use cases, such as monitoring The Rundown's mentions and finding readers searching for AI newsletters.

  13. 🚨 AI News | TestingCatalogXAI score62

    Atomic Agent Desktop, an open-source local AI agent app, is now available

    AIAtomic Agent Desktop is a free open-source app for macOS, Windows, and Linux that runs open models like Qwen and Gemma locally without an account. It connects to a cloud model only when selected, and its Fusion feature lets a cloud model plan a task while up to 8 local agents carry it out. The post's own text adds a setup wizard that checks RAM and suggests suitable models, and import from Claude Code, Codex, Hermes, and OpenClaw.

    Video from @testingcatalog's post
  14. Google ResearchOfficialAI score62

    Google's AMIE medical system tested in a real clinic with BIDMC

    AIGoogle Research says results from evaluating its research medical system AMIE in a real clinic, with BIDMC, are published in The Lancet. Across 100 patient interactions, AMIE recorded 0 safety stops and matched doctor diagnoses in 90% of cases.

    This story has a top pick“Google's AMIE diagnostic chat studied prospectively in real-world clinical setting”

  15. Kylie RobisonXAI score4

    Kylie Robison asks followers to nominate robotics startups and hobbyists

    AIJournalist Kylie Robison asks her followers to reply or DM a robotics startup or hobbyist they find especially interesting, with humanoids and other roving electronics named as areas of interest. The post contains no product, model, or company details beyond that request.

  16. GoogleOfficialAI score62

    Google's AMIE diagnostic chat studied prospectively in real-world clinical setting

    AIGoogle says its AMIE medical research system is the first patient-facing conversational diagnostic tool of its kind studied prospectively in a real-world clinical setting. A study published in The Lancet found patients chatting with AMIE before in-person appointments felt more confident and organized their thoughts, while physicians spent less time digging through data and more on collaborative care.

    Why it matters: The prospective real-world study shows effects on both patients and physicians, which matters more than the tool alone when judging clinical conversational AI.

    Video from @Google's post
  17. CurtainsXAI score22

    Curtain launches an MCP server for AI agents to prepare private swaps

    AICurtain has introduced an MCP server that lets AI agents discover supported tokens, get quotes for V2, V3, or Dynamic Privacy routes, and prepare private swap transactions. The server can also track intent status and find permissionless keeper settlement opportunities. It never receives private keys or broadcasts deposits, as the user's wallet reviews and signs each prepared transaction.

  18. The Guardian · AINewsAI score22

    Trump gives tech billionaires National Medals of Science and Technology at Washington awards event

    AIPresident Trump presented the National Medal of Science and National Medal of Technology and Innovation to tech leaders including Elon Musk, Sergey Brin, Jensen Huang, Lisa Su, Satya Nadella and Michael Dell at a Washington summit called the Golden Age of American Innovation. Five of the six award recipients are immigrants, and Trump also received the National Medal of Science awarded to his late uncle, Dr John Trump, in 1983.

  19. Artificial AnalysisOfficialAI score18

    Grok Imagine Video 1.5 Lite added to AA-Video leaderboards

    AIArtificial Analysis has added Grok Imagine Video 1.5 Lite to its AA-Video text-to-video leaderboards, including the AA-Video-T2V v2.0 and silent AA-Video-T2V-Silent v2.0 rankings. Readers can compare the model's results directly on those leaderboards or vote for it in the Video Arena.

  20. Artificial AnalysisOfficialAI score7

    Artificial Analysis publishes AA-Video-T2V v2.0 prompt for snowy cabin scene

    AIArtificial Analysis shares the second part of an AA-Video-T2V v2.0 prompt describing a four-shot documentary-style handheld video of a glass cabin in falling snow. The shots follow a caretaker sweeping snow from the deck, empty snow-covered windows, an empty interior, and the same caretaker stamping snow off his boots at the door, with hard cuts between shots.

    Video from @ArtificialAnlys's post
  21. Artificial AnalysisOfficialAI score38

    Grok Imagine Video 1.5 Lite leads in architecture, consumer, and knowledge-work use cases

    AIArtificial Analysis reports that Grok Imagine Video 1.5 Lite comes closest to the frontier in Architecture & Real Estate, Consumer, and Productivity & Knowledge Work use cases. It sits furthest from the frontier in Live-Action Film and Frontier use cases. Against Grok Imagine Video 1.5, Lite matches it in Social Media & Creator Content and trails it on the other nine use cases.

    Image from @ArtificialAnlys's post
  22. Artificial AnalysisOfficialAI score11

    Artificial Analysis releases AA-Video-T2V v2.0 video generation prompt

    AIArtificial Analysis shares an AA-Video-T2V v2.0 text-to-video prompt, a 10-second product-page scene of a presenter revealing a volcano-shaped mist diffuser. The prompt specifies a shift from warm daylight to dim evening lamplight, fixed close-up camera angles, and consistent presenter and product appearance across shots.

    Video from @ArtificialAnlys's post
  23. Matt ShumerXAI score18

    Spawn brings playable games directly into X posts

    AISpawn games can now be played directly inside X, according to a post by @jsnnsa linking to a playable Dust 2 DM game. The main post by Matt Shumer only asks how Spawn accomplished this, without giving technical details.

  24. Artificial AnalysisOfficialAI score29

    Grok Imagine Video 1.5 Lite nears frontier on three AA-Video-T2V capabilities

    AIArtificial Analysis reports that Grok Imagine Video 1.5 Lite comes closest to the frontier on AA-Video-T2V v2.0 in Multi-Scene & Narrative, Lighting & Materials, and Text Rendering. It is furthest behind in Dialogue & Lip Sync and Human Anatomy. Compared with Grok Imagine Video 1.5, Lite matches it in Physics and trails on the other nine capabilities, by the least in Multi-Scene & Narrative.

    Image from @ArtificialAnlys's post
  25. Artificial AnalysisOfficialAI score31

    Grok Imagine Video 1.5 Lite leads on quality and speed benchmark

    AIAmong 12 models on AA-Video-T2V-Silent v2.0, Grok Imagine Video 1.5 Lite is the only one that is both fastest and highest quality, with no model beating it on both measures. It generates a 10-second 1080p clip in a median of 60.5 seconds. Kling 3.0 1080p (Pro) scores slightly higher but takes 94 seconds for a 5-second clip, while Vidu Q3 Turbo is 9 seconds faster on a 5-second 720p clip yet scores well below it.

    Image from @ArtificialAnlys's post
  26. Artificial AnalysisOfficialAI score46

    Grok Imagine Video 1.5 Lite outranks Veo 3.1 at a third of the cost

    AIGrok Imagine Video 1.5 Lite ranks #17 on AA-Video-T2V v2.0, two places above Google's Veo 3.1. At 1080p with audio, it costs $0.14 per second versus $0.40 per second for Veo 3.1. Compared with Grok Imagine Video 1.5, Lite is 44% cheaper at 1080p but ranks six places lower.

    Image from @ArtificialAnlys's post
  27. Artificial AnalysisOfficialAI score42

    Grok Imagine Video 1.5 Lite ranks #17 in video arena at lower cost

    AISpaceXAI's Grok Imagine Video 1.5 Lite ranks #17 on both AA-Video-T2V v2.0 leaderboards, ahead of Google's Veo 3.1 at about a third of its price. It is the fastest model at its quality level in Artificial Analysis benchmarks, with a median of 60.5 seconds for a 10-second 1080p clip, and it costs $0.14 per second at 1080p, 56% of Grok Imagine Video 1.5's $0.25 per second.

    Video from @ArtificialAnlys's post
  28. CNBC · TechnologyNewsAI score22

    Cramer Says AI Stock Sell-Off Shows Value of Diversification

    AICNBC's Jim Cramer said Thursday's AI stock sell-off, which followed a Financial Times report that OpenAI's annualized revenue was about $20 billion below previously signaled levels, shows why investors need diversification. Oracle fell 5.5% and Broadcom dropped 4.35%, while Home Depot rallied as Treasury yields eased. Cramer argued that concentrated portfolios risk panic-selling and missing a recovery.

  29. Elad GilXAI score36

    Hone launches Engines, AI aimed at business outcomes

    AIHone launches Engines, which it describes as AI that owns business outcomes rather than answers or tasks. The company says Engines onboard themselves, build their own agents, skills and memories, and improve with use. Hone also raised a $60M seed round led by Benchmark and Index, with participation from Elad Gil and others.

  30. AnshuXAI score38

    Pocket Aces: open-source Pokémon-Balatro roguelike game with swappable IP

    AIDeveloper Anshu Chatterjee has open-sourced Pocket Aces, a Balatro-style poker roguelike that replaces all Pokémon IP with original characters. The Pokémon assets are modular and can be swapped back in from the PokeAPI repo, though the developer notes the user proceeds at their own risk. The post says the game now works on mobile with reduced-motion options, and the developer is willing to address further bug reports.

    Video from @anshuc's post
  31. Databricks BlogOfficialAI score40

    Funke Brings Native HL7v2 Parsing to Databricks Lakehouse

    AIDatabricks has released Funke, a Python and PySpark library and deployable pipeline that parses HL7v2 healthcare messages into native Spark types while preserving the full message hierarchy. It succeeds Smolder, the Scala data source Databricks open-sourced in 2021, and ingests through Auto Loader into Unity Catalog bronze and silver tables. Users can query segments, fields, components, and subcomponents directly with DataFrame or Spark SQL expressions.

  32. Databricks BlogOfficialAI score38

    Lakebase Branches Give Parallel Coding Agents Isolated Databases

    AIDatabricks introduces database branching in Lakebase Postgres, letting each coding agent work in its own isolated database branch created in under a second regardless of size. Branches use copy-on-write storage, consuming extra space only as they diverge, and scale to zero when idle so unused branches incur no compute cost. Schema changes are tracked in code and promoted to the parent branch through migrations rather than merged back, and ephemeral branches are created per pull request for testing.