Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. Alex HeathXAI score38

    Qualcomm CEO Cristiano Amon on AI phones, glasses, and 6G

    AIQualcomm CEO Cristiano Amon discusses the coming AI smartphone supercycle, arguing phones will not disappear as agents use personal context. He also expects smart glasses to become the largest AI wearable category, and covers Qualcomm's Modular acquisition as an alternative to Nvidia's CUDA software and its data center strategy. The conversation, recorded live at the Snapdragon Summit in Hawaii, also covers 6G being designed for AI.

    Video from @alexeheath's post
  2. TiboXAI score62

    OpenAI rolls out GPT-6.1 Sol ultrafast with faster steering

    AITibo, an OpenAI team member, says GPT-6.1 Sol ultrafast is rolling out today in the API, Codex, and ChatGPT Work. He says it offers near-Astra intelligence at up to 8x the speed of Sol Standard. The post also says improved steering now lets the model react faster to user adjustments in real time.

    Why it matters: The post specifies the new Ultrafast option, its availability across API, Codex, and ChatGPT Work, and its speed claim relative to Sol Standard.

    Video from @thsottiaux's post
  3. GoodfireOfficialAI score25

    Goodfire launches a challenge to build AI models on a dataset

    AIParticipants will use a dataset to build AI models, evaluated through a series of evals ranging from general benchmarks to more complex tasks. Top teams will be shortlisted and have their experimental hypotheses tested in Prima Mente's wet lab.

  4. GoodfireOfficialAI score44

    Alzheimer's Translation Challenge Built on 150M-Cell Atlas

    AIThe Alzheimer's Translation Challenge is built on a new atlas of 150M cells, covering neurons, astrocytes, and microglia across different genetic backgrounds under combinatorial perturbations with multi-modal readouts. The data will be made available through the AD workbench and Prima Mente's modeling platform.

  5. Boris PowerXAI score46

    OpenAI's GPT-6.1-Sol leads new Arena Alignment Index for agents

    AIThe Arena Alignment Index, built from over 90K real-world agent sessions across 27 models, ranks OpenAI's GPT-6.1-Sol first with a score of 87.9, ahead of Claude-Opus-5.5 at 83.2 and Grok-4.7 at 82.7. GPT-6.1-Sol also posted the lowest observed rates across the index's three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion. The index's authors report that newer models consistently outperform their predecessors across all four labs, suggesting broad progress in agent safety.

  6. Harrison ChaseXAI score28

    LangChain's Sam explains decision models and using Jev in harnesses

    AISam from LangChain discusses where decision models fit inside an agent harness and how to use Jev with LangChain. The post links to a LangChain blog on building a harness with Jev, which the background post says Jev from typesafeai popularized alongside OpenAI's Decisions API and Databricks' ai_decide function.

  7. The DecoderNewsAI score62

    Anthropic updates Claude usage policy to ban sustained abusive behavior toward Claude

    AIAnthropic has updated its Claude usage policy for the first time in over a year, banning sustained and needless abusive or cruel behavior toward Claude. Violations can lead to warnings, throttling, restriction, suspension, or termination, and the company says the rule applies only in extreme cases. The update also expands bans on propaganda, weapons, and surveillance, and it allows exceptions for government contracts.

  8. ClaudeOfficialAI score46

    Claude Motion turns reports into editable code-based animations

    AIAnthropic's Claude Motion converts reports, charts, or product walkthroughs into short animations. Claude writes each animation as code rather than using a video model, so users can edit any word, number, or timing and export an MP4. The feature is in beta on Team and Enterprise plans.

    Image from @claudeai's post
  9. Comfy BlogOfficialAI score34

    How I Generated Live Video with MiniMax H3 on a Single GPU

    AIA ComfyUI developer generated 15-second 448×256 video in 15 seconds or less on one RTX 5090 using MiniMax H3 with FastVideo's FastH3 V2 checkpoint in four sampling steps. The setup combined sparse attention, a smaller ClipProj text encoder, a pruned INT8 checkpoint, and a fused FP4 MLP, cutting VRAM needs from 80GB to under 30GB. The custom ComfyUI node is open source.

  10. Andrew CurranXAI score62

    Three fired OpenAI safety researchers publish open letter to leadership

    AIThree OpenAI safety and alignment employees, Tomek Korbak, Jasmine Wang, and Mikita Balesni, were fired last week and have published an open letter to OpenAI's safety and governance committees. The letter argues that OpenAI cannot make AI safe on its own, calls for open debate, third-party collaboration, and clear internal procedures, and says the firing and its handling bear directly on safety oversight.

    Image from @AndrewCurran_'s post
  11. Codex · GitHub ReleasesOfficialAI score36

    Codex 0.162.0 adds managed worktree tools and clickable URLs in the TUI

    AIOpenAI's Codex 0.162.0 release adds tools for creating and listing managed Git worktrees from trusted local projects when the worktrees feature is enabled. The update also lets users pin tasks in the agent Command Center, copy transcript blocks with /copy, and make URLs clickable in approval headers, questions, and warnings, along with several Linux and Windows sandbox fixes.

  12. ZDNet · AINewsAI score50

    Microsoft 365 Family and Premium plans will switch to shared storage and AI credit pools

    AIMicrosoft will replace per-user OneDrive allowances on Microsoft 365 Family and Premium plans with a shared 2 TB storage pool, down from 6 TB across six accounts. Heavy users could face bills that double or triple, since each extra 1 TB adds $10 a month, while family members will be able to share a single AI usage allowance. New and upgraded subscriptions get shared storage starting Oct. 8, 2026, and existing subscribers move at their first renewal on or after May 2, 2027.

  13. Gizmodo · AINewsAI score18

    REK Stages Human-Versus-Robot Fights, Then Puts Swords on Robots

    AIRobot Entertainment Kombat, founded by Cix Liv, hosted a September 18 San Francisco event where a human fought a remotely controlled robot and lost each match, leaving with a hand injury. The California State Athletic Commission sent a cease-and-desist, saying the fights were unsanctioned and require approval before any bout involving a human. REK plans a robot deathmatch later this month with remotely controlled bots equipped with weapons.

  14. 🚨 AI News | TestingCatalogXAI score36

    Gemini Agent for Business may add Claude Opus 5 and Sonnet 5.5

    AIGoogle's recently announced Gemini Agent for Gemini Business is reportedly set to offer Gemini Argon 4, Gemini Flash 3.8, Claude Opus 5, and Claude Sonnet 5.5. If accurate, it would mark the first time Claude models appear on Google's platform alongside Google's own models, which the post frames as a way for Google to compete for enterprise customers.

    Video from @testingcatalog's post
  15. CNBC · TechnologyNewsAI score50

    US suspends Microsoft, Adobe from green card labor program amid foreign worker crackdown

    AIThe U.S. Department of Labor said it suspended Microsoft and Adobe from its Permanent Labor Certification program, citing multiple active federal investigations. Labor Secretary Keith Sonderling also said no new applications will be accepted for Cognizant, Infosys, Capgemini, Tata, Wipro and HCL. Microsoft said the vast majority of its U.S. employees are Americans and that 80% of its roughly 6,000 H-1B petitions last fiscal year were to extend or change the status of existing employees.

  16. Arena.aiOfficialAI score44

    Arena raises $200M Series B led by Lightspeed, launches Alignment Index

    AIArena has secured a $200M Series B, with Lightspeed doubling down on its investment. The company is also launching the Alignment Index, which measures how closely AI behavior aligns with human values in real-world settings. Arena reports annualized revenue above $100M since its Series A, with millions of people helping evaluate frontier models through real-world use.

  17. Tessl BlogOfficialAI score29

    One Brain Means Owning Your Organizational Memory

    AILeapfrog, a small team doing high-volume AI visual and production work for fashion and brand clients, is building a "one brain" system that makes company knowledge and client context searchable through natural-language agents. The starter stack described is OpenClaw in a sandbox, a GitHub repository, Obsidian on the local machine, and Telegram as the access point. The system's research structure had roughly 1,200 files at the time of the talk.

  18. Tessl BlogOfficialAI score42

    Agent Skills Should Be Treated as Supply Chain Components

    AITessl's talk at AI Native DevCon London argues that agent skills, which can be markdown files with instructions and bundled material, act as supply chain components that can shape agent behavior. The author says reading SKILL.md once is insufficient because risks can sit in supporting files, updates, and workspace trust settings. He identifies the danger as the combination of private context, untrusted content, and external communication, and cites research scanning roughly 4,000 public skills for issues including malware-like behavior.

  19. elvisXAI score42

    Voyager: an open harness for creative AI work across video and games

    AIElvis Saravia argues that creative work needs domain-specific agent harnesses rather than coding-oriented ones, and he highlights Voyager as an open harness for video, graphics, and games. According to the quoted post, Voyager lets agents work with local files and drive apps such as Blender, DaVinci Resolve, and Unity, and it is designed to work with models like Opus, Astra, and DeepSeek.

    Video from @omarsar0's post
  20. Google ResearchOfficialAI score22

    Google Research livestreams EmbeddingGemma 2 demo at COLM 2026 today

    AIGoogle Research is hosting a live demonstration of EmbeddingGemma 2 at its COLM booth #107 today at 12:00pm. The open multimodal model unifies text, image, audio, and video representations, with Sahil Dua available to connect with attendees.

    Video from @GoogleResearch's post
  21. Tessl BlogOfficialAI score38

    Mozilla.ai's cq Aims to Give Agents a Shared, Reviewable Knowledge Commons

    AIMozilla.ai's cq project proposes a shared knowledge layer where AI agents capture lessons from non-obvious fixes as structured knowledge units that other agents can later query. The default setup is local-first, using a local SQLite database so nothing leaves the machine, with an option to connect to a remote team server that adds review.

  22. Tessl BlogOfficialAI score52

    Cisco engineer argues agent skills need a context pipeline with evals

    AIJohn Groetzinger, writing in a personal capacity rather than for Cisco, argues that enterprise skills need packaging, evaluation, syncing, and distribution rather than scattered markdown files. He describes using skills to make cheaper models viable, converting curated TAC knowledge-base articles into maintained skills, and rolling out an eval framework across teams. He also describes syncing a repository README to Confluence with a deterministic script.

  23. LiveKitOfficialAI score22

    LiveKit Simulations lets teams test voice agents before customers do

    AILiveKit is offering free access to its Simulations product through October, letting teams check what their agent can do and find gaps before deployment. The product also lets teams test any model against their own scenarios before switching models.

    Video from @livekit's post
  24. 🚨 AI News | TestingCatalogXAI score50

    OpenAI rolls out GPT-6.1 Sol Ultrafast at 8x standard speed

    AIOpenAI is rolling out GPT-6.1 Sol Ultrafast on ChatGPT Work, Codex, and the API. The Ultrafast mode is priced at $12 per million input tokens and $60 per million output tokens, and it runs 8x faster than Sol Standard.

    Video from @testingcatalog's post
  25. Vaibhav (VB) SrivastavOfficialAI score46

    GPT-6.1 Sol Ultrafast rolls out with up to 8x faster token generation

    AIOpenAI is rolling out GPT-6.1 Sol Ultrafast, which generates tokens up to 8x faster than Sol Standard. On the API it is priced at $12 per 1M input tokens and $60 per 1M output tokens. The mode is also available today in Codex and ChatGPT Work.

  26. Artificial AnalysisOfficialAI score34

    Harvey LAB-AA: Artificial Analysis benchmark for legal AI agents

    AIArtificial Analysis has released Harvey LAB-AA, an evaluation built on Harvey's LAB dataset and developed in collaboration with Harvey. Full results are published on the Artificial Analysis evaluations page, alongside Harvey's commentary on the benchmark and human expert preferences.