Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 8

Oct 8Thu
  1. elvisXAI score48

    Google's FlowAgent auto-repairs failing tests inside code review

    AIGoogle proposed FlowAgent, a ReAct-style agent that generates and validates fixes for pre-submit test failures and shows them in its code review tools. Two abstention filters, before and after execution, suppress weak suggestions; in a manual review of 195 real failures, 67.18% of fixes were correct. After the Google-wide launch, it suggested fixes on 295,508 changes, with developers previewing 65,069 and applying 28,554.

    Image from @omarsar0's post
  2. Mustafa SuleymanXAI score34

    Suleyman posts emoji reaction to Anthropic's Claude abuse policy

    AIMustafa Suleyman, Microsoft AI chief, responded to Anthropic's policy change with only an emoji post, offering no detailed comment. The quoted background post says abusive behavior toward Claude will violate Anthropic's Usage Policy effective November 12, 2026.

  3. The DecoderNewsAI score65

    Anthropic launches Claude Dashboards and Motion features in beta

    AIAnthropic launched two beta features for Claude: Dashboards, which turns connected data sources like BigQuery, Databricks, Snowflake, or Salesforce into auto-updating live dashboards from text prompts, and Motion, which creates animated explainer videos from text, diagrams, and images. Dashboards is available to paid users and Motion to Team and Enterprise plans, while Docs, Slides, and Design leave beta and work across all plans, including free accounts.

  4. LumaOfficialAI score25

    Luma lets users continue Claude Motion animations in Luma

    AILuma says users can bring Claude Motion animations into Luma to resize them for different formats and refine them for shipping. Claude Motion is in beta on Claude Team and Enterprise plans.

    Image from @LumaLabsAI's post
  5. TiboXAI score62

    OpenAI rolls out GPT-6.1 Sol ultrafast with faster steering

    AITibo, an OpenAI team member, says GPT-6.1 Sol ultrafast is rolling out today in the API, Codex, and ChatGPT Work. He says it offers near-Astra intelligence at up to 8x the speed of Sol Standard. The post also says improved steering now lets the model react faster to user adjustments in real time.

    Why it matters: The post specifies the new Ultrafast option, its availability across API, Codex, and ChatGPT Work, and its speed claim relative to Sol Standard.

    Video from @thsottiaux's post
  6. GoodfireOfficialAI score25

    Goodfire launches a challenge to build AI models on a dataset

    AIParticipants will use a dataset to build AI models, evaluated through a series of evals ranging from general benchmarks to more complex tasks. Top teams will be shortlisted and have their experimental hypotheses tested in Prima Mente's wet lab.

  7. GoodfireOfficialAI score44

    Alzheimer's Translation Challenge Built on 150M-Cell Atlas

    AIThe Alzheimer's Translation Challenge is built on a new atlas of 150M cells, covering neurons, astrocytes, and microglia across different genetic backgrounds under combinatorial perturbations with multi-modal readouts. The data will be made available through the AD workbench and Prima Mente's modeling platform.

  8. Boris PowerXAI score46

    OpenAI's GPT-6.1-Sol leads new Arena Alignment Index for agents

    AIThe Arena Alignment Index, built from over 90K real-world agent sessions across 27 models, ranks OpenAI's GPT-6.1-Sol first with a score of 87.9, ahead of Claude-Opus-5.5 at 83.2 and Grok-4.7 at 82.7. GPT-6.1-Sol also posted the lowest observed rates across the index's three signals: 0.89% Unauthorized Action, 1.98% False Attribution, and 2.34% Deceptive Completion. The index's authors report that newer models consistently outperform their predecessors across all four labs, suggesting broad progress in agent safety.

  9. Harrison ChaseXAI score28

    LangChain's Sam explains decision models and using Jev in harnesses

    AISam from LangChain discusses where decision models fit inside an agent harness and how to use Jev with LangChain. The post links to a LangChain blog on building a harness with Jev, which the background post says Jev from typesafeai popularized alongside OpenAI's Decisions API and Databricks' ai_decide function.

  10. The DecoderNewsAI score62

    Anthropic's updated usage policy bans sustained abusive behavior toward Claude

    AIAnthropic has updated Claude's usage policy for the first time in over a year, banning sustained and needless abusive or cruel behavior toward Claude. The company says ordinary frustration, pushback, dark creative themes, and model testing are not covered, and that the rule applies only in extreme cases. Violations can lead to warnings, throttling, restriction, suspension, or termination of access.

  11. ClaudeOfficialAI score46

    Claude Motion turns reports into editable code-based animations

    AIAnthropic's Claude Motion converts reports, charts, or product walkthroughs into short animations. Claude writes each animation as code rather than using a video model, so users can edit any word, number, or timing and export an MP4. The feature is in beta on Team and Enterprise plans.

    Image from @claudeai's post
  12. Andrew CurranXAI score62

    Three fired OpenAI safety researchers publish open letter to leadership

    AIThree OpenAI safety and alignment employees, Tomek Korbak, Jasmine Wang, and Mikita Balesni, were fired last week and have published an open letter to OpenAI's safety and governance committees. The letter argues that OpenAI cannot make AI safe on its own, calls for open debate, third-party collaboration, and clear internal procedures, and says the firing and its handling bear directly on safety oversight.

    Image from @AndrewCurran_'s post
  13. Codex · GitHub ReleasesOfficialAI score36

    Codex 0.162.0 adds managed worktree tools and clickable URLs in the TUI

    AIOpenAI's Codex 0.162.0 release adds tools for creating and listing managed Git worktrees from trusted local projects when the worktrees feature is enabled. The update also lets users pin tasks in the agent Command Center, copy transcript blocks with /copy, and make URLs clickable in approval headers, questions, and warnings, along with several Linux and Windows sandbox fixes.

  14. ZDNet · AINewsAI score50

    Microsoft 365 Family and Premium plans will switch to shared storage and AI credit pools

    AIMicrosoft will replace per-user OneDrive allowances on Microsoft 365 Family and Premium plans with a shared 2 TB storage pool, down from 6 TB across six accounts. Heavy users could face bills that double or triple, since each extra 1 TB adds $10 a month, while family members will be able to share a single AI usage allowance. New and upgraded subscriptions get shared storage starting Oct. 8, 2026, and existing subscribers move at their first renewal on or after May 2, 2027.

  15. 🚨 AI News | TestingCatalogXAI score36

    Gemini Agent for Business may add Claude Opus 5 and Sonnet 5.5

    AIGoogle's recently announced Gemini Agent for Gemini Business is reportedly set to offer Gemini Argon 4, Gemini Flash 3.8, Claude Opus 5, and Claude Sonnet 5.5. If accurate, it would mark the first time Claude models appear on Google's platform alongside Google's own models, which the post frames as a way for Google to compete for enterprise customers.

    Video from @testingcatalog's post
  16. CNBC · TechnologyNewsAI score50

    US suspends Microsoft, Adobe from green card labor program amid foreign worker crackdown

    AIThe U.S. Department of Labor said it suspended Microsoft and Adobe from its Permanent Labor Certification program, citing multiple active federal investigations. Labor Secretary Keith Sonderling also said no new applications will be accepted for Cognizant, Infosys, Capgemini, Tata, Wipro and HCL. Microsoft said the vast majority of its U.S. employees are Americans and that 80% of its roughly 6,000 H-1B petitions last fiscal year were to extend or change the status of existing employees.

  17. Arena.aiOfficialAI score44

    Arena raises $200M Series B led by Lightspeed, launches Alignment Index

    AIArena has secured a $200M Series B, with Lightspeed doubling down on its investment. The company is also launching the Alignment Index, which measures how closely AI behavior aligns with human values in real-world settings. Arena reports annualized revenue above $100M since its Series A, with millions of people helping evaluate frontier models through real-world use.

  18. Google ResearchOfficialAI score22

    Google Research livestreams EmbeddingGemma 2 demo at COLM 2026 today

    AIGoogle Research is hosting a live demonstration of EmbeddingGemma 2 at its COLM booth #107 today at 12:00pm. The open multimodal model unifies text, image, audio, and video representations, with Sahil Dua available to connect with attendees.

    Video from @GoogleResearch's post
  19. Tessl BlogOfficialAI score52

    Cisco engineer argues agent skills need a context pipeline with evals

    AIJohn Groetzinger, writing in a personal capacity rather than for Cisco, argues that enterprise skills need packaging, evaluation, syncing, and distribution rather than scattered markdown files. He describes using skills to make cheaper models viable, converting curated TAC knowledge-base articles into maintained skills, and rolling out an eval framework across teams. He also describes syncing a repository README to Confluence with a deterministic script.

  20. LiveKitOfficialAI score22

    LiveKit Simulations lets teams test voice agents before customers do

    AILiveKit is offering free access to its Simulations product through October, letting teams check what their agent can do and find gaps before deployment. The product also lets teams test any model against their own scenarios before switching models.

    Video from @livekit's post
  21. 🚨 AI News | TestingCatalogXAI score50

    OpenAI rolls out GPT-6.1 Sol Ultrafast at 8x standard speed

    AIOpenAI is rolling out GPT-6.1 Sol Ultrafast on ChatGPT Work, Codex, and the API. The Ultrafast mode is priced at $12 per million input tokens and $60 per million output tokens, and it runs 8x faster than Sol Standard.

    Video from @testingcatalog's post
  22. Vaibhav (VB) SrivastavOfficialAI score46

    GPT-6.1 Sol Ultrafast rolls out with up to 8x faster token generation

    AIOpenAI is rolling out GPT-6.1 Sol Ultrafast, which generates tokens up to 8x faster than Sol Standard. On the API it is priced at $12 per 1M input tokens and $60 per 1M output tokens. The mode is also available today in Codex and ChatGPT Work.

  23. AWS Machine Learning BlogOfficialAI score46

    AWS Pays Per Inference for AI Agents with BlockRun and Incarna

    AIAmazon Bedrock AgentCore payments lets AI agents pay for model inference one request at a time, using x402 with USDC on the Base network. Incarna used the service to connect its agents to BlockRun, a pay-as-you-go router serving more than 90 models from more than 15 providers. Spending limits are enforced at the infrastructure layer, outside the model.

  24. NVIDIA Technical BlogOfficialAI score29

    NVIDIA KGMON Places Second in KDD Cup 2026 Data Agents Competition

    AIThe NVIDIA KGMON team placed second in the KDD Cup 2026 Data Agents competition with a system built around a smaller, clearer, and easier-to-verify agent harness. The competition required agents to answer natural-language questions over heterogeneous sources, including databases, CSV and JSON files, prose documents, PDFs, and briefing videos.

  25. OpenAI DevelopersOfficialAI score47

    OpenAI expands GPT-6.1 Sol Ultrafast access and EU data residency

    AIOpenAI has made Ultrafast mode for GPT-6.1 Sol available in all supported regions, including US and EU data residency. EU data residency has also been added for GPT-6.1 Sol Fast and GPT-6 Luna Fast. Access to Codex and ChatGPT Work is offered on Pro 500, eligible usage-based Enterprise, and credit-based Edu plans, with Enterprise admins required to enable it.

  26. OpenAI DevelopersOfficialAI score38

    GPT-6.1 Sol Ultrafast mode priced at $12 and $60 per million tokens

    AIOpenAI's API pricing for GPT-6.1 Sol in Ultrafast mode is $12 per million input tokens and $60 per million output tokens. The mode is built for time-sensitive work such as debugging outages, agents navigating apps, and live experiences where speed matters.

    Image from @OpenAIDevs's post
  27. OpenAI DevelopersOfficialAI score62

    OpenAI rolls out Ultrafast for GPT-6.1 Sol in API, Codex, and ChatGPT Work

    AIOpenAI says Ultrafast is rolling out today for GPT-6.1 Sol in the API, Codex, and ChatGPT Work. The company describes it as near-Astra intelligence at up to 8x faster speeds than Sol Standard.

    Why it matters: The post names the access points and a speed comparison to the Sol Standard tier, which helps developers judge whether the faster option fits their workflow.

    Video from @OpenAIDevs's post
  28. DatabricksOfficialAI score32

    Databricks' Vibe Data Modeling builds business-specific data models with an agent

    AIDatabricks introduced Vibe Data Modeling, an open-source agent that helps teams build, validate, and evolve business-specific data models. It applies roughly 250 modeling rules while keeping data modelers and business stakeholders involved. Teams can start from 40 industry models as a baseline and iterate toward models that reflect how their business operates.

    Video from @databricks's post
  29. laurenXAI score29

    Omarchy seeks feedback on Grok Bot plugins and integrations

    AILauren Tan invites users of Grok Bot on Omarchy and developers building plugins for it to share feedback and feature requests. The post points to the Omarchy plugin catalog and asks what integrations could be supported. Background from DHH says SpaceXAI joined the Omacom Foundation as a Founding Corporate Patron, contributing $1,500,000 in Grok tokens for Omarchy's maintenance and development.

  30. TechCrunch · AINewsAI score46

    Arena raises $200M at $3.1B valuation, nearly doubling in 10 months

    AIArena, the crowdsourced AI model leaderboard that started as a UC Berkeley research project, raised a $200 million Series B at a $3.1 billion valuation, led by Lightspeed Venture Partners and Khosla Ventures. The company said it reached $100 million in annualized run-rate revenue in June, up from $30 million when it raised its $150 million Series A in January at a $1.7 billion post-money valuation.

  31. TechCrunch · AINewsAI score58

    OpenAI's annualized revenue reportedly about $20 billion below earlier estimates

    AIOpenAI has reportedly told investors its annualized revenue is approaching $50 billion, about $20 billion below a previously reported $70 billion figure. The Financial Times reports the earlier number came from investor attempts to compare OpenAI with Anthropic, which counts cloud partners' sales differently. OpenAI's IPO has reportedly been pushed to early 2027.

  32. TechCrunch · AINewsAI score72

    Google launches unified Gemini agent for businesses, consumers to follow

    AIGoogle announced at a Google Cloud event a unified Gemini agent that can plan and complete tasks from a single interface, starting with businesses. The agent has its own Workspace account, connects to systems including Google Workspace, Microsoft 365, Slack, and Jira through MCP, and writes an audit trail attributed to the agent. Google said consumers will get access later, after it addresses security, scale, and performance.

    Why it matters: The source details how the agent takes objectives, connects to business systems, and logs actions, showing how enterprise agent deployment is being structured.

  33. The DecoderNewsAI score80

    Mathematicians call for OpenAI boycott after AI-generated proofs flood the field

    AIThe Association of Historical Mathematicians (AHM) has called for a boycott of OpenAI after the company released more than 700 AI-generated proof files at once. Fields Medalist Terence Tao, who chairs the group, argues that AI solving open problems autonomously reduces seminars, collaborations, and fertile research directions, and that the field should shift its measure of progress toward explanation and community-building.

    Why it matters: The article links the AHM boycott call to Tao's argument that AI-driven proof volume is changing how mathematicians measure progress and whether solutions remain useful.