Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Sep 12

Sep 12Sat
  1. Jakub PachockiAI score62

    Dario Amodei essay calls for AI industry to pace the frontier

    AIDario Amodei has written an essay arguing that the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first step by giving third-party evaluators permanent, employee-level access to its systems. The evaluators can verify adherence to safety measures, report incidents, and assess model alignment during training.

  2. Dario AmodeiAI score59

    Dario Amodei Calls for AI Industry to Slow Down and Pace the Frontier

    AIDario Amodei announced a new essay arguing the AI industry should slow down, with a three-part plan for doing so. Anthropic is unilaterally committing to the first step by giving third-party evaluators permanent, employee-level access to its systems to verify safety measures, report incidents, and assess model alignment during training.

Sep 11

Sep 11Fri
  1. Thinking MachinesAI score42

    John Schulman on where human judgment still matters as AI self-improves

    AIThinking Machines shared a Dwarkesh Patel podcast episode with John Schulman discussing where human judgment remains essential as models improve and self-improve. Schulman highlights teaching models to handle messy real-world tasks, applying taste to what works in the long run, and specifying what people actually want. The episode also covers recursive self-improvement, long-horizon RL, and the sim-to-real gap.

  2. Dwarkesh PatelAI score18

    Dwarkesh Patel on why Sonnet 5 and Opus 5 trail GLM 5.3

    AIDwarkesh Patel said a discussion with John, Beren, and Charlie questioned why Sonnet 5 and Opus 5 feel weaker than GLM 5.3, even though Anthropic could use raw logit distillation from Fable and train on Fable's environments. The discussion raised questions about the value of distillation, what makes it effective, and which model behaviors are hard to extract through it.

    Video from @dwarkesh_sp's post
  3. Mckay WrigleyAI score18

    Mckay Wrigley mocks mathematicians' concerns over AI curing cancer

    AIMckay Wrigley dismissed concerns that AI curing cancer could disrupt mathematicians' "process of understanding" and raise attribution questions, calling the objection one of the dumbest things he has read. He argued that building AI to solve humanity's greatest problems would be a miraculous achievement. The reply responds to a quoted post noting that 25 Fields Medal winners issued a joint declaration warning of severe misalignment between AI companies and the mathematics community.

  4. Dwarkesh PatelAI score42

    Dwarkesh Patel releases podcast with AI researchers on frontier progress

    AIDwarkesh Patel announced a new episode featuring John Schulman, Chris O'Neill, and Beren Millidge, three AI researchers from openish companies. The discussion covers the case against recursive self-improvement, drivers of Chinese labs' progress, training of automated AI researchers, long-horizon RL, the sim-to-real gap, and the role of data and RL in recent progress.

    Video from @dwarkesh_sp's post
  5. Interconnects (Nathan Lambert)AI score38

    Open-Source AI & Open Models Reading List Is Updated for Research and Policy Writing

    AINathan Lambert has compiled a reading list of open-model writing covering why labs release open weights, the open-versus-closed debate, and US-China competition, last updated 15 September 2026. The list includes pieces on open-model economics, safety and marginal-risk research, and recent Chinese releases such as Kimi K3 and GLM-5.2. It also cites lawmaker inquiries into Western companies' use of Chinese models.

Sep 10

Sep 10Thu
  1. PlatformerAI score57

    Anthropic and OpenAI researchers' superintelligence warnings reshape AI safety debate

    AIA former Anthropic researcher's resignation post and a senior Anthropic alignment leader's comments that AI could kill all humans drew wide attention. The column argues public and congressional concern about superintelligence risk is growing, citing the Ban Artificial Superintelligence Act and a Senate probe into an OpenAI-related incident.

  2. Sebastian RaschkaAI score62

    Raschka reviews DeepSeek V4.1-Flash's encoder-decoder architecture overhaul

    AISebastian Raschka says DeepSeek V4.1 contains a major architecture overhaul using an encoder-decoder setup, and he argues it could have been named V5. The attached diagrams compare DeepSeek V4-Flash (284B) with DeepSeek V4.1-Flash (552B), which has 1M supported context and a 10-layer encoder. The attached charts report a global KV cache per token of 890 bytes for V4.1-Flash, versus 3,514 for V4-Flash and 48,068 for DeepSeek-V3.2.

    Image from @rasbt's post
  3. Redwood Research BlogAI score52

    Redwood Research proposes tracking how architecture affects AI monitorability

    AIRedwood Research argues that AI companies should regularly report whether their architectures allow latent reasoning or latent communication between agents, and that such reporting should be externally verified. It proposes opaque serial depth as a minimally invasive proxy, with third-party evaluators reviewing near-frontier models, including internal R&D prototypes. The post also calls for published monitorability policies and stress tests on chain-of-thought monitoring.

  4. John SchulmanAI score40

    Schulman says user data gains in math are unlikely; disclosure norms needed

    AIJohn Schulman argues that training on user data contributes little to frontier math gains, which come mainly from scaling pretraining and RLVR. He says user data is more likely used to find failure modes that hired annotators struggle to recreate. He calls for stronger norms on disclosing how companies train on user data, including the methods and capabilities targeted.

  5. Interconnects (Nathan Lambert)AI score55

    Nathan Lambert on how one AI safety resignation went viral and why he doubts fast takeoff

    AINathan Lambert argues that a resignation post by AI researcher Jacob Coxon spread widely because public fear of AI extinction risk had been building. He says concrete risks such as cyber attacks and bio-risks deserve debate, while he assigns extinction risk a probability too low to discuss and expects recursive self-improvement to produce only lossy, jagged gains rather than a rapid takeoff.

  6. The Algorithmic BridgeAI score27

    Jacob Coxon's viral resignation tweet warns AI companies are gambling with lives

    AIFormer OpenAI and Anthropic employee Jacob Coxon resigned and posted a viral tweet, which has gathered over 700k likes and 140 million views, accusing AI companies of gambling with our lives. Coxon said people building AI earnestly believe it could kill us all by the end of the decade. The article argues that more insiders may leave, leaving the industry's remaining staff to accelerate development.

Sep 9

Sep 9Wed