Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Jun 25

Jun 25Thu

Jun 24

Jun 24Wed

Jun 23

Jun 23Tue

Jun 20

Jun 20Sat

Jun 19

Jun 19Fri
  1. Andrew NgAI score72

    Andrew Ng says Anthropic and U.S. export controls on Fable expose AI access risks

    AIAndrew Ng argues that Anthropic's restrictions on building competing LLMs and a U.S. Commerce Department license requirement for foreign nationals led Anthropic to disable Fable access worldwide. He says this shows governments and providers can quickly cut off access to frontier AI, which may push nations and businesses toward sovereignty efforts and open-source alternatives, though training frontier models remains difficult.

    Image from @AndrewYNg's post

Jun 18

Jun 18Thu

Jun 17

Jun 17Wed
  1. John SchulmanAI score40

    PPO's LLM-era revival and the unexpected reasons behind it

    AIJohn Schulman says PPO gained a second wave in the LLM era for reasons not anticipated in the original paper. He points to the importance-ratio objective, which corrects biases from numeric error, asynchronous training, and forward-pass noise, and to the clipping objective, whose effect on entropy was unknown at publication, citing DAPO's arXiv paper.

Jun 16

Jun 16Tue
  1. Mckay WrigleyAI score15

    Mckay Wrigley congratulates Cursor team on three-year milestone and SpaceX-xAI compute

    AIMckay Wrigley congratulated the Cursor team on more than three years of work and said he is excited to see what they build with compute from SpaceX and xAI. He noted he still keeps the original open-source Cursor repo on his laptop. Background posts from him describe Cursor as his full-time IDE, citing codebase search as a major productivity gain.

  2. HyperdimensionalAI score63

    Dean Ball argues the Anthropic Fable dispute shows frontier AI needs a governance framework

    AIDean Ball analyzes the Trump Administration's export controls on Anthropic's Fable and Mythos models after a jailbreak and a refused de-deployment request. He argues that the episode shows the need for a technocratic framework that separates political judgments about fairness from technical judgments about threats, in place of ad hoc executive action.

  3. Arthur MenschAI score13

    Mistral's Mensch says AI could decide abundance or extractive power

    AIMistral CEO Arthur Mensch compares AI to 20th-century oil, saying it will become a major source of global leverage and power. He argues the coming years could produce either broad wealth and abundance or the most extractive economies ever seen, and says Mistral is working toward the former by advancing AI research and accelerating its global diffusion.

  4. Arthur MenschAI score44

    Mistral says its upcoming models will all be open-weight

    AIMistral states that this model and upcoming ones will be open-weight. The company argues that open weights are critical for customer confidence and for research and developer communities. It contends that systems reachable only through someone else's interface cannot be owned, inspected, audited, or improved, especially if data recording can no longer be turned off.

Jun 15

Jun 15Mon

Jun 13

Jun 13Sat

Jun 12

Jun 12Fri

Jun 11

Jun 11Thu

Jun 10

Jun 10Wed
  1. AI Snake OilAI score70

    Why AI hasn't replaced software engineers, and why it likely won't

    AIThe essay argues that AI compresses the execution layer of software work while decision-making and accountability remain human, so AI is not yet replacing software engineers. It cites AI-attributed layoffs at Block, Snap, and Intuit that the authors say were not driven by AI, and WARN Act filings in which only one company checked an AI box. A Federal Reserve analysis is cited as finding software engineer employment growing about 3 percentage points per year more slowly after ChatGPT than a no-AI counterfactual.

Jun 9

Jun 9Tue
  1. Andrej KarpathyAI score65

    Karpathy Calls Claude Fable 5 a Major Step Forward for Long Tasks

    AIAndrej Karpathy says Claude Fable 5 is the same underlying model as Mythos with added safeguards, and that it leads on nearly all benchmarks. He describes it as a step change, especially for long, difficult problem-solving sessions where it handles more ambitious tasks without close supervision. He notes that its safeguards are set a bit too aggressively at launch and may be tuned over time.

  2. One Useful Thing (Ethan Mollick)AI score72

    Ethan Mollick tests Claude 5 Fable and finds it runs long projects with little user input

    AIEthan Mollick, who had early access to Claude 5 Fable, reports that it outperformed other public models in his tests, including an isochrone travel-time map and a nine-and-a-half-hour software build called Concord. He says the model delegated work to other agents and made many design choices he could not see or weigh in on, leaving him closer to a client than a hands-on operator. He also notes high token usage, frequent fallback to Claude 4.8 Opus under security guardrails, and persistent quirks in its writing style.

Jun 8

Jun 8Mon

Jun 5

Jun 5Fri

Jun 4

Jun 4Thu
  1. One Useful Thing (Ethan Mollick)AI score44

    Ethan Mollick Announces Co-Existence, a Sequel Book on Working Alongside AI

    AIEthan Mollick is releasing Co-Existence on October 20, a new book about working with AI systems that are sometimes, but not always, better than humans. The book follows his 2024 title Co-Intelligence, which he says was written about an era of chatbots rather than autonomous agents. Mollick also reports writing every chapter draft himself while using AI readers and fact-checkers, and building the book's website with Claude Code using Opus 4.8.

Jun 3

Jun 3Wed
  1. Mark ChenAI score25

    Mark Chen says OpenAI's models could match Mythos on cyber vulnerabilities

    AIOpenAI's Mark Chen said that after Mythos showed AI models can prove 80-year-old theorems, he expected them to also find cyber vulnerabilities, and they did. He added that researchers in math may now be thinking the same idea in reverse, applying cybersecurity-style capability to mathematics. The post offers no specific models, benchmarks, or figures.