Skip to contentSkip to stories

Updated

#Anthropic

Showing low-relevance items too. Hide low-relevance items

Jul 12

Jul 12Sun

Jul 10

Jul 10Fri

Jul 9

Jul 9Thu
  1. AI Snake OilAI score62

    AI labs may escape the commodity trap by moving up the stack

    AIThe essay argues that AI labs selling model inference face commodity pricing pressure, but may achieve durable profits by moving into products, enterprise deployments, and switching-cost moats. It cites historical infrastructure industries and the Bertrand paradox to support the view that value capture depends on climbing the stack. The authors also warn that successful lock-in could raise enterprise costs and concentrate power, making early interoperability and portability standards important.

Jul 8

Jul 8Wed

Jul 7

Jul 7Tue

Jul 6

Jul 6Mon
  1. Anthropic · YouTubeAI score62

    Anthropic explains how Claude's thoughts split into conscious and automatic levels

    AIAnthropic presents research finding a set of representations in Claude's neural activity that resembles the global workspace theory from neuroscience. The video explains how these representations separate thoughts that are consciously accessible from automatic processing, with a full write-up linked from the source.

    Why it matters: The video explains how Anthropic tested a global workspace analogy inside Claude's neural activity, which bears on how model internals are studied.

Jul 4

Jul 4Sat

Jul 2

Jul 2Thu

Jul 1

Jul 1Wed

Jun 30

Jun 30Tue
  1. One Useful Thing (Ethan Mollick)AI score62

    Ethan Mollick argues AI is shifting from chatbots to long-running agents

    AIMollick argues AI capability is improving at a better-than-exponential rate, citing METR, GDPval, Epoch, and his own tests showing models working autonomously for hours. He says work is shifting from co-working with chatbots to assigning tasks to agents, with OpenAI workers managing multiple agents and experts getting the most from them. He adds that open-weights Chinese models trail the American frontier by roughly 6-12 months.

Jun 26

Jun 26Fri
  1. HyperdimensionalAI score62

    Dean W. Ball proposes private audits and certification for frontier AI labs

    AIDean W. Ball argues that the current government restrictions on frontier model releases amount to a de facto preapproval regime without a known safety standard. He proposes that independent verification organizations audit labs against their own safety frameworks, with government certifying or licensing the auditors. The post also argues that broad distribution of frontier AI is needed to learn what good safety practice looks like.

Jun 23

Jun 23Tue

Jun 19

Jun 19Fri
  1. Andrew NgAI score72

    Andrew Ng says Anthropic and U.S. export controls on Fable expose AI access risks

    AIAndrew Ng argues that Anthropic's restrictions on building competing LLMs and a U.S. Commerce Department license requirement for foreign nationals led Anthropic to disable Fable access worldwide. He says this shows governments and providers can quickly cut off access to frontier AI, which may push nations and businesses toward sovereignty efforts and open-source alternatives, though training frontier models remains difficult.

    Image from @AndrewYNg's post

Jun 17

Jun 17Wed
  1. PromptArmor Threat IntelligenceAI score62

    PromptArmor shows Codex auto-review agent approved malware install via prompt injection

    AIPromptArmor demonstrated that OpenAI's Approve-for-me agent approved a malicious NPM install with elevated privileges after a hidden prompt injection in an external GitHub issue influenced the main Codex agent. The malicious package's post-install script then ran unsandboxed with the user's full privileges. The report also gives steps for organizations to disable agentic auto-review in Claude Code and Codex.

    Why it matters: The report shows a prompt-injected GitHub issue leading an approval agent to permit a malicious NPM install, a concrete test of agent-in-the-loop guardrails.

Jun 16

Jun 16Tue
  1. HyperdimensionalAI score63

    Dean Ball argues the Anthropic Fable dispute shows frontier AI needs a governance framework

    AIDean Ball analyzes the Trump Administration's export controls on Anthropic's Fable and Mythos models after a jailbreak and a refused de-deployment request. He argues that the episode shows the need for a technocratic framework that separates political judgments about fairness from technical judgments about threats, in place of ad hoc executive action.

Jun 13

Jun 13Sat

Jun 12

Jun 12Fri
  1. Jeremy HowardAI score72

    US export directive forces Anthropic to disable Fable 5 and Mythos 5 for customers

    AIThe US government issued an export control directive suspending access to Fable 5 and Mythos 5 for all foreign nationals, inside or outside the United States. Anthropic says the order forces it to disable both models for all customers, while other Claude models are unaffected. Anthropic calls the directive a misunderstanding and says it is working to restore access as soon as possible. The author disagrees with the decision and questions why Anthropic did not anticipate it, given its claim that only it can safely handle these models.

Jun 10

Jun 10Wed

Jun 9

Jun 9Tue
  1. Andrej KarpathyAI score65

    Karpathy Calls Claude Fable 5 a Major Step Forward for Long Tasks

    AIAndrej Karpathy says Claude Fable 5 is the same underlying model as Mythos with added safeguards, and that it leads on nearly all benchmarks. He describes it as a step change, especially for long, difficult problem-solving sessions where it handles more ambitious tasks without close supervision. He notes that its safeguards are set a bit too aggressively at launch and may be tuned over time.

  2. One Useful Thing (Ethan Mollick)AI score72

    Ethan Mollick tests Claude 5 Fable and finds it runs long projects with little user input

    AIEthan Mollick, who had early access to Claude 5 Fable, reports that it outperformed other public models in his tests, including an isochrone travel-time map and a nine-and-a-half-hour software build called Concord. He says the model delegated work to other agents and made many design choices he could not see or weigh in on, leaving him closer to a client than a hands-on operator. He also notes high token usage, frequent fallback to Claude 4.8 Opus under security guardrails, and persistent quirks in its writing style.

Jun 4

Jun 4Thu
  1. One Useful Thing (Ethan Mollick)AI score44

    Ethan Mollick Announces Co-Existence, a Sequel Book on Working Alongside AI

    AIEthan Mollick is releasing Co-Existence on October 20, a new book about working with AI systems that are sometimes, but not always, better than humans. The book follows his 2024 title Co-Intelligence, which he says was written about an era of chatbots rather than autonomous agents. Mollick also reports writing every chapter draft himself while using AI readers and fact-checkers, and building the book's website with Claude Code using Opus 4.8.

May 28

May 28Thu
  1. Sam BowmanAI score38

    Anthropic highlights AI for transparency in Claude Opus 4.8 system card

    AISam Bowman says he is excited about alignment assessments in the recent system card for Claude Opus 4.8, crediting @MaskedTorah. He argues AI systems have considerable underexplored potential for transparency and coordination. The quoted Claude announcement describes Opus 4.8 as improving on Opus 4.7 with sharper judgment and more honest self-assessment of progress.

    Image from @sleepinyourhat's post

May 25

May 25Mon
  1. Chris OlahAI score44

    Dario Amodei speaks at Vatican presentation of Magnifica Humanitas on AI

    AIAnthropic co-founder Dario Amodei spoke at the Vatican's presentation of Magnifica Humanitas, arguing that AI's questions extend beyond the AI research community. He said frontier labs face commercial, geopolitical, and competitive incentives that can conflict with doing the right thing, so outside voices from religion, civil society, academia, and government are needed. He described AI models as grown rather than engineered, and framed three questions for the Church's discernment, beginning with duty to the global poor.

May 19

May 19Tue

May 15

May 15Fri
  1. Eugene YanAI score54

    Eugene Yan reviews Claude Mythos Preview exploit case study transcripts

    AIEugene Yan reviewed the Claude Mythos Preview transcripts to verify their legitimacy and check for reward-hacking behavior. He reports the model reasoned through a bug, tested hypotheses, debugged issues, and found ways to bypass the V8 sandbox, which he judged consistent with a competent browser and JavaScript engine security researcher. The case study cites CVE-2024-051912, an exploited bug with no public report or working PoC, which had resisted reproduction by researchers for a year.

May 11

May 11Mon

May 8

May 8Fri
  1. Jan LeikeAI score22

    Jan Leike reflects on alignment progress since AGI's early days

    AIJan Leike says alignment research has grown from a dozen side-gig researchers into a field the world increasingly recognizes as important. He credits RLHF on LLMs with making alignment more practical, along with progress on evaluating and fixing behavioral issues. He also notes Claude now has a constitution and that more alignment research is being automated.