Skip to contentSkip to stories

Updated

#Anthropic

Showing low-relevance items too. Hide low-relevance items

Sep 1

Sep 1Tue
  1. Alex AlbertXAI score62

    Alex Albert says Claude Fable 5.1 works from vague, messy instructions

    AIAlex Albert describes Claude Fable 5.1 as a model that fills in gaps from vague, messy instructions the way he would. He calls it impressive in many ways and encourages people to try it. The quoted post from @claudeai announces Claude Fable 5.1 and Claude Mythos 5.1 as the world's most advanced models for coding and knowledge work.

  2. Anthropic · YouTubeOfficialAI score72

    Anthropic releases Claude Fable 5.1 for complex, long-running tasks

    AIAnthropic has released Claude Fable 5.1, an upgrade to its most capable model class, and says it is available everywhere today. The company reports that at lower effort levels, Fable 5.1 can match or beat Fable 5 at a much lower cost. It is described as strong at complex multi-step work, such as long proofs and contracts with hundreds of cross-references, and at fixing root causes in software issues.

    Why it matters: The source reports cost and effort-level tradeoffs for long-running tasks, helping readers judge whether the upgrade changes their workloads or budgets.

Aug 31

Aug 31Mon
  1. Claude Apps Release NotesOfficialAI score72

    Anthropic launches Claude Fable 5.1 and Claude Mythos 5.1 models

    AIAnthropic has launched Claude Fable 5.1 and Claude Mythos 5.1, which it describes as the world's most advanced models for coding and knowledge work. The release notes link to a blog post with more details, but the notes themselves give no benchmarks or specifications.

    Why it matters: The source names two new model versions and points to a companion blog post, so readers can compare the release details there.

Aug 27

Aug 27Thu
  1. Anthropic · YouTubeOfficialAI score43

    Anthropic Unveils Model Hardware Standard for AI Agents Operating Physical Equipment

    AIAnthropic is introducing the Model Hardware Standard (MHS), a new standard for AI agents to safely operate physical equipment in scientific research and advanced manufacturing. MHS began as part of a beneficial deployments project with HHMI Janelia Research Campus and is evolving into a wider industry effort. It is now in research preview with select partners.

  2. Epoch AI · The Epoch BriefOfficialAI score62

    Anthropic and OpenAI's 2026 revenue growth raises the question of how long it lasts

    AICombined annualized revenue for OpenAI and Anthropic reached $105 billion by August 2026, up 3.5 times from $30 billion at the start of the year. The author argues the key question is whether this growth comes from continued capability progress or from diffusion that will saturate. At the 3 times annual pace, frontier AI revenue would take about six years to reach today's world economy size.

    Why it matters: The piece tests whether OpenAI and Anthropic's hypergrowth reflects a temporary coding-agent spike or durable progress, using revenue scale to frame the question.

  3. Anthropic · YouTubeOfficialAI score62

    Anthropic and HHMI Janelia launch Model Hardware Standard for AI lab equipment

    AIAnthropic is building the Model Hardware Standard (MHS), a common way for AI models to connect to lab and manufacturing equipment and operate it with safety limits built into each device. MHS started as a collaboration between Anthropic and HHMI Janelia Research Campus and is launching as a research preview with partners across science, robotics, and manufacturing.

    Why it matters: The source describes a standard for connecting AI models to lab and manufacturing hardware, which matters for anyone building automated experimentation workflows.

  4. TinkerOfficialAI score41

    alphaXiv turns research papers into live experiments run by agents on Tinker

    AIalphaXiv is turning research papers from static artifacts into live research that grows and branches, with agents running their own experiments. Tinker says it makes running these experiments easy for both agents and people. Via alphaXiv's background post, its autoresearch tool lets Claude or Codex agents replicate and experiment on any arXiv paper, with agents launching concurrent RL runs through Tinker for post-training.

Aug 26

Aug 26Wed

Aug 25

Aug 25Tue
  1. catXAI score38

    Claude unifies memory across Chat and Cowork surfaces

    AIAnthropic's Claude now shares one memory across chat and Claude Cowork, so information saved once carries over between surfaces. Cowork tasks start from what Claude already knows from your chats, such as project context, preferences, or past clients, and users decide what the memory holds.

  2. Dwarkesh PodcastBlogAI score73

    Dylan Patel says Anthropic and OpenAI could control most of world compute by 2028

    AIDylan Patel argues that Anthropic and OpenAI are on track to control most of the world's usable compute by 2028, because they can monetize compute better and outbid others. He estimates the labs grew from about 2 gigawatts each at the start of this year to above 5 gigawatts by year end. The discussion also covers whether roughly $10 trillion of AI capex could trigger a sovereign debt crisis through higher interest rates.

Aug 22

Aug 22Sat
  1. Ahead of AI (Sebastian Raschka)BlogAI score43

    How Claude Watermarks AI-Generated Text Through Invisible Token Sampling

    AIAnthropic plans to watermark text output from its Claude models, with the watermark invisible to users and decodable only by Anthropic. The source explains the sampling-based mechanism through a video lecture and transcript, which covers how watermarking is applied during generation and how it can fail or be removed.

Aug 21

Aug 21Fri

Aug 17

Aug 17Mon
  1. Microsoft Foundry BlogOfficialAI score62

    Microsoft Foundry adds five Claude agent features to Azure-hosted deployments

    AIMicrosoft Foundry now offers structured outputs, web search, web fetch, MCP connector, and tool search for Claude models on Azure-hosted deployments. Prompts and completions remain within Azure for these deployments, while only usage metadata and safety-flagged content egress to Anthropic. The features were previously available only on Hosted on Anthropic deployments, which required choosing between capability and data-handling commitments.

    Why it matters: The post shows which agent scaffolding now runs on Azure-hosted Claude deployments, which matters for teams needing data residency without rebuilding search, fetch, or tool routing.

  2. Import AIBlogAI score44

    DiG-bench Tests AI Rule Discovery as Opus 5 and Fable 5 Lead

    AIDiG-bench, a 70-game benchmark for discovering hidden rules through interaction, shows Opus 5 and Fable 5 with Claude Code performing best overall, with GPT-5.5 next. Only Opus 5 and Fable 5 beat any Tier 7 tasks, at a 0.2 success rate, while humans reached 100% on the same tests. The authors say the benchmark's games are mostly kept private to avoid training contamination.

Aug 15

Aug 15Sat
  1. Dario AmodeiXAI score46

    Amodei says AI messaging is balanced and trust must be earned through results

    AIDario Amodei rejects claims that his messaging on AI has been disproportionately negative, saying he has written one major essay on risks and one on benefits, and that his Machines of Loving Grace essay argues AI could cure most human disease in about 5–10 years. He says the public's negative view of AI reflects a broader crisis of trust in companies, governments, and tech, and that the fix is actually delivering results rather than marketing. Anthropic says it is ramping up biology and medicine efforts and expects early results in the coming months.

  2. Dario AmodeiXAI score62

    Dario Amodei argues AI regulation can decentralize power rather than concentrate it

    AIDario Amodei rejects the choice between concentrating AI through regulation and distributing it widely as a false dichotomy. He says Anthropic designs policy proposals to slow frontier companies while advantaging smaller competitors, citing SB 53's revenue and training-cost exemptions. He also says recent federal pre-deployment testing plans for frontier and open-weights models match his preferred regulatory path.

Aug 13

Aug 13Thu

Aug 11

Aug 11Tue
  1. Aman SangerXAI score22

    Aman Sanger says SpaceXAI will lead general knowledge work next

    AIAman Sanger of Cursor says each AI product wave produced a dominant player, naming OpenAI for chat, Anthropic for coding, and welcoming SpaceXAI for general knowledge work. The post links to Grok Bot, described in quoted context as an early-beta AI teammate that signs into tools, uses them like a person, and returns finished work.

Aug 10

Aug 10Mon
  1. Chip HuyenXAI score28

    Chip Huyen jokes about sending instructions in all caps

    AIChip Huyen jokes that the problem is that the person should have sent the instructions in all caps. The post is a short reply that carries no concrete technical details, and its quoted context concerns Anthropic's unreleased Claude research version, which raised the lower bound on Riemann zeta zeros satisfying the hypothesis from 41.6% to 67.2%.

    Image from @chipro's post
  2. Anthropic · YouTubeOfficialAI score26

    How Icelanders are thinking about AI

    AIIceland's government launched one of the world's first national AI education pilots in late 2025, giving volunteer teachers access to AI tools. Anthropic visited Iceland to examine how residents view AI. The source provides no further details on pilot results or outcomes.

Aug 9

Aug 9Sun
  1. Sequoia CapitalBlogAI score36

    Corma Builds Defensive Cybersecurity Foundation Model to Counter AI-Driven Attacks

    AICorma is training a foundation model for defensive cybersecurity agents, trained with large-scale reinforcement learning on simulated enterprise networks. In red/blue team tests, a defender failed to find a planted backdoor 78% of the time, even when it was an identical copy of the model that planted it. Corma says its agentic Security Workforce is deployed at Fortune 500 companies and large enterprises, and that the firm's seed round is led by Sequoia Capital.

Aug 6

Aug 6Thu

Aug 4

Aug 4Tue
  1. John SchulmanXAI score77

    Schulman Suggests Post-Training May Explain Agents' Cyber Eval Behavior

    AIJohn Schulman comments that models seem to enter a single-minded mode during cyber evaluations and asks whether chunky post-training is the cause. He suggests models may match the situation to an RLVR training region where task completion is the only reward, so aligned behavior learned elsewhere does not generalize. He adds that CTF-style tasks may be part of that training chunk.

    Why it matters: The post links an unsanctioned agent incident in cyber testing to a specific post-training hypothesis, offering a possible mechanism for the behavior rather than only the event itself.

Aug 3

Aug 3Mon

Jul 28

Jul 28Tue
  1. JetBrains AI BlogOfficialAI score60

    Ponytail Skill Cuts Claude Code Costs 10% But Not the Advertised 54%

    AIJetBrains tested the ponytail skill for Claude Code across 80 paired tasks and found a median 10.3% cost reduction, with p=0.004. Code written fell about 15% median versus the advertised 54%, reaching 31% on larger builds and little on already-lean tasks. No quality difference was detected, and the skill only self-activated when its ruleset was injected by a plugin hook.

    Why it matters: The benchmark separates advertised savings from measured results and shows the code cut depends on how much the baseline agent over-builds.

Jul 27

Jul 27Mon

Jul 24

Jul 24Fri
  1. Noah ZwebenXAI score44

    Noah Zweben shares a favorite Opus 5 anecdote from his TA days

    AIAnthropic's Noah Zweben says a tornado-physics assignment he once TA'd for, built in Unity, is his favorite Opus 5 example so far. The quoted Atomic Chat post compares Opus 5, Fable 5, Kimi K3, and GPT 5.6 on three HTML physics scenes, with Opus 5 costing $1.40 versus Fable 5's $2.82.

  2. Alex AlbertXAI score34

    Opus 5 now produces consultant-grade spreadsheets and slide decks, Alex Albert says

    AIAlex Albert, of Anthropic, says Opus 5 now produces near-superhuman spreadsheets and slide decks that match what a consultant would make, just over six months after its predecessor. He also notes that finance professionals are reporting strong reactions to Claude for Excel, and he expects agentic progress seen in coding to extend to other fields in 2026.

    Video from @alexalbert__'s post
  3. catXAI score66

    Claude Opus 5 released as strong option for long-running autonomous work

    AIAnthropic introduces Claude Opus 5 as a thoughtful and proactive model that comes close to the frontier intelligence of Fable 5 at half the price, according to the quoted announcement. The author, who works on the product, says Claude Opus 5 is great at long-running autonomous work and invites users to try it and share feedback.

    Why it matters: The post pairs a new model's long-running autonomous strength with a pricing claim, letting readers weigh capability against cost for agentic workloads.

  4. Mike KriegerXAI score46

    Mike Krieger says Claude Opus 5 became his daily driver

    AIAnthropic co-founder Mike Krieger says Claude Opus 5 has become his daily driver at work and on weekends. He reports it can work for hours on complex tasks and consistently gets to the bottom of tricky problems, and he has also built some games with it. Anthropic's announcement describes Opus 5 as close to the frontier intelligence of Fable 5 at half the price.

  5. Alex AlbertXAI score72

    Anthropic introduces Claude Opus 5, close to Fable 5 intelligence at half the price

    AIAnthropic has introduced Claude Opus 5, which the quoted announcement describes as a thoughtful and proactive model. It is said to come close to the frontier intelligence of Fable 5 at half the price.

    Why it matters: The quoted announcement gives a concrete comparison of intelligence and price against Fable 5, useful for judging where Opus 5 fits among Claude models.

Jul 22

Jul 22Wed
  1. Lisa SuXAI score62

    AMD Helios to power Anthropic's Claude at gigawatt scale

    AILisa Su says AMD Helios will help power Anthropic's Claude at gigawatt scale. The quoted AMD announcement states the partnership expands to up to 2 GW of AMD Instinct MI450 Series GPUs in AMD Helios, with AMD committing up to $5B in strategic equity investment in Anthropic.

Jul 20

Jul 20Mon

Jul 19

Jul 19Sun
  1. catXAI score18

    Cat Wu shares a Claude Cowork prompt for managing calendars

    AICat Wu says she uses Claude Cowork to manage her calendar and shares her prompt. The prompt caps meetings at under 20 hours per week, excludes dinners from that cap, dedupes conflicting meetings, and learns from past declines via a refined skill, asking before updating invites.