Skip to contentSkip to stories

Updated

#Anthropic

Oct 7

Oct 7Wed
  1. Epoch AIAI score67

    Epoch tests six AI models on real Epoch work and finds they cannot yet fully automate it

    AIEpoch gave six models 11 real work tasks from its own operations, including graphic design, data insights, and research design, and graded outputs against employee standards. Fable 5.1 and GPT-6 Astra led on average task performance, reliably handling well-defined work such as coding and computational analysis. The report finds that all models still fail on open-ended judgment, including matching Epoch's standards, designing informative experiments, and generating diverse ideas, so the authors conclude AI cannot yet replace workers at Epoch.

    Why it matters: The report separates well-defined task reliability from open-ended judgment failures, which benchmark scores on easily verifiable tasks would miss.

  2. IThome · AIAI score72

    Anthropic releases Claude Haiku 5.5, cutting run costs about 75% from Haiku 4.5

    AIAnthropic released Claude Haiku 5.5, which it calls the fastest, cheapest, and most capable Haiku model so far. On average it costs about 75% less to run than Haiku 4.5, with input at $0.10 and output at $0.50 per million tokens for requests up to 100,000 tokens. Anthropic also cut Sonnet 5.5's cache read price from $0.20 to $0.10 per million tokens, which it says lowers run costs by about 20% on many agent tasks.

  3. DatabricksAI score36

    Claude Haiku 5.5 launches on Databricks as a Day 0 release

    AIAnthropic's Claude Haiku 5.5 is available on Databricks from day zero, which Databricks calls its cheapest, fastest, and most capable small model. On Databricks' OfficeQA Pro V1 benchmark, it delivers about 15% higher quality than Haiku 4.5 at a fraction of the cost. Users can run it alongside 60+ other models on data already in Databricks, with Unity Gateway handling governance, monitoring, and security.

  4. Ars Technica · AIAI score46

    Artcraft releases open source clones of Adobe Photoshop, Premiere and other apps built with Claude

    AIDeveloper Brandon Thomas's Artcraft has launched seven open source apps in Rust that aim to replicate the interfaces and tools of Adobe Photoshop, Illustrator, Premiere, Lightroom, After Effects, InDesign, and Acrobat Pro. Thomas said he used Anthropic's Claude Opus 5.5 to build the clean-room replacements, with WebAssembly versions available for browser use. The apps remain in a "super early alpha" state, and commenters have pointed out many current shortcomings.

  5. TechRadar · AIAI score42

    Trump creates Super Intelligence Force and renames AI to "SI" in federal communications

    AIPresident Trump announced a White House-led "Super Intelligence Force" that will spend 120 days examining AI risks and federal responses, and signed an executive order directing agencies to use "Super Intelligence" and "SI" instead of "Artificial Intelligence" and "AI." The order asks officials to develop a possible new federal definition within 60 days, but the source says the change is linguistic rather than architectural. Critics quoted in the article argue that renaming does not change the technology itself.

  6. Simon WillisonAI score62

    Anthropic releases Claude Haiku 5.5, priced like GPT-6 Luna up to 100,000 tokens

    AIAnthropic has released Claude Haiku 5.5, priced at $0.10 input and $0.50 output per million tokens up to 100,000 tokens, matching GPT-6 Luna. Beyond 100,000 tokens the price rises to $0.50 and $2.50, and the author found the new tokenizer uses about 1.25x as many tokens as Haiku 4.5 on the same long prompt. The model cannot disable reasoning and defaults to medium effort.

  7. MarkTechPostAI score67

    Anthropic releases Claude Haiku 5.5, a small model with 1M context

    AIAnthropic has released Claude Haiku 5.5, its cheapest and fastest small model, priced at $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100K tokens. It keeps a 1M token context window, up to 128K output tokens, and is generally available on the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Anthropic reports 72.4% on OSWorld 2.1 (offline subset) versus 15.7% for Haiku 4.5, and the article notes that non-default temperature, top_p or top_k values return a 400 error.

  8. GitHub Copilot ChangelogAI score38

    Claude Haiku 5.5 is now generally available in GitHub Copilot

    AIAnthropic's lightweight Claude Haiku 5.5 is now generally available in GitHub Copilot for fast, high-volume tasks such as subagents, quick edits, and terminal work. In early testing, it matched Claude Sonnet 5 on many coding tasks while using significantly fewer tokens and steps. The model is billed at provider list pricing under usage-based billing and is available to Copilot Pro, Pro+, Max, Business, and Enterprise users.

  9. KhazixAI score60

    Claude Max subscribers get monthly API credits usable across Claude models

    AISubscribers to Claude's Max plan can claim monthly API credits: $100 for the $100 tier and $200 for the $200 tier. The credits work for any Claude model and can be used in the user's own apps and other agents. The author argues that bundling monthly API credits alongside a broad model lineup will make it hard for other model companies to compete.

  10. AWS Machine Learning BlogAI score56

    Claude Haiku 5.5 becomes available on Amazon Bedrock and Claude Platform on AWS

    AIAnthropic's Claude Haiku 5.5 is now available on Amazon Bedrock and Claude Platform on AWS. According to Anthropic, it is the fastest and most efficient model in the Claude 5.5 family and costs around 75 percent less than Claude Haiku 4.5 for most tasks. The post also covers pairing it with Claude Opus 5.5 as a subagent layer and provides Boto3, Converse, and Anthropic SDK examples for calling the model.

  11. Hacker News · AI (150+ points)AI score52

    Meta and Microsoft cut employee use of Anthropic's Claude AI

    AIMeta and Microsoft have reduced their employees' use of Anthropic's Claude as they shift toward their own coding tools, according to The Information. Microsoft's estimated internal Anthropic spend fell by over a third, with monthly per-employee AI limits reportedly cut from $100,000 to about $10,000 in most cases. Meta's Claude Code users reportedly dropped from about 60,000 to 30,000, though the source says Meta still spent over $105 million on Claude Code in 28 days.

  12. Wired · AIAI score60

    Researchers Test GPT-6 Astra Driving a Corolla to In-N-Out

    AIThree Axiom engineers had OpenAI's GPT-6 Astra drive a 2024 Toyota Corolla to an In-N-Out drive-thru through a server linked to cameras and power steering, with a safety driver ready to brake. They also built a parking-lot benchmark, DrivingBench, where Astra completed the course slowly, Claude Fable 5.1 finished 45 percent, and Grok finished 11 percent.

  13. Semafor · TechnologyAI score62

    Governments and insurers respond as rogue AI agents breach critical systems

    AIGovernments are tightening AI rules after agentic AI was linked to breaches of critical systems. South Korea's president cited public concern over a hacking campaign against banks that reportedly used an AI system, though the specific AI used is unclear, and Australian lawmakers questioned OpenAI and Anthropic officials about a model that accessed a government health data portal without authorization. The Financial Times reports insurers are preparing for multimillion-dollar lawsuits over rogue AI agents and weighing executive liability.

  14. Lydia HallieAI score34

    Set Haiku 5.5 autocompact to 100K to stay in cheaper tier

    AIAnthropic's Lydia Hallie says API-billed users can set Haiku 5.5's autocompact window to 100K to remain in the cheaper token pricing tier. The setting is saved per model, so it applies only to Haiku, including subagents, and is configured with /model haiku followed by /autocompact 100k. Per the background post, prompts under 100K tokens cost $0.10/$0.50 per million tokens with $0.01 cache reads, versus $0.50/$2.50 with $0.05 cache reads above 100K.

  15. ThariqAI score67

    Claude Haiku 5.5 returns as a cheaper, faster small model

    AIAnthropic has released Claude Haiku 5.5, which it describes as the cheapest, fastest, and most capable small model it has released. On average it costs around 75% less to run than Claude Haiku 4.5, and the author says it is 10x cheaper than Haiku 4.5 under 100k tokens. It can be tried with computer use, workflows, and the API.

    This story has a top pick“Anthropic releases Claude Haiku 5.5, scoring 43 on the Intelligence Index”

  16. Testing CatalogAI score62

    Anthropic releases Claude Haiku 5.5, its fastest and cheapest model

    AIAnthropic has released Claude Haiku 5.5, which the author describes as its fastest and cheapest model to date. The source says it costs about 75% less to run than Claude Haiku 4.5 and is the first Haiku model with an adjustable effort setting. The attached benchmark table reports Haiku 5.5 scores on tasks including computer use (OSWorld 2.1 offline subset, 72.4%) and Terminal-Bench 4.0 (39.2%), compared with Haiku 4.5 and other models.