Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Oct 9

TodayOct 9Fri289 items
  1. TechCrunch · AIAI score36

    Amazon drops data center NDAs as AI agents seek consumer credit cards

    AIAmazon says it will stop using NDAs when negotiating data center deals with local governments, following a similar move from Microsoft earlier this year. Secrecy has fueled community backlash against AI infrastructure, with hundreds of proposed and enacted moratoriums from New York to San Francisco. The episode also covers startups seeking consumer access to inboxes, files, and credit cards for AI agents.

  2. LangChain BlogAI score40

    LangChain adds emoji reactions to Managed Deep Agents Slack channels

    AILangChain's Managed Deep Agents v0.9 adds a reactions attribute for Slack channels that accepts either an emoji string or a callable returning one. The article shows a function that returns a bug emoji when a message contains "broken" and eyes otherwise. It also shows a TypeSafe Classifier that picks from a seven-emoji vocabulary and falls back to eyes below 25% confidence.

  3. NewcomerAI score42

    OpenAI's ARR figure is $20 billion below reported levels amid revenue math confusion

    AIThe Financial Times reported that OpenAI's annualized revenue run-rate was $20 billion below a widely cited $70 billion figure, a gap the newsletter attributes to confusion between gross and net revenue. OpenAI has said its figure is net, while Anthropic's includes partner sales such as those made through Amazon and Microsoft. Bloomberg reported that OpenAI expects to reach $70 billion annualized revenue by year end, assuming a net basis.

  4. a16z NewsAI score40

    a16z leads investment in TypeSafe AI, maker of Jev System One model

    AIa16z says it is leading an investment in TypeSafe AI, whose Jev model hands decisions to code as typed values and reached 1 trillion tokens generated three days after launch. The company says Jev costs roughly 1/100 to 1/500 of frontier models and runs 100x faster on classification tasks at comparable accuracy. TypeSafe says 25% of the Fortune 500 have integrated Jev.

  5. a16z NewsAI score33

    Prediction markets showed no significant partisan or demographic bias in election pricing, study finds

    AIAn NBER working paper by Prof. Zitzewitz, drawing on over 100 years of prediction market data, found no dimension of political affiliation, gender, race, or age where returns differed from zero beyond the standard error. The only exception was non-US elections, where markets appeared to overrate right-leaning candidates, but that bias was not statistically significant. The findings are preliminary.

  6. O'Reilly RadarAI score40

    US AI oversight debate, OpenAI Dots, and Gemini 4 Argon featured in This Week in AI

    AIThe Trump administration announced a voluntary agreement with major AI companies calling for internal safety monitoring, external audits, and independent board reviews, and the Federal Trade Commission launched an investigation into OpenAI, Anthropic, and other AI companies over potential consumer risks. OpenAI released Dots, a proactive assistant that retains context, works across applications, and acts without waiting for prompts. Google says Gemini 4 Argon can generate up to a million output tokens in a single response.

  7. 🚨 AI News | TestingCatalogAI score41

    Pine AI launches Pine Computer, a cloud runtime for agentic tasks

    AIPine AI launched Pine Computer, a cloud computer, harness, and runtime layer built for agentic tasks. On the publisher's SaaS-Bench v1.1, it posts a 78.3% checkpoint score against 74.3% for Opus 5 with Claude Code, but completes fewer whole tasks, 27.4% against 31.1%. Instead of simulating clicks and screenshots, it reads web pages as structured data, and access is through a private beta waitlist.

    Image from @testingcatalog's post
  8. 🚨 AI News | TestingCatalogAI score62

    Anthropic moves dynamic workflows in Claude Managed Agents into public beta

    AIAnthropic has expanded dynamic workflows in Claude Managed Agents into a public beta, according to Testing Catalog. Users can configure their agents for multiagent orchestration, with Claude planning and operating a fleet of agents to achieve a goal. The post also links a video from Anthropic's ClaudeDevs account, which the author describes as a new SWE norm.

    Video from @testingcatalog's post
  9. ClaudeDevsAI score60

    Claude Managed Agents adds dynamic workflows in public beta

    AIAnthropic's ClaudeDevs account announces that dynamic workflows for Claude Managed Agents are now available in public beta. The feature is a new type of multiagent orchestration in which a lead agent writes a plan that runs across many agents in phases, then combines their results at the end.

    Why it matters: The post describes how a lead agent plans work across many agents in phases and merges their results, a structure useful for understanding complex agent orchestration.

    Video from @ClaudeDevs's post
  10. elvisAI score34

    Elvis Saravia urges builders to focus on agent harnesses and environments

    AIElvis Saravia says AI models are already smart, but they need better harnesses and environments, with major cost implications. He recommends reading a report on how Pine Computer can help teams, and says he will test it himself and share more later. The quoted post from Stanley Wei argues that real-world AI tasks remain slow, expensive and unreliable because AI runs on computers built for humans, and announces Pine Computer.

    Image from @omarsar0's post
  11. dexAI score40

    Dex Horthy says small tasks should skip heavy planning workflows

    AIDex Horthy says the share of tasks that can be one-shot without strict process has grown, but alignment, grilling, and planning workflows still matter. He argues that heavy planning on small tasks makes developers feel slower, and predicts tools will add escape hatches so humans or models can decide to ship directly. He adds that as model capabilities improve, the "smart zone" has grown to roughly 200k–400k tokens, and HumanLayer is prototyping research-to-implement and research-to-short-design-to-implement workflows.

  12. SiliconANGLE · AIAI score40

    OpenAI reports $18 billion less revenue, hitting AI stocks Thursday

    AIOpenAI told investors it had $18 billion less revenue than the $68 billion it reported last month, and AI-linked stocks including Nvidia, CoreWeave, Oracle and Nebius fell Thursday. Google debuted a Gemini assistant that can act autonomously, generate code and complete work across web, mobile and desktop. Anthropic released Claude Haiku 5.5 and halved Sonnet 5.5 cache read prices.

  13. The Guardian · AIAI score38

    Britain's advertising directors fear AI will cut off the next generation's training

    AIWPP has opened a flagship AI-enabled production facility in east London as part of a £600m WPP Production business, with CEO Cindy Rose saying AI is ushering in a "golden age of marketing". Some AI-led ads can be made up to 60% cheaper, one industry source says, while critics including Luke Scott warn that a lost generation of directors may miss the hands-on training that built careers like Ridley Scott's.

  14. elvisAI score60

    Meta researchers propose agent plasticity to measure self-improvement efficiency

    AIResearchers from UC Berkeley, Meta Superintelligence Labs, and other institutions introduce agent plasticity, the gain on held-out tasks per dollar of learning cost, with model weights frozen. The paper reports that in chess, Go, and Hex, Claude Fable 5 reaches the highest final score while GPT-5.6 Sol gains the most per dollar, and in NetHack only Claude Opus 5.5 improves significantly.

    Image from @omarsar0's post
  15. TechCrunch · AIAI score44

    a16z's Olivia Moore says consumer AI revenue is mostly prosumer and many categories lack AI apps

    AIAndreessen Horowitz partner Olivia Moore released a report on the top 100 consumer AI apps, finding ChatGPT still leads by a wide margin while smaller players like Suno and ElevenLabs show staying power. Moore says almost all AI revenue comes from subscriptions and token usage, and that most consumer AI is prosumer AI. The report finds no top-100 entrants in social, dating, marketplace, retail, travel, finance, or health categories.

  16. Mike KnoopAI score62

    Tufa Labs hits 88.06% on ARC-AGI-2, clearing the Kaggle bonus threshold

    AIMike Knoop says the 85% Grand Prize bonus threshold has been reached on Kaggle. The ARC Prize 2026 leaderboard lists Tufa Labs first at 88.06%, followed by Rabbithole at 80.56% and Yi-Chia Chen at 77.22%. Knoop says this will be the final year for ARC-AGI-2 on Kaggle and expects an open-source, low-cost, offline reproducible solution and model.

  17. The Algorithmic BridgeAI score40

    Meta's AI comeback follows heavy Anthropic Claude spending and a new Muse Spark model

    AIMeta spent heavily on Anthropic's Claude models, with internal use reaching up to 60,000 employees and a projected $10 billion yearly spend, according to The Algorithmic Bridge. The author says Meta then released Muse Spark, which scored 52 on the Artificial Analysis intelligence benchmark, on par with Claude Opus 4.6.

  18. Tessl BlogAI score36

    Tessl's agentic code review splits PR checks into standards, lenses, and memory

    AITessl Blog describes an agentic code review workflow built for teams whose coding agents produce pull requests faster than humans can review them. The workflow runs review against a written standard in the repository, applies four parallel perspectives covering correctness, security and privacy, scale and resilience, and maintainability, then records each finding, verdict, and response. Tessl Code Review, which the post says is free to start, runs these perspectives as skills, and the team's memory of past decisions is fed back into the standard.