Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri293 items
  1. dexAI score40

    Dex Horthy says small tasks should skip heavy planning workflows

    AIDex Horthy says the share of tasks that can be one-shot without strict process has grown, but alignment, grilling, and planning workflows still matter. He argues that heavy planning on small tasks makes developers feel slower, and predicts tools will add escape hatches so humans or models can decide to ship directly. He adds that as model capabilities improve, the "smart zone" has grown to roughly 200k–400k tokens, and HumanLayer is prototyping research-to-implement and research-to-short-design-to-implement workflows.

  2. SiliconANGLE · AIAI score40

    OpenAI reports $18 billion less revenue, hitting AI stocks Thursday

    AIOpenAI told investors it had $18 billion less revenue than the $68 billion it reported last month, and AI-linked stocks including Nvidia, CoreWeave, Oracle and Nebius fell Thursday. Google debuted a Gemini assistant that can act autonomously, generate code and complete work across web, mobile and desktop. Anthropic released Claude Haiku 5.5 and halved Sonnet 5.5 cache read prices.

  3. The Wall Street Journal · TechAI score4

    Extreme weather is everywhere now: how to prepare your home

    AIThe Wall Street Journal's tech newsletter covers how to prepare homes for extreme weather, how AI could help short-staffed fire departments, the electronic battlefield in warfare, and a startup seeking to double the world's compute. The source provides only a teaser, so no further figures or product details are available.

  4. The Guardian · AIAI score38

    Britain's advertising directors fear AI will cut off the next generation's training

    AIWPP has opened a flagship AI-enabled production facility in east London as part of a £600m WPP Production business, with CEO Cindy Rose saying AI is ushering in a "golden age of marketing". Some AI-led ads can be made up to 60% cheaper, one industry source says, while critics including Luke Scott warn that a lost generation of directors may miss the hands-on training that built careers like Ridley Scott's.

  5. elvisAI score60

    Meta researchers propose agent plasticity to measure self-improvement efficiency

    AIResearchers from UC Berkeley, Meta Superintelligence Labs, and other institutions introduce agent plasticity, the gain on held-out tasks per dollar of learning cost, with model weights frozen. The paper reports that in chess, Go, and Hex, Claude Fable 5 reaches the highest final score while GPT-5.6 Sol gains the most per dollar, and in NetHack only Claude Opus 5.5 improves significantly.

    Image from @omarsar0's post
  6. 404 MediaAI score11

    Behind the Blog: 404 Media discusses AI and spirituality

    AI404 Media's Behind the Blog column discusses AI and spirituality, with Jason saying the outlet writes about AI's current capabilities and harms rather than dismissing it outright. He says reporters sometimes test AI tools while working on stories to write from an informed perspective. The excerpt does not say more about the spirituality discussion.

  7. TechCrunch · AIAI score44

    a16z's Olivia Moore says consumer AI revenue is mostly prosumer and many categories lack AI apps

    AIAndreessen Horowitz partner Olivia Moore released a report on the top 100 consumer AI apps, finding ChatGPT still leads by a wide margin while smaller players like Suno and ElevenLabs show staying power. Moore says almost all AI revenue comes from subscriptions and token usage, and that most consumer AI is prosumer AI. The report finds no top-100 entrants in social, dating, marketplace, retail, travel, finance, or health categories.

  8. Mike KnoopAI score62

    Tufa Labs hits 88.06% on ARC-AGI-2, clearing the Kaggle bonus threshold

    AIMike Knoop says the 85% Grand Prize bonus threshold has been reached on Kaggle. The ARC Prize 2026 leaderboard lists Tufa Labs first at 88.06%, followed by Rabbithole at 80.56% and Yi-Chia Chen at 77.22%. Knoop says this will be the final year for ARC-AGI-2 on Kaggle and expects an open-source, low-cost, offline reproducible solution and model.

  9. The Algorithmic BridgeAI score40

    Meta's AI comeback follows heavy Anthropic Claude spending and a new Muse Spark model

    AIMeta spent heavily on Anthropic's Claude models, with internal use reaching up to 60,000 employees and a projected $10 billion yearly spend, according to The Algorithmic Bridge. The author says Meta then released Muse Spark, which scored 52 on the Artificial Analysis intelligence benchmark, on par with Claude Opus 4.6.

  10. Tessl BlogAI score36

    Tessl's agentic code review splits PR checks into standards, lenses, and memory

    AITessl Blog describes an agentic code review workflow built for teams whose coding agents produce pull requests faster than humans can review them. The workflow runs review against a written standard in the repository, applies four parallel perspectives covering correctness, security and privacy, scale and resilience, and maintainability, then records each finding, verdict, and response. Tessl Code Review, which the post says is free to start, runs these perspectives as skills, and the team's memory of past decisions is fed back into the standard.

  11. Ai2 (Allen Institute for AI)AI score46

    Ai2 describes GPU time budgets that replaced its priority-based cluster scheduler

    AIAi2's AI Infrastructure team replaced its priority-based scheduler for GPU clusters with GPU time budgets, hierarchical fair-share allocation, and a time-slicing contract. The team says the change moved debates over how much GPU time each research project deserves from case-by-case operational decisions into a transparent budgeting process. The clusters range from 88 to 1024 GPUs across NVIDIA H100, B200, and B300 hardware, and serve about 150 internal researchers.

  12. AWS Machine Learning BlogAI score67

    How Postman runs Agent Mode for 40 million developers on Amazon Bedrock

    AIPostman describes the architecture behind Agent Mode, its AI agent for API testing, documentation, discovery, and implementation. The post covers limiting tools per task, using schema-based queries, building purpose-shaped context handlers, and running on Amazon Bedrock with cross-Region inference and prompt caching. Postman reports that tool-selection errors rose once the visible toolset exceeded about 40 tools.

    Why it matters: The post shows concrete patterns for tool scoping, context handling, and Bedrock routing and caching, which apply to any team moving an agent past a prototype.

  13. AWS Machine Learning BlogAI score36

    AWS recaps September 2026 Bedrock, AgentCore, and Strands updates for AI builders

    AIAmazon Bedrock Managed Agents, powered by OpenAI, entered public preview, and OpenAI's GPT-6 Astra, GPT-6.1 Sol, and GPT-6.1 Luna became generally available on Amazon Bedrock. AWS also released Strands Decider 2B, a 2B-parameter open source decision model that answers in about 115ms locally, and said the Strands harness uses 28 percent fewer tokens than popular harnesses while matching their accuracy.

  14. Hugging Face BlogAI score38

    Ai2 replaces priority scheduler with GPU time budgets for cluster allocation

    AIAi2's AI Infrastructure team replaced its priority-based GPU cluster scheduler with a system using GPU time budgets, hierarchical fair-share allocation, and a time-slicing contract. The team says the change turns decisions about how much GPU time each research project receives into a transparent administrative budgeting process. Its clusters, which range from 88 to 1024 GPUs including H100, B200, and B300 units, serve about 150 researchers facing demand two to three times available capacity.

  15. ElevenLabsAI score32

    ElevenLabs partners with Banner Health on AI voice agents for patient calls

    AIElevenLabs says it is partnering with Banner Health to answer patient calls with AI voice agents, starting with primary care scheduling. The ElevenAgents system books, reschedules, or cancels appointments directly in Banner's electronic medical record at any hour, and transfers calls to a Banner team member with context when a patient asks for a person.

    Image from @ElevenLabs's post
  16. elvisAI score40

    Syren Video learns your style to build AI videos from prompts

    AISyren Video, a new agentic video tool, learns preferred graphics, motion, and editing rhythm from a user's library and generates new videos from a prompt. Users refine the results through chat, and the tool is free to try in a browser or through Claude MCP, per the company's announcement. The post's author says the education sector is exploring it.