Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 3

Oct 3Sat
  1. roonXAI score8

    Post mocks flat AI progress on graphs as invisible

    AIroon (@tszzl) jokes that graphs show straight lines where AI progress should be visible, saying the AI breakthrough cannot even be seen. The post offers no specific models, benchmarks, or figures.

    Image from @tszzl's post
  2. Amjad MasadXAI score42

    Amjad Masad and Alex Atallah discuss AI independence and specialized agents

    AIAmjad Masad of Replit and Alex Atallah of OpenRouter discuss why AI independence and model diversification matter for enterprises. They argue that depending on a single lab risks lock-in and that specialized agents may outperform one general superagent. The post presents the conversation as a podcast episode, the first Atallah has done since Stripe acquired OpenRouter.

  3. Felix RiesebergXAI score12

    Felix Rieseberg praises a meta Claude-made music video

    AIFelix Rieseberg, who posts from Anthropic's account, says this is his favorite Claude music video so far. He prefers Claude as a human-directed tool for making art rather than generating art unattended, though he enjoyed this one, which he calls a little meta.

  4. Vaibhav (VB) SrivastavXAI score23

    OpenAI's Vaibhav Srivastav urges people to slow down and connect

    AIIn a reflective X post, Vaibhav Srivastav encourages readers to enjoy a local restaurant meal, have a relaxed drink somewhere new, and talk with people they may not agree with. He closes by asking friends to be checked on and to enjoy the ride.

    Image from @reach_vb's post
  5. Nathan LambertXAI score22

    Lambert doubts frontier AI pacing is practical, favors preparedness instead

    AINathan Lambert argues that pacing frontier AI is a good idea in principle but unworkable in practice, asking who would decide which capabilities or benchmarks to slow down. He warns that halting capability work could shift research toward swarms and efficiency, which bring their own risks. He contends most AI risk comes from diffusing existing models, so investment should go to preparedness and pressing labs to be more careful.

  6. Boris PowerXAI score20

    OpenAI's Boris Power says AI model value is rising exponentially

    AIBoris Power calls the progress "amazing news for the world" and says the value derived from using these models continues to climb exponentially. Quoted context from Ramp's AI Index says AI spend fell, driven by frontier price cuts and competition between OpenAI and Anthropic, with open source models under 5% of business spend.

  7. Yuchen JinXAI score22

    Yuchen Jin says terminals are wrong for coding agents

    AIYuchen Jin argues that the terminal is the wrong interface for coding agents, since managing many tabs creates cognitive overhead while context should persist. He says he rarely needs an IDE like Cursor because he seldom navigates the whole codebase now, calling the agent rather than the file the new primitive. He names the Codex desktop app as the best agentic UI for now, while noting the space is still early.

  8. Max ZeffXAI score45

    Former OpenAI safety staffer says culture, not rules, needs fixing

    AIMax Zeff quotes former OpenAI safety team member David Robinson, who resigned this week, saying he regrets not staying to push for staffing and culture changes. The quoted passage says colleagues were too busy sprinting to consider or make major changes. The Atlantic piece argues that the fix lies in culture rather than specific rules or new laws.

  9. Joshua AchiamXAI score35

    Achiam says OpenAI must earn public trust on superintelligence safety

    AIJoshua Achiam praises former colleague David Robinson's critique that AI safety has not adopted professional safety-engineering practices from other fields. He argues OpenAI must meet a higher bar, earning public trust for a path to superintelligence through high-reliability engineering, candid incident disclosure, and unimpeachable third-party verification.

  10. Joshua AchiamXAI score12

    Joshua Achiam backs criticism of WBE and calls him destructive

    AIJoshua Achiam says he is normally reluctant to endorse such takes, but he argues that WBE, whom Austin is friendly with, describes himself as egregiously evil, manipulative, and destructive. He says the case for mitigating factors is hard to find, moving beyond showing grace to flawed people. The post quotes @astridwilde calling for WBE's removal from leadership and a boycott of Mox and Manifund events.

  11. Exponential ViewBlogAI score28

    Weekend reads on effective altruism, Anthropic, and machine consciousness debates

    AIThe Economist argues that effective altruism's belief that only its adherents can be trusted with powerful AI is an alarming idea, and the newsletter links to a response from Coefficient Giving CEO Alexander Berger. The New York Times reports that Anthropic consulted religious scholars and theologians on machine consciousness, and the newsletter notes Anthropic's team proposed withdrawing from a Vatican event before the Pope's encyclical Magnifica Humanitas said AIs do not possess a moral conscience.

  12. SantiagoXAI score23

    Consultant reports engineering teams gain speed by validating agent output

    AIA consultant helping several companies adopt AI in engineering workflows says teams become much more productive and ship better software faster once they ramp up. The shift he recommends is from prioritizing human-maintainable code to building strong processes that validate what agents do, and he rejects the view that such software will later prove worthless.

  13. Orange AIXAI score55

    Local Qwen Flash inference on consumer GPUs jumps roughly tenfold in a week

    AIThe author reports that a dual RTX 5070 Ti setup running Qwen Flash rose from 200 prefill and 10 decode to 2200 prefill and 67 decode, now on a single card, using Strata and a custom PR. The post argues that such consumer-hardware speeds, once limited to top-end machines, could pressure the economics of selling model compute via API.

  14. Dongxi NLPXAI score14

    Dongxi NLP Advises Against Buying DGX Spark Too Early

    AIThe author advises against buying an NVIDIA DGX Spark for a personal AI setup, arguing that single products like the DGX Spark can pull buyers toward costly multi-unit clusters. The post warns that combining such hardware with expensive coding agent plans could drain a moderate cash flow quickly. It instead recommends building up one's own capabilities before engaging with NVIDIA's higher-tier offerings.

    Image from @dongxi_nlp's post
  15. Yuchen JinXAI score4

    SF rents surge as AI boom pushes new grads' affordability limits

    AIA former tenant reports their San Francisco one-bedroom rose from $4,000 to a $6,000 monthly relisting, rented instantly after they moved out. A friend pays $5,000 for a studio without a dishwasher, prompting the question of how new CS graduates can afford the city. The post frames the trend as part of the AI economy.

Oct 2

Oct 2Fri
  1. Hamel HusainXAI score35

    Hamel Husain criticizes a Claude Code mod demo as hard to follow

    AIHamel Husain says he cannot understand a demo video for a new Claude Code modding feature, calling it visual slop. He suggests the feature may be cool but argues demos should be understandable to humans. The background post says Claude Code can now be modded to change behavior, customize the UI, or add features via TypeScript or Claude-built mods installed through /plugin.

  2. IThome · AINewsAI score36

    Analyst Dumps Airbnb, Buys Meta After Testing Meta's Muse AI Agent

    AIIndependent analyst Mostly Borrowed Ideas said he sold his Airbnb stake and added to Meta after testing Meta's Muse AI agent for about 10 days. He said Muse browsed Airbnb like a human, then found a farmhouse stay about 60% cheaper by booking directly with the host, suggesting AI agents could bypass booking platforms. He acknowledged Muse is slow, with a five-hotel price comparison taking 14 minutes.

  3. Stanford HAIOfficialAI score15

    Stanford HAI's James Landay on keeping people central to AI policy

    AIStanford HAI Denning Director James Landay discussed educating lawmakers, working with industry while staying independent, and what the AI Index reveals beyond headlines. He spoke with Imagination in Action in a video interview about how people can remain at the center of AI development.

  4. MIT News · AIOfficialAI score14

    MIT's Cathy Wu Uses Reinforcement Learning to Tackle Transportation Challenges

    AIMIT associate professor Cathy Wu is applying machine learning and reinforcement learning (RL) to design safer, more efficient transportation systems. Her team found RL can train effectively on about 10 percent of related problems, and a selection algorithm improved training efficiency by up to 30 times. Her recent work estimates eco-driving measures could cut vehicle emissions by 11 to 22 percent.

  5. Epoch AI · The Epoch BriefOfficialAI score62

    Epoch AI estimates 2026 compute could run hundreds of millions of AI agents

    AIEpoch AI estimates that compute built from projected 2025 to 2027 high-bandwidth memory shipments could support tens to hundreds of millions of frontier AI agents, or billions of cheaper ones. Running nonstop, the top-tier agents would match the working hours of 140 million to 700 million full-time employees, and the central DeepSeek V4 Pro estimate of about 1.9 billion agents would match 8 billion workers.

    Why it matters: The estimate converts memory shipments into agent capacity and revenue ranges, showing how hardware supply could translate into labor and sales if demand keeps up.

  6. Vaibhav (VB) SrivastavXAI score13

    Vaibhav Srivastav says OpenAI's Dot feels like Codex-level product-market fit

    AIVaibhav Srivastav says OpenAI's Dot has achieved product-market fit comparable to the Codex App launch, citing proactive fixes and recommendations that improve with use. Sam Altman, quoted in the background, calls Dot his favorite OpenAI product and says it feels better each day as it learns his workflow and handles tasks he dislikes.

  7. Hamel HusainXAI score7

    Hamel Husain Questions Keeping Laptops Open Instead of Remote SSH Access

    AIHamel Husain says it is baffling that people keep their laptops open for long-running work instead of using SSH into a machine or Codex's remote feature with a Mac mini. He argues that after being stuck in this situation more than three times, continuing it is crazy.

  8. Harrison ChaseXAI score53

    Google Research's Cogentic uses multi-agent proof search to produce verified results

    AIGoogle Research's Cogentic is a multi-agent harness running on Gemini that searches for proofs of open theoretical computer science problems without expert hints. It runs rounds where an orchestrator launches provers, two adversarial verifiers must both accept each draft, and shared disk documents store attempts and verified lemmas. The system produced new results on five open problems in online learning, auction theory, and mechanism design, each checked by domain experts.

  9. Thomas WolfXAI score34

    Thomas Wolf backs Sigil's Underdog for on-device private personal AI

    AIThomas Wolf says most personal AI should run on users' devices and praises Sigil's Underdog, which is powered by the Hugging Face Hub. Underdog's announcement says it aims to deliver free, private AI that runs on consumer hardware, with backing from a16z, Khosla Ventures, and other investors.