Skip to contentSkip to stories

Updated

#Expert opinion

Showing low-relevance items too. Hide low-relevance items

Aug 3

Aug 3Mon
  1. Intern Large ModelsOfficialAI score34

    Legal and AI meanings of "agent" diverge over accountability for machines

    AIThe post contrasts AI agents, systems that perceive, plan, and act, with legal agents who receive authority and assume fiduciary duties and accountability. Mark Nitzberg of Berkeley AI Research says closing this gap requires AI that is well-founded, legible, and steerable, while Lan Xue of Tsinghua notes that because machines cannot be punished, responsibility must be redistributed across design, development, deployment, and use.

    Video from @intern_lm's post

Aug 2

Aug 2Sun

Aug 1

Aug 1Sat
  1. Andrej KarpathyXAI score66

    Karpathy tests Opus 5 by rendering Lord of the Rings opening in 3D

    AIAndrej Karpathy gave Claude Opus 5 the first paragraph of Lord of the Rings with a 1M token budget and asked for a Three.js render. Opus spent about two hours writing 5500 lines of code that procedurally renders the story, which Karpathy calls janky but fun. He notes the model struggled to audit its work because it cannot efficiently perceive video or play the resulting game, relying on slow screenshots that led to several errors.

    Video from @karpathy's post
  2. Amanda AskellXAI score3

    Askell mocks fear of a "permanent underclass" in a Padmé reference

    AIAmanda Askell, who is affiliated with Anthropic, posted a brief reaction to people discussing how to avoid "the permanent underclass," comparing her response to a Padmé moment. The post offers no further argument or data, so its substance is limited to this comment.

    Image from @AmandaAskell's post
  3. Kevin Weil 🇺🇸XAI score38

    OpenAI's Kevin Weil touts ten major mathematics advances from OpenAI

    AIKevin Weil of OpenAI called ten linked results in mathematics "major" and linked to an OpenAI article titled "ten advances in mathematics." He congratulated researchers Sébastien Bubeck, Noam Brown, Mark Chen, and Meredith Tanner, and referred to a future model available to the whole world.

  4. Werner VogelsXAI score22

    Werner Vogels praises conversation with Clare Liguori on Kiro and agent support

    AIWerner Vogels called his conversation with Clare Liguori an excellent discussion of developer support for agents and Kiro. The quoted InfoQ podcast covers moving agents from demo to production, including why extra if statements can hurt agent performance, achieving high accuracy and low cost with small models, and observability within agent hops.

Jul 31

Jul 31Fri
  1. Thinking MachinesOfficialAI score44

    Thinking Machines argues for staged access to capable open-weight models

    AIThinking Machines says indiscriminately releasing model weights is unsafe, but keeping capable models inside a few labs is also not the answer. Its new post describes how it assessed its model Inkling and argues that access should widen in stages. The company says it has not mapped the full path, only the portion it can currently see.

Jul 30

Jul 30Thu
  1. Thinking Machines LabOfficialAI score65

    Thinking Machines proposes staged, evidence-based release path for open-weight models

    AIThinking Machines argues that safe open-weight releases depend on both model safety testing and readiness of the surrounding ecosystem, and that release should proceed in iterative stages. For its Inkling and Inkling-Small models, internal evaluations, four external red-teaming groups, and adversarial fine-tuning tests led the company to conclude that releasing the weights was not likely to add material risk beyond existing open-weight models.

    Why it matters: The post lays out a staged, evidence-gated path to releasing open weights, with concrete safety tests and the ecosystem measures behind each stage.

  2. Jeff DeanXAI score38

    Jeff Dean thanks Diana Hu after Startup School conversation at Chase Center

    AIJeff Dean, Google's Chief Scientist, thanked YC partner Diana Hu for an engaging conversation at Chase Center last weekend, his first in a basketball arena. The post is a brief acknowledgment, with the surrounding context describing a Startup School 2026 discussion on AI inference hardware, the origins of TPUs, and advice for founders.

  3. Microsoft AI BlogOfficialAI score14

    Leaders share how AI transformation depends on mindset, team adoption, and culture

    AILeaders interviewed for Alysa Taylor's "What's the Tea?" series, including executives at Adobe, Lumen, and Sitecore, say the shift from AI apprehension to expected adoption is the precondition for transformation. Behavioral scientist Jon Levy argues the goal is raising a team's collective intelligence, not just cutting costs, with leadership and continuous training driving scale.

Jul 29

Jul 29Wed
  1. Ahmad Al-DahleXAI score52

    Ahmad Al-Dahle argues AI capex is both short on compute and overbuilt

    AIAhmad Al-Dahle argues that AI infrastructure faces both a compute shortage and overbuilding, with the four largest hyperscalers planning roughly $725 billion of capex in 2026, up 77 percent from last year. He describes a "mutually assured construction" dynamic in which every well-capitalized player buys the same insurance against falling behind, so the industry overbuilds by construction.

Jul 28

Jul 28Tue
  1. Ali GhodsiXAI score5

    Ali Ghodsi Endorses Democratizing AI as a Positive Vision

    AIDatabricks CEO Ali Ghodsi endorsed democratizing AI, reacting to a post by finkd arguing that the future of superintelligence should be for everyone. The main post gives no specific product, model, or figure, so the summary stays limited to that stated position.

  2. Rowan CheungXAI score40

    Zuckerberg says Meta's superintelligence lab should stay small and elite

    AIMark Zuckerberg said Meta's superintelligence lab should have 50 to 100 people who can keep the whole project in their heads at once. He said he personally recruits top AI researchers because underperformers have an outsized negative effect, and he rejects top-down deadlines and non-technical management layers.

    Video from @rowancheung's post
  3. Kimi.aiOfficialAI score14

    Kimi launches Global Ambassador Program for K3 community builders

    AIKimi has launched a Global Ambassador Program to recruit influential individuals who have implemented Kimi K3 into their products, agents, workflows, or communities. Applicants are expected to share their experiences and passions with a broader audience. Applications are open at

    Image from @Kimi_Moonshot's post
  4. METR BlogOfficialAI score58

    METR outlines how independent researchers could investigate AI agent misalignment incidents

    AIMETR proposes that AI companies track agent misalignment incidents and have independent researchers investigate the most serious ones, focusing on the motives behind the behavior. The post lists core investigation questions covering incident surveys, root causes, and remediation, along with the model access, transcripts, employee interviews, and training-data tools such investigators would need. It also calls for results to go to company boards and oversight bodies and be published with disclosed redaction terms.

Jul 27

Jul 27Mon
  1. Lilian WengXAI score7

    Lilian Weng reflects on curiosity and cofounder lessons on strategy

    AILilian Weng says her curiosity-driven nature gives her joy in learning and tackling ill-defined problems. She describes how cofounding pushed her to develop new perspectives on company strategy and team building, and how these abstract ideas connect to daily actions and narratives. She adds that real-world experience makes once-theoretical ideas more approachable.

  2. Andrew NgXAI score34

    Andrew Ng urges open models for AI defense, rejecting closed-model safety claims

    AIAndrew Ng praised Nvidia's letter and argued that open models and harnesses are needed for defense, citing the OpenAI-Hugging Face hack. He said claims that closed models are safer are regulatory capture. Jensen Huang's background post says closed AI blocked forensics during the Hugging Face incident, while an open-weight frontier model helped contain it, leading to the Open Secure AI Alliance.

Jul 26

Jul 26Sun
  1. Jeremy HowardXAI score16

    Jeremy Howard criticizes employees who ignore their employer's interests

    AIJeremy Howard argues that many people with otherwise sound judgment seem unable to think clearly about actions that affect the company paying them. The post is a general observation and cites no specific company, product, or event. Its quoted context about Nvidia's CUDA and GPU driver open source release is not the post's main subject.

Jul 25

Jul 25Sat
  1. Kevin Weil 🇺🇸XAI score4

    Kevin Weil says there really is nothing like X

    AIOpenAI's Kevin Weil posted "There really is nothing like X" with no further details in the post itself. A reply tagging NVIDIA CEO Jensen Huang and NVIDIA suggests the remark relates to NVIDIA, but the source does not state the subject.

  2. Ali GhodsiXAI score26

    Longer-running AI agents often perform worse than faster ones, says Ghodsi

    AIAli Ghodsi argues that AI agents which take longer to work through a task are often worse, while Genie reaches results faster. He adds that ontology will be key to giving agents the context they need to answer correctly and quickly. The related post reports that Genie Code outperformed three general-purpose coding agents on more than 400 real user data tasks.

  3. Sebastien BubeckXAI score25

    Bubeck asks what Erdős would do with GPT-5.6 Sol

    AISebastien Bubeck asked what the mathematician Paul Erdős would have done with access to GPT-5.6 Sol. The post offers no benchmark results, prices, or capability details beyond the question itself.

  4. LangChain BlogOfficialAI score39

    What does it mean for companies to "own their intelligence" with AI?

    AILangChain Blog argues that companies need to own their AI intelligence rather than rely on generic models, because general models do not know company-specific policies, workflows, or risk tolerances. Ownership means controlling the agent system (model optionality, harness, and context), the economics, quality, and risk of AI work, and how intelligence compounds over time. The post uses an insurer's claims processing as an example of why off-the-shelf models fall short.

Jul 24

Jul 24Fri
  1. Noah ZwebenXAI score44

    Noah Zweben shares a favorite Opus 5 anecdote from his TA days

    AIAnthropic's Noah Zweben says a tornado-physics assignment he once TA'd for, built in Unity, is his favorite Opus 5 example so far. The quoted Atomic Chat post compares Opus 5, Fable 5, Kimi K3, and GPT 5.6 on three HTML physics scenes, with Opus 5 costing $1.40 versus Fable 5's $2.82.

  2. Alex AlbertXAI score34

    Opus 5 now produces consultant-grade spreadsheets and slide decks, Alex Albert says

    AIAlex Albert, of Anthropic, says Opus 5 now produces near-superhuman spreadsheets and slide decks that match what a consultant would make, just over six months after its predecessor. He also notes that finance professionals are reporting strong reactions to Claude for Excel, and he expects agentic progress seen in coding to extend to other fields in 2026.

    Video from @alexalbert__'s post
  3. Mike KriegerXAI score22

    Mike Krieger says models now build games from brief, dynamic prompts

    AIMike Krieger, who is associated with Anthropic, says two games were built from prompts of about four sentences that used dynamic /workflows extensively. He contrasts this with earlier in the year, when he relied on a bespoke harness and verification system, noting that current models accomplish much more with far less instruction.

  4. Mira MuratiXAI score16

    Murati says useful AI knowledge must be distributed, backing Jensen Huang's vision

    AIMira Murati argues that the knowledge making AI useful is spread across scientists, engineers, clinicians, and firms, so AI must itself be distributed to benefit from it. She says she agrees with Jensen Huang that this is a future worth building. The post accompanies Huang's shared NVIDIA letter arguing that open models strengthen safety, cybersecurity, innovation, and sovereignty alongside frontier closed models.

  5. Mike KriegerXAI score46

    Mike Krieger says Claude Opus 5 became his daily driver

    AIAnthropic co-founder Mike Krieger says Claude Opus 5 has become his daily driver at work and on weekends. He reports it can work for hours on complex tasks and consistently gets to the bottom of tricky problems, and he has also built some games with it. Anthropic's announcement describes Opus 5 as close to the frontier intelligence of Fable 5 at half the price.