Skip to contentSkip to stories

Updated

#Expert opinion

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. Lucas Beyer (bl16)XAI score10

    Lucas Beyer questions crediting a human "meatproxy" in a math paper

    AILucas Beyer, who says he is not a mathematician, finds it bizarre that an arXiv paper credits a "meatproxy" and lists an author "in personal capacity." He comments on the awkward form of the attribution while congratulating the authors on the content, which is linked in a reply from @mahdi_tcs_.

    Image from @giffmana's post
  2. Sakana AIOfficialAI score16

    Sakana AI's CEO argues AI advantage now lies in system orchestration

    AIAt the STS forum's 23rd annual meeting in Kyoto on October 5, Sakana AI president Ito spoke at the Koji Omi Memorial Plenary Session on AI's lights and shadows. He argued that competitive advantage in AI deployment is shifting from single-model performance to the ability to orchestrate entire systems, and that using multiple models alongside an independent capability to evaluate AI is key to a new AI sovereignty.

    Image from @SakanaAILabs's post
  3. Gergely OroszXAI score10

    Pragmatic Engineer receives about 100 non-AI essays on industry change

    AIGergely Orosz says the Pragmatic Engineer received about 100 submissions in a single day for an essay contest on how developers see the industry changing. He praised the human-written, non-AI essays as deeply thoughtful and said many will soon be published for readers.

  4. dexXAI score3

    Dex Horthy says "real slop" has never been tried yet

    AIDex Horthy's post claims "real slop has never been tried," a brief remark with no details or figures. The quoted reply, from @emmanuel_2m, jokingly says they are no longer sending their deck, which suggests the comment is a humorous take on AI-generated or low-effort presentation content.

  5. Harrison ChaseXAI score20

    Harrison Chase praises a take on agent harnesses

    AIHarrison Chase, founder of LangChain, endorsed a post on harnesses with the brief comment "Good take on harnesses." The post, from @zeeg, argues that general coding harnesses like Codex will be superseded by specialized ones and that local models will handle most daily tasks within five years.

Oct 5

Oct 5Mon
  1. dexXAI score14

    Founder pitches for human-in-the-loop AI guardrails draw skeptical feedback

    AIDex Horthy says he repeatedly gets founder requests for feedback on human-in-the-loop notification, guardrail, or audit products, and lessons he learned in late 2024 and early 2025 apply to them. Akio Nuernberger, linked as background, reports receiving multiple monthly inbound messages from such startups without a single Langfuse customer showing interest.

  2. roonXAI score10

    Asking whether California's SB53 whistleblower protections have ever been used

    AIroon asks whether SB53's whistleblower protections have ever been successfully invoked, such as by contacting the California Attorney General, and whether they have achieved anything. The post says the only AI-safety whistleblower-driven government action it can recall is Andy Jassy calling Scott Bessent, which it presents as a joke.

  3. SemiAnalysisXAI score22

    Subscription plan value depends on model and workload credit costs

    AISemiAnalysis argues a subscription's worth cannot be judged as a fixed dollar amount, because monthly payments buy credits that each model and token type consume at different rates. Since credit cost ratios can differ sharply from API price ratios, the API-equivalent value of the same $200/month Claude plan shifts with the model and workload.

    Image from @SemiAnalysis_'s post
  4. Mike KnoopXAI score62

    Dust pretrains transformers with zeroth-order optimization, approaching backprop results

    AIDust is a zeroth-order method that pretrains transformers and sometimes matches or exceeds backprop given large compute. The authors report it is about 1,000 to 10,000x more compute efficient than EGGROLL, the state-of-the-art ES method, for training transformers. The post also cites the gradient-alignment result up to 1B tokens and the virtual population idea for scaling.

  5. Thomas WolfXAI score14

    Thomas Wolf hopes Claude Opus 4.6 stays available for a long time

    AIThomas Wolf, who runs Hugging Face, said he hopes Claude Opus 4.6 remains available for a long time. The post is a brief expression of preference, supported by a quoted post in which David Holz reported that in a self-run "have fun" test across LLMs, Opus 4.6 repeatedly won by imagining brief worlds of contradictions inside falling water droplets, while he felt newer models seemed to have less fun.

  6. Nous ResearchOfficialAI score18

    Nous Research argues AI agents should give users full control

    AINous Research says users should control their agent's models, data, memory, compute location, prompts, tools, and code. The post lists choices such as switching models mid-conversation, running fully offline, and exporting the agent. It frames these freedoms as the standard an agent should meet, calling it "yours."

  7. Harrison ChaseXAI score50

    Cognition's Devin adds "Dreaming" offline memory cleanup, open-sourced as a standard

    AIHarrison Chase praises Cognition's "Dreaming" feature, which lets Devin clean stale memory records and surface latent information offline. He argues agent memory needs an offline cleanup loop rather than only better retrieval, and questions how inferred memories get validated before use. He also welcomes Cognition's plan to release Agent Memory Repo as an open standard.

  8. Ethan MollickXAI score10

    Mollick Criticizes Labs for Deferring AI Policy to Superintelligence

    AIEthan Mollick argues that AI labs deferring hard policy decisions on building AI well by assuming superintelligence will soon arrive and resolve them is a flawed stance. He frames this as a cautionary poem that ends with a warning about what happens if superintelligence does not arrive.

    Image from @emollick's post
  9. LlamaIndex 🦙OfficialAI score36

    Agentic OCR replaces single-pass text extraction with verified parsing loops

    AILlamaIndex argues that traditional OCR, which makes one pass and returns unchecked text, is being replaced by agentic OCR. The approach treats parsing as a loop with layout-aware reading order, routing of hard elements such as tables and charts to suitable models, and multi-pass self-correction.

    Image from @llama_index's post
  10. Gergely OroszXAI score35

    Gergely Orosz says coding agent product strategy feels like "YOLO"

    AIGergely Orosz says many coding agents seem to follow a "YOLO" product strategy, with rapid week-over-week change learned about through random social media posts. He notes this makes some sense given how quickly the industry and capabilities keep changing. Quoted context reports that Anthropic is removing Cowork's local option for Pro/Max users, with new tasks running in the cloud while existing local tasks stay on the computer.

  11. Hamel HusainXAI score15

    Hamel Husain on how often to run AI evals

    AIHamel Husain advises weighing the cost of running an eval and how saturated it is against the business value of catching errors. The post links to a full FAQ on evals, with a dedicated answer on how often to run them.

    Image from @HamelHusain's post
  12. IEEE Spectrum · AINewsAI score36

    Six Guidelines for Governing AI Agents in Enterprise Operations

    AILowe's enterprise AI transformation leader outlines six guidelines for governing AI systems, arguing that people must set principles, decision rights, and escalation thresholds rather than only building the technology. The author, who coauthored The Enterprise Brain, cites a 2025 MIT Media Lab Project NANDA report estimating that about 5 percent of integrated generative-AI pilots generated substantial value.

  13. roonXAI score12

    roon argues capital realism and successionism are the same view

    AIRoon, posting as @tszzl, says the "capital realism" and "successionism" views are identical and points to Elon Musk as drifting toward them over the decades. He hopes Musk, one of the few people able to direct capital toward human-centered ends, will try to steer it that way.

  14. The Next PlatformNewsAI score42

    Gartner Forecasts AI Spending Rising to $3.64 Trillion by 2027 as IT Shrinks

    AIGartner's latest forecast projects AI spending rising 49.7 percent to $2.67 trillion in 2026 and 36 percent to $3.64 trillion in 2027, following a 2.6X jump to $1.79 trillion in 2025. Traditional IT spending is in recession, shrinking 12.5 percent in 2025 and projected to fall further, which the article says means AI spending will exceed traditional IT spending in 2027.

  15. LumaOfficialAI score12

    Luma says research must reach creators to become part of craft

    AILuma argues that a model alone has never made a film, and that the real work lies in getting new capabilities into creators' hands and ensuring they can use them. The post frames this delivery and usability as how research becomes part of the creative craft.

    Video from @LumaLabsAI's post