Skip to contentSkip to stories

Updated

#Expert opinion

Oct 8

Oct 8Thu
  1. Andrew CurranAI score28

    Association for Human Mathematics sets three vows against AI in math

    AIThe Association for Human Mathematics requires members to take three vows opposing AI use in mathematics. Members must not provide technical labor, knowledge, consultation, or publicity to commercial AI companies, and must not publish AI-generated mathematical texts, including papers, referee reports, and lecture notes. Members of its AI-free caucus also must not use AI models in research.

  2. Lauren TanAI score38

    Lauren Tan argues PR volume matters now that agents make coding machines universal

    AILauren Tan argues that with frontier AI agents, anyone can produce code at machine speed, so PR volume now signals productivity alongside impact. She says the bottleneck is trust in agent output, and that higher token costs are worth it compared with hiring many engineers. She frames the engineer's job as building the software-producing machine rather than writing code directly.

  3. Sherwin WuAI score62

    Harvey LAB-AA v1.1 adds hallucination gate, reshaping legal benchmark rankings

    AIArtificial Analysis and Harvey released LAB-AA v1.1, which credits a legal task only when deliverables pass every rubric criterion with no material hallucinations. Grok 4.7 (xhigh) leads at 9.4%, ahead of Muse Spark 1.3 (max) at 8.9% and GPT-6 Astra (max) at 8.6%, while over 60% of otherwise passing results contained a material hallucination. The sharper reordering appears in the hallucination counts, where GPT-6 Astra averages 0.03 material hallucinations per task against 13.96 for Gemini 3.8 Flash (high).

  4. Alex HeathAI score38

    Qualcomm CEO Cristiano Amon on AI phones, glasses, and 6G

    AIQualcomm CEO Cristiano Amon discusses the coming AI smartphone supercycle, arguing phones will not disappear as agents use personal context. He also expects smart glasses to become the largest AI wearable category, and covers Qualcomm's Modular acquisition as an alternative to Nvidia's CUDA software and its data center strategy. The conversation, recorded live at the Snapdragon Summit in Hawaii, also covers 6G being designed for AI.

  5. Tessl BlogAI score42

    Agent Skills Should Be Treated as Supply Chain Components

    AITessl's talk at AI Native DevCon London argues that agent skills, which can be markdown files with instructions and bundled material, act as supply chain components that can shape agent behavior. The author says reading SKILL.md once is insufficient because risks can sit in supporting files, updates, and workspace trust settings. He identifies the danger as the combination of private context, untrusted content, and external communication, and cites research scanning roughly 4,000 public skills for issues including malware-like behavior.

  6. Artificial AnalysisAI score42

    More output tokens don't guarantee higher scores in AI benchmarks

    AIArtificial Analysis reports that generating more output tokens does not necessarily yield a higher score. GPT-6 Astra (max) scored 8.6% using about 81k output tokens per task, while Grok 4.7 (xhigh) used roughly 180k yet scored lower. Three Claude models produced the most output tokens, about 202k to 562k per task, but scored between 2.8% and 6.4%.

  7. TechCrunch · AIAI score36

    Ben Affleck's AI expertise goes viral as he explains neural networks and fine-tuning

    AIActor Ben Affleck drew attention this week for explaining machine learning concepts, including convolutional neural networks, tensors, and transformers, in several recent interviews. He said he fine-tuned open video models by unfreezing weights and training only the last cinematic layer, using a dataset he built over about eight months for his startup. Affleck said he worries about students and learned helplessness more than Skynet, and predicted AI will be additive to the movie business.

  8. The DecoderAI score75

    Mathematicians call for OpenAI boycott after AI-generated proofs flood their field

    AIA group of mathematicians led by Terence Tao has called for a boycott of OpenAI after the company released more than 700 AI-generated proof files at once. Tao and other Fields Medalists argue that mass-produced solutions undermine the discipline's focus on conceptual understanding, while Scott Aaronson contrasts this batch release with Anthropic's collaborative approach. The article reports that the internal model tested about 8,000 problems with roughly a five percent success rate.

  9. Jerry LiuAI score22

    LlamaIndex argues Markdown is the universal format for agents

    AILlamaIndex says Markdown has become a universal representation between humans and agents, preserving headings, lists, and tables while remaining readable to models. Since most unstructured documents are not natively in Markdown, the main challenge is the translation layer, which the company addresses with models that convert document containers into Markdown. The quoted post adds that Markdown keeps table columns intact, with HTML used for tables with merged headers.

  10. SemiAnalysisAI score72

    SemiAnalysis Finds China's AI Safety Rules Target Applications, Not Frontier Models

    AISemiAnalysis argues China's AI safety regime is speed-first, with rules covering content and public-facing services but no frontier-risk duties tied to training compute or capability. Its dataset of 857 releases from nine leading Chinese developers found only 31 (3.6%) ever had a published safety result, and just 9 at launch. The analysis also finds that technical experts favor binding frontier rules while the top leadership's development-first preference settled the policy debate.

  11. Miles BrundageAI score22

    Miles Brundage suspects Anthropic's Claude abuse policy aims at IPO and regulatory capture

    AIMiles Brundage speculates that Anthropic's new rule, making abusive behavior toward Claude a Usage Policy violation effective November 12, 2026, is meant to help its IPO and win favor with the administration as part of a regulatory capture strategy. The post offers this as a guess about motive rather than a confirmed fact, and it relies on the policy change flagged in the quoted post by Andrew Curran.

  12. Lewis TunstallAI score62

    Lewis Tunstall Shares a Physics Paper Proof Developed with OpenAI's Astra Model

    AILewis Tunstall quotes Kyle Cranmer's post about a paper by Nate Gunnarsson on a non-perturbative approach to chiral fermions in the Standard Model, extending Lüscher's abelian result. The paper's acknowledgments state that OpenAI's GPT-6 Astra model was essential, proposing refinement strategies, writing rewrites of the proof, and carrying out Lean verification.

  13. SiliconANGLE · AIAI score30

    Liquid AI Builds On-Device Personal AI Around Device-Level Context

    AILiquid AI is building personal AI that runs on devices such as phones, wearables, PCs, and cars, using its Liquid Context layer, which is optimized for Snapdragon processors, to sit between models, agents, and hardware. The company's agent harness uses its own models to decide which user context to retain and how to compress it within fixed compute limits. Liquid AI is also collaborating with Mercedes-Benz Group AG to bring on-device AI to its cars and plans observability and continuous improvement loops for self-improving agents.

  14. Tessl BlogAI score44

    Continuous AI Brings Agentic Automation to Repository Workflows

    AITessl's blog post argues that repository automation needs Continuous AI, a third pillar alongside CI and CD for scheduled, auditable AI workflows that improve repositories over time. The article describes GitHub Agentic Workflows, which harden agentic workflow specifications into GitHub Actions that can run coding agents such as Claude Code, Copilot CLI, Gemini CLI, or Codex-style agents. It emphasizes read-only agent steps, restricted outputs, and human review of pull requests.

  15. Meta NewsroomAI score22

    Meta Debunks Three Common Myths About Its Data Centers

    AIMeta says its closed-loop liquid cooling recirculates water in a sealed system, so its data centers use less water annually than an average US golf course. The company also says it pays for the new generation and transmission its facilities require, including in Louisiana under its Entergy agreement, and that data centers create construction and operations jobs.

  16. Stanford HAIAI score22

    Stanford HAI leaders urge keeping people central to AI-driven research

    AIStanford HAI associate directors Risa Wechsler and Russ Altman, speaking at a Stanford orientation, argued that AI agents can deepen scientific research but must be paired with interdisciplinary collaboration. They stressed rigorous, reproducible methods and clearly measured uncertainty, since convincing AI answers are not enough. They also said labs must weigh agent costs and preserve mentorship so that automation supports human participation in research.

  17. Elvis SaraviaAI score22

    Interface ring lets users control AI agents by voice from hand

    AINatura AI's Interface is a ring that lets users press and hold to speak requests to AI agents such as Claude Code, Codex, or Hermes, then release to send them. The post argues that screenless interfaces may define the next phase of agent use, since handing work to agents is currently slowed by pulling out a phone. Early-adopter pricing is $99, with shipping slated for January.

  18. The Robot ReportAI score34

    Jabil Says Humanoid Robots Are Moving Toward Tens-of-Thousands Production Volumes

    AIJabil senior director Thomas Brown says humanoid robots are entering a phase of tens of thousands of units, where manufacturability, cost structure, and quality become central. He says Jabil works with developers to cut costs for scale, while compute and memory prices remain a pain point, and that humanoids make sense in factories and warehouses while mobile arms still suit high-speed tasks.