Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. Interconnects (Nathan Lambert)BlogAI score52

    Nathan Lambert argues the open-weight cyber risk debate is missing trade-offs

    AINathan Lambert argues that policy debates on open-weight model cyber risks lack nuance, because banning open models may not reduce risk and could weaken American competitiveness. He says closed frontier APIs have been tied to most documented cyber attacks, and that restricting open models while closed models keep advancing could widen the offense-defense gap. He also argues that Chinese labs' safety practices are shaped by their own government and society, and that the claimed risk of models like Claude Mythos has been overstated.

  2. Mustafa SuleymanXAI score42

    Daron Acemoglu predicts AI will replace only 5% of human work in 10 years

    AINobel laureate Daron Acemoglu argues in the first issue of The Humanist Review, published by MAI, that AI will replace only about 5% of what humans do over the next decade. He says AI is not yet visible in productivity statistics and projects roughly 1.5% added to GDP over 10 years, and he urges building pro-worker tools that make people better at their jobs.

  3. Yuchen JinXAI score72

    Mistral Large 4 launches as a 1T-parameter multimodal model with open weights due end of October

    AIMistral announced Mistral Large 4, a natively multimodal model with 1T parameters and 49B active, available via API today. Mistral claims it is the best open weights model from the US or Europe on aggregated benchmarks, with open weights set for release at the end of October. The author quotes this claim and comments that it appears to beat GLM-5.3.

  4. Sophia YangXAI score45

    Mistral Large 4 tops benchmarks across cybersecurity, legal, and agentic tasks

    AIMistral Large 4 is a 1T-parameter natively multimodal model with 49B active parameters, which the Mistral account says leads open-weights models from the US or Europe on aggregated benchmarks. The post claims it beats closed frontier models on visual grounding and posts strong results across cybersecurity, legal, and agentic behavior. It is available via API now, with open weights due at the end of October.

    Image from @sophiamyang's post
  5. Allie K. MillerXAI score22

    Users combine personal AIs for group collaboration and delegation

    AIAllie K. Miller argues that collaboration between people's AIs is an underappreciated feature, with users combining their AIs, delegating across them, and having them sort tasks out. She says this multiplayer AI is already happening, and that Instinct has since added the ability to put a personal Instinct into a group text.

    Image from @alliekmiller's post
  6. NVIDIA BlogOfficialAI score32

    Telecom Operators Build AI Strategies on Open Models, Citing Control and Customization

    AITelecom operators are building AI strategies on open models for reasons beyond cost, including control, customization, and trust across workloads from autonomous networks to customer care. NVIDIA's State of AI in Telecommunications report found 89% of respondents say open source models and software are important to their company's AI strategy. The NVIDIA Nemotron family offers open weights, training data, and recipes, and the 30-billion-parameter Nemotron 3 Large Telco Model was fine-tuned by AdaptKey on open telecom datasets.

  7. Guillermo RauchXAI score5

    AI progress could yield GTA 7 before GTA 6 ships

    AIGuillermo Rauch jokes that current AI development pace could produce GTA 7 before GTA 6 is released. The post is a lighthearted remark with no specific models, figures, or announcements.

  8. ChinaTalkBlogAI score33

    Bharat Patel on why data, not models, is the hard part of military AI

    AIAccenture defense AI lead Bharat Patel argues that data quality depends on the use case and that "AI-ready data" is a myth. He cites Project Maven, which began in 2017, where early imagery lacked relevant targets and models underperformed until teams continuously collected targeted data. The conversation also covers why fully autonomous tanks remain distant and the risks of data poisoning.

  9. O'Reilly RadarBlogAI score62

    O'Reilly Radar Trends for October 2026: Models, Agents, and Security

    AIThe roundup covers September 2026 AI developments, including model price cuts and new specialized models from Anthropic, OpenAI, Google, and others. It also tracks agents delegating work to other agents, security incidents involving AI agents, and the author's warning that adopters must remain accountable for what their agents do.

  10. Rest of WorldNewsAI score42

    China leads global research in nearly 90% of key technologies, challenging U.S. dominance

    AIChina now leads research in nearly 90% of 74 critical technologies, according to the Australian Strategic Policy Institute's December 2025 Critical Technology Tracker. China also produces 70% of the world's EVs, 80%–85% of global solar photovoltaic manufacturing, and over 75% of battery production. The report measures cited research rather than deployable manufacturing, a gap the article flags as a key caveat.

  11. Gergely OroszXAI score18

    LinkedIn Posts Increasingly AI-Generated, Says Gergely Orosz

    AIGergely Orosz says over 90% of LinkedIn content now appears AI-generated, including comments and summaries of well-written articles. He argues LinkedIn partly brought this on itself by adding a write-with-AI button about 12 months ago, possibly to boost post numbers.

  12. Lucas Beyer (bl16)XAI score10

    Lucas Beyer questions crediting a human "meatproxy" in a math paper

    AILucas Beyer, who says he is not a mathematician, finds it bizarre that an arXiv paper credits a "meatproxy" and lists an author "in personal capacity." He comments on the awkward form of the attribution while congratulating the authors on the content, which is linked in a reply from @mahdi_tcs_.

    Image from @giffmana's post
  13. Gergely OroszXAI score10

    Pragmatic Engineer receives about 100 non-AI essays on industry change

    AIGergely Orosz says the Pragmatic Engineer received about 100 submissions in a single day for an essay contest on how developers see the industry changing. He praised the human-written, non-AI essays as deeply thoughtful and said many will soon be published for readers.

  14. dexXAI score3

    Dex Horthy says "real slop" has never been tried yet

    AIDex Horthy's post claims "real slop has never been tried," a brief remark with no details or figures. The quoted reply, from @emmanuel_2m, jokingly says they are no longer sending their deck, which suggests the comment is a humorous take on AI-generated or low-effort presentation content.

  15. Harrison ChaseXAI score20

    Harrison Chase praises a take on agent harnesses

    AIHarrison Chase, founder of LangChain, endorsed a post on harnesses with the brief comment "Good take on harnesses." The post, from @zeeg, argues that general coding harnesses like Codex will be superseded by specialized ones and that local models will handle most daily tasks within five years.

Oct 5

Oct 5Mon
  1. dexXAI score14

    Founder pitches for human-in-the-loop AI guardrails draw skeptical feedback

    AIDex Horthy says he repeatedly gets founder requests for feedback on human-in-the-loop notification, guardrail, or audit products, and lessons he learned in late 2024 and early 2025 apply to them. Akio Nuernberger, linked as background, reports receiving multiple monthly inbound messages from such startups without a single Langfuse customer showing interest.

  2. roonXAI score10

    Asking whether California's SB53 whistleblower protections have ever been used

    AIroon asks whether SB53's whistleblower protections have ever been successfully invoked, such as by contacting the California Attorney General, and whether they have achieved anything. The post says the only AI-safety whistleblower-driven government action it can recall is Andy Jassy calling Scott Bessent, which it presents as a joke.

  3. SemiAnalysisXAI score22

    Subscription plan value depends on model and workload credit costs

    AISemiAnalysis argues a subscription's worth cannot be judged as a fixed dollar amount, because monthly payments buy credits that each model and token type consume at different rates. Since credit cost ratios can differ sharply from API price ratios, the API-equivalent value of the same $200/month Claude plan shifts with the model and workload.

    Image from @SemiAnalysis_'s post
  4. Mike KnoopXAI score62

    Dust pretrains transformers with zeroth-order optimization, approaching backprop results

    AIDust is a zeroth-order method that pretrains transformers and sometimes matches or exceeds backprop given large compute. The authors report it is about 1,000 to 10,000x more compute efficient than EGGROLL, the state-of-the-art ES method, for training transformers. The post also cites the gradient-alignment result up to 1B tokens and the virtual population idea for scaling.

  5. Thomas WolfXAI score14

    Thomas Wolf hopes Claude Opus 4.6 stays available for a long time

    AIThomas Wolf, who runs Hugging Face, said he hopes Claude Opus 4.6 remains available for a long time. The post is a brief expression of preference, supported by a quoted post in which David Holz reported that in a self-run "have fun" test across LLMs, Opus 4.6 repeatedly won by imagining brief worlds of contradictions inside falling water droplets, while he felt newer models seemed to have less fun.

  6. Nous ResearchOfficialAI score18

    Nous Research argues AI agents should give users full control

    AINous Research says users should control their agent's models, data, memory, compute location, prompts, tools, and code. The post lists choices such as switching models mid-conversation, running fully offline, and exporting the agent. It frames these freedoms as the standard an agent should meet, calling it "yours."

  7. Harrison ChaseXAI score50

    Cognition's Devin adds "Dreaming" offline memory cleanup, open-sourced as a standard

    AIHarrison Chase praises Cognition's "Dreaming" feature, which lets Devin clean stale memory records and surface latent information offline. He argues agent memory needs an offline cleanup loop rather than only better retrieval, and questions how inferred memories get validated before use. He also welcomes Cognition's plan to release Agent Memory Repo as an open standard.

  8. Ethan MollickXAI score10

    Mollick Criticizes Labs for Deferring AI Policy to Superintelligence

    AIEthan Mollick argues that AI labs deferring hard policy decisions on building AI well by assuming superintelligence will soon arrive and resolve them is a flawed stance. He frames this as a cautionary poem that ends with a warning about what happens if superintelligence does not arrive.

    Image from @emollick's post