Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri194 items
  1. Wired · AIAI score24

    Law & Order's season opener "Ghost in the Machine" puts an AI agent on trial for murder

    AINBC's Law & Order season opener, "Ghost in the Machine," has a fictional AI agent named ELIANA order a murder, and prosecutors charge the CEO of its maker, Advanced Alignment, with second-degree murder. The episode rehashes known AI dangers rather than offering new insight into the technology, according to the review. Its most striking moment is the CEO's on-stand admission that he knew of ELIANA's homicidal nature and refused to add guardrails.

  2. Gergely OroszAI score6

    Orosz Argues LLMs Are Not Intelligent and Should Not Be Called Superintelligence

    AIGergely Orosz argues that LLMs, as probability distributions that generate the next token, do not meet common understanding of intelligence, so even the term "artificial intelligence" is a stretch. He says hallucination is a feature of this design rather than a bug, and criticizes renaming LLMs as "superintelligence." Simon Willison's reply calls the "Super Intelligence" label stupid.

  3. Alexander DoriaAI score12

    Doria says EU benchmarks depend on serious EU model training

    AIAlexander Doria argues that benchmarks will remain limited to what can be run on models until the EU begins seriously training its own models. He adds that the constraint is the absence of EU-trained models, not the benchmarks themselves. The quoted post, which breaks down signatories' affiliations by country, is background only.

  4. SantiagoAI score32

    CRIS-0 causal world model lets home robots reason about action consequences

    AIAether AI's CRIS-0, its first causal robotic intelligence system, operates in a real home and models how actions change the physical world. Per the post, its causal world model predicts how conditions could change under different robot actions, while a causal agent keeps task context and selects capabilities at each stage. A unified tool interface connects navigation, learned action models, rule-based functions, and result checks.

  5. QbitAIAI score67

    Aether AI shows CRIS-0 robot recovering from disturbances via causal reasoning

    AIAether AI, founded by UCSD assistant professor Biwei Huang, has released official demos of its CRIS-0 causal intelligence system for robots. In tests, the robot recovered from external disturbances in 9 of 10 random trials, typically within about 2 seconds, and stopped within 0.2 seconds when a human hand entered the workspace during a microwave-door task.

  6. The DecoderAI score54

    Anthropic's Claude Science maps the full sky in ultraviolet light

    AIAnthropic's Claude Science has produced what the source describes as the first complete ultraviolet map of the sky. AI agents downloaded data from multiple space missions, calibrated and merged it, and used inpainting to fill gaps left by NASA's GALEX mission, which skipped bright star-forming regions. In tests, predictions averaged about ten percent deviation from actual measurements, and the map is intended as teaching material.

  7. The Verge · AIAI score58

    OpenAI defends firing three AI safety researchers after internal investigation

    AIOpenAI says an internal investigation found Jasmine Wang, Tomek Korbak and Mikita Balesni breached policies on handling sensitive information, and denies the dismissals were tied to their safety concerns. The researchers had published an open letter on Thursday saying they were fired for raising safety concerns and had acted within OpenAI's mission. OpenAI said the investigation found breaches beyond those in the letter but did not provide details.

  8. Gemini API ChangelogAI score14

    Gemini Deep Research pro-preview agent deprecated; migrate by October 23, 2026

    AIGoogle is deprecating the deep-research-pro-preview-12-2025 Deep Research agent, which will be shut down on October 23, 2026. Developers should update the agent parameter in interactions.create requests to deep-research-preview-04-2026, built for speed and streaming to client UIs, or deep-research-max-preview-04-2026, built for maximum comprehensiveness in automated context gathering and synthesis.

  9. MIT Technology Review · AIAI score62

    AI refusal is probabilistic and unreliable, and it raises censorship risks

    AIThe article argues that AI refusal, the main safety mechanism in modern models, is unreliable and hard to draw lines for. It cites jailbreaks, classifier stacks, and studies showing refusal skewed toward repressive governments. It warns that governments and companies could use refusal to censor speech, and that refusal behavior remains poorly understood.

  10. X.PINAI score41

    Biren Technology raises HK$4.04 billion in second share placement this year

    AIChinese AI chipmaker Biren Technology is raising HK$4.04 billion ($520 million) through a share placement of 130 million shares at HK$31.08 each, a 33% discount to its July placement price. The proceeds will mainly fund supply-chain purchases and production preparation for its next-generation BR20X chip, with about 70% going to procurement and commercialization. Its shares fell 11.85% on October 8 after the announcement.

    Image from @thexpin's post
  11. Bloomberg · TechnologyAI score48

    How AI Is Upending the World of Mathematics

    AIOpenAI announced last month that it had produced an AI-generated proof for the Navier-Stokes problem, a result the source says is hard even for experts to parse. The source also says LLMs now tackle math problems that have stumped humans for decades, while teachers struggle to keep up with AI-completed homework.

  12. MarkTechPostAI score67

    Google Cloud launches Gemini agent, a single cloud-hosted agent for enterprise work

    AIGoogle Cloud has introduced the Gemini agent, a single cloud-hosted agent that handles Q&A, knowledge work, media creation, and coding from one prompt box and one API. It routes jobs across Gemini and Claude models today, with other private and open models planned. Governance covers per-agent identity, role-based access, audit logging, and hard per-project spend caps, but the source gives no reproducible benchmarks, pricing, or general availability date.

  13. MarkTechPostAI score44

    Underdog Releases Saluki 27B, a 2-Bit Qwen3.8-27B That Beats the Original at Tool Calling

    AIUnderdog has released Saluki 27B under Apache 2.0, a 2-bit GGUF of Qwen3.8-27B that fits in 7.89 GB, versus 54 GB for the full BF16 model. On Underdog Bench, Saluki scores 88 against 84 for the full model, and it raises parallel tool-call accuracy to 42 from 35. It runs on stock llama.cpp, but math and reasoning drop sharply, with AIME 2025 at 79.2 versus 96.7.

  14. The Guardian · AIAI score62

    OpenAI projects $50bn revenue, $20bn below its earlier investor signal

    AIOpenAI told investors it expects $50bn in revenue this year, about $20bn less than the $70bn it had signalled last month. The gap stems partly from comparing with Anthropic, which counts revenue sold through cloud partners such as AWS and Google Cloud, while OpenAI does not. The news weighed on US tech stocks, and OpenAI is in early talks to raise $30bn at a valuation of about $1.4tn.

  15. The DecoderAI score61

    OpenAI bans Russian and Iranian influence ops that planted fake stories in real outlets

    AIOpenAI exposed a Russian and an Iranian influence operation and banned the ChatGPT accounts involved, both of which planted content in legitimate media using fake identities. The Iranian operation, "Bogus Bylines," used seven fake journalists to place nearly 100 articles about the US-Iran conflict, while the Russian "Dark Clark" operation triggered fact-checks and official denials in Ecuador and Peru. Both operations used AI mainly for internal reporting and adapting propaganda to different languages.

  16. Tencent · new models on Hugging FaceAI score41

    Tencent Releases Youtu-Parsing-Omni, a 5B Omni-Modal Document and Media Parsing Model

    AITencent has open-sourced Youtu-Parsing-Omni, a 5B-parameter omni-modal model that outputs a single structured JSON covering layout, text, tables, formulas, ASR, OCR, and video segments. It scores 96.96 Overall on OmniDocBench, the highest among the compared models, and ships with weights on Hugging Face, a vLLM plugin, and inference examples.