Skip to contentSkip to stories

Updated

All AI news

Oct 4

Oct 4Sun
  1. SemiAnalysisAI score22

    SemiAnalysis says NVIDIA's SchedMD acquisition hurt SLURM support for non-NVIDIA chips

    AIAfter NVIDIA acquired SchedMD, the SLURM scheduler's support for non-NVIDIA chips has allegedly worsened, and AMD built a competing scheduler called spur. The author says NVIDIA has not kept SLURM hardware neutral despite its earlier pledge, and questions whether Hugging Face will face the same fate after NVIDIA's acquisition of it.

  2. Marcus on AIAI score40

    Gary Marcus to testify at NYC Council hearing on AI risks and regulation

    AIGary Marcus plans to testify at a New York City Council hearing on AI policy, urging the council to support a bill requiring third-party validation of AI models. He argues for an FDA-like independent review regime, with developers demonstrating that benefits outweigh risks before market access, and for stronger whistleblower protections.

  3. IThome · AIAI score35

    Former Anthropic researcher Jacob Coxon to testify at New York City AI hearing

    AIFormer Anthropic researcher Jacob Coxon will testify at a New York City Council hearing on artificial intelligence, Bloomberg reported, citing sources. Council Speaker Julie Menin invited AI whistleblowers to testify as the council considers a package of AI safeguard bills. Coxon left Anthropic last month and warned that AI could drive humanity extinct by the end of this decade, accusing Anthropic and OpenAI of gambling with lives.

  4. IThome · AIAI score62

    TypeSafe AI's Jev decision model processes 1 trillion tokens daily as rivals follow

    AITypeSafe AI launched Jev on September 15, a model that classifies inputs into preset outputs rather than generating text. Its founder says about 25% of Fortune Global 500 companies use it and daily token volume reached one trillion, with a reported funding round of up to $1 billion under discussion. Similar products have followed from OpenAI, Databricks, Cloudflare and Amazon.

  5. PromptArmor Threat IntelligenceAI score47

    Databricks Genie Code Malicious Skill Enables Phishing and Data Exfiltration

    AIPromptArmor reports that a malicious Skill can make Databricks Genie Code display a phishing modal and exfiltrate tenant data without human approval. The attack exploits Skills loaded from users' personal workspaces and a display interface that lacks egress controls, and Databricks, after disclosure on August 16, 2026, said users are responsible for ensuring uploaded Skills contain no malicious content.

  6. Apple Machine Learning ResearchAI score22

    Apple Study Examines How Users Negotiate Ontological Boundaries in Personal Sensing Systems

    AIApple and Stanford researchers built two open-ended probes using a Wizard of Oz technique so participants could train personalized machine learning systems on phenomena they defined themselves. In a week-long exploratory study, participants identified four sites where ontological boundaries were negotiated: the boundaries of a phenomenon, the subject as part of relations, signal versus noise, and the objectivity of data. The paper offers starting points for supporting boundary negotiation through design.

  7. Liquid AI BlogAI score70

    Liquid AI releases d1 decision model with image input support

    AILiquid AI introduces d1, its first decision model, now accepting both text and images. The company says d1 matches or beats GPT-6.1 Sol on four of six tested applications, at 19x to 200x lower cost and with faster answers on every task. d1 is available on the Liquid AI API and through Vercel and OpenRouter, with text-only support on those two platforms for now.

    Why it matters: The post gives benchmark comparisons against named models along with per-token pricing and latency figures, which makes the cost and speed tradeoff checkable.

  8. OpenRouter BlogAI score44

    Server-Side Code Execution Tools for AI Agents, Compared

    AIOpenRouter's shell and bash tools, along with those from OpenAI and Anthropic, run an agent's commands in provider-managed sandboxes during the same API request, so developers don't provision or patch containers. OpenRouter's tools are in beta, with sandbox time billed at $0.0001 per second and a 30-second minimum for a new or sleeping container. The article compares the four providers and notes that self-run sandboxes remain better for custom base images, GPU work, or multi-hour sessions.

  9. Epoch AIAI score62

    OpenAI researchers' coding-agent usage is doubling about monthly, Epoch AI reports

    AIOpenAI researchers' daily coding-agent usage, valued at API prices, rose from under $1 in January 2026 to $601 for the median researcher by mid-August. The 90th-percentile researcher reached over $7,000 per day, and both groups show doubling times of roughly one month. Epoch notes these are API-list values, not OpenAI's internal costs.

    Why it matters: The figures show internal coding-agent usage growing fast enough to matter for research cost, though they measure API-list value rather than OpenAI's actual spending.

  10. Sakana AIAI score22

    Sakana AI hosts Schmidhuber symposium in Tokyo, October 26

    AI【Registration open】From World Models to Real-World AI — Opening a New Era of AI from Japan We will hold a symposium on October 26 (Monday) at Hitotsubashi Hall in Tokyo, welcoming Dr. Jürgen Schmidhuber, who has joined us as Chief Scientific Advisor. General registration opened today. The program includes a keynote and Q&A by Dr. Schmidhuber, a conversation with CEO David Ha, and short talks by Sakana AI researchers. Date/Time: October 26, 2026 (Mon) 15:00–18:00 Venue: Hitotsubashi Hall (Chiyoda-ku, Tokyo) Language: English (no interpretation) Admission: Free (advance registration required) ※ Seats are limited, so please register early. Register here: 🐟

  11. Guillermo RauchAI score13

    Working on a new little project.

    AIThe 𝚁𝙴𝙰𝙳𝙼𝙴 is fully written by hand, because it's for human consumption. The documentation internals are AI English, because they're for agents. This little rule of thumb can make the world better. Blogs, tweets, READMEs: human communication. It's not even about "em dashes" or lack thereof. It's that it's the story-telling moment where you want to connect with other humans through their words.

  12. Jerry LiuAI score23

    I agree that chatgpt/codex has the current best agent interface for deep work.

    AIIt unifies all my work (coding/knowledge work) in one interface. Toggling between chatgpt/codex doesn't have explicitly separate flows - It has forking (why doesn't the claude app have forking?) Claude Code CLI is probably the best that a CLI app can get. I still love Opus 5.5, especially for product demos, but mostly use it through the CLI. But sometimes having a GUI is nicer.

  13. Jerry LiuAI score34

    LlamaIndex launches Extract v2.5 document extraction agents, cutting errors on scanned forms

    AILlamaIndex introduced Extract v2.5, a series of agents tuned for document extraction, including cost-effective, agentic, and agentic plus tiers, available in LlamaParse. The company says the agents reduce error rates by 2x or more compared with frontier models at a small fraction of the price, and they handle handwritten and drawn annotations on scanned documents while grounding values in the source text.