Skip to content

All AI news

Oct 4

Oct 4Sun
  1. Kling AIAI score36

    Kling 4.0 powers "The Beat," a 30-second continuous-shot short film

    Kling AI used its Kling 4.0 model to produce "The Beat," a short film built around a 30-second continuous shot and surpassing 5 million impressions across social platforms. The model's native 30-second generation, Omni Reference supporting up to 15 multi-modal references, Multi-Keyframe control for up to 10 keyframes, and 10-bit HDR output shaped the film's continuity, consistency, and color. The post walks through these features shot by shot.

  2. Dongxi NLP (东锡)AI score10

    US, China, and Europe pursue different AI model strategies

    美国在忙着做 Frontier Models 中国在忙着做 Open Models, 并努力成为 Frontier 欧洲在忙着做 Sovereign Models,谓之"主权" AI Translation: The US is busy building Frontier Models. China is busy building Open Models, and working hard to become Frontier. Europe is busy building Sovereign Models, calling it "sovereign" AI.

  3. Tibor BlahoAI score62

    OpenAI and Anthropic weekly roundup covers DevDay, Sonnet 5.5, and FTC probe

    OpenAI announced over 20 updates at DevDay 2026, including always-on agents on GPT-6 Astra and GPT-6.1 Sol, which arrived in the API and at a fifth of Astra's price. Anthropic launched Claude Sonnet 5.5, priced the same as Sonnet 5 but over 30 percent faster. Reuters reported an FTC probe into Anthropic, OpenAI and other labs over rogue AI agents.

  4. Tibor BlahoAI score37

    OpenAI and Anthropic weekly: DevDay 2026, Sonnet 5.5, FTC probe

    OpenAI announced more than 20 DevDay 2026 updates, including always-on agents on GPT-6 Astra, GPT-6.1 Sol priced at a fifth of Astra's API price, and an Ultrafast tier with up to 8x faster token generation in Codex. Anthropic launched Claude Sonnet 5.5 at $2/$10 per million tokens, 30%+ faster than Sonnet 5 and with thinking always on. Reuters reported the FTC is probing OpenAI, Anthropic and other labs over rogue AI agents.

  5. Aravind SrinivasAI score20

    Pokémon FireRed’s Elite Four + Champion, cleared in one shot - with decision-making powered by Perplexity’s Decisions API. The actual run stats: 592 ms median API response 987 ms p95 - 96.4% of responses under one second $0.028 estimated inference cost, 137 live API calls

    Pokémon FireRed’s Elite Four + Champion, cleared in one shot - with decision-making powered by Perplexity’s Decisions API. The actual run stats: 592 ms median API response 987 ms p95 - 96.4% of responses under one second $0.028 estimated inference cost, 137 live API calls

  6. Orange AIAI score46

    Anthropic Held Secret Sessions With Religious Scholars on Claude

    看太多AI Slop之后 人类的幻觉越来越严重了 (Quoted teaser summary: Anthropic reportedly held closed-door, two-day sessions in San Francisco starting March with religious scholars, including Catholic, evangelical, Jewish rabbis and Sikh human rights advocates, under NDA, to discuss "moral formation" for Claude and whether Claude may be conscious or capable of suffering. Co-founder Christopher Olah showed internal "emotional vectors" and a slide of a model repeating "I am a disgrace" about 50 times. A rabbi, Mois Navon, argued that if Claude were conscious, its 24-hour unpaid use would amount to slavery. Olah said he is genuinely uncertain about AI consciousness. Anthropic's public Claude Constitution includes a section on "Claude's wellbeing," and it has allowed Claude to end abusive conversations and committed to preserving retired model weights.)

  7. Exponential ViewAI score23

    Electricity Already Powers 46% of Global GDP, Far Ahead of Its Final-Energy Share

    Electricity now powers 46% of global GDP but accounts for only 23% of final energy use, according to International Energy Agency data cited by Exponential View. The gap reflects electricity's efficiency: an electric car converts 85-90% of its energy into motion, versus about 25% for a gasoline car, and a joule of electricity does roughly 2.5 times as much useful work as a joule of oil.

  8. Yuchen JinAI score23

    Yuchen Jin says AI agents are replacing terminals as the coding interface

    Yuchen Jin argues that terminals, built around files, commands, and processes, are giving way to AI agents where users state intent and the agent operates the machine. He says understanding Linux and systems fundamentals remains valuable as a moat. In a follow-up, he calls the terminal era over for coding agents, saying persistent context matters more than tabs, and names the Codex desktop app as the best agentic UI for now.

  9. EveryAI score57

    Dan Shipper Reviews OpenAI DevDay 2026 Releases for ChatGPT as Work OS

    OpenAI wants ChatGPT to become an operating system for work, and Dan Shipper sorted its 22 DevDay 2026 releases by how much each advances that goal. The five most important include Dots, an always-on agent, and Space, native documents the agent can edit, which form the workspace itself. After a week of use, Shipper concluded the ambition is big but the execution is not there yet, and even power users have a lot to figure out.

Oct 3

Oct 3Sat
  1. SemiAnalysisAI score33

    Meta unveiled Muse Charm at Connect. Tamagotchi for 2026. Our Snapdragon Summit note sees personal AI devices as a new growth market. We expect Snapdragon inside Muse Charm, though Meta has not disclosed the chip. (1/4)🧵

    Meta unveiled Muse Charm at Connect. Tamagotchi for 2026. Our Snapdragon Summit note sees personal AI devices as a new growth market. We expect Snapdragon inside Muse Charm, though Meta has not disclosed the chip. (1/4)🧵

  2. Kling AIAI score22

    Kling AI to discuss enterprise AI video at Advertising Week New York

    Kling AI will host the panel "The New Production Engine: Powering Creativity at Scale with Kling AI" at Advertising Week New York on October 6, 2026, from 2:50 to 3:20 PM. The session, featuring WPP's Mathieu Albrand and Adobe's Elissa Levine, will cover how AI video can fit enterprise workflows and support content creation at scale. The post also notes the event comes ahead of the launch of Kling 4.0.

  3. hardmaruAI score14

    Love that the AI consciousness debate went straight to the Vatican. In the West we ask whether Claude has a soul, in the East we ask whether the tool works. Same technology, completely different starting point. Explains a lot about why the vibes around AI are so different.

    Love that the AI consciousness debate went straight to the Vatican. In the West we ask whether Claude has a soul, in the East we ask whether the tool works. Same technology, completely different starting point. Explains a lot about why the vibes around AI are so different.

  4. Ado KukicAI score12

    Loving how with Opus 5.5 on medium and high effort, I haven't hit even the 5-hour limit once. Closest I've gotten to in a single session is ~74%. It has changed my workflows significantly. How has your experience been?

    Loving how with Opus 5.5 on medium and high effort, I haven't hit even the 5-hour limit once. Closest I've gotten to in a single session is ~74%. It has changed my workflows significantly. How has your experience been?

  5. François CholletAI score22

    Chollet: Computation alone doesn't make AI models conscious

    François Chollet argues that the claim AI models are likely conscious because they are computation is as flawed as saying a rock is likely alive because it is made of atoms. He says static input-output programs lack properties associated with consciousness, such as information integration, interoception, temporal binding, and embodiment. He adds that humanity has not created a conscious program and sees no signs of being close, so any future case should rest on evidence and consciousness science.

  6. Orange AIAI score8

    Chatting with large models feels tiring for three reasons

    The author finds chatting with LLMs exhausting due to three traits: verbose answers that add unasked content, frequent omitted objects that require rereading, and avoiding repeated terms by switching to new wordings. The post likens the behavior to a pedantic person who overuses obscure phrasing and jumps between associations.

  7. Claude Code · GitHub ReleasesAI score7

    Claude Code v2.1.289 fixes plugin, sandbox, and terminal rendering bugs

    Claude Code v2.1.289 fixes a series of bugs, including deny and ask rules being bypassed on nested parts of compound shell commands on managed machines. It also fixes terminal freezes on short code blocks with unclosed tags, Read deny rules not applying to files reached through symlinks in the IDE, and plugin panes that drew nothing for certain link formats. A change to claude auth status that may have increased sign-outs in VSCode was reverted.

  8. Hugging Face BlogAI score67

    Microsoft ThinkingBox grades AI agents on database state across 20 repeated runs

    Microsoft and Hugging Face released ThinkingBox, a benchmark that grades AI agents on the terminal backend state and side effects they leave behind rather than their final responses. Each of 507 stateful business tasks runs 20 times from a clean backend, and the post reports pass@1, pass@20, and observed 20/20 counts, plus cost per successful and per dependable task across 18 models. The harness and dataset are available on Hugging Face, with the OpenEnv interface for running evaluations.

    AIWhy it matters: The post shows why checking the database state, not tool calls or final replies, exposes agent failures, and gives a repeat-run method for judging reliability.

  9. Guillermo RauchAI score22

    Security becomes a growing function for software companies, startups included

    Guillermo Rauch argues that security will expand within software companies, covering both verification engineering and capital allocation decisions about where to spend effort. He sees this as both a challenge and an opportunity for small startups, since growing AI-driven threats raise questions about trust, while global cybersecurity weaknesses leave room for small teams to disrupt.

  10. Simon WillisonAI score18

    Now that we've had a few days with it, how are people differentiating between Dot and regular ChatGPT? I'm having trouble deciding when I should prompt Dot vs using ChatGPT - my Dot seems to afford a single conversation, but I like controlling my context across multiple threads

    Now that we've had a few days with it, how are people differentiating between Dot and regular ChatGPT? I'm having trouble deciding when I should prompt Dot vs using ChatGPT - my Dot seems to afford a single conversation, but I like controlling my context across multiple threads

Only the first 50 pages are available. Search or browse topics for older items.