Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 6

Oct 6Tue
  1. GitHub Copilot ChangelogOfficialAI score32

    Update your IDE to restore Copilot agent activity in usage metrics

    AIGitHub says some IDEs that moved Copilot agent sessions to the Copilot SDK left that activity unattributed in usage metrics, and a fix is rolling out by IDE. Visual Studio Code 1.139.0 and later has the fix now, while Visual Studio 18.12, JetBrains, Eclipse, and Xcode are expected between October and November 2026. Billing is unaffected, and missing data from affected versions cannot be backfilled.

  2. Dongxi NLPXAI score22

    OpenAI releases Openai/math, suggesting verifiable problems are being solved

    AIOpenAI has published a repository called Openai/math, which the author reads as a sign that math problems, or any verifiable problems, are being solved. The author says OpenAI's tools exhausted their Pro token allowance on subagent tests unrelated to their main task, concluding that the work was aimed at verification for its own sake.

    Image from @dongxi_nlp's post
  3. Thomas WolfXAI score22

    Ben Affleck jokes about convolutions and his AI background

    AIThomas Wolf's post is a short, playful reply: "how do you like them convolutions," apparently referencing Ben Affleck's comments on convolutional neural networks. The quoted context reports Affleck describing his Python scripting, understanding of CNNs and tensors, GPU work, and private looks at Google and OpenAI's video models.

  4. Thomas WolfXAI score38

    OpenAI releases new mathematical results from internal frontier model

    AIOpenAI is releasing a broad range of new mathematical results produced by an internal frontier model, with release guidance from the Institute for Advanced Study's Advisory Group on Mathematics and Artificial Intelligence. The results are available on GitHub at openai/math. The post itself is brief and emphasizes the results rather than hype.

  5. Simon WillisonBlogAI score34

    llm-openai-decisions 0.1a0 Adds OpenAI Decisions API Support to LLM Tool

    AISimon Willison released llm-openai-decisions 0.1a0, a plugin that adds OpenAI's new Decisions API to the LLM command-line tool. The plugin supports yes/no, choices, and score question types, and works with the gpt-6-luna decision model, which accepts both text and image input. OpenAI charges 10 cents per million input tokens for gpt-6-luna, while Jev's rate is 4.2 cents per million, and output is not charged.

  6. PyTorch BlogOfficialAI score46

    PyTorch Introduces FBTriton Kernels to Speed Table Batched Embedding Operations

    AIPyTorch's blog describes a Triton-based implementation of Table Batched Embedding (TBE) forward and backward kernels for recommendation-system embedding lookups, which the post says outperforms legacy CUDA kernels on these workloads. On B200, an updated CUDA bounds-check step reaches up to 1.24x speedup on that component, and an optional forward-side preprocessing path cuts combined latency from 79.537 ms to 66.183 ms (−16.8%) on a large configuration.

  7. will depueXAI score62

    Will DePue's list claims AI resolved dozens of famous open math problems

    AIA post by Will DePue titled "Fable 5.1's list" presents 100 mathematical results and says 59% were released today, 87% AI and 13% human. The list includes items attributed to OpenAI, Anthropic, Google DeepMind and human mathematicians, each marked by a colored indicator, and it describes many entries as formalized in Lean or as openai/math family numbers. The post supplies no independent verification of these claims.

    Why it matters: The list catalogs claimed AI-assisted results across famous open problems, with the source's own color codes separating AI-generated items from human ones, useful for gauging how far such claims extend.

    Image from @willdepue's post
  8. will depueXAI score7

    Will Depue says GPT 6 Pro and Fable 5.1 ranked recent discoveries

    AIWill Depue, an OpenAI account, asked GPT 6 Pro and Fable 5.1 to rank all discoveries from the last three years. He color-coded them by origin: human, AI before October 6, and AI from OpenAI's math repo. He said 81% of the listed discoveries were released today.

    Image from @willdepue's post
  9. Alex HeathXAI score42

    Reflection CEO argues only open models let users truly own intelligence

    AIReflection CEO Misha Laskin argues that closed AI models are like renting an apartment, while open models let users own intelligence as AI adoption grows. He says the only way to own intelligence is if it is open. Reflection is preparing to release Beam, its first open-weight model, in a podcast discussion with its co-founders.

    Video from @alexeheath's post
  10. OpenAIOfficialAI score62

    OpenAI releases new mathematical results from an internal frontier model

    AIOpenAI is releasing a broad range of new mathematical results produced by an internal frontier model. The company says it consulted the independent Advisory Group on Mathematics and Artificial Intelligence at the Institute for Advanced Study and drew on its advice and public recommendations for how the results are released. The results are available at

    Why it matters: The release shows how a lab is handling mathematical results from an internal model, following advice from an external advisory group on mathematics and AI.

  11. Teknium 🪽XAI score33

    Hermes Index launches to rank models for Hermes Agent users

    AITeknium announced Hermes Index, which combines scores from the new HermesBench and three other agent benchmarks. The index aims to help Hermes Agent users find the best model at a given time and at a given price point. It was introduced by Nous Research as a way to inform model choice and show labs their performance in Hermes.

  12. Hacker News · AI (150+ points)BlogAI score46

    OpenAI shares AI progress in mathematics research

    AIOpenAI published a post on AI progress in mathematics, according to the Hacker News listing, which shows 1334 points and 1520 comments. The source text provided contains no body content, so no specific model, result, or figure can be confirmed from it.

  13. Nous ResearchOfficialAI score34

    Claude Opus 5.5 tops new benchmark at 63.31 per-task score

    AIClaude Opus 5.5 leads the benchmark with a score of 63.31 at $4.99 per task, ahead of GPT 6 Astra at 56.25 ($11.61) and Sonnet 5.5 at 53.14 ($2.82). At the low end, DeepSeek V4.1 Flash scores 36.91 at $0.26, and Ling 3.0 Flash scores 21.56 at $0.054.

    Image from @NousResearch's post
  14. Replit ⠕OfficialAI score7

    Replit invites SF Tech Week event on AI and creator partnerships

    AIReplit is hosting a San Francisco Tech Week discussion tomorrow with Passion Fruit, ElevenLabs, and Gamma on the creator economy. The panel will cover how AI and tech have changed creator-brand partnerships and what true influence looks like. RSVP is available via a Partiful link.

    Image from @Replit's post
  15. Vaibhav (VB) SrivastavXAI score46

    OpenAI's Decisions API enters public beta with GPT-6 Luna

    AIOpenAI has released its Decisions API in public beta, using GPT-6 Luna to classify text and images, route requests, and score inputs. It is reported to run about 10× faster than the Responses API, starting at $0.10 per million input tokens with no output or cache charges.

  16. GitHubOfficialAI score72

    GitHub rebuilds Git infrastructure to handle agent-scale write volume

    AIGitHub reports that Git events on the platform rose from 218.2 billion to 473.3 billion per month between September 2025 and August 2026. It says agent workloads push write throughput and merge contention beyond what its current replica-based architecture handles well, so it is separating durable storage from compute while GitHub keeps running. The article states internal benchmarks reached up to 35 times higher write throughput.

    Why it matters: The post links rising Git event volume to specific architectural bottlenecks, showing why agent workloads strain write paths and how GitHub plans to separate storage from compute.

  17. AdoXAI score62

    Claude now works inside Google Docs, Sheets, and Slides

    AIand those files can also open inside Claude. In Google Workspace, Claude appears in a sidebar next to the open file, reads its contents, and edits it in place, with the option to approve each edit before it is applied.

    Why it matters: The quoted post shows Claude moving into Google Docs, Sheets, and Slides, with per-edit approval, which matters for anyone editing documents with AI.

  18. Google AntigravityOfficialAI score18

    Google invites users to build in Antigravity today

    AIGoogle Antigravity is promoting its platform, urging users to start building in Antigravity now via a link to antigravity.google. The post gives no further details about features, pricing, or availability.

  19. Google AntigravityOfficialAI score36

    Antigravity builds and tests native Android apps from prompt to phone

    AIGoogle's Antigravity agent can take an Android app from prompt to a real device, using the Stitch MCP and Android CLI plugin. The agent pulls designs, builds native Jetpack Compose components, verifies them in the emulator, and runs the final build on a physical phone.

    Video from @antigravity's post
  20. ollamaOfficialAI score55

    Google DeepMind's EmbeddingGemma 2 is now available on Ollama

    AIOllama announced that Google DeepMind's EmbeddingGemma 2 is now available on Ollama. The author describes it as made for consumer devices and multimodal, and gives the command ollama pull embeddinggemma-2 to download it. The quoted DeepMind post says the model is a natively multimodal open model for on-device embeddings that unifies code, images, audio, and video in a shared space.

  21. falOfficialAI score13

    fal launches Nano Banana 2.1 for text-to-image and image editing

    AIfal announced that Google's Nano Banana 2.1 is available today on its platform for both text-to-image generation and image editing. The post links separate pages for each mode but gives no benchmark, pricing, or capability details.