Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 7

Oct 7Wed
  1. Amazon ScienceOfficialAI score35

    Amazon's AI smart glasses guide delivery drivers hands-free to doorsteps

    AIAmazon has developed AI-powered smart glasses that guide delivery drivers from their van to the customer's doorstep without using their hands. Amazon applied scientist Yelin Kim will discuss the computer vision and edge AI behind the system at COLM 2026.

    Video from @AmazonScience's post
  2. FireworksOfficialAI score16

    Fireworks invites developers to a Forge talk on building your own evals

    AIFireworks promotes a Forge session in San Francisco on November 3, where Notion's Head of AI, Sarah Sachs, argues teams should build their own evals to judge which new models and agents work for their product. Attendance requires applying through the linked registration page.

    Image from @FireworksAI_HQ's post
  3. catXAI score14

    Cat Wu shares using Claude to find top feature users for feedback

    AICat Wu, who works at Anthropic, says a favorite product manager use case for Claude is asking who used a given feature most last week. She suggests having Claude build an artifact of the top 10 users by usage, then reaching out to schedule 15-minute chats, which she calls the fastest way to get user feedback.

    Image from @_catwu's post
  4. Ethan MollickXAI score60

    Mathematicians react to hundreds of AI-generated proofs released by OpenAI

    AIEthan Mollick shares early first-hand accounts from mathematicians grappling with hundreds of AI proofs released by OpenAI. He highlights problems solved in ways no human has yet understood, raising questions about what it means to know something. The linked Scott Aaronson post quotes a researcher, Dana, describing the proofs as unclear and hard to read without AI help, with some possibly verified by a Lean certificate.

    Image from @emollick's post
  5. Teknium 🪽XAI score28

    Teknium posts a "hello" greeting on X

    AITeknium, a Nous Research affiliate, posted only the word "hello" on X. The quoted post from Nous Research announces a Series B raise to advance Hermes Agent and build a mobile app, with investors including NVIDIA and Samsung Next.

  6. Guillermo RauchXAI score18

    Hardening and optimizing code never ends; know when to stop

    AIGuillermo Rauch argues that any program can be hardened and optimized almost endlessly, which the engineering community will rediscover. He notes that these efforts carry real costs in time, attention, and opportunity, and that agents will keep drilling without knowing when to stop.

  7. Google ResearchOfficialAI score23

    Google Research invites COLM visitors to ContinuousBench walkthrough on DP synthetic data

    AIGoogle Research is hosting a walkthrough at its COLM booth #107 today at 5:00 PM of ContinuousBench, a standardized benchmark for measuring knowledge transfer in differentially private synthetic data. The session, led by Alex Bie, asks whether DP synthetic data preserve actual information or only style. A paper is linked on arXiv.

    Image from @GoogleResearch's post
  8. IThome · AINewsAI score72

    Anthropic releases Claude Haiku 5.5, cutting run costs about 75% from Haiku 4.5

    AIAnthropic released Claude Haiku 5.5, which it calls the fastest, cheapest, and most capable Haiku model so far. On average it costs about 75% less to run than Haiku 4.5, with input at $0.10 and output at $0.50 per million tokens for requests up to 100,000 tokens. Anthropic also cut Sonnet 5.5's cache read price from $0.20 to $0.10 per million tokens, which it says lowers run costs by about 20% on many agent tasks.

    This story has a top pick“Anthropic releases Claude Haiku 5.5 as its cheapest, fastest small model”

  9. Aravind SrinivasXAI score13

    Perplexity Decider demoed in a Millionaire quiz speedrun

    AIPerplexity shared a demo of its Decider model answering a "Who Wants to Be a Millionaire?" speedrun. Sanchit Monga, whose EVE inference stack uses Decider for decision models, called it the best decision model for that task.

  10. TypeSafe AIOfficialAI score25

    Jev-killer OpenAI Decisions API benchmarked against Jev for HiringCafe

    AIThe main post is a short reply saying reports of a company's death have been greatly exaggerated, with no details about products or figures. The background post from @h_nilforoshan reports that OpenAI's Decisions API, billed as a "Jev-killer," was benchmarked against Jev for HiringCafe, which serves 2.5 million users. On the task of scoring job-description relevance from 1 to 10, the author reports OpenAI costing 2x more and performing 5-10% worse.

  11. IThome · AINewsAI score44

    Economist Acemoglu estimates AI will automate only about 5% of jobs within 10 years

    AINobel economist Daron Acemoglu estimates AI could technically automate about 20% of work, but adoption limits will cut actual automation to roughly 5% within 10 years. He said the figure is admittedly only an estimate, noting AI models excel in lab settings but underperform when enterprises deploy them in real environments. Microsoft AI CEO Mustafa Suleyman shared the forecast on X on October 6.

  12. ThariqXAI score20

    Anthropic's Thariq says game dev content creation is booming

    AIThariq says this is an incredible time to be a game dev content creator because many people make games for fun and will pay for it. Background context from Tim Sweeney notes that Fab seller revenue rose significantly in September, as AI acceleration increased content demand and developers used AI assistants to build scenes with Fab assets.

  13. IThome · AINewsAI score75

    OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users

    AIOpenAI announced on X that GPT-6 and Intelligent UI are now rolling out to all ChatGPT users, after GPT-6 Astra, Sol and Luna were previously limited to ChatGPT Work and Codex. Intelligent UI lets GPT-6 combine text, images and interactive elements such as charts, clickable buttons and forms, with a mahjong learning example shown.

  14. DatabricksOfficialAI score36

    Claude Haiku 5.5 launches on Databricks as a Day 0 release

    AIAnthropic's Claude Haiku 5.5 is available on Databricks from day zero, which Databricks calls its cheapest, fastest, and most capable small model. On Databricks' OfficeQA Pro V1 benchmark, it delivers about 15% higher quality than Haiku 4.5 at a fraction of the cost. Users can run it alongside 60+ other models on data already in Databricks, with Unity Gateway handling governance, monitoring, and security.

    Video from @databricks's post
  15. Ars Technica · AINewsAI score46

    Artcraft releases open source clones of Adobe Photoshop, Premiere and other apps built with Claude

    AIDeveloper Brandon Thomas's Artcraft has launched seven open source apps in Rust that aim to replicate the interfaces and tools of Adobe Photoshop, Illustrator, Premiere, Lightroom, After Effects, InDesign, and Acrobat Pro. Thomas said he used Anthropic's Claude Opus 5.5 to build the clean-room replacements, with WebAssembly versions available for browser use. The apps remain in a "super early alpha" state, and commenters have pointed out many current shortcomings.

  16. OpenAI · YouTubeOfficialAI score85

    OpenAI introduces GPT-6 in ChatGPT with Intelligent UI

    AIOpenAI says ChatGPT now offers Intelligent UI, which returns answers with fully interactive user interfaces. GPT-6 is rolling out in ChatGPT, powered by GPT-6 Sol for Plus, Pro, Business, and Enterprise tiers and GPT-6 Luna for Free and Go tiers. The rollout begins globally today in the Chat tab for paid tiers, with Free and Go tiers following tomorrow.

    Why it matters: The source names the new interactive interface and the model tiers that receive it, showing how access differs across ChatGPT plans.

  17. ZDNet · AINewsAI score36

    Managers using AI for performance reviews draws mixed employee reactions, survey finds

    AIA Highwire survey of 1,034 corporate employees found 78% of managers used AI to help draft, edit, summarize, or inform performance reviews in the past year. Fifty-four percent of employees said feedback became more specific and actionable, while 34% found it more generic and 32% said it was less useful. Nearly one in four employees rehearse difficult workplace conversations with an AI tool, according to the survey.

  18. SemiAnalysisXAI score32

    Claude $200 plan offers far more token value than ChatGPT

    AIA $200 Claude subscription can deliver about $12,000 of Opus 5.5 tokens at API pricing, according to comments quoted in the post. The post argues that when comparing against ChatGPT's subscription, the value gap is not close.

    Video from @SemiAnalysis_'s post
  19. eric zakariassonXAI score20

    Grok Bot can search X feedback and propose plans, no connector needed

    AIThe post says users can ask the Grok bot to find all feedback about what they are building, summarize it, and propose a plan to address it, without needing an X account or connector. The background post notes that Grok Bot can now search, read, and monitor X.

  20. Visual StudioXAI score34

    Claude Haiku 5.5 is now available in Visual Studio

    AIAnthropic's lightweight Claude Haiku 5.5 is now available in Visual Studio through the Copilot Chat model picker. The model is designed for fast, high-volume work such as subagents, quick edits, and terminal tasks.

    Image from @VisualStudio's post
  21. Mark ChenOfficialAI score62

    OpenAI rolls out Intelligent UI in ChatGPT, generating custom interfaces for answers

    AIMark Chen recommends trying Intelligent UI, a ChatGPT feature where every completion can create a custom interface. He calls it useful for learning new things. The quoted OpenAI post says GPT-6 and Intelligent UI are rolling out in ChatGPT for everyone, providing interactive answers and on-the-spot tools.

    This story has a top pick“OpenAI rolls out GPT-6 and Intelligent UI to all ChatGPT users”

  22. Noam BrownXAI score46

    LLMs now surpass top human experts on some research problems

    AINoam Brown says LLMs have crossed a threshold by surpassing top human experts on some research problems, a jump that makes the recent surge in math results feel sudden. He expects breakthroughs in other domains to follow as models keep improving, though capabilities remain jagged and often still weaker than humans.

  23. CognitionOfficialAI score26

    Cognition shares a blog post on Claude Haiku 5.5

    AICognition's X post links to a blog post at devin.ai about Claude Haiku 5.5, but the text provides no further details. The post itself offers no benchmark scores, prices, or capabilities to report.

  24. falOfficialAI score46

    fal Launches H3 Max Relight for Changing Video Lighting Without Reshoots

    AIfal introduced H3 Max Relight, a tool that changes the lighting of uploaded videos through a built-in Lighting studio where users pick colors, orbit lights around subjects, and adjust intensity and softness. It preserves the original subjects, motion, camera movement, and audio while relighting every frame, and the post says it is powered by H3 Max, which it calls the #1 model for overall quality, prompt understanding, and aesthetics.

    Video from @fal's post