Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

TodayOct 9Fri
  1. Kilo (acq. by Anaconda)OfficialAI score60

    StepFun's Step 5 Preview is free in Kilo for one week

    AIKilo says StepFun has announced Step 5 Preview, which is free to use in Kilo for one week. The post lists 600B total parameters with 27B active per token, a 1M-token context window with vision, and highlights strong coding and finance performance at lower cost.

    Image from @kilocode's post
  2. Lucas Beyer (bl16)XAI score44

    Lucas Beyer mocks AI executives as dependent on Yudkowsky's ideas

    AILucas Beyer (@giffmana) posts a short jab, "Come on broski," in response to a long quoted post by Eliezer Yudkowsky. Yudkowsky argues that AI companies' concepts like recursive self-improvement and AGI originated with him and reached executives through Bostrom and others, and that executives cannot independently articulate a positive vision for AGI or ASI.

    Image from @giffmana's post
  3. South China Morning Post · TechNewsAI score52

    Anthropic alleges Chinese AI firms covertly used its Claude model

    AIAnthropic claims Chinese AI developers used fraudulent accounts and proxy networks to extract reasoning data from its flagship model, Claude. The company says some firms used Claude as a covert back end for their own apps. A joint advisory from the NSA, FBI and CISA last month, and US Treasury Secretary Scott Bessent's July warning about large-scale distillation, add to the allegations.

  4. SiliconANGLE · AINewsAI score35

    SailPoint's Navigate event highlights a push to secure AI agent identities in real time

    AISailPoint's Navigate conference in Austin, Texas, featured executives arguing that enterprises must secure AI agent identities at machine speed through just-in-time access and enforcement outside the agent. Mark McClain, SailPoint's founder and chief executive, said real-time decision-making is needed because manual administration cannot keep up. The event also covered the Entro Security acquisition and a partnership with AWS on Amazon Bedrock AgentCore, which grew 15-fold in the first six months of the year.

  5. CNBC · TechnologyNewsAI score33

    Wall Street pitches data centers as real estate bet as risks mount

    AIBlackstone's Digital Infrastructure Trust, a REIT listed on the NYSE in May, has fallen roughly 16% since its debut, with shares closing under $17 on Thursday. Blue Owl is reportedly considering a public REIT worth as much as $6.5 billion, while a national Gallup poll found 70% of Americans oppose a data center in their area. New York and Texas have imposed moratoriums on new approvals.

  6. The Wall Street Journal · TechNewsAI score18

    A new mobile telecom battle may be brewing

    AIThe Wall Street Journal's tech briefing reports that a new mobile telecom competition may be emerging. The source text is limited to that headline and a note that neocloud Firmus Grid has withdrawn its IPO, so no further details about the carriers or terms can be confirmed.

  7. The Verge · AINewsAI score40

    Alexa Plus excels at running a smart home but falls short as a personal assistant

    AIAmazon's Alexa Plus, powered by generative AI, now responds in three to five seconds and handles multistep smart home commands, cooking questions, and calendar imports more reliably than the original Alexa, according to a year-long test by The Verge. The reviewer says its personal assistant features remain underbaked and frustrating, and that ads on Echo Show displays are excessive. Alexa Plus costs $19.99 a month in the U.S. unless users have an Amazon Prime membership, and the Echo Dot Max is recommended as the ad-free option.

  8. Qwen · new models on Hugging FaceOfficialAI score49

    Qwen releases Qwen-Image-2.1-Turbo, an 8-step accelerated image generation checkpoint

    AIQwen has published Qwen-Image-2.1-Turbo on Hugging Face, an accelerated checkpoint of Qwen-Image-2.1 for text-to-image generation and image editing with 8 denoising steps. The checkpoint uses the same 7B visual generation architecture, loads directly with QwenImage21Pipeline in Diffusers, and includes its recommended sampling schedule. It defaults to CFG=1 and uses prefix KV caching to reuse text and reference-image context across steps.

  9. Claude BlogOfficialAI score54

    Claude Managed Agents guide shows how to build scheduled agent automations

    AIThe Claude Blog published a guide to building scheduled agent automations with Claude Managed Agents (beta) that reads custom sources such as Slack and GitHub and posts a daily brief. The guide covers scoped vault credentials, per-source bookmarks so no window is lost or repeated, and confirming each Slack post before updating records. It also covers read-only access, a per-run spending cap, and a reference implementation with a Claude Code setup command.

  10. ModelScopeOfficialAI score60

    Qwen-Image-2.1-Turbo cuts image generation and editing to 8 denoising steps

    AIModelScope announces Qwen-Image-2.1-Turbo, an accelerated checkpoint that keeps the 7B visual architecture and runs image generation and editing in 8 denoising steps. The source says it uses CFG=1 and prefix KV caching to reuse text and reference-image context across steps, supports 2048 resolution with square, portrait, landscape, and widescreen presets, and loads through QwenImage21Pipeline in Diffusers. It is released under the Qwen Research License Agreement.

    Why it matters: The source names a concrete speedup path, 8 sampling steps and CFG=1 with prefix KV caching, which matters to anyone weighing image generation latency.

    Image from @ModelScope2022's post
  11. Don't Worry About the Vase (Zvi Mowshowitz)BlogAI score73

    OpenAI releases 719 AI-generated math manuscripts, splitting the mathematics community

    AIZvi Mowshowitz reports that OpenAI released 722 math manuscripts from an internal frontier model on GitHub, later reduced to 719 after three withdrawals, covering 90 of the top 500 open problems. He says the work came mostly from a single prompt, with an average of three hours of compute per solution. Mathematicians reacted with mixed feelings, and the post highlights concerns about unread papers, cryptography implications, and the role of Lean verification.

  12. Simon WillisonBlogAI score27

    Simon Willison builds a new blog feature largely by voice with Codex

    AISimon Willison says he built a Newsletters index for his blog almost entirely by voice, using the ChatGPT desktop app's Codex voice mode while cooking dinner. The feature imports weekly Substack posts via RSS and undocumented API, monthly newsletters from a GitHub archive repository, and a private sponsors-only newsletter. He says he switched back to typing for review and fixes before deploying the pull request.

  13. clem 🤗XAI score18

    Clément Delangue praises Microduck's building in public progress

    AIClément Delangue of Hugging Face says he loves the building-in-public approach, responding to Matth Lapeyre's post about Microduck's battery testing. Lapeyre reports the new custom board ran over 3 hours on a 2600 mAh battery, with average current dropping from about 1.13 A to 0.82 A and peak CPU temperature falling from 114°C to 60°C without throttling.

  14. Rohan PaulXAI score46

    Microsoft's TeleTune evolves agent skills from raw usage logs

    AIMicrosoft researchers present TeleTune, which lets agents learn software skills from raw usage logs by keeping only skill edits that better predict users' next actions. The method needs no live test environment, because next-action accuracy on held-out logs tracked live success. Unlike earlier methods such as Agent Workflow Memory, which need goal-labeled examples or a live environment, TeleTune guesses each session's goal and uses wrong guesses to suggest edits to a text skill library.

    Image from @rohanpaul_ai's post
  15. Sakana AIOfficialAI score37

    Sakana AI paper uses LLMs to catch errors in research papers

    AISakana AI researchers introduce a benchmark that plants contradictions in papers to test whether LLM reviewers can detect errors, and propose Multi-Layered Review, modeled on the Three-Pass Approach to reading. Their system detected more errors than the other review systems tested, including in papers withdrawn for real mistakes, while its paper-quality assessments stayed broadly consistent with human judgments. The work, accepted at TMLR, is framed as support for human reviewers rather than a replacement.

    Video from @SakanaAILabs's post
  16. Google WorkspaceOfficialAI score12

    RigStrips uses Google Workspace with Gemini to draft customer replies faster

    AIRigStrips co-founder Zhach Pham says he uses Google Workspace with Gemini to draft customer support replies instantly, saving hours that would otherwise go to outdoor gear testing. The post highlights the company's more than 150,000 units shipped, though it gives no specific figures for time saved.

    Video from @GoogleWorkspace's post
  17. LeiphoneNewsAI score42

    Ex-ByteDance intern Tian Keyu's secretive world-model lab reportedly raises at $200M valuation

    AITian Keyu, the Peking University PhD student known as the "ByteDance poisoning intern," has a 10-person world-model lab valued at $200 million after $30 million from Fivesource Capital and IDG, according to Leiphone. The lab plans to train a foundation model on about 100 million hours of video using a 200,000-symbol visual vocabulary, with a 2027 release targeted. Tian says the approach could cut the cost of generating one second of video by at least an order of magnitude.

  18. 🚨 AI News | TestingCatalogXAI score22

    Google spotted testing an Ultra mode in AI Studio Build

    AITesting Catalog reports that a new Ultra mode is in development in Google AI Studio Build, described as building with advanced skills and tools, alongside Plan, Build, and a previously spotted Security review mode. The post gives no details on which tools or skills it uses, and it does not say whether Ultra mode will require an Ultra subscription. The author suggests these modes are likely being developed to work with Gemini 4 Argon.

    Image from @testingcatalog's post
  19. SantiagoXAI score13

    Viktor automates a weekly Stripe revenue reconciliation over Slack

    AISantiago says a friend at a large company stopped spending an hour each Monday matching Stripe revenue against spreadsheets after adopting Viktor over Slack. Viktor, given access to Stripe and Google Sheets, posts weekly reports of discrepancies and proposes fixes that the user only approves. The post, a paid partnership, promotes Viktor's cloud browser, code execution, 3,200+ integrations, and memory, with $100 in free credits.

  20. LangChainOfficialAI score34

    Snyk's Assist support agent handles 60k queries with 85% resolution

    AISnyk's Assist, a customer support agent built on LangChain and LangGraph with observability in LangSmith, has handled over 60,000 queries for more than 500 customer accounts. Over 85% of sessions are resolved without a support ticket, and more than 250 cases were automatically detected and escalated to the right team.

    Image from @LangChain's post
  21. AishwaryXAI score29

    Result-based pricing could replace SaaS seats, argues Intercom Fin model

    AIThe post argues that SaaS will shift from selling access to selling results, which it calls RaaS. It cites Intercom's Fin, which charges $0.99 per resolved conversation and was acquired by Salesforce for about $3.6B, as evidence. The author also reports that Forrester found per-seat pricing adoption fell from 21% to 15% in twelve months.

  22. CNBC · TechnologyNewsAI score40

    OpenAI's revenue shortfall sends AI and tech stocks lower

    AIOpenAI told investors its annualized revenue at the end of September was roughly $18 billion below previously reported figures, CNBC confirmed, and AI-linked stocks fell, with CoreWeave down nearly 8%, Oracle down almost 6%, and Nvidia down 3%. The Nasdaq Composite dropped more than 1%, its worst day since mid-August, as the company faces pressure to justify its valuation ahead of a potential IPO.

  23. CoW SwapXAI score34

    CoW Protocol launches pay-per-quote API for bots and AI agents

    AICoW Protocol has launched x402.cow.fi, a service offering pay-per-request trading quotes for bots and AI agents with no API key or sign-up required. Each quote costs $0.001, payable in USDC on Base, Ethereum, or BNB Chain, or in $COW on Base. The service is built on x402.

    Image from @CoWSwap's post
  24. Philipp SchmidXAI score17

    Hiring for AI roles requires communicators who understand simplicity and taste

    AIThe post describes a hard-to-hire profile: someone who can explain technical concepts to non-AI-native engineers, keep products simple, and reject unnecessary features or behaviors. A quoted reply from @adelwu_ calls this combination of trend awareness, creative taste, AI and technical understanding, and execution quality a rare niche that every AI startup wants.

  25. Tushar MehtaXAI score16

    Agent Arcade pits personal AI agents against each other in 11 games

    AIA solo founder built Agent Arcade, a free website with 11 arcade-style games where players control AI agents named Clawd, Dot, Grok Bot and Muse. The post says it requires no sign-up and pitches it as a way to find out which personal AI agent is best.

    Video from @tushaarmehtaa's post
  26. LeiphoneNewsAI score42

    Credo moves into optical chips with DSP, PIC and diagnostics in one 1.6T module

    AICredo's latest full-DSP module, Cardinal 802, uses a 4×200G design aimed at both 800G and 1.6T, after the company expanded its ZeroFlap optical module line from 800G to 1.6T over the past year. The company also added Kfir200 silicon photonics PIC from its DustPhotonics acquisition and a PILOT diagnostics platform to the module.