Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Oct 9

Oct 9Fri
  1. South China Morning Post · TechNewsAI score52

    Anthropic alleges Chinese AI firms covertly used its Claude model

    AIAnthropic claims Chinese AI developers used fraudulent accounts and proxy networks to extract reasoning data from its flagship model, Claude. The company says some firms used Claude as a covert back end for their own apps. A joint advisory from the NSA, FBI and CISA last month, and US Treasury Secretary Scott Bessent's July warning about large-scale distillation, add to the allegations.

  2. SiliconANGLE · AINewsAI score35

    SailPoint's Navigate event highlights a push to secure AI agent identities in real time

    AISailPoint's Navigate conference in Austin, Texas, featured executives arguing that enterprises must secure AI agent identities at machine speed through just-in-time access and enforcement outside the agent. Mark McClain, SailPoint's founder and chief executive, said real-time decision-making is needed because manual administration cannot keep up. The event also covered the Entro Security acquisition and a partnership with AWS on Amazon Bedrock AgentCore, which grew 15-fold in the first six months of the year.

  3. The Verge · AINewsAI score40

    Alexa Plus excels at running a smart home but falls short as a personal assistant

    AIAmazon's Alexa Plus, powered by generative AI, now responds in three to five seconds and handles multistep smart home commands, cooking questions, and calendar imports more reliably than the original Alexa, according to a year-long test by The Verge. The reviewer says its personal assistant features remain underbaked and frustrating, and that ads on Echo Show displays are excessive. Alexa Plus costs $19.99 a month in the U.S. unless users have an Amazon Prime membership, and the Echo Dot Max is recommended as the ad-free option.

  4. Qwen · new models on Hugging FaceOfficialAI score49

    Qwen releases Qwen-Image-2.1-Turbo, an 8-step accelerated image generation checkpoint

    AIQwen has published Qwen-Image-2.1-Turbo on Hugging Face, an accelerated checkpoint of Qwen-Image-2.1 for text-to-image generation and image editing with 8 denoising steps. The checkpoint uses the same 7B visual generation architecture, loads directly with QwenImage21Pipeline in Diffusers, and includes its recommended sampling schedule. It defaults to CFG=1 and uses prefix KV caching to reuse text and reference-image context across steps.

  5. Claude BlogOfficialAI score54

    Claude Managed Agents guide shows how to build scheduled agent automations

    AIThe Claude Blog published a guide to building scheduled agent automations with Claude Managed Agents (beta) that reads custom sources such as Slack and GitHub and posts a daily brief. The guide covers scoped vault credentials, per-source bookmarks so no window is lost or repeated, and confirming each Slack post before updating records. It also covers read-only access, a per-run spending cap, and a reference implementation with a Claude Code setup command.

  6. ModelScopeOfficialAI score60

    Qwen-Image-2.1-Turbo cuts image generation and editing to 8 denoising steps

    AIModelScope announces Qwen-Image-2.1-Turbo, an accelerated checkpoint that keeps the 7B visual architecture and runs image generation and editing in 8 denoising steps. The source says it uses CFG=1 and prefix KV caching to reuse text and reference-image context across steps, supports 2048 resolution with square, portrait, landscape, and widescreen presets, and loads through QwenImage21Pipeline in Diffusers. It is released under the Qwen Research License Agreement.

    Why it matters: The source names a concrete speedup path, 8 sampling steps and CFG=1 with prefix KV caching, which matters to anyone weighing image generation latency.

    Image from @ModelScope2022's post
  7. Don't Worry About the Vase (Zvi Mowshowitz)BlogAI score73

    OpenAI releases 719 AI-generated math manuscripts, splitting the mathematics community

    AIZvi Mowshowitz reports that OpenAI released 722 math manuscripts from an internal frontier model on GitHub, later reduced to 719 after three withdrawals, covering 90 of the top 500 open problems. He says the work came mostly from a single prompt, with an average of three hours of compute per solution. Mathematicians reacted with mixed feelings, and the post highlights concerns about unread papers, cryptography implications, and the role of Lean verification.

  8. Simon WillisonBlogAI score27

    Simon Willison builds a new blog feature largely by voice with Codex

    AISimon Willison says he built a Newsletters index for his blog almost entirely by voice, using the ChatGPT desktop app's Codex voice mode while cooking dinner. The feature imports weekly Substack posts via RSS and undocumented API, monthly newsletters from a GitHub archive repository, and a private sponsors-only newsletter. He says he switched back to typing for review and fixes before deploying the pull request.

  9. Rohan PaulXAI score46

    Microsoft's TeleTune evolves agent skills from raw usage logs

    AIMicrosoft researchers present TeleTune, which lets agents learn software skills from raw usage logs by keeping only skill edits that better predict users' next actions. The method needs no live test environment, because next-action accuracy on held-out logs tracked live success. Unlike earlier methods such as Agent Workflow Memory, which need goal-labeled examples or a live environment, TeleTune guesses each session's goal and uses wrong guesses to suggest edits to a text skill library.

    Image from @rohanpaul_ai's post
  10. Sakana AIOfficialAI score37

    Sakana AI paper uses LLMs to catch errors in research papers

    AISakana AI researchers introduce a benchmark that plants contradictions in papers to test whether LLM reviewers can detect errors, and propose Multi-Layered Review, modeled on the Three-Pass Approach to reading. Their system detected more errors than the other review systems tested, including in papers withdrawn for real mistakes, while its paper-quality assessments stayed broadly consistent with human judgments. The work, accepted at TMLR, is framed as support for human reviewers rather than a replacement.

    Video from @SakanaAILabs's post
  11. Google WorkspaceOfficialAI score12

    RigStrips uses Google Workspace with Gemini to draft customer replies faster

    AIRigStrips co-founder Zhach Pham says he uses Google Workspace with Gemini to draft customer support replies instantly, saving hours that would otherwise go to outdoor gear testing. The post highlights the company's more than 150,000 units shipped, though it gives no specific figures for time saved.

    Video from @GoogleWorkspace's post
  12. 🚨 AI News | TestingCatalogXAI score22

    Google spotted testing an Ultra mode in AI Studio Build

    AITesting Catalog reports that a new Ultra mode is in development in Google AI Studio Build, described as building with advanced skills and tools, alongside Plan, Build, and a previously spotted Security review mode. The post gives no details on which tools or skills it uses, and it does not say whether Ultra mode will require an Ultra subscription. The author suggests these modes are likely being developed to work with Gemini 4 Argon.

    Image from @testingcatalog's post
  13. LangChainOfficialAI score34

    Snyk's Assist support agent handles 60k queries with 85% resolution

    AISnyk's Assist, a customer support agent built on LangChain and LangGraph with observability in LangSmith, has handled over 60,000 queries for more than 500 customer accounts. Over 85% of sessions are resolved without a support ticket, and more than 250 cases were automatically detected and escalated to the right team.

    Image from @LangChain's post
  14. LeiphoneNewsAI score42

    Credo moves into optical chips with DSP, PIC and diagnostics in one 1.6T module

    AICredo's latest full-DSP module, Cardinal 802, uses a 4×200G design aimed at both 800G and 1.6T, after the company expanded its ZeroFlap optical module line from 800G to 1.6T over the past year. The company also added Kfir200 silicon photonics PIC from its DustPhotonics acquisition and a PILOT diagnostics platform to the module.

  15. TechRadar · AINewsAI score36

    Google Playground turns plain-language prompts into playable AI-generated games

    AIGoogle's Playground experiment lets users describe a game in ordinary language and have generative AI build a playable browser-based result that can be revised through further prompts. TechRadar's reviewer built a dragon platformer, Mystic Dragon Glide, from a couple of sentences and a satirical puzzle RPG, Red Tape Hero, from a longer prompt. Playground produced working controls, objectives and music, but the reviewer found the results impressive as prototypes rather than games they would want to play for dozens of hours.

  16. The New York Times · TechnologyNewsAI score20

    Anthropic's quest to give AI morals

    AIThe New York Times reports on Anthropic's effort to instill moral values in its AI systems, which the excerpt describes as part research and part evangelism. The source text provided is only one sentence, so no further details about methods, models, or results can be confirmed.

  17. X.PINXAI score36

    Tencent considers up to $5 billion offshore bond sale for AI

    AITencent is reportedly considering an offshore bond sale of up to $5 billion in US dollars and offshore yuan, possibly as early as this month, according to Bloomberg. The move follows its $4.66 billion bond offering in June, its largest debt deal since 2020, with proceeds earmarked for general corporate purposes including AI development. The post notes that major tech firms are increasingly turning to debt to fund AI infrastructure spending beyond their existing cash flows.

    Image from @thexpin's post
  18. QbitAINewsAI score67

    TRAE merges Code and Work into one platform with Agent and IDE modes

    AITRAE has merged its TraeCode and TraeWork products into a unified new TRAE with an Agent mode and an IDE mode. In hands-on tests, multiple agents handled planning, design, coding, testing, and fixes within one project, with outputs saved in a shared 'My Artifacts' area. The tests also found that agents working in parallel produced conflicting specifications, so someone had to coordinate them.

    Why it matters: The hands-on tests show how parallel agents split planning, design, coding, testing, and fixing inside one project, and where their outputs conflicted.

  19. Rest of WorldNewsAI score40

    Insta360 opens its first U.S. flagship store as DJI faces FCC drone and camera sales ban

    AIShenzhen-based Insta360 opened its first U.S. flagship store in New York City's Times Square in September, marketing its pocket-sized cameras to content creators, travelers, and sports enthusiasts. Co-founder Max Richter said the U.S. accounts for 30% to 40% of the company's revenue, while the FCC has blocked DJI from selling new drones and cameras in the country. Insta360 also launched drone brand Antigravity, whose A1 drone received FCC approval just before a new foreign drone rule took effect.

  20. Bloomberg · TechnologyNewsAI score34

    Bending Spoons CEO Sees Acquisition Opportunity in Falling Software Valuations

    AIBending Spoons CEO Luca Ferrari says falling software valuations are creating acquisition opportunities as AI reshapes the industry. He says the company benefits from cheaper targets and from using AI to write code, improve products and scale its acquisition model. Bending Spoons says AI now writes at least 90% of its code.

  21. Semafor · TechnologyNewsAI score44

    OpenAI reportedly tells investors it expects $70 billion annualized revenue this year

    AIOpenAI reportedly assured investors it expects to reach a $70 billion annualized revenue target this year, an apparent effort to ease concerns that it is missing earlier projections. The figures matter because they are seen as a gauge of underlying demand for cutting-edge AI models and a justification for huge spending on tech infrastructure. Some analysts, however, say the annualized revenue metric itself is flawed.

  22. Amazon Web ServicesOfficialAI score10

    Meliá Hotels International adopts AI approach, per AWS post

    AIMeliá Hotels International, a hotel chain with more than 380 hotels across four continents, is cited in an AWS post as having taken a step the post describes as the approach needed. The source gives no further details on the specific technology, implementation, or results.

  23. Amazon Web ServicesOfficialAI score33

    AWS helps migrate COBOL reservation system to microservices

    AIA company used AWS to migrate its entire COBOL-based central reservation system to microservices. The migration handles 50 million daily availability requests, up from 26 million, with faster response times.

  24. Amazon Web ServicesOfficialAI score14

    AWS customer cuts feature delivery to one month and compute costs 60%

    AIAn AWS customer reports that new features now ship in one month instead of four, with 60% compute cost savings worth seven figures. The post also cites a 75% faster time to market, near 99.99% availability, and a four-year project finished in two years.

  25. QbitAINewsAI score38

    Lenovo's TianxiCode Agent Tops SWE-bench-Live Lite Leaderboard at 71%

    AILenovo's TianxiCode, paired with DeepSeek-v4.1-Flash, ranked first on the SWE-bench-Live Lite leaderboard with a 71% issue resolution rate and passed official Verified review. The framework combines multi-hop retrieval, autonomous planning with multi-turn tool calling, and test-driven self-correction, and will be applied to Lenovo AI hardware products.

  26. The DecoderNewsAI score54

    Anthropic's Claude Science maps the full sky in ultraviolet light

    AIAnthropic's Claude Science has produced what the source describes as the first complete ultraviolet map of the sky. AI agents downloaded data from multiple space missions, calibrated and merged it, and used inpainting to fill gaps left by NASA's GALEX mission, which skipped bright star-forming regions. In tests, predictions averaged about ten percent deviation from actual measurements, and the map is intended as teaching material.