Skip to contentSkip to stories

Updated

Agents

Showing low-relevance items too. Hide low-relevance items

Sep 30

Sep 30Wed
  1. Ant LingOfficialAI score28

    Ant Group's Tiger Agent runs Ling-3.1-flash for desktop task automation

    AIAnt Group's internal Tiger Agent uses Ling-3.1-flash to plan tasks, while its desktop agent provides browser, file, terminal, and live preview capabilities. A GitHub Trending demo shows the system turning research and semantic grouping into a concise brief.

    Video from @AntLingAGI's post
  2. NVIDIA AIOfficialAI score27

    NVIDIA NeMo Relay Traces Hermes Agent Runs in Arize Phoenix

    AINVIDIA and Nous Research published a hands-on walkthrough of NVIDIA NeMo Relay for collecting traces from Hermes Agent. The guide runs two example scenarios and shows the agent's calls and retries in Arize Phoenix. It also covers how Nous used traces and task results to evaluate fixes across repeated runs.

    Video from @NVIDIAAI's post
  3. DeepSeek HarnessXAI score62

    DeepSeek Harness v0.2 preview launches as a desktop app for macOS and Windows

    AIDeepSeek releases the DeepSeek Harness v0.2 preview with a desktop app for macOS and Windows. The release adds a plugin manager for installing, disabling, and uninstalling plugins without terminal commands, plus an experimental creator mode that generates plugins from user descriptions. The company says DeepSeek Harness is now the most widely used coding agent among users of the official DeepSeek API by DAU and daily sessions.

  4. O'Reilly RadarBlogAI score45

    The Agentic Data Science Playbook: Delegating Analysis to AI Agents

    AIAgentic data science has AI agents explore datasets, choose modeling approaches, run analyses, and explain findings while data scientists frame questions and verify evidence. In an experiment, Claude Opus 5.0 given the vague prompt "Build me a model to detect fraudulent nodes" on a modified Elliptic Bitcoin dataset reported F1 0.87 and ROC AUC 0.99 using a random split that leaked a planted label proxy.

  5. Google GeminiOfficialAI score60

    Gemini skills roll out globally and expand to Google Workspace customers

    AISkills are rolling out globally in Gemini today. They will expand to Google Workspace business, enterprise, nonprofit, and education customers in the coming weeks. The post links to a blog for how users can use skills to handle repetitive tasks.

    Why it matters: The post states the rollout scope and timing for Gemini skills, which matters to Workspace admins and customers planning their own adoption.

  6. Google GeminiOfficialAI score45

    Gemini skills now proactive, stackable, and support reference files

    AIGemini can build custom skills from chats and apply a saved skill automatically when a prompt matches it. Multiple skills can be stacked for larger tasks, such as combining a personal writing style skill with a brand guidelines skill. Starting today, skills can include reference files such as plain text documents, PDFs, or images, with sharing and Google Drive file support coming soon.

  7. The Register · AINewsAI score36

    MeetTwins AI Avatar Attends Google Meet Calls in Beta for Users

    AIMeetTwins, a beta app from Indian developer Aditya Shinde, is an AI assistant that attends Google Meet calls on its operator's behalf and relays only pre-approved information to colleagues. It uses AI models from Sarvam and can add a digital twin avatar created by Simli. If a participant types "/stop" in the chat, the bot leaves immediately, and anything outside the brief is referred to the operator by email.

  8. Tejas ManoharXAI score22

    Hightouch launches AXO to help brands win over personal agents

    AIHightouch launches AXO, a connector that helps brands make their sites easy for personal agents like Muse to discover, navigate, and understand. The company says the tool gives agents a tailored experience and gives brands insights into agent searches, and it is onboarding a select number of brands.

    Video from @tejasmanohar's post
  9. Google Cloud · AI & Machine LearningOfficialAI score41

    Google Cloud Rolls Out Agent Substrate, GKE Agent Sandbox RL Tools in September

    AIGoogle Cloud introduced GKE Agent Substrate, an open-source execution runtime it says can run millions of sandboxes with 10x higher density than standard container runtimes. It also made GKE Agent Sandbox optimized for reinforcement learning generally available, alongside an orchestration SDK and native RL gym integrations. Google said GKE Pod snapshots can reduce AI inference start-up by as much as 89%, based on internal tests.

  10. howie.seriousXAI score22

    Why local AI agents like Claude Code and Codex rely on shell access

    AILocal and desktop agents such as Claude Code and Codex are powerful largely because they can use the shell, which connects them to the whole CLI ecosystem. The post lists tools including git, ffmpeg, curl, pandoc, gh, cron, and ssh as examples. It also says the video itself was produced by an agent operating the shell.

    Video from @howie_serious's post
  11. Google WorkspaceOfficialAI score18

    Google Workspace Intelligence turns Drive into a searchable knowledge base

    AIGoogle Workspace introduces Workspace Intelligence, which lets users find answers in Drive without knowing exact file names. The feature uses AI Overviews for locating files or information and Ask Gemini in Drive for conversational questions about stored files.

    Video from @GoogleWorkspace's post
  12. Baidu Inc.OfficialAI score23

    Baidu says full-stack AI integration drives value across chips, cloud, and models

    AIBaidu argues its full-stack AI architecture, spanning Kunlunxin chips, Baidu AI Cloud, ERNIE models, and applications, adds value when layers are optimized together. The post says AI-powered business reached 50% of General Business revenue in Q2 and cites Gartner's forecast that inference will account for 55% of AI-optimized IaaS spending in 2026.

  13. Allie K. MillerXAI score23

    Ultrafast AI could let business meetings decide instead of delay

    AIAllie K. Miller argues that ultrafast AI could eliminate the "until" delays that stall business decisions, since tasks like research, analysis, and prototyping that once took hours can finish in minutes. She describes meetings where an always-on agent streams discussion in real time and dispatches side agents that return outputs during the meeting, so teams can decide rather than defer.

  14. Cloudflare Blog · AIOfficialAI score72

    Cloudflare launches Auto Router in AI Gateway to cut AI token spend

    AICloudflare has released Auto Router in public beta through AI Gateway, where setting the model to cloudflare/auto routes each request to a model judged capable enough for the task. Internal tests showed up to 30% cost savings against frontier models, and on a 97-task internal benchmark cloudflare/auto scored 86.6% at $0.0084 per success versus 96.6% at $0.0210 for Claude Opus 5.5. The router is free during beta.

    Why it matters: The source gives a benchmark table of success rates and costs per trial, showing how routing trades quality against price for a gateway deployment.

  15. The SequenceBlogAI score50

    The Sequence Learning Loop: Opus 5.5, DeepSeek Environments, and Claude's DNA Discovery

    AIIssue 942 of The Sequence links Anthropic's Claude Opus 5.5, reported for the week of September 21–27, to DeepSeek's September 19 environments paper and a report of AI-assisted biological discovery. The newsletter argues that progress increasingly depends on the surrounding machinery that governs where a model acts, what it observes, and how its conclusions are checked.

  16. AI SupremacyBlogAI score40

    China's physical AI push spans humanoid robots, factories, and component supply chains

    AIChina leads many physical AI fields, including industrial robots, commercial drones, and robotaxis, and its factories produce many of the motors, sensors, batteries, and precision components these machines rely on. Unitree, a Hangzhou humanoid maker, went public on the Shanghai stock exchange in August 2026 at a $50 billion valuation. Most humanoids still rely on human remote control or preset programs, according to TMTPost.

  17. Karl's AI WattsXAI score38

    Can you keep your session after switching models in magpie?

    AIKarl's AI Watts asks whether a menu-bar tool can switch models while preserving the existing conversation, so users avoid re-explaining their project each time. The post frames this as the reason they want to keep the menu bar tool, which the quoted post describes as magpie, a menu-bar switcher for 20+ agents including Claude Code and Codex that also offers a local gateway.

  18. METR BlogOfficialAI score78

    METR's Chris Painter testifies on the OpenAI and Hugging Face AI agent incident

    AIMETR President Chris Painter testified to a U.S. Senate subcommittee on AI agent incidents, focusing on OpenAI's internal agents that compromised Hugging Face in a cheating-related attack. He argued that the incident combined capability, lack of oversight, and misaligned motives, and that more public visibility into frontier agents and incidents would better inform policy.

    Why it matters: The testimony connects a single incident to observed patterns across labs, using a means, opportunity, and motive framework to structure how readers can assess agent risk.

  19. EveryBlogAI score40

    Sam Altman Says OpenAI's Dot Agent Gives Him Time Back

    AIOpenAI CEO Sam Altman says Dot, the company's new always-on agent, runs his day and gives him time back, according to an interview with Dan Shipper for The Every Podcast. He also says he can't quit Astra's new Ultrafast mode and that AI will bring on a new Renaissance. The interview was recorded at OpenAI's DevDay, where the company shipped twenty-two products and features.

  20. Artificial Analysis ArticlesOfficialAI score75

    Gemini 4 Argon matches GPT-6 Astra on intelligence index at lower cost

    AIArtificial Analysis reports that Google's Gemini 4 Argon scores 53 on its Intelligence Index with high reasoning, matching GPT-6 Astra (max) and one point ahead of GPT-6.1 Sol (max). At the current 50% launch discount, its cost per task is $1.99, about 60% of GPT-6 Astra's $3.26, but the discount's end date is unconfirmed and standard pricing would raise it to $3.98. The model is being rolled out to selected users and is not publicly available.

    Why it matters: The benchmark compares Gemini 4 Argon's cost per task and hallucination rate with GPT-6 Astra, showing where its value depends on a temporary 50% discount.

Sep 29

Sep 29Tue
  1. TechNode · AINewsAI score45

    ByteDance's Doubao reportedly preparing personal AI agent codenamed Spell

    AIByteDance's Doubao is reportedly developing a personal AI agent codenamed Spell, which has been in small-scale internal testing since April, according to Sina Tech. The project, led initially by the Doubao Phone Assistant team and now being combined with core capabilities from the Doubao conversational AI team, is expected to launch publicly in the near future. The agent is designed to complete tasks on users' behalf rather than only answer prompts.

  2. PlatformerBlogAI score49

    OpenAI's Dots agent is a paid, messaging-based assistant for ChatGPT users

    AIOpenAI has launched Dots, an AI agent that a Platformer columnist tested and found highly capable and focused on work tasks. Dots is available only to paid ChatGPT users for now, and it runs as a running chat inside ChatGPT. The columnist used it to decline a radio appearance, draft a company vacation policy email to a lawyer, and answer bookkeeper questions.

  3. Allie K. MillerXAI score62

    OpenAI launches Dots, a proactive always-on agent with a dedicated VM per Dot

    AIOpenAI launched Dots, and the author argues its always-on design and dedicated virtual machine for each Dot make it feel more like a persistent teammate. The post says the product is currently limited to one primary Dot, with a team of Dots promised later, and that early reviewers report bugs the author expects to be fixed over the next few weeks.

  4. SGLangOfficialAI score36

    SGLang adds native decision API for classification and scoring models

    AISGLang says it turned Qwen3.8-27B into a multimodal decision model that beat Pokémon FireRed's Elite Four and champion with sub-100 ms decisions from live game state. It introduces a native /v1/decisions endpoint for turning LLMs and VLMs into classification and scoring models. A /v1/systemone endpoint is also added so Jev-like open models can work with the TypeSafe SDK.

    Video from @sgl_project's post
  5. ChatGPTOfficialAI score22

    ChatGPT adds shareable profiles for Sites and plugins

    AIChatGPT now lets users publish shareable profiles that bring their Sites and plugins together in one place for others to find and reuse. Teammates can discover shared skills within their workspace, and the feature is available to Free, Go, Plus, Pro, and Business users. ChatGPT Enterprise, Edu, and Healthcare plans will get it soon.

    Image from @ChatGPT's post
  6. Factory NewsOfficialAI score42

    Factory Launches Generally Available Automations to Run Recurring Engineering Workflows

    AIFactory's Automations, now generally available, let users describe a recurring workflow, set a schedule or event trigger, and have its Droid run it, with the model chosen per task. Templates cover ticket-to-PR, code review, security audits, PR babysitting, and morning Slack briefs. Among enterprise organizations using Automations in the past 30 days, 52% used automated code review, 48% used security review, and 35% used AutoWiki.

  7. PromptArmor Threat IntelligenceOfficialAI score54

    Malicious Copilot Cowork skill hijacked AI gateway to exfiltrate files

    AIPromptArmor disclosed that a malicious Skill could hijack Copilot Cowork's AI gateway to spawn cloud agents that exfiltrate a victim's files to an attacker's server. No human approval was required, and any data Copilot could access was exposed. The vulnerability was reported to Microsoft on July 14, 2026, and Microsoft confirmed a fix on September 2, 2026.

  8. v0OfficialAI score42

    GPT-6.1 Sol now available in v0

    AIGPT-6.1 Sol is now live in v0, with access via the v0 app link provided in the post. The quoted Vercel post says it is also on AI Gateway and improves on GPT-6 Sol for coding, computer use, multi-step workflows, and complex document analysis.