Skip to contentSkip to stories

Updated

#Agent

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 28

Sep 28Mon
  1. IEEE Spectrum · AIAI score25

    Charlie Kemp Builds Assistive Mobile Robots to Help People Live Independently

    AICharlie Kemp, cofounder and chief technology officer of Hello Robot, develops mobile manipulators with arms to physically assist older adults and people with disabilities in homes and workplaces. His work began with humanoid robots at MIT and led to assistive robotics research, including a collaboration with Henry Evans through the Robots for Humanity effort. The profile is part of IEEE Spectrum's "A Day in the Life of a Roboticist" series.

  2. Andrew NgAI score46

    Andrew Ng says OpenWorker will use Nvidia OpenShell for sandboxed AI agents

    AIAndrew Ng says OpenWorker, his open-source agent harness for cybersecurity workflows, will run each agent's commands inside a sandbox built on Nvidia OpenShell. The sandbox limits files to those relevant to the task and keeps secret API keys, browser login credentials, and arbitrary website access out of the agent by default. Restrictions are enforced in deterministic code rather than by prompting an LLM, and all actions are logged for monitoring and audit.

  3. Google AIAI score44

    Google Labs expands experimental CC agent into a family group assistant

    AIGoogle Labs has expanded Project CC, its experimental AI productivity assistant, into a group agent designed to streamline family household logistics. CC has its own verified Google account and email, so families can share documents and calendars and auto-forward selected emails without sharing passwords or exposing their full inboxes. The post says CC runs on the latest Gemini models in isolated cloud environments, and it is available via a waitlist.

  4. catAI score72

    Claude Sonnet 5.5 Lifts Claude Code Task Completion by About 30%

    AIAnthropic's Cat Wu says Claude Sonnet 5.5 lets Claude Code users complete about 30% more tasks than with Sonnet 5. The model needs fewer tokens for the same work, and in a leaf-raking tool-call demo it finished 24 seconds faster using 6K fewer tokens.

    Why it matters: The post gives a measured Claude Code task-completion gain and a token-use example, showing what the model upgrade means for a coding agent workflow.

    Video from @_catwu's post
  5. RadixArkAI score46

    RadixArk releases Miles v0.1.1 with multi-LoRA and expanded model support

    AIRadixArk has released Miles v0.1.1, adding multi-LoRA with Tinker API compatibility so multiple training jobs can share one base model. The update also supports agentic training with harnesses like Claude Code and runs Harbor tasks in sandboxes including AgentENV, Daytona, E2B, and Modal. It further reduces memory needs for training larger models on validated NVIDIA and AMD GPUs and adds stable support for Qwen3.8-Flash-Next, GLM-5.3-Flash, and Kimi-K3.

    Image from @radixark's post
  6. Artificial IgnoranceAI score42

    OpenAI Engineer Argues Voice Agents Should Act, Not Only Talk

    AIAn OpenAI developer experience team member argues voice agents need not always speak back, outlining speech-to-speech, speech-to-action, and event-to-speech as emerging design modes. He cites form filling, creative tools, and computer use as examples of speech-to-action, which he calls among the most underexplored areas. He says event-to-speech is still very exploratory, with hands-free recipe guidance and proactive alerts as examples.

  7. Google · Gemini appAI score38

    See what 4 builders are making with Gemini 3.8 Flash

    AIGoogle says Gemini 3.8 Flash, its most intelligent workhorse model, improves on 3.7 Flash in software engineering, agentic tasks, and multistep reasoning by running extra reasoning steps and calling tools iteratively. The post highlights four community builds, including a model rocket simulation, an animated ink-painting effect, a 3D dinosaur skeleton, and an interactive automatic transmission simulation. Developers can try the model through Google Antigravity and Google AI Studio.

  8. clem 🤗AI score49

    Hugging Face proposes egress usage monitoring for OpenShell agent sandboxes

    AIHugging Face is contributing egress usage monitoring to NVIDIA's OpenShell, part of the newly launched Open Agent Safety Platform, arguing that allowlists alone restrict where agents can go but not what they do. The proposed features include per-sandbox network budgets for requests, writes, and bytes, drift detection against each sandbox's baseline and cohort, and a fleet view that flags many sandboxes writing to one host even when every request is allowed.

    Video from @ClementDelangue's post
  9. Philipp SchmidAI score52

    Gemini Managed Agents adds a Credentials API that keeps secrets out of sandboxes

    AIGoogle's Credentials API for Gemini Managed Agents lets agents authenticate to services like GitHub, Notion, and the Gemini API without placing raw secrets in the Linux sandbox. Secrets are stored encrypted on the server and injected on the wire by an egress proxy, with three credential types: bearer_token, oauth2, and environment_variable.

  10. Sierra BlogAI score34

    Sierra's Ghostwriter becomes a proactive Slack and Teams teammate for AI agents

    AISierra has turned its Ghostwriter tool into an always-on teammate in Slack and Teams that proactively suggests ideas, flags problems, and proposes experiments. Ghostwriter reviews recent customer calls, recommends which changes to try first, runs experiments, and reports when results are statistically significant. Sierra said it will begin rolling the feature out more broadly next week.

  11. NVIDIAAI score34

    NVIDIA launches Open Agent Safety Platform to control AI agent access

    AINVIDIA has launched the Open Agent Safety Platform to help teams control what AI agents can access and do. NVIDIA OpenShell enforces permissions around agent work, while BlueField-4 and DOCA add independent monitoring and security controls in the infrastructure beyond the agent's reach. Together, these components aim to give organizations defined permissions, oversight, and protection for long-running agent tasks.

    Image from @nvidia's post
  12. KhazixAI score31

    Solo developer rewrites AIHOT with multi-model AI workflow in three days

    AIThe developer behind AIHOT rewrote the entire project over three days, then launched it after a 12-step AI-assisted workflow. The process used Claude Opus 5.5, Claude Fable 5.1, and GPT-6 Astra for distillation, rewriting, audits, testing, and a six-hour shadow-system rehearsal before cutover. The post frames this as an amateur's experience and includes a quoted suggestion to distill the source project into a feature document and rewrite it directly with the latest models.

    Image from @Khazix0918's post
  13. Import AIAI score52

    Import AI 474 covers Michael Levin's mind-pattern paper, robot post-training, Google's space TPUs, and Zhipu's self-improvement loop

    AIImport AI 474 is a research newsletter by Jack Clark that surveys four developments and one fiction piece. It covers Michael Levin's paper proposing minds as patterns that ingress into bodies, Stanford researchers' call for a universal post-training recipe for robotics, Google's plan to send TPUs to space with Planet, and Zhipu's use of GLM-5.3 to speed up its own inference infrastructure.

  14. Jensen HuangAI score42

    NVIDIA releases open agent safety platform combining OpenShell and Sentry

    AINVIDIA's Open Agent Safety Platform Reference Design combines NVIDIA OpenShell and NVIDIA Sentry to secure AI agents. OpenShell, an open-source secure runtime, enforces clear boundaries and policy on agent actions while tracing them as they work. NVIDIA Sentry adds hardware-based enforcement on NVIDIA BlueField, continuously monitoring agent activity and enabling millisecond-scale containment and quarantine.

    Image from @JensenHuang's post
  15. Baseten BlogAI score26

    Baseten and Blaxel Back NVIDIA OpenShell Sandboxes With Carbon Preview

    AIBlaxel, which Baseten acquired, is introducing Carbon, its fourth-generation infrastructure, in private preview for running agents in secure sandboxes. Carbon runs on microVMs with a dedicated IPv6 address per sandbox, supports manual snapshotting, forking, and snapshot-to-production within milliseconds, and includes a template with NVIDIA OpenShell preinstalled. Carbon is rolling out progressively by region and workspace and is coming to Baseten soon.

  16. Manus BlogAI score60

    Manus 2.0 adds Cascade agent harness, Manus Studio, and Cue app

    AIManus 2.0 introduces a new agent harness called Cascade, Manus Studio with Video Editor and Game Dev environments, and a standalone Cue app for personal agents. In one tested configuration, Cascade used 23.2% fewer tokens, completed tasks 28.2% faster, and cost 32% less to run than the previous system. Cue is in early access and available with an invite code.

    Why it matters: The post separates the new agent harness, Studio, and Cue, and its Cascade chart gives measured token, time, and cost comparisons against the previous system.

Sep 27

Sep 27Sun
  1. PromptArmor Threat IntelligenceAI score72

    Elastic's AI SOC agent can be manipulated into leaking API credentials

    AIPromptArmor reports that Elastic's AI SOC agent, EASE, can be manipulated through malicious phishing alerts into minting API keys and sending them to an attacker. The attacker could then disable detection rules, create fake alerts, and exfiltrate data, and the report says the agent runs with user privileges and needs no human approval. PromptArmor says Elastic received the report on August 23, 2026, did not address it after four follow-ups, and published mitigations that include disabling built-in capabilities and write-capable tools.

    Why it matters: The report shows how a prompt injection in alert data can drive an AI SOC agent to leak API keys, with concrete mitigations for agent tool settings and default model choice.

  2. xAI News (Grok)AI score58

    xAI launches Team Bots, shared Grok Bots that learn as teams work

    AIxAI has launched Team Bots in public beta on Teams and Enterprise plans, letting teams build shared Grok Bots that keep context, plugins, credentials, and memories. Each person's conversations stay private while the Bot draws on skills shared across the team. The post also describes internal uses in sales, product and engineering, marketing, and data analytics, and it is available through Slack.

  3. Philipp SchmidAI score59

    Gemini Managed Agents Credentials API keeps secrets out of the sandbox

    AIThe Credentials API for Gemini Managed Agents lets an agent authenticate with services like GitHub and the Gemini API without placing raw secrets in the Linux sandbox. Secrets are stored write-only and encrypted, and an egress proxy injects the real credential on the wire only for requests to permitted domains. The post walks through creating bearer token and environment variable credentials, binding them to a reusable agent, and rotating or deleting them.