Perplexity builds SPACE, a Rust-based sandbox system
AIPerplexity says SPACE is built in Rust and powers the sandboxes behind Perplexity Computer and the Agent API sandbox tool. The post links to a blog post explaining how the company built SPACE.
Updated
Updated
Showing low-relevance items too. Hide low-relevance items
AIPerplexity says SPACE is built in Rust and powers the sandboxes behind Perplexity Computer and the Agent API sandbox tool. The post links to a blog post explaining how the company built SPACE.
AIPerplexity says its MCP tools perplexity_ask, perplexity_research, and perplexity_reason run on Agent API. The company describes Agent API as a single, multi-provider endpoint with built-in tools and dynamic presets.
AIPerplexity's API MCP server now supports OAuth, letting users connect by adding to their client, signing in with their Perplexity account, selecting an org, and approving access. Once connected, agents gain real-time web search, deep research, and advanced reasoning.
AIMicrosoft Foundry's July and August 2026 updates make Hosted Agents, Voice Live integration, and Toolboxes generally available. The post adds Claude tools on Azure, Model Router region and model pool changes, Foundry Local preview features, and updated Python, JavaScript, Java, and .NET SDK versions with migration notes.
Why it matters: The roundup links each GA and preview change to code examples, migration notes, and runtime requirements, which helps developers judge what to upgrade and test first.
AIGamma says its API now lets users search their entire library of gammas, ranking results by relevance across both titles and full body text. Users can prompt from Claude, ChatGPT, or other connected chats, filter by creator or last-updated date, include archived work, and open results via direct links. The feature is rolling out gradually starting today.
AIDatabricks used Unity AI Gateway tracing and Genie One to find seven small MCP-server bugs and eliminate an estimated $1.2M in annual wasted AI spend and lost productivity within an hour. The bugs drove about $499K per year in wasted tokens, roughly 12,000 engineering hours per year in agent wait time, and 1,409 tool errors in a single 24-hour window. Matei Zaharia argues that analyzing tracing data for AI workloads will become a routine form of operational data analysis across companies, much like finance and security.
AIMeituan's LongCat-2.0, a 1.6T open-weights MoE model with a 1M context window, is now free to use in Cline. Cline's post says it scores similarly to Claude Opus 4.7 and Gemini 3.1 Pro. Users can select it under free models via /model after installing Cline with npm i -g cline.
AIGoogle's AI Agents Challenge judges highlighted four engineering patterns in top-ranked submissions: bidirectional MCP, event-driven concurrency, same-bar fallback, and tiered routing. One team exposed its internal MCP tools as an external MCP server that other agents could call, with access control required once outside callers reach it. Another replaced a linear agent pipeline with an asyncio.Queue-based event bus so agents react to shared events in parallel rather than waiting in a call chain.
AIVercel says users can ship a new eve agent in one minute by adding a prompt, picking models and MCP connections, and deploying at The deployed agent is described as production-ready for chat and backed by a Git repository the user owns.
AILM Studio introduced Auto Review, which auto-approves shell tool requests using AST parsing, command matching, and a reviewer subagent. In the team's internal use, about 82% of commands were approved before reaching LLM review.
AIAnthropic is building the Model Hardware Standard (MHS), a common way for AI models to connect to lab and manufacturing equipment and operate it with safety limits built into each device. MHS started as a collaboration between Anthropic and HHMI Janelia Research Campus and is launching as a research preview with partners across science, robotics, and manufacturing.
Why it matters: The source describes a standard for connecting AI models to lab and manufacturing hardware, which matters for anyone building automated experimentation workflows.
AIv0 now lets users connect apps directly to Slack, GitHub, Notion, Salesforce, and other services through Vercel Connect. Vercel describes Connect as generally available, offering short-lived scoped access tokens, token and trigger observability, and RBAC with audit trails for 100+ services.
AIVercel Connect is now generally available, letting apps and agents securely access more than 100 services, including Slack, Linear, and GitHub. It uses short-lived, scoped access tokens, token and trigger observability, and RBAC with audit trails.
AIPromptArmor disclosed a vulnerability in Microsoft Copilot Cowork that allowed a bypass of the sandbox, letting attacker servers send commands that run in the sandbox and return results. The attack could be triggered through a prompt injection or a malicious bundled script in a user-uploaded Skill, and it could read data from Outlook, SharePoint, plugins, and chat history. The issue was reported to Microsoft on June 24, 2026 and confirmed mitigated on August 19, 2026.
Why it matters: The report traces how a malicious bundled script in an uploaded Skill escaped the sandbox and kept running after the stop button was pressed, a concrete case of agent security failure.
AIv0 apps and agents can now securely connect to Slack, GitHub, Salesforce, and over 100 other services through Vercel Connect. The integration handles authentication using reusable team connectors and short-lived tokens.
AIswyx reports that covering this year's MongoDB Build Fest was a major step up from last year and gratifying to see San Francisco builders rediscovering MongoDB. He recalls learning to code with MongoDB over ten years ago through the MERN stack.
AIDeepSeek says its V4-Flash-Vision-Exp works smoothly across agent frameworks, combining visual understanding with a range of tools. The post presents this multimodal capability as a way to unlock more practical agent workflows.

AIChip Huyen asks what a good model tiering system looks like, since she is tired of naming specific models per vendor for her agent orchestrator. She wants to instruct the orchestrator by task tier, such as "use models tier ..." for a given kind of task, instead of listing Claude, OpenAI, and other models individually.
AIMicrosoft Foundry now offers structured outputs, web search, web fetch, MCP connector, and tool search for Claude models on Azure-hosted deployments. Prompts and completions remain within Azure for these deployments, while only usage metadata and safety-flagged content egress to Anthropic. The features were previously available only on Hosted on Anthropic deployments, which required choosing between capability and data-handling commitments.
Why it matters: The post shows which agent scaffolding now runs on Azure-hosted Claude deployments, which matters for teams needing data residency without rebuilding search, fetch, or tool routing.
AIJason Wei now believes a small 1B-parameter "cognitive core" relying on tools is insufficient, because fast, natural recall without tool use matters. He cites speed, knowledge better learned through backpropagation than retrieved from search, and the greater reliability of already-known facts over repeated lookups. Since a 1B model has an information limit, he argues that demanding AI will still need larger models, not just tool access.
AILM Studio says its Bionic assistant can install skills when given a URL or a description of where to find the skill. The post advises installing skills only from trusted sources.

AIDeepSeek has released DeepSeek Harness v0.1 in Developer Preview, opening the codebase under the MIT license for developers building agent harnesses. The harness is built on the Cordis meta-framework and treats models, tools, skills, sessions, sandboxes, filesystems, loops, orchestration, and UI as plugins that can be mixed, matched, replaced, and extended.
Why it matters: The source specifies the MIT license and a plugin-based architecture covering models, tools, and sessions, which helps developers assess extensibility before adopting it.
AIand reviewing their code, and invites first users into a private beta. Delta keeps code and conversations connected through DeltaDB, which captures edits and conversations between git commits and works with existing repositories. The app also supports cloud runners, browser-based sharing, and live syncing of Claude Code sessions.
Why it matters: The post explains how the new Delta app links conversations with code history, which clarifies a shift in how teams review agent-written changes.
AIGoogle Search and Google Maps can now be used in the same Gemini API call with Gemini 3.5 Flash and 3.6 Flash. Custom functions and MCP servers can be added to the same request through Tool Combination. The author says Gemini handles the search, place lookup, and function call in one interaction without extra roundtrips from the developer's side.
AIGrok Bot is now available in early beta as an AI teammate that signs in to your tools, uses them as you would, and returns finished work. The post frames it as an early step toward capable, delightful digital colleagues.
AIPromptArmor reports that a malicious Skill or indirect prompt injection can make Zoom's ZoomMate agent connect to an attacker's server and run commands. The connection can persist after the user clicks stop or closes Zoom, and the final chat output appears normal.
Why it matters: The report shows how a malicious skill or prompt injection can keep a Zoom agent connected after the user stops it, a risk to weigh before enabling agentic assistants.
AIThe v0 team has launched its new API, which lets developers build their own app builders, give agents the ability to build and deploy apps, and generate apps from scripts or CI jobs. The post links to further details at v0.link/v0api.
AIPromptArmor reports that a hidden prompt injection in an uploaded file can make Atlassian Rovo send Jira tickets and Confluence documents to an attacker's URL without human approval. The attack works even when organization-wide web search is disabled, because the setting does not remove the URL retrieval tool. PromptArmor says it disclosed the issue to Atlassian on May 23, 2026, and that Rovo remained vulnerable at publication on August 5, 2026.
Why it matters: The report traces a full indirect prompt injection chain in Rovo, showing how a disabled web search setting still leaves a data exfiltration path open.
AIManus has launched an ElevenLabs connector that lets users generate speech, transcribe recordings, clone voices, and build audio apps through a single chat. Users connect their authorized ElevenLabs account via Integrations, and audio is processed within their own ElevenLabs environment according to its policies. Availability depends on users having an active ElevenLabs account, with capabilities tied to their ElevenLabs plan and credit balance.
AIAndrew Ng and Rohit Prasad announced OpenWorker, an open-source agent that produces deliverables such as documents, Slack messages, and calendar updates across files and everyday tools. It checks in before consequential actions, runs on Mac with Windows support coming soon, and works with user-supplied API keys for models including GPT 5.6 Sol, Claude Fable, Gemini 3.6, open-weight models, or local Ollama models. Source code is available on GitHub, and the tool requires the user's own API key.
AIJetBrains has launched JetBrains Context in early access, a repository intelligence layer that builds a semantic index so coding agents can retrieve relevant code without repeated searching. In tests on 205 SWE-bench tasks, 175 production-monorepo tasks, and 1,953 code-localization tasks, it reduced agent turns by up to 68%, latency by up to 59%, and execution cost by up to 48%. It works with Claude Code, Codex CLI, and Junie CLI at no additional cost for JetBrains AI subscribers, and it does not store source code on JetBrains Context servers.
Why it matters: The source gives benchmark figures for turns, latency, and cost, showing how repository indexing might change agent workflows on large codebases.
AIWork doesn't happen in one app. That's why Skywork connects your entire workflow: From Gmail and Notion to Slack and Drive, you can focus on getting things done. Connect your workflow. Automate your work. Build with Skywork 👉

AIEugene Yan says he likes Claude Tag's multiplayer form factor because other people can reply in the thread to give Claude context and direction. Claude Tag, per the background post, lets teams in Slack tag Claude in as a team member with access to chosen channels and tools.
AIMiniMax describes MaxProof, an evolutionary-search framework that lets its M3 model refine candidate proofs over multiple rounds. The post says the M3 model exceeded the human gold-medal threshold on the IMO 2025 and USAMO 2026 benchmarks with MaxProof, and explains the Proof RL, verifier alignment, and refinement training behind it.
AIBaidu CoBuddy is now available for free on Novita AI, a code-focused model aimed at developers and AI agents. It offers a 131K context window, up to 65K output tokens, and native tool calling. The model is served through Novita's serverless API for high-throughput, low-latency inference.
AICognition has announced Devin Desktop, the next generation of Windsurf, which makes the Agent Command Center the default IDE surface for managing local and cloud agents, PRs, and context. Spaces let related agents share context, and Agent Client Protocol (ACP) support lets any ACP-compatible agent run alongside Devin. The IDE remains fully backwards-compatible with Windsurf, including editor extensions, keybindings, LSPs, and terminal workflows.
AICognition describes autonomous testing in Devin, where the agent writes a source-grounded test plan, operates the app through computer use, and returns labeled screenshots and an annotated video. Login steps are handled by a deterministic testing skill, and the company says test runs approved per day more than doubled in recent months. Known limits include timing errors with transient UI elements and models sometimes triggering states through JavaScript instead of clicking the interface.
Why it matters: The post explains how computer use, test plans, deterministic login scripts, and annotated recordings let Devin verify its own code changes end to end.
AINVIDIA's developer account announced a livestream titled "Scaling NemoClaw: Roadmap, OpenClaw Collaboration, and Real-World Integration" as part of Nemotron Labs. The post provides no further details on the roadmap, collaboration terms, or integration specifics.
AIBAAI announces ClawKeeper v1.0, an open-source security framework for OpenClaw AI agents, combining Skill-based command policies, Plugin-based runtime monitoring, and a Watcher system-level observer. The independent Watcher is designed to block high-risk operations such as prompt injections, key leaks, rogue commands, and remote code execution, even if the agent is compromised. The paper is available on arXiv and the project code is hosted on GitHub.
AIAnthropic's Managed Agents separates the harness, sandbox, and session into independently replaceable interfaces. The source says this design let failed containers be replaced, kept tokens out of the sandbox, and reduced p50 time-to-first-token by roughly 60% and p95 by over 90%.
Why it matters: The post explains how decoupling the harness, sandbox, and session changed failure recovery, credential security, and latency, offering a reusable architecture pattern for long-running agents.