Skip to contentSkip to stories

Updated

Agents

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 26

Sep 26Sat
  1. MetaOfficialAI score22

    Meta unveils Muse, a personal AI agent for everyday life

    AIMeta introduced Muse, a personal AI agent that learns the user's goals and works across different areas of their life to return time to them. The post says it was built with privacy and security from day one, and it was shared under #MetaConnect.

    Video from @Meta's post
  2. Liquid AIOfficialAI score20

    Liquid AI Explains Post-Training for On-Device Agentic Models

    AILiquid AI's post-training team, including Maxime Labonne, Edoardo Mosca, and Jiahui Wang, discusses what makes an on-device agentic model useful. The post says post-training shapes how models learn to use tools, follow instructions, handle longer contexts, and recover when tasks become complex.

    Video from @liquidai's post

Sep 25

Sep 25Fri
  1. Boris ChernyXAI score49

    Anthropic launches a portal for submitting and tracking Claude plugins

    AIAnthropic has launched a new portal where developers can submit Claude plugins, track review status, and monitor usage. Plugins package MCP and skills, and the company says MCP usage across Claude products is up 110x this year. Boris Cherny said he is eager to see what developers build.

  2. Lydia Hallie ✨XAI score22

    Claude Code's prompt-audit command renamed from /claude-api to /checkup

    AIAnthropic's Lydia Hallie says the prompt-audit command is now also available as /checkup, replacing the API-specific name that suggested it only worked with the API. The command checks CLAUDE.md, skills, and agents for instructions the model no longer needs, and it has always worked on Claude Code setups.

    Video from @lydiahallie's post
  3. Google AntigravityOfficialAI score34

    Antigravity 2.0 adds planning mode with /plan command

    AIGoogle Antigravity 2.0 now includes a dedicated planning mode, matching the Antigravity CLI. Typing /plan makes the agent research the task and generate an implementation plan for user review before execution, requiring approval to proceed. Users can also request a lighter plan through a natural prompt.

    Video from @antigravity's post
  4. KreaOfficialAI score22

    Krea launches an agent for creative workflows

    AIKrea published a post linking to presenting an agent-focused product page. The post itself provides no further details on features, models, pricing, or availability.

  5. WorkBuddyOfficialAI score36

    Grok-4.7 now available on WorkBuddy for user tasks

    AIWorkBuddy has added Grok-4.7 to its model lineup, making it available for users to try on their next task. The post gives no benchmark, pricing, context length, or other technical details.

    Image from @WorkBuddy_AI's post
  6. GitHub Blog · AI & MLOfficialAI score33

    How to build custom workflows with canvases in the GitHub Copilot app

    AICanvases in the GitHub Copilot app are customizable interfaces that you and the agent share, such as kanban boards, dashboards, or checklists. You create one by running /create-canvas and describing the workflow, what you can do in the interface, and what the agent can do. Changes made by either you or the agent appear immediately in the shared canvas, and completed canvases can be saved as reusable extensions.

  7. Boris ChernyXAI score62

    Claude Tag in Slack gains personal connectors for channel workflows

    AIClaude Tag in Slack can now use users' personal connectors, such as Drive, Salesforce, and warehouse access, within channels. The author says Tag writes over 50% of their PRs daily and handles nearly all their data analysis and many product bug fixes. The personal connector feature is available on Teams today and Enterprise next week.

    Why it matters: The post gives concrete usage examples for an in-Slack agent handling PRs, bug reproduction, and data analysis, showing how a team might fold such tooling into daily engineering work.

  8. Cognition Blog (Devin, Windsurf)OfficialAI score30

    Cognition Reaches $1B in Annualized Revenue Run Rate

    AICognition has crossed $1 billion in annualized revenue run rate, according to a company blog post dated September 25, 2026. The company says Devin, which became generally available less than two years ago, now works alongside engineering teams at GE Aerospace, Rivian, Rohlik, and Exa.

  9. Alex HeathXAI score42

    Satya Nadella says AI agents will create a market orders of magnitude bigger than cloud

    AIMicrosoft CEO Satya Nadella told Alex Heath that AI agents could create a market "orders of magnitude" bigger than the cloud, during an interview tied to the unveiling of the new Copilot. The conversation covers Autopilot, Microsoft's OpenClaw-based agent that works on users' behalf, along with AI safety, public trust, Microsoft's relationship with OpenAI, and Xbox's path back to growth.

    Video from @alexeheath's post
  10. Meituan LongCatOfficialAI score62

    Meituan LongCat-2.5-Preview Launches with 1.6T Parameters and 1M-Token Context

    AIMeituan's LongCat team has released LongCat-2.5-Preview, a natively multimodal model with 1.6T total parameters, about 48B active, and a 1M-token context window. The model is built for long-horizon tasks spanning terminals, browsers, GUIs, spreadsheets, and design tools. It is available now through an API on the LongCat platform and a chat interface.

    Why it matters: The announcement lists concrete scale, context, and multimodal specs, plus a named range of agent tasks, which helps readers gauge the preview's scope against other long-context models.

    Image from @Meituan_LongCat's post
  11. Microsoft CopilotOfficialAI score40

    Microsoft Copilot app refreshed to unify chat, agents, app building, and workflows

    AIMicrosoft has refreshed its Copilot app to bring chat, task delegation, app building, and workflow automation into one place. The update is positioned as an AI built for work, with Satya Nadella describing Copilot as a new OS for work spanning models, form factors, and tasks. The announcement includes Autopilot, an enterprise agent, Code for building apps hosted within a company's tenant, Home combining Chat and Cowork, and Office fully embedded in Copilot.

  12. Satya NadellaXAI score52

    Satya Nadella announces Copilot update with Autopilot, Code, Home, and Office

    AIMicrosoft CEO Satya Nadella announced what he called the biggest Copilot update to date, positioning Copilot as a new operating system for work. The update bundles Autopilot, a proactive long-running enterprise agent; Code, for building apps hosted inside a company's tenant; Home, combining Chat and Cowork; and Office, now fully embedded in Copilot. Copilot can also be invoked in Teams, and a new proactive experience called Today surfaces key information from across M365 without a prompt.

    Video from @satyanadella's post
  13. François CholletXAI score32

    Chollet: Software engineering difficulty stays constant across abstraction levels

    AIFrançois Chollet argues that the difficulty of software engineering stays essentially constant regardless of abstraction level, because human cognition adapts to new tools. He says tools are affordances rather than magic wands that eliminate work, and that great software engineering remains immensely challenging despite changed workflows. Simon Willison's background post similarly argues that coding agents make software engineering harder, requiring extraordinary discipline and knowledge.

Sep 24

Sep 24Thu
  1. PlatformerBlogAI score55

    Meta's Muse agent and VR Glasses reflect a shift from the metaverse

    AICasey Newton argues that Meta's focus on Muse, a personal AI agent under a month old, partly conveys momentum as the company plans up to $145 billion in capital spending this year. He contrasts Muse's early reported usage with Meta's earlier metaverse claims and calls the new Meta VR Glasses a notable engineering step, while urging testing beyond demos. The column also covers an OpenAI agent that accessed an Australian Medicare portal without authorization.

  2. Noah ZwebenXAI score30

    Claude adds personal connectors in channels with two risk safeguards

    AIAnthropic's Noah Zweben says the team addressed two key risks before launching personal connectors in channels. In a shared environment, one user's connectors could otherwise be unusable by anyone else, and private data could leak into the channel without the user's review. Tools now let the user and Claude prevent that data from entering the channel without review.

  3. GitHub Blog · AI & MLOfficialAI score46

    GitHub Copilot app's canvases argue chat is the wrong AI interface

    AIGitHub argues that chat is often the wrong interface for AI work and proposes customizable "canvases" inside the GitHub Copilot app. Canvases are full-stack applications running without browser chrome that can communicate bi-directionally with the Copilot agent and execute code locally. The post cites examples including a Connect 4 game, a Winget package manager UI, and a SQLite database interface.

  4. Google ResearchOfficialAI score60

    Google Research details four agentic frameworks for coherent long-form video generation

    AIGoogle Research introduces four multi-agent frameworks for generating minutes-long videos with consistent characters and environments across shots. The frameworks include AI video co-director, CANVAS, A²RD, and VQQA, which are built as orchestration layers on Gemini and Veo and use SynthID watermarking. The post reports measured gains on benchmarks such as GenAD-Bench, HardContinuityBench, and LVBench-C, with the full architectures described in the linked papers.

    Why it matters: The post links four frameworks to specific failure modes in long video generation, such as semantic drift and cascading errors, making the design choices easier to compare.

  5. Baseten BlogOfficialAI score44

    LangSmith Fine-Tuning Trains Open Models on Agent Traces via Baseten Loops

    AILangChain launched LangSmith Fine-Tuning, which lets users fine-tune open models on their LangSmith agent traces using the open-source smithtune CLI. Training runs on Baseten Loops in the user's own workspace, and smithtune deploy places the evaluated checkpoint on a Baseten Dedicated Inference deployment. Loops is in early access, so users may need to request access for their workspace.

  6. GitHub Blog · AI & MLOfficialAI score66

    GitHub Security Lab shows an LLM agent running AI-driven fuzzing for C/C++ projects

    AIGitHub Security Lab describes the Fuzzing Taskflow, an LLM agent pipeline that identifies entrypoints, writes harnesses, runs AFL++, reads coverage reports, and triages crashes for C/C++ repositories. The agent makes decisions while MCP tools handle execution, and state is stored in a SQLite database. The post also warns that the taskflow runs AFL and build commands directly on the host, so it should be used only in disposable environments without elevated privileges.

    Why it matters: The post explains how an LLM agent automates fuzzing steps like harness writing, coverage gap chasing, and crash triage, with a runnable workflow and design tradeoffs.

  7. Azure BlogOfficialAI score67

    Microsoft Foundry adds voice agents and continuous optimization for production agents

    AIMicrosoft Foundry expands its agent platform with voice agents in public preview, long-running resilience for hosted agents, and tools for evaluating production agents. The post also says GPT-6 Sol, GPT-6 Luna, and Claude Opus 5.5 are now available in Foundry. Agent optimizer, Insights, and Rubric evaluator are described as tools for continuous improvement, with some reaching general availability later this month.

    Why it matters: The post shows how Foundry combines model choice, voice agents, long-running resilience, and production evaluation into one agent workflow, with a customer example.

  8. Microsoft Foundry BlogOfficialAI score40

    Foundry Agent Service adds egress policies to restrict hosted agent destinations in preview

    AIMicrosoft's Foundry Agent Service preview lets developers attach a named, ordered egress policy to a hosted agent, allowing only approved destination hostnames. The walkthrough uses an invoice agent, an Audit-mode RAI policy with a Deny default, and Allow rules for two finance and vendor hosts, configured outside the agent code. Network egress controls are preview features, not GA, with no preview SLA, and are not intended for production use.

  9. Google for DevelopersOfficialAI score37

    Gemma 4 now runs on-device in the Antigravity SDK

    AIGoogle says Gemma 4 can now run locally on-device within the Antigravity SDK. Developers can build fully local or hybrid multi-agent workflows that pair cloud models with Gemma 4 agents for auditing, patching, and testing code. The post emphasizes total data privacy and zero API fees, powered by LiteRT.

    Video from @googledevs's post