Claude Tag adds Group DMs for ad-hoc team collaboration
AIAnthropic's Noah Zweben announced that Claude Tag now supports Group DMs, letting users work with Claude in smaller ad-hoc groups. He encouraged users to try the feature.
Updated
Updated
Items with an AI score under 20 are hidden. Show low-relevance items
AIAnthropic's Noah Zweben announced that Claude Tag now supports Group DMs, letting users work with Claude in smaller ad-hoc groups. He encouraged users to try the feature.
AIMuse has released Muse Gadgets, open-source ESP32 firmware and a Linux SDK for building hardware devices that work with Muse. Developers can get an API token at gadgets.muse.ai and use a coding agent with the GitHub repository to build custom peripherals.
AIPerplexity has released several open source projects, including the pplx-decider-v1-27b multimodal decision model, the pplx-embed-v2-context-9b-preview contextual embeddings model, and the Lily local inference engine for Apple silicon. The post also lists the 0.6B on-device PII-Tracer classifier with its PII-TRACE benchmark, the WANDR research agent benchmark, and the Numbat and Bumblebee security tools, and says more open source releases are coming soon.
AIAnthropic's ClaudeDevs describes a "You should know" mod that spins off a sideagent to observe Claude's output. The mechanism is detailed in a newer blog post on getting started with Claude Code mods.
AIAnthropic is adding a new Claude Code plugin called "You should Know" that scans Claude's output for important information users might otherwise miss. It can be enabled with the command /plugin enable cc-plugin-you-should-know@builtin.
AIAnthropic released Claude Code v2.1.288, adding $.ui.selection() for mods, a built-in gh api for cloud sessions without the GitHub CLI, and --max-findings for /code-review. The release also fixes many issues, including mid-response API timeouts, resume and compaction bugs, and auto mode denials and model switching on Bedrock and Mantle.
AISGLang has released v0.5.20, bringing Intel XPU into standard releases alongside RL sampling masks that make rollouts more reliable with up to 52% faster decode. The update also adds Unified Radix Tree SWA branching-point caching, which the project says lifts cache hit rate about 20 points and cuts TTFT by roughly one-third, plus up to 12.5× faster ROCm model loading. New models named in the release include GLM-5.3-Flash, Qwen3.8-Flash-Next, K2 Horizon, Hy4-Preview, FastH3, and VDN-H3.
AISGLang's update adds a /v1/score endpoint that returns scores for requested labels such as Yes/No or A/B/C, avoiding the label loss of generate with top-k logprobs. Its multi-item scoring computes shared context once and keeps each candidate isolated, with 16-candidate p95 on Qwen3-8B dropping from 54.1 ms (Generate) to 20.6 ms.
AISGLang demonstrated Qwen3.8-27B as a multimodal decision model that beat Pokémon FireRed's Elite Four and champion with sub-100 ms decisions from live game state. The company says its native /v1/decisions API lets LLMs and VLMs be used for classification and scoring. It also announced /v1/systemone for running Jev-like open models with the TypeSafe SDK.
AISGLang has released v0.5.21 with a native Decisions API that turns an LLM or VLM into a low-latency classifier and scorer. The release also lets /v1/score rerank search or RAG results in one call, lets PD instances switch between prefill and decode without restarting, and adds support for models including DeepSeek-V4.1 Flash, Kimi K3, and GLM-5.3-Flash on AMD MI355X. The announcement reports a 22% faster first token on long prompts for DeepSeek-V4.1 Flash and 20.6% higher prefill throughput for Kimi K3 in PD serving.

AIAssemblyAI's Universal 3.6 Pro speech-to-text model is now available in LiveKit Inference, with 45% fewer wrong yes/no confirmations and about 30% less background speech transcribed. It supports 32 languages plus code-switching and endpointing that waits out phone numbers and emails, at the same $0.45/hr price, accessible by switching to universal-3-6-pro.

AIThe Linux Foundation Education has launched a PyTorch Certified Associate (PTCA) Certification Pathway that combines four self-paced learning modules with the PTCA exam. The pathway includes 15–17 hours of self-paced learning and hands-on labs covering tensors, data handling, model development, and performance optimization. The source recommends additional hands-on practice before taking the exam.
AIGitHub Copilot code review can now be requested through the REST and GraphQL APIs, with an optional review effort level set per request. Balanced became the default review effort level for new and existing repositories and organizations as of September 28, 2026, while users who explicitly selected Lite keep that setting. The changes are generally available to Copilot Pro, Pro+, Max, Business, and Enterprise plans.
AIElevenLabs has released Eleven v4, its most expressive voice model, in the ElevenReader app. Users can listen to articles, ebooks, or PDFs in a chosen voice across more than 90 languages.
AIAdaption Labs has made Invent, its tool for generating AI training datasets from a plain-language description, available through an API. Developers can reportedly produce AI-ready training datasets in minutes with a few lines of code, according to the quoted post. Documentation is available at docs.adaptionlabs.ai.

AIfal says H3 Max Reference-to-Video now supports first, middle, and last frame control. Users can set keyframes, choose when the middle frame appears, and add image or video references to guide the scene.
AIxAI has released an experimental TypeScript SDK, installable via npm install @xai-official/sdk, that covers text, voice, image, and video in one package. It provides access to the latest Grok models along with server-side tools including real-time X search, web search, code execution, and remote MCP.
AIDatabricks' new open-source meta-harness, Omnigent, lets multiple coding agents such as Claude Code and Codex share sessions, rules, and security policies in one system. A walkthrough by @leonvz demonstrates forking work across agents, multi-agent review and debate with Debby, and splitting implementation across subagents with Polly.

AICline Desktop now offers Connectors in beta, letting users link Gmail, Slack, Google Calendar, Linear, Sentry, Notion, and other apps in one click. Once connected, Cline can use these apps' tools to retrieve context and take actions on the user's behalf.
AIAs of October 2, 2026, GitHub deprecated Gemini 3.5 Flash, Gemini 3.6 Flash, Kimi K2.7 Code, and Claude Opus 4.7 across all Copilot experiences. Suggested replacements are Gemini 3.8 Flash, Kimi K3, and Claude Opus 5.5, and Enterprise administrators may need to enable them through model policies.
AIChatGPT now offers Finances, which can find forgotten subscriptions, flag unfamiliar or duplicate charges, and track recurring bills that have increased. It also provides weekly updates, monthly spending breakdowns, budget building, credit score tracking, debt payoff planning, emergency fund estimates, and investment mix and concentration views. Users can access it at
AIChatGPT's Finances feature is rolling out to Free and Go users in the U.S. Users can securely connect their accounts through Plaid and Experian to get answers based on their own financial information.
AILangChain's LangSmith Custom Apps lets agent teams build their own review UI over their traces and publish it directly into the workspace. Developers build the interface on their LangSmith data, while hosting, authentication, and permissions are handled by the platform.
AITypeSafe AI shows how to build long-document search by combining its Jev with PageIndex, with no vector database or embeddings required. The project is open source on GitHub under VectifyAI's jev-doc-search repository.
AICursor introduced Rollouts, a tool that writes a monitoring plan and watches changes as they deploy to catch regressions before users see them. When Rollouts detects a regression, it identifies the offending PR and opens an issue, and one click starts a cloud agent to fix it. Rollouts usage credits are included through Oct 3.
AIGitHub has made the Project HydraFusion research preview available in the GitHub Copilot app and @code, where it orchestrates multiple models rather than acting as a single model. New models from Anthropic (Fable 5.1 and Opus 5.5) and OpenAI (GPT-6.1 Sol) are also now selectable in the Copilot model picker.
AIOpenCode is offering inclusionAI's latest model, Ling-3.1-flash, for free. The model has 560B total parameters, 25B active parameters, and a 262K context window.
AIAnt Ling has made Ling-3.1-flash available on OpenRouter at no cost, inviting users to try it and share feedback. The post provides no details on model size, benchmarks, context length, or pricing beyond the free access.
AIHarrison Chase says LangChain is improving memory in managed Deepagents, noting that memory is difficult to get working well in company settings. The linked background post describes user memory in Managed Deep Agents 0.8, which lets an agent remember the people it works with.
AIMicrosoft's MAI-Transcribe-2-Streaming, MAI-Voice-2.1, and MAI-Voice-2.1-Flash are now live on LiveKit for building voice agents. LiveKit says MAI-Transcribe-2-Streaming debuts at #1 on the Artificial Analysis accuracy leaderboard, and suggests pairing it with MAI-Voice-2.1-Flash for efficient, expressive voice agents.
AIMicrosoft's Copilot Code is designed to help more people turn ideas into apps, workflows, and solutions for their work. Microsoft Copilot EVP Jacob Andreou discusses how the product expands who gets to build.

AIKilo Code is offering Ling 3.1 Flash for free until October 13, with the model served by Novita Labs. Ant Ling's background post describes the model as roughly 560B total parameters with about 25B active per token and a context window of up to 1M tokens. Ant Ling says it scores 1,673 Elo on GDPVal-AA v2.1, 75.16 on FrontierSWE, and 65.35 on HealthBench Professional, and plans to open-source it soon.
AIMicrosoft AI's MAI voice models are now accessible through OpenRouter, according to the announcement. The quoted OpenRouter post highlights MAI-Voice-2.1, a text-to-speech model that supports 23 languages with native accents and is priced at $22 per 1M characters.
AILlamaIndex released Extract v2.5, a set of document extraction agents that it says reach 93%–96%+ accuracy on long-list extraction, including records spanning pages. The post claims the agents outperform frontier VLMs, which it says stop early, miss repeated records, and struggle to attribute values to sources, while LlamaIndex attributes every extracted value to its source. The agents are available through LlamaParse.
AIGoogle Research introduced a next-generation Federated Learning system that uses Trusted Execution Environments to deliver verifiable differential privacy. The design moves computation server-side, which the post says cuts training times.

AIMicrosoft AI's Voice and Transcribe Streaming model is now available on Vercel for building agents, and the post claims it ranks first for quality and speed. The post says it is cheaper than other hyperscalers and 60% cheaper than Eleven Labs.
AIllama.cpp can now run decision models on-device, according to Clément Delangue of Hugging Face. He says the setup is free, fast, and private, and gives the command llama serve -hf ggml-org/Kev-4B-GGUF to start it.

AIGoogle's September 2026 roundup highlights Gemini 4 Argon, a frontier model with a 1-million-token output limit aimed at complex tasks such as cybersecurity defense. Argon is rolling out first to trusted cyber defenders through the Fairwind Program, with developer, enterprise, and consumer access to follow after guardrail feedback. The post also covers Gemini 3.8 Flash, Connected Apps in Gemini, and WeatherNext 3.
AIllama.cpp now supports decision models, which route tickets, moderate content, or choose an agent's next step by returning a probability for every option. Five open models from 144M to 27B parameters are supported at launch, and the team says more will follow in the coming days. Because most decision models do not need large GPUs, they are a good fit for llama.cpp, and a Hugging Face blog post explains how to set them up.
AIGeorgi Gerganov announced that decision models are now available in the latest llama.cpp builds through a new /v1/systemone endpoint. The endpoint supports local, private inference with multiple open models, and more are planned.