Skip to contentSkip to stories

Updated

#Deployment/Engineering

Showing low-relevance items too. Hide low-relevance items

Sep 16

Sep 16Wed
  1. Mike KriegerXAI score46

    Claude Cowork and Chat Merge into One Unified Claude

    AIAnthropic is merging Claude Cowork and Chat into a single Claude starting today, which Mike Krieger says removes the friction of choosing which product to start with. Per the @claudeai announcement, Claude will carry tasks forward even after the laptop is closed, asking for clarification when needed while users keep final say. The rollout to Pro and Max plans will take place over the coming weeks.

  2. Felix RiesebergXAI score10

    Anthropic's new feature rolls out to Pro and Max users

    AIAnthropic is rolling out an unnamed feature to Pro and Max subscribers over the next few weeks. The company says it is staging the release to ensure exceptional reliability, and asks users for patience.

  3. Felix RiesebergXAI score32

    Cowork users can resume chats, projects, and skills after update

    AIAnthropic's Felix Rieseberg says Cowork users need to take no action, since chats, projects, artifacts, connectors, and skills all remain available. Users can simply open the app and continue where they left off.

  4. hardmaruXAI score16

    Sakana AI expands go-to-market team to deploy products globally

    AISakana AI says it shipped Sakana Chat, Namazu, Sakana Translate, Sakana Marlin, Fugu, Fugu Cyber, and Fugu Max this year and is now deploying them to enterprises, manufacturers, financial institutions, and government agencies in Japan and internationally. The company is massively expanding its GTM team and recruiting a Product Sales & Account Executive and a Forward Deployed Engineer (GTM) for founding roles.

    Image from @hardmaru's post
  5. LlamaIndex 🦙OfficialAI score14

    LlamaIndex webinar on insurance document pipelines with LlamaParse and Extract

    AILlamaIndex Solutions Architect Abrar Mahi will host a webinar on turning insurance documents such as accord forms and policy documents into structured data for underwriting, policy review, and claims. The session covers extracting policy, property, and claims history into a defined schema, verifying values with citations and bounding boxes, and using confidence scores with validation rules to route items to human review.

    Image from @llama_index's post
  6. BAAI · new models on Hugging FaceOfficialAI score34

    BAAI and Peking University release Brainμ-Spike spike camera image reconstruction model

    AIPeking University's Yu Zhaofei team and the Beijing Academy of Artificial Intelligence (BAAI) released Brainμ-Spike, a small convolutional network for spike camera image reconstruction that is paired with the Brainμ model. The package includes weights, inference scripts, and evaluation tools, but the base large model and LoRA weights are not yet released, so the full generation pipeline cannot run from this repository alone.

  7. Varun MohanXAI score14

    Antigravity resets rate limits after fixing Gemini 3.8 Flash errors

    AIAntigravity has fixed the high-load errors affecting Gemini 3.8 Flash requests and rolled out quota resets for all users. The earlier post said the team was actively resolving the request failures before resetting rate limits.

Sep 15

Sep 15Tue
  1. Tencent · new models on Hugging FaceOfficialAI score44

    Tencent releases WeVisDoc-4B, a document parser that leads OmniDocBench v1.6

    AITencent's WeVisDoc-4B, fine-tuned from Qwen3-VL-4B-Instruct, converts page images into structured Markdown with LaTeX formulas and HTML tables. It scores 95.38 Overall on OmniDocBench v1.6 and a mean Overall of 75.54 across three PureDocBench tracks, ranking first among compared end-to-end parsers in all four reported settings. The model is available on Hugging Face and runs through vLLM, which requires version 0.11.1 or later.

  2. Noah ZwebenXAI score17

    Anthropic offers Claude Tag office hours for on-call triage feedback

    AIAnthropic is hosting office hours for teams interested in using Claude Tag for on-call work, and it is asking Team or Enterprise plan users to share triage feedback. Claude Tag can start investigating when a Slack alert fires by pulling metrics, diffing deploys, and checking flags to propose a likely cause and fix. Sign-up is through a Google Calendar booking link.

  3. Zed BlogOfficialAI score72

    Zed launches Delta public beta to replace pull requests with agent threads

    AIZed has launched the public beta of Delta, a multiplayer environment for coding with agents and reviewing their work, which replaces pull requests with shared threads. Delta is built on DeltaDB, which records edits and messages between Git commits, and it is free during the beta, with paid plans for individuals and teams to follow.

    Why it matters: The post explains how Delta replaces pull requests with shared agent threads and DeltaDB, showing a concrete alternative to the GitHub review workflow.

  4. Claude Apps Release NotesOfficialAI score72

    Claude Cowork moves into every conversation, adding designs, slides, and docs

    AIClaude now makes Cowork capabilities available from any conversation without choosing a mode first, with chats, tasks, projects, connectors, and skills carrying over. Users can also create designs, decks, and docs in any conversation, including Claude Code and the Artifacts tab, and edit them with Claude.

    Why it matters: The release merges Cowork tasks into ordinary chats and adds design, slide, and doc creation, changing how Claude users start larger work.

  5. Google Developers BlogOfficialAI score46

    Google Launches Agent Anomaly Detection in Private Preview on Gemini Enterprise Agent Platform

    AIGoogle has put Agent Anomaly Detection into Private Preview on the Gemini Enterprise Agent Platform, a reasoning-based audit layer that reviews agent reasoning traces, tool calls, and execution flow to flag behavioral anomalies and policy violations. It runs asynchronously without adding runtime latency and publishes findings to Security Command Center. The preview requires ADK 1.2 or later.

  6. xAI News (Grok)OfficialAI score43

    Grok Build Adds Memory That Saves Project Notes Between Sessions

    AIGrok Build now has memory, which records conventions, decisions, and project facts after each completed turn and reads them in later sessions. Notes are stored per project plus a global set, and the /dream command organizes them into topic files while /memory opens a read-only browser. The feature is available now and applies to new sessions.

  7. Google AntigravityOfficialAI score37

    Antigravity adds permissions system and sandboxed command execution

    AIGoogle Antigravity says its new permissions system reduces the number of commands users must approve, as improvements to its sandbox let commands run automatically in an isolated environment. The sandbox has no network access by default, keeping the user's machine protected while the agent can do more out of the box. The update is rolling out today on macOS and Linux.

    Image from @antigravity's post
  8. Chip HuyenXAI score40

    Jev model chooses from predefined outputs, promising very cheap inference

    AIChip Huyen praises an approach where models select from predefined values rather than generating freeform text, which she sees as useful for data labeling and fixed-action tasks. She notes that how reasoning would work is unclear, but the approach is very cheap because output tokens are free.

    Image from @chipro's post
  9. VercelOfficialAI score22

    Delphi ships 100+ deploys daily on Vercel's Python backend

    AIDelphi, which turns experts' knowledge into digital minds, runs its Python backend on Vercel with a 10-person team and no dedicated infrastructure role. The stack uses Vercel Workflows for long-running agents and Vercel Queues for background jobs. The team reports more than 100 production deploys a day.

  10. Microsoft Foundry BlogOfficialAI score32

    Microsoft Launches Foundry Dev Pack to Install Foundry Development Tools in One Command

    AIMicrosoft has launched Foundry Dev Pack, an all-in-one installer that sets up tools for Microsoft Foundry development across the terminal, IDE, and coding agents. Depending on the environment, it installs Azure CLI (az), Azure Developer CLI (azd) with the Microsoft Foundry Extension for azd, the Microsoft Foundry Skill, the Microsoft Foundry Toolkit for Visual Studio Code, and Foundry Canvas (preview), with the last two conditional on VS Code or GitHub Copilot App being present.

  11. LiveKitOfficialAI score31

    LiveKit Agents adds Gemini 3.8 Live for low-latency and extended-thinking voice agents

    AILiveKit Agents now supports Gemini 3.8 Live, letting developers choose gemini-3.8-live for low-latency audio or gemini-3.8-live-extended-thinking for longer asynchronous reasoning. Developers can switch between the two modes without changing their stack. The post points readers to LiveKit's documentation for more details.

    Video from @livekit's post
  12. Josh WoodwardXAI score15

    Google adds new study tools to Gemini notebooks in September 2026

    AIGoogle has announced new study tools for Gemini notebooks, detailed in a September 2026 blog post. The post itself contains only a link and the single word "The features," so no specific tools, capabilities, or figures can be confirmed from this source.

  13. Greg BrockmanXAI score46

    ChatGPT Work adds Data agent for dashboards and actions on company data

    AIOpenAI's Greg Brockman says ChatGPT Work can operate over and act on a company's data, including building dashboards, by connecting existing tools such as PowerBI, Tableau, Clickhouse, Oracle BI, and AWS Redshift. The linked ChatGPT announcement describes a Data agent with a Data Plugin that turns company data into answers, interactive dashboards, and actions through conversation.

  14. Google AIOfficialAI score72

    Google rolls out Gemini 3.8 Live and Extended Thinking across consumer, developer, and enterprise channels

    AIGoogle is rolling out Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking across several channels. Consumers get them in Search Live and Gemini Live, developers get public preview access through the Gemini API, and enterprises get private preview through Gemini Enterprise, with Customer Experience support coming soon.

    Why it matters: The post lays out where each Gemini 3.8 Live variant reaches consumers, developers, and enterprises, which clarifies access paths for a voice model release.

  15. Google DeepMindOfficialAI score72

    Google DeepMind releases Gemini 3.8 Live models for real-time voice agents

    AIGoogle DeepMind introduced Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking, two live dialogue models for voice agents. Extended Thinking scores 82.6 on Artificial Analysis' Speech to Speech Quality Index, 68.6% on τ-Voice, and 97.7% on Big Bench Audio. Gemini 3.8 Live is rolling out now in the Gemini API, Google AI Studio, and Search Live, with enterprise access in private preview.

    Why it matters: The release covers a voice model's benchmark results and availability across developer, enterprise, and consumer products, useful for judging voice agent options.

  16. Cognition Blog (Devin, Windsurf)OfficialAI score60

    Cognition and AWS sign multi-year deal to deploy Devin for enterprise modernization

    AICognition and AWS have entered a multi-year Strategic Collaboration Agreement to help enterprises deploy the Devin autonomous engineer in production. Devin can be purchased through AWS Marketplace, and the companies are exploring deeper engineering integrations within customers' AWS environments. Mercedes-Benz reportedly used Devin to analyze more than 200,000 lines of COBOL, reducing an estimated eight-month modernization project to eight days.

    Why it matters: The collaboration shows how an autonomous coding agent is being packaged for enterprise legacy modernization inside existing AWS environments, with concrete customer migration figures.

  17. NVIDIA · new models on Hugging FaceOfficialAI score34

    NVIDIA Releases RT-DETR Hand Detection v1.0 for Real-Time RGB Hand Localization

    AINVIDIA's RT-DETR Hand Detection v1.0 detects and localizes left and right hands in RGB images, outputting 2D bounding boxes with per-hand confidence scores in a single pass. The model, built on RT-DETRv2-S with HGNetv2-S backbone and about 20M parameters, is intended as a region-of-interest stage for downstream 3D hand pose estimation and is exported to ONNX. The source describes it as for demonstration purposes rather than production use, runs on NVIDIA Lovelace GPUs under Linux, and is licensed under the NVIDIA Software and Model Evaluation License.

  18. RadixArkOfficialAI score42

    Periodic Labs builds Neon on SGLang and Miles for 2.5x faster inference

    AIPeriodic Labs chose SGLang and Miles to build Neon, an open-source model it says surpasses GPT-6 Astra on its analysis benchmark after mid-training and RL on 1,300 H200s. RadixArk says Periodic extended both frameworks for scientific RL at trillion-parameter scale, delivering more efficient training, lower memory use, and 2.5x faster inference. The work has been contributed back to both projects.

  19. Dwarkesh PatelXAI score38

    Dwarkesh Patel on Materials Synthesis Search and Depth-First Experimentation

    AIDwarkesh Patel reports that materials synthesis experiment spaces are very wide but amenable to depth-first search, where each next experiment becomes better designed and more informative as data accumulates. The post is a lab visit reaction and does not name specific models, figures, or results.

  20. LlamaIndex 🦙OfficialAI score22

    LlamaIndex Moves Off Stainless for LlamaParse SDK Generation

    AILlamaIndex says Stainless helped it keep LlamaParse SDKs current and pushed it to make the API's names and schemas more consistent. With the Stainless team joining Anthropic, George He and Yong Park explain what worked, what they learned, and why changing SDK generators needs careful handling.

    Image from @llama_index's post
  21. Google · Innovation & AIOfficialAI score44

    Google says its language tools now support over 300 languages used by 7 billion people

    AIGoogle says its technologies now support more than 300 languages spoken by 7 billion people, representing 86% of the global population. The company also released its AI & Economy ATLAS, which it describes as a look at how people are using AI globally. The post highlights recent AI science work, including AlphaGenome Atlas, WeatherNext 3, and a Planetary Prediction Engine.

  22. Lovable BlogOfficialAI score44

    Lovable and Salesforce partner so teams can build apps inside Salesforce workflows

    AILovable and Salesforce are working together so apps and agents built with Lovable can read and write Salesforce data through Headless 360, using each user's own Salesforce permissions. Teams can publish read-only apps into Salesforce, mention @Lovable in Slack to build apps, and install agents into a Slack workspace.

  23. Google · AI blogOfficialAI score14

    Google spotlights AI projects for disease, disaster prediction, education, and economic opportunity

    AIGoogle is showcasing how partners are applying AI to societal challenges, including making disease detectable, treatable, and preventable, predicting natural disasters, expanding learning, and unlocking economic opportunities. The source describes these efforts as measurable real-world impact but provides no specific models, figures, or benchmarks.

  24. Baseten BlogOfficialAI score40

    LangChain uses Baseten Loops to train custom models for LangSmith Engine

    AILangChain is using Baseten Loops, a managed fine-tuning service, to train custom models for LangSmith Engine, its in-platform agent that debugs and improves AI agents. The article says LangChain fine-tunes large open-weight models on agent traces and trains smaller open-weight models such as Qwen for tasks like failure-mode categorization. Baseten Loops supports supervised fine-tuning, reinforcement learning, and long-context workloads, and lets checkpoints be evaluated and deployed directly to inference.

  25. Air Street PressBlogAI score39

    Air Street Capital leads $40 million Series A in Jack & Jill, an AI career agent platform

    AIAir Street Capital led Jack & Jill's $40 million Series A, with Madrona joining and Creandum and Entrepreneurs First investing again, following a $20 million seed round less than a year earlier. Jack & Jill uses AI agents named Jack, which helps candidates plan career moves and search job postings, and Jill, which helps companies recruit from opted-in candidates. The company says it has arranged 25,000 interviews and plans 5,000 more each month.

  26. Kilo (acq. by Anaconda)OfficialAI score22

    Kilo App launches on Product Hunt for iOS and Android

    AIKilo announces that its Kilo App is live on Product Hunt, letting users start coding agents, check sessions, and review pull requests from iOS and Android. The company asks supporters to upvote or comment on its Product Hunt listing.

    Image from @kilocode's post