Skip to contentSkip to stories

Updated

#Deployment/Engineering

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 30

Sep 30Wed
  1. Google · Gemini appAI score42

    Google Gemini Adds Reusable Skills to Replace Gems Over Coming Months

    AIGoogle is rolling out skills in Gemini chat globally, letting users save frequently used instructions and invoke them by typing a forward slash and the skill name. Skills will replace Gems, which Google will remove starting in November for personal accounts, March 2027 for Workspace business, enterprise and nonprofit customers, and June 2027 for Workspace education customers. Gems will be automatically migrated into skills.

  2. Google Cloud · AI & Machine LearningAI score45

    Google Cloud Launches Preview of CLI Remote MCP Server for AI Agents

    AIGoogle Cloud has introduced the Google Cloud CLI remote MCP server in preview, giving AI agents access to gcloud and bq command-line operations through two tools, run_gcloud_command and run_bq_command. The server runs in an isolated, network-restricted execution sandbox on Google Cloud, so teams need no local CLI binaries, and calls are authenticated through Agent Identity, OAuth 2.0, and IAM, with Model Armor screening and Audit Logs available.

  3. Microsoft ResearchAI score46

    Machine learning system forecasts space-weather grid risk for 66,935 U.S. substations

    AIMicrosoft Research intern-developed machine learning pipeline forecasts location-specific geomagnetic risk for 66,935 substations in the continental United States. It combines solar-wind observations, AE and Dst forecasts, geological conductivity and grid data to estimate risk 30 to 60 minutes ahead. The pipeline detected nearly 80% of major space-weather events during the evaluation period.

  4. Google Cloud · AI & Machine LearningAI score41

    Google Cloud Rolls Out Agent Substrate, GKE Agent Sandbox RL Tools in September

    AIGoogle Cloud introduced GKE Agent Substrate, an open-source execution runtime it says can run millions of sandboxes with 10x higher density than standard container runtimes. It also made GKE Agent Sandbox optimized for reinforcement learning generally available, alongside an orchestration SDK and native RL gym integrations. Google said GKE Pod snapshots can reduce AI inference start-up by as much as 89%, based on internal tests.

  5. Lovable BlogAI score47

    Lovable Discloses TanStack Start Vulnerability CVE-2026-102989 and Protects Hosted Apps

    AILovable's security team found a vulnerability (CVE-2026-102989) in TanStack Start, which allows attackers to run unwanted JavaScript in visitors' browsers via crafted links. Lovable reported it to TanStack and deployed firewall protections for hosted apps while a fix was prepared, and affected projects will be automatically updated on their next change or via the Security page. Lovable says it found no evidence of exploitation in reviewed logs, and apps hosted elsewhere must apply the upstream update themselves.

  6. Google DeepMindAI score62

    Google DeepMind introduces SynthID Bio to watermark AI-designed proteins

    AIGoogle DeepMind introduced SynthID Bio, a watermarking method that embeds a detectable signature into AI-generated protein sequences and predicted structures. In wet-lab tests across three target proteins, watermarked binders matched unwatermarked versions in hit rate, binding affinity, and sequence diversity. The team is publishing its methods paper, open-sourcing code and in vitro data, and releasing weights to the research community.

    Why it matters: The report shows watermarks surviving wet-lab testing with unchanged binding and folding accuracy, offering a concrete tool for tracking AI-designed proteins in biosecurity screening.

  7. Google DeepMind · The KeywordAI score46

    Google DeepMind introduces SynthID Bio to watermark AI-designed proteins

    AIGoogle DeepMind has introduced SynthID Bio, a technology that embeds an imperceptible, verifiable watermark into AI-designed protein sequences and predicted 3D structures. In laboratory tests across target proteins, watermarked designs matched the performance and natural diversity of unwatermarked versions. The company says the watermark provides a provenance layer intended to strengthen biosecurity and preserve the integrity of open scientific databases.

  8. Azure BlogAI score36

    Azure Circular Centers recover value from retired hyperscale hardware

    AIMicrosoft says its Circular Centers now operate eight facilities across North America, Europe, and Asia Pacific to decide the next life of decommissioned Azure hardware. Last year, Microsoft achieved a 92% reuse and recycling rate for decommissioned servers and components. Since 2014, Azure cores per rack have increased about 13-fold while power for the same task fell roughly 90%.

  9. Baidu Inc.AI score23

    Baidu says full-stack AI integration drives value across chips, cloud, and models

    AIBaidu argues its full-stack AI architecture, spanning Kunlunxin chips, Baidu AI Cloud, ERNIE models, and applications, adds value when layers are optimized together. The post says AI-powered business reached 50% of General Business revenue in Q2 and cites Gartner's forecast that inference will account for 55% of AI-optimized IaaS spending in 2026.

  10. Allie K. MillerAI score23

    Ultrafast AI could let business meetings decide instead of delay

    AIAllie K. Miller argues that ultrafast AI could eliminate the "until" delays that stall business decisions, since tasks like research, analysis, and prototyping that once took hours can finish in minutes. She describes meetings where an always-on agent streams discussion in real time and dispatches side agents that return outputs during the meeting, so teams can decide rather than defer.

  11. Cloudflare Blog · AIAI score72

    Cloudflare launches Auto Router in AI Gateway to cut AI token spend

    AICloudflare has released Auto Router in public beta through AI Gateway, where setting the model to cloudflare/auto routes each request to a model judged capable enough for the task. Internal tests showed up to 30% cost savings against frontier models, and on a 97-task internal benchmark cloudflare/auto scored 86.6% at $0.0084 per success versus 96.6% at $0.0210 for Claude Opus 5.5. The router is free during beta.

    Why it matters: The source gives a benchmark table of success rates and costs per trial, showing how routing trades quality against price for a gateway deployment.

  12. AI SupremacyAI score40

    China's physical AI push spans humanoid robots, factories, and component supply chains

    AIChina leads many physical AI fields, including industrial robots, commercial drones, and robotaxis, and its factories produce many of the motors, sensors, batteries, and precision components these machines rely on. Unitree, a Hangzhou humanoid maker, went public on the Shanghai stock exchange in August 2026 at a $50 billion valuation. Most humanoids still rely on human remote control or preset programs, according to TMTPost.

  13. Hamel HusainAI score42

    Hamel Husain Tests Anthropic's Claude Eval Plugin on Leasing Assistant Traces

    AIHamel Husain reviewed Anthropic's new build_eval and hill-climb commands in the claude-api plugin for Claude Code, finding it useful for discovering issues like human handoff, formatting, and voice agent problems. He criticized it for pushing evaluator creation before data review, asking for label validation in Markdown files, and bundling four failure checks into one broad call-transfer evaluator. Husain says he would hold off on using it for now.

  14. X.PINAI score72

    DeepSeek releases Ascend versions of its core kernel toolkit

    AIDeepSeek has released an Ascend toolkit that mirrors its Nvidia components, including TileLang, DeepGEMM, DeepEP, TileKernels, FlashMLA and DeepSelect. It says every TileLang kernel used in its training now has a high-performance Ascend implementation. The post also reports that a 128-card Ascend 950 supernode, jointly optimized with Huawei, has key compute and communication tests approaching hardware limits.

    Image from @thexpin's post
  15. Mastra BlogAI score22

    Mastra Factory adds Jira, GitLab, and incident.io work intake integrations

    AIMastra Factory now supports Jira, GitLab, and incident.io as work intake sources, joining GitHub, Linear, and Slack. The intake column can be filtered by source, and each item can be moved through the pipeline until the work is complete. The integrations ship in the @mastra/factory package, with credentials configured under Settings → Work Intake.

Sep 29

Sep 29Tue
  1. TechNode · AIAI score54

    ByteDance's Doubao reportedly preparing personal AI agent codenamed Spell

    AIByteDance's Doubao is reportedly accelerating work on a personal AI agent codenamed Spell, which entered small-scale internal testing in April. According to Sina Tech, the project is being combined with core capabilities from Doubao's conversational AI team, with a public launch expected in the near future. The report places it alongside Doubao Work, an enterprise agent launched August 25, as a sign Doubao is pursuing parallel enterprise and consumer agent tracks.

  2. Allie K. MillerAI score62

    OpenAI launches Dots, a proactive always-on agent with a dedicated VM per Dot

    AIOpenAI launched Dots, and the author argues its always-on design and dedicated virtual machine for each Dot make it feel more like a persistent teammate. The post says the product is currently limited to one primary Dot, with a team of Dots promised later, and that early reviewers report bugs the author expects to be fixed over the next few weeks.

  3. SGLangAI score36

    SGLang adds native decision API for classification and scoring models

    AISGLang says it turned Qwen3.8-27B into a multimodal decision model that beat Pokémon FireRed's Elite Four and champion with sub-100 ms decisions from live game state. It introduces a native /v1/decisions endpoint for turning LLMs and VLMs into classification and scoring models. A /v1/systemone endpoint is also added so Jev-like open models can work with the TypeSafe SDK.

    Video from @sgl_project's post
  4. Factory NewsAI score42

    Factory Launches Generally Available Automations to Run Recurring Engineering Workflows

    AIFactory's Automations, now generally available, let users describe a recurring workflow, set a schedule or event trigger, and have its Droid run it, with the model chosen per task. Templates cover ticket-to-PR, code review, security audits, PR babysitting, and morning Slack briefs. Among enterprise organizations using Automations in the past 30 days, 52% used automated code review, 48% used security review, and 35% used AutoWiki.

  5. Fireworks AI BlogAI score51

    Fireworks explains how numerical mismatch and MoE routing can derail RL training

    AINumerical differences between a rollout engine and a trainer can make reinforcement learning collapse even when algorithm and data stay identical. In a GLM 5.2 experiment, reward fell from about 0.9 to under 0.2 around step 20 without alignment, while aligned numerics kept reward stable over 25 steps. A Qwen3.5-MoE investigation traced a significant mismatch to how expert outputs were combined, and router replay alone was judged insufficient.

  6. PromptArmor Threat IntelligenceAI score54

    Malicious Copilot Cowork skill hijacked AI gateway to exfiltrate files

    AIPromptArmor disclosed that a malicious Skill could hijack Copilot Cowork's AI gateway to spawn cloud agents that exfiltrate a victim's files to an attacker's server. No human approval was required, and any data Copilot could access was exposed. The vulnerability was reported to Microsoft on July 14, 2026, and Microsoft confirmed a fix on September 2, 2026.

  7. Google Developers BlogAI score47

    Google Details Sparse Attention Speedup for Video Diffusion on TPUs

    AIGoogle Developers Blog describes how Sparse VideoGen (SVG) routes video diffusion attention heads into spatial or temporal sparse masks and implements them as custom JAX and Pallas Splash Attention kernels on TPU v6e. In isolated single-chip tests with 75.6K tokens and 10 heads, the sparse variants retain about 38.87% of query-key pairs. The article argues that theoretical sparsity must be converted into hardware tile skipping to yield real speedups.

  8. vLLMAI score23

    vLLM presents keynote and talks at PyTorchCon North America

    AIThe vLLM project announced a strong presence at PyTorchCon North America, with core maintainer and Inferact CEO Simon Mo giving the keynote on scaling open frontier inference infrastructure. Other vLLM maintainers, including Nick Hill and Red Hat AI engineers, will lead a developer session and talks on agentic inference and attention.

  9. Microsoft Foundry BlogAI score30

    Why content extraction still matters in the GenAI era

    AIMicrosoft's Azure AI team argues that better models do not eliminate the need for a dedicated content extraction layer, since agents need trustworthy, structured, and auditable inputs. The post notes that building extraction directly on an LLM quickly demands chunking, layout parsing, grounding, normalization, and evaluation infrastructure. Microsoft positions Azure Document Intelligence and Azure Content Understanding in Foundry Tools as managed options for that layer.

  10. DatabricksAI score34

    Databricks adds GPT-6.1 Sol and Grok 4.7 on Unity Gateway

    AIDatabricks has made OpenAI's GPT-6.1 Sol and xAI's Grok 4.7 available on Unity Gateway the day they launched. The post says GPT-6.1 Sol leads the cost-quality Pareto frontier on OfficeQA Pro v2, while Grok 4.7 reaches the frontier on enterprise document parsing. Unity Gateway also offers access to 60+ other frontier and open models on Databricks.

    Video from @databricks's post