Skip to contentSkip to stories

Updated

#Deployment/Engineering

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 17

Sep 17Thu
  1. Google AI StudioAI score58

    Google AI Studio open-sources Speakeasy's OpenAPI SDK generator suite

    AIGoogle AI Studio announced that Speakeasy is open sourcing its OpenAPI client generation suite under AGPLv3, following a May 2026 vendor shutdown that disrupted Google's SDK pipeline. The suite covers SDK generation for 7 languages, an agent-native CLI generator, and a documentation MCP server generator. Google says its pipeline now serves six targets with roughly one engineer maintaining it.

  2. Google AI StudioAI score80

    Google updates Gemini managed agents with Files and Credentials APIs

    AIGoogle AI Studio released antigravity-preview-09-2026, an updated harness for Gemini managed agents, now live in the Interactions API and AI Studio and running on Gemini 3.8 Flash. The release adds a Files API for moving data into and out of the agent's sandbox and a Credentials API that stores secrets encrypted so the model never sees them.

    Why it matters: The post shows what changed in the agent harness and how the new Files and Credentials APIs keep secrets out of the model's context, useful for developers building agents.

  3. Sierra BlogAI score38

    Sierra Achieves AIUC-1 Certification for Its AI Agent Platform

    AISierra has become AIUC-1 certified after an independent audit by Schellman and testing by the Artificial Intelligence Underwriting Company (AIUC), a new standard for AI agents that tests resistance to manipulation and unauthorized access. Schellman found that Sierra met all applicable AIUC-1 requirements, and the technical evaluations recur at least quarterly with a full audit each year. The certification complements Sierra's existing SOC 2 Type II, ISO 27001, and ISO 42001 attestations.

  4. OpenBMBAI score40

    OpenMed and MiniCPM5-2B demo local agentic clinical AI workflow

    AIOpenMed paired with MiniCPM5-2B to demonstrate a local clinical AI workflow combining privacy-preserving data processing with a compact model's tool use and long-context reasoning. OpenMed masks sensitive identifiers and extracts clinical context before MiniCPM5-2B calls tools, compares lab results, and generates clinical handoffs with source references. The post presents this as an example of keeping inference on local, resource-constrained hardware.

    Image from @OpenBMB's post
  5. inclusionAI (Ant Ling) · new models on Hugging FaceAI score42

    inclusionAI releases Ming-Image-0.1-Design, a 6B text-to-image model for text-rich designs

    AIinclusionAI has released Ming-Image-0.1-Design, a 6B text-to-image model for UI, infographics, and posters that outputs RGBA images with transparent backgrounds. The model is available on Hugging Face and ModelScope under the MIT License. It runs at 2048 x 2048 with 12 sampling steps and a CFG scale of 1.0, validated on one CUDA GPU with 80 GiB VRAM.

  6. Z.aiAI score40

    GLM-5.3 helped build the inference stack serving GLM-5.3-Flash

    AIZ.ai reports that GLM-5.3 helped build and optimize the inference infrastructure for GLM-5.3-Flash. The system went from first successful run to production readiness in under two weeks, with end-to-end throughput tripling over the initial baseline. The team credited dense feedback from local correctness tests, execution traces, microbenchmarks, and end-to-end measurements for enabling targeted hypothesis testing.

  7. WanAI score44

    Wan3.0 generates single 30-second video shots with director-level control

    AIAlibaba's Wan3.0 video model now produces a single 30-second shot directly, up from 15 seconds and a year ago's 5-second clips. It adds director-level control and omni-reference input accepting up to five videos, letting creators generate long takes instead of stitching short clips. A filmmaker used Wan3.0 in a production workflow to make Soulscape and Johnny Mai.

    Video from @Alibaba_Wan's post

Sep 16

Sep 16Wed
  1. Amp NewsAI score50

    Amp Runners Now Serve Multiple Directories and Update Themselves

    AIAmp runners can now serve multiple directories, specified with repeated --dir flags or found automatically with --discover-dirs, which scans Git checkouts up to two levels deep by default. Runners also check for new releases about once an hour, install them, and restart into the new version once no thread is running, at most once every 12 hours, with auto-update disabled via amp.runner.autoUpdate.enabled: false.

  2. Google Developers BlogAI score38

    Google and Speakeasy open-source OpenAPI SDK generator suite under AGPLv3 license

    AISpeakeasy is open-sourcing its full OpenAPI client suite under the AGPLv3 license, including generators for seven languages (Python, TypeScript, Go, Java, C#, PHP, Ruby), an agent-native CLI generator, and a documentation MCP server generator. Google said the move followed the May 2026 shutdown of the SDK generation provider it had been using, which it cited as evidence that closed-source generators pose platform risk. Google's new Google GenAI SDKs for the Interactions, Agents, and Webhooks APIs were built with this pipeline across six targets.

  3. Greg BrockmanAI score62

    Databricks rolls out Astra to all engineers, reports 60% higher coding spend

    AIDatabricks rolled out Astra to every engineer, about 3,500 people, after a pilot with around 200 users. Engineers given Astra increased coding spend by roughly 60% compared to baseline. The company reports Astra outperforms Opus 5 and Sol 5.6 on highly complex system design tasks, but sees no clear gain on medium or low complexity coding. Astra gets a separate sub-budget in Unity Gateway to encourage selective use.

  4. Microsoft AI BlogAI score22

    Microsoft commits to AI in education with safeguards, educator control and student learning focus

    AIMicrosoft signed a landmark agreement with the American Federation of Teachers and introduced a Privacy & Safety Standard for Schools covering Microsoft Education products. The standard limits how student and educator data is used, requires human oversight for consequential decisions and keeps school-created knowledge owned by schools. Microsoft also introduced Teach in Microsoft 365 Copilot, an education-first AI experience for educators.

  5. Midjourney UpdatesAI score34

    Midjourney Alpha Site Adds Editing, Mobile, and Folder Fixes in 9/16/26 Update

    AIMidjourney's September 16, 2026 alpha site update fixes the v8.2 editor, adds a Korean language option, and makes its mobile and tablet layouts fill the screen. Users can now drag images into folders with the sidebar collapsed, use Add to Folder without existing folders, and delete uploads from a menu. Default parameters and sref previews are listed as upcoming work.

  6. Matei ZahariaAI score58

    Databricks reports engineering measurements from rolling out Astra to 3,500 engineers

    AIDatabricks rolled out Astra to all of its engineers and reported internal measurements. Engineers given Astra increased overall coding spend by around 60% compared with baseline, and Astra outperformed prior top models on highly complex system design tasks, while the gain on medium or low complexity tasks was unclear.

  7. TinkerAI score32

    Sundial trains Inkling-Small to fix LaTeX errors in under a second

    AISundial fine-tuned Thinking Machines' Inkling-Small with RLVR on 3,978 verified TeX.StackExchange fixes, using rewards for compilation and PDF match and penalties for removed content. The trained model fixes 83.7% of LaTeX errors in under one second at $0.0013 per fix, according to the post. Sundial says it is rolling out the model in its editor, applying fixes as suggestions and rebuilding the PDF.

  8. Google for DevelopersAI score38

    Three companies use Gemini agentic video understanding to cut token costs

    AIMosaic, Ponder Studio, and Revyl used early access to Google's Gemini Flash models to test agentic video understanding on long footage. Mosaic reports a 97% cut in median token usage and nearly double the ability to handle complex edits, while Ponder Studio reports a 0.967 F1 score and about 72% lower token costs for B-roll selection. Revyl says the approach improved mobile UI bug-catching accuracy by 65%. The capability is available now for video uploads and YouTube videos via the Gemini API.

  9. Baseten BlogAI score54

    Baseten launches Hosted Tools with web search for open-source models

    AIBaseten has launched Hosted Tools, starting with Baseten Grounded Inference, a server-side web search capability for models hosted on Baseten. Developers enable it by adding a hosted search tool to a Messages, Chat Completions, or Responses request, and the platform runs the search loop with partners Exa, Keenable, Parallel, and You.com. In Baseten's benchmarks, agents using the hosted tools saw a 15% reduction in end-to-end latency compared with client-side tools, and the feature is in playground preview with 25 RPM rate limits and $2 of free credits.

  10. Mike KriegerAI score46

    Claude Cowork and Chat Merge into One Unified Claude

    AIAnthropic is merging Claude Cowork and Chat into a single Claude starting today, which Mike Krieger says removes the friction of choosing which product to start with. Per the @claudeai announcement, Claude will carry tasks forward even after the laptop is closed, asking for clarification when needed while users keep final say. The rollout to Pro and Max plans will take place over the coming weeks.

  11. BAAI · new models on Hugging FaceAI score34

    BAAI and Peking University release Brainμ-Spike spike camera image reconstruction model

    AIPeking University's Yu Zhaofei team and the Beijing Academy of Artificial Intelligence (BAAI) released Brainμ-Spike, a small convolutional network for spike camera image reconstruction that is paired with the Brainμ model. The package includes weights, inference scripts, and evaluation tools, but the base large model and LoRA weights are not yet released, so the full generation pipeline cannot run from this repository alone.

Sep 15

Sep 15Tue
  1. Tencent · new models on Hugging FaceAI score44

    Tencent releases WeVisDoc-4B, a document parser that leads OmniDocBench v1.6

    AITencent's WeVisDoc-4B, fine-tuned from Qwen3-VL-4B-Instruct, converts page images into structured Markdown with LaTeX formulas and HTML tables. It scores 95.38 Overall on OmniDocBench v1.6 and a mean Overall of 75.54 across three PureDocBench tracks, ranking first among compared end-to-end parsers in all four reported settings. The model is available on Hugging Face and runs through vLLM, which requires version 0.11.1 or later.

  2. Zed BlogAI score72

    Zed launches Delta public beta to replace pull requests with agent threads

    AIZed has launched the public beta of Delta, a multiplayer environment for coding with agents and reviewing their work, which replaces pull requests with shared threads. Delta is built on DeltaDB, which records edits and messages between Git commits, and it is free during the beta, with paid plans for individuals and teams to follow.

    Why it matters: The post explains how Delta replaces pull requests with shared agent threads and DeltaDB, showing a concrete alternative to the GitHub review workflow.