Skip to contentSkip to stories

Updated

#Deployment/Engineering

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 10

Sep 10Thu
  1. DeepSeekOfficialAI score72

    DeepSeek V4.1-Flash goes live on its API with native multimodal support

    AIDeepSeek says V4.1-Flash is now live on its API with native multimodal support, accessed through the model name deepseek-flash. The older V4-Flash and V4-Flash-Vision-Exp are retired, while deepseek-v4-flash and deepseek-v4-flash-vision-exp temporarily route to V4.1-Flash. Requests to deepseek-v4-pro will route to V4.1-Flash at V4.1-Flash rates starting 04:00 UTC on Sept 14, 2026, until V4.1-Pro launches.

  2. DeepSeekOfficialAI score46

    DeepSeek V4.1-Flash cuts KV cache to 1/4 HBM and 1/8 SSD

    AIDeepSeek says its V4.1-Flash model needs only 1/4 the HBM and 1/8 the SSD storage for its KV cache compared with the previous generation. Because cache-hit charges often make up a large share of agent costs, the company says the compressed cache significantly reduces those costs.

    Image from @deepseek_ai's post
  3. DeepSeekOfficialAI score38

    DeepSeek unveils 552B MoE model with asymmetric encoder-decoder design

    AIDeepSeek has introduced a 552B-parameter MoE model built on a new Causal Encoder–Decoder architecture, activating 8B parameters for input and 16B for output. The company says new pre-training methods and larger-scale RL post-training deliver benchmark results ahead of flagship models, including DeepSeek-V4-Pro.

    Image from @deepseek_ai's post
  4. The Register · AINewsAI score38

    Microsoft launches AI-powered Dynamics 365 Activate tool to migrate Salesforce customers

    AIMicrosoft has released Dynamics 365 Activate as a public preview, an AI-powered tool that converts Salesforce implementations to Dynamics 365. The company said it will add more CRM and ERP migration scenarios later this year. The tool profiles data, entities, relationships, and customizations to give implementation teams a migration blueprint.

  5. DeepSeek API NewsOfficialAI score72

    DeepSeek releases V4.1-Flash with native multimodal support and API updates

    AIDeepSeek officially released DeepSeek-V4.1-Flash, the smallest model in its new architecture family, with native multimodal visual understanding. The API now serves it under the model name deepseek-flash, while V4 Flash and V4 Flash Vision Exp were retired and routed to V4.1 Flash. API prices were reduced with the release, and V4 Pro remains available after September 14, 2026.

    Why it matters: The release lists benchmark results alongside API model-name changes and retirements, so developers can check both capability claims and migration steps.

Sep 9

Sep 9Wed
  1. BAAI · new models on Hugging FaceOfficialAI score24

    BAAI open-sources EPT, UniPath, and MiSI AIDD molecular and crystal modeling resources

    AIBAAI released open-source resources for three AIDD projects on Hugging Face: EPT, an equivariant pretrained transformer for unified 3D molecular representation learning, and UniPath, a learnable-time flow matching method for crystal structure and energy prediction. The repository mirrors their GitHub source code and READMEs, with setup, preprocessing, training, and evaluation documentation. The MiSI benchmark is released separately on Hugging Face.

  2. Cursor ChangelogOfficialAI score73

    Cursor launches Projects for long-running, multi-agent coding work

    AICursor is launching Projects, a beta feature for larger work such as a feature, migration, or full app, rolling out to all users starting today. A coordinator agent plans the work, delegates it to implementing agents that can run in parallel, and runs on a cloud computer so it continues when the laptop is closed. Each Project keeps shared context files synced across cloud and local machines, and subscriptions let the coordinator act on Slack channels, schedules, or PRs without a prompt.

    Why it matters: The source details how a coordinator agent plans, delegates, and syncs shared context across cloud and local machines, useful for judging how long-running agent work might fit a team's workflow.

  3. Together AI BlogOfficialAI score42

    Together AI Launches Preemptible GPU Compute at 50% of On-Demand Price

    AITogether AI has launched a public preview of preemptible compute for Together GPU Clusters on Kubernetes in all regions, billed sub-hourly at a flat 50% of the on-demand rate. Preemptible nodes can be reclaimed when capacity is needed elsewhere, with a five-minute drain window for checkpointing before removal. The rate stays fixed rather than tracking a spot market, and the cluster automatically refills its preemptible target as capacity becomes available.

  4. Fireworks AI BlogOfficialAI score60

    Genspark's Gen-1 Slides matches Opus 5 decks at about one-tenth the cost per deck

    AIGenspark and Fireworks Lab post-trained the open-weight MiniMax M3 into Gen-1 Slides, a model that plans, writes, and checks slide decks end-to-end. On Genspark's evaluation it matches Claude Opus 5 at about 1/17 of its input-token list price, roughly 90% less per finished deck. In production it cut low-rated decks from 18% to 3.6% over the base model.

    Why it matters: The post explains a post-training pipeline with reward design, curriculum, and numerical fixes, showing how a cheaper model was tuned toward a frontier quality bar.

  5. Claude Apps Release NotesOfficialAI score39

    Anthropic launches Smart reports beta for Claude Enterprise teams

    AIAnthropic has launched Smart reports in beta for Claude Enterprise plans. The reports analyze how a team uses Claude, covering work completed, costs, session friction, and repeated patterns worth packaging as shared skills.

  6. Greg BrockmanXAI score32

    OpenAI's Astra helps medical students link anatomy diagrams and CT scans

    AIThe main post only says "Astra for medical education:" with no details, so the source offers little beyond that label. The quoted post claims a GPT-6 Astra service links anatomical diagrams, CT cross-sections, and 3D relationships in one spatial tool for medical students.

  7. Perplexity DevelopersOfficialAI score20

    Perplexity's MCP tools now run on its Agent API

    AIPerplexity says its MCP tools perplexity_ask, perplexity_research, and perplexity_reason run on Agent API. The company describes Agent API as a single, multi-provider endpoint with built-in tools and dynamic presets.

  8. Microsoft Foundry BlogOfficialAI score62

    Microsoft Foundry's July and August 2026 updates bring Hosted Agents and Toolboxes to GA

    AIMicrosoft Foundry's July and August 2026 updates make Hosted Agents, Voice Live integration, and Toolboxes generally available. The post adds Claude tools on Azure, Model Router region and model pool changes, Foundry Local preview features, and updated Python, JavaScript, Java, and .NET SDK versions with migration notes.

    Why it matters: The roundup links each GA and preview change to code examples, migration notes, and runtime requirements, which helps developers judge what to upgrade and test first.

  9. Sherwin WuXAI score46

    ChatGPT Voice adds GPT-6 Astra for Pro users

    AIOpenAI's ChatGPT Voice can now use GPT-6 Astra, available to Pro users, for search and reasoning tasks. Users can select any model and effort level, and voice will use the chosen model when it needs to search or reason.

  10. LlamaIndex 🦙OfficialAI score23

    LlamaParse now available as a ChatGPT connector for document parsing

    AILlamaIndex has made LlamaParse available in the ChatGPT plugin directory, following its earlier Claude integration. The connector parses scanned, table-heavy, and chart-filled documents into Markdown, JSON, or HTML, extracts fields into a user-defined schema, searches document collections, and classifies and splits files into sections.

    Video from @llama_index's post
  11. RadixArkOfficialAI score38

    RadixArk's Miles integrates SGLang for fast, aligned post-training rollouts

    AIRadixArk says its Miles framework natively supports SGLang for fast rollouts while keeping rollout and training aligned for reliable post-training at scale. The post thanks the community for contributions and feedback shaping Miles. A related post from @adarshxs describes Miles v0.1 running fully async agentic RL on a 744B MoE across 64 GB300 GPUs.

  12. Google DeepMind · YouTubeOfficialAI score38

    How AI is transforming weather prediction, featuring WeatherNext 3

    AIGoogle DeepMind's Peter Battaglia discusses how machine learning is changing global weather forecasting, including early warnings for storms such as Hurricane Melissa. The episode covers traditional physics-based models versus AI models and probabilistic forecasting, and highlights WeatherNext 3 as Google DeepMind's most advanced global weather AI model yet.

  13. Mistral AIOfficialAI score54

    Mistral details how AI agents migrated 40,000 lines of Fortran to C++

    AIMistral AI helped a European energy operator migrate 40,000 lines of Fortran 77 to C++ for a reservoir simulator with no test suite. The post explains a parity harness that checks numerical agreement between the two codebases, and a workflow where agents coder, tester, and reviewer migrate modules under human review. Its authors note the approach covered the self-contained first sprint of 40,000 of 300,000 lines and that dependent systems would bring additional challenges.

  14. Ahead of AI (Sebastian Raschka)BlogAI score46

    GPT-6 Astra Leads Coding and Math Benchmarks, Shows Strong Computer Use

    AIOpenAI's GPT-6 Astra scores 99.9% on ARC-AGI-3, versus 7.8% for GPT-5.6 Sol, and leads Raschka's coding and math tests. Its strongest showing is in graphics and computer-use tasks, such as redrawing an image in a browser-based Paint app. The author notes that Artificial Analysis shows Astra at the frontier but not pulling far ahead on its Coding Agent Index.

  15. The Register · AINewsAI score38

    Microsoft Edge Team Says AI-Assisted Coding Is Overloading Extension Reviews

    AIMicrosoft's Edge team said rapid adoption of AI-assisted coding has increased the volume of browser extension submissions, slowing its review pipeline. The team is adding automation to validation checks and says review standards will not change. It also plans to refresh the "Featured" badge every 15 days.

Sep 8

Sep 8Tue
  1. Perplexity DevelopersOfficialAI score30

    Perplexity Search API now available in Hermes Agent

    AIPerplexity says its Search API is now available in Hermes Agent, giving it access to an index of more than 400 billion URLs. The API returns real-time results with snippets ranked by relevance.

    Video from @perplexitydevs's post
  2. Google Developers BlogOfficialAI score72

    Google releases ADK for Kotlin 1.0 for building production AI agents

    AIGoogle announced general availability of ADK for Kotlin 1.0, a Kotlin Multiplatform framework for building AI agents on servers and Android. Version 1.0 reaches feature parity with ADK 1.0 Core and adds Android extensions for on-device models, cloud Gemini via Firebase AI Logic, and persistent sessions and memory with Room and AppSearch. The post includes a server-side incident triage example using KSP-generated tools and skills, plus an Android financial assistant example with human confirmation for transfers.

    Why it matters: The post names the new Android and server-side capabilities and the code setup, helping Kotlin developers judge whether ADK fits their agent projects.

  3. Factory NewsOfficialAI score34

    Factory Now on Claude Marketplace for Enterprise Autonomous Software Development

    AIFactory is now available on the Claude Marketplace, letting enterprise customers apply their committed Anthropic spend toward its autonomous software development platform. The platform automates the software development lifecycle, covering planning, implementation, testing, and security within one system, with enterprise deployment options that keep execution close to customers' code and infrastructure.

  4. Jazzyear · InsightsNewsAI score62

    Arm expands from mobile IP into cloud, edge, and physical AI at Shanghai event

    AIAt Arm Everywhere China on September 8, 2026, Arm launched products spanning data center CPUs, mobile compute subsystems, and robotics platforms. The article says CSS for Mobile 2 integrates CPU, GPU, and neural accelerator for agent AI on phones, and Arm's Neoverse CSS N4 and AGI CPU target agent sandboxes in data centers. It also reports that Arm's Total Design ecosystem now covers over 80 partners for physical AI.

  5. Gemini NotebookOfficialAI score22

    Gemini Notebook rolls out upgraded mobile experience and new study features

    AIGemini Notebook's upgraded experience is rolling out across all mobile devices this week, according to the account. Web users can opt in via Settings to be notified when their artifacts are ready, and after a Quiz, Chat can recommend what to study next or generate targeted flashcards for weak spots.

  6. InferactOfficialAI score42

    Inferact reports open models hit 130K tokens/GPU-sec on agentic workloads

    AIInferact says months of vLLM tuning for agentic workloads, validated on SemiAnalysis's AgentX benchmark, let open-source models reach up to 130K tokens per GPU-second. The company claims this is 106 times cheaper than Opus 5 API pricing. The work is described as part of a vLLM blog post covering architecture, framework, and runtime optimizations.

  7. ReplicateOfficialAI score46

    GPT Image 2.5 Launches on Replicate for Precise Image Editing

    AIOpenAI's GPT Image 2.5 is now available on Replicate, offering precise image editing, sharper details, and higher consistency across multi-turn edits. Two model pages are provided, openai/gpt-image-2.5-sunburst and openai/gpt-image-2.5-flare.

    Image from @replicate's post
  8. GranolaOfficialAI score34

    Granola adds webhooks to its API for note and folder events

    AIGranola has added webhooks to its API, letting developers notify their systems or apps when a note is generated or edited. The notifications also cover notes shared with a user or added to a folder.

    Image from @meetgranola's post
  9. Werner VogelsXAI score50

    Werner Vogels highlights Kiro Crew's memory system drawing on brain evolution

    AIWerner Vogels says that after spending time with Kiro Crew since its launch, its memory system stands out for deciding what to keep, compress, and let go. He notes that Amazon engineers, starting from engineering constraints, arrived at an approach resembling the brain's evolved architecture. Per the referenced post, Kiro Crew is a persistent workspace that retains project context across sessions and runs scheduled jobs.

  10. Demis HassabisXAI score73

    Google DeepMind launches AlphaGenome Atlas to predict impact of human DNA variants

    AIDemis Hassabis announced AlphaGenome Atlas, a searchable AI database that maps the predicted impact of all 9 billion possible single-letter DNA changes. The post says it can help scientists better understand disease and is freely available for academic research.

    Why it matters: The post describes a searchable database of predicted effects for all 9 billion single-letter DNA variants, which is useful for researchers tracing disease-related genetic changes.

  11. Google DeepMind · YouTubeOfficialAI score60

    Google DeepMind launches AlphaGenome Atlas for mapping genetic variant effects

    AIGoogle DeepMind introduced AlphaGenome Atlas, an AI-powered database charting the molecular impact of every possible genetic variant. Scientists are already using it to investigate unsolved rare diseases and map rare mutations linked to complex traits.

    Why it matters: The source names a concrete use case, finding disease-causing DNA variants, which shows how the database could support rare disease research.