Updated
#Deployment/Engineering
Updated
Showing low-relevance items too. Hide low-relevance items
Sep 9
v0@v0OfficialAI score32
v0@v0OfficialAI score33Vercel integrations like Resend and MongoDB Atlas now available in v0
AIVercel integrations are coming directly to v0, with Resend, Amazon OpenSearch, MongoDB Atlas, Algolia, and Clerk now one click away from the chat. Setup is automatic and skills load automatically, so users can build without leaving v0.

LlamaIndex 🦙@llama_indexOfficialAI score23LlamaParse now available as a ChatGPT connector for document parsing
AILlamaIndex has made LlamaParse available in the ChatGPT plugin directory, following its earlier Claude integration. The connector parses scanned, table-heavy, and chart-filled documents into Markdown, JSON, or HTML, extracts fields into a user-defined schema, searches document collections, and classifies and splits files into sections.

RadixArk@radixarkOfficialAI score38RadixArk's Miles integrates SGLang for fast, aligned post-training rollouts
AIRadixArk says its Miles framework natively supports SGLang for fast rollouts while keeping rollout and training aligned for reliable post-training at scale. The post thanks the community for contributions and feedback shaping Miles. A related post from @adarshxs describes Miles v0.1 running fully async agentic RL on a 744B MoE across 64 GB300 GPUs.
RadixArk@radixarkOfficialAI score43RadixArk publishes full technical report for Miles post-training framework
AIRadixArk has released the full technical report for Miles, an open-source RL framework for LLMs and multimodal models built for production post-training. The report covers Miles' system design and how it works in practice, with a focus on stability, efficiency, and flexibility.

Granola@meetgranolaOfficialAI score16Granola lets teams teach it company jargon for better transcription
AIGranola now lets teams add project names, acronyms, and other internal terms so it transcribes them correctly across everyone's meetings. Workspace admins can set this up under Settings → Workspace.

Google DeepMind · YouTubeOfficialAI score38 How AI is transforming weather prediction, featuring WeatherNext 3
AIGoogle DeepMind's Peter Battaglia discusses how machine learning is changing global weather forecasting, including early warnings for storms such as Hurricane Melissa. The episode covers traditional physics-based models versus AI models and probabilistic forecasting, and highlights WeatherNext 3 as Google DeepMind's most advanced global weather AI model yet.
Mistral AIOfficialAI score54 Mistral details how AI agents migrated 40,000 lines of Fortran to C++
AIMistral AI helped a European energy operator migrate 40,000 lines of Fortran 77 to C++ for a reservoir simulator with no test suite. The post explains a parity harness that checks numerical agreement between the two codebases, and a workflow where agents coder, tester, and reviewer migrate modules under human review. Its authors note the approach covered the self-contained first sprint of 40,000 of 300,000 lines and that dependent systems would bring additional challenges.
Ahead of AI (Sebastian Raschka)BlogAI score46 GPT-6 Astra Leads Coding and Math Benchmarks, Shows Strong Computer Use
AIOpenAI's GPT-6 Astra scores 99.9% on ARC-AGI-3, versus 7.8% for GPT-5.6 Sol, and leads Raschka's coding and math tests. Its strongest showing is in graphics and computer-use tasks, such as redrawing an image in a browser-based Paint app. The author notes that Artificial Analysis shows Astra at the frontier but not pulling far ahead on its Coding Agent Index.
Kimi.ai@Kimi_MoonshotOfficialAI score36Kimi Work adds Remote Control for phone access to your computer
AIKimi has launched Remote Control in Kimi Work, letting users leave the app running on their computer and keep tasks moving from their phone. The post frames this as enabling users to work anywhere, anytime.

The Register · AINewsAI score38 Microsoft Edge Team Says AI-Assisted Coding Is Overloading Extension Reviews
AIMicrosoft's Edge team said rapid adoption of AI-assisted coding has increased the volume of browser extension submissions, slowing its review pipeline. The team is adding automation to validation checks and says review standards will not change. It also plans to refresh the "Featured" badge every 15 days.
Manus BlogOfficialAI score30 Sister Builds AI Voice Aid for Brother After Cancer Surgery Removes His Vocal Cords
AIAmy, a 58-year-old with no coding background, built Wally, an AI voice aid app using Manus, for her brother Jamie after cancer surgery at Brigham and Women's Hospital left him unable to speak. The app combines one-tap urgent and everyday phrases with golf news, live tournament scores, and family messages, all placed on his phone's home screen.
Sep 8
Perplexity Developers@perplexitydevsOfficialAI score4Perplexity offers $10 in free API credits to new developers
AIPerplexity Developers says new users who create an account get $10 in free API credits to start building. The post links to a get-started page for its API.
Perplexity Developers@perplexitydevsOfficialAI score30Perplexity Search API now available in Hermes Agent
AIPerplexity says its Search API is now available in Hermes Agent, giving it access to an index of more than 400 billion URLs. The API returns real-time results with snippets ranked by relevance.

Google Developers BlogOfficialPickAI score72 Google releases ADK for Kotlin 1.0 for building production AI agents
AIGoogle announced general availability of ADK for Kotlin 1.0, a Kotlin Multiplatform framework for building AI agents on servers and Android. Version 1.0 reaches feature parity with ADK 1.0 Core and adds Android extensions for on-device models, cloud Gemini via Firebase AI Logic, and persistent sessions and memory with Room and AppSearch. The post includes a server-side incident triage example using KSP-generated tools and skills, plus an Android financial assistant example with human confirmation for transfers.
Why it matters: The post names the new Android and server-side capabilities and the code setup, helping Kotlin developers judge whether ADK fits their agent projects.
Factory NewsOfficialAI score34 Factory Now on Claude Marketplace for Enterprise Autonomous Software Development
AIFactory is now available on the Claude Marketplace, letting enterprise customers apply their committed Anthropic spend toward its autonomous software development platform. The platform automates the software development lifecycle, covering planning, implementation, testing, and security within one system, with enterprise deployment options that keep execution close to customers' code and infrastructure.
Jazzyear · InsightsNewsAI score62 Arm expands from mobile IP into cloud, edge, and physical AI at Shanghai event
AIAt Arm Everywhere China on September 8, 2026, Arm launched products spanning data center CPUs, mobile compute subsystems, and robotics platforms. The article says CSS for Mobile 2 integrates CPU, GPU, and neural accelerator for agent AI on phones, and Arm's Neoverse CSS N4 and AGI CPU target agent sandboxes in data centers. It also reports that Arm's Total Design ecosystem now covers over 80 partners for physical AI.
Gemini Notebook@Gemini_NotebookOfficialAI score22Gemini Notebook rolls out upgraded mobile experience and new study features
AIGemini Notebook's upgraded experience is rolling out across all mobile devices this week, according to the account. Web users can opt in via Settings to be notified when their artifacts are ready, and after a Quiz, Chat can recommend what to study next or generate targeted flashcards for weak spots.
Inferact@inferactOfficialAI score42Inferact reports open models hit 130K tokens/GPU-sec on agentic workloads
AIInferact says months of vLLM tuning for agentic workloads, validated on SemiAnalysis's AgentX benchmark, let open-source models reach up to 130K tokens per GPU-second. The company claims this is 106 times cheaper than Opus 5 API pricing. The work is described as part of a vLLM blog post covering architecture, framework, and runtime optimizations.
Replicate@replicateOfficialAI score46GPT Image 2.5 Launches on Replicate for Precise Image Editing
AIOpenAI's GPT Image 2.5 is now available on Replicate, offering precise image editing, sharper details, and higher consistency across multi-turn edits. Two model pages are provided, openai/gpt-image-2.5-sunburst and openai/gpt-image-2.5-flare.

Aidan Gomez@aidangomezXAI score13Cohere promotes Model Vault for private deployments with no data access
AIAidan Gomez, owner of the Cohere account, promotes Cohere as a lab offering private deployments where it cannot see customer data. He points readers to its Model Vault page for details.
Google AI Studio@GoogleAIStudioOfficialAI score28Google AI Studio Promotes Building Apps With Gemini 3.8 Flash
AIGoogle AI Studio is encouraging developers to start building apps with Gemini 3.8 Flash, linking to a guide on ai.dev. The post offers no benchmarks, pricing, or feature details beyond that invitation.
Granola@meetgranolaOfficialAI score34Granola adds webhooks to its API for note and folder events
AIGranola has added webhooks to its API, letting developers notify their systems or apps when a note is generated or edited. The notifications also cover notes shared with a user or added to a folder.

Cognition Blog (Devin, Windsurf)OfficialAI score58 Cognition raises over $2B at $48B valuation for Devin agent push
AICognition has raised over $2B at a $48B valuation, led by Andreessen Horowitz and Accel. The company says run-rate revenue grew from $492M to almost $900M since its May round, and Devin now offers Auto-Triage, Security Swarm, and Automations.
Werner Vogels@WernerXAI score50Werner Vogels highlights Kiro Crew's memory system drawing on brain evolution
AIWerner Vogels says that after spending time with Kiro Crew since its launch, its memory system stands out for deciding what to keep, compress, and let go. He notes that Amazon engineers, starting from engineering constraints, arrived at an approach resembling the brain's evolved architecture. Per the referenced post, Kiro Crew is a persistent workspace that retains project context across sessions and runs scheduled jobs.
Demis Hassabis@demishassabisXPickAI score73Google DeepMind launches AlphaGenome Atlas to predict impact of human DNA variants
AIDemis Hassabis announced AlphaGenome Atlas, a searchable AI database that maps the predicted impact of all 9 billion possible single-letter DNA changes. The post says it can help scientists better understand disease and is freely available for academic research.
Why it matters: The post describes a searchable database of predicted effects for all 9 billion single-letter DNA variants, which is useful for researchers tracing disease-related genetic changes.
Sundar Pichai@sundarpichaiXAI score60Google DeepMind launches AlphaGenome Atlas for predicting DNA variant effects
AIGoogle DeepMind has launched AlphaGenome Atlas, an AI-powered searchable database mapping the predicted impact of all 9 billion possible single-letter DNA changes. It runs in a regular web browser without coding and is free for academic researchers.
Google DeepMind · YouTubeOfficialPickAI score60 Google DeepMind launches AlphaGenome Atlas for mapping genetic variant effects
AIGoogle DeepMind introduced AlphaGenome Atlas, an AI-powered database charting the molecular impact of every possible genetic variant. Scientists are already using it to investigate unsolved rare diseases and map rare mutations linked to complex traits.
Why it matters: The source names a concrete use case, finding disease-causing DNA variants, which shows how the database could support rare disease research.
Google DeepMindOfficialPickAI score74 Google DeepMind launches AlphaGenome Atlas to predict 9 billion DNA variant effects
AIGoogle DeepMind has introduced AlphaGenome Atlas, a platform with predicted molecular effects for 9 billion single-nucleotide variants in the human genome. It is free for academic research through a web portal, and the AlphaGenome Variant Impact score condenses predictions from AlphaGenome and AlphaMissense into one number for ranking variants. The source says collaborators used it to identify variants in unsolved rare disease cases and to find rare non-coding variants linked to traits.
Why it matters: The source details how precomputed variant predictions, a single impact score, and linked feature attributions make genome-wide mutation effects searchable for researchers without coding skills.
Google DeepMind · The KeywordOfficialPickAI score72 Google DeepMind launches AlphaGenome Atlas, a database of DNA variant effect predictions
AIGoogle DeepMind has released AlphaGenome Atlas, a web portal that predicts the regulatory effects of all 9 billion possible single-letter genetic changes in the human genome. The Atlas provides an AlphaGenome Variant Impact (AVI) score that combines coding and non-coding predictions to help researchers prioritize variants. The source says the portal requires no coding skills and is available to researchers and biologists worldwide.
Why it matters: The source details how the Atlas's AVI score is used in real rare disease and UK Biobank analyses, showing a practical route for prioritizing non-coding variants.
Google DeepMind · YouTubeOfficialPickAI score78 DeepMind releases AlphaGenome Atlas, a predictive map of every possible DNA letter change
AIGoogle DeepMind has used AlphaGenome to predict the molecular impact of every possible single-letter change in the human genome, around nine billion variants. The resulting AlphaGenome Atlas is a 1PB dataset that assigns each variant an AlphaGenome Variant Impact (AVI) score, covering both coding and non-coding variations, and is available to researchers worldwide. The video notes that AlphaGenome has not been validated or approved for any clinical use.
Why it matters: The release supplies a precomputed impact score for every possible single-letter genome change, which lets researchers look up variants without running the model themselves.
Leandro von Werra@lvwerraXAI score22Leandro von Werra builds interactive star map simulator with Astra
AIHugging Face's Leandro von Werra asked Astra to build an interactive star map for his old telescope, and Astra also built a full simulator while a missing cable delays connecting the telescope to the dashboard. The simulator is available as a Hugging Face Space at The post does not specify what Astra is.

Mistral AI@MistralAIOfficialAI score25Mistral says its open-weight models give organisations real control over AI deployment
AIMistral AI says its open-weight models, products, and infrastructure let organisations choose how and where they run AI, not just access a model. The company frames this as frontier performance without vendor lock-in.
Sep 6
Noam Brown@polynoamialXPickAI score67Noam Brown Shares OpenAI Data on Models Accelerating Internal Research
AINoam Brown shares an OpenAI blog post with details on internal research acceleration and says he expects these trends to continue. The post also says OpenAI has paced model development to prioritize monitoring, alignment, and security. A chart shows median daily spend per researcher on internal coding agents rising from near zero in early 2026 to about $600 by August 2026.
Why it matters: The post links an OpenAI blog on internal research acceleration with a chart of rising daily coding agent spend per researcher, useful for judging how fast internal AI use is growing.

Satya Nadella@satyanadellaXAI score22Opal powers new Copilot Autopilot experiences, now in Frontier rings
AISatya Nadella highlights Opal, the technology behind one of Microsoft's new Autopilot experiences, now available in Frontier rings and coming to Copilot soon. The post links to a Microsoft Tech Community blog introducing Project Opal as a new way to complete task-based work.
Satya Nadella@satyanadellaXAI score34Copilot Autopilots now complete long-running multi-step work tasks
AISatya Nadella says Microsoft is bringing new models into Copilot to handle increasingly complex work, from quick questions to delegated tasks and complete long-running jobs via Autopilots. As an example, an Opal-powered Autopilot on a secure Windows 365 Cloud PC sorts a month of trail cam footage, extracts species sightings, and builds a highlight reel, spreadsheet, PowerPoint, and Teams share.

Sebastian Raschka@rasbtXAI score22Raschka's Reasoning From Scratch video covers LLM text generation and KV caching
AISebastian Raschka released a video in his Reasoning From Scratch series covering text generation in LLMs and KV caching. The walkthrough uses a pretrained Qwen3 model from the Reasoning From Scratch package, covering tokenization, greedy decoding, end-of-sequence handling, and a benchmarked KV caching speedup. It prepares the base model for reasoning techniques in later episodes.

OpenBMB (MiniCPM) · new models on Hugging FaceOfficialAI score32 MiniCPM5-2B-DSpark draft model released for speculative decoding with MiniCPM5-2B
AIOpenBMB released MiniCPM5-2B-DSpark, a 323,776,001-parameter DSpark draft checkpoint with five layers that proposes seven draft tokens per forward pass for the MiniCPM5-2B target model. The model, trained on 7,054,154,509 tokens with an average acceptance length of 5.5174 at T=0 and 4.0514 at T=1.0, is served through SGLang with DSPARK speculative decoding. It is released in BF16 under the Apache-2.0 License.
Sep 5
AI at Meta@AIatMetaOfficialAI score46AIRA₃ cuts GPU kernel latency 27% and reaches Kaggle gold level
AIMeta's AIRA₃ system generalizes across domains by changing only the task specification, according to the post. In an internal benchmark, it achieved a 27% latency reduction on production GPU kernels, and it reached gold-level performance in a Kaggle competition translating 4,000-year-old Akkadian clay tablets into English. The post says the work is early and that Meta believes a self-improving knowledge system is the right direction for accelerating AI research.
AI at Meta@AIatMetaOfficialAI score43AIRA₃ coordinates long-running agents through a shared forum and filesystem
AIMeta's AIRA₃ replaces a central controller with many long-running agents, each pairing a model with a coding harness in its own isolated environment. The agents coordinate asynchronously through a shared forum for hypotheses and findings and a shared filesystem for solution artifacts. According to the post, performance gains compound over time as agents build on each other's discoveries.
