Ollama says plugins, web search, and computer use work out of the box
AIOllama states that plugins, web search, and computer use work out of the box. The post gives no further detail on which models, versions, or setup steps are involved.

Updated
Updated
Showing low-relevance items too. Hide low-relevance items
AIOllama states that plugins, web search, and computer use work out of the box. The post gives no further detail on which models, versions, or setup steps are involved.

AIThe ChatGPT Desktop app, referred to as the Codex app, can now be configured to run Ollama models. Users can download or update to Ollama 0.34 to get started.

AITogether AI expanded its Fine-Tuning service with support for newer open-weight models, live metrics tracking, and finer training controls. Expert LoRA adapters can be applied to Mixture-of-Experts expert layers, and early stopping keeps the checkpoint with the best validation loss. Dataset previews, sample weights, pre-flight validation, and lower prices on selected models are also included.
AIOllama announces that the DeepSeek V4.1 Flash model is available on its library, with a link to the model page. The post provides no further details about capabilities, parameters, or pricing.
AIOllama is rolling out DeepSeek-V4.1-Flash on its cloud, starting with Max and Team accounts. The company says it is quickly adding capacity to extend access to all subscribers.
AIGPT-Live-1 is now available in the API, letting developers bring ChatGPT-style back-and-forth conversation into their apps. The quoted announcement says the voice agents can listen while they speak and can work with the models and harness developers choose.
Why it matters: The quoted announcement describes a real-time voice model entering the API, which matters for builders weighing voice agents against existing stacks.
AIOpenAI's Greg Brockman announced GPT-6 Astra Pro for clinicians, and a quoted post says verified U.S. clinicians can access it free today through ChatGPT for Clinicians. Interested clinicians can sign up at
AIThe Gemini app is now available for Windows, letting users open it with the Alt + Space shortcut. From any app, it can polish drafts, summarize long documents, brainstorm ideas, and create custom images and videos.
AIOpenAI announced ChatGPT for Financial Services, a product aimed at giving bankers and financial professionals higher-quality financial data. The post says the company is excited to double down on this focus, and a linked Business Insider article reports on the launch.
AIPerplexity says SPACE is built in Rust and powers the sandboxes behind Perplexity Computer and the Agent API sandbox tool. The post links to a blog post explaining how the company built SPACE.
AIOpenAI has made ChatGPT for Financial Services available, a tailored ChatGPT Work experience that combines built-in financial data with GPT-6 Astra's reasoning. Teams can use it to develop research, build financial models, and create customized client materials. The author says it integrates financial data sources including Daloopa, PitchBook, and LSEG.
AIMicrosoft released Azure AI Speech LLM 2607, which improves multilingual recognition, mixed-language audio handling, and domain-specific entity accuracy, and runs up to 3x faster than the previous 2605 release. A new dedicated phrase list parameter lets developers supply domain vocabulary, supporting 2,000+ entities, without embedding it in a prompt. The model is available through the Fast API and Real-Time API, is testable in the Foundry Playground, and is deployed automatically with no customer action required.
AIGoogle Labs' Dreambeans is now available free to all US users aged 18 and older on iOS and Android, with no subscription required. Users can connect the Gemini app to Dreambeans, which will use their Gemini chat context to surface more personalized daily stories.
AIGoogle Labs has made Dreambeans, its experimental app that creates personalized daily story collections, available to all U.S. accounts aged 18 and over on Android and iOS. Each daily collection combines personalized topics with information distilled from connected Google apps, including Calendar, Gmail, Photos, Search, YouTube, and Gemini. Users can dive deeper into stories, bookmark them, share them, and give feedback to improve future collections.
AIGoogle Antigravity says users can maximize the right-side pane to view artifacts in full screen. The post contains no further details about the feature's availability or behavior.
AIGoogle Antigravity says users can split the terminal view by pressing CMD + \ on macOS or CTRL + \ on other systems. The post offers no further details about the feature.
AIGoogle Antigravity announces a new capability to generate audio files. The post provides no further details on supported formats, models, or limits.
AIGoogle Antigravity says users can press CMD + L on Mac or CTRL + L on Windows to quote copied content into their next prompt. The post provides no further details about the feature's behavior or availability.
AIGoogle Gemini product lead Logan Kilpatrick shares a link to AI Studio's documentation and encourages readers to try it out. The post itself gives no details on what the documentation covers or what changed.
AIOpenAI introduces a Data agent in ChatGPT Work that turns a company's data into answers, interactive dashboards, and actions from a plain question. Users add the Data Plugin, connect their existing data sources and context, and start the conversation. Sherwin Wu says non-data scientists are now running their own quick analyses and building their own dashboards.
AIGoogle AI Studio now offers a fully integrated documentation experience designed for both humans and agents, according to Logan Kilpatrick of Google Gemini. He described it as a first step toward a further reimagined experience and credited the team for the work.
AIGoogle AI Studio is previewing Gemini API documentation directly within its development environment. The docs are designed to consolidate resources so developers can read them where they build, and are available at ai.studio/docs.
AIBlack Forest Labs has released a FLUX video editing tool, with documentation and a public trial page linked in its post. The post provides no further details on capabilities, pricing, or limits.
AIThomas Dohmke says Entire is already approving merges this way, with the capability coming to iPhone Duo in October. The background post notes the new Trails split view was designed with iPhone Duo in mind.
AIReplit's integration with Databricks is now generally available, adding native Databricks Lakebase support that lets Replit Agent automatically provision a Lakebase database when an app is ready to deploy. Apps built with Replit can read live Databricks warehouse data while storing new app data in Lakebase, inheriting existing Unity Catalog security and governance controls. The update also adds automated preview deploys that keep test data isolated from live business data.
AIRadixArk has published a cookbook on its Miles documentation site covering how to run DeepSeek V4.1 Flash. The post itself contains only a link to the cookbook page, so no further details about features, figures, or setup steps are available.
AIRadixArk says Miles brings day-0 RL support to DeepSeek-V4.1-Flash, with SGLang providing inference support. The post says quantization-aware training mirrors SGLang's FP4/FP8 rounding, and that colocated training and rollout fit full-parameter RL on 16 GPUs. In a DAPO run over steps 0–80, per-token trainer–rollout KL stayed at 0.0012–0.0017 while reward rose from 0.51 to 0.78.
AISGLang and Miles ship day-0 inference and RL support for DeepSeek V4.1 Flash, with weights now available. The model is natively multimodal with 552B backbone parameters, 16B active during decode and 8B during prefill, and supports up to 1M context. V4.1 adds shared compressed KV across layers, a two-stage sparse indexer, and a 196B Engram lookup memory.
AIDeepSeek says its more efficient V4.1-Flash architecture lets it serve more users at lower cost and pass the savings on through lower API prices. Peak/off-peak pricing continues, with off-peak rates at 50% of peak rates. The new pricing takes effect at 04:00 UTC on September 10, 2026.

AIDeepSeek says V4.1-Flash is now live on its API with native multimodal support, accessed through the model name deepseek-flash. The older V4-Flash and V4-Flash-Vision-Exp are retired, while deepseek-v4-flash and deepseek-v4-flash-vision-exp temporarily route to V4.1-Flash. Requests to deepseek-v4-pro will route to V4.1-Flash at V4.1-Flash rates starting 04:00 UTC on Sept 14, 2026, until V4.1-Pro launches.
AIMicrosoft has released Dynamics 365 Activate as a public preview, an AI-powered tool that converts Salesforce implementations to Dynamics 365. The company said it will add more CRM and ERP migration scenarios later this year. The tool profiles data, entities, relationships, and customizations to give implementation teams a migration blueprint.
AIBAAI released open-source resources for three AIDD projects on Hugging Face: EPT, an equivariant pretrained transformer for unified 3D molecular representation learning, and UniPath, a learnable-time flow matching method for crystal structure and energy prediction. The repository mirrors their GitHub source code and READMEs, with setup, preprocessing, training, and evaluation documentation. The MiSI benchmark is released separately on Hugging Face.
AICursor is launching Projects, a beta feature for larger work such as a feature, migration, or full app, rolling out to all users starting today. A coordinator agent plans the work, delegates it to implementing agents that can run in parallel, and runs on a cloud computer so it continues when the laptop is closed. Each Project keeps shared context files synced across cloud and local machines, and subscriptions let the coordinator act on Slack channels, schedules, or PRs without a prompt.
Why it matters: The source details how a coordinator agent plans, delegates, and syncs shared context across cloud and local machines, useful for judging how long-running agent work might fit a team's workflow.
AITogether AI has launched a public preview of preemptible compute for Together GPU Clusters on Kubernetes in all regions, billed sub-hourly at a flat 50% of the on-demand rate. Preemptible nodes can be reclaimed when capacity is needed elsewhere, with a five-minute drain window for checkpointing before removal. The rate stays fixed rather than tracking a spot market, and the cluster automatically refills its preemptible target as capacity becomes available.
AIAnthropic has launched Smart reports in beta for Claude Enterprise plans. The reports analyze how a team uses Claude, covering work completed, costs, session friction, and repeated patterns worth packaging as shared skills.
AIThe main post only says "Astra for medical education:" with no details, so the source offers little beyond that label. The quoted post claims a GPT-6 Astra service links anatomical diagrams, CT cross-sections, and 3D relationships in one spatial tool for medical students.
AIPerplexity says its MCP tools perplexity_ask, perplexity_research, and perplexity_reason run on Agent API. The company describes Agent API as a single, multi-provider endpoint with built-in tools and dynamic presets.
AIPerplexity's API MCP server now supports OAuth, letting users connect by adding to their client, signing in with their Perplexity account, selecting an org, and approving access. Once connected, agents gain real-time web search, deep research, and advanced reasoning.
AIMicrosoft Foundry's July and August 2026 updates make Hosted Agents, Voice Live integration, and Toolboxes generally available. The post adds Claude tools on Azure, Model Router region and model pool changes, Foundry Local preview features, and updated Python, JavaScript, Java, and .NET SDK versions with migration notes.
Why it matters: The roundup links each GA and preview change to code examples, migration notes, and runtime requirements, which helps developers judge what to upgrade and test first.
AISatya Nadella says NFL analysts and coaches, including Seahawks analyst Brian Eayrs, are using new Copilot and Excel tools to support decision-making in the broadcast booths and on the sidelines. The post frames this as the NFL season returning, with no specific features, metrics, or availability details provided.
