Skip to contentSkip to stories

Updated

All AI news

Showing low-relevance items too. Hide low-relevance items

Sep 24

Sep 24Thu
  1. Baseten BlogOfficialAI score44

    LangSmith Fine-Tuning Trains Open Models on Agent Traces via Baseten Loops

    AILangChain launched LangSmith Fine-Tuning, which lets users fine-tune open models on their LangSmith agent traces using the open-source smithtune CLI. Training runs on Baseten Loops in the user's own workspace, and smithtune deploy places the evaluated checkpoint on a Baseten Dedicated Inference deployment. Loops is in early access, so users may need to request access for their workspace.

  2. GitHub Blog · AI & MLOfficialAI score66

    GitHub Security Lab shows an LLM agent running AI-driven fuzzing for C/C++ projects

    AIGitHub Security Lab describes the Fuzzing Taskflow, an LLM agent pipeline that identifies entrypoints, writes harnesses, runs AFL++, reads coverage reports, and triages crashes for C/C++ repositories. The agent makes decisions while MCP tools handle execution, and state is stored in a SQLite database. The post also warns that the taskflow runs AFL and build commands directly on the host, so it should be used only in disposable environments without elevated privileges.

    Why it matters: The post explains how an LLM agent automates fuzzing steps like harness writing, coverage gap chasing, and crash triage, with a runnable workflow and design tradeoffs.

  3. Azure BlogOfficialAI score67

    Microsoft Foundry adds voice agents and continuous optimization for production agents

    AIMicrosoft Foundry expands its agent platform with voice agents in public preview, long-running resilience for hosted agents, and tools for evaluating production agents. The post also says GPT-6 Sol, GPT-6 Luna, and Claude Opus 5.5 are now available in Foundry. Agent optimizer, Insights, and Rubric evaluator are described as tools for continuous improvement, with some reaching general availability later this month.

    Why it matters: The post shows how Foundry combines model choice, voice agents, long-running resilience, and production evaluation into one agent workflow, with a customer example.

  4. Google DeepMindOfficialAI score62

    Google DeepMind adds Live Avatar to Gemini 3.8 Live for enterprise

    AIGoogle DeepMind has launched Gemini 3.8 Live with Live Avatar, which adds near real-time visual presence to its native live dialogue models. The feature is available today in Gemini Enterprise, supports 97 languages with adaptive lip-sync, and allows custom avatars through enterprise allowlisting. All output carries an imperceptible SynthID watermark.

    Why it matters: The post specifies the new avatar capabilities, the Gemini Enterprise access path, and the SynthID watermark, which helps readers judge its enterprise deployment fit.

  5. Google for DevelopersOfficialAI score37

    Gemma 4 now runs on-device in the Antigravity SDK

    AIGoogle says Gemma 4 can now run locally on-device within the Antigravity SDK. Developers can build fully local or hybrid multi-agent workflows that pair cloud models with Gemma 4 agents for auditing, patching, and testing code. The post emphasizes total data privacy and zero API fees, powered by LiteRT.

    Video from @googledevs's post
  6. OdysseyOfficialAI score18

    Odyssey introduces Agora-2, a multi-agent world model

    AIOdyssey launched Agora-2, a multi-agent world model that the company is making available for public experimentation. The post predicts such models will increasingly power applications in AI training, AI safety, robotics, autonomous vehicles, defense, energy, cybersecurity, and gaming.

  7. Philipp SchmidXAI score56

    Gemini 3.8 TTS adds custom voice creation from a short recording or prompt

    AIGemini 3.8 TTS lets users replicate their own voice or design a custom voice from a text prompt. The workflow is to record about 20 seconds of speech with a consent sentence, create the voice through an API call, then use it in any request with styles set in speech_metadata. The author also points readers to a guide for setting up and testing the process with an agent.

  8. Philipp SchmidXAI score62

    Gemini 3.8 Flash TTS adds custom voice creation from recordings or a sentence

    AIGemini 3.8 Flash TTS and Flash-Lite TTS are now available in the Gemini API and AI Studio, with a new option to replicate a user's own voice from two recordings or design one from a sentence. The guide says the reusable voice ID can be passed in later requests, or an encrypted voicekey that expires after 7 days can be used if nothing is stored server-side. Prompting changed from gemini-3.1-flash-tts-preview: input text is spoken word for word, delivery goes in speech_metadata.style, and non-streaming responses are now real WAV.

  9. Google · Gemini appOfficialAI score62

    Google launches Gemini 3.8 Live with Live Avatar for enterprises

    AIGoogle introduced Gemini 3.8 Live with Live Avatar, which adds a visual persona with lip-syncing and expressions to its live dialogue models. The feature is available in Gemini Enterprise and supports 97 languages, with custom avatars available through enterprise allowlisting. Google says all output is watermarked with SynthID.

    Why it matters: The post specifies enterprise availability, custom avatar allowlisting, and 97-language support, which clarifies who can use the feature and how far it reaches.

  10. vLLMOfficialAI score34

    vLLM and RL-Kernel achieve bit-exact logprob match on AMD MI300X

    AIThe RLKernel team integrated RL-Align/RL-Kernel with vllm-project/vime, and a 200-step Qwen3-8B GRPO run on 8× AMD MI300X recorded zero logprob mismatches between Megatron training and vLLM rollout. The strict path aligns reduction order, intermediate precision, rounding points, and math primitives across both sides to achieve bit-for-bit matching on ROCm.

  11. Microsoft Foundry BlogOfficialAI score61

    Microsoft Foundry Routines reach general availability for scheduled and event-driven agents

    AIMicrosoft announced general availability of Routines in Foundry Agent Service, a managed way to run agents on a timer, on a recurring schedule, or in response to GitHub issue events and new Microsoft Teams channel messages. Routines keep the trigger, agent action, identity, connections, and run history in the Foundry project, and each routine can run under the creator's identity or the agent's own Microsoft Entra ID identity. A preview reminder tool lets a Hosted Agent schedule itself to resume later on the same conversation.

    Why it matters: The post explains how scheduled, event-based, and self-reminding agent runs are managed in one place, along with the creator versus agent identity choice for unattended tasks.

  12. Google Cloud · AI & Machine LearningOfficialAI score55

    Gemini 3.8 Live with Live Avatar becomes generally available in Gemini Enterprise

    AIGoogle says Gemini 3.8 Live with Live Avatar is now generally available in Gemini Enterprise, with US and EU endpoints, provisioned throughput, and enterprise compliance. Its video avatars use synchronized lip-syncing, custom avatars are limited to an allowlist, and generated audio and video carry SynthID watermarks. The model also understands and speaks 97 languages and can run tool calls in the background while the conversation continues.

  13. Liquid AI NewsletterOfficialAI score38

    Liquid AI's Liquid Context now optimized for Snapdragon NPUs; LFM Longevity models released

    AILiquid AI announced its on-device Liquid Context layer is now optimized for Snapdragon processors using the Qualcomm Hexagon NPU, letting edge agents learn user routines and share context across devices. Separately, Liquid AI released LFM2-1.2B-Longevity and LFM2-2.6B-Longevity, which the company says often match or outperform much larger frontier LLMs on longevity prediction tasks.

  14. OpenBMBOfficialAI score34

    FIT-GGUF enables size-targeted mixed-precision quantization of MiniCPM5-2B

    AIDeveloper @Scorp1o_117 used FIT-GGUF to build four MiniCPM5-2B GGUF variants, ranging from about 1.14 GiB to 1.46 GiB, tuned to target file sizes or fidelity tiers. Instead of fixed presets, FIT-GGUF allocates precision tensor by tensor, with Quality, Balanced, Compact, and Mini options, and its generated files matched predicted sizes. Builds are evaluated with KL Divergence and Same-top metrics and are available on Hugging Face.

    Image from @OpenBMB's post
  15. Philipp SchmidXAI score22

    Gemini 3.8 Flash launched for multimodal understanding tasks

    AIGoogle's Philipp Schmid announced Gemini 3.8 Flash, recommending Gemini for multimodal understanding. A quoted post by Spencer Schiff reported that frontier models struggled to match correct names to people in a drawing, offering a visual test for future models.

    Image from @_philschmid's post
  16. ModelScopeOfficialAI score38

    Qwen-Image-2.1-Fun-Controlnet-Union adds eight controls and inpainting

    AIModelScope released Qwen-Image-2.1-Fun-Controlnet-Union, a single checkpoint adding eight structural controls, including Canny, Depth, Pose, and Scribble, plus inpainting to Qwen-Image 2.1. Control and inpainting share one branch with 16 injection points across every second Transformer block, keeping the base model frozen and requiring no checkpoint switching. It runs at guidance scale 1.0 with CFG-distilled sampling and prefix KV caching, and is available under the Qwen Research License with base Qwen-Image 2.1 weights required.

    Image from @ModelScope2022's post
  17. KrASIA · Big TechNewsAI score55

    Mind Lab launches Mint Recursive, a post-training platform for companies

    AIMind Lab unveiled Mint Recursive, a post-training and inference platform for industry use, alongside Macaron-V1.1, a model post-trained entirely on it. Macaron-V1.1 is a 752-billion-parameter model built from GLM-5.3 with four two-billion-parameter LoRA expert modules for chat, agents, coding, and generation. The platform is serverless and bills by token usage, and it collects feedback from models in use to support continued training.

  18. Lovable BlogOfficialAI score44

    Lovable Now Offers Free Chat for Planning and App Work

    AILovable now lets users chat for free to explore app ideas, review existing projects, and draft business materials before making changes. The chat can connect to tools like Notion, Granola, and Linear, and Free, Pro, and Business workspaces include a daily free chat allowance. Chats that generate images or video, or hand work off to Plan or Build, use credits as usual, and current chat pricing applies through October 31, 2026.

  19. MiniMax (official)OfficialAI score34

    MiniMax-H3 video generation accelerated on AMD MI355X by Nunchux

    AINunchux runs MiniMax-H3 on AMD MI355X GPUs, generating 5 seconds of video in 1.3 seconds with up to 26.7x faster inference than SGLang on 8 GPUs. The stack supports streaming generation, letting users change prompts while the video plays. Free access to MiniMax-H3 through Nunchux is coming soon, with a waitlist open.

  20. LangChain BlogOfficialAI score50

    LangSmith Engine v2 adds red teaming and pre-validated agent fixes

    AILangChain released LangSmith Engine v2, an in-platform agent that scans production traces to detect agent issues and validates proposed fixes before human review. Engine v2 adds Red Teaming, currently in Private Beta for LangSmith Deployment users, which tests agents for weaknesses such as hallucinations and system-prompt violations before they reach production. Engine v2 is available in SaaS deployments for LangSmith Plus and Enterprise plans, with Self-Hosted support and BYOK for Engine coming later.

  21. LangChain BlogOfficialAI score50

    LangSmith Fine-Tuning and smithtune Turn Agent Trajectories Into Custom Models

    AILangChain launched LangSmith Fine-Tuning and smithtune, a CLI that turns LangSmith agent trajectories into fine-tuned models through dataset creation, training with Fireworks or Baseten, and evaluation in LangSmith. smithtune currently supports supervised fine-tuning, training models on recorded examples of good agent behavior by updating model weights. The tool lets teams train specialized models without building the data pipeline by hand.

  22. LangChain BlogOfficialAI score44

    LangSmith Launches Trajectories for Readable, Chronological Agent Session Views

    AILangChain has launched Trajectories in LangSmith, a chronological, conversational view that aggregates human, AI, and tool messages across an agent and its subagents. Trajectories work with traces from LangChain, LangGraph, Deep Agents, OpenAI and Claude agent SDKs, and coding agents like Codex, Claude Code, and Cursor. The feature is available now on all plans in the US.

Sep 23

Sep 23Wed
  1. OpenClaw🦞OfficialAI score12

    OpenClaw adds live meeting notes and saved transcript tabs

    AIOpenClaw lets users follow notes while a Google Meet, Teams, Zoom, or voice capture is still running. The saved transcript can be opened in its own tab, though transcription may lag. Generated notes use the user's configured model and incur its usual charges.

  2. WanOfficialAI score15

    Wan3.0 generates cinematic 1080p video with sound from one prompt

    AIAlibaba Cloud and Venice AI demonstrated Wan3.0 producing a finished scene from a single text prompt, with cinematic motion, 1080p output, and generated sound. The post presents the end-to-end workflow as a sign that the gap between an idea and a finished video is shrinking.

    Video from @Alibaba_Wan's post
  3. Midjourney UpdatesOfficialAI score31

    Midjourney Alpha changelog adds style previews, default parameters, and a new Create feed

    AIMidjourney's alpha site now lets users preview their current prompt across styles with "Live previews" in the Styles sidebar and save prompt-bar settings as defaults via Settings → Advanced → Your defaults. The Create feed received a full-width masonry redesign with hover-based prompts and buttons, and Korean is now live for all users on midjourney.com.