Haiku 5.5 to complete the Claude model family in coming weeks
AIAnthropic's Mike Krieger says Haiku 5.5 will round out the model family in the coming weeks. The post gives no details on capabilities, pricing, or exact release timing.
Updated
Updated
Showing low-relevance items too. Hide low-relevance items
AIAnthropic's Mike Krieger says Haiku 5.5 will round out the model family in the coming weeks. The post gives no details on capabilities, pricing, or exact release timing.
AIAnthropic's Cat Wu says Claude Sonnet 5.5 lets Claude Code users complete about 30% more tasks than with Sonnet 5. The model needs fewer tokens for the same work, and in a leaf-raking tool-call demo it finished 24 seconds faster using 6K fewer tokens.
Why it matters: The post gives a measured Claude Code task-completion gain and a token-use example, showing what the model upgrade means for a coding agent workflow.
AIAnthropic's Claude Sonnet 5.5, the second model in the Claude 5.5 family, is shown fixing a bug in Claude Code. Boris Cherny says it runs 30% faster and uses 30% less usage, and Anthropic's announcement says it runs over 30% faster and costs up to 30% less for most work.
AIFelix Rieseberg, who is affiliated with Anthropic, shared a sailing game generated by Sonnet 5.5 at medium effort. The post links to a Claude artifact containing the game and gives no further details on its features or performance.
AIFelix Rieseberg of Anthropic shared a flight simulator built by Sonnet 5.5 at medium effort, linking to a Claude artifact. The post provides no further details on its features or performance.
AIAnthropic launched Sonnet 5.5, which the post says is smarter and more tasteful than Sonnet 5. It is positioned for work that does not need the extra capability of Opus or Fable.

AIwhich links to a page for trying the model. The quoted announcement describes it as the second model in the Claude 5.5 family, a clear upgrade over Sonnet 5 that runs more than 30% faster and costs up to 30% less for most work.
AIAnthropic has released Claude Sonnet 5.5, the second model in the Claude 5.5 family. The quoted announcement says it is a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.
AIReplicate has added P-Video-2-Pro, the latest video model from Pruna AI, which sits on the edge of the preference-speed and preference-price Pareto frontiers. Design Arena ranks its Quality and Speed variants tied for #2 on the Image to Video leaderboard with an Elo of 1325, with the Quality version generating in 8.0 seconds and the Speed version in 4.5 seconds.
AINormal Factory joins the Specialized Intelligence Index with CAD Arena, which tests whether AI agents can turn engineering drawings into accurate, editable CAD parts. The benchmark evaluates agents across five CAD platforms, extending the SII into engineering design.

AIJosh Woodward, Google's Gemini leader, called a Yosemite-themed browser project "awesome" after returning from the park. The project, by @trondw, is a three.js and WebGL photography simulator featuring real sun and Moon positions, real lidar data, and a tripod-mounted film camera for virtual shooting at Yosemite locations.
AIRadixArk has released Miles v0.1.1, adding multi-LoRA with Tinker API compatibility so multiple training jobs can share one base model. The update also supports agentic training with harnesses like Claude Code and runs Harbor tasks in sandboxes including AgentENV, Daytona, E2B, and Modal. It further reduces memory needs for training larger models on validated NVIDIA and AMD GPUs and adds stable support for Qwen3.8-Flash-Next, GLM-5.3-Flash, and Kimi-K3.

AIUnsloth Desktop can now serve local Laya decision models through a Jev-compatible API, shown in a real-time packing demo where suitcase items update as the user types. The demo runs through Unsloth's Decision API, and the team says more optimizations are coming to speed up local hardware performance.
AIGoogle Workspace says users can prompt Gemini directly in Gmail to extract goals, timelines, and next steps from email threads. Gemini then generates a formatted Doc automatically based on the user's current work, while they keep working through their inbox.
AIGoogle says Gemini 3.8 Flash, its most intelligent workhorse model, improves on 3.7 Flash in software engineering, agentic tasks, and multistep reasoning by running extra reasoning steps and calling tools iteratively. The post highlights four community builds, including a model rocket simulation, an animated ink-painting effect, a 3D dinosaur skeleton, and an interactive automatic transmission simulation. Developers can try the model through Google Antigravity and Google AI Studio.
AIHugging Face is contributing egress usage monitoring to NVIDIA's OpenShell, part of the newly launched Open Agent Safety Platform, arguing that allowlists alone restrict where agents can go but not what they do. The proposed features include per-sandbox network budgets for requests, writes, and bytes, drift detection against each sandbox's baseline and cohort, and a fleet view that flags many sandboxes writing to one host even when every request is allowed.
AIKling AI says its Kling 4.0 Flash is live now for Ultra Yearly subscribers, with the full Kling 4.0 coming this October. The update advertises up to 4K resolution, 10-bit HDR output, stereo audio, and native 30-second generation. It also adds Omni Reference supporting up to 15 multimodal references and multi-keyframe control with up to 10 keyframes.

AIUnsloth AI says Laya Decision models can run locally on just 4GB of RAM, on CPU, Mac, Windows, Linux, and GPU setups. The post adds that Laya can be served through a Jev-compatible API via Unsloth Desktop.

AIGoogle's Credentials API for Gemini Managed Agents injects secrets on the wire only for trusted domains, so sandboxed code cannot read raw tokens. It supports environment variables, CLIs, and MCP servers. Passing API keys as plain environment variables lets any sandboxed dependency read and potentially leak them.
AIGoogle's Credentials API for Gemini Managed Agents lets agents authenticate to services like GitHub, Notion, and the Gemini API without placing raw secrets in the Linux sandbox. Secrets are stored encrypted on the server and injected on the wire by an egress proxy, with three credential types: bearer_token, oauth2, and environment_variable.
AISierra has turned its Ghostwriter tool into an always-on teammate in Slack and Teams that proactively suggests ideas, flags problems, and proposes experiments. Ghostwriter reviews recent customer calls, recommends which changes to try first, runs experiments, and reports when results are statistically significant. Sierra said it will begin rolling the feature out more broadly next week.
AIKling AI released THE BEAT, a short film created with Kling 4.0 that follows a drummer's journey back to the stage. The post promotes the piece with the tagline "WATCH ME PLAY" and provides no further technical details.

AILovable announced a partnership with Microsoft that lets users publish apps into their company's Microsoft Entra tenant using Copilot Managed Runtime. Apps can connect to Microsoft 365, Fabric, Dataverse, and SQL data, and staff sign in with their work login. Copilot Managed Runtime is in public preview, and Microsoft 365 connectors, Fabric, and Microsoft sign-in are available on every Lovable plan, while Entra workspace sign-in is included on Business and Enterprise.
AIReplicate's X account shared only an eye emoji, with no details about the announcement. A quoted post from Kling AI says "CLING ON! We've got news," indicating an upcoming Kling announcement without specifying its nature.
AIMeta's AI account says it has released a series of Muse-branded models and products in the past six months, listing Muse Spark, Muse Image, Muse Video, Muse Spark 1.1, Meta Model API, Muse Glimmer, Muse Spark 1.2, Muse Code, Muse Spark 1.3, and Muse. The post says the company is "just getting started" without giving specifications, benchmarks, or pricing.
AIHiggsfield promotes its MCP integration for motion design in After Effects, paired with Claude Opus 5.5. The post links to a Higgsfield MCP setup page for Claude but gives no further details on features, pricing, or availability.
AINVIDIA has launched the Open Agent Safety Platform to help teams control what AI agents can access and do. NVIDIA OpenShell enforces permissions around agent work, while BlueField-4 and DOCA add independent monitoring and security controls in the infrastructure beyond the agent's reach. Together, these components aim to give organizations defined permissions, oversight, and protection for long-running agent tasks.

AIGoogle Workspace announces Polly, a tool that brings quick polls, surveys, and Q&A directly into Google Chat to speed up team decision-making. The post says it lets teams get instant input and stay aligned without leaving the chat, and points readers to a link to get Polly for Google Chat.

AIA community-built 4.0 bpw EXL3 quantization of MiniCPM5-2B reduces the quantized model weights to 1.61 GB for local inference. The author reports roughly 68–70 tokens/s on an NVIDIA Tesla T4, and the model runs with ExLlamaV3 and TabbyAPI.

AIMeta is launching Meta Enterprise Platform to bring its full technology stack to businesses and developers. The initial offering includes the Muse agent, Meta Business Agent, Muse API, and Muse Code, which the company says will help businesses grow.
AIOpenCode has launched Go Plus, a $40 monthly plan with higher usage limits. The post gives no further details on what the limits or features include.
AICreator CHunye produced the Japanese-style zombie short drama DUHAI entirely on his own from Episode 3 onward using SenseTime's Seko AI video creation agent. The series reportedly passed 74 million cumulative views on Douyin and beyond by Episode 9, with Seko handling workflows, characters, scenes, and props on one canvas.
AINVIDIA's Open Agent Safety Platform Reference Design combines NVIDIA OpenShell and NVIDIA Sentry to secure AI agents. OpenShell, an open-source secure runtime, enforces clear boundaries and policy on agent actions while tracing them as they work. NVIDIA Sentry adds hardware-based enforcement on NVIDIA BlueField, continuously monitoring agent activity and enabling millisecond-scale containment and quarantine.

AINVIDIA announced the NVIDIA Open Agent Safety Platform with more than 100 industry partners to set security boundaries for AI agents. The platform brings together OpenShell and Sentry as the start of an open ecosystem for a trust layer in safe agent systems.
AINVIDIA introduced the Open Agent Safety Platform with more than 100 industry partners, bringing together OpenShell and Sentry. The company describes it as the beginning of an open ecosystem to build a trust layer for safe agent systems.

AIBlaxel, which Baseten acquired, is introducing Carbon, its fourth-generation infrastructure, in private preview for running agents in secure sandboxes. Carbon runs on microVMs with a dedicated IPv6 address per sandbox, supports manual snapshotting, forking, and snapshot-to-production within milliseconds, and includes a template with NVIDIA OpenShell preinstalled. Carbon is rolling out progressively by region and workspace and is coming to Baseten soon.
AIModelScope released two Qwen-Image-2.1 LoRAs, LayerExtract and LayerRemove, for layer-based image editing. LayerExtract isolates a prompt-specified subject onto a transparent background, while LayerRemove deletes the matching object from the source image and reconstructs the scene behind it. Both can be hot-swapped within the same DiffSynth-Studio pipeline, and the LoRA weights are licensed under Apache 2.0, with Qwen-Image-2.1 base-model terms also applying.
