Skip to contentSkip to stories

Updated

All AI news

Items with an AI score under 20 are hidden. Show low-relevance items

Sep 28

Sep 28Mon
  1. Grok BotOfficialAI score45

    SpaceXAI launches Team Bots public beta for Teams and Enterprise

    AISpaceXAI says its Team Bots, which prep account teams, coordinate engineering work, answer data questions, triage customer feedback, and run hiring loops, are now in public beta for Teams and Enterprise customers. The post links to a full announcement at

  2. SemiAnalysisBlogAI score43

    How GLM-5.3 Sparse Attention Affects HBM and Serving Costs on GB200, GB300, and MI355X

    AISparse attention cuts per-operation KV cache reads but does not reduce overall memory capacity, so top-k cache misses still depend on HBM. SemiAnalysis's InferenceX estimates GB200 at about $0.044 per million total tokens at 150 tokens per second, roughly 12% below MI355X running ATOM at $0.049. Neither system holds a uniform cost advantage across the tested 100, 125, and 150 tokens-per-second targets.

  3. Ali GhodsiXAI score62

    Databricks finds Opus 5.5 cheaper and better, GPT-6 Luna 20x cheaper per task

    AIDatabricks tested recent AI models across 2,400 engineers and found Opus 5.5 offers the highest quality mid-tier performance, with about 20% lower same-task costs than Opus 4.8. The company is now encouraging Opus 5.5 as a default model for coding, and reports that GPT-6 Luna is at least 20 times cheaper per task than Opus 5.5, roughly matching Opus 4.6 on one difficult evaluation suite. The Luna findings are preliminary.

  4. Google · AI blogOfficialAI score26

    Future Vision XPRIZE grand prize goes to The Gifted, a film about a boy and an AI voice

    AIIndependent filmmaker Jeff Synthesized won the Future Vision XPRIZE grand prize for The Gifted, a film about an 11-year-old who recreates his late mother's voice and essence using code. The solo-developed project was chosen from more than 2,500 worldwide entries and receives $100,000 plus $2.5 million in feature production funding. Google is partnering with Range Media Partners through its 100 ZEROS initiative to bring the film to the big screen.

  5. Perplexity DevelopersOfficialAI score44

    Perplexity adds reusable custom agents to its Agent API

    AIPerplexity says developers can now build custom reusable agents in its Agent API using Profiles, Skills, and managed connectors. Agents are configured once in the API Portal and can then be reused across applications and workflows.

    Video from @perplexitydevs's post
  6. Google AIOfficialAI score44

    Google Labs expands experimental CC agent into a family group assistant

    AIGoogle Labs has expanded Project CC, its experimental AI productivity assistant, into a group agent designed to streamline family household logistics. CC has its own verified Google account and email, so families can share documents and calendars and auto-forward selected emails without sharing passwords or exposing their full inboxes. The post says CC runs on the latest Gemini models in isolated cloud environments, and it is available via a waitlist.

  7. Mike KriegerXAI score67

    Anthropic releases Claude Sonnet 5.5, 30% faster and up to 30% cheaper than Sonnet 5

    AIAnthropic has released Claude Sonnet 5.5, the second model in the Claude 5.5 family. The company says it is more than 30% faster than Sonnet 5 and costs up to 30% less for most work.

    Why it matters: The post gives concrete speed and price changes against Sonnet 5, which helps readers judge whether the upgrade fits their workloads and budgets.

  8. catXAI score72

    Claude Sonnet 5.5 Lifts Claude Code Task Completion by About 30%

    AIAnthropic's Cat Wu says Claude Sonnet 5.5 lets Claude Code users complete about 30% more tasks than with Sonnet 5. The model needs fewer tokens for the same work, and in a leaf-raking tool-call demo it finished 24 seconds faster using 6K fewer tokens.

    Why it matters: The post gives a measured Claude Code task-completion gain and a token-use example, showing what the model upgrade means for a coding agent workflow.

    Video from @_catwu's post
  9. Boris ChernyXAI score38

    Claude Sonnet 5.5 runs 30% faster with 30% less usage

    AIAnthropic's Claude Sonnet 5.5, the second model in the Claude 5.5 family, is shown fixing a bug in Claude Code. Boris Cherny says it runs 30% faster and uses 30% less usage, and Anthropic's announcement says it runs over 30% faster and costs up to 30% less for most work.

    Video from @bcherny's post
  10. Alex AlbertXAI score62

    Claude Sonnet 5.5 Is Faster and Cheaper Than Sonnet 5, Per Anthropic

    AIAnthropic has introduced Claude Sonnet 5.5, the second model in the Claude 5.5 family, as a clear upgrade over Sonnet 5. The announcement says it runs more than 30% faster and costs up to 30% less for most work. Alex Albert, quoting the announcement, says the model writes clearly, is very fast, and makes a major capabilities jump over Sonnet 5.

    Why it matters: The quoted announcement gives concrete speed and cost figures for Sonnet 5.5, which helps readers weigh it against Sonnet 5 for everyday work.

  11. v0OfficialAI score62

    Claude Sonnet 5.5 is now available in v0

    AIwhich links to a page for trying the model. The quoted announcement describes it as the second model in the Claude 5.5 family, a clear upgrade over Sonnet 5 that runs more than 30% faster and costs up to 30% less for most work.

    Why it matters: The post shows the model's availability inside v0 and links a test entry point, though it gives no performance details beyond the quoted claims.

  12. AnthropicOfficialAI score61

    Claude Sonnet 5.5 is now available from Anthropic

    AIAnthropic has released Claude Sonnet 5.5, the second model in the Claude 5.5 family. The quoted announcement says it is a clear upgrade over Sonnet 5, runs more than 30% faster, and costs up to 30% less for most work.

    Why it matters: The quoted post gives concrete speed and cost gains over Sonnet 5 for most work, which helps readers weigh upgrade decisions against their current model.

  13. Google WorkspaceOfficialAI score34

    Gemini in Gmail can turn email threads into structured briefs

    AIGoogle Workspace says users can prompt Gemini directly in Gmail to extract goals, timelines, and next steps from email threads. Gemini then generates a formatted Doc automatically based on the user's current work, while they keep working through their inbox.

    Video from @GoogleWorkspace's post
  14. Google Cloud · AI & Machine LearningOfficialAI score40

    Why startups should pair open models like Gemma 4 with frontier APIs

    AIGoogle Cloud argues startups should combine open-weight models with frontier APIs rather than routing every request to one frontier model. It cites Gemma 4, which spans five sizes including a 31B dense model and a 26B A4B Mixture-of-Experts model, released under Apache 2.0. The article's examples report a 44% latency drop for Cue, from 876 ms to 488 ms, and a $0 server cost for BetterSpeak's on-device Gemma 4 E2B.

  15. Google · Gemini appOfficialAI score38

    See what 4 builders are making with Gemini 3.8 Flash

    AIGoogle says Gemini 3.8 Flash, its most intelligent workhorse model, improves on 3.7 Flash in software engineering, agentic tasks, and multistep reasoning by running extra reasoning steps and calling tools iteratively. The post highlights four community builds, including a model rocket simulation, an animated ink-painting effect, a 3D dinosaur skeleton, and an interactive automatic transmission simulation. Developers can try the model through Google Antigravity and Google AI Studio.

  16. Mistral AIOfficialAI score40

    Mistral Opens Munich Hub to Advance Physics and Industrial AI in Germany

    AIMistral AI has opened a new hub in Munich, housing research teams focused on Physics AI and Industrial AI alongside applied engineers serving enterprise partners. The company plans to build one gigawatt of European compute capacity by 2030 and says its open-weight models can run on customers' own infrastructure. Mistral is working with BMW on crash simulations and with Siemens Energy on industrial AI applications.

  17. clem 🤗XAI score49

    Hugging Face proposes egress usage monitoring for OpenShell agent sandboxes

    AIHugging Face is contributing egress usage monitoring to NVIDIA's OpenShell, part of the newly launched Open Agent Safety Platform, arguing that allowlists alone restrict where agents can go but not what they do. The proposed features include per-sandbox network budgets for requests, writes, and bytes, drift detection against each sandbox's baseline and cohort, and a fleet view that flags many sandboxes writing to one host even when every request is allowed.

    Video from @ClementDelangue's post
  18. Kling AIOfficialAI score42

    Kling 4.0 Flash launches now for Ultra Yearly subscribers; Kling 4.0 arrives October

    AIKling AI says its Kling 4.0 Flash is live now for Ultra Yearly subscribers, with the full Kling 4.0 coming this October. The update advertises up to 4K resolution, 10-bit HDR output, stereo audio, and native 30-second generation. It also adds Omni Reference supporting up to 15 multimodal references and multi-keyframe control with up to 10 keyframes.

    Video from @Kling_ai's post
  19. Philipp SchmidXAI score36

    Gemini Managed Agents' Credentials API keeps secrets out of sandboxed code

    AIGoogle's Credentials API for Gemini Managed Agents injects secrets on the wire only for trusted domains, so sandboxed code cannot read raw tokens. It supports environment variables, CLIs, and MCP servers. Passing API keys as plain environment variables lets any sandboxed dependency read and potentially leak them.

  20. Philipp SchmidXAI score52

    Gemini Managed Agents adds a Credentials API that keeps secrets out of sandboxes

    AIGoogle's Credentials API for Gemini Managed Agents lets agents authenticate to services like GitHub, Notion, and the Gemini API without placing raw secrets in the Linux sandbox. Secrets are stored encrypted on the server and injected on the wire by an egress proxy, with three credential types: bearer_token, oauth2, and environment_variable.

  21. Sierra BlogOfficialAI score34

    Sierra's Ghostwriter becomes a proactive Slack and Teams teammate for AI agents

    AISierra has turned its Ghostwriter tool into an always-on teammate in Slack and Teams that proactively suggests ideas, flags problems, and proposes experiments. Ghostwriter reviews recent customer calls, recommends which changes to try first, runs experiments, and reports when results are statistically significant. Sierra said it will begin rolling the feature out more broadly next week.

  22. Azure BlogOfficialAI score34

    Azure Introduces VM Lifecycle Policy with Current, Extended, End of Life, and Retired Stages

    AIMicrosoft has introduced an Azure Virtual Machine Lifecycle Policy defining four stages, Current, Extended, End of Life, and Retired, for managing VM transitions. Azure General Purpose, Memory Optimized, Compute Optimized, and Storage Optimized VMs are the first families covered, with Current VMs recommended for new deployments. Retired VMs can no longer be provisioned, are not covered by an SLA, and are no longer eligible for Microsoft support.

  23. Lovable BlogOfficialAI score57

    Lovable apps can now run inside a company's Microsoft tenant

    AILovable announced a partnership with Microsoft that lets users publish apps into their company's Microsoft Entra tenant using Copilot Managed Runtime. Apps can connect to Microsoft 365, Fabric, Dataverse, and SQL data, and staff sign in with their work login. Copilot Managed Runtime is in public preview, and Microsoft 365 connectors, Fabric, and Microsoft sign-in are available on every Lovable plan, while Entra workspace sign-in is included on Business and Enterprise.

  24. Higgsfield AI 🧩OfficialAI score34

    Claude Opus 5.5 Drives 12 Laptops to Produce a Launch Video

    AIHiggsfield AI gave Claude Opus 5.5 access to 12 laptops, and from one prompt it split the work across machines using Computer Use and Higgsfield MCP. The system generated the visuals, built the animations, and assembled a fully editable After Effects project.

    Video from @higgsfield's post
  25. NVIDIAOfficialAI score34

    NVIDIA launches Open Agent Safety Platform to control AI agent access

    AINVIDIA has launched the Open Agent Safety Platform to help teams control what AI agents can access and do. NVIDIA OpenShell enforces permissions around agent work, while BlueField-4 and DOCA add independent monitoring and security controls in the infrastructure beyond the agent's reach. Together, these components aim to give organizations defined permissions, oversight, and protection for long-running agent tasks.

    Image from @nvidia's post
  26. Together AIOfficialAI score34

    Together AI Launches as Partner for NVIDIA Open Agent Safety Platform

    AITogether AI is a launch partner for NVIDIA's Open Agent Safety Platform, which brings together OpenShell and Sentry with over 100 industry partners. Together AI says it has built platform capabilities for secure agent development and deployment and will keep investing in this area, including its work with NVIDIA on OpenShell.

  27. Mark ZuckerbergXAI score27

    Meta hires Chirantan "CJ" Desai as Chief Enterprise Platform Officer

    AIMark Zuckerberg announced that Chirantan "CJ" Desai will join Meta as Chief Enterprise Platform Officer, reporting directly to him. Desai previously served as CEO and President of MongoDB, led product and engineering at Cloudflare, and spent nearly eight years at ServiceNow, including as President and COO. The post describes his experience across AI, infrastructure, business applications, and security as the basis for leading Meta's enterprise effort.