Skip to contentSkip to stories

Updated

#Google

Showing low-relevance items too. Hide low-relevance items

Oct 8

Oct 8Thu
  1. Sundar PichaiAI score62

    Google introduces Gemini agent as a single universal agent for work

    AIGoogle introduced a new Gemini agent that combines question answering, knowledge work, image and media creation, and code writing in one prompt box. The agent connects to personal workflows, systems of record, and enterprise controls, and runs in the cloud with a shared memory and personalization graph. It can create sub-agents for multi-step tasks, act as a coworker agent with its own identity, and orchestrate across multiple models to balance quality and cost.

    Image from @sundarpichai's post
  2. Google ResearchAI score14

    Google Research demos EnvHarness for co-evolving LLM agents and environments at COLM 2026

    AIGoogle Research is presenting EnvHarness, a flexible framework that enables co-evolution between LLM agents and their training environments, at the #COLM2026 Google booth #107 today at 11:00 AM PT. The post notes that static environments limit agent growth, and EnvHarness is described as a plug-in architecture that dynamically reshapes environment behaviors to improve reinforcement learning and adaptability.

  3. Google GemmaAI score44

    Google publishes a developer guide for EmbeddingGemma 2 multimodal embeddings

    AIGoogle Gemma announces a developer guide showing how to embed text, code, images, video, audio, and interleaved inputs with EmbeddingGemma 2 using the sentence-transformers library. The guide outlines a four-step workflow: loading the model, embedding text and code with task prompts, embedding multimodal inputs, and optionally truncating dimensions with Matryoshka.

    Image from @googlegemma's post
  4. The Verge · AIAI score52

    Google's experimental AI Edge Foresight transcribes meetings fully offline on Mac

    AIGoogle has released AI Edge Foresight, a free experimental note-taking app that transcribes meetings and audio files entirely offline on macOS. It runs on the on-device EmbeddingGemma 2 model and turns shorthand notes into polished notes based on the transcript. Google says files, meeting audio, and notes never leave the computer, and the app is currently optimized only for Macs with Apple Silicon.

  5. The Next PlatformAI score43

    How Distributed AI Training Changes the Network Between Datacenters

    AILarge-scale AI training is spreading across multiple datacenters, with Google, Microsoft, AWS, Meta, and CoreWeave cited as examples. Because synchronized GPU clusters must exchange data in bursts, inter-site links can become a bottleneck, which Cisco estimates may require aggregate bandwidth about 14x a conventional DCI baseline.

  6. Philipp SchmidAI score46

    SynthID Detector now publicly available for verifying AI-generated content

    AIGoogle's SynthID Detector is now publicly available, letting users check whether an image, video, or audio file was generated by supported tools. Per the post, it scans for watermarks from Google and partners, including Nano Banana 2.1, OpenAI, NVIDIA, and Kakao, with Apple support coming soon. Uploaded files are deleted right after scanning.

    Video from @_philschmid's post
  7. The Verge · AIAI score58

    Google launches a universal Gemini agent for enterprise work tasks

    AIGoogle is launching a "universal" Gemini agent that works across apps and devices in the background, available in private preview to enterprise customers. Users can chat with it and assign tasks from the Gemini Enterprise app, and it works inside Gmail, Drive, Docs, Sheets, and Calendar as well as third-party apps like Slack and Microsoft 365. The source notes it runs in the cloud, keeps the same context across devices, and can use job-specific sub-agents.

  8. 🚨 AI News | TestingCatalogAI score46

    Google announces a unified Gemini agent for Gemini Enterprise work

    AIGoogle has announced a single, universal Gemini agent for Gemini Enterprise as part of its Gemini at Work updates. The agent answers questions, handles knowledge work, creates images and media, and writes and runs code. It works inline in Gmail, Drive, Docs, Slides, Sheets, Chat, and Calendar, with new data and analytics skills for plain-language insights and industry-specific tools for financial services and legal teams.

    Image from @testingcatalog's post
  9. SiliconANGLE · AIAI score62

    Google Cloud launches Gemini agent for enterprise work across devices and apps

    AIGoogle Cloud introduced Gemini agent, a unified AI assistant that acts autonomously, generates code, and completes work across web, mobile, desktop, and third-party apps. It runs jobs on models matched to each task, including Gemini Flash and a flagship frontier model, with Anthropic Claude models also available. Hard spend limits per project let companies enforce budgets and charge AI costs to departments.

  10. ZDNet · AIAI score24

    Google Maps adds Ask Maps food ordering via Square and Uber Eats, plus fan-favorite dining list

    AIGoogle Maps now lets users place restaurant orders through its Gemini-powered Ask Maps chat tool, with Square and Uber Eats joining Toast as partners. Google also released its first fan-favorite dining list, which tracks trending food and drink interest across 10 cities, including a 216% rise in cheeseburger interest in Tokyo over the past year.

  11. Google Cloud · AI & Machine LearningAI score19

    Irish Brands Scale AI Operations with Gemini Enterprise, from Ryanair to Startups

    AIIrish organizations including Ryanair, Smyths Toys, the Irish Revenue Commissioners, Virgin Media Ireland and startups Hexis, Kitman Labs, Spryt and IMPT are moving agentic AI from prototypes to production using Gemini Enterprise and Google Cloud. Ryanair is deploying Gemini Enterprise and Google Workspace for about 35,000 employees, while Smyths Toys' AI agent Codie has resolved more than 60% of web inquiries. The article also states that Google operations added an estimated 10 billion euros to Irish GDP in 2025.

  12. Google Cloud · AI & Machine LearningAI score78

    Google Cloud launches Gemini agent as single universal work agent

    AIGoogle Cloud announced the Gemini agent, a single agent that answers questions, handles knowledge work, creates media, and writes and runs code from one prompt box. It runs in the cloud with persistent memory, uses multi-agent orchestration, and adds Workspace integration, domain skills for data and industries, identity-based governance through Agent Gateway, and spend caps. The source also cites customer deployments and says nearly 80% of Google Cloud customers use its AI products.

    Why it matters: The announcement shows how a single work agent spans chat, Workspace, data analysis, governance, and cost controls, useful for judging enterprise agent deployment scope.

  13. 🚨 AI News | TestingCatalogAI score23

    Antigravity's agent renamed "Chief of stuff" in latest update

    AIGoogle's Antigravity agent has been renamed "Chief of stuff" in its latest update, which the poster reads as a promotion. The poster wonders whether Antigravity could become a home for Google's own agents, and background notes that Google is prototyping a voice agent internally called "Concierge," which appears to be a very early version.

    Image from @testingcatalog's post
  14. MIT Technology Review · AIAI score44

    AI advances won't quickly make robots useful in everyday life, researchers say

    AIResearchers at robotics labs say that AI advances behind chatbots like ChatGPT and Claude will not quickly produce robots that are useful in everyday life. Many skeptics argue that using language- and image-based intelligence to master the physical world is far harder than it sounds, despite bold predictions from Elon Musk about Tesla's Optimus. Progress is real but incremental, as shown by Google DeepMind's Gemini Robotics controlling ALOHA 2 arms to pack a lunchbox.

  15. Artificial Analysis ArticlesAI score50

    Harvey LAB-AA v1.1 adds hallucination checks to legal AI benchmark

    AIHarvey LAB-AA v1.1 adds hallucination checks that audit every model deliverable against task source documents, with material hallucinations zeroing a task's score. GPT-6 Astra averaged 0.03 material hallucinations per task across 120 tasks, while Gemini 3.8 Flash averaged 13.96. Harvey uses GPT-6 Sol (high) as the hallucination checker, separate from its three-judge rubric panel.

Oct 7

Oct 7Wed
  1. elvisAI score67

    Tool-using multimodal models refuse harmful requests less often, NVIDIA study finds

    AIA NVIDIA study accepted at NeurIPS 2026 reports that multimodal models refuse harmful requests less reliably when they call tools. Refusal failures rise by up to 68.7% relative and by 17.7% on average across the models tested, including Claude Opus 4.6 and 4.7 and Gemini Agentic Vision. The authors attribute this to tool outputs crowding out the original harmful intent and to attention shifting toward describing tool results. Re-inserting the original request and image before the final response restores part of the lost refusals.

    Image from @omarsar0's post
  2. Google Developers BlogAI score62

    Google open-sources ML Drift, a cross-platform GPU engine for on-device AI

    AIGoogle's AI Edge Team open-sourced ML Drift under Apache 2.0, a GPU compute engine for on-device AI inference across OpenGL ES, OpenCL, Metal, and WebGPU. It serves as the core GPU acceleration engine within LiteRT and succeeds the legacy TFLite GPU delegate, which will no longer receive new features. The post cites benchmarks showing up to 40% lower frame latency in YouTube Shorts and up to 30% faster on-device performance in Adobe Lightroom and Photoshop.

    Why it matters: The post explains how ML Drift unifies GPU shaders across platforms and replaces the TFLite GPU delegate, which matters for developers deploying on-device models.

  3. Google Developers BlogAI score62

    Google's AQuA agent diagnoses production failures in a multi-agent travel concierge

    AIGoogle Developers Blog introduces AQuA, an ambient quality agent that runs in a customer's Google Cloud project and samples production sessions to find recurring agent failures. In a 32-session travel-concierge sweep, it verified six issues and traced two of them to specific prompt lines, and a replay after the fixes raised full-session passes from 5/32 to 13/32. The post notes that verification and diagnosis are model-based, and that the tool proposes edits without applying them.

    Why it matters: The post walks through a concrete production workflow, from sweep and verification to a code-anchored fix and replay, that shows how to diagnose silent agent failures.

  4. Google ResearchAI score23

    Google Research invites COLM visitors to ContinuousBench walkthrough on DP synthetic data

    AIGoogle Research is hosting a walkthrough at its COLM booth #107 today at 5:00 PM of ContinuousBench, a standardized benchmark for measuring knowledge transfer in differentially private synthetic data. The session, led by Alex Bie, asks whether DP synthetic data preserve actual information or only style. A paper is linked on arXiv.

    Image from @GoogleResearch's post
  5. MarkTechPostAI score67

    Anthropic releases Claude Haiku 5.5, a small model with 1M context

    AIAnthropic has released Claude Haiku 5.5, its cheapest and fastest small model, priced at $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100K tokens. It keeps a 1M token context window, up to 128K output tokens, and is generally available on the Claude API, Amazon Bedrock, Google Cloud, Microsoft Foundry and Claude Platform on AWS. Anthropic reports 72.4% on OSWorld 2.1 (offline subset) versus 15.7% for Haiku 4.5, and the article notes that non-default temperature, top_p or top_k values return a 400 error.

  6. Google ResearchAI score62

    Google Research finds AI boosts patent drafting but junior lawyers' gains vanish without it

    AIA Google Research field experiment with 133 patent lawyers found AI tool access raised drafting scores by 0.34 to 0.38 standard deviations over three months. When the tool was removed for a redlining task, only senior lawyers kept an advantage of 0.45 SD, while junior lawyers showed no discernible improvement. The authors argue that tools which boost current output must not stop junior professionals from building the judgment that senior experts rely on.

    Why it matters: The field experiment separates AI's short-term productivity gains from skill retained after the tool is removed, which matters for training junior professionals.