Skip to content

Companies & models · Latest news

Google / Gemini

Follow Google and DeepMind AI news: Gemini models, Veo video models, research, and products.

98 top picks all-time · 57 in the past 30 days · chosen from 670 items collected all-time

Latest pick

Top picks archive · Page 5

Top picks 81–98 of 98

Aug 13

Aug 13Thu
  1. Google AI DevelopersOfficialAI score75

    Google releases Gemini 3.7 Flash for coding and agentic tasks

    AIGoogle AI Developers announced Gemini 3.7 Flash as its most intelligent workhorse model yet for coding and agents, citing higher instruction adherence, first-pass code accuracy, and high-quality agentic execution. The post shows the model building a complex 3D web game in Antigravity, covering Three.js engine logic, asset orchestration with PBR textures and Nano Banana sprite sheets, and procedural sound effects.

    Why it matters: The post shows a concrete build workflow across engine logic, assets, and audio, which helps readers judge how the model handles multi-step agentic coding.

    Video from @googleaidevs's post
  2. Demis HassabisXAI score67

    Google releases Gemini 3.7 Flash with coding and web development upgrades

    AIGoogle DeepMind has released Gemini 3.7 Flash, which the post says is stronger for coding, knowledge work, and web development. Its introductory price is half the original cost of Gemini 3.6 Flash.

    Why it matters: The post names concrete upgrade areas and a price change against the prior version, which helps readers compare it with earlier Flash releases.

  3. koray kavukcuogluXAI score72

    Google launches Gemini 3.7 Flash for coding and agentic workflows

    AIGoogle launches Gemini 3.7 Flash, its latest Flash model for coding and agentic workflows, with an introductory price at half the original cost of 3.6 Flash. The post reports gains from 3.5 to 3.7 Flash, including DeepSWE v1.1 rising from 37.0% to 65.3%, Code Arena Elo from 1506 to 1588, and AutomationBench from 13.4% to 30.4%.

    Why it matters: The post pairs a launch with specific before-and-after benchmark gains and an introductory price, letting readers weigh capability against cost for coding and agent work.

    Image from @koraykv's post

Jul 29

Jul 29Wed
  1. Google LabsOfficialAI score60

    Google Launches Lyria 3.5 in Flow Music With Better Vocals and Lyrics

    AIGoogle is rolling out Lyria 3.5, its newest music generation model, in Google Flow Music today. The update improves musicality, lyric quality and prompt adherence, and vocal expressiveness and pronunciation, and gives users more control over tempo and duration.

    Why it matters: The post names the specific capability changes and where users can access them, which helps readers judge fit for music creation workflows.

Jul 21

Jul 21Tue
  1. koray kavukcuogluXAI score72

    Google releases Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber

    AIGoogle introduces Gemini 3.6 Flash as its workhorse model, with better coding, knowledge work, and multimodal performance while reducing token usage. It also launches Gemini 3.5 Flash-Lite, described as the fastest and most cost-effective 3.5-class model for high-throughput applications, and 3.5 Flash Cyber, a version of 3.5 Flash fine-tuned to find and fix cybersecurity vulnerabilities.

    Why it matters: The post lists three distinct models, each aimed at a different job, so readers can map which one fits coding, high-volume, or security workloads.

    Video from @koraykv's post
  2. Jeff DeanXAI score60

    Google's Gemini 3.6 Flash uses fewer tokens than 3.5 Flash at the same cost

    AIGoogle released Gemini 3.6 Flash alongside Gemini 3.5 Flash-Lite and Gemini 3.5 Flash Cyber, aimed at faster and cheaper AI agents. Jeff Dean says Gemini 3.6 Flash is much more token efficient than Gemini 3.5 Flash and shares a side-by-side video demonstration.

    Why it matters: The post contrasts token use between two Flash models at the same cost, which is useful for judging efficiency gains in Gemini's agent-oriented releases.

    Video from @JeffDean's post

May 19

May 19Tue
  1. koray kavukcuogluXAI score72

    Google's Gemini 3.5 Flash beats Gemini 3.1 Pro on coding and agentic benchmarks

    AIGoogle's Gemini 3.5 Flash outperforms Gemini 3.1 Pro on Terminal-Bench 2.1 (76.2%), GDPval-AA (1656 Elo), and MCP Atlas (83.6%). The post also claims it is 4x faster than other frontier models, or 12x in Antigravity, and reports 83.6% on MMMU-Pro for multimodal performance.

    Why it matters: The post gives specific benchmark scores against Gemini 3.1 Pro, letting readers compare coding, agentic, and multimodal results directly.

    Image from @koraykv's post
  2. koray kavukcuogluXAI score72

    Google rolls out Gemini 3.5 Flash globally across consumer, developer, and enterprise platforms

    AIGoogle is rolling out Gemini 3.5 Flash globally for consumers in the Gemini app and Search AI Mode. It is also available to developers through the Gemini API, Google Antigravity, and Google AI Studio, and to businesses on the Gemini Enterprise Agent Platform.

    Why it matters: The post shows where each Gemini 3.5 Flash access path goes, from consumer apps to developer and enterprise platforms, which helps readers pick the right entry point.

  3. koray kavukcuogluXAI score62

    Google introduces Gemini 3.5 Flash, used with agents to rebuild AlphaZero

    AIAt Google I/O, Google introduced Gemini 3.5 Flash, which the author says has become part of the daily research cycle. The author says a team of agents in Antigravity 2.0 recreated the original AlphaZero paper and built a playable web version from two prompts, coding the reinforcement learning pipeline in JAX/Flax and training a ResNet model via self-play on multi-TPU pods.

    Why it matters: The post shows Gemini 3.5 Flash applied to an end-to-end agent task, recreating and training AlphaZero from two prompts, which indicates practical coding and research use.

    Video from @koraykv's post
  4. Oriol VinyalsXAI score62

    Gemini 3.5 Flash is now available globally

    AIGoogle's Gemini 3.5 Flash is available today globally, according to Oriol Vinyals. The post invites developers to build agents and apps with it and links to Google's blog for more details.

    Why it matters: The post confirms Gemini 3.5 Flash's global availability and points to the blog for details, the main new information for developers.

Mar 11

Mar 11Wed
  1. Nano Banana 2.1OfficialAI score62

    How to get the most out of Nano Banana 2 for image generation

    AINano Banana 2, also called Gemini 3.1 Flash Image, adds visual grounding with Google Search, 512px resolutions, and extreme aspect ratios of 1:8 and 1:4. The guide advises using it as the default for new projects, with Nano Banana Pro reserved for complex prompts it fails, and keeping Thinking mode off by default.

    Why it matters: The guide compares Nano Banana 1, 2, and Pro with concrete routing advice, which helps developers decide which model to default to and how to control cost.

Feb 26

Feb 26Thu
  1. Oriol VinyalsXAI score60

    Nano Banana 2 debuts at #1 in Image Arena text-to-image ranking

    AINano Banana 2, officially released as Gemini 3.1 Flash Image Preview, ranks first in Image Arena text-to-image with a score of 1279. The quoted post says it also ties for first in single-image editing at 1407 and costs $0.067 per image, about half the price of Nano Banana Pro.

    Why it matters: The quoted leaderboard figures and per-image price give concrete reference points for comparing this image model against Nano Banana Pro and GPT-Image-1.5.

    Image from @OriolVinyalsML's post
  2. Nano Banana 2.1OfficialAI score67

    Google introduces Nano Banana 2, its best image generation and editing model

    AINano Banana announces Nano Banana 2, which it describes as its best image generation and editing model yet. The model can be tried in the Gemini app, Google AI Studio, and other places the post does not specify.

    Why it matters: The post names the access points for Nano Banana 2, which helps readers see where the image generation and editing model can be tried.

Feb 25

Feb 25Wed
  1. Quoc LeXAI score65

    Aletheia Agent Solves 6 of 10 FirstProof Math Problems Autonomously

    AIGoogle researchers used the Aletheia agent, powered by Gemini 3 Deep Think, to attempt 10 FirstProof challenge problems without modification. The agent operated fully autonomously and solved 6 of the 10 problems, according to the post, with methodology and expert evaluations described in the linked arXiv paper.

    Why it matters: The post gives the autonomous setup and expert-evaluated results for an AI agent on FirstProof math problems, useful for judging how far such systems go on research-level math.

    Image from @quocleix's post

Feb 19

Feb 19Thu
  1. Yi TayXAI score78

    Google releases Gemini 3.1 Pro, reporting 77.1% on ARC-AGI-2

    AIGoogle has released Gemini 3.1 Pro, reporting 77.1% on ARC-AGI-2 and more than twice the score of Gemini 3 Pro on that benchmark. The model is rolling out to developers in preview through the Gemini API and Google AI Studio, to enterprises via Vertex AI and Gemini Enterprise, and to consumers in the Gemini app and NotebookLM.

    Why it matters: The post pairs the release with a benchmark table comparing Gemini 3.1 Pro against Gemini 3 Pro, Claude Sonnet 4.6, Claude Opus 4.6, and GPT-5.2 on reasoning and coding tasks.

Feb 11

Feb 11Wed
  1. Quoc LeXAI score62

    Aletheia, a Gemini Deep Think agent, tackles PhD-level math and open Erdős conjectures

    AIQuoc Le announced a paper describing Aletheia, an agent built on Gemini Deep Think that goes beyond Olympiad problems to PhD-level mathematics. The post says Aletheia iteratively generates and verifies proofs, collaborates on human-AI research, autonomously generates a paper on eigenweights, and solves open Erdős conjectures.

    Why it matters: The post details an agent's iterative proof checking and open-problem results, which show how AI research workflows are being tested beyond competition math.

Feb 2

Feb 2Mon
  1. Quoc LeXAI score60

    Gemini Helps Address 13 Open Erdős Problems in Math Case Study

    AIQuoc Le announced a case study using Gemini to systematically evaluate 700 conjectures labeled open in the Erdős Problems database. The team addressed 13 problems, finding 5 novel autonomous solutions and identifying 8 existing solutions missed by previous literature.

    Why it matters: The case study shows how a systematic AI sweep found new solutions and missed prior literature across 700 open Erdős conjectures, offering a concrete look at AI-assisted math research.

    Image from @quocleix's post

Dec 4, 2025

Dec 4, 2025Thu
  1. Quoc LeXAI score62

    Gemini 3 Deep Think mode goes live in the Gemini app for Ultra users

    AIGoogle's Gemini 3 Deep Think mode is now available in the Gemini app for Ultra users. The post says it uses parallel thinking for difficult coding and scientific tasks and builds on technology that reached gold-medal level at the ICPC World Finals and IMO.

    Why it matters: The post names the access tier and the coding and scientific task focus, which helps readers judge whether the mode fits their work.