ModelRadar
/

What shipped in AI.

Releases from frontier labs, major companies and research institutes — grouped by model, deduplicated across HuggingFace, OpenRouter and official blogs.

Releases per weekFrontierMajor labResearch
10Jun 15Jul 13Aug 10Sep 7Sep 28

Click a week to filter

Last week1 model

Gemini 3.8 Live

Google DeepMind🇺🇸Frontier

Text

Gemini 3.8 Live erweitert ein Live-Dialytemodell um eine Videoavatar-Funktion, die Sprache und visuelles Feedback simultan verarbeitet — mit Lippen-Sync, multi-modalen Eingaben (Video plus Audio) und asynchronem Tool-Calling waehrend des Dialoges. Es laeuft nahtlos ueber 97 Sprachen hinweg ohne Abbruch der Videofidelitaet.

Week of August 31, 20262 models

WeatherNext 3

Google DeepMind🇺🇸Frontier

Other

AI weather model that uses real-time satellite data instead of traditional physics simulations to produce hourly forecasts at five times the resolution of its predecessor. Also tracks precipitation and includes clean-energy variables.

Gemini 3.8 Flash

Google DeepMind🇺🇸Frontier

TextReasoningVisionAPI

Gemini 3.8 Live adds near real-time streaming video with expressive avatars — it processes audio and visual inputs simultaneously, supports precise lip-syncing across 97 languages, and can execute tool calls asynchronously while maintaining an uninterrupted conversation.

1 variantBatch

Week of August 10, 20261 model

Gemini 3.7 Flash

Google DeepMind🇺🇸Frontier

TextReasoningVisionAPI

Gemini 3.7 is a model developed by Google DeepMind, featuring two variants: Flash and Flash Batch. It has a context window of 1,048,576 tokens.

1 variantBatch

Week of July 20, 20262 models

Gemini 3.6 Flash

Google DeepMind🇺🇸Frontier

TextReasoningVisionAPI

Gemini 3.6 Flash consumes 17 % fewer output tokens than its predecessor and completes multi-step workflows with fewer reasoning steps and tool calls. It targets coding, knowledge work, and agentic automation at lower cost.

1 variantBatch

Gemini 3.5 Flash Lite

Google DeepMind🇺🇸Frontier

TextReasoningVisionAPI

Fastest and most cost-effective 3.5-class model delivering 350 output tokens per second according to the Artificial Analysis Index, significantly outperforming prior Flash-Lite generations in agentic workflows.

1 variantBatch

Week of July 13, 20261 model

Gemini 3.5 Flash Cyber

Google DeepMind🇺🇸Frontier

Text

A lightweight cybersecurity model built on Gemini 3.5 Flash and fine-tuned to find, validate, and patch software vulnerabilities in code more efficiently than the mainline Flash models. It is designed for use by security agents scanning large codebases at scale.

Week of June 29, 20262 models

Nano Banana 2 Lite

Google DeepMind🇺🇸Frontier

TextReasoningVisionImageAPI

A text-to-image generation and editing model used for creating graphics, posters, and social-media content. It supports object segmentation, direct in-image text editing, and is integrated into Google Workspace apps.

Tabfm 1.0.0

Google DeepMind🇺🇸Frontier

Otheropen

TabFM is a zero-shot tabular foundation model from Google Research.

2 variantsJaxPytorch

Week of June 15, 20262 models

Nano Banana Pro

Google DeepMind🇺🇸Frontier

TextReasoningVisionImageAPI

Nano Banana Pro is Google’s most advanced image-generation and editing model, built on Gemini 3 Pro. It extends the original Nano Banana with significantly improved multimodal reasoning, real-world grounding, and...

Gemini 3.1 Flash Image

Google DeepMind🇺🇸Frontier

VisionImageAPI

Google's latest Gemini Flash generation with native image input/output. Available on OpenRouter ($0.50/$3 per MTok). Gemini 3.x is a standalone series alongside Gemini 2.5.

Week of June 8, 20261 model

Diffusiongemma

Google DeepMind🇺🇸Frontier

TextVisionopen

DiffusionGemma is a generative model built by Google DeepMind. Based on the 26B A4B Mixture-of-Experts (MoE) Gemma 4 architecture, DiffusionGemma generates tokens using discrete diffusion.

1 variant26B A4B IT

Week of May 18, 20261 model

Gemini 3.5 Flash

Google DeepMind🇺🇸Frontier

ReasoningVisionAPI

Google's current Flash generation: 1M context, all modalities, $1.50/$9 per MTok on OpenRouter. Gemini 3.5 Flash is the speed tier of the Gemini 3.x series that replaces Gemini 2.5.

Week of May 4, 20261 model

Gemini 3.1 Flash Lite

Google DeepMind🇺🇸Frontier

VisionTextAPI

Google's cheapest Gemini 3.x option: 1M context, $0.25/$1.50 per MTok. The 'Lite' variant is the first choice when token costs are critical and multimodality is needed.

Week of February 23, 20261 model

Gemma 3n

◆ Underrated

Google DeepMind🇺🇸Frontier

Visionopen

Google's on-device multimodal model: handles text, image, video AND audio at just 4B effective parameters. MatFormer architecture enables sub-models within the same checkpoint. Designed for edge deployments.

1 variantE4B

Week of March 10, 20251 model

Gemma 3

Google DeepMind🇺🇸Frontier

Visionopen

Google's open-weights flagship: 27B (also 1B, 4B, 12B), 128K context, text+image, 140+ languages. Gemma 3 is the basis for many community fine-tunes. Superseded by Gemma 4.

1 variant27B