ModelRadar
/

What shipped in AI.

Releases from frontier labs, major companies and research institutes — grouped by model, deduplicated across HuggingFace, OpenRouter and official blogs.

Releases per weekFrontierMajor labResearch
10Jun 15Jul 13Aug 10Sep 7Sep 28

Click a week to filter

Last week1 model

NV-Reason-CT Open 3D CT VLM

NVIDIA🇺🇸Major lab

TextVisionopen

NV-Reason-CT processes 3D CT volumes natively (not slice-by-slice) and generates structured diagnostic reports with step-by-step chain-of-thought reasoning validated by NIH radiologists. It scores state-of-the-art on CT-RATE (Macro-F1 0.614).

Week of August 24, 20261 model

Nemotron-3-Diarization

NVIDIA🇺🇸Major lab

Otheropen

Nemotron 3 Diarization is an open-weight speaker diarization model designed to determine "who spoke when" in real-world audio. It supports both streaming and offline inference and handles up to eight speakers.

1 variantPreview

Week of August 10, 20261 model

Cmd

NVIDIA🇺🇸Major lab

Videoopen

Hmrishav Bandyopadhyay 1,2 , Xuanchi Ren 1 , Zijian Huang 1 , Jay Zhangjie Wu 1 , Tianshi Cao 1 , Ruilong Li 1 , Bryan Chu 1 , Sanja Fidler 1 , Yi-Zhe Song 2 , Zian Wang 1

Week of July 27, 20262 models

Nemotron 3.5 Lightning

NVIDIA🇺🇸Major lab

TextReasoningopenAPI

Nemotron 3.5 Lightning ist ein Textmodell von NVIDIA mit einem Kontextfenster von 262 K Tokens. Es ist seit dem 11. August 2026 auf OpenRouter verfuegbar.

6 variantsFree30B A3B Base BF1630B A3B NVFP4 Dspark30B A3B NVFP4 Dflash30B A3B NVFP4

NemotronLabs-VoiceChat

NVIDIA🇺🇸Major lab

Otheropen

▶ Hear it first. Natural turn-taking, barge-in and live tool calling.

1 variant11B

Week of July 20, 20261 model

Cosmos-H-Dreams

NVIDIA🇺🇸Major lab

Videoopen

Cosmos-H-Dreams is a real-time, action-conditioned generative surgical world model that lets a human operator or a learned surgical-robotics policy act inside a synthesized surgical scene and observe the interactions live.

Week of July 13, 20261 model

Ising-Calibration-1.5

NVIDIA🇺🇸Major lab

TextVisionopen

NVIDIA-Ising-Calibration-1.5-31B-BF16 is a dense multimodal vision-language model built on Gemma 4 31B.

2 variants31B NVFP431B BF16

Week of July 6, 20263 models

Cosmos3-Super-Image2Video-4Step

NVIDIA🇺🇸Major lab

Videoopen

NVIDIA Cosmos™ is a world foundation model platform designed to accelerate the development of Physical AI by enabling machines to understand, simulate, and interact with the physical world across robotics, autonomous driving, and smart space environments, including industrial and factory-scale applications.

Cosmos3-Super-Text2Image-4Step

NVIDIA🇺🇸Major lab

Imageopen

NVIDIA Cosmos™ is a world foundation model platform designed to accelerate the development of Physical AI by enabling machines to understand, simulate, and interact with the physical world across robotics, autonomous driving, and smart space environments, including industrial and factory-scale applications.

Nemotron-Labs-Audex

NVIDIA🇺🇸Major lab

Textopen

We're excited to introduce Nemotron-Labs-Audex-2B, a unified audio-text LLM with a similar recipe as Nemotron-Labs-Audex-30B-A3B. Audex-2B extends the vocabulary for discrete audio tokens used for speech and general audio outputs, as well as an audio encoder for speech and general audio inputs.

2 variants2B30B A3B

Week of June 29, 20262 models

Cosmos3-Edge

NVIDIA🇺🇸Major lab

Otheropen

NVIDIA Cosmos3-Edge is a 4B-parameter omni-modal world model that generates text, images, video, and action commands from multimodal inputs. Its Mixture-of-Transformers architecture combines autoregressive decoding for text with diffusion-based denoising for other modalities. Optimized for physical AI, it serves robotics, autonomous vehicles, and smart space simulations.

Nemotron-Parse-2.0

NVIDIA🇺🇸Major lab

TextVisionopen

NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information.

Week of June 22, 20261 model

Nemotron-Labs-3-Puzzle

NVIDIA🇺🇸Major lab

Textopen

Nemotron-Labs-3-Puzzle-75B-A9B is a deployment-optimized large language model developed by NVIDIA, derived from Nemotron-3-Super-120B-A12B.

3 variants75B A9B FP875B A9B BF1675B A9B NVFP4

Week of June 1, 20263 models

Nemotron 3 Ultra

NVIDIA🇺🇸Major lab

ReasoningTextopenAPI

NVIDIA's powerful open-weights MoE model: hybrid LatentMoE + MTP layers, 1M context, 550B/55B active. Free on OpenRouter. Focus: complex multi-agent workflows, code, math, science.

3 variants550B550B A55B NVFP4550B A55B Base BF16

Nemotron 3.5 Content

NVIDIA🇺🇸Major lab

TextReasoningVisionAPI

NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...

2 variantsSafetySafety Free

ArtiFixer

NVIDIA🇺🇸Major lab

Otheropen

ArtiFixer is a few-step causal auto-regressive model that enhances and extends 3D reconstruction. The related source code provides implementations for training, evaluation, and inference, supporting various stages including bidirectional training, diffusion forcing, and Self-Forcing-style DMD distillation.

Week of April 27, 20261 model

Nemotron 3 Nano Omni

◆ Underrated

NVIDIA🇺🇸Major lab

ReasoningVisionopenAPI

NVIDIA's small multimodal reasoning model: 30B MoE, only 3B active, text+image+audio input. Free on OpenRouter. Designed for on-device, 4x faster than its predecessor, reasoning ON/OFF mode.

1 variant30B