Releases from frontier labs, major companies and research institutes — grouped by model, deduplicated across HuggingFace, OpenRouter and official blogs.
Releases per weekFrontierMajor labResearch
Click a week to filter
Tier
Type
Source17 models
Last week1 model
NV
NV-Reason-CT Open 3D CT VLM
NVIDIA·🇺🇸·Major lab
TextVisionopen
NV-Reason-CT processes 3D CT volumes natively (not slice-by-slice) and generates structured diagnostic reports with step-by-step chain-of-thought reasoning validated by NIH radiologists. It scores state-of-the-art on CT-RATE (Macro-F1 0.614).
Week of August 24, 20261 model
NV
Nemotron-3-Diarization
NVIDIA·🇺🇸·Major lab
Otheropen
Nemotron 3 Diarization is an open-weight speaker diarization model designed to determine "who spoke when" in real-world audio. It supports both streaming and offline inference and handles up to eight speakers.
1 variantPreview
Week of August 10, 20261 model
NV
Cmd
NVIDIA·🇺🇸·Major lab
Videoopen
Hmrishav Bandyopadhyay 1,2 , Xuanchi Ren 1 , Zijian Huang 1 , Jay Zhangjie Wu 1 , Tianshi Cao 1 , Ruilong Li 1 , Bryan Chu 1 , Sanja Fidler 1 , Yi-Zhe Song 2 , Zian Wang 1
Week of July 27, 20262 models
NV
Nemotron 3.5 Lightning
NVIDIA·🇺🇸·Major lab
TextReasoningopenAPI
Nemotron 3.5 Lightning ist ein Textmodell von NVIDIA mit einem Kontextfenster von 262 K Tokens. Es ist seit dem 11. August 2026 auf OpenRouter verfuegbar.
▶ Hear it first. Natural turn-taking, barge-in and live tool calling.
1 variant11B
Week of July 20, 20261 model
NV
Cosmos-H-Dreams
NVIDIA·🇺🇸·Major lab
Videoopen
Cosmos-H-Dreams is a real-time, action-conditioned generative surgical world model that lets a human operator or a learned surgical-robotics policy act inside a synthesized surgical scene and observe the interactions live.
Week of July 13, 20261 model
NV
Ising-Calibration-1.5
NVIDIA·🇺🇸·Major lab
TextVisionopen
NVIDIA-Ising-Calibration-1.5-31B-BF16 is a dense multimodal vision-language model built on Gemma 4 31B.
2 variants31B NVFP431B BF16
Week of July 6, 20263 models
NV
Cosmos3-Super-Image2Video-4Step
NVIDIA·🇺🇸·Major lab
Videoopen
NVIDIA Cosmos™ is a world foundation model platform designed to accelerate the development of Physical AI by enabling machines to understand, simulate, and interact with the physical world across robotics, autonomous driving, and smart space environments, including industrial and factory-scale applications.
NV
Cosmos3-Super-Text2Image-4Step
NVIDIA·🇺🇸·Major lab
Imageopen
NVIDIA Cosmos™ is a world foundation model platform designed to accelerate the development of Physical AI by enabling machines to understand, simulate, and interact with the physical world across robotics, autonomous driving, and smart space environments, including industrial and factory-scale applications.
NV
Nemotron-Labs-Audex
NVIDIA·🇺🇸·Major lab
Textopen
We're excited to introduce Nemotron-Labs-Audex-2B, a unified audio-text LLM with a similar recipe as Nemotron-Labs-Audex-30B-A3B. Audex-2B extends the vocabulary for discrete audio tokens used for speech and general audio outputs, as well as an audio encoder for speech and general audio inputs.
2 variants2B30B A3B
Week of June 29, 20262 models
NV
Cosmos3-Edge
NVIDIA·🇺🇸·Major lab
Otheropen
NVIDIA Cosmos3-Edge is a 4B-parameter omni-modal world model that generates text, images, video, and action commands from multimodal inputs. Its Mixture-of-Transformers architecture combines autoregressive decoding for text with diffusion-based denoising for other modalities. Optimized for physical AI, it serves robotics, autonomous vehicles, and smart space simulations.
NV
Nemotron-Parse-2.0
NVIDIA·🇺🇸·Major lab
TextVisionopen
NVIDIA Nemotron Parse 2.0 transforms document images into structured, machine-readable representations with text, layout classes, bounding boxes, and reading-order information.
Week of June 22, 20261 model
NV
Nemotron-Labs-3-Puzzle
NVIDIA·🇺🇸·Major lab
Textopen
Nemotron-Labs-3-Puzzle-75B-A9B is a deployment-optimized large language model developed by NVIDIA, derived from Nemotron-3-Super-120B-A12B.
NVIDIA Nemotron 3.5 Content Safety is a compact 4B-parameter multimodal guardrail model from NVIDIA, fine-tuned from Google Gemma-3-4B. It moderates both inputs to and responses from LLMs and VLMs, accepting...
2 variantsSafetySafety Free
NV
ArtiFixer
NVIDIA·🇺🇸·Major lab
Otheropen
ArtiFixer is a few-step causal auto-regressive model that enhances and extends 3D reconstruction. The related source code provides implementations for training, evaluation, and inference, supporting various stages including bidirectional training, diffusion forcing, and Self-Forcing-style DMD distillation.
Week of April 27, 20261 model
NV
Nemotron 3 Nano Omni
◆ Underrated
NVIDIA·🇺🇸·Major lab
ReasoningVisionopenAPI
NVIDIA's small multimodal reasoning model: 30B MoE, only 3B active, text+image+audio input. Free on OpenRouter. Designed for on-device, 4x faster than its predecessor, reasoning ON/OFF mode.