ModelRadar
/

VibeVoice-ASR-Streaming

Microsoft🇺🇸 USMajor lab· Sep 2, 2026

Audioopen weights

VibeVoice-ASR-Streaming is a unified streaming ASR model that transcribes Who (Speaker) said What (Content), with support for Customized Hotwords and 10 languages.

From the model card

License
mit
Modalities
audio → text

2 variants

1.5B1.5BHF
7B7BHF

Sources

HuggingFacePaper

Provenance

Detected
Sep 27, 2026
First source
huggingface
Description
From the model card
Editorially reviewed
—