Model Detail
VibeVoice-ASR-Streaming-7B
▲ 7062.5%VibeVoice-ASR-Streaming-7B is an audio model with 7B parameters released by Microsoft. The model is registered under the automatic-speech-recognition pipeline tag on Hugging Face, distributed under the permissive mit license.
VibeVoice-ASR-Streaming-7B ships with 7B parameters. Total weight footprint is approximately 8.7 GB, which is the relevant figure when planning local-inference VRAM. The mit license is permissive, allowing commercial deployment and derivative work without per-seat fees, though attribution requirements still apply.
Downloads of VibeVoice-ASR-Streaming-7B have moved +7062.5% over the past 24 hours. That is a slight downtrend, consistent with normal cooling as newer models compete for the same workloads. These numbers are signal, not guarantee — week-over-week download counts on Hugging Face also reflect mirror traffic, CI scrapes, and one-off benchmarking runs.
VibeVoice-ASR-Streaming-7B is best fit for speech recognition, transcription, or speech synthesis depending on the task head. Treat this as a starting matrix rather than a benchmark verdict — the right deployment usually depends on the specific evaluation suite that mirrors your workload.