Model Detail
Kokoro-82M
▼ 0.2%Kokoro-82M is an audio model released by hexgrad. The model is registered under the text-to-speech pipeline tag on Hugging Face, distributed under the permissive apache-2.0 license.
The apache-2.0 license is permissive, allowing commercial deployment and derivative work without per-seat fees, though attribution requirements still apply.
Downloads of Kokoro-82M have moved -0.2% over the past 24 hours, +1.0% over the trailing seven days, -1.2% over the trailing thirty days. The trend is mildly positive, consistent with a model that is being picked up incrementally rather than going viral. These numbers are signal, not guarantee — week-over-week download counts on Hugging Face also reflect mirror traffic, CI scrapes, and one-off benchmarking runs.
Kokoro-82M is best fit for speech recognition, transcription, or speech synthesis depending on the task head. Treat this as a starting matrix rather than a benchmark verdict — the right deployment usually depends on the specific evaluation suite that mirrors your workload.