GitHubPyPI
FunASR
AI Agent 指数 第 25(共 971)↓ 下载分享海报
74 综合分
真实使用 68
动量 76
关注度 83
项目介绍
Industrial speech recognition. 170x faster than Whisper. 50+ languages.
No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.
Flagship model — Fun-ASR-Nano (LLM-ASR, 31 languages; the default recommendation, needs a GPU):
On CPU (or for multilingual + emotion in one pass), use SenseVoice — which also returns speaker diarization and timestamps:
Output — structured text with speaker labels, timestamps, and punctuation:
That's it. One model, one call — VAD segmentation, speech recognition, punctuation, speaker diarization all happen automatically.
At scale, accelerate Fun-ASR-Nano with vLLM (batch processing):
Deploy as API…
各数据源
469.9k 月下载量
- 月下载量 469.9k
- 周下载量 116.8k
- 日下载量 14.2k
19.6k Star
- Star 19.6k
- Fork 2.0k
- 提交 5.5k
- 发布 41