aiagent.club
English
PyPI 包

funasr

更新于 2026-09-12 查看源 ↗

在全部 PyPI 包 中按月下载量排名第 40(共 54)

355.6k 月下载量

OpenAI-compatible speech recognition toolkit with WebSocket streaming, vLLM acceleration, and llama.cpp/GGUF edge runtime.

项目介绍

Industrial speech recognition. 170x faster than Whisper. 50+ languages.

No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.

Flagship model — Fun-ASR-Nano (LLM-ASR, 31 languages; the default recommendation, needs a GPU):

On CPU (or for multilingual + emotion in one pass), use SenseVoice — which also returns speaker diarization and timestamps:

Output — structured text with speaker labels, timestamps, and punctuation:

That's it. One model, one call — VAD segmentation, speech recognition, punctuation, speaker diarization all happen automatically.

At scale, accelerate Fun-ASR-Nano with vLLM (batch processing):

Deploy as API…

摘自 pypi.org/project/funasr/

最新指标

月下载量 355.6k 2026-09-12
周下载量 58.7k 2026-09-12
日下载量 9.5k 2026-09-12