aiagent.club
English
GitHubPyPI

FunASR

AI Agent 指数 第 28(共 51) ↓ 下载分享海报

51 综合分
可观测采用度当前有效 · 2026-09-12
51
动量当前有效 · 2026-09-12
53
关注度当前有效 · 2026-09-12
49
信号可信度 依据当前有数据的独立评分维度数量计算。
3/3 · 2 种数据源

方法论 v2.0 · 快照 2026-09-12 · 超过 2 天视为过期

项目介绍

Industrial speech recognition. 170x faster than Whisper. 50+ languages.

No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.

Flagship model — Fun-ASR-Nano (LLM-ASR, 31 languages; the default recommendation, needs a GPU):

On CPU (or for multilingual + emotion in one pass), use SenseVoice — which also returns speaker diarization and timestamps:

Output — structured text with speaker labels, timestamps, and punctuation:

That's it. One model, one call — VAD segmentation, speech recognition, punctuation, speaker diarization all happen automatically.

At scale, accelerate Fun-ASR-Nano with vLLM (batch processing):

Deploy as API…

各数据源

355.6k 月下载量
  • 月下载量 355.6k
  • 周下载量 58.7k
  • 日下载量 9.5k
20.3k Star
  • Star 20.3k
  • Fork 2.0k
  • 提交 5.9k
  • 发布 64