Skip to content
Agent Search Engine.

Record · funasrSDK / libraryOpen sourceVerified Jul 3, 2026

FunASR

Industrial-grade speech recognition toolkit: 170x realtime, 50+ languages, speaker diarization, emotion detection, streaming, and OpenAI-compatible API.

Best for

Self-hosted speech recognition with VAD, punctuation, and speaker pipelines.

Avoid if

You need one blanket licence — the toolkit is MIT but model licences vary.

About FunASR

Industrial speech recognition. Up to 340x realtime, 26x faster than Whisper. 50+ languages. Speaker diarization · Emotion detection · Streaming · One API call

Quick Start · Colab · Benchmark · Model selection · Migration guide · Use cases · Deployment matrix · Models · Agent Integration · Docs · Contribute

No local setup? Open the Colab quickstart to transcribe a public sample or upload your own audio in a browser.

From the project's README

FunASR is an open-source project written primarily in Python, with 21k stars on GitHub. It was last updated in September 2026.

Install

pip install torch torchaudio
For agent buildersMake your product part of the discovery.Explore advertising

FunASR vs. the alternatives

All 163 alternatives →
RecordStarsPricing
FunASRSDK / librarythis listing21kOpen source
MCP Reference ServersMCP server91kOpen source
context7MCP server62kOpen source
chrome-devtools-mcpMCP server53kOpen source
Playwright MCPMCP server38kOpen source
github-mcp-serverMCP server33kOpen source

See all 163 FunASR alternatives, compared →