Project overview
funasr-transcribe runs FunASR and SenseVoiceSmall locally to transcribe m4a, mp3, wav, and other audio, with support for 50-plus languages, voice-activity segmentation, punctuation, filler removal, and automatic device selection. It needs no API key and uploads nothing, but the first transcription downloads roughly 1.2 GB of models and requires Python 3.9–3.11. A follow-up workflow can organize the transcript into meeting notes.
Repository facts
- Primary language
- Python
- License
- MIT
- Repository updated
- Jul 8, 2026
- Default branch
- main
Resource types
General skill
Use cases
Media, audio and video
Platforms
Claude Code, Codex, and more
Runtime
Local
Capabilities
Audio processingTranscription
Audience
General users