Ffvoice MCP server
Offline speech-to-text & speaker diarization MCP server: transcribe audio on-device, no cloud
3 stars186 downloads/wk
Reviews
Write oneNobody has reviewed Ffvoice yet.
If you have run it, two minutes of your experience saves the next person an afternoon.
Ffvoice tools (5)
write = sends, deletes, buys or postsRead from the package source without running it. The installed server may list more.
capture_and_captionCapture live microphone audio and produce real-time captions using LiveCaptioner.
capture_and_transcribeCapture live microphone audio for a given duration and transcribe it.
list_audio_devicesList all audio input/output devices available on this machine.
transcribe_fileTranscribe a local audio file to text using offline Whisper ASR.
transcribe_file_with_diarizationTranscribe a local audio file and label each segment with a speaker.
Public scan report
scanner v0.1.9 · 2026-09-27 · same rubric, same numbers if you re-run it
- Code scan9 source files scanned25/25
- –Live reliabilityno gateway calls yet and no remote to proben/a
- –Tool poisoningtools not inspected (local package is not executed); not countedn/a
- Auth qualitylocal package, no credentials required12/15
- Maintenancelast push 13 days ago15/15
- Maintainer identityregistry namespace matches repository owner; GitHub account older than a year8/10
Install Ffvoice in Claude Code, Cursor or VS Code
claude mcp add ffvoice -- uvx ffvoice
What the publisher says
From the Ffvoice repository's README, as published. We do not edit it. Read it on GitHub
ffvoice-engine
<!-- mcp-name: io.github.chicogong/ffvoice -->
<!-- Build & CI Status -->
<!-- License & Language -->
<!-- Platform Support -->
<!-- Version & Community -->
<!-- Dependencies -->
<!-- Code Quality -->
🎙️ Offline speech-to-text & speaker diarization for AI agents — Whisper ASR, live captioning, an MCP server, a CLI and Python bindings. Fully on-device, no cloud API.
🎙️ 离线语音识别 + 说话人分离,AI Agent 开箱即用 —— Whisper 实时转写 · 实时字幕 · MCP server · CLI · Python 绑定 · 100% 本地运行,音频不上云。
Why ffvoice? / 为什么用 ffvoice?
The honest pitch: ffvoice is an integration layer, not a new ASR engine. It embeds whisper.cpp as-is and makes no changes to its accuracy or inference speed. What ffvoice adds is a batteries-included, pre-wired pipeline — microphone capture → RNNoise denoising → VAD segmentation → Whisper ASR → speaker diarization → live captions / WAV / FLAC / subtitles — delivered as a single C++ SDK with Python bindings, a CLI, and an MCP server that lets AI agents (Claude and others) transcribe audio out of the box — all in one pip install or cmake build.
诚实定位: ffvoice 是一个集成层,而非新的 ASR 引擎。它内嵌 whisper.cpp,不修改其识别精度或推理速度。ffvoice 带来的是一条开箱即用、预连接的完整管道——麦克风采集 → RNNoise 降噪 → VAD 分段 → Whisper ASR → 说话人分离 → 实时字幕 / WAV / FLAC / 字幕输出——打包成 C++ SDK + Python 绑定 + CLI,外加一个 MCP server,让 AI agent(Claude 等)开箱即用地转写音频——一条 pip install 或 cmake 即可完成。
Pain points it addresses / 解决的痛点
Shortened. The full README is on GitHub.
Nothing above is checked by us. What we check is on the safety report.
Ffvoice: common questions
- Is Ffvoice MCP server safe?
- Yes, by our scan: it is graded A (92/100). Read the Ffvoice safety report
- How do I install Ffvoice?
- It runs on your machine. Copy the Claude Code, Cursor, VS Code or Claude Desktop config from the install section.
- Does Ffvoice need an API key?
- Not as far as the registry entry and our scan can tell: no credentials are declared or required.
- Is Ffvoice maintained?
- The last commit was 13 days ago (2026-09-15). The latest release is v0.8.3.
- What can I use instead of Ffvoice?
- Servers from other publishers that do the same job: FunASR MCP server, three.ws Audio MCP server and Justtranscribe MCP server. Compare all Ffvoice alternatives.
Alternatives to Ffvoice
Same job from other publishers: the closest match first, then the best rated.
- FunASRTranscribe local audio with FunASR and SenseVoice using private, on-device inference.not reviewedEstablishedA
- three.ws AudioText-to-speech, speech-to-text, audio-to-face lipsync, and motion-capture clips for 3D agents.not reviewedEstablishedB
- JusttranscribeTranscribe public videos & audio (YouTube, TikTok, IG) into accurate, timestamped text via API.not reviewedGrowingA
ElevenlabsConvert text to speech, transcribe audio, and dub videos with AI voicesnot reviewedGrowingB- Transloadit Media ProcessingProcess video, audio, images, and documents with 86+ cloud media processing robots.not reviewedEstablishedA