Text to Speech API alternatives
23 media MCP servers from other publishers do the same job as Text to Speech API. The closest match comes first, then the best rated, with at most two per publisher.
Text-to-speech, speech-to-text, audio-to-face lipsync, and motion-capture clips for 3D agents.
No reviews yetlocal · npm20638/wkFreeOffline speech-to-text & speaker diarization MCP server: transcribe audio on-device, no cloud
No reviews yetlocal · pypi3186/wkFreeConvert text to speech, transcribe audio, and dub videos with AI voices
No reviews yetlocal · mixed31154/wkFreeGenerate video, images, audio and speech with Vidofy — Veo 3.1, Kling 3.0, Flux 2 and 570+ models.
No reviews yetgateway-ready1298/wkFreeTranscribe public videos & audio (YouTube, TikTok, IG) into accurate, timestamped text via API.
No reviews yetgateway-ready37/wkFreeTranscribe local audio with FunASR and SenseVoice using private, on-device inference.
No reviews yetlocal · oci21kFreeProcess video, audio, images, and documents with 86+ cloud media processing robots.
No reviews yetgateway-ready732.1k/wkFreeAutomate Google NotebookLM — Q&A with citations, audio, video, content generation
No reviews yetlocal · npm179672/wkFreeAny file → clean Markdown for AI agents: PDF, Office, EPUB, HTML, images, audio/video. Local MCP.
No reviews yetlocal · npm935FreeTransform video, audio and images, and generate media from prompts. FFmpeg, captions, models.
No reviews yetgateway-ready1512/wkFreeMistral AI MCP server: chat, OCR, Voxtral audio, Codestral FIM, vision, agents, batch.
No reviews yetlocal · npm15106/wkFreeMusic, image, video and audio generation across top AI providers - one key, one credit pool.
No reviews yetlocal · npm687/wkFreeElevenLabs MCP server: TTS, music, sound effects, voices, and audio transcription
No reviews yetlocal · npm1276/wkFreeInspect local audio files — playback, metadata, loudness, spectrogram.
No reviews yetlocal · npm4346/wkFreeChat with 300+ LLMs via OpenRouter. Analyze and generate images, audio, and video from MCP.
No reviews yetlocal · mixed92FreeAI co-pilot for Foundry VTT: runs combat, scenes, voices and audio. Requires the Familiar module.
No reviews yetlocal · npm1371/wkFreeParallel video rendering tools: detect GPU encoders, render, color grade, merge audio, concat.
No reviews yetlocal · npm368/wkFreeImage & PDF tools for AI agents: compress, convert, resize, PDF, AI vision, pipeline.
No reviews yetgateway-ready7Free10 TikTok creator tools — video/MP3 download, daily trends for 16 countries, hashtags, hooks.
No reviews yetlocal · npm199/wkFreeControl macOS system settings, apps, windows, audio, displays, screenshots, and Focus mode via MCP.
No reviews yetlocal · npm3120/wkFreeAuthor Rive (.riv/.rev) and convert Lottie, After Effects, Figma and Spline into it, offline.
No reviews yetlocal · npm344/wkFreeAudio mastering for AI agents: LUFS/True Peak targets, Suno/Udio AI-fingerprint removal.
No reviews yetgateway-ready29/wkFreeLocal image tools: circle crop, crop, resize, compress, convert JPG/PNG/WebP/AVIF. By RoundCut.
No reviews yetlocal · npm226/wkFree
Text to Speech API alternatives by what you need
Open source Text to Speech API alternatives (19)
Published under an open licence such as MIT or Apache-2.0.
three.ws Audio, Ffvoice, Vidofy, Justtranscribe, FunASR, Transloadit Media Processing, NotebookLM MCP, anymd and 11 more below.
Hosted Text to Speech API alternatives (6)
Nothing to install: a remote endpoint you add by URL. Also callable through the mcp.market gateway.
Vidofy, Justtranscribe, Transloadit Media Processing, Rendobar, SammaPix — Image & PDF tools for AI agents, Magic Master — Audio Mastering.
Local Text to Speech API alternatives (17)
A package you install and run on your own machine, so your data stays on your computer.
three.ws Audio, Ffvoice, Elevenlabs, FunASR, NotebookLM MCP, anymd, Mistral, Aetherwave and 9 more below.
Safest Text to Speech API alternatives (15)
Graded A: passed every check the scanner runs on the latest release.
Ffvoice, Vidofy, Justtranscribe, FunASR, Transloadit Media Processing, anymd, Rendobar, ElevenLabs and 7 more below.
Text to Speech API and its alternatives compared
| Server | Grade | Rating | Licence | Runs | Stars |
|---|---|---|---|---|---|
| Text to Speech API(this one) | D47 | No reviews yet | MIT | Hosted | 0 |
| three.ws Audio | B83 | No reviews yet | Apache-2.0 | Local | 206 |
| Ffvoice | A92 | No reviews yet | MIT | Local | 3 |
| Elevenlabs | B80 | No reviews yet | NOASSERTION | Local | 31 |
| Vidofy | A98 | No reviews yet | MIT | Hosted | 1 |
| Justtranscribe | A95 | No reviews yet | MIT | Hosted | 0 |
| FunASR | A90 | No reviews yet | MIT | Local | 21k |
| Transloadit Media Processing | A92 | No reviews yet | MIT | Hosted | 73 |
| NotebookLM MCP | C63 | No reviews yet | MIT | Local | 179 |
| anymd | A94 | No reviews yet | MIT | Local | 935 |
| Rendobar | A99 | No reviews yet | MIT | Hosted | 1 |
| Mistral | B83 | No reviews yet | MIT | Local | 15 |
| Aetherwave | B83 | No reviews yet | MIT | Local | 6 |
| ElevenLabs | A85 | No reviews yet | NOASSERTION | Local | 12 |
| Audio File MCP App | A89 | No reviews yet | ISC | Local | 43 |
| Openrouter Multimodal | B83 | No reviews yet | Apache-2.0 | Local | 92 |
| Familiar | B83 | No reviews yet | NOASSERTION | Local | 1 |
| Ffmpeg Render Pro | A89 | No reviews yet | MIT | Local | 3 |
| SammaPix — Image & PDF tools for AI agents | A87 | No reviews yet | MIT | Hosted | 7 |
| TikTapDown | A92 | No reviews yet | MIT | Local | 1 |
| Macos | A92 | No reviews yet | Apache-2.0 | Local | 3 |
| Polymation | C68 | No reviews yet | NOASSERTION | Local | 0 |
| Magic Master — Audio Mastering | A92 | No reviews yet | MIT | Hosted | 0 |
| Roundcut | A89 | No reviews yet | MIT | Local | 0 |