MuseDirectory
All connectors / Lifestyle
Lifestyle

Speech AI

Pronunciation scoring, speech-to-text, and text-to-speech for English language learning

Assess English pronunciation quality from audio with phoneme-level scoring. Transcribe speech to text with timestamps. Generate natural speech from text using 12 English voices. Supports multiple audio formats and includes multilingual transcription via Whisper.

Try asking Muse: "Score my pronunciation of this English sentence and tell me which words I need to practice."

Working now?
Working
Worked last 7 days
93.8% of checks
How to get it
Extra setup
Account
No account needed
Price
Not stated

Source: Found in the official MCP Registry (io.github.fasuizu-br/speech-ai) · First listed September 25, 2026

Last 24 hours

Each bar is one check, every 15 minutes. Green means it answered. Last checked 53 min ago.

Is it safe to connect?

What Muse can see: It does not ask you to sign in, so it cannot see your accounts. It only sees what Muse sends it from your request.

Before it acts: Read what Muse plans to do before you approve it, and remove the app from Muse when you stop using it.

musedirectory.ai is not part of Meta. More about how Muse handles your information

How to add it to Muse

Not in Muse's Connectors list yet, but Muse can still use it. Copy the request below and paste it into Muse. Muse asks before it shares anything with the app's site. Meta does not review apps used this way, so only use ones you trust. We tested this in the Muse app on September 24, 2026: Muse used an app's link directly this way and returned a live answer.

Open Muse
https://apim-ai-apis.azure-api.net/mcp/pronunciation/mcp

Not in Muse's Connectors list yet. Muse can still use it: its page gives you a request to paste into Muse.

Technical details: safety screening, tools, response time and recent checks

Response time at the last check: 1900ms. Worked in 93.8% of checks over 30 days.

Screening · September 25, 2026

Screened, no issues found

MCP handshake
Pass
Answered in 527ms
Domain against threat feeds (Cloudflare security DNS)
Pass
apim-ai-apis.azure-api.net, brainiall.com not flagged
Published packages against the OSV malicious-package database
n/a
No npm or PyPI package published
Hidden instructions or invisible characters in tool text
Pass
10 tools read, nothing found
Inputs asking for passwords, card numbers or seed phrases
Pass
None found
Domain and redirects
Pass
No redirects off the domain
AI review of purpose and tool behavior
Pass
No concerns

Screening looks for known threats and hidden instructions at the time of the check, and runs again weekly and whenever the tool list changes. It cannot see the server's code, so only connect what you need and review what Muse asks to do. How screening works.

Tools · 10

assess_pronunciationAssess English pronunciation quality from audio. Scores pronunciation at four levels: overall, sentence, word, and phoneme. Each score is 0-100. Phonemes are returned in both IPA and ARPAbet notation. Sub-300ms inference latency. Args: audio_base64: Base64-encoded audio data. Sup
check_pronunciation_serviceCheck if the pronunciation assessment service is healthy and ready. Returns: dict with keys: - status (str): 'healthy' or error state - modelLoaded (bool): Whether the scoring model is loaded - version (str): API version
get_phoneme_inventoryGet the full phoneme inventory supported by the pronunciation scorer. Returns a list of all English phonemes the engine can assess, including ARPAbet symbol, IPA equivalent, example word, and phoneme category (vowel, consonant, diphthong). Returns: list of dicts, each with keys:
transcribe_audioTranscribe audio to text with word-level timestamps. Converts spoken English audio into text with optional word-level timestamps and per-word confidence scores. Args: audio_base64: Base64-encoded audio data (WAV, MP3, OGG, FLAC, WebM). audio_format: Audio format hint. Auto-detect
check_stt_serviceCheck if the speech-to-text service is healthy and ready. Returns: dict with keys: - status (str): 'healthy' or error state - modelLoaded (bool): Whether the STT model is loaded - version (str): API version
synthesize_speechGenerate natural speech audio from English text. Produces high-quality speech with 12 English voices. Returns base64-encoded WAV audio (16-bit PCM, 24kHz mono) along with metadata. Available voices: - af_heart (default), af_bella, af_nicole, af_sarah, af_sky (American female) - a
list_tts_voicesList all available text-to-speech voices with metadata. Returns: dict with keys: - voices (list): Available voices, each with id, name, gender, accent, grade - defaultVoice (str): Default voice ID
check_tts_serviceCheck if the text-to-speech service is healthy and ready. Returns: dict with keys: - status (str): 'healthy' or error state - modelLoaded (bool): Whether the TTS model is loaded - version (str): API version
transcribe_audio_proTranscribe audio with Whisper Large V3 Turbo — multilingual STT. Supports 99 languages with automatic language detection, word-level timestamps, per-word confidence scores, and optional speaker diarization (identifies who spoke each word). Best-in-class WER (~2%). Args: audio_bas
check_whisper_serviceCheck if the Whisper STT Pro service is healthy and ready. Returns: dict with keys: - status (str): 'healthy' or error state - modelLoaded (bool): Whether the Whisper model is loaded - diarizeLoaded (bool): Whether the diarization pipeline is loaded - version (str): API version -

Recent checks

2026-09-28 08:16:13live · HTTP 2001900ms
2026-09-28 06:01:02live · HTTP 200459ms
2026-09-28 03:31:21live · HTTP 200601ms
2026-09-28 01:16:12live · HTTP 200874ms
2026-09-27 23:00:56live · HTTP 2002035ms
2026-09-27 20:31:26live · HTTP 2001523ms
2026-09-27 18:16:11live · HTTP 2002355ms
2026-09-27 16:00:58down · no response8001ms
2026-09-27 13:46:50live · HTTP 2003086ms
2026-09-27 11:31:37live · HTTP 2001584ms
2026-09-27 09:15:52live · HTTP 2001570ms
2026-09-27 06:46:31live · HTTP 2001515ms

Link

https://apim-ai-apis.azure-api.net/mcp/pronunciation/mcp

Questions about Speech AI in Muse

How do I connect Speech AI to Muse?

Not in Muse's Connectors list yet, but Muse can still use it. Paste this into Muse: "Use Speech AI to help me. It is a free service with an MCP server at https://apim-ai-apis.azure-api.net/mcp/pronunciation/mcp. It does not need an API key. Ask me before you share anything with it." Muse asks before it shares anything with the app's site. Meta does not review apps used this way, so only use ones you trust. We tested this in the Muse app on September 24, 2026: Muse used an app's link directly this way and returned a live answer.

Is Speech AI working right now?

At the last check (September 28, 2026, 08:16 UTC) the endpoint was working, answering in 1900ms. Over the last 7 days it answered 93.8% of health checks. It is checked every 15 minutes.

Is Speech AI safe to connect to Muse?

It was screened on September 25, 2026 with the result "screened, no issues found". Screening checks the domain against threat feeds and reads the tools for hidden instructions and requests for passwords or card numbers. It cannot see the server's code, so grant only the access you need.

What can Speech AI do in Muse?

It exposes 10 tools, including assess_pronunciation, check_pronunciation_service, get_phoneme_inventory, transcribe_audio. For example, you could ask Muse: "Score my pronunciation of this English sentence and tell me which words I need to practice."

Follow Speech AI

Tell me if Speech AI goes down, comes back, or changes its tools
One email per change. You confirm first, and every email has a link to stop.
Made this app? Claim it or get a status badge

Is this yours?

This listing was added from public sources (Found in the official MCP Registry (io.github.fasuizu-br/speech-ai)). If you build Speech AI, claim it to correct the details and get your badge.

Claim this listing

Speech AI status badge

Paste this on your site or README. It always shows the latest check.

<a href="https://musedirectory.ai/connector/speech-ai"><img src="https://musedirectory.ai/badge/speech-ai.svg" alt="Speech AI on musedirectory.ai" width="236" height="40"></a>

More for app makers