🤖

AI Backends

Connect cloud or local services for speech recognition, synthesis, and text generation.

ASR · Speech-to-Text

Used for auto-transcribing reference audio. The app calls these when you click Re-recognise text.

Free tiers available
Groq Whisper
whisper-large-v3-turbo · OpenAI-compatible
Free
2 000 req / day Fastest inference OpenAI API format
Get key ↗
Endpoint https://api.groq.com/openai/v1
whisper-large-v3-turbo whisper-large-v3 distil-whisper-large-v3-en
🤗
HuggingFace Inference
Serverless Whisper models
Free
~1 000 req / day Slower cold starts Many model variants
Get key ↗
Endpoint https://api-inference.huggingface.co/models/openai/whisper-large-v3
📋
AssemblyAI
High-accuracy transcription + speaker diarization
Free
100 h lifetime Speaker labels Auto-chapters
Get key ↗
Endpoint https://api.assemblyai.com/v2/transcript