Local and cloud services for language generation, speech recognition, and synthesis.
Pick the Active Language Model below — it is used by default for every LLM task (persona rewriting, transcription refinement, conversation, and rehearser character analysis). The cards below let you connect engines and apply a URL with Use as LLM.
Your Docker stack containers appear at the top. Local runners and cloud APIs below.
Docker-detected engines and local presets are listed together here. Click Use as TTS to apply a URL to this app’s backend settings.
Default endpoint & model for all LLM tasks. Per-task dropdowns can still override it.
https://api.groq.com/openai/v1
https://openrouter.ai/api/v1
https://generativelanguage.googleapis.com/v1beta/openai/
gemini-2.5-flashhttps://api.mistral.ai/v1
Automatic text cleanup after transcription
https://api.groq.com/openai/v1
https://api-inference.huggingface.co/models/openai/whisper-large-v3
https://api.assemblyai.com/v2/transcript
Set up endpoints for Speech-to-Text backends.
POST /v1/audio/transcriptions.
Default language and preferred STT backend for captures.
https://api.elevenlabs.io/v1/text-to-speech
https://api.fish.audio/v1/tts
Controls how previews play and how OpenAI-compatible requests are shaped.
Pre-selected voice in the STT-TTS panel