Skip to main content
Admin › Settings › Audio
Full Audio settings tab

Full Audio settings tab

From the top, the screen is laid out as STT Settings → TTS Settings → Avatar.

STT (Speech-to-Text)

Configure the engine that converts user speech input into text.
Transcribes with the built-in faster-whisper model on the server. This is the default and needs no external API key.

TTS (Text-to-Speech)

Configure the engine that converts AI responses into speech.
Uses the browser’s built-in Web Speech API. Only a TTS Voice needs to be chosen; no key is required.

Avatar

Configure the avatar shown alongside AI responses. Selecting Azure AI Speech reveals the following fields.