Skip to content

Speech Synthesis & Transcription

Speech Synthesis & Transcription

AMC WebUI features expressively rich voice generation and speech-to-text recognition.


1. Text-to-Speech (TTS)

  • Default Model: gemini-3.1-flash-tts-preview.
  • 30 Character Voices:
    • Broad palette of expressive timbres, cadences, and styles (including Zephyr, Aoede, Charon, Fenrir, Kore, Puck, and Orpheus).
    • Preview voice samples and configure your default voice under Settings -> Language & Voice.
  • Message Readout: Click the speaker icon on any message bubble to synthesize and stream natural voice audio with an interactive media player.

2. Speech-to-Text Transcription

  • Default Model: gemini-3.5-transcribe.
  • Voice Typing: Click the microphone icon to dictate messages. Audio is converted into structured text with high accuracy across technical jargon and multi-lingual accents.