- OpenAI
- OpenRouter
- ElevenLabs
- Groq
1
Set your API key
Store your Fish Audio API key
as an environment variable:
2
Swap the base URL and generate
From a framework
Frameworks that speak the OpenAI protocol work the same way: point them at/compat/v1:
Models
transcribe-1 needs no language hint. The language field in a transcription
response echoes what you sent: empty when you sent none, never a detection
result (the ElevenLabs protocol
echoes language_code the same way).
For speaker turns, long recordings, and inline emotion cues, use
transcribe-1-pro on the native
Speech to Text API.
Voices
The examples above use the model’s default voice. To pick a specific one, pass a voice ID: browse the Voice Library and copy the id of any voice, or make your own with Voice Cloning or Voice Design. Vendor preset names (nova, echo,
Rachel, …) don’t exist here.

