Skip to content
Dashboard

S1

Fish Audio S1 is trained on over 2 million hours of audio with online RLHF (GRPO). It achieves 0.8% WER and 0.4% CER on Seed TTS Eval. S1 supports open-domain emotion, tone, and special effect markers. Your use is subject to Fish Audio's Terms & Privacy Policies.

import { experimental_generateSpeech as generateSpeech } from 'ai';
import { gateway } from '@ai-sdk/gateway';
import { writeFile } from 'node:fs/promises';
const result = await generateSpeech({
model: gateway.speechModel('fish-audio/s1'),
text: 'Hello from the Vercel AI Gateway!',
// Browse voices at https://fish.audio/app/discovery
// Open a voice, then use "Copy Model Id" in its "..." menu.
voice: '933563129e564b19a115bedd57b7406a',
});
await writeFile('speech.mp3', result.audio.uint8Array);
Read docs

More models by Fish Audio

Model
Context
Latency
Throughput
Input
Output
Cache
Web Search
Capabilities
Providers
ZDR
No Training
Release Date
$15/M chars
fish-audio logo
07/28/2026
$15/M chars
fish-audio logo
03/09/2026
$0.36/hr
fish-audio logo
03/01/2026