Fish Audio
Products
Products
Product overview
Explore the all-in-one creative suite
Voice library
Voices for any role or character
Create
Text to Speech
Generate human-like AI speech
Multi-speaker Dialogue
Create dialogue audio from multi-character scripts
Speech to Text
Transcribe audio and video
Voice Design
Generate custom voices
Voice Changer
Output audio in any voice
Voice Isolation
Extract clear speech
Voice Cloning
Clone your voice
Sound Effects
Coming soon
Generate any sound
Dubbing
Coming soon
Localize audio content
Music
Coming soon
Turn ideas into songs
Images
Generate images from text
Video
Generate video from text or images
S2
Fish Audio S2.1 Pro
Multi-speaker, multi-turn generation with natural language control over voice performance.
API
Platform
OverviewDocsAPI referenceAPI keysAPI pricingAPI Playground
API
Text to Speech
Generate speech through the API
Music
Coming soon
Create songs through the API
Speech to Text
Batch transcribe speech
Sound Effects
Coming soon
Generate sound effects through the API
Real-time Speech to Text
Transcribe speech in real time
Voice Cloning
Clone voices for TTS
Speech Engine
Coming soon
Give agents voice capabilities
Agents
Coming soon
Deploy voice agents in minutes
Dubbing
Coming soon
Translate video and audio through the API
API
Fish Audio API quick start
Debug voice generation, transcription, and account keys online
Text to Speech
Convert text to natural speech with Fish Audio, MiniMax, Qwen, and more
Speech to Text
High-accuracy transcription from uploaded audio
Voice Cloning
Clone your voice in about a minute from short samples
Voice Gallery
Browse public models and pick a reference voice
AI Image
Generate images from prompts with leading models
AI Video
Create video from text descriptions and styles
Lip-sync & digital human
Align speech to video for avatars and presenters
Voice Workspace
Voice synthesis workspace to create and manage your voice projects
Short video & dubbing
Fast voiceover for social, ads, and UGC
Audiobooks & podcasts
Long-form narration with natural pacing
Education & training
Clear narration for courses and internal comms
Company
AboutBlog
Resources
Coze
Tavo
SillyTavern
Dify
Open WebUI
AnythingLLM
Home Assistant
n8n
Affiliate program
Coming soon
API Playground
Try REST endpoints online with your API key
API keys
Create and manage API keys in your account
Pricing
Home/Blog

Blog

Product updates, model guides, and practical workflows for AI voice.

AI voiceover tutorial2026-07-03

How to make AI voiceover for video with Fish Audio

Choose a calm female voice, paste your script, generate natural AI voiceover, and use the audio in Reels, Shorts, TikTok, or CapCut.

Research2026-07-01

Fish Audio S2.1 Pro: high-quality AI text to speech and voice cloning

A next-generation multilingual TTS model for creators and developers.

TTS Model Guide2026-03-19

Xiaomi MiMo-V2-TTS: Text-Driven Expressive TTS

From free-form style instructions to non-verbal events and singing capability: MiMo-V2-TTS brings “expression” into speech generation.

Prompt Templates2026-03-19

MiMo-V2-TTS Style Prompts

Write emotion, pacing, pauses, voice texture, and performance intent in 1-2 sentences-copyable templates for voice-over workflows.

Non-verbal events2026-03-19

MiMo-V2-TTS Non-Verbal Events

Pauses, breaths, sighs, coughs, and laughter often decide whether it sounds robotic or performed.

Singing2026-03-19

MiMo-V2-TTS Singing Capability & Prompts

Spell out singing intent-rhythm, emotional intensity, and line endings can drastically change the result.

Tutorial2026-03-13

How to Use Fish Audio API for Audiobook Production

Fish Audio is one of the most natural AI voice models available. This guide walks through producing a complete audiobook using Fish Audio, which is built on the Fish Audio API.

Creator Guide2026-03-13

Fish Audio Voice Cloning for YouTube Creators

Clone your voice once, generate voiceovers for every video — no recording sessions needed. This guide covers how to build an efficient YouTube voiceover workflow using Fish Audio, powered by Fish Audio API.

Product Launch2026-03-02

Fish Audio S2 Model Launch: AI Voice Enters the 2.0 Era

More natural emotions, finer control, lower latency — Fish Audio S2 opens a new chapter in AI voice

Comparison2025-03-14

Best ElevenLabs Alternatives in 2025

ElevenLabs is popular — but it's not the only option. Here are the top AI voice tools worth considering, and why Fish Audio stands out.

Use Cases2025-03-14

Voice Cloning for Audiobooks: Produce Professional Narration with AI

AI voice cloning is transforming audiobook production — here's how to use it to create natural, engaging narration at a fraction of the traditional cost.

Use Cases2025-03-14

AI Dubbing: How to Dub Videos into Any Language with AI

Traditional dubbing is expensive and slow. AI dubbing tools let you localize video content into dozens of languages in hours, not weeks.

Product

  • Text to Speech
  • Multi-speaker Dialogue
  • Speech to Text
  • Voice Design
  • Voice Changer
  • Voice Isolation
  • Voice Cloning
  • Sound EffectsComing soon
  • DubbingComing soon
  • MusicComing soon
  • Images
  • Video

Solutions

  • Short video & dubbing
  • Audiobooks & podcasts
  • Education & training

Research

  • Fish Audio S1
  • Fish Audio S2 Pro
  • Fish Audio S2.1 Pro

Resources

  • Docs
  • API reference
  • Model library
  • Voice clone tutorial
  • Product comparison

Company

  • About
  • Blog
Kitta AudioPowered by Fish Audio
© 2026 Kitta AI. All rights reserved.|Privacy Policy|Terms of Service|Report abuse|support@fishaudio.org
support@fishaudio.org

Kitta Audio is independently operated by Kitta AI, not the official Fish Audio service.