Fish Audio
Products
Products
Product overview
Explore the all-in-one creative suite
Voice library
Voices for any role or character
Create
Text to Speech
Generate human-like AI speech
Multi-speaker Dialogue
Create dialogue audio from multi-character scripts
Speech to Text
Transcribe audio and video
Voice Design
Generate custom voices
Voice Changer
Output audio in any voice
Voice Isolation
Extract clear speech
Voice Cloning
Clone your voice
Sound Effects
Coming soon
Generate any sound
Dubbing
Coming soon
Localize audio content
Music
Coming soon
Turn ideas into songs
Images
Generate images from text
Video
Generate video from text or images
S2
Fish Audio S2.1 Pro
Multi-speaker, multi-turn generation with natural language control over voice performance.
API
Platform
OverviewDocsAPI referenceAPI keysAPI pricingAPI Playground
API
Text to Speech
Generate speech through the API
Music
Coming soon
Create songs through the API
Speech to Text
Batch transcribe speech
Sound Effects
Coming soon
Generate sound effects through the API
Real-time Speech to Text
Transcribe speech in real time
Voice Cloning
Clone voices for TTS
Speech Engine
Coming soon
Give agents voice capabilities
Agents
Coming soon
Deploy voice agents in minutes
Dubbing
Coming soon
Translate video and audio through the API
API
Fish Audio API quick start
Debug voice generation, transcription, and account keys online
Text to Speech
Convert text to natural speech with Fish Audio, MiniMax, Qwen, and more
Speech to Text
High-accuracy transcription from uploaded audio
Voice Cloning
Clone your voice in about a minute from short samples
Voice Gallery
Browse public models and pick a reference voice
AI Image
Generate images from prompts with leading models
AI Video
Create video from text descriptions and styles
Lip-sync & digital human
Align speech to video for avatars and presenters
Voice Workspace
Voice synthesis workspace to create and manage your voice projects
Short video & dubbing
Fast voiceover for social, ads, and UGC
Audiobooks & podcasts
Long-form narration with natural pacing
Education & training
Clear narration for courses and internal comms
Company
AboutBlog
Resources
Coze
Tavo
SillyTavern
Dify
Open WebUI
AnythingLLM
Home Assistant
n8n
Affiliate program
Coming soon
API Playground
Try REST endpoints online with your API key
API keys
Create and manage API keys in your account
Pricing
Prompt Templates2026-03-19·4 min read

MiMo-V2-TTS Style Prompts

Write emotion, pacing, pauses, voice texture, and performance intent in 1-2 sentences-copyable templates for voice-over workflows.

Back to MiMo-V2-TTS Overview ->Try TTS Tool ->

Writing Rules (Fast and Reliable)

✓

Write intent, not tags: 'holding back anger', 'trying not to cry', 'pretending to be calm'.

✓

Define rhythm first: pace, sentence length, pauses, and how to end each line.

✓

Then define texture: breathiness, hoarseness, brightness, close/whisper vs projecting.

✓

Optionally add non-verbal cues: sighs, breaths, soft laugh, cough.

Copyable Examples

Reusable template (recommended)

Emotion + pacing/pauses + intensity; voice texture (breathy/hoarse/bright/close) + how to end each sentence.

"Angry but holding back"

Angry but trying to stay calm. Slightly faster pace, short sentences, micro-pauses mid-sentence. Keep the ending restrained, no outburst.

"Exhausted comfort"

Extremely tired yet gentle. Slow pacing, heavier breath, slightly hoarse as if after a long day.

"Whisper"

Almost whispering, softer volume, slow pacing. A hint of warmth and a subtle smile, not exaggerated.

"Formal announcement"

Formal, clear, professional. Medium pace, crisp articulation. Avoid casual phrasing and keep a stable tone.

Next Reads

Non-verbal Events (pauses/breath/laughter) ->Singing Capability & Use Cases ->API Playground ->

Product

  • Text to Speech
  • Multi-speaker Dialogue
  • Speech to Text
  • Voice Design
  • Voice Changer
  • Voice Isolation
  • Voice Cloning
  • Sound EffectsComing soon
  • DubbingComing soon
  • MusicComing soon
  • Images
  • Video

Solutions

  • Short video & dubbing
  • Audiobooks & podcasts
  • Education & training

Research

  • Fish Audio S1
  • Fish Audio S2 Pro
  • Fish Audio S2.1 Pro

Resources

  • Docs
  • API reference
  • Model library
  • Voice clone tutorial
  • Product comparison

Company

  • About
  • Blog
Kitta AudioPowered by Fish Audio
© 2026 Kitta AI. All rights reserved.|Privacy Policy|Terms of Service|Report abuse|support@fishaudio.org
support@fishaudio.org

Kitta Audio is independently operated by Kitta AI, not the official Fish Audio service.