Kitta Audio
Products
Products
Product overview
Explore the all-in-one creative suite
Voice library
Voices for any role or character
Create
Text to Speech
Generate human-like AI speech
Multi-speaker Dialogue
Create dialogue audio from multi-character scripts
Speech to Text
Transcribe audio and video
Voice Design
Generate custom voices
Voice Changer
Output audio in any voice
Voice Isolation
Extract clear speech
Voice Cloning
Clone your voice
Sound Effects
Coming soon
Generate any sound
Dubbing
Coming soon
Localize audio content
Music
Coming soon
Turn ideas into songs
Images
Generate images from text
Video
Generate video from text or images
S2
Fish Audio S2.1 Pro
Multi-speaker, multi-turn generation with natural language control over voice performance.
API
Platform
OverviewDocsAPI referenceAPI keysAPI pricingAPI Playground
API
Text to Speech
Generate speech through the API
Music
Coming soon
Create songs through the API
Speech to Text
Batch transcribe speech
Sound Effects
Coming soon
Generate sound effects through the API
Real-time Speech to Text
Transcribe speech in real time
Voice Cloning
Clone voices for TTS
Speech Engine
Coming soon
Give agents voice capabilities
Agents
Coming soon
Deploy voice agents in minutes
Dubbing
Coming soon
Translate video and audio through the API
API
Kitta Audio API quick start
Debug voice generation, transcription, and account keys online
Text to Speech
Convert text to natural speech with Fish Audio, MiniMax, Qwen, and more
Speech to Text
High-accuracy transcription from uploaded audio
Voice Cloning
Clone your voice in about a minute from short samples
Voice Gallery
Browse public models and pick a reference voice
AI Image
Generate images from prompts with leading models
AI Video
Create video from text descriptions and styles
Lip-sync & digital human
Align speech to video for avatars and presenters
Voice Workspace
Voice synthesis workspace to create and manage your voice projects
Short video & dubbing
Fast voiceover for social, ads, and UGC
Audiobooks & podcasts
Long-form narration with natural pacing
Education & training
Clear narration for courses and internal comms
Company
AboutBlog
Resources
Coze
Tavo
SillyTavern
Dify
Open WebUI
AnythingLLM
Home Assistant
n8n
Affiliate program
Coming soon
API Playground
Try REST endpoints online with your API key
API keys
Create and manage API keys in your account
Pricing

Kitta Audio API key & playground

Create an API key in your account, check the authentication guide, then send a test request for text-to-speech, voice cloning, or speech recognition.

Use an API key created in your Kitta Audio account, including when calling Fish Audio models through this platform. Kitta Audio keys and Fish Audio's own API keys are separate and cannot be used interchangeably.

Create or manage API keyAPI key authenticationAPI quickstart
  • Text to Speech (HTTP)

    REST synthesis with your voice model ID and engine options.

    POST/api/open/v1/speech/tts
  • TTS WebSocket

    Streaming speech over WebSocket for realtime use cases.

    WS/v1/tts/live
  • TTS WebSocket v2

    Updated WebSocket protocol for TTS.

    WS/v2/tts/live
  • Async TTS — create job

    Create a long-running TTS job and poll for the result.

    POST/api/open/v1/speech/tts/jobs
  • Async TTS — list jobs

    List recent async TTS jobs for your API key.

    GET/api/open/v1/speech/tts/jobs
  • Async TTS — query job

    Fetch async TTS job status and output by task ID.

    GET/api/open/v1/speech/tts/jobs/{taskId}
  • Speech to Text

    Transcribe audio from a public URL.

    POST/api/open/v1/speech/transcriptions
  • Voice clone — create model

    Upload reference audio to create a voice model.

    POST/api/open/v1/voices
  • Voice clone — get model

    Fetch a single voice model by ID.

    GET/api/open/v1/voices/{voiceId}
  • Voice clone — delete model

    Remove a voice model by ID.

    DELETE/api/open/v1/voices/{voiceId}
  • Voice clone — list models

    List public and personal voice models.

    GET/api/open/v1/voices
  • Voice design — create

    Generate voice design preview candidates from a prompt.

    POST/api/open/v1/voice-designs
  • Voice design — save voice

    Save a design candidate as a permanent voice model.

    POST/api/open/v1/voice-designs/{designId}/voices
  • Voice design — preview audio

    Download authenticated preview audio for a candidate.

    GET/api/open/v1/voice-designs/{designId}/candidates/{candidateId}/audio
  • Lip sync — create task

    Create a lip-sync video generation task.

    POST/api/open/v1/media/lip-sync/jobs
  • Lip sync — query task

    Poll task status and results by ID.

    GET/api/open/v1/media/lip-sync/jobs/{jobId}
  • Lip sync — list tasks

    List lip-sync tasks and statistics.

    GET/api/open/v1/media/lip-sync/jobs
  • Media models

    List verified AI image and video models.

    GET/api/open/v1/media/models
  • Image — create job

    Create an AI image generation or edit job.

    POST/api/open/v1/media/images/jobs
  • Image — query job

    Fetch an AI image job by ID.

    GET/api/open/v1/media/images/jobs/{jobId}
  • Video — create job

    Create an AI video generation job.

    POST/api/open/v1/media/videos/jobs
  • Video — query job

    Fetch an AI video job by ID.

    GET/api/open/v1/media/videos/jobs/{jobId}
  • User profile (API)

    Remaining API quota and basic account info.

    GET/api/open/v1/profile
  • User profile — update

    Update Open API profile settings.

    POST/api/open/v1/profile