KIE.AI
All Model
  • All Model
  • Old Model
language
language
  • 🇺🇸 English
  • 🇨🇳 Chinese
language
language
  • 🇺🇸 English
  • 🇨🇳 Chinese
Support
All Model
  • All Model
  • Old Model
All Model
  • All Model
  • Old Model
Market
File Upload APICommon API
Market
File Upload APICommon API
  1. Music Generation
  • Getting Started with KIE API (Important)
  • Market
  • Image Models
    • Seedream
      • Seedream4.0 - Text to Image
      • Seedream4.0 - Edit
      • Seedream4.5 - Text to Image
      • Seedream4.5 - Edit
      • Seedream5.0 Lite - Text to Image
      • Seedream5.0 Lite - Image to Image
      • Seedream5.0 Pro - Text to Image
      • Seedream5.0 Pro - Image to Image
      • Seedream 5.0 Pro - Layer Decomposition
      • Seedream3.0 - Text to Image
    • Z-image
      • Z-Image
    • Google
      • Google - imagen4-fast
      • Google - imagen4-ultra
      • Google - imagen4
      • Google - Nano Banana Edit
      • Google - Nano Banana
      • Google - Nano Banana Pro
      • Google - Nano Banana 2
      • Google - Nano Banana 2 Lite
    • Flux-2
      • Flux-2 - Pro Image to Image
      • Flux-2 - Pro Text to Image
      • Flux-2 - Image to Image
      • Flux-2 - Text to Image
    • Grok Imagine
      • Grok Imagine Image 2.0 Text To Image
      • Grok Imagine Image 2.0 Segment Map
      • Grok Imagine Image 2.0 Segment Edit
      • Grok Imagine - Text to Image
      • Grok Imagine Image 2.0 Image Edit
      • Grok Imagine - image to image
    • GPT Image
      • GPT Image 2.5 Flare - Text to Image
      • GPT Image 2.5 Flare - Image To Image
      • GPT Image 2.5 Sunburst - Text to Image
      • GPT Image 2.5 Sunburst - Image To Image
      • GPT Image-1.5 - Text to Image
      • GPT Image-1.5 - Image to Image
      • GPT Image-2 - Text to Image
      • GPT Image 2 - Image To Image
    • Topaz
      • Topaz - Image Upscale
    • Recraft
      • Recraft - Remove Background
      • Recraft - Crisp Upscale
    • Ideogram
      • Ideogram - Character Edit
      • Ideogram - Character Remix
      • Ideogram - Character
      • Ideogram V3 Text to Image
      • Ideogram V3 Edit
      • Ideogram V3 Remix
    • Qwen
      • Qwen - Text to Image
      • Qwen - Image to Image
      • Qwen - Image Edit
      • Qwen2 - Image Edit
      • Qwen2 - Text To Image
      • Qwen 2.1 - Text to Image
      • Qwen 2.1 - Image to Image
      • Qwen3 Pro Text to Image
      • Qwen3 Text to Image
      • Qwen3 Pro Image to Image
      • Qwen3 Image to Image
    • Wan
      • Wan 2.7 Image
      • Wan 2.7 Image Pro
    • 4o Image API
      • 4o Image Generation Callbacks
      • Generate 4o Image
    • Flux Kontext API
      • Image Generation or Editing Callbacks
      • Generate or Edit Image
  • Video Models
    • Grok Imagine
      • Grok Imagine Text to Video
      • Grok Imagine Image to Video
      • Grok Imagine - Video Upscale
      • Grok Imagine - Video Extend
      • Grok Imagine Video 1.5 Preview
    • Kling
      • Kling 2.6 Text to Video
      • Kling 2.6 Image to Video
      • Kling - V2.5 Turbo Image to Video Pro
      • Kling - V2.5 Turbo Text to Video Pro
      • Kling AI Avatar Standard
      • Kling AI Avatar Pro
      • Kling V2.1 Master Image to Video
      • Kling V2.1 Master Text to Video
      • Kling V2.1 Pro
      • Kling V2.1 Standard
      • Kling 2.6 motion-control
      • Kling-3.0 motion-control
      • Kling 3.0
      • Kling - V3 Turbo Text to Video
      • Kling - V3 Turbo Image to Video
      • Kling 3.0 Omni Reference To Video
      • Kling 3.0 Omni Transformation
      • Kling 3.0 Omni Image To Video
      • Kling 3.0 Omni Text to Video
    • Bytedance
      • Bytedance Seedance 2.0
      • Bytedance Seedance 2.0 Fast
      • Bytedance Seedance 2.0 Mini
      • Bytedance Seedance 2.5
      • Bytedance Seedance 1.5 Pro
      • Bytedance V1 Pro Fast Image to Video
      • Bytedance V1 Pro Image to Video
      • Bytedance - V1 Pro Text to Video
      • Bytedance - V1 Lite Image to Video
      • Bytedance - V1 Lite Text to Video
    • Hailuo
      • Hailuo 2.3 Pro Image to Video
      • Hailuo 2.3 Standard Image to Video
      • Hailuo Pro Text to Video
      • Hailuo Pro Image to Video
      • Hailuo Standard Text to Video
      • Hailuo Standard Image to Video
    • Wan
      • Wan - 2.2 A14B Image to Video Turbo
      • Wan - 2.2 A14B Speech to Video Turbo
      • Wan - 2.2 A14B Text to Video Turbo
      • Wan - Animate Move
      • Wan - Animate Replace
      • Wan 2.6 - Image to Video
      • Wan 2.6 - Text to Video
      • Wan 2.6 - Video to Video
      • Wan - 2.6-flash-image-to-video
      • Wan - 2-6-flash-video-to-video
      • Wan 2.5 - Image to Video
      • Wan 2.5 - Text to Video
      • Wan 2.7 - Text to Video
      • Wan 2.7 - Image to Video
      • Wan 2.7 - Video Edit
      • Wan 2.7 - Reference to Video
      • Wan 3.0 - Video
      • Wan 3.0 - Video Prime
    • Topaz
      • Topaz - Video Upscale
    • Infinitalk
      • Infinitalk - From Audio
    • PixVerse
      • PixVerse V6 Text-to-Video
      • PixVerse V6 Image-to-Video
      • PixVerse V6 First & Last Frame Transition
      • PixVerse V6 Video Extension
      • PixVerse V6 Fusion / Reference-to-Video
    • MiniMax H3
      • MiniMax H3 Text-to-Video
      • MiniMax H3 Image-to-Video
      • MiniMax H3 Reference-to-Video
    • Runway API
      • AI Video Generation Callbacks
      • AI Video Extension Callbacks
      • Aleph
        • Aleph Video Generation Callbacks
        • Generate Aleph Video
      • Generate AI Video
      • Extend AI Video
    • HappyHorse
      • HappyHorse - text-to-video
      • HappyHorse - image-to-video
      • HappyHorse - reference-to-video
      • HappyHorse - video-edit
      • HappyHorse-1-1 image-to-video
      • HappyHorse-1-1 text-to-video
      • HappyHorse-1-1 reference-to-video
    • Gemini Omni
      • Gemini Omni 1.1 Flash
      • Gemini Omni Video
      • Gemini Omni Audio
      • Gemini Omni Character
    • OmniHuman
      • Omnihuman 1.5
      • Omnihuman 1.5 Human Identification
      • OmniHuman 1.5 Subject Detection
    • Volcengine
      • Volcengine video to video lip sync
    • Veo3.1 API
      • Veo3.1 Video Generation Callbacks
      • Get 4K Video Callbacks
      • Generate Veo3.1 Video
      • Get 1080P Video
      • Get 4K Video
      • Extend Veo3.1 Video
  • Music Models
    • ElevenLabs
      • elevenlabs/audio-isolation
      • elevenlabs/text-to-dialogue-v3
      • elevenlabs/text-to-speech-multilingual-v2
      • elevenlabs/text-to-speech-turbo-2-5
    • Gemini
      • Gemini 3.1 Flash Text to speech
      • Gemini 2.5 Pro Text to Speech
    • Suno
      • Music Generation
        • Music Cover Generation Callbacks
        • Music Generation Callbacks
        • Music Extension Callbacks
        • Audio Upload and Cover Callbacks
        • Audio Upload and Extension Callbacks
        • Add Instrumental Callbacks
        • Add Vocals Callbacks
        • Replace Music Section Callbacks
        • Generate Music
          POST
        • Extend Music
          POST
        • Upload And Cover Audio
          POST
        • Upload And Extend Audio
          POST
        • Add Instrumental to Music
          POST
        • Add Vocals to Music
          POST
        • Get Timestamped Lyrics
          POST
        • Boost Music Style
          POST
        • Generate Music Cover
          POST
        • Replace Music Section
          POST
        • Generate Persona
          POST
        • Generate Mashup Music
          POST
        • Recovery Audio
          POST
      • WAV Conversion
        • Convert to WAV Callbacks
        • Convert to WAV Format
      • Music Video Generation
        • Music Video Generation Callbacks
        • Create Music Video
      • Lyrics Generation
        • Lyrics Generation Callbacks
        • Generate Lyrics
      • voice
        • Suno Voice Generation Callback
        • Suno Voice Validation Phrase Callback
        • Suno Voice Generate Verification Phrase API
        • Suno Voice Create Custom Voice API
        • Suno Voice Regenerate Verification Phrase
        • Suno Voice Check Availability API
      • Vocal Removal
        • Audio Separation Callbacks
        • MIDI Generation Callbacks
        • Vocal & Instrument Stem Separation
        • Generate MIDI from Audio
      • Sounds Generation
        • Generate sounds
  • Chat Models
    • GPT
      • GPT 5.2
      • GPT 5.4 (response)
      • GPT 5.5 (response)
      • GPT 5.6 Luna
      • Gpt 6 Astra
      • GPT 5.6 Terra
      • GPT 5.6 Sol
    • Claude
      • Claude Code + kie.ai Integration Guide
      • Claude Opus 4.7
      • Claude Opus 4.8
      • Claude Fable 5
      • Claude Sonnet 5
      • Claude Haiku 4.5
      • Claude Opus 4.5
      • Claude Opus 4.6
      • Claude Opus 5
      • Claude Sonnet 4.5
      • Claude Sonnet 4.6
    • Codex
      • GPT Codex
    • Gemini
      • Gemini 2.5 Pro (openai)
      • Gemini 3 Pro (openai)
      • Gemini 3.1 Pro (openai)
      • Gemini 2.5 Flash (openai)
      • Gemini 3 Flash (openai)
      • Gemini 3.5 Flash
      • Gemini 3.5 Flash (openai)
      • Gemini 3.6 Flash
      • Gemini 3.6 Flash (openai)
      • Gemini 3.7 Flash
      • Gemini 3.7 Flash (openai)
      • Gemini 3.8 Flash
      • Gemini 3.8 Flash (openai)
      • Gemini 3 Flash
    • Grok
      • Grok 4.7
      • Grok 4.3
      • Grok 4.5
      • Grok 4.6
  • Get Task Details
    GET
  1. Music Generation

Upload And Cover Audio

POST
/api/v1/jobs/createTask
Document updated
If you have already completed the integration process previously, you can still access the old version document at: Old version address (https://docs.kie.ai/old-model/suno-api/upload-and-cover-audio)
This API creates a cover version of an audio track by transforming it into a new style while retaining its core melody. It incorporates Suno's upload capability, enabling users to upload an audio file for processing. The expected result is a refreshed audio track with a new style, keeping the original melody intact.

Parameter Usage Guide#

Character Limits
Character limits vary depending on the model version:
V6, V6_WILD, and V6_MINI: style (max 1000 chars), title (max 80 chars), lyrics / prompt (max 5000 chars)
Model V5_5 and V5: style (max 1000 chars), title (max 100 chars), prompt (max 5000 chars)
Models V4.5PLUS and V4.5: style (max 1000 chars), title (max 100 chars), prompt (max 5000 chars)
Model V4.5ALL: style (max 1000 chars), title (max 80 chars), prompt (max 5000 chars)
Model V4: style (max 200 chars), title (max 80 chars), prompt (max 3000 chars)
Do not pass image_urls. Cover does not support simple-mode reference images.
Always required: upload_url, instrumental, model
With source audio (upload_url), generation can proceed; lyrics (formerly used via prompt), title, and style are optional
lyrics takes priority over prompt as lyrics. If lyrics is omitted, prompt is used as lyrics when provided
If instrumental is true: generate instrumental music (no vocals). Do not pass lyrics, prompt, or vocal_gender
upload_url specifies the source audio file; ensure the uploaded audio does not exceed 8 minutes in length

Developer Notes#

1.
Generated files will be deleted after 15 days.
2.
Pay attention to character limits for lyrics / prompt, style, and title to ensure successful processing.
3.
Callback Process Stages: The callBackUrl process has three stages: text (text generation complete), first (first track complete), and complete (all tracks complete).
4.
Active Status Check: You can use the Get Music Generation Details endpoint to actively check the task status instead of waiting for callbacks.
5.
The upload_url parameter is used to specify the source audio file; please provide a valid URL.

Optional Parameters#

ParameterTypeDescription
lyricsstringLyrics for the covered audio. Optional. V6 maximum 5000 characters. Takes priority over prompt as lyrics.
promptstringUsed as lyrics only when lyrics is not provided. Optional. V6 maximum 5000 characters.
stylestringMusic style specification. Optional. V6 maximum 1000 characters.
titlestringTrack title. Optional. Maximum 80 characters.
negative_tagsstringMusic styles or traits to exclude from the generated audio.
vocal_genderstringVocal gender preference. Use m for male, f for female.
style_weightnumberStrength of adherence to style. Range 0–1, up to 2 decimal places. Example: 0.65.
weirdness_constraintnumberControls creative deviation. Range 0–1, up to 2 decimal places. Example: 0.65.
audio_weightnumberBalance weight for audio features. Range 0–1, up to 2 decimal places. Example: 0.65.
varietyintegerDiversity of generated results. Integer from 0–4, default 1. 0 off, 1 normal, 2 high, 3 extra, 4 max.
persona_idstringPersona ID or Voice ID to apply to the generated music. To create one, use Generate Persona.
durationintegerAudio duration in seconds. Range 10–360, default 20. Valid when model is V5_5, V6, V6_MINI, or V6_WILD.

Callbacks

audioGenerated

Request

Authorization
Bearer Token
Provide your bearer token in the
Authorization
header when making requests to protected resources.
Example:
Authorization: Bearer ********************
or
Body Params application/jsonRequired

Examples

Responses

🟢200
application/json
Request successful
Bodyapplication/json

Request Request Example
Shell
JavaScript
Java
Swift
curl --location 'https://api.kie.ai/api/v1/jobs/createTask' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data '{
  "model": "ai-music-api/upload-and-cover-audio",
  "callBackUrl": "https://api.example.com/callback",
  "input": {
    "upload_url": "https://storage.example.com/upload",
    "prompt": "A calm and relaxing piano track with soft melodies",
    "instrumental": true,
    "model": "V4",
    "style": "Classical",
    "title": "Peaceful Piano Meditation",
    "negative_tags": "Heavy Metal, Upbeat Drums",
    "vocal_gender": "m",
    "style_weight": 0.65,
    "weirdness_constraint": 0.65,
    "audio_weight": 0.65,
    "persona_id": "persona_123",
    "persona_model": "style_persona",
    "lyrics": "[Verse]\nKeep the melody flowing\nInto a gentle bridge\n[Chorus]\nLet the piano rise and fall",
    "variety": 1
  }
}'
Response Response Example
{
    "code": 200,
    "msg": "success",
    "data": {
        "taskId": "5c79****be8e"
    }
}
Previous
Extend Music
Next
Upload And Extend Audio
Built with