docs.emix.ai
language
language
  • 🇺🇸 English
  • 🇨🇳 Chinese
language
language
  • 🇺🇸 English
  • 🇨🇳 Chinese
Market
File Upload API
Market
File Upload API
  1. Music Generation
  • Getting Started with Emix API (Important)
  • Market
  • Image Models
    • Topaz
      • Topaz - Image Upscale
    • Seedream
      • Seedream3.0 - Text to Image
      • Seedream4.0 - Text to Image
      • Seedream4.0 - Edit
      • Seedream4.5 - Text to Image
      • Seedream4.5 - Edit
      • Seedream5.0 Lite - Text to Image
      • Seedream5.0 Lite - Image to Image
      • Seedream5.0 Pro - Text to Image
      • Seedream5.0 Pro - Image to Image
      • Seedream 5.0 Pro - Layer Decomposition
    • Z-image
      • Z-Image
    • Google
      • Google - imagen4-fast
      • Google - imagen4-ultra
      • Google - imagen4
      • Google - Nano Banana
      • Google - Nano Banana Pro
      • Google - Nano Banana 2
      • Google - Nano Banana 2 Lite
    • Flux-2
      • Flux-2 - Pro Image to Image
      • Flux-2 - Pro Text to Image
      • Flux-2 - Image to Image
      • Flux-2 - Text to Image
    • Grok Imagine
      • Grok Imagine - Text to Image
      • Grok Imagine - image to image
      • Grok Imagine Image 2.0 Text To Image
      • Grok Imagine Image 2.0 Segment Map
      • Grok Imagine Image 2.0 Image Edit
    • GPT Image
      • GPT Image 2.5 Sunburst Text to Image
      • GPT Image 2.5 Sunburst Image to Image
      • GPT Image 2.5 Flare Text to Image
      • GPT Image 2.5 Flare Image to Image
      • GPT Image-1.5 - Text to Image
      • GPT Image-1.5 - Image to Image
      • GPT Image-2 - Text to Image
      • GPT Image 2 - Image To Image
    • Recraft
      • Recraft - Remove Background
      • Recraft - Crisp Upscale
    • Ideogram
      • Ideogram - V3 Reframe
      • Ideogram - Character Edit
      • Ideogram - Character Remix
      • Ideogram - Character
      • Ideogram V3 Text to Image
      • Ideogram V3 Edit
      • Ideogram V3 Remix
    • 4o Image API
      • 4o Image API Quickstart
      • 4o Image Generation Callbacks
      • Generate 4o Image
      • Get 4o Image Details
      • Get Direct Download URL
    • Flux Kontext API
      • Flux Kontext API Quickstart
      • Image Generation or Editing Callbacks
      • Generate or Edit Image
      • Get Image Details
    • Qwen
      • Qwen - Text to Image
      • Qwen - Image to Image
      • Qwen - Image Edit
      • Qwen2 - Image Edit
      • Qwen2 - Text To Image
      • Qwen3 Pro Text to Image
      • Qwen3 Text to Image
      • Qwen3 Pro Image to Image
      • Qwen3 Image to Image
      • Qwen 2.1 - Text to Image
      • Qwen 2.1 - Image to Image
    • Wan
      • Wan 2.7 Image
      • Wan 2.7 Image Pro
  • Video Models
    • OmniHuman
      • Omnihuman 1.5
      • Omnihuman 1.5 Human Identification
      • OmniHuman 1.5 Subject Detection
    • Hailuo
      • Hailuo 2.3 Pro Image to Video
      • Hailuo 2.3 Standard Image to Video
      • Hailuo Pro Text to Video
      • Hailuo Pro Image to Video
      • Hailuo Standard Text to Video
      • Hailuo Standard Image to Video
    • Topaz
      • Topaz - Video Upscale
    • Runway API
      • Runway Image To Video
      • Runway Text To Video
      • Runway Video Extension
    • HappyHorse
      • HappyHorse - text-to-video
      • HappyHorse - image-to-video
      • HappyHorse - reference-to-video
      • HappyHorse - video-edit
      • HappyHorse-1-1 image-to-video
      • HappyHorse-1-1 text-to-video
      • HappyHorse-1-1 reference-to-video
    • Wan
      • Wan - 2.2 A14B Image to Video Turbo
      • Wan - 2.2 A14B Speech to Video Turbo
      • Wan - 2.2 A14B Text to Video Turbo
      • Wan - Animate Move
      • Wan - Animate Replace
      • Wan 2.6 - Image to Video
      • Wan 2.6 - Text to Video
      • Wan 2.6 - Video to Video
      • Wan 2.5 - Image to Video
      • Wan 2.5 - Text to Video
      • Wan 2.7 - Text to Video
      • Wan 2.7 - Image to Video
      • Wan 2.7 - Video Edit
      • Wan 2.7 - Reference to Video
      • Wan 3.0 - Video
      • Wan 3.0 - Video Prime
    • Grok Imagine
      • Grok Imagine Text to Video
      • Grok Imagine Image to Video
      • Grok Imagine - Video Upscale
      • Grok Imagine - Video Extend
      • Grok Imagine Video 1.5 Preview
    • Kling
      • Kling 2.6 Text to Video
      • Kling 2.6 Image to Video
      • Kling - V2.5 Turbo Image to Video Pro
      • Kling - V2.5 Turbo Text to Video Pro
      • Kling AI Avatar Standard
      • Kling AI Avatar Pro
      • Kling V2.1 Master Image to Video
      • Kling V2.1 Master Text to Video
      • Kling V2.1 Pro
      • Kling V2.1 Standard
      • Kling 2.6 motion-control
      • Kling-3.0 motion-control
      • Kling 3.0
      • Kling - V3 Turbo Text to Video
      • Kling - V3 Turbo Image to Video
      • Kling 3.0 Omni Reference To Video
      • Kling 3.0 Omni Transformation
      • Kling 3.0 Omni Image To Video
      • Kling 3.0 Omni Text to Video
    • Bytedance
      • Bytedance Seedance 2.0
      • Bytedance Seedance 2.0 Fast
      • Bytedance Seedance 2.0 Mini
      • Bytedance Seedance 1.5 Pro
      • Bytedance V1 Pro Fast Image to Video
      • Bytedance V1 Pro Image to Video
      • Bytedance - V1 Pro Text to Video
      • Bytedance - V1 Lite Image to Video
      • Bytedance - V1 Lite Text to Video
      • Bytedance Seedance 2.5
    • Veo3.1 API
      • Get 4K Video Callbacks
      • Veo3.1 API Quickstart
      • Veo3.1 Video Generation Callbacks
      • Generate Veo3.1 Video
      • Get Veo3.1 Video Details
      • Get 1080P Video
      • Get 4K Video
      • VEO 3.1 Extend Video
      • VEO 3.1 Text to video
      • VEO 3.1 Image to video
      • VEO 3.1 Reference to vidoe
    • Gemini Omni
      • Gemini Omni 1.1 Flash
      • Gemini Omni Video
      • Gemini Omni Audio
      • Gemini Omni Character
    • Volcengine
      • Volcengine video to video lip sync
    • PixVerse
      • PixVerse V6 Text-to-Video
      • PixVerse V6 Image-to-Video
      • PixVerse V6 First & Last Frame Transition
      • PixVerse V6 Video Extension
      • PixVerse V6 Fusion / Reference-to-Video
    • MiniMax H3
      • MiniMax H3 Text-to-Video
      • MiniMax H3 Image-to-Video
      • MiniMax H3 Reference-to-Video
  • Music Models
    • ElevenLabs
      • elevenlabs/audio-isolation
      • elevenlabs/sound-effect-v2
      • elevenlabs/speech-to-text
      • elevenlabs/text-to-dialogue-v3
      • elevenlabs/text-to-speech-multilingual-v2
      • elevenlabs/text-to-speech-turbo-2-5
    • Suno API
      • Music Generation
        • Music Generation Callbacks
        • Music Extension Callbacks
        • Add Instrumental Callbacks
        • Add Vocals Callbacks
        • Music Cover Generation Callbacks
        • Replace Music Section Callbacks
        • Audio Upload and Extension Callbacks
        • Audio Upload and Cover Callbacks
        • Generate Music
          POST
        • Extend Music
          POST
        • Upload And Cover Audio
          POST
        • Upload And Extend Audio
          POST
        • Add Instrumental to Music
          POST
        • Add Vocals to Music
          POST
        • Get Timestamped Lyrics
          POST
        • Boost Music Style
          POST
        • Generate Music Cover
          POST
        • Replace Music Section
          POST
        • Generate Persona
          POST
        • Generate Mashup Music
          POST
      • Lyrics Generation
        • Lyrics Generation Callbacks
        • Generate Lyrics
      • WAV Conversion
        • Convert to WAV Format
      • Vocal Removal
        • Audio Separation Callbacks
        • Vocal & Instrument Stem Separation
        • Generate MIDI from Audio
      • Music Video Generation
        • Music Video Generation Callbacks
        • Create Music Video
      • Sounds Generation
        • Generate sounds
      • voice
        • Suno Voice Generation Callback
        • Suno Voice Validation Phrase Callback
        • Suno Voice Generate Verification Phrase API
        • Suno Voice Create Custom Voice API
        • Suno Voice Regenerate Verification Phrase
        • Suno Voice Check Availability API
    • Gemini
      • Gemini 3.1 Flash Text to speech
      • Gemini 2.5 Pro Text to Speech
  • Chat Models
    • Grok
      • Grok 4.5
      • Grok 4.3
      • Grok 4.6
    • Codex
      • GPT Codex
    • GPT
      • GPT 5.2
      • GPT 5.4 (response)
      • GPT 5.6 Luna
      • GPT 5.6 Terra
      • GPT 5.6 Sol
      • GPT 5.5 (response)
    • Claude
      • Claude Code + emix.ai Integration Guide
      • Claude Sonnet 5
      • Claude Sonnet 4.5
      • Claude Opus 4.7
      • Claude Opus 4.8
      • Claude Opus 5
      • Claude Fable 5
      • Claude Haiku 4.5
      • Claude Opus 4.5
      • Claude Opus 4.6
      • Claude Sonnet 4.5
      • Claude Sonnet 4.6
    • Gemini
      • Gemini 3.6 Flash
      • Gemini 3.6 Flash (openai)
      • Gemini 3.7 Flash (openai)
      • Gemini 2.5 Pro (openai)
      • Gemini 3 Pro (openai)
      • Gemini 3.1 Pro (openai)
      • Gemini 2.5 Flash (openai)
      • Gemini 3 Flash (openai)
      • Gemini 3.5 Flash
      • Gemini 3.5 Flash (openai)
      • Gemini 3 Flash
      • Gemini 3.7 Flash
      • Gemini 3.8 Flash
      • Gemini 3.8 Flash (openai)
  • Get Task Details
    GET
  1. Music Generation

Add Instrumental to Music

POST
/api/v1/generate/add-instrumental
Generate instrumental accompaniment based on uploaded audio files. This interface allows you to upload audio files and add instrumental tracks to them.

Usage Guide#

Use this interface to add instrumental tracks to existing audio
With source audio (uploadUrl), generation can proceed; title, negativeTags, and tags are required; other parameters are optional
Supports generation of various music style accompaniments
Allows customization of style, exclusion of specific elements, etc.

Parameter Details#

Always required: uploadUrl, title, negativeTags, tags
uploadUrl specifies the audio file URL to be processed
model specifies the AI model version
title specifies the title for the generated music
tags and negativeTags are used to control music style
lyrics is optional lyrics content (V6 maximum 5000 characters)
variety controls result diversity (integer 0–4, default 1)
Supports various optional parameters for fine-tuning generation effects

Developer Notes#

Generated files will be retained for 14 days
Callback process has three stages: text (text generation), first (first track completed), complete (all completed)

Callbacks

audioGenerated

Request

Authorization
Bearer Token
Provide your bearer token in the
Authorization
header when making requests to protected resources.
Example:
Authorization: Bearer ********************
or
Body Params application/jsonRequired

Examples

Responses

🟢200
application/json
Request successful
Bodyapplication/json

Request Request Example
Shell
JavaScript
Java
Swift
curl --location 'https://api.emix.ai/api/v1/generate/add-instrumental' \
--header 'Authorization: Bearer <token>' \
--header 'Content-Type: application/json' \
--data '{
    "model": "suno/add-instrumental",
    "callBackUrl": "https://example.com/callback",
    "input": {
        "uploadUrl": "https://example.com/music.mp3",
        "model": "V6",
        "title": "Relaxing Piano",
        "negativeTags": "heavy metal, fast drums",
        "tags": "relaxing, piano, soothing",
        "lyrics": "[Verse]\nKeep the melody flowing\nInto a gentle bridge\n[Chorus]\nLet the piano rise and fall",
        "vocalGender": "m",
        "styleWeight": 0.61,
        "weirdnessConstraint": 0.72,
        "audioWeight": 0.65,
        "variety": 1
    }
}'
Response Response Example
{
    "msg": "Playground task completed successfully.",
    "code": 200,
    "data": {
        "resultJson": "{\"sunoData\":[{\"duration\":172.16,\"stream_audio_url\":\"https://cdn1.suno.ai/f54b75dc-6dab-41d5-a947-bc1433d4040e.mp3\",\"image_suno_watermark\":false,\"model_name\":\"chirp-fenix\",\"image_url\":\"https://tempfile.aiquickdraw.com/r/83c4a74510eb4445b65a3633f99ea24e.jpeg\",\"audio_url\":\"https://tempfile.aiquickdraw.com/r/c1333b5343de474e9bbfaa10ec40c9be.mp3\",\"id\":\"f54b75dc-6dab-41d5-a947-bc1433d4040e\",\"title\":\"test2\",\"prompt\":\"\",\"tags\":\"pop\"},{\"duration\":171.4,\"stream_audio_url\":\"https://cdn1.suno.ai/f6d5cfe8-64f4-4376-ae62-514c5076bd87.mp3\",\"image_suno_watermark\":false,\"model_name\":\"chirp-fenix\",\"image_url\":\"https://tempfile.aiquickdraw.com/r/72a3ab41bbb941dd9c0dfe93e3abb3c6.jpeg\",\"audio_url\":\"https://tempfile.aiquickdraw.com/r/dcfbbf9a392b407393d15788af5a8ac5.mp3\",\"id\":\"f6d5cfe8-64f4-4376-ae62-514c5076bd87\",\"title\":\"test2\",\"prompt\":\"\",\"tags\":\"pop\"}]}",
        "param": "{\"model\":\"suno/add-instrumental\",\"callBackUrl\":\"http://localhost:8305/test\",\"input\":\"{\\\"model\\\":\\\"V5_5\\\",\\\"uploadUrl\\\":\\\"https://tempfileb.aiquickdraw.com/kieai/market/1786500031698_0JQTSDn6.mp3\\\",\\\"title\\\":\\\"test2\\\",\\\"negativeTags\\\":\\\"noise\\\",\\\"tags\\\":\\\"pop\\\",\\\"vocalGender\\\":\\\"f\\\",\\\"styleWeight\\\":0.65,\\\"weirdnessConstraint\\\":0.65,\\\"audioWeight\\\":0.65}\"}",
        "costTime": 252,
        "createTime": 1786500268000,
        "completeTime": 1786500528000,
        "model": "suno/add-instrumental",
        "updateTime": 1786500528000,
        "state": "success",
        "creditsConsumed": 12,
        "taskId": "8bfdf5fba96719572196a689b0133137"
    }
}
Previous
Upload And Extend Audio
Next
Add Vocals to Music
Built with