SandBase is live — $1 in free credits on signupStart free ›

PixVerse modelsvideo generation api

pixverse/sound-effects

Sound Effects is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.

Input
Description of the sound effect to generate. If empty, a random sound effect will be generated
URL of the input video to add sound effects to
Whether to keep the original audio from the video
Idle

Example output — click Run to generate your own

API README

PixVerse

PixVerse is the PixVerse route for video sound design, turning a picture-locked source clip into a production-ready result while keeping the operation distinct from neighboring endpoints. Sound Effects is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification. The workflow is designed for creators who need the model’s specific transformation to remain visible in the request, so the source, intended change, and finished artifact can be reviewed as one coherent creative decision.

In practical use, this route exposes prompt, sync_mode, video_url, safety_tolerance, original_sound_switch to shape the exact deliverable. Those controls let a team preserve the important source constraints, state subject behavior or material treatment precisely, choose supported timing or output characteristics, and reproduce successful settings across alternate takes. The result fits an iterative pipeline: establish the core brief, compare controlled variations, then pass the selected asset into editorial, design, localization, visualization, or publishing work.

Highlights

  • Visual-to-audio timing. Generate sound effects that follow the action and timing of a supplied video.
  • Prompt-guided soundscape. Describe the desired ambience, impacts, textures, and emphasis without rebuilding the picture edit.
  • Picture-locked workflow. Keep the existing video as the timing reference while producing a matching audio layer.
  • Post-production handoff. Return generated audio for mixing, editorial review, or delivery alongside the source clip.

Pricing

ConfigurationPrice
Per request$0.100000

When to Use

ScenarioWhy it fits
Exact workflow fitChoose this route when the required deliverable is video sound design, rather than a related route with different source media.
Directed creative iterationUse it when subject, motion, material, speech, framing, or finish should be expressed explicitly and compared across controlled variants.
Existing-asset continuityUse it when supplied images, video, audio, references, or styles must remain the anchor for the generated result.
Repeatable productionUse it when successful inputs need to be saved and rerun across a campaign, asset set, localization pass, or batch.
Pipeline handoffUse it when the returned artifact will move into editorial, compositing, visualization, review, storage, or publishing.

Prompt Guide

Start with the source or subject, then describe the intended transformation, movement or behavior, camera and composition, and the desired finish. Keep media URLs reachable, use only fields documented for this exact route, and change one major control at a time when comparing results. For source-led tasks, describe what should change as well as what must remain recognizable.

{
  "prompt": ""
}

Technical Specs

PropertyValue
Model IDpixverse/sound-effects
Execution modeasync
Required inputsprompt
promptstring; default:
sync_modeboolean; default: false
video_urlstring
safety_tolerancestring; default: 6
original_sound_switchboolean; default: false

Related Models

  • pixverse/extend
  • pixverse/lipsync
  • pixverse/swap
  • pixverse/c1/image-to-video

Related Models

pixverse/c1/image-to-videoC1 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.pixverse/c1/reference-to-videoC1 is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.pixverse/c1/text-to-videoC1 by PixVerse - generate cinematic videos from text descriptions with AI. Create high-quality video content with natural motion, camera control, and optional audio generation.pixverse/c1/transitionC1 Transition is PixVerse's image-to-video AI model. Bring static images to life with fluid animation, consistent character motion, and professional-grade video output.pixverse/extendExtend by PixVerse - AI-powered video editing and transformation. Apply style transfer, motion control, lip-sync, and visual effects to existing videos with natural language instructions.pixverse/extend/fastExtend Fast is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.pixverse/lipsyncLipsync is PixVerse's video-to-video AI model. Transform, enhance, and edit video content using text prompts - from style changes to object manipulation and scene modification.pixverse/swapSwap by PixVerse - animate still images into dynamic videos with AI. Transform photos into cinematic clips with natural motion, camera movement, and optional audio generation.