VidMage

VidMage is an AI video and image creation platform for creators. Turn your ideas and images into videos and fresh visuals.

Open the creator tools
Explore APIs
AI Video APIsAI Image APIsAI Audio APIsAI 3D APIsAI Face Swap APIsAI Effects APIs
Build with VidMage
QuickstartAPI keysVidMage MCPError reference
Your account
Developer consoleAbout VidMagePlans and creditsAPI billingContact support
© 2026 VidMage. All rights reserved.
Privacy PolicyTerms of ServiceReport Abuse
Skip to content
VidMage/Developers
Overview
APIs
All APIsAI Video APIsAI Image APIsAI Audio APIsAI 3D APIsAI Face Swap APIsAI Effects APIs
DocumentationVidMage MCPOpen console
Developers/Documentation/Lip Sync API documentation
Start here
DocumentationQuickstartAuthentication
Core workflows
File uploadsTask lifecycleCredits and billingErrors and recoveryVidMage MCP
Browse documentation
All API documentation
Video documentation 36AI Video Head SwapAI Video Face SwapAI Multiple Face Swap VideoAI Text to VideoAI Image to VideoAI Video to VideoAI Video ExtenderAI Video to Anime ConverterAI Video Background RemoverAI Video Watermark RemoverSora Link Watermark RemoverSora 2 Video GeneratorAI Photo DanceAI Motion ControlAI Talking PhotoAI Lip SyncAI Subtitle GeneratorAI Video UpscalerKling AI Video GeneratorPixVerse AI Video GeneratorHailuo AI Video GeneratorMiniMax H3 AI Video GeneratorGrok Video GeneratorSora 2 AI Video GeneratorWan AI Video GeneratorWan 3.0 AI Video GeneratorSeedance 2.0 AI Video GeneratorSeedance 2.5 AI Video GeneratorSeedance AI Video GeneratorMidjourney Video GeneratorVidu AI Video GeneratorVeo 3.1 AI Video GeneratorKling 3.0 AI Video GeneratorSkyReels AI Video GeneratorHappyHorse AI ModelRunway AI Video Generator
Image documentation 24AI Photo Face SwapAI Head SwapAI Multiple Face SwapAI Text to ImageAI Image to ImageAI Girl GeneratorAI Hairstyle ChangerAI Clothes ChangerAI Object RemoverAI Image Watermark RemoverAI Image UpscalerGPT Image 2 Image to ImageAI GIF Face SwapMidjourney AI Image GeneratorGrok AI Image GeneratorNano Banana AI Image GeneratorGPT Image GeneratorSeedream AI Image GeneratorZ-Image AI ModelWan Image GeneratorQwen Image GeneratorQwen 3.0 Image GeneratorGPT Image 2 GeneratorGPT Image 2.5 Generator
Audio documentation 3AI Voice CloneAI Voice DesignAI Text to Music
3d documentation 3AI Image to 3DAI Four-View to 3DAI Text to 3D

Lip Sync API documentation

Synchronize lip movements in a photo or video. Image + audio costs 15 credits/second (minimum 75), video + audio costs 25 credits/second (minimum 125), and video-to-video costs 2 credits/second (minimum 20). Audio modes support 2–15 seconds; video-to-video supports up to 20 seconds.

On this page01 / 07
01Endpoints02Parameters and input rules03Measured media limits04Example request05Task results and recovery06Credits07Integration guides

Endpoints

Contract itemValue
Capabilitylip-sync
SubmitPOST https://vidmage.ai/api/v1/lip-sync/submit
QueryPOST https://vidmage.ai/api/v1/lip-sync/query
Task identifiertaskId
Result fieldvideoUrl

Authenticate with Authorization: Bearer YOUR_API_KEY. Save the task identifier and query the same capability. Use a stable Idempotency-Key for submission retries.

Product guide: AI Lip Sync API for Photos and Videos

Parameters and input rules

FieldTypeRequiredDefaultMeaning and limits
visualUrlstringYesNot specifiedURL of the source image (JPG/JPEG/PNG/WebP, up to 30 MB) or video (MP4/MOV, up to 50 MB). Video duration limits depend on inputType.
inputTypestringNo"image"Type of source.
Values: "image", "video", "video-to-video"
audioUrlstringNoNot specifiedDriving-audio fileUrl returned by upload_files (MP3 or WAV, required unless inputType is video-to-video).
targetVideoUrlstringNoNot specifiedTarget-video fileUrl returned by upload_files (MP4/MOV; required when inputType is video-to-video).

audioUrl is required when inputType is not "video-to-video".

targetVideoUrl is required when inputType equals "video-to-video".

Measured media limits

InputRule
audioUrlThe server measures this media for billing. Maximum 15 seconds. Minimum 2 seconds. Do not send audioDuration in the submission body.
targetVideoUrlUsed when inputType is video-to-video.
inputType = video-to-videoMaximum 20 seconds.
inputType = video-to-videoMinimum 0.1 seconds.

Example request

{
  "visualUrl": "https://vidmage.ai/assets/images/samples/blue-eyed-woman-sunlight.webp",
  "inputType": "image",
  "audioUrl": "https://vidmage.ai/templates/voices/alice.mp3"
}
# Set VIDMAGE_IDEMPOTENCY_KEY to a unique value for this task; preserve it for transport retries.
curl -X POST "https://vidmage.ai/api/v1/lip-sync/submit" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Idempotency-Key: ${VIDMAGE_IDEMPOTENCY_KEY}" \
  -H "Content-Type: application/json" \
  -d '{"visualUrl":"https://vidmage.ai/assets/images/samples/blue-eyed-woman-sunlight.webp","inputType":"image","audioUrl":"https://vidmage.ai/templates/voices/alice.mp3"}'
# -> { "success": true, "taskId": "...", "creditsConsumed": ... }

Task results and recovery

URL of the lip-synced video. Read videoUrl from the completed query response. Preserve taskId while the task is running.

curl -X POST "https://vidmage.ai/api/v1/lip-sync/query" \
  -H "Authorization: Bearer YOUR_API_KEY" \
  -H "Content-Type: application/json" \
  -d '{ "taskId": "TASK_ID_FROM_SUBMIT" }'
# -> { "success": true, "data": { "status": "...", "videoUrl": "https://..." } }

Submission and task queries report the available task and usage information. A timeout is not a confirmed failure. Recover the existing task instead of resubmitting.

For face selection and error recovery, follow Task lifecycle. MCP uses the normalized taskId argument in get_task_result, including capabilities whose REST identifier is requestId.

Credits

Component / optionRateMinimum
Image + audio15 credits / second75 credits
Video + audio25 credits / second125 credits
Video to Video2 credits / second20 credits

The server measures the uploaded driving media; audioDuration is quote-only input.

Image/video plus audio is billed at a minimum five-second output duration.

The example uses the duration shown above. The server measures uploaded media for billing; this duration is only supplied to the estimate.

The selected input workflow determines the rate. Duration is rounded up to whole seconds and the workflow minimum applies.

Use the Playground or the MCP credit estimator for your exact inputs. Credits and billing.

Integration guides

Authentication · File uploads · Task lifecycle · Error recovery · MCP

Reference generated from the current API contractDocumentation home ↗