LTX 2.3 唇形同步: AI Lipsync

LTX-2.3 LipSync 可将嘴部动作与इनपुट音频片段同步,生成逼真的说话视频。它会保留面部身份、头部位置、光照和自然表情,同时呈现准确的唇部运动、细微眨眼和稳定的时间一致性,由升级的 LTX-2.3 架构提供支持。

📝

Overview

About this model

LTX-2.3 LipSync 可通过将嘴部动作与इनपुट音频片段同步,生成逼真的说话视频。它会保留面部身份、头部位置、光照和自然表情。

1AI 数字化身
2视频配音
3对白替换
4角色旁白
💰

Pricing & Value

Cost analysis

muapiapp$0.26 avg

基于 1.3 倍系数

** Competitor pricing is estimated based on similar model architectures and usage tiers.

⚙️

Technical Details

Configuration schema

प्रॉम्प्टstring

Optional prompt to guide lipsync generation.

Default ValueAnimate natural lip-sync to the provided audio, add subtle blinking and gentle head motion, maintain the original lighting and facial identity.
इमेज URLstring

इनपुट人像इमेज的 URL。

Default Valuehttps://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/ltx-2.3-lipsync.png
ऑडियो URLstring

URL of the input audio file.

Default Valuehttps://d3adwkbyhxyrtq.cloudfront.net/webassets/videomodels/ltx-2.3-lipsync.wav
रिज़ॉल्यूशनEnum (3 options)

生成वीडियो的रिज़ॉल्यूशन。

Default Value720p
种子int

रैंडम सीड。-1 表示随机。

Default Value-1
📖

Implementation Guide

Developer documentation

提供清晰的人像图像和高质量音频文件。你还可以提供प्रॉम्प्ट,用于指导细微的面部表情和头部运动。

Common Questions

Frequently asked

音频片段最长可以多长?

目前支持最长 20 秒的音频片段。