模型概览
模型规格
- 上下文
- 暂未提供
- 最大输出
- 暂未提供
- 架构
- text+image+audio->video
- 分词器
- Other
- 知识截止时间
- 暂未提供
- 内容审核
- 否
能力与模态
输入 → 输出
输入
textimageaudio
输出
video
API
支持的 API 参数
可用提供方
1 提供方
HeyGen
unknown- 上下文
- 暂未提供
- 最大输出
- 暂未提供
- 输入
- 免费
- 输出
- 免费
- 缓存读取
- 暂未提供
- 缓存写入
- 暂未提供
HeyGen: Avatar IV is an image-to-video model that animates a single photo into an expressive, lip-synced talking-head video. Rather than only matching mouth shapes to words, it interprets the vocal...