Skip to main content

OmniHuman-1 with API / High quality human video generation!

基于 ,可通过图片、音频或视频生成高拟真的人物动画!

OmniHuman API Pricing

Pay-as-you-go access through PiAPI.

Features

Capabilities from the production model page.

卓越的人物动画能力

基于扩散 Transformer 框架,OmniHuman 可融合多种输入,生成高拟真且富有表现力的人物动画。

多模态输入灵活性

支持图片、音频、视频输入,适配多样风格与场景,满足多种视频生成需求。

更强的音频驱动生成

提升口型同步与自然肢体动作,显著增强音频驱动视频的真实感。

动态交互能力

支持复杂的人物与物体交互表现,让演奏、持物等场景更自然生动。

更自然的头部与面部动作

面部表情与头部动作更贴合音频输入,进一步提升角色运动真实性。

灵活的视频比例与风格

支持多种画幅比例与人像风格,可覆盖近景到全身等不同构图需求。

高精度姿态驱动动画

支持任意姿态生成,细节还原优秀,适合人物视频创作与制作流程。

逼真的唱歌与说话片段

动作与表达同步性更高,适用于唱歌、口播、虚拟角色演绎等场景。

可商用

高级套餐用户可将生成视频用于合法商业用途。

Get started

Create your API key

Sign up and grab an API key from the PiAPI workspace — free credits are included on sign-up.

Top up credits

Add credits on the billing page when you are ready to scale beyond the free tier.

Call the API

POST your first task following the API docs, then poll the task until the result is ready.

Iterate with the API

Use the API docs and the request example above to iterate on prompts and settings before wiring them into your product.

Frequently asked questions

什么是 OmniHuman AI 视频生成器?

OmniHuman 是由 字节跳动 研发的人物动画生成模型。它基于扩散 Transformer 框架,支持音频、姿态、图片、视频等输入,生成高质量人物动画。

什么是 OmniHuman API?

OmniHuman API 由 PiAPI 提供,帮助开发者以高可用、可扩展的方式接入 OmniHuman 能力,并快速集成到应用中。

OmniHuman 可以生成哪些类型的视频?

支持音频驱动、姿态驱动、交互类、人像风格化、角色表演等多种人物视频类型。

OmniHuman 支持哪些输入类型?

支持图片、音频、视频,以及音视频组合输入。

OmniHuman 的典型应用场景有哪些?

适用于内容创作、营销、影视制作与开发者应用,包括虚拟角色、广告、讲解视频、娱乐内容等。

More questions? See the API docs.