Overview
UniAnimate is a framework that enables efficient and long-term human video generation by taming unified video diffusion models for consistent human image animation.
Unified Video Diffusion Model
Maps the reference image along with the posture guidance and noise video into a common feature space.
Unified Noise Input
Supports random noised input as well as first frame conditioned input, enhancing the ability to generate long-term video.
Temporal Modeling Architecture
An alternative temporal modeling architecture based on state space model to replace the original computation-consuming temporal Transformer.
Get started
- Open the official website and confirm the service is available in your region.
- Check the current plan, usage limits and terms for your intended use.
- Try a small task with sample data before committing to a paid plan.
Editorial note
Pricing checked: Sep 20, 2026. Official website is active. Please verify current pricing and terms directly on the official site.
Record updated: Sep 20, 2026
Suggest a correction