Overview
Stable Video 3D is a generative model that advances the field of 3D technology, delivering greatly improved quality and multi-view consistency. It can be used for commercial purposes with a Stability AI Membership and is available for non-commercial use on Hugging Face.
Multi-View Consistency
Stable Video 3D generates multi-view videos of an object, allowing for the creation of 3D video along specified camera paths.
Video Diffusion Models
Stable Video 3D uses video diffusion models to generate multi-view videos of an object, providing major benefits in generalization and view-consistency of generated outputs.
3D Optimization
Stable Video 3D leverages its multi-view consistency to optimize 3D Neural Radiance Fields (NeRF) and mesh representations to improve the quality of 3D meshes generated directly from novel views.
Disentangled Illumination Model
Stable Video 3D employs a disentangled illumination model that is jointly optimized along with 3D shape and texture to reduce the issue of baked-in lighting.
Get started
- Open the official website and confirm the service is available in your region.
- Check the current plan, usage limits and terms for your intended use.
- Try a small task with sample data before committing to a paid plan.
Editorial note
Pricing checked: Sep 20, 2026. Tiered plans available. Review current pricing and subscription options on the official website.
Record updated: Sep 20, 2026
Suggest a correction