Overview
Discover the power of MoMask, a novel masked modeling framework for text-driven 3D human motion generation, leveraging a hierarchical quantization scheme for high-fidelity motion representation.
Hierarchical Quantization Scheme
Represents human motion as multi-layer discrete motion tokens with high-fidelity details.
Masked Transformer
Predicts randomly masked motion tokens conditioned on text input at training stage.
Residual Transformer
Learns to progressively predict the next-layer tokens based on the results from current layer.
Text-Driven Motion Generation
Generates 3D human motions based on text input, leveraging the hierarchical quantization scheme.
Get started
- Open the official website and confirm the service is available in your region.
- Check the current plan, usage limits and terms for your intended use.
- Try a small task with sample data before committing to a paid plan.
Editorial note
Pricing checked: Sep 20, 2026. Official website is active. Please verify current pricing and terms directly on the official site.
Record updated: Sep 20, 2026
Suggest a correction