Back to tools

Audio2Photoreal

This repository contains a PyTorch implementation of the paper 'From Audio to Photoreal Embodiment: Synthesizing Humans in Conversations' by Facebook Research. It generates photorealistic humans from audio inputs using advanced algorithms and machine learning models.

Free + paidPricing checked: Sep 20, 2026
Visit website

Best for

Users seeking ai avatar generator tools for their daily workflows.

Things to know

Usage limits, quotas, and premium capabilities are determined by the provider.

About this tool

Overview

Generate photorealistic humans from audio inputs using Audio2Photoreal, a PyTorch implementation of the paper 'From Audio to Photoreal Embodiment: Synthesizing Humans in Conversations' by Facebook Research.

Face Diffusion Model

Generates photorealistic faces from audio inputs using a face diffusion model.

Body Diffusion Model

Generates photorealistic bodies from audio inputs using a body diffusion model.

Body VQ VAE

Generates photorealistic bodies from audio inputs using a body VQ VAE model.

Body Guide Transformer

Generates photorealistic bodies from audio inputs using a body guide transformer model.

Get started

  1. Open the official website and confirm the service is available in your region.
  2. Check the current plan, usage limits and terms for your intended use.
  3. Try a small task with sample data before committing to a paid plan.

Editorial note

Pricing checked: Sep 20, 2026. Tiered plans available. Review current pricing and subscription options on the official website.

Record updated: Sep 20, 2026

Suggest a correction