ant-research/4DAnyone

[SIGGRAPH Asia 2026] 4DAnyone: Create Anyone in 4D from a Casual Monocular Video

What it solves

4DAnyone transforms a single, casual monocular video of a person into a set of multi-view videos. This solves the problem of needing expensive multi-camera rigs to capture 4D (3D + time) human reconstruction, allowing users to generate a dense set of target views from a simple portrait video.

How it works

The system takes a portrait video (720p or higher) of a person and generates multiple synthetic views of that person across different camera configurations. Users can specify the number of views per layer, pitch angles for different height levels, and the horizontal yaw span to create a custom camera rig. The output includes generated target view videos and camera metadata, which can then be used as input for downstream 4D Gaussian Splatting (4DGS) reconstruction.

Who it’s for

This tool is designed for researchers and developers working in 4D human reconstruction, computer vision, and digital human creation, as well as those interested in integrating synthetic multi-view data with 4DGS pipelines.

Highlights

  • Flexible Camera Layouts: Supports various configurations, from a simple 6-view orbit to a dense 48-view layout across three pitch layers.
  • 4DGS Integration: Provides scripts to export data for Nerfstudio and train foreground-only 3D Gaussian Splatting models.
  • Customizable Viewpoints: Allows precise control over the number of views, pitch, and yaw coverage.
  • Automated Workflow: Automatically downloads required models and examples upon first use.

Related

  • Project
  • Project
  • Project
  • Project
  • Project