← Capture & Create4DAnyone

4DAnyone Turns One Video Into a Moving 3D Person

The research system generates 16 synchronized views, then builds a free-viewpoint Gaussian-splat scene without a camera rig, known lens settings, or known camera positions.

ShareXFacebookLinkedIn
A guitarist surrounded by generated video views and a free-viewpoint volumetric reconstruction
4DAnyone generates views around a person, then uses them to build a moving scene that can be viewed from new angles.Image: 4DAnyone research team

A new research system called 4DAnyone can turn one casually recorded video of a person into a moving 3D scene that a viewer can circle from new angles.

The system does not need a multi-camera studio, calibrated lenses, known camera positions, or a tripod. Its nine researchers released the paper, code, and model on August 20, and the official project page lists the work for SIGGRAPH Asia 2026.

How It Works

4DAnyone first estimates a moving 3D skeleton from the source video. It renders that skeleton from target angles, giving its video model a guide for where the person’s body should appear in each new view.

The model then generates video from 16 viewpoints around the person. Two new techniques help those views stay consistent. One compresses earlier views into a fixed reference, while the other lets groups of views exchange information as the video takes shape.

Those generated videos become the input for 4D Gaussian splatting. This represents the person with many tiny colored points that move over time, so a viewer can watch the performance from a freely chosen angle.

Diagram showing a source video becoming 3D skeleton guides, generated target-view videos, and a 4D Gaussian-splat reconstruction
The system estimates a 3D skeleton, generates target-view videos, and uses those views to train a moving Gaussian-splat scene.Image: 4DAnyone research team

Results

The researchers tested 4DAnyone on 10 scenes from the DNA-Rendering benchmark and three from DyMVHumans. It led every reported image-quality and consistency measure across both datasets when compared with MV-Performer, TrajectoryCrafter, and a fine-tuned version of ReCamMaster.

The official demonstrations also include ordinary portrait videos recorded outside a capture studio. The paper presents those examples as evidence of broader use, but they do not have the ground-truth camera views needed for the same numerical comparison.

Release Details

The public code currently expects a portrait video showing one full or upper body, at least 121 frames, and only mild camera movement. The team lists support for lower-memory inference below 32 GB and a complete open-source 4D reconstruction path as future work.

The full research pipeline is still compute-heavy. The paper reports about two minutes to prepare a clip on an RTX 4090, about seven minutes to generate each four-video group on one H20 GPU, and about 30 minutes to train the final 4D scene on an RTX 4090.

Loose clothing can change across views because a body skeleton does not describe flowing fabric. An incorrect pose estimate also carries into every generated angle. The authors separately warn that realistic human synthesis can enable deepfakes and call for consent and clear disclosure.

For creators, the released model makes the first part available now. One simple video can become a ring of consistent moving views, with a more accessible end-to-end 4D workflow as the next step.

More from Capture & Create.

Browse all stories
Justin’s 3D character with visible skeleton bones in Tripo Studio

Tripo AI · 4 min read

Tripo Adds 3D Animation From Text Prompts

Creators can generate short actions from prompts and combine up to five motion segments, with skeleton options for Unreal Engine, Unity, and avatar workflows.

MultiSet illustration showing industrial machinery as a purple point cloud on the left and solid geometry on the right

3D mapping · 2 min read

MultiSet Adds 3D Scan Exports for NVIDIA’s Robot Simulator

The latest update also lets devices find nearby maps automatically, combines separate survey scans, and makes mapped spaces shareable in a browser.

OpenAI DevDay 2026 title from the official event teaser

AI tools · 2 min read

OpenAI Teases 20+ Launches at DevDay Today

An always-on AI agent and new model access are reportedly on the way, with Sam Altman’s keynote streaming free at 10 a.m. Pacific.

Preparing the Spatial Insider studio