Presentation

Interspatial Attention for Efficient 4D Human Video Generation
DescriptionWe introduce a novel interspatial attention (ISA) for diffusion transformers, which maintains identity and ensures motion consistency while allowing precise control of camera and body poses. Combined with a custom video variation autoencoder, our model achieves state-of-the-art performance for photorealistic 4D human video generation.
Event Type
Technical Paper
TimeTuesday, 12 August 202511:35am - 11:45am PDT
LocationWest Building, Rooms 118-120
Session Time & Location
Sunday, 10 August 20256:00pm - 8:45pm PDTWest Building, Ballroom AB
Tuesday, 12 August 202510:45am - 12:35pm PDTWest Building, Rooms 118-120
Recordings
Livestreamed
Recorded