Predicting ego-centric video from human actions (PEVA) generates the next frame from past frames and actions; model supports atomic actions, counterfactuals, and long video generation
Read the original at bair.berkeley.edu→× Predicting Ego-centric Video from human Actions (PEVA). Given past video frames and an action specifying a desired change in 3D pose, PEVA predicts the next video frame. Our results show...
Original headline: "Whole-Body Conditioned Egocentric Video Prediction"