Vid2Game: Controllable Characters Extracted from Real-World Videos

Oran Gafni,Lior Wolf,Yaniv Taigman
DOI: https://doi.org/10.48550/arXiv.1904.08379
IF: 5.414
2019-04-17
Machine Learning
Abstract:We are given a video of a person performing a certain activity, from which we extract a controllable model. The model generates novel image sequences of that person, according to arbitrary user-defined control signals, typically marking the displacement of the moving body. The generated video can have an arbitrary background, and effectively capture both the dynamics and appearance of the person. The method is based on two networks. The first network maps a current pose, and a single-instance control signal to the next pose. The second network maps the current pose, the new pose, and a given background, to an output frame. Both networks include multiple novelties that enable high-quality performance. This is demonstrated on multiple characters extracted from various videos of dancers and athletes.
What problem does this paper attempt to address?