Skip to main navigation Skip to search Skip to main content

Video-to-video translation with global temporal consistency

  • Xingxing Wei
  • , Sitong Feng
  • , Jun Zhu*
  • , Hang Su
  • *Corresponding author for this work
  • Tsinghua University
  • Macau University of Science and Technology

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

Abstract

Although image-to-image translation has been widely studied, the video-to-video translation is rarely mentioned. In this paper, we propose an unified video-to-video translation framework to accomplish different tasks, like video super-resolution, video colourization, and video segmentation, etc. A consequent question within video-to-video translation lies in the flickering appearance along with the varying frames. To overcome this issue, a usual method is to incorporate the temporal loss between adjacent frames in the optimization, which is a kind of local frame-wise temporal consistency. We instead present a residual error based mechanism to ensure the video-level consistency of the same location in different frames (called дlobal temporal consistency). The global and local consistency are simultaneously integrated into our video-to-video framework to achieve more stable videos. Our method is based on the GAN framework, where we present a two-channel discriminator. One channel is to encode the video RGB space, and another is to encode the residual error of the video as a whole to meet the global consistency. Extensive experiments conducted on different video-to-video translation tasks verify the effectiveness and flexibleness of the proposed method.

Original languageEnglish
Title of host publicationMM 2018 - Proceedings of the 2018 ACM Multimedia Conference
PublisherAssociation for Computing Machinery, Inc
Pages18-25
Number of pages8
ISBN (Electronic)9781450356657
DOIs
StatePublished - 15 Oct 2018
Externally publishedYes
Event26th ACM Multimedia conference, MM 2018 - Seoul, Korea, Republic of
Duration: 22 Oct 201826 Oct 2018

Publication series

NameMM 2018 - Proceedings of the 2018 ACM Multimedia Conference

Conference

Conference26th ACM Multimedia conference, MM 2018
Country/TerritoryKorea, Republic of
CitySeoul
Period22/10/1826/10/18

Keywords

  • Generative Adversarial Network
  • Temporal Consistency
  • Video-to-Video Translation

Fingerprint

Dive into the research topics of 'Video-to-video translation with global temporal consistency'. Together they form a unique fingerprint.

Cite this