Dual In-painting Model for Unsupervised Gaze Correction and Animation in the Wild

We address the problem of unsupervised gaze correction in the wild, presenting a solution that works without the need of precise annotations of the gaze angle and the head pose. We created a new dataset called CelebAGaze consisting of two domains X, Y, where the eyes are either staring at the camera or somewhere else. Our method consists of three novel modules: the Gaze Correction module(GCM), the Gaze Animation module(GAM), and the Pretrained Autoencoder module (PAM). Specifically, GCM and GAM separately train a dual in-painting network using data from the domain X for gaze correction and data from the domain Y for gaze animation. Additionally, a Synthesis-As-Training method is proposed when training GAM to encourage the features encoded from the eye region to be correlated with the angle information, resulting in gaze animation achieved by interpolation in the latent space. To further preserve the identity information e.g., eye shape, iris color, we propose the PAM with an Autoencoder, which is based on Self-Supervised mirror learning where the bottleneck features are angle-invariant and which works as an extra input to the dual in-painting models. Extensive experiments validate the effectiveness of the proposed method for gaze correction and gaze animation in the wild and demonstrate the superiority of our approach in producing more compelling results than state-of-the-art baselines. Our code, the pretrained models and supplementary results are available at:https://github.com/zhangqianhui/GazeAnimation.

Dual In-painting Model for Unsupervised Gaze Correction and Animation in the Wild / Zhang, Jichao; Chen, Jingjing; Tang, Hao; Wang, Wei; Yan, Yan; Sangineto, Enver; Sebe, Nicu. - (2020), pp. 1588-1596. (Intervento presentato al convegno 28th ACM International Conference on Multimedia, MM 2020 tenutosi a online nel October 2020) [10.1145/3394171.3413981].

Dual In-painting Model for Unsupervised Gaze Correction and Animation in the Wild

Zhang, Jichao;Chen, Jingjing;Tang, Hao;Wang, Wei;Yan, Yan;Sangineto, Enver;Sebe, Nicu

2020

Abstract

We address the problem of unsupervised gaze correction in the wild, presenting a solution that works without the need of precise annotations of the gaze angle and the head pose. We created a new dataset called CelebAGaze consisting of two domains X, Y, where the eyes are either staring at the camera or somewhere else. Our method consists of three novel modules: the Gaze Correction module(GCM), the Gaze Animation module(GAM), and the Pretrained Autoencoder module (PAM). Specifically, GCM and GAM separately train a dual in-painting network using data from the domain X for gaze correction and data from the domain Y for gaze animation. Additionally, a Synthesis-As-Training method is proposed when training GAM to encourage the features encoded from the eye region to be correlated with the angle information, resulting in gaze animation achieved by interpolation in the latent space. To further preserve the identity information e.g., eye shape, iris color, we propose the PAM with an Autoencoder, which is based on Self-Supervised mirror learning where the bottleneck features are angle-invariant and which works as an extra input to the dual in-painting models. Extensive experiments validate the effectiveness of the proposed method for gaze correction and gaze animation in the wild and demonstrate the superiority of our approach in producing more compelling results than state-of-the-art baselines. Our code, the pretrained models and supplementary results are available at:https://github.com/zhangqianhui/GazeAnimation.

Scheda breve

Scheda completa

Scheda completa (DC)

	Anno di pubblicazione
	
				2020
			
	Titolo del Convegno
	
				28th ACM International Conference on Multimedia, MM 2020
			
	Luogo del Convegno
	
				online
			
	Data del Convegno
	
				October 2020
			
	Codice DOI
	
				https://dx.doi.org/10.1145/3394171.3413981
			
	Codice WoS
	
				WOS:000810735001072
			
	Codice Scopus
	
				2-s2.0-85106685325
			
	Pagina iniziale
	
				1588
			
	Pagina finale
	
				1596
			
	Tutti gli autori
	
						Zhang, Jichao; Chen, Jingjing; Tang, Hao; Wang, Wei; Yan, Yan; Sangineto, Enver; Sebe, Nicu
					
	Citazione
	
				Dual In-painting Model for Unsupervised Gaze Correction and Animation in the Wild / Zhang, Jichao; Chen, Jingjing; Tang, Hao; Wang, Wei; Yan, Yan; Sangineto, Enver; Sebe, Nicu. - (2020), pp. 1588-1596. (Intervento presentato al  convegno 28th ACM International Conference on Multimedia, MM 2020 tenutosi a online nel October 2020) [10.1145/3394171.3413981].
			
	Tipologia
	
				Relazione in Atti di Convegno

File in questo prodotto:

File	Dimensione	Formato
3394171.3413981.pdf Accesso riservato Dimensione 6.44 MB Formato Adobe PDF Visualizza/Apri Richiedi una copia	6.44 MB	Adobe PDF	Visualizza/Apri Richiedi una copia

Pubblicazioni consigliate

I metadati presenti in IRIS UNIMORE sono rilasciati con licenza Creative Commons CC0 1.0 Universal, mentre i file delle pubblicazioni sono rilasciati con licenza Attribuzione 4.0 Internazionale (CC BY 4.0), salvo diversa indicazione.
In caso di violazione di copyright, contattare Supporto Iris

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11380/1264567

Citazioni

ND

11

2

social impact