CATALOGO DEI PRODOTTI DELLA RICERCA

Recent approaches on trajectory forecasting use tracklets to predict the future positions of pedestrians exploiting Long Short Term Memory (LSTM) architectures. This paper shows that adding vislets, that is, short sequences of head pose estimations, allows to increase significantly the trajectory forecasting performance. We then propose to use vislets in a novel framework called MX-LSTM, capturing the interplay between tracklets and vislets thanks to a joint unconstrained optimization of full covariance matrices during the LSTM backpropagation. At the same time, MX-LSTM predicts the future head poses, increasing the standard capabilities of the long-term trajectory forecasting approaches. With standard head pose estimators and an attentional-based social pooling, MX-LSTM scores the new trajectory forecasting state-of-the-art in all the considered datasets (Zara01, Zara02, UCY, and TownCentre) with a dramatic margin when the pedestrians slow down, a case where most of the forecasting approaches struggle to provide an accurate solution.

MX-LSTM: Mixing Tracklets and Vislets to Jointly Forecast Trajectories and Head Poses

Irtiza Hasan;Francesco Setti;Theodore Tsesmelis;Alessio Del Bue;Fabio Galasso;Marco Cristani^{Formal Analysis}

2018-01-01

Abstract

Recent approaches on trajectory forecasting use tracklets to predict the future positions of pedestrians exploiting Long Short Term Memory (LSTM) architectures. This paper shows that adding vislets, that is, short sequences of head pose estimations, allows to increase significantly the trajectory forecasting performance. We then propose to use vislets in a novel framework called MX-LSTM, capturing the interplay between tracklets and vislets thanks to a joint unconstrained optimization of full covariance matrices during the LSTM backpropagation. At the same time, MX-LSTM predicts the future head poses, increasing the standard capabilities of the long-term trajectory forecasting approaches. With standard head pose estimators and an attentional-based social pooling, MX-LSTM scores the new trajectory forecasting state-of-the-art in all the considered datasets (Zara01, Zara02, UCY, and TownCentre) with a dramatic margin when the pedestrians slow down, a case where most of the forecasting approaches struggle to provide an accurate solution.

Scheda breve

Scheda completa

Scheda completa (DC)

	Anno
	
				2018
			
	Presenza di coautori internazionali
	
				sì
			
	Lingua/e di pubblicazione
	
				Inglese
			
	Formato della pubblicazione
	
				ELETTRONICO
			
	Pubblicazione con Referee
	
				Comitato scientifico
			
	Titolo del Convegno
	
				IEEE CONFERENCE ON COMPUTER VISION AND PATTERN RECOGNITION
			
	Luogo del Convegno
	
				Salt Lake City
			
	Periodo del Convegno
	
				Giugno 2018
			
	Rilevanza del Convegno
	
				Internazionale
			
	Convegno su invito
	
				contributo
			
	Titolo del libro
	
				Prooceedings of the 2018 IEEE/CVF Conference on Computer Vision and Pattern Recognition
			
	Pagina iniziale
	
				6067
			
	Pagina finale
	
				6076
			
	Numero di Pagine
	
				10
			
	Specificare l'identificativo della risorsa su Web of Science
	
				WOS:000457843606023
			
	Specificare l'identificativo della risorsa su Scopus
	
				2-s2.0-85062825158
			
	Indicare il codice DOI della risorsa
	
				https://dx.doi.org/10.1109/CVPR.2018.00635
			
	Parole Chiave
	
				Video surveillance, forecasting, deep learning
			
	Fulltext
	
				open
			
	Tutti gli autori
	
						Hasan, Irtiza; Setti, Francesco; Tsesmelis, Theodore; Del Bue, Alessio; Galasso, Fabio; Cristani, Marco
					
	Numero autori
	
				6
			
	Tipologia
	
				04 Contributo in atti di convegno::04.01 Contributo in atti di convegno
			
	Tipologia sito docente
	
				273
			
	Tipologia
	
				info:eu-repo/semantics/conferenceObject
			
	Appare nelle tipologie:
	
				04.01 Contributo in atti di convegno

File in questo prodotto:

File	Dimensione	Formato
Hasan_MX-LSTM_Mixing_Tracklets_CVPR_2018_paper.pdf accesso aperto Tipologia: Versione dell'editore Licenza: Dominio pubblico Dimensione 1.19 MB Formato Adobe PDF Visualizza/Apri	1.19 MB	Adobe PDF	Visualizza/Apri

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11562/988634

Citazioni

ND

123

103

social impact