Multi-task Deep Learning models for real-time deployment in embedded systems

Martí I Rabadán, Miquel

dc.contributor	Pardàs Feliu, Montse
dc.contributor.author	Martí I Rabadán, Miquel
dc.contributor.other	Universitat Politècnica de Catalunya. Departament de Teoria del Senyal i Comunicacions
dc.date.accessioned	2017-06-20T07:56:55Z
dc.date.available	2017-06-20T07:56:55Z
dc.date.issued	2017-05-23
dc.identifier.uri	http://hdl.handle.net/2117/105629
dc.description.abstract	Multitask Learning (MTL) was conceived as an approach to improve the generalization ability of machine learning models. When applied to neural networks, multitask models take advantage of sharing resources for reducing the total inference time, memory footprint and model size. We propose MTL as a way to speed up deep learning models for applications in which multiple tasks need to be solved simultaneously, which is particularly useful in embedded, real-time systems such as the ones found in autonomous cars or UAVs. In order to study this approach, we apply MTL to a Computer Vision problem in which both Object Detection and Semantic Segmentation tasks are solved based on the Single Shot Multibox Detector and Fully Convolutional Networks with skip connections respectively, using a ResNet-50 as the base network. We train multitask models for two different datasets, Pascal VOC, which is used to validate the decisions made, and a combination of datasets with aerial view images captured from UAVs. Finally, we analyse the challenges that appear during the process of training multitask networks and try to overcome them. However, these hinder the capacity of our multitask models to reach the performance of the best single-task models trained without the limitations imposed by applying MTL. Nevertheless, multitask networks benefit from sharing resources and are 1.6x faster, lighter and use less memory compared to deploying the single-task models in parallel, which turns essential when running them on a Jetson TX1 SoC as the parallel approach does not fit into memory. We conclude that MTL has the potential to give superior performance as far as the object detection and semantic segmentation tasks are concerned in exchange of a more complex training process that requires overcoming challenges not present in the training of single-task models.
dc.language.iso	eng
dc.publisher	Universitat Politècnica de Catalunya
dc.rights	S'autoritza la difusió de l'obra mitjançant la llicència Creative Commons o similar 'Reconeixement-NoComercial- SenseObraDerivada'
dc.rights.uri	http://creativecommons.org/licenses/by-nc-nd/3.0/es/
dc.subject	Àrees temàtiques de la UPC::Enginyeria de la telecomunicació
dc.subject.lcsh	Machine learning
dc.subject.lcsh	Pattern recognition systems
dc.subject.other	Multi-task Learning
dc.subject.other	Deep Learning
dc.subject.other	Computer vision
dc.subject.other	Object detection
dc.subject.other	Semantic segmentation
dc.title	Multi-task Deep Learning models for real-time deployment in embedded systems
dc.title.alternative	Models multi-tasca basats en xarxes neuronals d’aprenentage profund per al desplegament en sistemes encastats i en temps real
dc.type	Master thesis
dc.subject.lemac	Aprenentatge automàtic
dc.subject.lemac	Reconeixement de formes (Informàtica)
dc.identifier.slug	ETSETB-230.123501
dc.rights.access	Open Access
dc.date.updated	2017-06-02T05:51:41Z
dc.audience.educationlevel	Màster
dc.audience.mediator	Escola Tècnica Superior d'Enginyeria de Telecomunicació de Barcelona
dc.audience.degree	MÀSTER UNIVERSITARI EN ENGINYERIA DE TELECOMUNICACIÓ (Pla 2013)
dc.contributor.covenantee	Kungl. tekniska högskolan. Skolan för elektroteknik och datavetenskap

Fitxers d'aquest items

Nom:: thesis_upc.pdf
Mida:: 55,14Mb
Format:: PDF

Visualitza/Obre

Aquest ítem apareix a les col·leccions següents

Master's degree in Telecommunications Engineering (MET) [393]

Mostra el registre d'ítem simple

UPCommons. Portal del coneixement obert de la UPC

Multi-task Deep Learning models for real-time deployment in embedded systems

Fitxers d'aquest items

Aquest ítem apareix a les col·leccions següents

Explora