Leannet : uma arquitetura que utiliza o contexto da cena para melhorar o reconhecimento de objetos

Export this record:

Please use this identifier to cite or link to this item: https://tede2.pucrs.br/tede2/handle/tede/8168

Document type:	Dissertação
Title:	Leannet : uma arquitetura que utiliza o contexto da cena para melhorar o reconhecimento de objetos
Author:	Silva, Leandro Pereira da
Advisor:	Ruiz, Duncan Dubugras Alcoba
Abstract (native):	A visão computacional é a ciência que permite fornecer aos computadores a ca- pacidade de verem o mundo em sua volta. Entre as tarefas, o reconhecimento de objetos pretende classificar objetos e identificar a posição onde cada objeto está em uma imagem. Como objetos costumam ocorrer em ambientes particulares, a utilização de seus contex- tos pode ser vantajosa para melhorar a tarefa de reconhecimento de objetos. Para utilizar o contexto na tarefa de reconhecimento de objetos, a abordagem proposta realiza a iden- tificação do contexto da cena separadamente da identificação do objeto, fundindo ambas informações para a melhora da detecção do objeto. Para tanto, propomos uma nova arquite- tura composta de duas redes neurais convolucionais em paralelo: uma para a identificação do objeto e outra para a identificação do contexto no qual o objeto está inserido. Por fim, a informação de ambas as redes é concatenada para realizar a classificação do objeto. Ava- liamos a arquitetura proposta com os datasets públicos PASCAL VOC 2007 e o MS COCO, comparando o desempenho da abordagem proposta com abordagens que não utilizam o contexto. Os resultados mostram que nossa abordagem é capaz de aumentar a probabili- dade de classificação para objetos que estão em contexto e reduzir para objetos que estão fora de contexto.
Abstract (english):	Computer vision is the science that aims to give computers the capability of see- ing the world around them. Among its tasks, object recognition intends to classify objects and to identify where each object is in a given image. As objects tend to occur in particular environments, their contextual association can be useful to improve the object recognition task. To address the contextual awareness on object recognition task, the proposed ap- proach performs the identification of the scene context separately from the identification of the object, fusing both information in order to improve the object detection. In order to do so, we propose a novel architecture composed of two convolutional neural networks running in parallel: one for object identification and the other to the identification of the context where the object is located. Finally, the information of the two-streams architecture is concatenated to perform the object classification. The evaluation is performed using PASCAL VOC 2007 and MS COCO public datasets, by comparing the performance of our proposed approach with architectures that do not use the scene context to perform the classification of the ob- jects. Results show that our approach is able to raise in-context object scores, and reduces out-of-context objects scores.
Keywords:	Detecção de Objetos Rede Neural Convolucional Rede Neural Aprendizagem Profunda Objetos em Contexto Object Detection Convolutional Neural Network Neural Network Deep Learning Object in Context
CNPQ Knowledge Areas:	CIENCIA DA COMPUTACAO::TEORIA DA COMPUTACAO
Language:	por
Country:	Brasil
Publisher:	Pontifícia Universidade Católica do Rio Grande do Sul
Institution Acronym:	PUCRS
Department:	Escola Politécnica
Program:	Programa de Pós-Graduação em Ciência da Computação
Access type:	Acesso Aberto
Fulltext access restriction:	Trabalho não apresenta restrição para publicação
URI:	http://tede2.pucrs.br/tede2/handle/tede/8168
Issue Date:	27-Mar-2018
Appears in Collections:	Programa de Pós-Graduação em Ciência da Computação

Files in This Item:

File	Description	Size	Format
DIS_LEANDRO_PEREIRA_DA_SILVA_COMPLETO.pdf	LEANDRO_PEREIRA_DA_SILVA_DIS	1.7 MB	Adobe PDF	Download/Open Preview ×

Show full item record Recommend this item

PUCRS

Digital Library of Theses and Dissertations