Using complex networks and Deep Learning to model and learn context

Júnior, Edilson Anselmo Corrêa

doi:10.11606/T.55.2020.tde-16022021-151616

Home

Facilities

Doctoral Thesis

DOI

https://doi.org/10.11606/T.55.2020.tde-16022021-151616

Document

Doctoral Thesis

Author

Júnior, Edilson Anselmo Corrêa (Catálogo USP)

Full name

Edilson Anselmo Corrêa Júnior

E-mail

Institute/School/College

Instituto de Ciências Matemáticas e de Computação

Knowledge Area

Computer Science and Computational Mathematics

Date of Defense

2020-12-15

Published

São Carlos, 2020

Supervisor

Amancio, Diego Raphael (Catálogo USP)

Committee

Amancio, Diego Raphael (President)
Comin, César Henrique
Liang, Zhao
Travieso, Gonzalo

Title in English

Using complex networks and Deep Learning to model and learn context

Keywords in English

Ambiguity
Complex networks
Context
Deep learning

Abstract in English

The structure of language is strongly influenced by the context, whether it is the social setting, of discourse (spoken and written) or the context of words itself. This fact allowed the creation of several techniques of Natural Language Processing (NLP) that take advantage of this information to tackle a myriad of tasks, including machine translation, summarization and classification of texts. However, in most of these applications, the context has been approached only as a source of information and not as an element to be explored and modeled. In this thesis, we explore the context on a deeper level, bringing new representations and methodologies. Throughout the thesis, we considered context as an important element that must be modeled in order to better perform NLP tasks. We demonstrated how complex networks can be used both to represent and learn context information while performing word sense disambiguation. In addition, we proposed a context modeling approach that combines word embeddings and a network representation, this approach allowed the induction of senses in an unsupervised way using community detection methods. Using this representation we further explored its application in text classification, we expanded the approach to allow the extraction of text features based on the semantic flow, which were later used in a supervised classifier trained to discriminate texts by genre and publication date. The studies carried out in this thesis demonstrate that context modeling is important given the interdependence between language and context, and that it can bring benefits for different NLP tasks. The framework proposed, both for modeling and textual feature extraction can be further used to explore other aspects and mechanisms of language.

Title in Portuguese

Modelagem e aprendizado de contexto usando redes complexas e Deep Learning

Keywords in Portuguese

Ambiguidade
Contexto
Deep learning
Redes complexas

Abstract in Portuguese

A estrutura da língua é fortemente influenciada pelo contexto, seja ele social, do discurso (falado e escrito) ou o próprio contexto de palavras. Este preceito propiciou a criação de várias técnicas de Processamento de Língua Natural (PLN) que tiram vantagem dessa informação para realizar uma miríade de tarefas, incluindo tradução automática, sumarização e classificação de textos. Entretanto, em grande parte dessas aplicações o contexto tem sido abordado apenas como uma informação de entrada e não como um elemento a ser explorado e modelado. Nesta tese, exploramos o contexto em um nível mais profundo, trazendo novas representações e metodologias. Ao longo da tese, consideramos o contexto como um elemento importante que deve ser modelado para melhor desempenhar as tarefas da PLN. Demonstramos como redes complexas podem ser usadas para representar e aprender informações de contexto durante a desambiguação do sentido das palavras. Além disso, propusemos uma abordagem de modelagem de contexto que combina word embeddings e uma representação de rede, esta abordagem permitiu a indução de sentidos de uma forma não supervisionada usando métodos de detecção de comunidade. Usando essa representação exploramos sua aplicação na classificação de textos, expandimos a abordagem para permitir a extração de características de texto com base no fluxo semântico, que foram posteriormente usadas em um classificador supervisionado treinado para discriminar textos por gênero e data de publicação. Os estudos realizados nesta tese demonstram que a modelagem de contexto é importante dada a interdependência entre linguagem e contexto, e que pode trazer benefícios para diferentes tarefas de PLN. O framework proposto, tanto para modelagem quanto para extração de características textuais, pode ser posteriormente utilizado para explorar outros aspectos e mecanismos da linguagem.

WARNING - Viewing this document is conditioned on your acceptance of the following terms of use:
This document is only for private use for research and teaching activities. Reproduction for commercial use is forbidden. This rights cover the whole data about this document as well as its contents. Any uses or copies of this document in whole or in part must include the author's name.

EdilsonAnselmoCorreaJunior_revisada.pdf (2.78 Mbytes)

Publishing Date

2021-02-16

Derived works

WARNING: Learn what derived works are clicking here.