Luận Văn Thạc Sĩ: Mô Hình Học Sâu Nâng Cao và Ứng Dụng trong Khai Thác Quan Hệ Ngữ Nghĩa

Luận văn thạc sĩ nghiên cứu vnu uet advanced deep learning models and applications in semantic relation extraction, khảo sát thực trạng, phân tích nguyên nhân, đề xuất giải pháp

Chuyên ngành

Computer Science

Người đăng

Ẩn danh

Thể loại

master thesis

2019

82
1
0

Phí lưu trữ

30 Point

Mục lục chi tiết

Abstract

Acknowledgements

Declaration

1. Chapter 1: Introduction

1.1. Motivation

1.2. Problem Statement

1.3. Formal Definition

1.4. Examples

Tóm tắt

I. Tổng Quan về Mô Hình Học Sâu Nâng Cao trong Khai Thác Quan Hệ Ngữ Nghĩa

Mô hình học sâu nâng cao đã trở thành một công cụ quan trọng trong lĩnh vực khai thác quan hệ ngữ nghĩa. Các mô hình này không chỉ giúp nhận diện các thực thể mà còn phân loại mối quan hệ giữa chúng. Việc áp dụng các thuật toán học sâu như mạng nơ-ron tích chập (CNN) và mạng nơ-ron hồi tiếp dài ngắn (LSTM) đã mở ra nhiều cơ hội mới trong việc xử lý ngôn ngữ tự nhiên. Nghiên cứu này sẽ đi sâu vào các phương pháp và ứng dụng của mô hình học sâu trong khai thác quan hệ ngữ nghĩa.

1.1. Khái Niệm và Tầm Quan Trọng của Khai Thác Quan Hệ Ngữ Nghĩa

Khai thác quan hệ ngữ nghĩa (RE) là quá trình nhận diện và phân loại mối quan hệ giữa các thực thể trong văn bản. Điều này rất quan trọng trong việc trích xuất thông tin từ các tài liệu không có cấu trúc. Các ứng dụng của RE bao gồm tìm kiếm thông tin, phân tích văn bản và nhiều lĩnh vực khác.

1.2. Các Mô Hình Học Sâu Nổi Bật trong Khai Thác Quan Hệ

Các mô hình học sâu như CNN và LSTM đã chứng minh hiệu quả trong việc khai thác quan hệ ngữ nghĩa. CNN giúp trích xuất đặc trưng từ văn bản, trong khi LSTM hỗ trợ trong việc xử lý các chuỗi dữ liệu. Sự kết hợp của chúng tạo ra những mô hình mạnh mẽ cho việc phân loại quan hệ.

II. Thách Thức trong Khai Thác Quan Hệ Ngữ Nghĩa với Mô Hình Học Sâu

Mặc dù mô hình học sâu mang lại nhiều lợi ích, nhưng vẫn tồn tại nhiều thách thức trong việc khai thác quan hệ ngữ nghĩa. Một trong những vấn đề chính là việc thiếu dữ liệu huấn luyện chất lượng cao. Ngoài ra, việc xử lý các mối quan hệ phức tạp và đa dạng trong ngữ nghĩa cũng là một thách thức lớn.

2.1. Thiếu Dữ Liệu Huấn Luyện Chất Lượng

Việc thiếu dữ liệu huấn luyện chất lượng cao có thể dẫn đến việc mô hình không học được các đặc trưng quan trọng. Điều này ảnh hưởng trực tiếp đến độ chính xác của các kết quả khai thác quan hệ.

2.2. Xử Lý Các Mối Quan Hệ Phức Tạp

Các mối quan hệ ngữ nghĩa thường rất phức tạp và đa dạng. Việc nhận diện và phân loại chính xác các mối quan hệ này đòi hỏi các mô hình phải có khả năng hiểu ngữ cảnh và các yếu tố liên quan.

III. Phương Pháp Nâng Cao Mô Hình Học Sâu trong Khai Thác Quan Hệ

Để cải thiện hiệu suất của các mô hình học sâu trong khai thác quan hệ ngữ nghĩa, nhiều phương pháp đã được đề xuất. Một trong số đó là việc kết hợp các đặc trưng ngữ nghĩa với các mô hình học sâu. Phương pháp này giúp tăng cường khả năng nhận diện và phân loại các mối quan hệ.

3.1. Kết Hợp Các Đặc Trưng Ngữ Nghĩa

Việc kết hợp các đặc trưng ngữ nghĩa vào mô hình học sâu giúp cải thiện khả năng nhận diện các mối quan hệ. Các đặc trưng này có thể bao gồm thông tin ngữ cảnh và các yếu tố ngữ nghĩa khác.

3.2. Sử Dụng Các Kỹ Thuật Attention

Kỹ thuật attention giúp mô hình tập trung vào các phần quan trọng của văn bản, từ đó cải thiện độ chính xác trong việc phân loại quan hệ. Kỹ thuật này đã được áp dụng thành công trong nhiều mô hình học sâu hiện đại.

IV. Ứng Dụng Thực Tiễn của Mô Hình Học Sâu trong Khai Thác Quan Hệ Ngữ Nghĩa

Mô hình học sâu đã được áp dụng rộng rãi trong nhiều lĩnh vực khác nhau, từ y tế đến tài chính. Việc khai thác quan hệ ngữ nghĩa giúp trích xuất thông tin quan trọng từ các tài liệu lớn, từ đó hỗ trợ ra quyết định và phân tích dữ liệu.

4.1. Ứng Dụng trong Y Tế

Trong lĩnh vực y tế, việc khai thác quan hệ ngữ nghĩa giúp nhận diện các mối quan hệ giữa các bệnh và thuốc. Điều này hỗ trợ trong việc phát hiện các tác dụng phụ và tương tác giữa các loại thuốc.

4.2. Ứng Dụng trong Tài Chính

Trong lĩnh vực tài chính, khai thác quan hệ ngữ nghĩa giúp phân tích các mối quan hệ giữa các công ty, sản phẩm và thị trường. Điều này giúp các nhà đầu tư đưa ra quyết định chính xác hơn.

V. Kết Luận và Tương Lai của Mô Hình Học Sâu trong Khai Thác Quan Hệ Ngữ Nghĩa

Mô hình học sâu đã mở ra nhiều cơ hội mới trong việc khai thác quan hệ ngữ nghĩa. Tuy nhiên, vẫn còn nhiều thách thức cần phải vượt qua. Tương lai của lĩnh vực này hứa hẹn sẽ có nhiều tiến bộ với sự phát triển của công nghệ và các phương pháp mới.

5.1. Tiềm Năng Phát Triển

Với sự phát triển không ngừng của công nghệ học sâu, tiềm năng phát triển trong khai thác quan hệ ngữ nghĩa là rất lớn. Các mô hình mới sẽ ngày càng chính xác và hiệu quả hơn.

5.2. Hướng Nghiên Cứu Tương Lai

Hướng nghiên cứu tương lai có thể tập trung vào việc cải thiện độ chính xác của các mô hình và phát triển các phương pháp mới để xử lý các mối quan hệ phức tạp hơn.

22/07/2025
Luận văn thạc sĩ vnu uet advanced deep learning models and applications in semantic relation extraction

Trích đoạn nội dung tài liệu

VIETNAM NATIONAL UNIVERSITY, HANOI UNIVERSITY OF ENGINEERING AND TECHNOLOGY CAN DUY CAT ADVANCED DEEP LEARNING MODELS AND APPLICATIONS IN SEMANTIC RELATION EXTRACTION MASTER THESIS Major: Computer Science HA NOI - 2019 LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com VIETNAM NATIONAL UNIVERSITY, HANOI UNIVERSITY OF ENGINEERING AND TECHNOLOGY Can Duy Cat ADVANCED DEEP LEARNING MODELS AND APPLICATIONS IN SEMANTIC RELATION EXTRACTION MASTER THESIS Major: Computer Science Supervisor: Assoc. Ha Quang Thuy Assoc. Chng Eng Siong HA NOI - 2019 LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com Abstract Relation Extraction (RE) is one of the most fundamental task of Natural Language Pro- cessing (NLP) and Information Extraction (IE). To extract the relationship between two entities in a sentence, two common approaches are (1) using their shortest dependency path (SDP) and (2) using an attention model to capture a context-based representation of the sentence.

Each approach suffers from its own disadvantage of either missing or redundant information. In this work, we propose a novel model that combines the ad- vantages of these two approaches. This is based on the basic information in the SDP enhanced with information selected by several attention mechanisms with kernel filters, namely RbSP (Richer-but-Smarter SDP). To exploit the representation behind the RbSP structure effectively, we develop a combined Deep Neural Network (DNN) with a Long Short-Term Memory (LSTM) network on word sequences and a Convolutional Neural Network (CNN) on RbSP.

Furthermore, experiments on the task of RE proved that data representation is one of the most influential factors to the model’s performance but still has many limitations. We propose (i) a compositional embedding that combines several dominant linguistic as well as architectural features and (ii) dependency tree normalization techniques for generating rich representations for both words and dependency relations in the SDP. Experimental results on both general data (SemEval-2010 Task 8) and biomedical data (BioCreative V Track 3 CDR) demonstrate the out-performance of our proposed model over all compared models. Keywords: Relation Extraction, Shortest Dependency Path, Convolutional Neural Net- work, Long Short-Term Memory, Attention Mechanism.

iii LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com Acknowledgements I would first like to thank my thesis supervisor Assoc. Ha Quang Thuy of the Data Science and Knowledge Technology Laboratory at University of Engineering and Technology. He consistently allowed this paper to be my own work, but steered me in the right the direction whenever he thought I needed it. I also want to acknowledge my co-supervisor Assoc.Prof Chng Eng Siong from Nanyang Technological University, Singapore for offering me the internship opportuni- ties at NTU, Singapore and leading me working on diverse exciting projects.

Furthermore, I am very grateful to my external advisor MSc. Le Hoang Quynh, for insightful comments both in my work and in this thesis, for her support, and for many motivating discussions. In addition, I have been very privileged to get to know and to collaborate with many other great collaborators. I would like to thank BSc.

Nguyen Minh Trang and BSc. Nguyen Duc Canh for inspiring discussion, and for all the fun we have had over the last two years. I thank to MSc. Ho Thi Nga and MSc.

Vu Thi Ly for continuous support during the time in Singapore. Finally, I must express my very profound gratitude to my family for providing me with unfailing support and continuous encouragement throughout my years of study and through the process of researching and writing this thesis. This accomplishment would not have been possible without them. iv LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com Declaration I declare that the thesis has been composed by myself and that the work has not be submitted for any other degree or professional qualification.

I confirm that the work submitted is my own, except where work which has formed part of jointly-authored publications has been included. My contribution and those of the other authors to this work have been explicitly indicated below. I confirm that appropriate credit has been given within this thesis where reference has been made to the work of others. The model presented in Chapter 3 and the results presented in Chapter 4 was pre- viously published in the Proceedings of ACIIDS 2019 as “Improving Semantic Relation Extraction System with Compositional Dependency Unit on Enriched Shortest Depen- dency Path” and NAACL-HTL 2019 as “A Richer-but-Smarter Shortest Dependency Path with Attentive Augmentation for Relation Extraction” by myself et al.

This study was conceived by all of the authors. I carried out the main idea(s) and implemented all the model(s) and material(s). I certify that, to the best of my knowledge, my thesis does not infringe upon any- one’s copyright nor violate any proprietary rights and that any ideas, techniques, quota- tions, or any other material from the work of other people included in my thesis, pub- lished or otherwise, are fully acknowledged in accordance with the standard referencing practices. Furthermore, to the extent that I have included copyrighted material, I certify that I have obtained a written permission from the copyright owner(s) to include such material(s) in my thesis and have fully authorship to improve these materials.

Master student Can Duy Cat v LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com Table of Contents Abstract. v Table of Contents. ix List of Figures. xi List of Tables .3 Difficulties and Challenges .5 Contributions and Structure of the Thesis .1 Rule-Based Approaches .1 Feature-Based Machine Learning .2 Deep Learning Methods .4 Distant and Semi-Supervised Methods.

18 vi LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com 3 Materials and Methods .2 Convolutional Neural Network .3 Long Short-Term Memory .2 Overview of Proposed System .3 Richer-but-Smarter Shortest Dependency Path .1 Dependency Tree and Dependency Tree Normalization .2 Shortest Dependency Path and Dependency Unit .3 Richer-but-Smarter Shortest Dependency Path .4 Multi-layer Attention with Kernel Filters .2 Multi-layer Attention .5 Deep Learning Model for Relation Classification .2 CNN on Shortest Dependency Path .3 Training objective and Learning method .4 Model Improvement Techniques. 41 4 Experiments and Results .1 Implementation and Configurations .2 Training and Testing Environment .2 Datasets and Evaluation methods .2 Metrics and Evaluation .3 Performance of Proposed model .2 System performance on General domain .3 System performance on Biomedical data .4 Contribution of each Proposed Component. 56 vii LUAN VAN CHAT LUONG download : add luanvanchat@agmail. 60 List of Publications.

62 viii LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com Acronyms Adam Adaptive Moment Estimation ANN Artificial Neural Network BiLSTM Bidirectional Long Short-Term Memory CBOW Continuous Bag-Of-Words CDR Chemical Disease Relation CID Chemical-Induced Disease CNN Convolutional Neural Network DNN Deep Neural Network DU Dependency Unit GD Gradient Descent IE Information Extraction LSTM Long Short-Term Memory MLP Multilayer Perceptron NE Named Entity NER Named Entity Recognition NLP Natural Language Processing POS Part-Of-Speech ix LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com RbSP Richer-but-Smarter Shortest Dependency Path RC Relation Classification RE Relation Extraction ReLU Rectified Linear Unit RNN Recurrent Neural Network SDP Shortest Dependency Path SVM Suport Vector Machine x LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com List of Figures 1.1 A typical pipeline of Relation Extraction system.2 Two examples from SemEval 2010 Task 8 dataset.3 Example from SemEval 2017 ScienceIE dataset.4 Examples of (a) cross-sentence relation and (b) intra-sentence relation.5 Examples of relations with specific and unspecific location.6 Examples of directed and undirected relation from Phenebank corpus.1 Sentence modeling using Convolutional Neural Network.2 Convolutional approach to character-level feature extraction.3 Traditional Recurrent Neural Network.4 Architecture of a Long Short-Term Memory unit.5 The overview of end-to-end Relation Classification system.6 An example of dependency tree generated by spaCy.7 Example of normalized dependency tree.8 Dependency units on the SDP.9 Examples of SDPs and attached child nodes.10 The multi-layer attention architecture to extract the augmented informa- tion.11 The architecture of RbSP model for relation classification.1 Contribution of each compositional embeddings component.2 Comparing the contribution of augmented information by removing these components from the model .3 Comparing the effects of using RbSP in two aspects, (i) RbSP improved performance and (ii) RbSP yielded some additional wrong results. 58 xi LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com List of Tables 4.1 Configurations and parameters of proposed model.2 Statistics of SemEval-2010 Task 8 dataset.3 Summary of the BioCreative V CDR dataset .4 The comparison of our model with other comparative models on SemEval 2010 Task 8 dataset.5 The comparison of our model with other comparative models on BioCre- ative V CDR dataset .6 The examples of error from RbSP and Baseline models. 59 xii LUAN VAN CHAT LUONG download : add luanvanchat@agmail.com Chapter 1 Introduction 1.1 Motivation With the advent of the Internet, we are stepping in to a new era, the era of information and technology where the growth and development of each individual, organization, and society is relied on the main strategic resource - information. There exists a large amount of unstructured digital data that are created and maintained within an enterprise or across the Web, including news articles, blogs, papers, research publications, emails, reports, governmental documents, etc.

Lot of important information is hidden within these doc- uments that we need to extract to make them more accessible for further processing. Many tasks of Natural Language Processing (NLP) would benefit from extracted information in large text corpora, such as Question Answering, Textual Entailment, Text Understanding, etc. For example, getting a paperwork procedure from a large collection of administrative documents is a complicated problem; it is far easier to get it from a structural database such as that shown above. Similarly, searching for the side effects of a chemical in the bio-medical literature will be much easier if these relations have been extracted from biomedical text.

We, therefore, have urge to turn unstructured text into structured by annotating semantic information. Normally, we are interested in relations between entities, such as person, organization, and location. However, it is impossible for human annotation because of sheer volume and heterogeneity of data. Instead, we would like to have a Relation Extraction (RE) system that annotate all data with the structure of our interest.

In this thesis, we will focus on the task of recognizing relations between entities in unstructured text. 1 LUAN VAN CHAT LUONG download : add luanvanchat@agmail.2 Problem Statement Relation Extraction task includes of detecting and classifying relationship between enti- ties within a set of artifacts, typically from text or XML documents.1 shows an overview of a typical pipeline for RE system. Here we have to sub-tasks: Named Entity Recognition (NER) task and Relation Classification (RC) task. Named Relation Unstructured Entity Classification Knowledge literature Recognition Figure 1.1: A typical pipeline of Relation Extraction system.

A Named Entity (NE) is a specific real-world object that is often represented by a word or phrase. It can be abstract or have a physical existence such as a person, a loca- tion, a organization, a product, a brand name, etc. For example, “Hanoi” and “Vietnam” are two named entities, and they are specific mentions in the following sentence: “Hanoi city is the capital of Vietnam”. Named entities can simply be viewed as entity instances (e., Hanoi is an instance of a city).

A named entity mention in a particular sentence can be using the name itself (Hanoi), nominal (capital of Vietnam), or pronominal (it). Named Entity Recognition is the task of seeking to locate and classify named entity mentions in unstructured text into pre-defined categories. A relation usually denotes a well-defined (having a specific meaning) relationship between two or more NEs. It can be defined as a labeled tuple R(e1 , e2 , ., en ) where the ei are entities in a predefined relation R within document D.

Most relation extrac- tion systems focus on extracting binary relations. Some examples of relations are the relation capital-of between a CITY and a COUNTRY, the relation author-of be- tween a PERSON and a BOOK, the relation side-effect-of between DISEASEs and a CHEMICAL, etc. It is also possible be the n-ary relation as well. For example, the relation diagnose between a DOCTOR, a PATIENT and a DISEASE.

Nội dung được bảo vệ bản quyền — Tải xuống đầy đủ

Tài liệu "Mô Hình Học Sâu Nâng Cao và Ứng Dụng trong Khai Thác Quan Hệ Ngữ Nghĩa" mang đến một cái nhìn chuyên sâu về cách các mô hình học sâu tiên tiến được ứng dụng để nhận diện và phân tích các mối quan hệ ngữ nghĩa phức tạp ẩn chứa trong văn bản. Đây là một khía cạnh cực kỳ quan trọng trong xử lý ngôn ngữ tự nhiên (NLP), giúp máy móc không chỉ "đọc" mà còn "hiểu" được ý nghĩa sâu sắc, ngữ cảnh và các mối liên kết giữa các thực thể. Độc giả sẽ nắm bắt được các kỹ thuật hiện đại để trích xuất thông tin có giá trị cao từ lượng lớn dữ liệu văn bản, mở ra tiềm năng lớn trong việc phát triển các hệ thống tìm kiếm thông minh hơn, phân tích dữ liệu chuyên sâu và tự động hóa các tác vụ dựa trên ngôn ngữ.

Nếu bạn muốn khám phá thêm về cách học sâu được áp dụng vào các bài toán cụ thể khác trong NLP, hãy đào sâu kiến thức với Luận văn thạc sĩ phân tích ý kiến người dùng theo khía cạnh bằng phương pháp học sâu để hiểu cách thức phân tích cảm xúc và ý kiến từ dữ liệu người dùng, một ứng dụng thiết thực của việc hiểu ngữ nghĩa. Hoặc để tìm hiểu về các phương pháp học sâu tiên tiến hơn và ứng dụng của chúng trong việc xây dựng các hệ thống hỏi đáp thông minh, giúp máy tính trả lời câu hỏi một cách tự nhiên và chính xác, bạn có thể tham khảo Luận văn thạc sĩ vnu uet advanced deep learning methods and applications in opendomain question answering các phương pháp học sâu tiên tiến và ứng dụng vào bài toán hệ hỏi đáp miền mở. Mỗi tài liệu là một cơ hội tuyệt vời để bạn mở rộng tầm nhìn và tăng cường kiến thức chuyên sâu về tiềm năng không giới hạn của học sâu trong lĩnh vực ngôn ngữ.