New community

Subscribe to the gold package and get unlimited access to Shamra Academy

Active Learning for Interactive Relation Extraction in a French Newspaper's Articles

التعلم النشط لاستخراج العلاقات التفاعلية في مقالات الصحف الفرنسية

227 0 0 0.0 ( 0 )

Download Cite

Added by Association for Computation Linguistics مقالة

Publication date 2021

fields Artificial Intelligence

and research's language is English

Created by Shamra Editor

interactive relation extraction french newspaper articles newspaper articles استخراج العلاقة التفاعلية مقالات الصحف الفرنسية مقالات الصحف صناعة حمض الفوسفور

visit our facebook page

‎Shamra Academia - شمرا أكاديميا‎

Ask ChatGPT about the research

Abstract in Arabic Abstract in English

استخراج العلاقات هو الترجمة الفرعية لمعالجة Langage الطبيعية التي شهدت العديد من التحسينات في السنوات الأخيرة، مع ظهور البنية المعقدة المدربة مسبقا. يتم اختبار العديد من هذه النهج من هذه النهج من المعايير مع الجمل المسماة التي تحتوي على كيانات الموسومة، وتتطلب التدريب المسبق الهامة والضبط بشكل جيد على البيانات الخاصة بالمهام. ومع ذلك، في سيناريو حقيقي للاستخدام، مثل في شركة صحيفة في الغالب مخصصة لمعلومات المحلية، فإن العلاقات هي من نوع متنوع للغاية، مع عدم وجود بيانات مشروح تقريبا لمثل هذه العلاقات، والعديد من الكيانات تعاني في جملة دون أن تكون ذات صلة. نشكك في استخدام النماذج الإشرفة من أحدث النماذج في هذا السياق، حيث توجد موارد مثل الوقت والحوسبة وقوة الحوسبة والنحاذج البشرية محدودة. للتكيف مع هذه القيود، نقوم بتجربة خط أنابيب استخراج التعلم في التعلم النشط، وتتألف من نموذج خفيف الوزن يستند إلى LSTM ثنائي للكشف عن العلاقات الموجودة، ونموذج أحدث لتصنيف العلاقة. قارن العديد من الخيارات لنماذج التصنيف في هذا السيناريو، من الكلمة الأساسية لتضمين المتوسط، على الرسم البياني للشبكات العصبية وتلك القائمة على برت، وكذلك العديد من استراتيجيات الاستحواذ النشطة للتعلم، من أجل إيجاد نهج الأكثر كفاءة من حيث التكلفة ولكن دقيقة في موقعنا أكبر حالة استخدام شركة صحيفة صحيفة الفرنسية.

Relation extraction is a subtask of natural langage processing that has seen many improvements in recent years, with the advent of complex pre-trained architectures. Many of these state-of-the-art approaches are tested against benchmarks with labelled sentences containing tagged entities, and require important pre-training and fine-tuning on task-specific data. However, in a real use-case scenario such as in a newspaper company mostly dedicated to local information, relations are of varied, highly specific type, with virtually no annotated data for such relations, and many entities co-occur in a sentence without being related. We question the use of supervised state-of-the-art models in such a context, where resources such as time, computing power and human annotators are limited. To adapt to these constraints, we experiment with an active-learning based relation extraction pipeline, consisting of a binary LSTM-based lightweight model for detecting the relations that do exist, and a state-of-the-art model for relation classification. We compare several choices for classification models in this scenario, from basic word embedding averaging, to graph neural networks and Bert-based ones, as well as several active learning acquisition strategies, in order to find the most cost-efficient yet accurate approach in our French largest daily newspaper company's use case.

References used

https://aclanthology.org/

rate research

Investigating Active Learning in Interactive Neural Machine Translation

391 - Association for Computation Linguistics 2021 مقالة

Interactive-predictive translation is a collaborative iterative process and where human translators produce translations with the help of machine translation (MT) systems interactively. Various sampling techniques in active learning (AL) exist to upd ate the neural MT (NMT) model in the interactive-predictive scenario. In this paper and we explore term based (named entity count (NEC)) and quality based (quality estimation (QE) and sentence similarity (Sim)) sampling techniques -- which are used to find the ideal candidates from the incoming data -- for human supervision and MT model's weight updation. We carried out experiments with three language pairs and viz. German-English and Spanish-English and Hindi-English. Our proposed sampling technique yields 1.82 and 0.77 and 0.81 BLEU points improvements for German-English and Spanish-English and Hindi-English and respectively and over random sampling based baseline. It also improves the present state-of-the-art by 0.35 and 0.12 BLEU points for German-English and Spanish-English and respectively. Human editing effort in terms of number-of-words-changed also improves by 5 and 4 points for German-English and Spanish-English and respectively and compared to the state-of-the-art.

interactive neural machine investigating active learning آلة العصبية التفاعلية التحقيق في التعلم النشط صناعة حمض الفوسفور

Active Learning for Argument Strength Estimation

401 - Association for Computation Linguistics 2021 مقالة

High-quality arguments are an essential part of decision-making. Automatically predicting the quality of an argument is a complex task that recently got much attention in argument mining. However, the annotation effort for this task is exceptionally high. Therefore, we test uncertainty-based active learning (AL) methods on two popular argument-strength data sets to estimate whether sample-efficient learning can be enabled. Our extensive empirical evaluation shows that uncertainty-based acquisition functions can not surpass the accuracy reached with the random acquisition on these data sets.

argument strength estimation strength estimation argument strength تقدير قوة الوسيطة تقدير القوة قوة الحجة صناعة حمض الفوسفور المزيد..

Gradient Imitation Reinforcement Learning for Low Resource Relation Extraction

277 - Association for Computation Linguistics 2021 مقالة

Low-resource Relation Extraction (LRE) aims to extract relation facts from limited labeled corpora when human annotation is scarce. Existing works either utilize self-training scheme to generate pseudo labels that will cause the gradual drift problem , or leverage meta-learning scheme which does not solicit feedback explicitly. To alleviate selection bias due to the lack of feedback loops in existing LRE learning paradigms, we developed a Gradient Imitation Reinforcement Learning method to encourage pseudo label data to imitate the gradient descent direction on labeled data and bootstrap its optimization capability through trial and error. We also propose a framework called GradLRE, which handles two major scenarios in low-resource relation extraction. Besides the scenario where unlabeled data is sufficient, GradLRE handles the situation where no unlabeled data is available, by exploiting a contextualized augmentation method to generate data. Experimental results on two public datasets demonstrate the effectiveness of GradLRE on low resource relation extraction when comparing with baselines.

imitation reinforcement learning gradient imitation reinforcement resource relation extraction التعزيز التقليد التعلم التعزيز التدريجي استخراج علاقة الموارد صناعة حمض الفوسفور المزيد..

Incorporating medical knowledge in BERT for clinical relation extraction

257 - Association for Computation Linguistics 2021 مقالة

In recent years pre-trained language models (PLM) such as BERT have proven to be very effective in diverse NLP tasks such as Information Extraction, Sentiment Analysis and Question Answering. Trained with massive general-domain text, these pre-traine d language models capture rich syntactic, semantic and discourse information in the text. However, due to the differences between general and specific domain text (e.g., Wikipedia versus clinic notes), these models may not be ideal for domain-specific tasks (e.g., extracting clinical relations). Furthermore, it may require additional medical knowledge to understand clinical text properly. To solve these issues, in this research, we conduct a comprehensive examination of different techniques to add medical knowledge into a pre-trained BERT model for clinical relation extraction. Our best model outperforms the state-of-the-art systems on the benchmark i2b2/VA 2010 clinical relation extraction dataset.

حقيقي clinical relation extraction استخراج العلاقة السريرية صناعة حمض الفوسفور

ActiveEA: Active Learning for Neural Entity Alignment

358 - Association for Computation Linguistics 2021 مقالة

Entity Alignment (EA) aims to match equivalent entities across different Knowledge Graphs (KGs) and is an essential step of KG fusion. Current mainstream methods -- neural EA models -- rely on training with seed alignment, i.e., a set of pre-aligned entity pairs which are very costly to annotate. In this paper, we devise a novel Active Learning (AL) framework for neural EA, aiming to create highly informative seed alignment to obtain more effective EA models with less annotation cost. Our framework tackles two main challenges encountered when applying AL to EA: (1) How to exploit dependencies between entities within the AL strategy. Most AL strategies assume that the data instances to sample are independent and identically distributed. However, entities in KGs are related. To address this challenge, we propose a structure-aware uncertainty sampling strategy that can measure the uncertainty of each entity as well as its impact on its neighbour entities in the KG. (2) How to recognise entities that appear in one KG but not in the other KG (i.e., bachelors). Identifying bachelors would likely save annotation budget. To address this challenge, we devise a bachelor recognizer paying attention to alleviate the effect of sampling bias. Empirical results show that our proposed AL strategy can significantly improve sampling quality with good generality across different datasets, EA models and amount of bachelors.

neural entity alignment محاذاة الكيان العصبي صناعة حمض الفوسفور

يمكنك البدء بجني المال وتحقيق ربح مادي من أبحاثك العلمية، المزيد

Active Learning for Interactive Relation Extraction in a French Newspaper's Articles

التعلم النشط لاستخراج العلاقات التفاعلية في مقالات الصحف الفرنسية

Ask ChatGPT about the research

Read More

suggested questions