New community

Subscribe to the gold package and get unlimited access to Shamra Academy

Say `YES' to Positivity: Detecting Toxic Language in Workplace Communications

قل "نعم" إلى الإيجابية: اكتشاف اللغة السامة في مجال الاتصالات

68 0 0 0.0 ( 0 )

Download Cite

Added by Association for Computation Linguistics مقالة

Publication date 2021

fields Artificial Intelligence

and research's language is English

Created by Shamra Editor

visit our facebook page

‎Shamra Academia - شمرا أكاديميا‎

Ask ChatGPT about the research

Abstract in Arabic Abstract in English

الاتصالات في مكان العمل (على سبيل المثال البريد الإلكتروني والدردشة، إلخ.) هو جزء أساسي من إنتاجية المؤسسة. المحادثات الصحية أمر حاسم لإنشاء بيئة شاملة والحفاظ على الوئام في منظمة. يمكن للاتصالات السامة في مكان العمل أن تؤثر سلبا على الرضا الوظيفي الإجمالي وغالبا ما تكون خفية أو مخفية أو إظهار تحيزات بشرية. جعلت الدقة اللغوية للمحادثات الخفيفة والأذى من الصعب على الباحثين تحديدها واستخراج المحادثات السامة تلقائيا. في حين أن اللغة الهجومية أو الكلام الكراهية قد درست على نطاق واسع في المجتمعات الاجتماعية، إلا أنه كان هناك القليل من العمل في دراسة الاتصالات السامة في رسائل البريد الإلكتروني. على وجه التحديد، فإن عدم وجود كوربوس، Sparsity من السمية في رسائل البريد الإلكتروني للمؤسسات، ومعايير محددة جيدا للتسجيل المحادثات السامة قد منع الباحثون من معالجة المشكلة على نطاق واسع. نأخذ الخطوة الأولى نحو دراسة السمية في رسائل البريد الإلكتروني في مكان العمل من خلال توفير (1) تصنيفا عاما وقابل للاستثناء بشكل خاص لدراسة اللغة السامة في مكان العمل (2) مجموعة بيانات لدراسة اللغة السامة في مكان العمل بناء على التصنيف و (3) تحليل لماذا لا تكون مجموعات البيانات الهجومية والكراهية مناسبة للكشف عن سمية مكان العمل.

Workplace communication (e.g. email, chat, etc.) is a central part of enterprise productivity. Healthy conversations are crucial for creating an inclusive environment and maintaining harmony in an organization. Toxic communications at the workplace can negatively impact overall job satisfaction and are often subtle, hidden, or demonstrate human biases. The linguistic subtlety of mild yet hurtful conversations has made it difficult for researchers to quantify and extract toxic conversations automatically. While offensive language or hate speech has been extensively studied in social communities, there has been little work studying toxic communication in emails. Specifically, the lack of corpus, sparsity of toxicity in enterprise emails, and well-defined criteria for annotating toxic conversations have prevented researchers from addressing the problem at scale. We take the first step towards studying toxicity in workplace emails by providing (1) a general and computationally viable taxonomy to study toxic language at the workplace (2) a dataset to study toxic language at the workplace based on the taxonomy and (3) analysis on why offensive language and hate-speech datasets are not suitable to detect workplace toxicity.

References used

https://aclanthology.org/

rate research

Mitigating Biases in Toxic Language Detection through Invariant Rationalization

172 - Association for Computation Linguistics 2021 مقالة

Automatic detection of toxic language plays an essential role in protecting social media users, especially minority groups, from verbal abuse. However, biases toward some attributes, including gender, race, and dialect, exist in most training dataset s for toxicity detection. The biases make the learned models unfair and can even exacerbate the marginalization of people. Considering that current debiasing methods for general natural language understanding tasks cannot effectively mitigate the biases in the toxicity detectors, we propose to use invariant rationalization (InvRat), a game-theoretic framework consisting of a rationale generator and a predictor, to rule out the spurious correlation of certain syntactic patterns (e.g., identity mentions, dialect) to toxicity labels. We empirically show that our method yields lower false positive rate in both lexical and dialectal attributes than previous debiasing methods.

toxic language detection toxic language toxic language plays كشف اللغة السامة اللغة السامة تلعب لغة سامة صناعة حمض الفوسفور المزيد..

SemEval-2021 Task 5: Toxic Spans Detection

182 - Association for Computation Linguistics 2021 مقالة

The Toxic Spans Detection task of SemEval-2021 required participants to predict the spans of toxic posts that were responsible for the toxic label of the posts. The task could be addressed as supervised sequence labeling, using training data with gol d toxic spans provided by the organisers. It could also be treated as rationale extraction, using classifiers trained on potentially larger external datasets of posts manually annotated as toxic or not, without toxic span annotations. For the supervised sequence labeling approach and evaluation purposes, posts previously labeled as toxic were crowd-annotated for toxic spans. Participants submitted their predicted spans for a held-out test set and were scored using character-based F1. This overview summarises the work of the 36 teams that provided system descriptions.

toxic spans detection spans detection task spans detection يمتد يمتد السامة يمتد مهمة الكشف عنها يمتد الكشف صناعة حمض الفوسفور المزيد..

NLP\_UIOWA at Semeval-2021 Task 5: Transferring Toxic Sets to Tag Toxic Spans

297 - Association for Computation Linguistics 2021 مقالة

We leverage a BLSTM with attention to identify toxic spans in texts. We explore different dimensions which affect the model's performance. The first dimension explored is the toxic set the model is trained on. Besides the provided dataset, we explore the transferability of 5 different toxic related sets, including offensive, toxic, abusive, and hate sets. We find that the solely offensive set shows the highest promise of transferability. The second dimension we explore is methodology, including leveraging attention, employing a greedy remove method, using a frequency ratio, and examining hybrid combinations of multiple methods. We conduct an error analysis to examine which types of toxic spans were missed and which were wrongly inferred as toxic along with the main reasons why they occurred. Finally, we extend our method via ensembles, which achieves our highest F1 score of 55.1.

tag toxic spans transferring toxic sets transferring toxic علامة السامة يمتد نقل مجموعات سامة نقل السامة صناعة حمض الفوسفور المزيد..

Assertion Detection in Clinical Notes: Medical Language Models to the Rescue?

157 - Association for Computation Linguistics 2021 مقالة

In order to provide high-quality care, health professionals must efficiently identify the presence, possibility, or absence of symptoms, treatments and other relevant entities in free-text clinical notes. Such is the task of assertion detection - to identify the assertion class (present, possible, absent) of an entity based on textual cues in unstructured text. We evaluate state-of-the-art medical language models on the task and show that they outperform the baselines in all three classes. As transferability is especially important in the medical domain we further study how the best performing model behaves on unseen data from two other medical datasets. For this purpose we introduce a newly annotated set of 5,000 assertions for the publicly available MIMIC-III dataset. We conclude with an error analysis that reveals situations in which the models still go wrong and points towards future research directions.

free-text clinical notes clinical notes medical language models النص السريري ملاحظات السريرية نماذج اللغة الطبية صناعة حمض الفوسفور المزيد..

WLV-RIT at SemEval-2021 Task 5: A Neural Transformer Framework for Detecting Toxic Spans

347 - Association for Computation Linguistics 2021 مقالة

In recent years, the widespread use of social media has led to an increase in the generation of toxic and offensive content on online platforms. In response, social media platforms have worked on developing automatic detection methods and employing h uman moderators to cope with this deluge of offensive content. While various state-of-the-art statistical models have been applied to detect toxic posts, there are only a few studies that focus on detecting the words or expressions that make a post offensive. This motivates the organization of the SemEval-2021 Task 5: Toxic Spans Detection competition, which has provided participants with a dataset containing toxic spans annotation in English posts. In this paper, we present the WLV-RIT entry for the SemEval-2021 Task 5. Our best performing neural transformer model achieves an 0.68 F1-Score. Furthermore, we develop an open-source framework for multilingual detection of offensive spans, i.e., MUDES, based on neural transformers that detect toxic spans in texts.

مطابقة الفهم القراءة toxic سامة صناعة حمض الفوسفور

يمكنك البدء بجني المال وتحقيق ربح مادي من أبحاثك العلمية، المزيد

Say `YES' to Positivity: Detecting Toxic Language in Workplace Communications

قل "نعم" إلى الإيجابية: اكتشاف اللغة السامة في مجال الاتصالات

Ask ChatGPT about the research

Read More

suggested questions