Skip to main navigation Skip to search Skip to main content

KEIS@JUST at SemEval-2020 Task 12: Identifying Multilingual Offensive Tweets Using Weighted Ensemble and Fine-Tuned BERT

  • Jordan University of Science and Technology

Research output: Chapter in Book/Report/Conference proceedingConference contributionpeer-review

6 Scopus citations

Abstract

This research presents our team KEIS@JUST participation at SemEval-2020 Task 12 which represents shared task on multilingual offensive language. We participated in all the provided languages for all subtasks except sub-task-A for the English language. Two main approaches have been developed the first is performed to tackle both languages Arabic and English, a weighted ensemble consists of Bi-GRU and CNN followed by Gaussian noise and global pooling layer multiplied by weights to improve the overall performance. The second is performed for other languages, a transfer learning from BERT beside the recurrent neural networks such as Bi-LSTM and Bi-GRU followed by a global average pooling layer. Word embedding and contextual embedding have been used as features, moreover, data augmentation has been used only for the Arabic language.

Original languageEnglish
Title of host publicationCOLING 2020 - The International Workshop on Semantic Evaluation, Proceedings of the 14th Workshop
EditorsAurelie Herbelot, Xiaodan Zhu, Alexis Palmer, Nathan Schneider, Jonathan May, Ekaterina Shutova
PublisherInternational Committee for Computational Linguistics
Pages2035-2044
Number of pages10
ISBN (Electronic)9781952148316
DOIs
StatePublished - 2020
Externally publishedYes
Event14th International Workshops on Semantic Evaluation, SemEval 2020, co-located with COLING 2020 - Virtual, Online, Spain
Duration: 12 Dec 202013 Dec 2020

Publication series

Name14th International Workshops on Semantic Evaluation, SemEval 2020 - co-located 28th International Conference on Computational Linguistics, COLING 2020, Proceedings

Conference

Conference14th International Workshops on Semantic Evaluation, SemEval 2020, co-located with COLING 2020
Country/TerritorySpain
CityVirtual, Online
Period12/12/2013/12/20

Fingerprint

Dive into the research topics of 'KEIS@JUST at SemEval-2020 Task 12: Identifying Multilingual Offensive Tweets Using Weighted Ensemble and Fine-Tuned BERT'. Together they form a unique fingerprint.

Cite this