Detecting Persuasion Attempts on Social Networks: Unearthing the Potential of Loss Functions and Text Pre-Processing in Imbalanced Data Settings

dc.contributor.authorTeimas, Rúben
dc.contributor.authorSaias, José
dc.date.accessioned2023-11-22T11:09:03Z
dc.date.available2023-11-22T11:09:03Z
dc.date.issued2023-10-29
dc.description.abstractThe rise of social networks and the increasing amount of time people spend on them have created a perfect place for the dissemination of false narratives, propaganda, and manipulated content. In order to prevent the spread of disinformation, content moderation is needed. However, manual moderation is unfeasible due to the large amount of daily posts. This paper studies the impact of using different loss functions on a multi-label classification problem with an imbalanced dataset, consisting of 20 persuasion techniques and only 950 samples, provided by SemEval’s 2021 Task 6. We used machine learning models, such as Naive Bayes and Decision Trees, and a custom deep learning architecture, based on DistilBERT and Convolutional Layers. Overall, the machine learning models achieved far worse results than the deep learning model, using Binary Cross Entropy, which we considered our baseline deep learning model. To address the class imbalance problem, we trained our model using different loss functions, such as Focal Loss and Asymmetric Loss. The latter providing the best results, particularly for the least represented classes.por
dc.identifier.authoremailruben.teimas@uevora.pt
dc.identifier.authoremailjsaias@uevora.pt
dc.identifier.citationRúben Teimas and José Saias. 2023. "Detecting Persuasion Attempts on Social Networks: Unearthing the Potential of Loss Functions and Text Pre-Processing in Imbalanced Data Settings" Electronics 12, no. 21: 4447.por
dc.identifier.doihttps://doi.org/10.3390/electronics12214447por
dc.identifier.issn2079-9292
dc.identifier.revistaElectronics
dc.identifier.scientificarea283por
dc.identifier.urihttps://www.mdpi.com/2539862
dc.identifier.urihttp://hdl.handle.net/10174/35708
dc.language.isoengpor
dc.peerreviewedyespor
dc.publisherMDPI - Electronicspor
dc.rightsopenAccesspor
dc.subjectNatural Language Processingpor
dc.subjectmachine learningpor
dc.subjectdeep learningpor
dc.subjectpersuasion attemptspor
dc.subjectsocial networkspor
dc.titleDetecting Persuasion Attempts on Social Networks: Unearthing the Potential of Loss Functions and Text Pre-Processing in Imbalanced Data Settingspor
dc.typearticlepor

Files

Original bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
electronics-12-04447.pdf
Size:
313.73 KB
Format:
Adobe Portable Document Format

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
3.89 KB
Format:
Item-specific license agreed upon to submission
Description: