preprint Open access

FairStream : Dynamic Bias Mitigation for Real-Time NLP in Social Media Moderation

Research footprint

At a glance

Citations
0
References
19
Comments
0
Paper overview

Abstract

Social media platforms demand real-time content moderation to curb toxic content, yet streaming NLP models often perpetuate biases, unfairly targeting specific demographics. This research introduces FairStream, a novel bias correction algorithm that dynamically adjusts model outputs using fairness-aware embeddings and adversarial training. Operating in microseconds, FairStream ensures equitable moderation across diverse user groups. Evaluated on a unique dataset of 50,000 social media posts, the approach achieves a 95% accuracy in toxic content detection while reducing bias by 70% compared to baseline models. This work advances fair and efficient moderation, fostering inclusive online environments.

Record transparency

Publication details

DOI
10.36227/techrxiv.174803925.58646463/v1
OpenAlex
W4410635503
Document type
preprint
Language
EN
Last metadata update
Community

Comments

Log in to join the discussion.

  1. No comments yet. Start the discussion.