DomainAdaptive Pre-training for Hate Speech Detection with BERT
Summary The core challenge presented is a common architectural dilemma: leveraging unlabelled in-domain data to augment a supervised learning task (hate speech detection) when the existing labelled dataset is already deemed sufficient for standard training. The goal is to move beyond simple supervised learning and utilize semi-supervised learning or domain adaptation techniques to improve model … Read more