This study introduces the task of predicting the incivility of conversations following replies to hate speech. We first propose a metric to measure conversation incivility based on the number of civil and uncivil comments as well as the unique authors involved in the discourse. Our metric approximates human judgments more accurately than previous metrics. We then use the metric to evaluate the outcomes of replies to hate speech. A linguistic analysis uncovers the differences in the language of replies that elicit follow-up conversations with high and low incivility. Experimental results show that forecasting incivility is challenging. We close with a qualitative analysis shedding light into the most common errors made by the best model.https://arxiv.org/abs/2312.04804Share this:FacebookXLike this:Like Loading... Post navigation Self-censorship among online harassment targets: the role of support at work, harassment characteristics, and the target’s public visibility (Information, Communication & Society) CReHate: Cross-cultural Re-annotation of English Hate Speech Dataset (arXiv Vanity)