Incivility Detection in Open Source Code Review and Issue Discussions

Ferreira, Isabella; Rafiq, Ahlaam; Cheng, Jinghui

doi:10.1016/j.jss.2023.111935

Computer Science > Software Engineering

arXiv:2206.13429 (cs)

[Submitted on 27 Jun 2022 (v1), last revised 19 Dec 2023 (this version, v2)]

Title:Incivility Detection in Open Source Code Review and Issue Discussions

Authors:Isabella Ferreira, Ahlaam Rafiq, Jinghui Cheng

View PDF HTML (experimental)

Abstract:Given the democratic nature of open source development, code review and issue discussions may be uncivil. Incivility, defined as features of discussion that convey an unnecessarily disrespectful tone, can have negative consequences to open source communities. To prevent or minimize these negative consequences, open source platforms have included mechanisms for removing uncivil language from the discussions. However, such approaches require manual inspection, which can be overwhelming given the large number of discussions. To help open source communities deal with this problem, in this paper, we aim to compare six classical machine learning models with BERT to detect incivility in open source code review and issue discussions. Furthermore, we assess if adding contextual information improves the models' performance and how well the models perform in a cross-platform setting. We found that BERT performs better than classical machine learning models, with a best F1-score of 0.95. Furthermore, classical machine learning models tend to underperform to detect non-technical and civil discussions. Our results show that adding the contextual information to BERT did not improve its performance and that none of the analyzed classifiers had an outstanding performance in a cross-platform setting. Finally, we provide insights into the tones that the classifiers misclassify.

Comments:	18 pages
Subjects:	Software Engineering (cs.SE)
Cite as:	arXiv:2206.13429 [cs.SE]
	(or arXiv:2206.13429v2 [cs.SE] for this version)
	https://doi.org/10.48550/arXiv.2206.13429
Related DOI:	https://doi.org/10.1016/j.jss.2023.111935

Submission history

From: Isabella Ferreira [view email]
[v1] Mon, 27 Jun 2022 16:26:18 UTC (5,033 KB)
[v2] Tue, 19 Dec 2023 02:39:41 UTC (6,173 KB)

Computer Science > Software Engineering

Title:Incivility Detection in Open Source Code Review and Issue Discussions

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Software Engineering

Title:Incivility Detection in Open Source Code Review and Issue Discussions

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators