On one SLR update dataset, ML classifiers achieved an F-score of 0.33 and could discard 33.9% of studies at 100% recall, but human-human reviewer pairs still outperformed human-ML pairs.
Externalising tacit knowledge of the system- atic review process,
1 Pith paper cite this work. Polarity classification is still indexing.
1
Pith paper citing it
citation-role summary
background 1
citation-polarity summary
fields
cs.SE 1years
2025 1verdicts
CONDITIONAL 1roles
background 1polarities
unclear 1representative citing papers
citing papers explorer
-
Can Machine Learning Support the Selection of Studies for Systematic Literature Review Updates?
On one SLR update dataset, ML classifiers achieved an F-score of 0.33 and could discard 33.9% of studies at 100% recall, but human-human reviewer pairs still outperformed human-ML pairs.