Use of Large Language Models to Assist Risk of Bias Assessment Supported by an Implementation Document
Introduction Recent studies showed poor performance of large language models (LLM) for assessing risk of bias (RoB) with the RoB2 tool. This is in line with the low reliability that humans have in assessing RoB. However, the use of an implementation document (ID) prepared by expert reviewers – i.e., a standardised doc...