BI-RADS Category Assignments by GPT-3.5, GPT-4, and Google Bard: A Multilanguage Study
Based on breast imaging reports written in three languages, GPT-3.5, GPT-4, and Google Bard achieved moderate agreement with human reader–assigned Breast Imaging Reporting and Data System categories but also yielded a high percentage of discordant category assignments that would negatively impact clinical management.
