This is the dataset of AbdomenCT-1K: Weakly Supervised Learning Benchmark. Related paper: https://ieeexplore.ieee.org/document/9497733/ Benchmark homepage: https://abdomenct-1k-weaklysupervisedlea
These scores were compiled as part of a study which compared ChatGPT’s performance with real doctors on the Swedish family medicine licensing exam. The scores from zero to ten for the cases of exam y