Resources


Database Credentialed Access

CXReasonBench: A Benchmark for Evaluating Structured Diagnostic Reasoning in Chest X-rays

Hyungyung Lee, Geon Choi, Jung Oh Lee, Hangyul Yoon, Hyuk Gi Hong, Edward Choi

CheXStruct is an automated pipeline that derives structured diagnostic reasoning steps from chest X-rays. CXReasonBench builds on this to evaluate whether models perform clinically grounded, multi-step reasoning beyond final diagnoses.

chest x-ray evaluation benchmark structured diagnostic pipeline structured chest x-ray qa diagnostic reasoning intermediate reasoning steps grounded reasoning structured reasoning

Published: Oct. 15, 2025. Version: 1.0.0


Database Credentialed Access

CXReasonBench: A Benchmark for Evaluating Structured Diagnostic Reasoning in Chest X-rays

Hyungyung Lee, Geon Choi, Jung Oh Lee, Hangyul Yoon, Hyuk Gi Hong, Edward Choi

CheXStruct is an automated pipeline that derives structured diagnostic reasoning steps from chest X-rays. CXReasonBench builds on this to evaluate whether models perform clinically grounded, multi-step reasoning beyond final diagnoses.

chest x-ray evaluation benchmark structured diagnostic pipeline structured chest x-ray qa diagnostic reasoning intermediate reasoning steps grounded reasoning structured reasoning

Published: Oct. 15, 2025. Version: 1.0.0