CORTEXA
← Browse

Kangzhe Hu

1 paper indexed

arxivcs.AI2026-07-03

Reinforcement Learning for Evidence-Seeking Diagnostic Reasoning with Large Language Models

Shengyi Hua, Kangzhe Hu, Conghui He, Xiaofan Zhang, Shaoting Zhang

Recent reasoning-centric Large Language Models (LLMs) have made significant strides, yet they predominantly operate on a passive-inference pattern that assumes complete information. In contrast, real-world clinical intelligence is inherently an iterative investigative process req…

View free PDFSource page