CORTEXA
← Browse

Andong Hua

1 paper indexed

arxivcs.CLcs.AI2026-07-14

From Critic to Confidence: PPO for Language-Based Quantitative Prediction with Confidence Estimation

Mehak Dhaliwal, Rasta Tadayon, Andong Hua, Haewon Jeong, Yao Qin

LLMs can perform language-based quantitative prediction from unstructured inputs, but remain susceptible to hallucinations and overconfident errors, making it critical to know not only what a model predicts, but when its predictions can be trusted. We introduce CARE-PPO, a reinfo…

View free PDFSource page