Agent systems based on large language models (LLMs) are increasingly deployed for autonomous tasks, yet existing evaluations mostly focus on task success rather than whether agents know when to abstain. This gap poses real risks: under ambiguity, conflicting constraints, or tool…
This study investigated the extent to which subjectively and objectively measured street-level perceptions complement or conflict with each other in explaining property value. Street-scene perceptions can be subjectively assessed from self-reported survey questions, or objectivel…