Abstract. Land models typically represent the entire fine-root system as a single homogeneous pool, collapsing structural and functional complexity into one set of parameters. This simplification weakens the belowground feedbacks that balance source-driven carbon dynamics. Here,…
Vision-language agents (VLAs) are increasingly used to interpret complex driving scenes and support safety-critical reasoning. This report presents the CVPR 2026@AdvML Workshop Challenge on adversarial multimodal attacks against autonomous-driving VLAs. Built on DriveLM-style mul…
In this paper, we propose MLLM-DataEngine, a novel closed-loop system that bridges data generation, model training, and evaluation. Within each loop iteration, the MLLM-DataEngine first analyzes the weakness of the model based on the evaluation results, then generates a proper in…
Depression screening from large-scale behavioral data is challenged by fragmented circadian indicators, limited interpretability, and the lack of intervention-oriented analysis. Existing approaches typically analyze sleep, activity, and social behaviors in isolation, failing to c…
3D Gaussian Splatting (3DGS) enables real-time novel view synthesis for static scenes. Extending it to dynamic scenes via deformation fields has recently attracted significant attention, particularly for dynamic scene reconstructionband distractor-free. However, existing deformat…
The traditional design method for terahertz metasurface biosensors is cumbersome and time-consuming, requires expertise, and often leads to significant discrepancies between expected and actual values. This paper presents a novel approach for the fast, efficient, and convenient i…