Camouflage image generation (CIG) focuses on generating visually concealed objects that seamlessly blend into their backgrounds. Existing methods typically follow either background-guided paradigms that adapt object appearance via style transfer, or foreground-guided strategies t…
Video multimodal large language models (MLLMs) can describe what happens in a video, but rarely identify when the supporting evidence occurs. We study generalist video temporal grounding, in which one model predicts a variable-cardinality set of evidence intervals across video le…
Recent advances in video understanding have spanned motion, long video, and streaming interaction, driving this field toward real-world applications. Despite this progress, current open-source models remain limited in several ways. They often struggle to generalize across diverse…
ABSTRACT The development of efficient catalysts for nitrogen conversion to ammonia is critical for a sustainable alternative to the energy‐intensive Haber–Bosch process. Yet, rational catalyst design remains highly challenging, compounded by complex structure–function relationshi…
As extreme weather events become more frequent, the icing of transmission lines in winter has become more common, causing significant economic losses to power systems and drawing increasing attention. However, owing to the complexity of the conductor icing process, establishing h…