CORTEXA
← Browse

Ming Hu

7 papers indexed

arxivcs.CV2026-07-23

WhereEdit: Mask-aware Local Latent Editing for One-Step Image Editing

Ming Hu, Mingyu Dou, Jianfu Yin, Miaomiao Zhang, Cong Hu, Yao Wang, et al.

Recent one-step text-to-image (T2I) models enable efficient image synthesis and provide new opportunities for real-time image editing. However, existing one-step editing methods primarily rely on text conditioning for semantic transformation, lacking explicit spatial control over…

View free PDFSource page
arxivcs.CV2026-07-15

SIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement Learning

Cheng Tang, Junzhi Ning, Min Cen, Wei Li, Xinyi Zeng, Pinxian Zeng, et al.

Reinforcement learning with verifiable rewards (RLVR) drives multimodal reasoning, but answer-level correctness does not guarantee that a vision-language model grounds its predictions in visual evidence. Existing visual-intervention methods contrast policy behavior on original an…

View free PDFSource page
arxivcs.AI2026-07-14

Evidence-Grounded AI for Musculoskeletal Care

Wenjie Li, Yujie Zhang, Fanrui Zhang, Haoran Sun, Renhao Yang, Junjun He, et al.

Musculoskeletal diseases are among the leading causes of disability and drive the greatest global need for rehabilitation. Because recovery, remodelling and degeneration of bones, joints and related tissues unfold over months to years, care requires longitudinal management rather…

View free PDFSource page
arxivcs.CV2026-07-09

UAV-OVVIS: Unmanned Aerial Vehicles Also Need Open-Vocabulary Video Instance Segmentation

Mingyu Dou, Shi Qiu, Ming Hu, Yifan Chen, Zhe Sun

Unmanned Aerial Vehicle (UAV) videos are widely used in traffic monitoring, urban management, and emergency rescue. However, existing UAV video perception is largely limited to box-level detection and tracking over predefined categories, making it difficult to jointly support fle…

View free PDFSource page
arxivcs.ROcs.AI2026-07-06

Geometry-Aware Motion Latents for Learning Robust Manipulation Policies

Yunchao Zhang, Yijia Weng, Ruizhe Liu, Ming Hu, Leonidas Guibas, Yanchao Yang

Learning motion latents for robotic manipulation heavily relies on extracting motion patterns from visual sequences, yet effective action abstractions require understanding three-dimensional geometric transformations. Here, we introduce GeoMoLa (Geometry-Aware Motion Latents), wh…

View free PDFSource page
arxivecon.THcs.AIcs.CYcs.GTcs.HC2026-07-06

Strategic Buying Agents

Mingyang Fu, Ming Hu

Agentic AI is shifting online shopping from search toward delegated purchasing, where autonomous buying agents monitor markets and decide when to buy on a consumer's behalf. We study the design of such strategic buying agents, which must decide when to purchase within a finite sh…

View free PDFSource page
crossrefAdvanced Intelligent Discovery2026-06-30

Probing Machine Learning Interatomic Potentials on Ion Transport Properties

Ogheneyoma Aghoghovbia, Ming Hu, Adji Bousso Dieng

Machine learning interatomic potentials (MLPs) are promising for accelerating the simulation of ion transport in all‐solid‐state battery materials, but their accuracy across diverse material compositions and symmetries remains unquantified. Here, we systematically benchmark six s…

View free PDFSource page