Vision-and-Language Navigation (VLN) enables embodied agents to follow natural-language instructions. However, route-level instructions commonly encode spatial priors, such as orientation, distance, and layout, that are not explicitly available from onboard sensing at deployment…
The two-tower model has been widely used for large-scale recommendation systems, particularly in the retrieval stage. Industry standards for training two-tower models typically involve in-batch and/or out-of-batch negative sampling. However, these methods often produce easy negat…
MG-SpaIR is a training-data-free framework for restoring a clean image from a single observation corrupted by a mixture of blur, downsampling, noise, and missing pixels. Building on implicit neural representations (INRs), we introduce a multi-grade coarse-to-fine residual hierarc…
ABSTRACT Calcium‐based liquid metal batteries are promising for large‐scale energy storage due to calcium abundance and low cost, yet their practical applications are impeded by high operating temperatures, severe self‐discharge, limited coulombic efficiency, and rapid capacity f…
Abstract Atmosphere‐ocean‐land coupled forecasting systems, despite their comprehensiveness, face substantial challenges in the “predictability desert” at subseasonal to seasonal (S2S) timescales, particularly for precipitation—a variable crucial for socioeconomic activities yet…