3D dense captioning, an emerging vision-language task, aims to generate descriptive sentences for each object in the 3D scene. Despite the impressive results achieved by previous methods, they suffer from two limitations. First, current research often employs global rigid transfo…
We present Qwen-Image-2.0-RL, a post-training pipeline that applies reinforcement learning from human feedback (RLHF) and on-policy distillation (OPD) to improve both the visual quality and instruction-following capability of the Qwen-Image-2.0 diffusion model. To provide reliabl…
The tight sandstone gas reservoirs in the LX area of the Ordos Basin are characterized by low porosity, poor permeability, and strong heterogeneity, which significantly complicate fluid type identification. Conventional methods based on petrophysical logging and core analysis hav…