Saved in:
| Main Authors: | Zheng, Dian, Zhang, Cheng, Wu, Xiao-Ming, Li, Cao, Lv, Chengfei, Hu, Jian-Fang, Zheng, Wei-Shi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2503.18420 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ProEdit: Inversion-based Editing From Prompts Done Right
by: Ouyang, Zhi, et al.
Published: (2025)
by: Ouyang, Zhi, et al.
Published: (2025)
Selective Hourglass Mapping for Universal Image Restoration Based on Diffusion Model
by: Zheng, Dian, et al.
Published: (2024)
by: Zheng, Dian, et al.
Published: (2024)
SpatialDreamer: Self-supervised Stereo Video Synthesis from Monocular Input
by: Lv, Zhen, et al.
Published: (2024)
by: Lv, Zhen, et al.
Published: (2024)
An Economic Framework for 6-DoF Grasp Detection
by: Wu, Xiao-Ming, et al.
Published: (2024)
by: Wu, Xiao-Ming, et al.
Published: (2024)
Progressive Human Motion Generation Based on Text and Few Motion Frames
by: Zeng, Ling-An, et al.
Published: (2025)
by: Zeng, Ling-An, et al.
Published: (2025)
MoRight: Motion Control Done Right
by: Liu, Shaowei, et al.
Published: (2026)
by: Liu, Shaowei, et al.
Published: (2026)
Test-Time Training Done Right
by: Zhang, Tianyuan, et al.
Published: (2025)
by: Zhang, Tianyuan, et al.
Published: (2025)
Classification Done Right for Vision-Language Pre-Training
by: Huang, Zilong, et al.
Published: (2024)
by: Huang, Zilong, et al.
Published: (2024)
VAR RL Done Right: Tackling Asynchronous Policy Conflicts in Visual Autoregressive Generation
by: Sun, Shikun, et al.
Published: (2026)
by: Sun, Shikun, et al.
Published: (2026)
IRG-MotionLLM: Interleaving Motion Generation, Assessment and Refinement for Text-to-Motion Generation
by: Li, Yuan-Ming, et al.
Published: (2025)
by: Li, Yuan-Ming, et al.
Published: (2025)
Taming Stable Diffusion for Text to 360° Panorama Image Generation
by: Zhang, Cheng, et al.
Published: (2024)
by: Zhang, Cheng, et al.
Published: (2024)
Decoupled Distillation to Erase: A General Unlearning Method for Any Class-centric Tasks
by: Zhou, Yu, et al.
Published: (2025)
by: Zhou, Yu, et al.
Published: (2025)
Elastic Weight Consolidation Done Right for Continual Learning
by: Liu, Xuan, et al.
Published: (2026)
by: Liu, Xuan, et al.
Published: (2026)
Refer-Agent: A Collaborative Multi-Agent System with Reasoning and Reflection for Referring Video Object Segmentation
by: Jiang, Haichao, et al.
Published: (2026)
by: Jiang, Haichao, et al.
Published: (2026)
ReferDINO-Plus: 2nd Solution for 4th PVUW MeViS Challenge at CVPR 2025
by: Liang, Tianming, et al.
Published: (2025)
by: Liang, Tianming, et al.
Published: (2025)
Siamese Learning with Joint Alignment and Regression for Weakly-Supervised Video Paragraph Grounding
by: Tan, Chaolei, et al.
Published: (2024)
by: Tan, Chaolei, et al.
Published: (2024)
PraNet-V2: Dual-Supervised Reverse Attention for Medical Image Segmentation
by: Hu, Bo-Cheng, et al.
Published: (2025)
by: Hu, Bo-Cheng, et al.
Published: (2025)
Single-View Scene Point Cloud Human Grasp Generation
by: Wang, Yan-Kang, et al.
Published: (2024)
by: Wang, Yan-Kang, et al.
Published: (2024)
PanoDiffusion: 360-degree Panorama Outpainting via Diffusion
by: Wu, Tianhao, et al.
Published: (2023)
by: Wu, Tianhao, et al.
Published: (2023)
Image-to-Video Transfer Learning based on Image-Language Foundation Models: A Comprehensive Survey
by: Li, Jinxuan, et al.
Published: (2025)
by: Li, Jinxuan, et al.
Published: (2025)
HORIZON: High-Resolution Semantically Controlled Panorama Synthesis
by: Yan, Kun, et al.
Published: (2022)
by: Yan, Kun, et al.
Published: (2022)
Modeling Multiple Normal Action Representations for Error Detection in Procedural Tasks
by: Huang, Wei-Jin, et al.
Published: (2025)
by: Huang, Wei-Jin, et al.
Published: (2025)
HarmoGS: Robust 3D Gaussian Splatting in the Wild via Conflict-Aware Gradient Harmonization
by: Kang, Yulei, et al.
Published: (2026)
by: Kang, Yulei, et al.
Published: (2026)
Ranking Distillation for Open-Ended Video Question Answering with Insufficient Labels
by: Liang, Tianming, et al.
Published: (2024)
by: Liang, Tianming, et al.
Published: (2024)
TechCoach: Towards Technical-Point-Aware Descriptive Action Coaching
by: Li, Yuan-Ming, et al.
Published: (2024)
by: Li, Yuan-Ming, et al.
Published: (2024)
EchoGen: Cycle-Consistent Learning for Unified Layout-Image Generation and Understanding
by: Zou, Kai, et al.
Published: (2026)
by: Zou, Kai, et al.
Published: (2026)
Omnidirectional Spatial Modeling from Correlated Panoramas
by: Zhang, Xinshen, et al.
Published: (2025)
by: Zhang, Xinshen, et al.
Published: (2025)
360DVD: Controllable Panorama Video Generation with 360-Degree Video Diffusion Model
by: Wang, Qian, et al.
Published: (2024)
by: Wang, Qian, et al.
Published: (2024)
PixelFade: Privacy-preserving Person Re-identification with Noise-guided Progressive Replacement
by: Zhang, Delong, et al.
Published: (2024)
by: Zhang, Delong, et al.
Published: (2024)
Diffusion Once and Done: Degradation-Aware LoRA for Efficient All-in-One Image Restoration
by: Tang, Ni, et al.
Published: (2025)
by: Tang, Ni, et al.
Published: (2025)
PanoWorld: A Generative Spatial World Model for Consistent Whole-House Panorama Synthesis
by: Jia, Jinrang, et al.
Published: (2026)
by: Jia, Jinrang, et al.
Published: (2026)
Radiology Report Generation for Low-Quality X-Ray Images
by: Zhu, Hongze, et al.
Published: (2026)
by: Zhu, Hongze, et al.
Published: (2026)
Seg-ReSearch: Segmentation with Interleaved Reasoning and External Search
by: Liang, Tianming, et al.
Published: (2026)
by: Liang, Tianming, et al.
Published: (2026)
TaoAvatar: Real-Time Lifelike Full-Body Talking Avatars for Augmented Reality via 3D Gaussian Splatting
by: Chen, Jianchuan, et al.
Published: (2025)
by: Chen, Jianchuan, et al.
Published: (2025)
HyperDet: Generalizable Detection of Synthesized Images by Generating and Merging A Mixture of Hyper LoRAs
by: Cao, Huangsen, et al.
Published: (2024)
by: Cao, Huangsen, et al.
Published: (2024)
ReferDINO: Referring Video Object Segmentation with Visual Grounding Foundations
by: Liang, Tianming, et al.
Published: (2025)
by: Liang, Tianming, et al.
Published: (2025)
Rethinking The Training And Evaluation of Rich-Context Layout-to-Image Generation
by: Cheng, Jiaxin, et al.
Published: (2024)
by: Cheng, Jiaxin, et al.
Published: (2024)
FHAvatar: Fast and High-Fidelity Reconstruction of Face-and-Hair Composable 3D Head Avatar from Few Casual Captures
by: Sun, Yujie, et al.
Published: (2026)
by: Sun, Yujie, et al.
Published: (2026)
Personalized Image Generation with Deep Generative Models: A Decade Survey
by: Wei, Yuxiang, et al.
Published: (2025)
by: Wei, Yuxiang, et al.
Published: (2025)
The Consistency Critic: Correcting Inconsistencies in Generated Images via Reference-Guided Attentive Alignment
by: Ouyang, Ziheng, et al.
Published: (2025)
by: Ouyang, Ziheng, et al.
Published: (2025)
Similar Items
-
ProEdit: Inversion-based Editing From Prompts Done Right
by: Ouyang, Zhi, et al.
Published: (2025) -
Selective Hourglass Mapping for Universal Image Restoration Based on Diffusion Model
by: Zheng, Dian, et al.
Published: (2024) -
SpatialDreamer: Self-supervised Stereo Video Synthesis from Monocular Input
by: Lv, Zhen, et al.
Published: (2024) -
An Economic Framework for 6-DoF Grasp Detection
by: Wu, Xiao-Ming, et al.
Published: (2024) -
Progressive Human Motion Generation Based on Text and Few Motion Frames
by: Zeng, Ling-An, et al.
Published: (2025)