Saved in:
| Main Authors: | Wang, Lei, Zhong, Yujie, Sun, Xiaopeng, Cheng, Jingchun, Feng, Chengjian, Cao, Qiong, Ma, Lin, Fan, Zhaoxin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2504.00394 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
UniMD: Towards Unifying Moment Retrieval and Temporal Action Detection
by: Zeng, Yingsen, et al.
Published: (2024)
by: Zeng, Yingsen, et al.
Published: (2024)
PoseBench: Benchmarking the Robustness of Pose Estimation Models under Corruptions
by: Ma, Sihan, et al.
Published: (2024)
by: Ma, Sihan, et al.
Published: (2024)
InstaGen: Enhancing Object Detection by Training on Synthetic Dataset
by: Feng, Chengjian, et al.
Published: (2024)
by: Feng, Chengjian, et al.
Published: (2024)
RoboTron-Sim: Improving Real-World Driving via Simulated Hard-Case
by: Xiao, Baihui, et al.
Published: (2025)
by: Xiao, Baihui, et al.
Published: (2025)
RFSR: Improving ISR Diffusion Models via Reward Feedback Learning
by: Sun, Xiaopeng, et al.
Published: (2024)
by: Sun, Xiaopeng, et al.
Published: (2024)
LatentStealth: Unnoticeable and Efficient Adversarial Attacks on Expressive Human Pose and Shape Estimation
by: Li, Zhiying, et al.
Published: (2025)
by: Li, Zhiying, et al.
Published: (2025)
Pose-RFT: Enhancing MLLMs for 3D Pose Generation via Hybrid Action Reinforcement Fine-Tuning
by: Li, Bao, et al.
Published: (2025)
by: Li, Bao, et al.
Published: (2025)
DiffPose-Animal: A Language-Conditioned Diffusion Framework for Animal Pose Estimation
by: Xiong, Tianyu, et al.
Published: (2025)
by: Xiong, Tianyu, et al.
Published: (2025)
DisTime: Distribution-based Time Representation for Video Large Language Models
by: Zeng, Yingsen, et al.
Published: (2025)
by: Zeng, Yingsen, et al.
Published: (2025)
InstructVEdit: A Holistic Approach for Instructional Video Editing
by: Zhang, Chi, et al.
Published: (2025)
by: Zhang, Chi, et al.
Published: (2025)
RoboTron-Drive: All-in-One Large Multimodal Model for Autonomous Driving
by: Huang, Zhijian, et al.
Published: (2024)
by: Huang, Zhijian, et al.
Published: (2024)
MIGC++: Advanced Multi-Instance Generation Controller for Image Synthesis
by: Zhou, Dewei, et al.
Published: (2024)
by: Zhou, Dewei, et al.
Published: (2024)
DisPose: Disentangling Pose Guidance for Controllable Human Image Animation
by: Li, Hongxiang, et al.
Published: (2024)
by: Li, Hongxiang, et al.
Published: (2024)
Probabilistic Prompt Distribution Learning for Animal Pose Estimation
by: Rao, Jiyong, et al.
Published: (2025)
by: Rao, Jiyong, et al.
Published: (2025)
STEP: Simultaneous Tracking and Estimation of Pose for Animals and Humans
by: Verma, Shashikant, et al.
Published: (2025)
by: Verma, Shashikant, et al.
Published: (2025)
PoseTraj: Pose-Aware Trajectory Control in Video Diffusion
by: Ji, Longbin, et al.
Published: (2025)
by: Ji, Longbin, et al.
Published: (2025)
AniMer: Animal Pose and Shape Estimation Using Family Aware Transformer
by: Lyu, Jin, et al.
Published: (2024)
by: Lyu, Jin, et al.
Published: (2024)
RepLDM: Reprogramming Pretrained Latent Diffusion Models for High-Quality, High-Efficiency, High-Resolution Image Generation
by: Cao, Boyuan, et al.
Published: (2024)
by: Cao, Boyuan, et al.
Published: (2024)
HOLO: Homography-Guided Pose Estimator Network for Fine-Grained Visual Localization on SD Maps
by: Zhong, Xuchang, et al.
Published: (2026)
by: Zhong, Xuchang, et al.
Published: (2026)
CAP-IQA: Context-Aware Prompt-Guided CT Image Quality Assessment
by: Rifa, Kazi Ramisa, et al.
Published: (2026)
by: Rifa, Kazi Ramisa, et al.
Published: (2026)
PAGCNet: A Pose-Aware and Geometry Constrained Framework for Panoramic Depth Estimation
by: Ning, Kanglin, et al.
Published: (2025)
by: Ning, Kanglin, et al.
Published: (2025)
RoboTron-Nav: A Unified Framework for Embodied Navigation Integrating Perception, Planning, and Prediction
by: Zhong, Yufeng, et al.
Published: (2025)
by: Zhong, Yufeng, et al.
Published: (2025)
Boosting Robotic Manipulation Generalization with Minimal Costly Data
by: Zheng, Liming, et al.
Published: (2025)
by: Zheng, Liming, et al.
Published: (2025)
UnCageNet: Tracking and Pose Estimation of Caged Animal
by: Dutta, Sayak, et al.
Published: (2025)
by: Dutta, Sayak, et al.
Published: (2025)
CAP-Net: A Unified Network for 6D Pose and Size Estimation of Categorical Articulated Parts from a Single RGB-D Image
by: Huang, Jingshun, et al.
Published: (2025)
by: Huang, Jingshun, et al.
Published: (2025)
Benchmarking Ultra-High-Definition Image Reflection Removal
by: Zhang, Zhenyuan, et al.
Published: (2023)
by: Zhang, Zhenyuan, et al.
Published: (2023)
DiTalker: A Unified DiT-based Framework for High-Quality and Speaking Styles Controllable Portrait Animation
by: Feng, He, et al.
Published: (2025)
by: Feng, He, et al.
Published: (2025)
MR.CAP: Multi-Robot Joint Control and Planning for Object Transport
by: Jaafar, Hussein Ali, et al.
Published: (2024)
by: Jaafar, Hussein Ali, et al.
Published: (2024)
FDCE-Net: Underwater Image Enhancement with Embedding Frequency and Dual Color Encoder
by: Cheng, Zheng, et al.
Published: (2024)
by: Cheng, Zheng, et al.
Published: (2024)
PoseCrafter: One-Shot Personalized Video Synthesis Following Flexible Pose Control
by: Zhong, Yong, et al.
Published: (2024)
by: Zhong, Yong, et al.
Published: (2024)
Always Clear Depth: Robust Monocular Depth Estimation under Adverse Weather
by: Jiang, Kui, et al.
Published: (2025)
by: Jiang, Kui, et al.
Published: (2025)
High-Level Synthesis of Efficient Pipelines with Visibility Control
by: Anonymous
Published: (2026)
by: Anonymous
Published: (2026)
Advancing Pose-Guided Image Synthesis with Progressive Conditional Diffusion Models
by: Shen, Fei, et al.
Published: (2023)
by: Shen, Fei, et al.
Published: (2023)
One-Shot Learning for Pose-Guided Person Image Synthesis in the Wild
by: Fan, Dongqi, et al.
Published: (2024)
by: Fan, Dongqi, et al.
Published: (2024)
Stochastic Optimal Control of Multi‐Machine Power Systems Driven by Fractional Gaussian Noise
by: Chengjian Zhang, et al.
Published: (2025)
by: Chengjian Zhang, et al.
Published: (2025)
CAP: Controllable Alignment Prompting for Unlearning in LLMs
by: Wang, Zhaokun, et al.
Published: (2026)
by: Wang, Zhaokun, et al.
Published: (2026)
CAP: Evaluation of Persuasive and Creative Image Generation
by: Aghazadeh, Aysan, et al.
Published: (2024)
by: Aghazadeh, Aysan, et al.
Published: (2024)
Pose-Robust Calibration Strategy for Point-of-Gaze Estimation on Mobile Phones
by: Zhao, Yujie, et al.
Published: (2025)
by: Zhao, Yujie, et al.
Published: (2025)
3-Dimensional CryoEM Pose Estimation and Shift Correction Pipeline
by: Shah, Kaishva Chintan, et al.
Published: (2025)
by: Shah, Kaishva Chintan, et al.
Published: (2025)
QualityFlow: An Agentic Workflow for Program Synthesis Controlled by LLM Quality Checks
by: Hu, Yaojie, et al.
Published: (2025)
by: Hu, Yaojie, et al.
Published: (2025)
Similar Items
-
UniMD: Towards Unifying Moment Retrieval and Temporal Action Detection
by: Zeng, Yingsen, et al.
Published: (2024) -
PoseBench: Benchmarking the Robustness of Pose Estimation Models under Corruptions
by: Ma, Sihan, et al.
Published: (2024) -
InstaGen: Enhancing Object Detection by Training on Synthetic Dataset
by: Feng, Chengjian, et al.
Published: (2024) -
RoboTron-Sim: Improving Real-World Driving via Simulated Hard-Case
by: Xiao, Baihui, et al.
Published: (2025) -
RFSR: Improving ISR Diffusion Models via Reward Feedback Learning
by: Sun, Xiaopeng, et al.
Published: (2024)