Saved in:
| Main Authors: | Zhu, Chen, Huang, Buzhen, Wu, Zijing, Zuo, Binghui, Wang, Yangang |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2508.06093 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Simultaneously Recovering Multi-Person Meshes and Multi-View Cameras with Human Semantics
by: Huang, Buzhen, et al.
Published: (2024)
by: Huang, Buzhen, et al.
Published: (2024)
Adapting Human Mesh Recovery with Vision-Language Feedback
by: Xu, Chongyang, et al.
Published: (2025)
by: Xu, Chongyang, et al.
Published: (2025)
Closely Interactive Human Reconstruction with Proxemics and Physics-Guided Adaption
by: Huang, Buzhen, et al.
Published: (2024)
by: Huang, Buzhen, et al.
Published: (2024)
Reconstructing Close Human Interaction with Appearance and Proxemics Reasoning
by: Huang, Buzhen, et al.
Published: (2025)
by: Huang, Buzhen, et al.
Published: (2025)
Synthesizing Physically Plausible Human Motions in 3D Scenes
by: Pan, Liang, et al.
Published: (2023)
by: Pan, Liang, et al.
Published: (2023)
Generalizable Human Gaussians from Single-View Image
by: Chen, Jinnan, et al.
Published: (2024)
by: Chen, Jinnan, et al.
Published: (2024)
TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization
by: Pan, Liang, et al.
Published: (2025)
by: Pan, Liang, et al.
Published: (2025)
ReactDiff: Latent Diffusion for Facial Reaction Generation
by: Li, Jiaming, et al.
Published: (2025)
by: Li, Jiaming, et al.
Published: (2025)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
by: Xu, Jiahao, et al.
Published: (2026)
by: Xu, Jiahao, et al.
Published: (2026)
Ready-to-React: Online Reaction Policy for Two-Character Interaction Generation
by: Cen, Zhi, et al.
Published: (2025)
by: Cen, Zhi, et al.
Published: (2025)
Knowledge NeRF: Few-shot Novel View Synthesis for Dynamic Articulated Objects
by: Cai, Wenxiao, et al.
Published: (2024)
by: Cai, Wenxiao, et al.
Published: (2024)
ShoeModel: Learning to Wear on the User-specified Shoes via Diffusion Model
by: Chen, Binghui, et al.
Published: (2024)
by: Chen, Binghui, et al.
Published: (2024)
ReGenNet: Towards Human Action-Reaction Synthesis
by: Xu, Liang, et al.
Published: (2024)
by: Xu, Liang, et al.
Published: (2024)
Toward Human Understanding with Controllable Synthesis
by: Cuevas-Velasquez, Hanz, et al.
Published: (2024)
by: Cuevas-Velasquez, Hanz, et al.
Published: (2024)
VirtualModel: Generating Object-ID-retentive Human-object Interaction Image by Diffusion Model for E-commerce Marketing
by: Chen, Binghui, et al.
Published: (2024)
by: Chen, Binghui, et al.
Published: (2024)
Artifact-Aware Evaluation for High-Quality Video Generation
by: Zhu, Chen, et al.
Published: (2026)
by: Zhu, Chen, et al.
Published: (2026)
Open-Set Video-based Facial Expression Recognition with Human Expression-sensitive Prompting
by: Liu, Yuanyuan, et al.
Published: (2024)
by: Liu, Yuanyuan, et al.
Published: (2024)
ReactDiff: Fundamental Multiple Appropriate Facial Reaction Diffusion Model
by: Cheng, Luo, et al.
Published: (2025)
by: Cheng, Luo, et al.
Published: (2025)
FLORA: Formal Language Model Enables Robust Training-free Zero-shot Object Referring Analysis
by: Chen, Zhe, et al.
Published: (2025)
by: Chen, Zhe, et al.
Published: (2025)
ReactFace: Online Multiple Appropriate Facial Reaction Generation in Dyadic Interactions
by: Luo, Cheng, et al.
Published: (2023)
by: Luo, Cheng, et al.
Published: (2023)
Smile upon the Face but Sadness in the Eyes: Emotion Recognition based on Facial Expressions and Eye Behaviors
by: Liu, Yuanyuan, et al.
Published: (2024)
by: Liu, Yuanyuan, et al.
Published: (2024)
More Is Better: A MoE-Based Emotion Recognition Framework with Human Preference Alignment
by: Xie, Jun, et al.
Published: (2025)
by: Xie, Jun, et al.
Published: (2025)
Strictly-ID-Preserved and Controllable Accessory Advertising Image Generation
by: Xue, Youze, et al.
Published: (2024)
by: Xue, Youze, et al.
Published: (2024)
MoReact: Generating Reactive Motion from Textual Descriptions
by: Xu, Xiyan, et al.
Published: (2025)
by: Xu, Xiyan, et al.
Published: (2025)
Controllable Human-Object Interaction Synthesis
by: Li, Jiaman, et al.
Published: (2023)
by: Li, Jiaman, et al.
Published: (2023)
AnyStory: Towards Unified Single and Multiple Subject Personalization in Text-to-Image Generation
by: He, Junjie, et al.
Published: (2025)
by: He, Junjie, et al.
Published: (2025)
ReactBench: A Cause-Driven Benchmark for Multimodal Hallucination via Systematic Evaluation
by: Zhou, Shizhe, et al.
Published: (2026)
by: Zhou, Shizhe, et al.
Published: (2026)
AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
by: Yang, Dejie, et al.
Published: (2025)
by: Yang, Dejie, et al.
Published: (2025)
Toward Realistic Camouflaged Object Detection: Benchmarks and Method
by: Xin, Zhimeng, et al.
Published: (2025)
by: Xin, Zhimeng, et al.
Published: (2025)
Video2Layout: Recall and Reconstruct Metric-Grounded Cognitive Map for Spatial Reasoning
by: Huang, Yibin, et al.
Published: (2025)
by: Huang, Yibin, et al.
Published: (2025)
Contrastive Multi-Modal Hypergraph Reasoning for 3D Crowd Mesh Recovery
by: Sun, Minghao, et al.
Published: (2026)
by: Sun, Minghao, et al.
Published: (2026)
Implicit Guidance and Explicit Representation of Semantic Information in Points Cloud: A Survey
by: Tang, Jingyuan, et al.
Published: (2025)
by: Tang, Jingyuan, et al.
Published: (2025)
MARRS: Masked Autoregressive Unit-based Reaction Synthesis
by: Wang, Yabiao, et al.
Published: (2025)
by: Wang, Yabiao, et al.
Published: (2025)
Expandable Residual Approximation for Knowledge Distillation
by: Yan, Zhaoyi, et al.
Published: (2025)
by: Yan, Zhaoyi, et al.
Published: (2025)
Investigating Domain Gaps for Indoor 3D Object Detection
by: Zhao, Zijing, et al.
Published: (2025)
by: Zhao, Zijing, et al.
Published: (2025)
Virtual Classification: Modulating Domain-Specific Knowledge for Multidomain Crowd Counting
by: Guo, Mingyue, et al.
Published: (2024)
by: Guo, Mingyue, et al.
Published: (2024)
AdaptiveDrag: Semantic-Driven Dragging on Diffusion-Based Image Editing
by: Chen, DuoSheng, et al.
Published: (2024)
by: Chen, DuoSheng, et al.
Published: (2024)
Smile on the Face, Sadness in the Eyes: Bridging the Emotion Gap with a Multimodal Dataset of Eye and Facial Behaviors
by: Liu, Kejun, et al.
Published: (2025)
by: Liu, Kejun, et al.
Published: (2025)
HumanDreamer: Generating Controllable Human-Motion Videos via Decoupled Generation
by: Wang, Boyuan, et al.
Published: (2025)
by: Wang, Boyuan, et al.
Published: (2025)
Seeing Through Reflections: Advancing 3D Scene Reconstruction in Mirror-Containing Environments with Gaussian Splatting
by: Guo, Zijing, et al.
Published: (2025)
by: Guo, Zijing, et al.
Published: (2025)
Similar Items
-
Simultaneously Recovering Multi-Person Meshes and Multi-View Cameras with Human Semantics
by: Huang, Buzhen, et al.
Published: (2024) -
Adapting Human Mesh Recovery with Vision-Language Feedback
by: Xu, Chongyang, et al.
Published: (2025) -
Closely Interactive Human Reconstruction with Proxemics and Physics-Guided Adaption
by: Huang, Buzhen, et al.
Published: (2024) -
Reconstructing Close Human Interaction with Appearance and Proxemics Reasoning
by: Huang, Buzhen, et al.
Published: (2025) -
Synthesizing Physically Plausible Human Motions in 3D Scenes
by: Pan, Liang, et al.
Published: (2023)