PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Gao, Mingju, Yang, Kaisen, Gao, Huan-ang, Li, Bohan, Ding, Ao, Li, Wenyi, Yu, Yangcheng, Liu, Jinkun, Xu, Shaocong, Niu, Yike, Chi, Haohan, Chen, Hao, Tang, Hao, Zhang, Yu, Yi, Li, Zhao, Hao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
von: Gao, Mingju, et al.
Veröffentlicht: (2025)
von: Gao, Mingju, et al.
Veröffentlicht: (2025)
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
von: Gao, Huan-ang, et al.
Veröffentlicht: (2024)
von: Gao, Huan-ang, et al.
Veröffentlicht: (2024)
Training-Free Model Merging for Multi-target Domain Adaptation
von: Li, Wenyi, et al.
Veröffentlicht: (2024)
von: Li, Wenyi, et al.
Veröffentlicht: (2024)
Benchmarking PhD-Level Coding in 3D Geometric Computer Vision
von: Li, Wenyi, et al.
Veröffentlicht: (2026)
von: Li, Wenyi, et al.
Veröffentlicht: (2026)
Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models
von: Chi, Haohan, et al.
Veröffentlicht: (2025)
von: Chi, Haohan, et al.
Veröffentlicht: (2025)
FairDiff: Fair Segmentation with Point-Image Diffusion
von: Li, Wenyi, et al.
Veröffentlicht: (2024)
von: Li, Wenyi, et al.
Veröffentlicht: (2024)
Alias-free 4D Gaussian Splatting
von: Chen, Zilong, et al.
Veröffentlicht: (2025)
von: Chen, Zilong, et al.
Veröffentlicht: (2025)
Challenger: Affordable Adversarial Driving Video Generation
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2025)
FB-4D: Spatial-Temporal Coherent Dynamic 3D Content Generation with Feature Banks
von: Li, Jinwei, et al.
Veröffentlicht: (2025)
von: Li, Jinwei, et al.
Veröffentlicht: (2025)
Dual-frame Fluid Motion Estimation with Test-time Optimization and Zero-divergence Loss
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
von: Zhang, Yifei, et al.
Veröffentlicht: (2024)
MagicPose4D: Crafting Articulated Models with Appearance and Motion Control
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
von: Zhang, Hao, et al.
Veröffentlicht: (2024)
TwinAligner: Visual-Dynamic Alignment Empowers Physics-aware Real2Sim2Real for Robotic Manipulation
von: Fan, Hongwei, et al.
Veröffentlicht: (2025)
von: Fan, Hongwei, et al.
Veröffentlicht: (2025)
One View, Many Worlds: Single-Image to 3D Object Meets Generative Domain Randomization for One-Shot 6D Pose Estimation
von: Geng, Zheng, et al.
Veröffentlicht: (2025)
von: Geng, Zheng, et al.
Veröffentlicht: (2025)
Real2Sim in HOI: Toward Physically Plausible HOI Reconstruction from Monocular Videos
von: Zhao, Yubo, et al.
Veröffentlicht: (2026)
von: Zhao, Yubo, et al.
Veröffentlicht: (2026)
An Infinite Family of Primitive Heron Triangles with Two Sides as Perfect Squares
von: Li, Yangcheng
Veröffentlicht: (2026)
von: Li, Yangcheng
Veröffentlicht: (2026)
A new perspective of arithmetic billiards
von: Li, Yangcheng
Veröffentlicht: (2023)
von: Li, Yangcheng
Veröffentlicht: (2023)
PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth
von: Jin, Bu, et al.
Veröffentlicht: (2025)
von: Jin, Bu, et al.
Veröffentlicht: (2025)
CubeBench: Diagnosing Interactive, Long-Horizon Spatial Reasoning Under Partial Observations
von: Gao, Huan-ang, et al.
Veröffentlicht: (2025)
von: Gao, Huan-ang, et al.
Veröffentlicht: (2025)
ArtHOI: Articulated Human-Object Interaction Synthesis by 4D Reconstruction from Video Priors
von: Huang, Zihao, et al.
Veröffentlicht: (2026)
von: Huang, Zihao, et al.
Veröffentlicht: (2026)
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2025)
von: Zhu, Xiaomeng, et al.
Veröffentlicht: (2025)
Ctrl-U: Robust Conditional Image Generation via Uncertainty-aware Reward Modeling
von: Zhang, Guiyu, et al.
Veröffentlicht: (2024)
von: Zhang, Guiyu, et al.
Veröffentlicht: (2024)
Event-based Motion & Appearance Fusion for 6D Object Pose Tracking
von: Li, Zhichao, et al.
Veröffentlicht: (2026)
von: Li, Zhichao, et al.
Veröffentlicht: (2026)
Shadows, Quasinormal Modes, and Optical Appearances of Black Holes in Horndeski Theory
von: Luo, Zhi, et al.
Veröffentlicht: (2024)
von: Luo, Zhi, et al.
Veröffentlicht: (2024)
AVD2: Accident Video Diffusion for Accident Video Description
von: Li, Cheng, et al.
Veröffentlicht: (2025)
von: Li, Cheng, et al.
Veröffentlicht: (2025)
Unifying Appearance Codes and Bilateral Grids for Driving Scene Gaussian Splatting
von: Wang, Nan, et al.
Veröffentlicht: (2025)
von: Wang, Nan, et al.
Veröffentlicht: (2025)
PersPose: 3D Human Pose Estimation with Perspective Encoding and Perspective Rotation
von: Hao, Xiaoyang, et al.
Veröffentlicht: (2025)
von: Hao, Xiaoyang, et al.
Veröffentlicht: (2025)
$(G,F)$-points on $\mathbb{Q}$-algebraic varieties
von: Li, Yangcheng, et al.
Veröffentlicht: (2025)
von: Li, Yangcheng, et al.
Veröffentlicht: (2025)
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
von: Liu, Jinkun, et al.
Veröffentlicht: (2026)
von: Liu, Jinkun, et al.
Veröffentlicht: (2026)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
Diffusion-based Visual Anagram as Multi-task Learning
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Xu, Zhiyuan, et al.
Veröffentlicht: (2024)
FlexAM: Flexible Appearance-Motion Decomposition for Versatile Video Generation Control
von: Sheng, Mingzhi, et al.
Veröffentlicht: (2026)
von: Sheng, Mingzhi, et al.
Veröffentlicht: (2026)
Digital Financial Inclusion and Entrepreneurship: A Spatial Analysis of Rural China
von: Wenyi Lyu, et al.
Veröffentlicht: (2025)
von: Wenyi Lyu, et al.
Veröffentlicht: (2025)
3D Magnetic Field Reconstruction and Mapping with Physics-Informed Neural Networks
von: Yu, Haohan, et al.
Veröffentlicht: (2026)
von: Yu, Haohan, et al.
Veröffentlicht: (2026)
PreAfford: Universal Affordance-Based Pre-Grasping for Diverse Objects and Environments
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
von: Ding, Kairui, et al.
Veröffentlicht: (2024)
CooHOI: Learning Cooperative Human-Object Interaction with Manipulated Object Dynamics
von: Gao, Jiawei, et al.
Veröffentlicht: (2024)
von: Gao, Jiawei, et al.
Veröffentlicht: (2024)
P-MapNet: Far-seeing Map Generator Enhanced by both SDMap and HDMap Priors
von: Jiang, Zhou, et al.
Veröffentlicht: (2024)
von: Jiang, Zhou, et al.
Veröffentlicht: (2024)
A Hybrid Approach for Closing the Sim2real Appearance Gap in Game Engine Synthetic Datasets
von: Pasios, Stefanos
Veröffentlicht: (2026)
von: Pasios, Stefanos
Veröffentlicht: (2026)
Morphodynamic Effects and Fish Habitat Assessment at an Idealized Bifurcation: A Numerical Experiment
von: Qianqian Wang, et al.
Veröffentlicht: (2025)
von: Qianqian Wang, et al.
Veröffentlicht: (2025)
SPRITE: From Static Mockups to Engine-Ready Game UI
von: Bai, Yunshu, et al.
Veröffentlicht: (2026)
von: Bai, Yunshu, et al.
Veröffentlicht: (2026)
Disintegration and Skipping Dynamics of Bilobate-shaped Meteoroids for Generating Ultra-Long Strewn Fields
von: Li, HaoYu
Veröffentlicht: (2025)
von: Li, HaoYu
Veröffentlicht: (2025)
Ähnliche Einträge
-
PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
von: Gao, Mingju, et al.
Veröffentlicht: (2025) -
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
von: Gao, Huan-ang, et al.
Veröffentlicht: (2024) -
Training-Free Model Merging for Multi-target Domain Adaptation
von: Li, Wenyi, et al.
Veröffentlicht: (2024) -
Benchmarking PhD-Level Coding in 3D Geometric Computer Vision
von: Li, Wenyi, et al.
Veröffentlicht: (2026) -
Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models
von: Chi, Haohan, et al.
Veröffentlicht: (2025)