PAM: A Pose-Appearance-Motion Engine for Sim-to-Real HOI Video Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Gao, Mingju, Yang, Kaisen, Gao, Huan-ang, Li, Bohan, Ding, Ao, Li, Wenyi, Yu, Yangcheng, Liu, Jinkun, Xu, Shaocong, Niu, Yike, Chi, Haohan, Chen, Hao, Tang, Hao, Zhang, Yu, Yi, Li, Zhao, Hao |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
por: Gao, Mingju, et al.
Publicado: (2025)
por: Gao, Mingju, et al.
Publicado: (2025)
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
por: Gao, Huan-ang, et al.
Publicado: (2024)
por: Gao, Huan-ang, et al.
Publicado: (2024)
Training-Free Model Merging for Multi-target Domain Adaptation
por: Li, Wenyi, et al.
Publicado: (2024)
por: Li, Wenyi, et al.
Publicado: (2024)
Benchmarking PhD-Level Coding in 3D Geometric Computer Vision
por: Li, Wenyi, et al.
Publicado: (2026)
por: Li, Wenyi, et al.
Publicado: (2026)
Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models
por: Chi, Haohan, et al.
Publicado: (2025)
por: Chi, Haohan, et al.
Publicado: (2025)
FairDiff: Fair Segmentation with Point-Image Diffusion
por: Li, Wenyi, et al.
Publicado: (2024)
por: Li, Wenyi, et al.
Publicado: (2024)
Alias-free 4D Gaussian Splatting
por: Chen, Zilong, et al.
Publicado: (2025)
por: Chen, Zilong, et al.
Publicado: (2025)
Challenger: Affordable Adversarial Driving Video Generation
por: Xu, Zhiyuan, et al.
Publicado: (2025)
por: Xu, Zhiyuan, et al.
Publicado: (2025)
FB-4D: Spatial-Temporal Coherent Dynamic 3D Content Generation with Feature Banks
por: Li, Jinwei, et al.
Publicado: (2025)
por: Li, Jinwei, et al.
Publicado: (2025)
Dual-frame Fluid Motion Estimation with Test-time Optimization and Zero-divergence Loss
por: Zhang, Yifei, et al.
Publicado: (2024)
por: Zhang, Yifei, et al.
Publicado: (2024)
MagicPose4D: Crafting Articulated Models with Appearance and Motion Control
por: Zhang, Hao, et al.
Publicado: (2024)
por: Zhang, Hao, et al.
Publicado: (2024)
TwinAligner: Visual-Dynamic Alignment Empowers Physics-aware Real2Sim2Real for Robotic Manipulation
por: Fan, Hongwei, et al.
Publicado: (2025)
por: Fan, Hongwei, et al.
Publicado: (2025)
One View, Many Worlds: Single-Image to 3D Object Meets Generative Domain Randomization for One-Shot 6D Pose Estimation
por: Geng, Zheng, et al.
Publicado: (2025)
por: Geng, Zheng, et al.
Publicado: (2025)
Real2Sim in HOI: Toward Physically Plausible HOI Reconstruction from Monocular Videos
por: Zhao, Yubo, et al.
Publicado: (2026)
por: Zhao, Yubo, et al.
Publicado: (2026)
An Infinite Family of Primitive Heron Triangles with Two Sides as Perfect Squares
por: Li, Yangcheng
Publicado: (2026)
por: Li, Yangcheng
Publicado: (2026)
A new perspective of arithmetic billiards
por: Li, Yangcheng
Publicado: (2023)
por: Li, Yangcheng
Publicado: (2023)
PosePilot: Steering Camera Pose for Generative World Models with Self-supervised Depth
por: Jin, Bu, et al.
Publicado: (2025)
por: Jin, Bu, et al.
Publicado: (2025)
CubeBench: Diagnosing Interactive, Long-Horizon Spatial Reasoning Under Partial Observations
por: Gao, Huan-ang, et al.
Publicado: (2025)
por: Gao, Huan-ang, et al.
Publicado: (2025)
ArtHOI: Articulated Human-Object Interaction Synthesis by 4D Reconstruction from Video Priors
por: Huang, Zihao, et al.
Publicado: (2026)
por: Huang, Zihao, et al.
Publicado: (2026)
Afford-X: Generalizable and Slim Affordance Reasoning for Task-oriented Manipulation
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
por: Zhu, Xiaomeng, et al.
Publicado: (2025)
Ctrl-U: Robust Conditional Image Generation via Uncertainty-aware Reward Modeling
por: Zhang, Guiyu, et al.
Publicado: (2024)
por: Zhang, Guiyu, et al.
Publicado: (2024)
Event-based Motion & Appearance Fusion for 6D Object Pose Tracking
por: Li, Zhichao, et al.
Publicado: (2026)
por: Li, Zhichao, et al.
Publicado: (2026)
Shadows, Quasinormal Modes, and Optical Appearances of Black Holes in Horndeski Theory
por: Luo, Zhi, et al.
Publicado: (2024)
por: Luo, Zhi, et al.
Publicado: (2024)
AVD2: Accident Video Diffusion for Accident Video Description
por: Li, Cheng, et al.
Publicado: (2025)
por: Li, Cheng, et al.
Publicado: (2025)
Unifying Appearance Codes and Bilateral Grids for Driving Scene Gaussian Splatting
por: Wang, Nan, et al.
Publicado: (2025)
por: Wang, Nan, et al.
Publicado: (2025)
PersPose: 3D Human Pose Estimation with Perspective Encoding and Perspective Rotation
por: Hao, Xiaoyang, et al.
Publicado: (2025)
por: Hao, Xiaoyang, et al.
Publicado: (2025)
$(G,F)$-points on $\mathbb{Q}$-algebraic varieties
por: Li, Yangcheng, et al.
Publicado: (2025)
por: Li, Yangcheng, et al.
Publicado: (2025)
Thinking in Text and Images: Interleaved Vision--Language Reasoning Traces for Long-Horizon Robot Manipulation
por: Liu, Jinkun, et al.
Publicado: (2026)
por: Liu, Jinkun, et al.
Publicado: (2026)
MoSA: Motion-Coherent Human Video Generation via Structure-Appearance Decoupling
por: Wang, Haoyu, et al.
Publicado: (2025)
por: Wang, Haoyu, et al.
Publicado: (2025)
Diffusion-based Visual Anagram as Multi-task Learning
por: Xu, Zhiyuan, et al.
Publicado: (2024)
por: Xu, Zhiyuan, et al.
Publicado: (2024)
FlexAM: Flexible Appearance-Motion Decomposition for Versatile Video Generation Control
por: Sheng, Mingzhi, et al.
Publicado: (2026)
por: Sheng, Mingzhi, et al.
Publicado: (2026)
Digital Financial Inclusion and Entrepreneurship: A Spatial Analysis of Rural China
por: Wenyi Lyu, et al.
Publicado: (2025)
por: Wenyi Lyu, et al.
Publicado: (2025)
3D Magnetic Field Reconstruction and Mapping with Physics-Informed Neural Networks
por: Yu, Haohan, et al.
Publicado: (2026)
por: Yu, Haohan, et al.
Publicado: (2026)
PreAfford: Universal Affordance-Based Pre-Grasping for Diverse Objects and Environments
por: Ding, Kairui, et al.
Publicado: (2024)
por: Ding, Kairui, et al.
Publicado: (2024)
CooHOI: Learning Cooperative Human-Object Interaction with Manipulated Object Dynamics
por: Gao, Jiawei, et al.
Publicado: (2024)
por: Gao, Jiawei, et al.
Publicado: (2024)
P-MapNet: Far-seeing Map Generator Enhanced by both SDMap and HDMap Priors
por: Jiang, Zhou, et al.
Publicado: (2024)
por: Jiang, Zhou, et al.
Publicado: (2024)
A Hybrid Approach for Closing the Sim2real Appearance Gap in Game Engine Synthetic Datasets
por: Pasios, Stefanos
Publicado: (2026)
por: Pasios, Stefanos
Publicado: (2026)
Morphodynamic Effects and Fish Habitat Assessment at an Idealized Bifurcation: A Numerical Experiment
por: Qianqian Wang, et al.
Publicado: (2025)
por: Qianqian Wang, et al.
Publicado: (2025)
SPRITE: From Static Mockups to Engine-Ready Game UI
por: Bai, Yunshu, et al.
Publicado: (2026)
por: Bai, Yunshu, et al.
Publicado: (2026)
Disintegration and Skipping Dynamics of Bilobate-shaped Meteoroids for Generating Ultra-Long Strewn Fields
por: Li, HaoYu
Publicado: (2025)
por: Li, HaoYu
Publicado: (2025)
Ejemplares similares
-
PartRM: Modeling Part-Level Dynamics with Large Cross-State Reconstruction Model
por: Gao, Mingju, et al.
Publicado: (2025) -
SCP-Diff: Spatial-Categorical Joint Prior for Diffusion Based Semantic Image Synthesis
por: Gao, Huan-ang, et al.
Publicado: (2024) -
Training-Free Model Merging for Multi-target Domain Adaptation
por: Li, Wenyi, et al.
Publicado: (2024) -
Benchmarking PhD-Level Coding in 3D Geometric Computer Vision
por: Li, Wenyi, et al.
Publicado: (2026) -
Impromptu VLA: Open Weights and Open Data for Driving Vision-Language-Action Models
por: Chi, Haohan, et al.
Publicado: (2025)