TokenHSI: Unified Synthesis of Physical Human-Scene Interactions through Task Tokenization
Fuente:
arXiv
Guardado en:
| Autores principales: | Pan, Liang, Yang, Zeshi, Dou, Zhiyang, Wang, Wenjia, Huang, Buzhen, Dai, Bo, Komura, Taku, Wang, Jingbo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation
por: Wang, Wenjia, et al.
Publicado: (2024)
por: Wang, Wenjia, et al.
Publicado: (2024)
Synthesizing Physically Plausible Human Motions in 3D Scenes
por: Pan, Liang, et al.
Publicado: (2023)
por: Pan, Liang, et al.
Publicado: (2023)
SceMoS: Scene-Aware 3D Human Motion Synthesis by Planning with Geometry-Grounded Tokens
por: Ghosh, Anindita, et al.
Publicado: (2026)
por: Ghosh, Anindita, et al.
Publicado: (2026)
TLControl: Trajectory and Language Control for Human Motion Synthesis
por: Wan, Weilin, et al.
Publicado: (2023)
por: Wan, Weilin, et al.
Publicado: (2023)
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
por: Lou, Yuke, et al.
Publicado: (2025)
por: Lou, Yuke, et al.
Publicado: (2025)
Pay Attention and Move Better: Harnessing Attention for Interactive Motion Generation and Training-free Editing
por: Chen, Ling-Hao, et al.
Publicado: (2024)
por: Chen, Ling-Hao, et al.
Publicado: (2024)
CoDA: Coordinated Diffusion Noise Optimization for Whole-Body Manipulation of Articulated Objects
por: Pi, Huaijin, et al.
Publicado: (2025)
por: Pi, Huaijin, et al.
Publicado: (2025)
MOSPA: Human Motion Generation Driven by Spatial Audio
por: Xu, Shuyang, et al.
Publicado: (2025)
por: Xu, Shuyang, et al.
Publicado: (2025)
EMDM: Efficient Motion Diffusion Model for Fast and High-Quality Motion Generation
por: Zhou, Wenyang, et al.
Publicado: (2023)
por: Zhou, Wenyang, et al.
Publicado: (2023)
Unified Human-Scene Interaction via Prompted Chain-of-Contacts
por: Xiao, Zeqi, et al.
Publicado: (2023)
por: Xiao, Zeqi, et al.
Publicado: (2023)
Closely Interactive Human Reconstruction with Proxemics and Physics-Guided Adaption
por: Huang, Buzhen, et al.
Publicado: (2024)
por: Huang, Buzhen, et al.
Publicado: (2024)
EgoReAct: Egocentric Video-Driven 3D Human Reaction Generation
por: Zhang, Libo, et al.
Publicado: (2025)
por: Zhang, Libo, et al.
Publicado: (2025)
RMD: A Simple Baseline for More General Human Motion Generation via Training-free Retrieval-Augmented Motion Diffuse
por: Liao, Zhouyingcheng, et al.
Publicado: (2024)
por: Liao, Zhouyingcheng, et al.
Publicado: (2024)
MotionWavelet: Human Motion Prediction via Wavelet Manifold Learning
por: Feng, Yuming, et al.
Publicado: (2024)
por: Feng, Yuming, et al.
Publicado: (2024)
EmbodMocap: In-the-Wild 4D Human-Scene Reconstruction for Embodied Agents
por: Wang, Wenjia, et al.
Publicado: (2026)
por: Wang, Wenjia, et al.
Publicado: (2026)
Motion2Motion: Cross-topology Motion Transfer with Sparse Correspondence
por: Chen, Ling-Hao, et al.
Publicado: (2025)
por: Chen, Ling-Hao, et al.
Publicado: (2025)
PhysHSI: Towards a Real-World Generalizable and Natural Humanoid-Scene Interaction System
por: Wang, Huayi, et al.
Publicado: (2025)
por: Wang, Huayi, et al.
Publicado: (2025)
GenHSI: Controllable Generation of Human-Scene Interaction Videos
por: Li, Zekun, et al.
Publicado: (2025)
por: Li, Zekun, et al.
Publicado: (2025)
EgoGrasp: World-Space Hand-Object Interaction Estimation from Egocentric Videos
por: Fu, Hongming, et al.
Publicado: (2026)
por: Fu, Hongming, et al.
Publicado: (2026)
MVTokenFlow: High-quality 4D Content Generation using Multiview Token Flow
por: Huang, Hanzhuo, et al.
Publicado: (2025)
por: Huang, Hanzhuo, et al.
Publicado: (2025)
E-React: Towards Emotionally Controlled Synthesis of Human Reactions
por: Zhu, Chen, et al.
Publicado: (2025)
por: Zhu, Chen, et al.
Publicado: (2025)
DNAMotifTokenizer: Towards Biologically Informed Tokenization of Genomic Sequences
por: Zhou, Xiaoxiao, et al.
Publicado: (2025)
por: Zhou, Xiaoxiao, et al.
Publicado: (2025)
ZeroHSI: Zero-Shot 4D Human-Scene Interaction by Video Generation
por: Li, Hongjie, et al.
Publicado: (2024)
por: Li, Hongjie, et al.
Publicado: (2024)
FantasyHSI: Video-Generation-Centric 4D Human Synthesis In Any Scene through A Graph-based Multi-Agent Framework
por: Mu, Lingzhou, et al.
Publicado: (2025)
por: Mu, Lingzhou, et al.
Publicado: (2025)
CBIL: Collective Behavior Imitation Learning for Fish from Real Videos
por: Wu, Yifan, et al.
Publicado: (2025)
por: Wu, Yifan, et al.
Publicado: (2025)
MemoSight: Unifying Context Compression and Multi Token Prediction for Reasoning Acceleration
por: Liu, Xinyu, et al.
Publicado: (2026)
por: Liu, Xinyu, et al.
Publicado: (2026)
Strips as Tokens: Artist Mesh Generation with Native UV Segmentation
por: Xu, Rui, et al.
Publicado: (2026)
por: Xu, Rui, et al.
Publicado: (2026)
SENC: Handling Self-collision in Neural Cloth Simulation
por: Liao, Zhouyingcheng, et al.
Publicado: (2024)
por: Liao, Zhouyingcheng, et al.
Publicado: (2024)
UniHOI: Unified Human-Object Interaction Understanding via Unified Token Space
por: Yang, Panqi, et al.
Publicado: (2025)
por: Yang, Panqi, et al.
Publicado: (2025)
Unified Pix Token And Word Token Generative Language Model
por: Leung, Haun, et al.
Publicado: (2026)
por: Leung, Haun, et al.
Publicado: (2026)
MergeDNA: Context-aware Genome Modeling with Dynamic Tokenization through Token Merging
por: Li, Siyuan, et al.
Publicado: (2025)
por: Li, Siyuan, et al.
Publicado: (2025)
InterControl: Zero-shot Human Interaction Generation by Controlling Every Joint
por: Wang, Zhenzhi, et al.
Publicado: (2023)
por: Wang, Zhenzhi, et al.
Publicado: (2023)
Unified Medical Image Tokenizer for Autoregressive Synthesis and Understanding
por: Ma, Chenglong, et al.
Publicado: (2025)
por: Ma, Chenglong, et al.
Publicado: (2025)
Simultaneously Recovering Multi-Person Meshes and Multi-View Cameras with Human Semantics
por: Huang, Buzhen, et al.
Publicado: (2024)
por: Huang, Buzhen, et al.
Publicado: (2024)
Learning Unified User Quantized Tokenizers for User Representation
por: He, Chuan, et al.
Publicado: (2025)
por: He, Chuan, et al.
Publicado: (2025)
UNIC: Neural Garment Deformation Field for Real-time Clothed Character Animation
por: Zhao, Chengfeng, et al.
Publicado: (2026)
por: Zhao, Chengfeng, et al.
Publicado: (2026)
CHOICE: Coordinated Human-Object Interaction in Cluttered Environments for Pick-and-Place Actions
por: Lu, Jintao, et al.
Publicado: (2024)
por: Lu, Jintao, et al.
Publicado: (2024)
Dynamic Realms: 4D Content Analysis, Recovery and Generation with Geometric, Topological and Physical Priors
por: Dou, Zhiyang
Publicado: (2024)
por: Dou, Zhiyang
Publicado: (2024)
MergeTok: Unified Continuous and Discrete Visual Tokenization via Token Merging
por: Zhang, Luyuan, et al.
Publicado: (2026)
por: Zhang, Luyuan, et al.
Publicado: (2026)
Tokenization Matters! Degrading Large Language Models through Challenging Their Tokenization
por: Wang, Dixuan, et al.
Publicado: (2024)
por: Wang, Dixuan, et al.
Publicado: (2024)
Ejemplares similares
-
SIMS: Simulating Stylized Human-Scene Interactions with Retrieval-Augmented Script Generation
por: Wang, Wenjia, et al.
Publicado: (2024) -
Synthesizing Physically Plausible Human Motions in 3D Scenes
por: Pan, Liang, et al.
Publicado: (2023) -
SceMoS: Scene-Aware 3D Human Motion Synthesis by Planning with Geometry-Grounded Tokens
por: Ghosh, Anindita, et al.
Publicado: (2026) -
TLControl: Trajectory and Language Control for Human Motion Synthesis
por: Wan, Weilin, et al.
Publicado: (2023) -
Zero-Shot Human-Object Interaction Synthesis with Multimodal Priors
por: Lou, Yuke, et al.
Publicado: (2025)