ActionPlan: Future-Aware Streaming Motion Synthesis via Frame-Level Action Planning
Fuente:
arXiv
Saved in:
| Main Authors: | Nazarenus, Eric, Li, Chuqiao, He, Yannan, Xie, Xianghui, Lenssen, Jan Eric, Pons-Moll, Gerard |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
InterTrack: Tracking Human Object Interaction without Object Templates
by: Xie, Xianghui, et al.
Published: (2024)
by: Xie, Xianghui, et al.
Published: (2024)
Template Free Reconstruction of Human-object Interaction with Procedural Interaction Generation
by: Xie, Xianghui, et al.
Published: (2023)
by: Xie, Xianghui, et al.
Published: (2023)
MoLingo: Motion-Language Alignment for Text-to-Motion Generation
by: He, Yannan, et al.
Published: (2025)
by: He, Yannan, et al.
Published: (2025)
FrankenMotion: Part-level Human Motion Generation and Composition
by: Li, Chuqiao, et al.
Published: (2026)
by: Li, Chuqiao, et al.
Published: (2026)
NRDF: Neural Riemannian Distance Fields for Learning Articulated Pose Priors
by: He, Yannan, et al.
Published: (2024)
by: He, Yannan, et al.
Published: (2024)
Unimotion: Unifying 3D Human Motion Synthesis and Understanding
by: Li, Chuqiao, et al.
Published: (2024)
by: Li, Chuqiao, et al.
Published: (2024)
GEARS: Local Geometry-aware Hand-object Interaction Synthesis
by: Zhou, Keyang, et al.
Published: (2024)
by: Zhou, Keyang, et al.
Published: (2024)
Human-3Diffusion: Realistic Avatar Creation via Explicit 3D Consistent Diffusion Models
by: Xue, Yuxuan, et al.
Published: (2024)
by: Xue, Yuxuan, et al.
Published: (2024)
Gen-3Diffusion: Realistic Image-to-3D Generation via 2D & 3D Diffusion Synergy
by: Xue, Yuxuan, et al.
Published: (2024)
by: Xue, Yuxuan, et al.
Published: (2024)
InfiniHuman: Infinite 3D Human Creation with Precise Control
by: Xue, Yuxuan, et al.
Published: (2025)
by: Xue, Yuxuan, et al.
Published: (2025)
Interaction Replica: Tracking Human-Object Interaction and Scene Changes From Human Motion
by: Guzov, Vladimir, et al.
Published: (2022)
by: Guzov, Vladimir, et al.
Published: (2022)
Hoi3DGen: Generating High-Quality Human-Object-Interactions in 3D
by: Sharma, Agniv, et al.
Published: (2026)
by: Sharma, Agniv, et al.
Published: (2026)
PhySIC: Physically Plausible 3D Human-Scene Interaction and Contact from a Single Image
by: Muralidhar, Pradyumna Yalandur, et al.
Published: (2025)
by: Muralidhar, Pradyumna Yalandur, et al.
Published: (2025)
Generating Continual Human Motion in Diverse 3D Scenes
by: Mir, Aymen, et al.
Published: (2023)
by: Mir, Aymen, et al.
Published: (2023)
MVGBench: Comprehensive Benchmark for Multi-view Generation Models
by: Xie, Xianghui, et al.
Published: (2025)
by: Xie, Xianghui, et al.
Published: (2025)
Neural Localizer Fields for Continuous 3D Human Pose and Shape Estimation
by: Sárándi, István, et al.
Published: (2024)
by: Sárándi, István, et al.
Published: (2024)
SemanticNVS: Improving Semantic Scene Understanding in Generative Novel View Synthesis
by: Chen, Xinya, et al.
Published: (2026)
by: Chen, Xinya, et al.
Published: (2026)
Paint-it: Text-to-Texture Synthesis via Deep Convolutional Texture Map Optimization and Physically-Based Rendering
by: Youwang, Kim, et al.
Published: (2023)
by: Youwang, Kim, et al.
Published: (2023)
Recent Trends in 3D Reconstruction of General Non-Rigid Scenes
by: Yunus, Raza, et al.
Published: (2024)
by: Yunus, Raza, et al.
Published: (2024)
Pose-Aware Multi-Level Motion Parsing for Action Quality Assessment
by: Zhu, Shuaikang, et al.
Published: (2025)
by: Zhu, Shuaikang, et al.
Published: (2025)
Permutation-Aware Action Segmentation via Unsupervised Frame-to-Segment Alignment
by: Tran, Quoc-Huy, et al.
Published: (2023)
by: Tran, Quoc-Huy, et al.
Published: (2023)
Improving 2D Feature Representations by 3D-Aware Fine-Tuning
by: Yue, Yuanwen, et al.
Published: (2024)
by: Yue, Yuanwen, et al.
Published: (2024)
CARI4D: Category Agnostic 4D Reconstruction of Human-Object Interaction
by: Xie, Xianghui, et al.
Published: (2025)
by: Xie, Xianghui, et al.
Published: (2025)
NICP: Neural ICP for 3D Human Registration at Scale
by: Marin, Riccardo, et al.
Published: (2023)
by: Marin, Riccardo, et al.
Published: (2023)
MixFlow: Mixed Source Distributions Improve Rectified Flows
by: Nayal, Nazir, et al.
Published: (2026)
by: Nayal, Nazir, et al.
Published: (2026)
UniAct: Unified Motion Generation and Action Streaming for Humanoid Robots
by: Jiang, Nan, et al.
Published: (2025)
by: Jiang, Nan, et al.
Published: (2025)
ControlEvents: Controllable Synthesis of Event Camera Datawith Foundational Prior from Image Diffusion Models
by: Hu, Yixuan, et al.
Published: (2025)
by: Hu, Yixuan, et al.
Published: (2025)
ActionDiffusion: An Action-aware Diffusion Model for Procedure Planning in Instructional Videos
by: Shi, Lei, et al.
Published: (2024)
by: Shi, Lei, et al.
Published: (2024)
OmniStream: Mastering Perception, Reconstruction and Action in Continuous Streams
by: Yan, Yibin, et al.
Published: (2026)
by: Yan, Yibin, et al.
Published: (2026)
SceneTok: A Compressed, Diffusable Token Space for 3D Scenes
by: Asim, Mohammad, et al.
Published: (2026)
by: Asim, Mohammad, et al.
Published: (2026)
SimNP: Learning Self-Similarity Priors Between Neural Points
by: Wewer, Christopher, et al.
Published: (2023)
by: Wewer, Christopher, et al.
Published: (2023)
Sp2360: Sparse-view 360 Scene Reconstruction using Cascaded 2D Diffusion Priors
by: Paul, Soumava, et al.
Published: (2024)
by: Paul, Soumava, et al.
Published: (2024)
MoonSeg3R: Monocular Online Zero-Shot Segment Anything in 3D with Reconstructive Foundation Priors
by: Du, Zhipeng, et al.
Published: (2025)
by: Du, Zhipeng, et al.
Published: (2025)
PersonaHOI: Effortlessly Improving Personalized Face with Human-Object Interaction Generation
by: Hu, Xinting, et al.
Published: (2025)
by: Hu, Xinting, et al.
Published: (2025)
Blendify -- Python rendering framework for Blender
by: Guzov, Vladimir, et al.
Published: (2024)
by: Guzov, Vladimir, et al.
Published: (2024)
Action Emergence from Streaming Intent
by: Jing, Pengfei, et al.
Published: (2026)
by: Jing, Pengfei, et al.
Published: (2026)
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
by: Cai, Zhejia, et al.
Published: (2025)
by: Cai, Zhejia, et al.
Published: (2025)
Dynamic Execution Commitment of Vision-Language-Action Models
by: Chen, Feng, et al.
Published: (2026)
by: Chen, Feng, et al.
Published: (2026)
Spatial Reasoners for Continuous Variables in Any Domain
by: Pogodzinski, Bart, et al.
Published: (2025)
by: Pogodzinski, Bart, et al.
Published: (2025)
Spatial Reasoning with Denoising Models
by: Wewer, Christopher, et al.
Published: (2025)
by: Wewer, Christopher, et al.
Published: (2025)
Similar Items
-
InterTrack: Tracking Human Object Interaction without Object Templates
by: Xie, Xianghui, et al.
Published: (2024) -
Template Free Reconstruction of Human-object Interaction with Procedural Interaction Generation
by: Xie, Xianghui, et al.
Published: (2023) -
MoLingo: Motion-Language Alignment for Text-to-Motion Generation
by: He, Yannan, et al.
Published: (2025) -
FrankenMotion: Part-level Human Motion Generation and Composition
by: Li, Chuqiao, et al.
Published: (2026) -
NRDF: Neural Riemannian Distance Fields for Learning Articulated Pose Priors
by: He, Yannan, et al.
Published: (2024)