Saved in:
| Main Authors: | Bai, Xiangyu, Liang, He, Galoaa, Bishoy, Nandi, Utsav, Moezzi, Shayda, He, Yuhang, Ostadabbas, Sarah |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.04221 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Motion-o: Trajectory-Grounded Video Reasoning
by: Galoaa, Bishoy, et al.
Published: (2026)
by: Galoaa, Bishoy, et al.
Published: (2026)
Look Around and Pay Attention: Multi-camera Point Tracking Reimagined with Transformers
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
Structure Over Scale: Learning Visual Reasoning from Pedagogical Video
by: Galoaa, Bishoy, et al.
Published: (2026)
by: Galoaa, Bishoy, et al.
Published: (2026)
Lang2Motion: Bridging Language and Motion through Joint Embedding Spaces
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
Track and Caption Any Motion: Query-Free Motion Discovery and Description in Videos
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
HORNet: Task-Guided Frame Selection for Video Question Answering with Vision-Language Models
by: Bai, Xiangyu, et al.
Published: (2026)
by: Bai, Xiangyu, et al.
Published: (2026)
PanoWorld: Geometry-Consistent Panoramic Video World Modeling
by: Jiang, Le, et al.
Published: (2026)
by: Jiang, Le, et al.
Published: (2026)
UniTrack: Differentiable Graph Representation Learning for Multi-Object Tracking
by: Galoaa, Bishoy, et al.
Published: (2026)
by: Galoaa, Bishoy, et al.
Published: (2026)
K-Track: Kalman-Enhanced Tracking for Accelerating Deep Point Trackers on Edge Devices
by: Galoaa, Bishoy, et al.
Published: (2025)
by: Galoaa, Bishoy, et al.
Published: (2025)
Broadening View Synthesis of Dynamic Scenes from Constrained Monocular Videos
by: Jiang, Le, et al.
Published: (2025)
by: Jiang, Le, et al.
Published: (2025)
STREAMS: An Assistive Multimodal AI Framework for Empowering Biosignal Based Robotic Controls
by: Rabiee, Ali, et al.
Published: (2024)
by: Rabiee, Ali, et al.
Published: (2024)
Re$^2$MoGen: Open-Vocabulary Motion Generation via LLM Reasoning and Physics-Aware Refinement
by: Zheng, Jiakun, et al.
Published: (2026)
by: Zheng, Jiakun, et al.
Published: (2026)
Uncertainty-Aware Ankle Exoskeleton Control
by: Tourk, Fatima Mumtaza, et al.
Published: (2025)
by: Tourk, Fatima Mumtaza, et al.
Published: (2025)
SnapMoGen: Human Motion Generation from Expressive Texts
by: Guo, Chuan, et al.
Published: (2025)
by: Guo, Chuan, et al.
Published: (2025)
MoTrans: Customized Motion Transfer with Text-driven Video Diffusion Models
by: Li, Xiaomin, et al.
Published: (2024)
by: Li, Xiaomin, et al.
Published: (2024)
MoLingo: Motion-Language Alignment for Text-to-Motion Generation
by: He, Yannan, et al.
Published: (2025)
by: He, Yannan, et al.
Published: (2025)
LaMoGen: Laban Movement-Guided Diffusion for Text-to-Motion Generation
by: Kim, Heechang, et al.
Published: (2025)
by: Kim, Heechang, et al.
Published: (2025)
CrowdMoGen: Zero-Shot Text-Driven Collective Motion Generation
by: Cao, Yukang, et al.
Published: (2024)
by: Cao, Yukang, et al.
Published: (2024)
ReVision: Refining Video Diffusion with Explicit 3D Motion Modeling
by: Liu, Qihao, et al.
Published: (2025)
by: Liu, Qihao, et al.
Published: (2025)
MoGenTS: Motion Generation based on Spatial-Temporal Joint Modeling
by: Yuan, Weihao, et al.
Published: (2024)
by: Yuan, Weihao, et al.
Published: (2024)
ReMoT: Reinforcement Learning with Motion Contrast Triplets
by: Wan, Cong, et al.
Published: (2026)
by: Wan, Cong, et al.
Published: (2026)
Learning Multimodal AI Algorithms for Amplifying Limited User Input into High-dimensional Control Space
by: Rabiee, Ali, et al.
Published: (2025)
by: Rabiee, Ali, et al.
Published: (2025)
Formalizing Linear Motion G-code for Invariant Checking and Differential Testing of Fabrication Tools
by: He, Yumeng, et al.
Published: (2025)
by: He, Yumeng, et al.
Published: (2025)
OmniMoGen: Unifying Human Motion Generation via Learning from Interleaved Text-Motion Instructions
by: Bu, Wendong, et al.
Published: (2025)
by: Bu, Wendong, et al.
Published: (2025)
UniMoGen: Universal Motion Generation
by: Khani, Aliasghar, et al.
Published: (2025)
by: Khani, Aliasghar, et al.
Published: (2025)
Please Make it Sound like Human: Encoder-Decoder vs. Decoder-Only Transformers for AI-to-Human Text Style Transfer
by: Paneru, Utsav
Published: (2026)
by: Paneru, Utsav
Published: (2026)
CoMo: Compositional Motion Customization for Text-to-Video Generation
by: Xu, Youcan, et al.
Published: (2025)
by: Xu, Youcan, et al.
Published: (2025)
VideoGen-Eval: Agent-based System for Video Generation Evaluation
by: Yang, Yuhang, et al.
Published: (2025)
by: Yang, Yuhang, et al.
Published: (2025)
Beyond the Buzzword: Rethinking Polycrises in Public Policy and Administration Research
by: Bishoy L. Zaki
Published: (2025)
by: Bishoy L. Zaki
Published: (2025)
Policy Learning and Policy Analysis Within European Union Institutions: A Systematic Review and Research Agenda
by: Bishoy L. Zaki
Published: (2026)
by: Bishoy L. Zaki
Published: (2026)
Toward a Theory of Value as Praxis: Linking Public Values and Public Value
by: Bishoy L. Zaki
Published: (2026)
by: Bishoy L. Zaki
Published: (2026)
AdSum: Two-stream Audio-visual Summarization for Automated Video Advertisement Clipping
by: Xie, Wen, et al.
Published: (2025)
by: Xie, Wen, et al.
Published: (2025)
ReGenNet: Towards Human Action-Reaction Synthesis
by: Xu, Liang, et al.
Published: (2024)
by: Xu, Liang, et al.
Published: (2024)
MoVideo: Motion-Aware Video Generation with Diffusion Models
by: Liang, Jingyun, et al.
Published: (2023)
by: Liang, Jingyun, et al.
Published: (2023)
Fleximo: Towards Flexible Text-to-Human Motion Video Generation
by: Zhang, Yuhang, et al.
Published: (2024)
by: Zhang, Yuhang, et al.
Published: (2024)
Searching Priors Makes Text-to-Video Synthesis Better
by: Cheng, Haoran, et al.
Published: (2024)
by: Cheng, Haoran, et al.
Published: (2024)
GenMAC: Compositional Text-to-Video Generation with Multi-Agent Collaboration
by: Huang, Kaiyi, et al.
Published: (2024)
by: Huang, Kaiyi, et al.
Published: (2024)
ReMoS: 3D Motion-Conditioned Reaction Synthesis for Two-Person Interactions
by: Ghosh, Anindita, et al.
Published: (2023)
by: Ghosh, Anindita, et al.
Published: (2023)
Overcoming Small Data Limitations in Video-Based Infant Respiration Estimation
by: Song, Liyang, et al.
Published: (2025)
by: Song, Liyang, et al.
Published: (2025)
Mechanisms of Multimodal Synchronization: Insights from Decoder-Based Video-Text-to-Speech Synthesis
by: Gupta, Akshita, et al.
Published: (2024)
by: Gupta, Akshita, et al.
Published: (2024)
Similar Items
-
Motion-o: Trajectory-Grounded Video Reasoning
by: Galoaa, Bishoy, et al.
Published: (2026) -
Look Around and Pay Attention: Multi-camera Point Tracking Reimagined with Transformers
by: Galoaa, Bishoy, et al.
Published: (2025) -
Structure Over Scale: Learning Visual Reasoning from Pedagogical Video
by: Galoaa, Bishoy, et al.
Published: (2026) -
Lang2Motion: Bridging Language and Motion through Joint Embedding Spaces
by: Galoaa, Bishoy, et al.
Published: (2025) -
Track and Caption Any Motion: Query-Free Motion Discovery and Description in Videos
by: Galoaa, Bishoy, et al.
Published: (2025)