CoMA: Compositional Human Motion Generation with Multi-modal Agents
Fuente:
arXiv
Saved in:
| Main Authors: | Sun, Shanlin, De Araujo, Gabriel, Xu, Jiaqi, Zhou, Shenghan, Zhang, Hanwen, Huang, Ziheng, You, Chenyu, Xie, Xiaohui |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Adaptive Image Registration: A Hybrid Approach Integrating Deep Learning and Optimization Functions for Enhanced Precision
by: De Araujo, Gabriel, et al.
Published: (2023)
by: De Araujo, Gabriel, et al.
Published: (2023)
Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering
by: Sun, Shanlin, et al.
Published: (2025)
by: Sun, Shanlin, et al.
Published: (2025)
CoMA: Complementary Masking and Hierarchical Dynamic Multi-Window Self-Attention in a Unified Pre-training Framework
by: Li, Jiaxuan, et al.
Published: (2025)
by: Li, Jiaxuan, et al.
Published: (2025)
Heisenberg uniqueness pairs and the wave equation
by: Shanlin Huang, et al.
Published: (2025)
by: Shanlin Huang, et al.
Published: (2025)
PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs
by: Zhang, Zhilin, et al.
Published: (2025)
by: Zhang, Zhilin, et al.
Published: (2025)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
by: Wang, Sen, et al.
Published: (2024)
by: Wang, Sen, et al.
Published: (2024)
Generative AI-Driven High-Fidelity Human Motion Simulation
by: Iyer, Hari, et al.
Published: (2025)
by: Iyer, Hari, et al.
Published: (2025)
Learning Emergent Modular Representations in Multi-modality Medical Vision Foundation Models
by: He, Yuting, et al.
Published: (2026)
by: He, Yuting, et al.
Published: (2026)
Integrating Efficient Optimal Transport and Functional Maps For Unsupervised Shape Correspondence Learning
by: Le, Tung, et al.
Published: (2024)
by: Le, Tung, et al.
Published: (2024)
MuMA-ToM: Multi-modal Multi-Agent Theory of Mind
by: Shi, Haojun, et al.
Published: (2024)
by: Shi, Haojun, et al.
Published: (2024)
FrankenMotion: Part-level Human Motion Generation and Composition
by: Li, Chuqiao, et al.
Published: (2026)
by: Li, Chuqiao, et al.
Published: (2026)
PE-MA: Parameter-Efficient Co-Evolution of Multi-Agent Systems
by: Deng, Yingfan, et al.
Published: (2025)
by: Deng, Yingfan, et al.
Published: (2025)
AgentRFC: Security Design Principles and Conformance Testing for Agent Protocols
by: Zheng, Shenghan, et al.
Published: (2026)
by: Zheng, Shenghan, et al.
Published: (2026)
Medical Image Registration via Neural Fields
by: Sun, Shanlin, et al.
Published: (2022)
by: Sun, Shanlin, et al.
Published: (2022)
TimeSeriesScientist: A General-Purpose AI Agent for Time Series Analysis
by: Zhao, Haokun, et al.
Published: (2025)
by: Zhao, Haokun, et al.
Published: (2025)
QuantAgent: Price-Driven Multi-Agent LLMs for High-Frequency Trading
by: Xiong, Fei, et al.
Published: (2025)
by: Xiong, Fei, et al.
Published: (2025)
LidaRF: Delving into Lidar for Neural Radiance Field on Street Scenes
by: Sun, Shanlin, et al.
Published: (2024)
by: Sun, Shanlin, et al.
Published: (2024)
Light Field Diffusion for Single-View Novel View Synthesis
by: Xiong, Yifeng, et al.
Published: (2023)
by: Xiong, Yifeng, et al.
Published: (2023)
Diffeomorphic Mesh Deformation via Efficient Optimal Transport for Cortical Surface Reconstruction
by: Le, Tung, et al.
Published: (2023)
by: Le, Tung, et al.
Published: (2023)
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
by: Liu, Shansong, et al.
Published: (2024)
by: Liu, Shansong, et al.
Published: (2024)
SlideGen: Collaborative Multimodal Agents for Scientific Slide Generation
by: Liang, Xin, et al.
Published: (2025)
by: Liang, Xin, et al.
Published: (2025)
Hypergraph-based Motion Generation with Multi-modal Interaction Relational Reasoning
by: Wu, Keshu, et al.
Published: (2024)
by: Wu, Keshu, et al.
Published: (2024)
Multi-modal Generative AI: Multi-modal LLMs, Diffusions, and the Unification
by: Wang, Xin, et al.
Published: (2024)
by: Wang, Xin, et al.
Published: (2024)
Calibrating Multi-modal Representations: A Pursuit of Group Robustness without Annotations
by: You, Chenyu, et al.
Published: (2024)
by: You, Chenyu, et al.
Published: (2024)
SAFE--MA--RRT: Multi-Agent Motion Planning with Data-Driven Safety Certificates
by: Esmaeili, Babak, et al.
Published: (2025)
by: Esmaeili, Babak, et al.
Published: (2025)
Dynamical versions of Morgan's Uncertainty Principle and Electromagnetic Schrödinger Evolutions
by: Huang, Shanlin, et al.
Published: (2025)
by: Huang, Shanlin, et al.
Published: (2025)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
by: Shi, Junyu, et al.
Published: (2025)
by: Shi, Junyu, et al.
Published: (2025)
AgentCTG: Harnessing Multi-Agent Collaboration for Fine-Grained Precise Control in Text Generation
by: Zhou, Xinxu, et al.
Published: (2025)
by: Zhou, Xinxu, et al.
Published: (2025)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
by: Xu, Jiahao, et al.
Published: (2026)
by: Xu, Jiahao, et al.
Published: (2026)
PhotoFramer: Multi-modal Image Composition Instruction
by: You, Zhiyuan, et al.
Published: (2025)
by: You, Zhiyuan, et al.
Published: (2025)
Generative RLHF-V: Learning Principles from Multi-modal Human Preference
by: Zhou, Jiayi, et al.
Published: (2025)
by: Zhou, Jiayi, et al.
Published: (2025)
On-device Large Multi-modal Agent for Human Activity Recognition
by: Siam, Md Shakhrul Iman, et al.
Published: (2025)
by: Siam, Md Shakhrul Iman, et al.
Published: (2025)
Computer-Vision-Enabled Worker Video Analysis for Motion Amount Quantification
by: Iyer, Hari, et al.
Published: (2024)
by: Iyer, Hari, et al.
Published: (2024)
ReCoM: Realistic Co-Speech Motion Generation with Recurrent Embedded Transformer
by: Xie, Yong, et al.
Published: (2025)
by: Xie, Yong, et al.
Published: (2025)
ChatMotion: A Multimodal Multi-Agent for Human Motion Analysis
by: Li, Lei, et al.
Published: (2025)
by: Li, Lei, et al.
Published: (2025)
MDG: Masked Denoising Generation for Multi-Agent Behavior Modeling in Traffic Environments
by: Huang, Zhiyu, et al.
Published: (2025)
by: Huang, Zhiyu, et al.
Published: (2025)
CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos
by: Zhao, Chengfeng, et al.
Published: (2026)
by: Zhao, Chengfeng, et al.
Published: (2026)
PlanAgent: A Multi-modal Large Language Agent for Closed-loop Vehicle Motion Planning
by: Zheng, Yupeng, et al.
Published: (2024)
by: Zheng, Yupeng, et al.
Published: (2024)
AgentClick: A Skill-Based Human-in-the-Loop Review Layer for Terminal AI Agents
by: Zhuang, Haomin, et al.
Published: (2026)
by: Zhuang, Haomin, et al.
Published: (2026)
CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment
by: Kang, Li, et al.
Published: (2026)
by: Kang, Li, et al.
Published: (2026)
Similar Items
-
Adaptive Image Registration: A Hybrid Approach Integrating Deep Learning and Optimization Functions for Enhanced Precision
by: De Araujo, Gabriel, et al.
Published: (2023) -
Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering
by: Sun, Shanlin, et al.
Published: (2025) -
CoMA: Complementary Masking and Hierarchical Dynamic Multi-Window Self-Attention in a Unified Pre-training Framework
by: Li, Jiaxuan, et al.
Published: (2025) -
Heisenberg uniqueness pairs and the wave equation
by: Shanlin Huang, et al.
Published: (2025) -
PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs
by: Zhang, Zhilin, et al.
Published: (2025)