CoMA: Compositional Human Motion Generation with Multi-modal Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sun, Shanlin, De Araujo, Gabriel, Xu, Jiaqi, Zhou, Shenghan, Zhang, Hanwen, Huang, Ziheng, You, Chenyu, Xie, Xiaohui |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adaptive Image Registration: A Hybrid Approach Integrating Deep Learning and Optimization Functions for Enhanced Precision
von: De Araujo, Gabriel, et al.
Veröffentlicht: (2023)
von: De Araujo, Gabriel, et al.
Veröffentlicht: (2023)
Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering
von: Sun, Shanlin, et al.
Veröffentlicht: (2025)
von: Sun, Shanlin, et al.
Veröffentlicht: (2025)
CoMA: Complementary Masking and Hierarchical Dynamic Multi-Window Self-Attention in a Unified Pre-training Framework
von: Li, Jiaxuan, et al.
Veröffentlicht: (2025)
von: Li, Jiaxuan, et al.
Veröffentlicht: (2025)
Heisenberg uniqueness pairs and the wave equation
von: Shanlin Huang, et al.
Veröffentlicht: (2025)
von: Shanlin Huang, et al.
Veröffentlicht: (2025)
PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs
von: Zhang, Zhilin, et al.
Veröffentlicht: (2025)
von: Zhang, Zhilin, et al.
Veröffentlicht: (2025)
MMoFusion: Multi-modal Co-Speech Motion Generation with Diffusion Model
von: Wang, Sen, et al.
Veröffentlicht: (2024)
von: Wang, Sen, et al.
Veröffentlicht: (2024)
Generative AI-Driven High-Fidelity Human Motion Simulation
von: Iyer, Hari, et al.
Veröffentlicht: (2025)
von: Iyer, Hari, et al.
Veröffentlicht: (2025)
Learning Emergent Modular Representations in Multi-modality Medical Vision Foundation Models
von: He, Yuting, et al.
Veröffentlicht: (2026)
von: He, Yuting, et al.
Veröffentlicht: (2026)
Integrating Efficient Optimal Transport and Functional Maps For Unsupervised Shape Correspondence Learning
von: Le, Tung, et al.
Veröffentlicht: (2024)
von: Le, Tung, et al.
Veröffentlicht: (2024)
MuMA-ToM: Multi-modal Multi-Agent Theory of Mind
von: Shi, Haojun, et al.
Veröffentlicht: (2024)
von: Shi, Haojun, et al.
Veröffentlicht: (2024)
FrankenMotion: Part-level Human Motion Generation and Composition
von: Li, Chuqiao, et al.
Veröffentlicht: (2026)
von: Li, Chuqiao, et al.
Veröffentlicht: (2026)
PE-MA: Parameter-Efficient Co-Evolution of Multi-Agent Systems
von: Deng, Yingfan, et al.
Veröffentlicht: (2025)
von: Deng, Yingfan, et al.
Veröffentlicht: (2025)
AgentRFC: Security Design Principles and Conformance Testing for Agent Protocols
von: Zheng, Shenghan, et al.
Veröffentlicht: (2026)
von: Zheng, Shenghan, et al.
Veröffentlicht: (2026)
Medical Image Registration via Neural Fields
von: Sun, Shanlin, et al.
Veröffentlicht: (2022)
von: Sun, Shanlin, et al.
Veröffentlicht: (2022)
TimeSeriesScientist: A General-Purpose AI Agent for Time Series Analysis
von: Zhao, Haokun, et al.
Veröffentlicht: (2025)
von: Zhao, Haokun, et al.
Veröffentlicht: (2025)
QuantAgent: Price-Driven Multi-Agent LLMs for High-Frequency Trading
von: Xiong, Fei, et al.
Veröffentlicht: (2025)
von: Xiong, Fei, et al.
Veröffentlicht: (2025)
LidaRF: Delving into Lidar for Neural Radiance Field on Street Scenes
von: Sun, Shanlin, et al.
Veröffentlicht: (2024)
von: Sun, Shanlin, et al.
Veröffentlicht: (2024)
Light Field Diffusion for Single-View Novel View Synthesis
von: Xiong, Yifeng, et al.
Veröffentlicht: (2023)
von: Xiong, Yifeng, et al.
Veröffentlicht: (2023)
Diffeomorphic Mesh Deformation via Efficient Optimal Transport for Cortical Surface Reconstruction
von: Le, Tung, et al.
Veröffentlicht: (2023)
von: Le, Tung, et al.
Veröffentlicht: (2023)
MuMu-LLaMA: Multi-modal Music Understanding and Generation via Large Language Models
von: Liu, Shansong, et al.
Veröffentlicht: (2024)
von: Liu, Shansong, et al.
Veröffentlicht: (2024)
SlideGen: Collaborative Multimodal Agents for Scientific Slide Generation
von: Liang, Xin, et al.
Veröffentlicht: (2025)
von: Liang, Xin, et al.
Veröffentlicht: (2025)
Hypergraph-based Motion Generation with Multi-modal Interaction Relational Reasoning
von: Wu, Keshu, et al.
Veröffentlicht: (2024)
von: Wu, Keshu, et al.
Veröffentlicht: (2024)
Multi-modal Generative AI: Multi-modal LLMs, Diffusions, and the Unification
von: Wang, Xin, et al.
Veröffentlicht: (2024)
von: Wang, Xin, et al.
Veröffentlicht: (2024)
Calibrating Multi-modal Representations: A Pursuit of Group Robustness without Annotations
von: You, Chenyu, et al.
Veröffentlicht: (2024)
von: You, Chenyu, et al.
Veröffentlicht: (2024)
SAFE--MA--RRT: Multi-Agent Motion Planning with Data-Driven Safety Certificates
von: Esmaeili, Babak, et al.
Veröffentlicht: (2025)
von: Esmaeili, Babak, et al.
Veröffentlicht: (2025)
Dynamical versions of Morgan's Uncertainty Principle and Electromagnetic Schrödinger Evolutions
von: Huang, Shanlin, et al.
Veröffentlicht: (2025)
von: Huang, Shanlin, et al.
Veröffentlicht: (2025)
GenM$^3$: Generative Pretrained Multi-path Motion Model for Text Conditional Human Motion Generation
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
von: Shi, Junyu, et al.
Veröffentlicht: (2025)
AgentCTG: Harnessing Multi-Agent Collaboration for Fine-Grained Precise Control in Text Generation
von: Zhou, Xinxu, et al.
Veröffentlicht: (2025)
von: Zhou, Xinxu, et al.
Veröffentlicht: (2025)
Stability-Driven Motion Generation for Object-Guided Human-Human Co-Manipulation
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
von: Xu, Jiahao, et al.
Veröffentlicht: (2026)
PhotoFramer: Multi-modal Image Composition Instruction
von: You, Zhiyuan, et al.
Veröffentlicht: (2025)
von: You, Zhiyuan, et al.
Veröffentlicht: (2025)
Generative RLHF-V: Learning Principles from Multi-modal Human Preference
von: Zhou, Jiayi, et al.
Veröffentlicht: (2025)
von: Zhou, Jiayi, et al.
Veröffentlicht: (2025)
On-device Large Multi-modal Agent for Human Activity Recognition
von: Siam, Md Shakhrul Iman, et al.
Veröffentlicht: (2025)
von: Siam, Md Shakhrul Iman, et al.
Veröffentlicht: (2025)
Computer-Vision-Enabled Worker Video Analysis for Motion Amount Quantification
von: Iyer, Hari, et al.
Veröffentlicht: (2024)
von: Iyer, Hari, et al.
Veröffentlicht: (2024)
ReCoM: Realistic Co-Speech Motion Generation with Recurrent Embedded Transformer
von: Xie, Yong, et al.
Veröffentlicht: (2025)
von: Xie, Yong, et al.
Veröffentlicht: (2025)
ChatMotion: A Multimodal Multi-Agent for Human Motion Analysis
von: Li, Lei, et al.
Veröffentlicht: (2025)
von: Li, Lei, et al.
Veröffentlicht: (2025)
MDG: Masked Denoising Generation for Multi-Agent Behavior Modeling in Traffic Environments
von: Huang, Zhiyu, et al.
Veröffentlicht: (2025)
von: Huang, Zhiyu, et al.
Veröffentlicht: (2025)
CoMoVi: Co-Generation of 3D Human Motions and Realistic Videos
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
von: Zhao, Chengfeng, et al.
Veröffentlicht: (2026)
PlanAgent: A Multi-modal Large Language Agent for Closed-loop Vehicle Motion Planning
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
von: Zheng, Yupeng, et al.
Veröffentlicht: (2024)
AgentClick: A Skill-Based Human-in-the-Loop Review Layer for Terminal AI Agents
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2026)
CoEnv: Driving Embodied Multi-Agent Collaboration via Compositional Environment
von: Kang, Li, et al.
Veröffentlicht: (2026)
von: Kang, Li, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Adaptive Image Registration: A Hybrid Approach Integrating Deep Learning and Optimization Functions for Enhanced Precision
von: De Araujo, Gabriel, et al.
Veröffentlicht: (2023) -
Ouroboros: Single-step Diffusion Models for Cycle-consistent Forward and Inverse Rendering
von: Sun, Shanlin, et al.
Veröffentlicht: (2025) -
CoMA: Complementary Masking and Hierarchical Dynamic Multi-Window Self-Attention in a Unified Pre-training Framework
von: Li, Jiaxuan, et al.
Veröffentlicht: (2025) -
Heisenberg uniqueness pairs and the wave equation
von: Shanlin Huang, et al.
Veröffentlicht: (2025) -
PosterGen: Aesthetic-Aware Multi-Modal Paper-to-Poster Generation via Multi-Agent LLMs
von: Zhang, Zhilin, et al.
Veröffentlicht: (2025)