Towards Customized Multimodal Role-Play
Fuente:
arXiv
Saved in:
| Main Authors: | Tang, Chao, Wu, Jianzong, Shi, Qingyu, Tian, Ye, Zhang, Aixi, Jiang, Hao, Zhang, Jiangning, Tong, Yunhai |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Decouple and Track: Benchmarking and Improving Video Diffusion Transformers for Motion Transfer
by: Shi, Qingyu, et al.
Published: (2025)
by: Shi, Qingyu, et al.
Published: (2025)
Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model
by: Shi, Qingyu, et al.
Published: (2025)
by: Shi, Qingyu, et al.
Published: (2025)
DreamRelation: Bridging Customization and Relation Generation
by: Shi, Qingyu, et al.
Published: (2024)
by: Shi, Qingyu, et al.
Published: (2024)
Does Hearing Help Seeing? Investigating Audio-Video Joint Denoising for Video Generation
by: Wu, Jianzong, et al.
Published: (2025)
by: Wu, Jianzong, et al.
Published: (2025)
MotionBooth: Motion-Aware Customized Text-to-Video Generation
by: Wu, Jianzong, et al.
Published: (2024)
by: Wu, Jianzong, et al.
Published: (2024)
DiffSensei: Bridging Multi-Modal LLMs and Diffusion Models for Customized Manga Generation
by: Wu, Jianzong, et al.
Published: (2024)
by: Wu, Jianzong, et al.
Published: (2024)
Towards Scalable and Deep Graph Neural Networks via Noise Masking
by: Liang, Yuxuan, et al.
Published: (2024)
by: Liang, Yuxuan, et al.
Published: (2024)
Read to Play (R2-Play): Decision Transformer with Multimodal Game Instruction
by: Jin, Yonggang, et al.
Published: (2024)
by: Jin, Yonggang, et al.
Published: (2024)
TriPlay-RL: Tri-Role Self-Play Reinforcement Learning for LLM Safety Alignment
by: Tan, Zhewen, et al.
Published: (2026)
by: Tan, Zhewen, et al.
Published: (2026)
Multimodal Large Language Models for Medicine: A Comprehensive Survey
by: Ye, Jiarui, et al.
Published: (2025)
by: Ye, Jiarui, et al.
Published: (2025)
DREAM-S: Speculative Decoding with Searchable Drafting and Target-Aware Refinement for Multimodal Generation
by: Liu, Zining, et al.
Published: (2026)
by: Liu, Zining, et al.
Published: (2026)
Efficient Group Lasso Regularized Rank Regression with Data-Driven Parameter Determination
by: Lin, Meixia, et al.
Published: (2025)
by: Lin, Meixia, et al.
Published: (2025)
Towards Robust Model-Based Reinforcement Learning Against Adversarial Corruption
by: Ye, Chenlu, et al.
Published: (2024)
by: Ye, Chenlu, et al.
Published: (2024)
Toward Non-Expert Customized Congestion Control
by: Zhang, Mingrui, et al.
Published: (2026)
by: Zhang, Mingrui, et al.
Published: (2026)
Towards Language-Driven Video Inpainting via Multimodal Large Language Models
by: Wu, Jianzong, et al.
Published: (2024)
by: Wu, Jianzong, et al.
Published: (2024)
RoleMAG: Learning Neighbor Roles in Multimodal Graphs
by: Zuo, Yilong, et al.
Published: (2026)
by: Zuo, Yilong, et al.
Published: (2026)
Prediction of Lane Change Intentions of Human Drivers using an LSTM, a CNN and a Transformer
by: De Cristofaro, Francesco, et al.
Published: (2025)
by: De Cristofaro, Francesco, et al.
Published: (2025)
Role Play: Learning Adaptive Role-Specific Strategies in Multi-Agent Interactions
by: Long, Weifan, et al.
Published: (2024)
by: Long, Weifan, et al.
Published: (2024)
RoleCraft-GLM: Advancing Personalized Role-Playing in Large Language Models
by: Tao, Meiling, et al.
Published: (2023)
by: Tao, Meiling, et al.
Published: (2023)
Learning to play: A Multimodal Agent for 3D Game-Play
by: Yue, Yuguang, et al.
Published: (2025)
by: Yue, Yuguang, et al.
Published: (2025)
Deep Hierarchical Graph Alignment Kernels
by: Tang, Shuhao, et al.
Published: (2024)
by: Tang, Shuhao, et al.
Published: (2024)
Instruction Backdoor Attacks Against Customized LLMs
by: Zhang, Rui, et al.
Published: (2024)
by: Zhang, Rui, et al.
Published: (2024)
RobGC: Towards Robust Graph Condensation
by: Gao, Xinyi, et al.
Published: (2024)
by: Gao, Xinyi, et al.
Published: (2024)
Towards Optimal Customized Architecture for Heterogeneous Federated Learning with Contrastive Cloud-Edge Model Decoupling
by: Chen, Xingyan, et al.
Published: (2024)
by: Chen, Xingyan, et al.
Published: (2024)
Towards Better Generalization via Distributional Input Projection Network
by: Hao, Yifan, et al.
Published: (2025)
by: Hao, Yifan, et al.
Published: (2025)
Training-free Heterogeneous Graph Condensation via Data Selection
by: Liang, Yuxuan, et al.
Published: (2024)
by: Liang, Yuxuan, et al.
Published: (2024)
Understanding Generalization in Role-Playing Models via Information Theory
by: Li, Yongqi, et al.
Published: (2025)
by: Li, Yongqi, et al.
Published: (2025)
Your Self-Play Algorithm is Secretly an Adversarial Imitator: Understanding LLM Self-Play through the Lens of Imitation Learning
by: Li, Shangzhe, et al.
Published: (2026)
by: Li, Shangzhe, et al.
Published: (2026)
Unified Multimodal Vessel Trajectory Prediction with Explainable Navigation Intention
by: Zhang, Rui, et al.
Published: (2025)
by: Zhang, Rui, et al.
Published: (2025)
Customizing Graph Neural Networks using Path Reweighting
by: Chen, Jianpeng, et al.
Published: (2021)
by: Chen, Jianpeng, et al.
Published: (2021)
Node Role-Guided LLMs for Dynamic Graph Clustering
by: Li, Dongyuan, et al.
Published: (2026)
by: Li, Dongyuan, et al.
Published: (2026)
Generalized Spherical Neural Operators: Green's Function Formulation
by: Tang, Hao, et al.
Published: (2025)
by: Tang, Hao, et al.
Published: (2025)
Customized Load Profiles Synthesis for Electricity Customers Based on Conditional Diffusion Models
by: Wang, Zhenyi, et al.
Published: (2023)
by: Wang, Zhenyi, et al.
Published: (2023)
Transformers as Multi-task Learners: Decoupling Features in Hidden Markov Models
by: Hao, Yifan, et al.
Published: (2025)
by: Hao, Yifan, et al.
Published: (2025)
ARMs: Adaptive Red-Teaming Agent against Multimodal Models with Plug-and-Play Attacks
by: Chen, Zhaorun, et al.
Published: (2025)
by: Chen, Zhaorun, et al.
Published: (2025)
Bag of Tricks for Multimodal AutoML with Image, Text, and Tabular Data
by: Tang, Zhiqiang, et al.
Published: (2024)
by: Tang, Zhiqiang, et al.
Published: (2024)
Online Iterative Reinforcement Learning from Human Feedback with General Preference Model
by: Ye, Chenlu, et al.
Published: (2024)
by: Ye, Chenlu, et al.
Published: (2024)
DarkMind: Latent Chain-of-Thought Backdoor in Customized LLMs
by: Guo, Zhen, et al.
Published: (2025)
by: Guo, Zhen, et al.
Published: (2025)
Structured Role-Aware Policy Optimization for Multimodal Reasoning
by: Jiang, Bingqing, et al.
Published: (2026)
by: Jiang, Bingqing, et al.
Published: (2026)
SPAC-Net: Rethinking Point Cloud Completion with Structural Prior
by: Wu, Zizhao, et al.
Published: (2024)
by: Wu, Zizhao, et al.
Published: (2024)
Similar Items
-
Decouple and Track: Benchmarking and Improving Video Diffusion Transformers for Motion Transfer
by: Shi, Qingyu, et al.
Published: (2025) -
Muddit: Liberating Generation Beyond Text-to-Image with a Unified Discrete Diffusion Model
by: Shi, Qingyu, et al.
Published: (2025) -
DreamRelation: Bridging Customization and Relation Generation
by: Shi, Qingyu, et al.
Published: (2024) -
Does Hearing Help Seeing? Investigating Audio-Video Joint Denoising for Video Generation
by: Wu, Jianzong, et al.
Published: (2025) -
MotionBooth: Motion-Aware Customized Text-to-Video Generation
by: Wu, Jianzong, et al.
Published: (2024)