PersonaVlog: Personalized Multimodal Vlog Generation with Multi-Agent Collaboration and Iterative Self-Correction
Fuente:
arXiv
Saved in:
| Main Authors: | Hou, Xiaolu, Ma, Bing, Cheng, Jiaxiang, Ren, Xuhua, Yu, Kai, Li, Wenyue, Zheng, Tianxiang, Lu, Qinglin |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
by: Wang, Yuanzhi, et al.
Published: (2026)
by: Wang, Yuanzhi, et al.
Published: (2026)
Implicit Preference Alignment for Human Image Animation
by: Wang, Yuanzhi, et al.
Published: (2026)
by: Wang, Yuanzhi, et al.
Published: (2026)
Phased One-Step Adversarial Equilibrium for Video Diffusion Models
by: Cheng, Jiaxiang, et al.
Published: (2025)
by: Cheng, Jiaxiang, et al.
Published: (2025)
Vlogger: Make Your Dream A Vlog
by: Zhuang, Shaobin, et al.
Published: (2024)
by: Zhuang, Shaobin, et al.
Published: (2024)
Antecedents and Consequences of Viewers' Engagement Toward Vlogging
by: Shanta Banik, et al.
Published: (2025)
by: Shanta Banik, et al.
Published: (2025)
LMVD: A Large-Scale Multimodal Vlog Dataset for Depression Detection in the Wild
by: He, Lang, et al.
Published: (2024)
by: He, Lang, et al.
Published: (2024)
Beyond Reward Margin: Rethinking and Resolving Likelihood Displacement in Diffusion Models via Video Generation
by: Xu, Ruojun, et al.
Published: (2025)
by: Xu, Ruojun, et al.
Published: (2025)
Young People's Experiences of Out‐of‐Home Care as Conveyed by Video Vlogs
by: Marjut Jokela, et al.
Published: (2025)
by: Marjut Jokela, et al.
Published: (2025)
Human Action Co-occurrence in Lifestyle Vlogs using Graph Link Prediction
by: Ignat, Oana, et al.
Published: (2023)
by: Ignat, Oana, et al.
Published: (2023)
Automatic Personality Evaluation from Transliterations of YouTube Vlogs Using Classical and State-of-the-Art Word Embeddings
by: Felipe O. López-Pabón
Published: (2022)
by: Felipe O. López-Pabón
Published: (2022)
VlogQA: Task, Dataset, and Baseline Models for Vietnamese Spoken-Based Machine Reading Comprehension
by: Ngo, Thinh Phuoc, et al.
Published: (2024)
by: Ngo, Thinh Phuoc, et al.
Published: (2024)
How Travel Vlogs Contribute to Destination Marketing: A Comparison with DMO Promotional Videos and the Moderating Role of Destination Competitiveness
by: Ying Zhou, et al.
Published: (2024)
by: Ying Zhou, et al.
Published: (2024)
Analysis of the Application of Grammar and Sentence Structure by Native English Speakers in English Talk Show, Sitcom, and Vlog on the YouTube Platform
by: Ridhwan, Moehamad, et al.
Published: (2025)
by: Ridhwan, Moehamad, et al.
Published: (2025)
Persona-based Multi-Agent Collaboration for Brainstorming
by: Straub, Nate, et al.
Published: (2025)
by: Straub, Nate, et al.
Published: (2025)
SkillGraph: Self-Evolving Multi-Agent Collaboration with Multimodal Graph Topology
by: Nie, Zheng, et al.
Published: (2026)
by: Nie, Zheng, et al.
Published: (2026)
MultiAgent Collaboration Attack: Investigating Adversarial Attacks in Large Language Model Collaborations via Debate
by: Amayuelas, Alfonso, et al.
Published: (2024)
by: Amayuelas, Alfonso, et al.
Published: (2024)
MorphAgent: Empowering Agents through Self-Evolving Profiles and Decentralized Collaboration
by: Lu, Siyuan, et al.
Published: (2024)
by: Lu, Siyuan, et al.
Published: (2024)
Persona-DB: Efficient Large Language Model Personalization for Response Prediction with Collaborative Data Refinement
by: Sun, Chenkai, et al.
Published: (2024)
by: Sun, Chenkai, et al.
Published: (2024)
Reflective Confidence: Correcting Reasoning Flaws via Online Self-Correction
by: Zeng, Qinglin, et al.
Published: (2025)
by: Zeng, Qinglin, et al.
Published: (2025)
Persona Inconstancy in Multi-Agent LLM Collaboration: Conformity, Confabulation, and Impersonation
by: Baltaji, Razan, et al.
Published: (2024)
by: Baltaji, Razan, et al.
Published: (2024)
A Novel Green Logistics Vehicle Scheduling Method Against Road Congestion Utilizing Vehicle–Road–Cloud Collaborative Technology
by: Rui Zheng, et al.
Published: (2025)
by: Rui Zheng, et al.
Published: (2025)
NarrativeLoom: Enhancing Creative Storytelling through Multi-Persona Collaborative Improvisation
by: Ma, Yuxi, et al.
Published: (2026)
by: Ma, Yuxi, et al.
Published: (2026)
PersonaVLM: Long-Term Personalized Multimodal LLMs
by: Nie, Chang, et al.
Published: (2026)
by: Nie, Chang, et al.
Published: (2026)
MCCD: Multi-Agent Collaboration-based Compositional Diffusion for Complex Text-to-Image Generation
by: Li, Mingcheng, et al.
Published: (2025)
by: Li, Mingcheng, et al.
Published: (2025)
Self-Evolving Multi-Agent Collaboration Networks for Software Development
by: Hu, Yue, et al.
Published: (2024)
by: Hu, Yue, et al.
Published: (2024)
Irreducible components of affine Lusztig varieties
by: He, Xuhua
Published: (2025)
by: He, Xuhua
Published: (2025)
On affine Lusztig varieties
by: He, Xuhua
Published: (2023)
by: He, Xuhua
Published: (2023)
Unleashing the Emergent Cognitive Synergy in Large Language Models: A Task-Solving Agent through Multi-Persona Self-Collaboration
by: Wang, Zhenhailong, et al.
Published: (2023)
by: Wang, Zhenhailong, et al.
Published: (2023)
AffectAgent: Collaborative Multi-Agent Reasoning for Retrieval-Augmented Multimodal Emotion Recognition
by: Wang, Zeheng, et al.
Published: (2026)
by: Wang, Zeheng, et al.
Published: (2026)
From Persona to Personalization: A Survey on Role-Playing Language Agents
by: Chen, Jiangjie, et al.
Published: (2024)
by: Chen, Jiangjie, et al.
Published: (2024)
CoAgent: Collaborative Planning and Consistency Agent for Coherent Video Generation
by: Zeng, Qinglin, et al.
Published: (2025)
by: Zeng, Qinglin, et al.
Published: (2025)
MAPS: Multi-Agent Personality Shaping for Collaborative Reasoning
by: Zhang, Jian, et al.
Published: (2025)
by: Zhang, Jian, et al.
Published: (2025)
MMedAgent-RL: Optimizing Multi-Agent Collaboration for Multimodal Medical Reasoning
by: Xia, Peng, et al.
Published: (2025)
by: Xia, Peng, et al.
Published: (2025)
Open-Vocabulary Scene Text Recognition via Pseudo-Image Labeling and Margin Loss
by: Ren, Xuhua, et al.
Published: (2024)
by: Ren, Xuhua, et al.
Published: (2024)
Transparent and Controllable Recommendation Filtering via Multimodal Multi-Agent Collaboration
by: Zhang, Chi, et al.
Published: (2026)
by: Zhang, Chi, et al.
Published: (2026)
HLER: Human-in-the-Loop Economic Research via Multi-Agent Pipelines for Empirical Discovery
by: Zhu, Chen, et al.
Published: (2026)
by: Zhu, Chen, et al.
Published: (2026)
MAC-SQL: A Multi-Agent Collaborative Framework for Text-to-SQL
by: Wang, Bing, et al.
Published: (2023)
by: Wang, Bing, et al.
Published: (2023)
A Survey of Personality, Persona, and Profile in Conversational Agents and Chatbots
by: Sutcliffe, Richard
Published: (2023)
by: Sutcliffe, Richard
Published: (2023)
From Visual Perception to Deep Empathy: An Automated Assessment Framework for House-Tree-Person Drawings Using Multimodal LLMs and Multi-Agent Collaboration
by: Wen, Shuide, et al.
Published: (2025)
by: Wen, Shuide, et al.
Published: (2025)
SOAR: Self-Correction for Optimal Alignment and Refinement in Diffusion Models
by: Qin, You, et al.
Published: (2026)
by: Qin, You, et al.
Published: (2026)
Similar Items
-
FaithfulFaces: Pose-Faithful Facial Identity Preservation for Text-to-Video Generation
by: Wang, Yuanzhi, et al.
Published: (2026) -
Implicit Preference Alignment for Human Image Animation
by: Wang, Yuanzhi, et al.
Published: (2026) -
Phased One-Step Adversarial Equilibrium for Video Diffusion Models
by: Cheng, Jiaxiang, et al.
Published: (2025) -
Vlogger: Make Your Dream A Vlog
by: Zhuang, Shaobin, et al.
Published: (2024) -
Antecedents and Consequences of Viewers' Engagement Toward Vlogging
by: Shanta Banik, et al.
Published: (2025)