LongCat-Video-Avatar 1.5 Technical Report
Fuente:
arXiv
Saved in:
| Main Authors: | Meituan LongCat Team, Cai, Xunliang, Cheng, Meng, Gao, Feng, Kong, Zhe, Li, Jiamu, Li, Le, Li, Weiheng, Liu, Hongyu, Tan, Shuai, Wei, Xiaoming, Yang, Tianyu, Zhang, Yong |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LongCat-Video Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Image Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Flash Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Flash-Omni Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Flash-Thinking-2601 Technical Report
by: Meituan LongCat Team, et al.
Published: (2026)
by: Meituan LongCat Team, et al.
Published: (2026)
Introducing LongCat-Flash-Thinking: A Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Next: Lexicalizing Modalities as Discrete Tokens
by: Meituan LongCat Team, et al.
Published: (2026)
by: Meituan LongCat Team, et al.
Published: (2026)
Efficient Context Scaling with LongCat ZigZag Attention
by: Zhang, Chen, et al.
Published: (2025)
by: Zhang, Chen, et al.
Published: (2025)
LongCat-AudioDiT: High-Fidelity Diffusion Text-to-Speech in the Waveform Latent Space
by: Xin, Detai, et al.
Published: (2026)
by: Xin, Detai, et al.
Published: (2026)
LongCat-Audio-Codec: An Audio Tokenizer and Detokenizer Solution Designed for Speech Large Language Models
by: Zhao, Xiaohan, et al.
Published: (2025)
by: Zhao, Xiaohan, et al.
Published: (2025)
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning
by: Wang, Jianing, et al.
Published: (2026)
by: Wang, Jianing, et al.
Published: (2026)
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
by: Yu, Haojie, et al.
Published: (2025)
by: Yu, Haojie, et al.
Published: (2025)
Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
G-SMOTE-CatBoost-Optuna
by: G-SMOTE-CatBoost-Optuna
Published: (2025)
by: G-SMOTE-CatBoost-Optuna
Published: (2025)
MeshAvatar: Learning High-quality Triangular Human Avatars from Multi-view Videos
by: Chen, Yushuo, et al.
Published: (2024)
by: Chen, Yushuo, et al.
Published: (2024)
InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing
by: Yang, Shaoshu, et al.
Published: (2025)
by: Yang, Shaoshu, et al.
Published: (2025)
STG-Avatar: Animatable Human Avatars via Spacetime Gaussian
by: Jiang, Guangan, et al.
Published: (2025)
by: Jiang, Guangan, et al.
Published: (2025)
Mind DeepResearch Technical Report
by: MindDR Team, et al.
Published: (2026)
by: MindDR Team, et al.
Published: (2026)
HunyuanVideo 1.5 Technical Report
by: Wu, Bing, et al.
Published: (2025)
by: Wu, Bing, et al.
Published: (2025)
KlingAvatar 2.0 Technical Report
by: Kling Team, et al.
Published: (2025)
by: Kling Team, et al.
Published: (2025)
Active Intelligence in Video Avatars via Closed-loop World Modeling
by: He, Xuanhua, et al.
Published: (2025)
by: He, Xuanhua, et al.
Published: (2025)
Capability at a Glance: Design Guidelines for Intuitive Avatars Communicating Augmented Actions in Virtual Reality
by: Lu, Yang, et al.
Published: (2026)
by: Lu, Yang, et al.
Published: (2026)
Qwen3.5-Omni Technical Report
by: Qwen Team
Published: (2026)
by: Qwen Team
Published: (2026)
SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention
by: Xu, Hongtao, et al.
Published: (2026)
by: Xu, Hongtao, et al.
Published: (2026)
360Zhinao Technical Report
by: 360Zhinao Team
Published: (2024)
by: 360Zhinao Team
Published: (2024)
GaussianAvatar-Editor: Photorealistic Animatable Gaussian Head Avatar Editor
by: Liu, Xiangyue, et al.
Published: (2025)
by: Liu, Xiangyue, et al.
Published: (2025)
Deblur-Avatar: Animatable Avatars from Motion-Blurred Monocular Videos
by: Luo, Xianrui, et al.
Published: (2025)
by: Luo, Xianrui, et al.
Published: (2025)
InstructAvatar: Text-Guided Emotion and Motion Control for Avatar Generation
by: Wang, Yuchi, et al.
Published: (2024)
by: Wang, Yuchi, et al.
Published: (2024)
TexVocab: Texture Vocabulary-conditioned Human Avatars
by: Liu, Yuxiao, et al.
Published: (2024)
by: Liu, Yuxiao, et al.
Published: (2024)
UI-Venus-1.5 Technical Report
by: Venus Team, et al.
Published: (2026)
by: Venus Team, et al.
Published: (2026)
Generating Editable Head Avatars with 3D Gaussian GANs
by: Li, Guohao, et al.
Published: (2024)
by: Li, Guohao, et al.
Published: (2024)
Vid2Avatar-Pro: Authentic Avatar from Videos in the Wild via Universal Prior
by: Guo, Chen, et al.
Published: (2025)
by: Guo, Chen, et al.
Published: (2025)
Long-RVOS: A Comprehensive Benchmark for Long-term Referring Video Object Segmentation
by: Liang, Tianming, et al.
Published: (2025)
by: Liang, Tianming, et al.
Published: (2025)
ReMamba: Equip Mamba with Effective Long-Sequence Modeling
by: Yuan, Danlong, et al.
Published: (2024)
by: Yuan, Danlong, et al.
Published: (2024)
2D Gaussian Splatting with Semantic Alignment for Image Inpainting
by: Li, Hongyu, et al.
Published: (2025)
by: Li, Hongyu, et al.
Published: (2025)
Refined Geometry-guided Head Avatar Reconstruction from Monocular RGB Video
by: Park, Pilseo, et al.
Published: (2025)
by: Park, Pilseo, et al.
Published: (2025)
AMemGym: Interactive Memory Benchmarking for Assistants in Long-Horizon Conversations
by: Jiayang, Cheng, et al.
Published: (2026)
by: Jiayang, Cheng, et al.
Published: (2026)
DAM-VSR: Disentanglement of Appearance and Motion for Video Super-Resolution
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
Zero-Shot Long-Form Video Understanding through Screenplay
by: Wu, Yongliang, et al.
Published: (2024)
by: Wu, Yongliang, et al.
Published: (2024)
How Green Innovation and Green Corporate Social Responsibility Transform Green Transformational Leadership Into Sustainable Performance? Evidence From an Emerging Economy
by: Thanh Tiep Le, et al.
Published: (2024)
by: Thanh Tiep Le, et al.
Published: (2024)
Similar Items
-
LongCat-Video Technical Report
by: Meituan LongCat Team, et al.
Published: (2025) -
LongCat-Image Technical Report
by: Meituan LongCat Team, et al.
Published: (2025) -
LongCat-Flash Technical Report
by: Meituan LongCat Team, et al.
Published: (2025) -
LongCat-Flash-Omni Technical Report
by: Meituan LongCat Team, et al.
Published: (2025) -
LongCat-Flash-Thinking-2601 Technical Report
by: Meituan LongCat Team, et al.
Published: (2026)