LongCat-Video Technical Report
Fuente:
arXiv
Saved in:
| Main Authors: | Meituan LongCat Team, Cai, Xunliang, Huang, Qilong, Kang, Zhuoliang, Li, Hongyu, Liang, Shijun, Ma, Liya, Ren, Siyu, Wei, Xiaoming, Xie, Rixu, Zhang, Tong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LongCat-Video-Avatar 1.5 Technical Report
by: Meituan LongCat Team, et al.
Published: (2026)
by: Meituan LongCat Team, et al.
Published: (2026)
LongCat-Image Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Flash-Omni Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Flash Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Flash-Thinking-2601 Technical Report
by: Meituan LongCat Team, et al.
Published: (2026)
by: Meituan LongCat Team, et al.
Published: (2026)
Introducing LongCat-Flash-Thinking: A Technical Report
by: Meituan LongCat Team, et al.
Published: (2025)
by: Meituan LongCat Team, et al.
Published: (2025)
LongCat-Next: Lexicalizing Modalities as Discrete Tokens
by: Meituan LongCat Team, et al.
Published: (2026)
by: Meituan LongCat Team, et al.
Published: (2026)
LongCat-AudioDiT: High-Fidelity Diffusion Text-to-Speech in the Waveform Latent Space
by: Xin, Detai, et al.
Published: (2026)
by: Xin, Detai, et al.
Published: (2026)
Efficient Context Scaling with LongCat ZigZag Attention
by: Zhang, Chen, et al.
Published: (2025)
by: Zhang, Chen, et al.
Published: (2025)
LongCat-Audio-Codec: An Audio Tokenizer and Detokenizer Solution Designed for Speech Large Language Models
by: Zhao, Xiaohan, et al.
Published: (2025)
by: Zhao, Xiaohan, et al.
Published: (2025)
Let Them Talk: Audio-Driven Multi-Person Conversational Video Generation
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
LongCat-Flash-Prover: Advancing Native Formal Reasoning via Agentic Tool-Integrated Reinforcement Learning
by: Wang, Jianing, et al.
Published: (2026)
by: Wang, Jianing, et al.
Published: (2026)
InfiniteTalk: Audio-driven Video Generation for Sparse-Frame Video Dubbing
by: Yang, Shaoshu, et al.
Published: (2025)
by: Yang, Shaoshu, et al.
Published: (2025)
G-SMOTE-CatBoost-Optuna
by: G-SMOTE-CatBoost-Optuna
Published: (2025)
by: G-SMOTE-CatBoost-Optuna
Published: (2025)
360Zhinao Technical Report
by: 360Zhinao Team
Published: (2024)
by: 360Zhinao Team
Published: (2024)
Qwen3.5-Omni Technical Report
by: Qwen Team
Published: (2026)
by: Qwen Team
Published: (2026)
Mind DeepResearch Technical Report
by: MindDR Team, et al.
Published: (2026)
by: MindDR Team, et al.
Published: (2026)
WildActor: Unconstrained Identity-Preserving Video Generation
by: Guo, Qin, et al.
Published: (2026)
by: Guo, Qin, et al.
Published: (2026)
Infinite-World: Scaling Interactive World Models to 1000-Frame Horizons via Pose-Free Hierarchical Memory
by: Wu, Ruiqi, et al.
Published: (2026)
by: Wu, Ruiqi, et al.
Published: (2026)
HyperCLOVA X THINK Technical Report
by: NAVER Cloud HyperCLOVA X Team
Published: (2025)
by: NAVER Cloud HyperCLOVA X Team
Published: (2025)
DynamicRad: Content-Adaptive Sparse Attention for Long Video Diffusion
by: Long, Yongji, et al.
Published: (2026)
by: Long, Yongji, et al.
Published: (2026)
Active Intelligence in Video Avatars via Closed-loop World Modeling
by: He, Xuanhua, et al.
Published: (2025)
by: He, Xuanhua, et al.
Published: (2025)
LLIA -- Enabling Low-Latency Interactive Avatars: Real-Time Audio-Driven Portrait Video Generation with Diffusion Models
by: Yu, Haojie, et al.
Published: (2025)
by: Yu, Haojie, et al.
Published: (2025)
U-Mind: A Unified Framework for Real-Time Multimodal Interaction with Audiovisual Generation
by: Deng, Xiang, et al.
Published: (2026)
by: Deng, Xiang, et al.
Published: (2026)
Discovery of 10,059 new three-dimensional periodic orbits of general three-body problem
by: Li, Xiaoming, et al.
Published: (2025)
by: Li, Xiaoming, et al.
Published: (2025)
Plasma‐Based Genomic Features Influencing Outcomes of T790M ‐Positive Non–Small Cell Lung Cancer Receiving Osimertinib
by: Heng Liu, et al.
Published: (2025)
by: Heng Liu, et al.
Published: (2025)
DAM-VSR: Disentanglement of Appearance and Motion for Video Super-Resolution
by: Kong, Zhe, et al.
Published: (2025)
by: Kong, Zhe, et al.
Published: (2025)
Arcee Trinity Large Technical Report
by: Singh, Varun, et al.
Published: (2026)
by: Singh, Varun, et al.
Published: (2026)
Monte Carlo Tree Search for Comprehensive Exploration in LLM-Based Automatic Heuristic Design
by: Zheng, Zhi, et al.
Published: (2025)
by: Zheng, Zhi, et al.
Published: (2025)
Enhancing CVRP Solver through LLM-driven Automatic Heuristic Design
by: Xie, Zhuoliang, et al.
Published: (2026)
by: Xie, Zhuoliang, et al.
Published: (2026)
SparseBalance: Load-Balanced Long Context Training with Dynamic Sparse Attention
by: Xu, Hongtao, et al.
Published: (2026)
by: Xu, Hongtao, et al.
Published: (2026)
Motif-Video 2B: Technical Report
by: Lim, Junghwan, et al.
Published: (2026)
by: Lim, Junghwan, et al.
Published: (2026)
Cut and paste invariants of moduli spaces of stable maps to toric surfaces
by: Rust, Cat
Published: (2026)
by: Rust, Cat
Published: (2026)
Phenotypic and genotypic evaluation of Turkish bread wheat (Triticum aestivum L.) varieties to stripe rust (Puccinia striiformis f.sp. tritici)
by: Ahmet Cat
Published: (2024)
by: Ahmet Cat
Published: (2024)
Scaling Multiagent Systems with Process Rewards
by: Li, Ed, et al.
Published: (2026)
by: Li, Ed, et al.
Published: (2026)
HunyuanVideo 1.5 Technical Report
by: Wu, Bing, et al.
Published: (2025)
by: Wu, Bing, et al.
Published: (2025)
Interpretable Cross-Sphere Multiscale Deep Learning Predicts ENSO Skilfully Beyond 2 Years
by: Hao, Rixu, et al.
Published: (2025)
by: Hao, Rixu, et al.
Published: (2025)
HoloBrain-0 Technical Report
by: Lin, Xuewu, et al.
Published: (2026)
by: Lin, Xuewu, et al.
Published: (2026)
Yi-Lightning Technical Report
by: Wake, Alan, et al.
Published: (2024)
by: Wake, Alan, et al.
Published: (2024)
UserGPT Technical Report
by: Xuan, Yunyi, et al.
Published: (2026)
by: Xuan, Yunyi, et al.
Published: (2026)
Similar Items
-
LongCat-Video-Avatar 1.5 Technical Report
by: Meituan LongCat Team, et al.
Published: (2026) -
LongCat-Image Technical Report
by: Meituan LongCat Team, et al.
Published: (2025) -
LongCat-Flash-Omni Technical Report
by: Meituan LongCat Team, et al.
Published: (2025) -
LongCat-Flash Technical Report
by: Meituan LongCat Team, et al.
Published: (2025) -
LongCat-Flash-Thinking-2601 Technical Report
by: Meituan LongCat Team, et al.
Published: (2026)