Megrez-Omni Technical Report
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Boxun, Li, Yadong, Li, Zhiyuan, Liu, Congyi, Liu, Weilin, Niu, Guowei, Tan, Zheyue, Xu, Haiyang, Yao, Zhuyu, Yuan, Tao, Zhou, Dong, Zhuang, Yueqing, Yan, Shengen, Dai, Guohao, Wang, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Megrez2 Technical Report
von: Li, Boxun, et al.
Veröffentlicht: (2025)
von: Li, Boxun, et al.
Veröffentlicht: (2025)
ReXMoE: Reusing Experts with Minimal Overhead in Mixture-of-Experts
von: Tan, Zheyue, et al.
Veröffentlicht: (2025)
von: Tan, Zheyue, et al.
Veröffentlicht: (2025)
LV-Eval: A Balanced Long-Context Benchmark with 5 Length Levels Up to 256K
von: Yuan, Tao, et al.
Veröffentlicht: (2024)
von: Yuan, Tao, et al.
Veröffentlicht: (2024)
Baichuan-Omni Technical Report
von: Li, Yadong, et al.
Veröffentlicht: (2024)
von: Li, Yadong, et al.
Veröffentlicht: (2024)
Kling-Omni Technical Report
von: Kling Team, et al.
Veröffentlicht: (2025)
von: Kling Team, et al.
Veröffentlicht: (2025)
SAIL-Embedding Technical Report: Omni-modal Embedding Foundation Model
von: Lin, Lin, et al.
Veröffentlicht: (2025)
von: Lin, Lin, et al.
Veröffentlicht: (2025)
Baichuan-Omni-1.5 Technical Report
von: Li, Yadong, et al.
Veröffentlicht: (2025)
von: Li, Yadong, et al.
Veröffentlicht: (2025)
EARL: Efficient Agentic Reinforcement Learning Systems for Large Language Models
von: Tan, Zheyue, et al.
Veröffentlicht: (2025)
von: Tan, Zheyue, et al.
Veröffentlicht: (2025)
Logics-Parsing-Omni Technical Report
von: An, Xin, et al.
Veröffentlicht: (2026)
von: An, Xin, et al.
Veröffentlicht: (2026)
LongCat-Flash-Omni Technical Report
von: Meituan LongCat Team, et al.
Veröffentlicht: (2025)
von: Meituan LongCat Team, et al.
Veröffentlicht: (2025)
BitSnap: Checkpoint Sparsification and Quantization in LLM Training
von: Peng, Yanxin, et al.
Veröffentlicht: (2025)
von: Peng, Yanxin, et al.
Veröffentlicht: (2025)
Qwen3-Omni Technical Report
von: Xu, Jin, et al.
Veröffentlicht: (2025)
von: Xu, Jin, et al.
Veröffentlicht: (2025)
OmniFusion Technical Report
von: Goncharova, Elizaveta, et al.
Veröffentlicht: (2024)
von: Goncharova, Elizaveta, et al.
Veröffentlicht: (2024)
CSKV: Training-Efficient Channel Shrinking for KV Cache in Long-Context Scenarios
von: Wang, Luning, et al.
Veröffentlicht: (2024)
von: Wang, Luning, et al.
Veröffentlicht: (2024)
Distilled Decoding 2: One-step Sampling of Image Auto-regressive Models with Conditional Score Distillation
von: Liu, Enshu, et al.
Veröffentlicht: (2025)
von: Liu, Enshu, et al.
Veröffentlicht: (2025)
Baichuan Alignment Technical Report
von: Lin, Mingan, et al.
Veröffentlicht: (2024)
von: Lin, Mingan, et al.
Veröffentlicht: (2024)
Evaluating Quantized Large Language Models
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
von: Li, Shiyao, et al.
Veröffentlicht: (2024)
Qwen3.5-Omni Technical Report
von: Qwen Team
Veröffentlicht: (2026)
von: Qwen Team
Veröffentlicht: (2026)
Qwen2.5-Omni Technical Report
von: Xu, Jin, et al.
Veröffentlicht: (2025)
von: Xu, Jin, et al.
Veröffentlicht: (2025)
Octopus v3: Technical Report for On-device Sub-billion Multimodal AI Agent
von: Chen, Wei, et al.
Veröffentlicht: (2024)
von: Chen, Wei, et al.
Veröffentlicht: (2024)
An Economical and Efficient Helium Recovery System for Vibration-Sensitive Applications
von: Yin, Zhiyuan, et al.
Veröffentlicht: (2024)
von: Yin, Zhiyuan, et al.
Veröffentlicht: (2024)
QuarkMed Medical Foundation Model Technical Report
von: Li, Ao, et al.
Veröffentlicht: (2025)
von: Li, Ao, et al.
Veröffentlicht: (2025)
FrameFusion: Combining Similarity and Importance for Video Token Reduction on Large Vision Language Models
von: Fu, Tianyu, et al.
Veröffentlicht: (2024)
von: Fu, Tianyu, et al.
Veröffentlicht: (2024)
X-OmniClaw Technical Report: A Unified Mobile Agent for Multimodal Understanding and Interaction
von: Ren, Xiaoming, et al.
Veröffentlicht: (2026)
von: Ren, Xiaoming, et al.
Veröffentlicht: (2026)
PM-KVQ: Progressive Mixed-precision KV Cache Quantization for Long-CoT LLMs
von: Liu, Tengxuan, et al.
Veröffentlicht: (2025)
von: Liu, Tengxuan, et al.
Veröffentlicht: (2025)
COPF: An Online Framework for Deployment-Stable Counterfactual Fairness in Evolving Graphs
von: Li, Sheng'en, et al.
Veröffentlicht: (2026)
von: Li, Sheng'en, et al.
Veröffentlicht: (2026)
CASTLE2026 Team WDL Technical Report
von: Li, Zhengyang, et al.
Veröffentlicht: (2026)
von: Li, Zhengyang, et al.
Veröffentlicht: (2026)
ABot-OCR Technical Report
von: Jiang, Kaitao, et al.
Veröffentlicht: (2026)
von: Jiang, Kaitao, et al.
Veröffentlicht: (2026)
(Debiased) Contrastive Learning Loss for Recommendation (Technical Report)
von: Jin, Ruoming, et al.
Veröffentlicht: (2023)
von: Jin, Ruoming, et al.
Veröffentlicht: (2023)
Causal Relationship Between Autism Spectrum Disorder and Inflammatory Bowel Disease: A Bidirectional Mendelian Randomization Study
von: Weilin Li, et al.
Veröffentlicht: (2024)
von: Weilin Li, et al.
Veröffentlicht: (2024)
asvspoof2015 in WebDataset Format
von: Yadong, Niu
Veröffentlicht: (2025)
von: Yadong, Niu
Veröffentlicht: (2025)
vocalimitationset in WebDataset Format
von: Yadong, Niu
Veröffentlicht: (2025)
von: Yadong, Niu
Veröffentlicht: (2025)
Scripts for DN1 single cell data analysis
von: Hu, Shengen
Veröffentlicht: (2025)
von: Hu, Shengen
Veröffentlicht: (2025)
MixDQ: Memory-Efficient Few-Step Text-to-Image Diffusion Models with Metric-Decoupled Mixed Precision Quantization
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
von: Zhao, Tianchen, et al.
Veröffentlicht: (2024)
R2R: Efficiently Navigating Divergent Reasoning Paths with Small-Large Model Token Routing
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
von: Fu, Tianyu, et al.
Veröffentlicht: (2025)
Yi-Lightning Technical Report
von: Wake, Alan, et al.
Veröffentlicht: (2024)
von: Wake, Alan, et al.
Veröffentlicht: (2024)
Omni-Scene: Omni-Gaussian Representation for Ego-Centric Sparse-View Scene Reconstruction
von: Wei, Dongxu, et al.
Veröffentlicht: (2024)
von: Wei, Dongxu, et al.
Veröffentlicht: (2024)
Nyonic Technical Report
von: Tian, Junfeng, et al.
Veröffentlicht: (2024)
von: Tian, Junfeng, et al.
Veröffentlicht: (2024)
Mogao: An Omni Foundation Model for Interleaved Multi-Modal Generation
von: Liao, Chao, et al.
Veröffentlicht: (2025)
von: Liao, Chao, et al.
Veröffentlicht: (2025)
GR-3 Technical Report
von: Cheang, Chilam, et al.
Veröffentlicht: (2025)
von: Cheang, Chilam, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Megrez2 Technical Report
von: Li, Boxun, et al.
Veröffentlicht: (2025) -
ReXMoE: Reusing Experts with Minimal Overhead in Mixture-of-Experts
von: Tan, Zheyue, et al.
Veröffentlicht: (2025) -
LV-Eval: A Balanced Long-Context Benchmark with 5 Length Levels Up to 256K
von: Yuan, Tao, et al.
Veröffentlicht: (2024) -
Baichuan-Omni Technical Report
von: Li, Yadong, et al.
Veröffentlicht: (2024) -
Kling-Omni Technical Report
von: Kling Team, et al.
Veröffentlicht: (2025)