MiniCPM4: Ultra-Efficient LLMs on End Devices
Fuente:
arXiv
Saved in:
Similar Items
MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
by: MiniCPM Team, et al.
Published: (2026)
by: MiniCPM Team, et al.
Published: (2026)
InfLLM-V2: Dense-Sparse Switchable Attention for Seamless Short-to-Long Adaptation
by: Zhao, Weilin, et al.
Published: (2025)
by: Zhao, Weilin, et al.
Published: (2025)
Effectively Enhancing Vision Language Large Models by Prompt Augmentation and Caption Utilization
by: Zhao, Minyi, et al.
Published: (2024)
by: Zhao, Minyi, et al.
Published: (2024)
Robust Multimodal Large Language Models Against Modality Conflict
by: Zhang, Zongmeng, et al.
Published: (2025)
by: Zhang, Zongmeng, et al.
Published: (2025)
BURN: Backdoor Unlearning via Adversarial Boundary Analysis
by: Su, Yanghao, et al.
Published: (2025)
by: Su, Yanghao, et al.
Published: (2025)
Character as a Latent Variable in Large Language Models: A Mechanistic Account of Emergent Misalignment and Conditional Safety Failures
by: Su, Yanghao, et al.
Published: (2026)
by: Su, Yanghao, et al.
Published: (2026)
Evaluating Accounting Reasoning Capabilities of Large Language Models
by: Zhou, Jie, et al.
Published: (2026)
by: Zhou, Jie, et al.
Published: (2026)
Always a good thing? The resource perspective to understand the double‐edged sword effect of empowering leadership on employee expediency
by: Jie Xiao, et al.
Published: (2026)
by: Jie Xiao, et al.
Published: (2026)
Look, Listen and Segment: Towards Weakly Supervised Audio-visual Semantic Segmentation
by: Li, Chengzhi, et al.
Published: (2026)
by: Li, Chengzhi, et al.
Published: (2026)
MiniCPM-V: A GPT-4V Level MLLM on Your Phone
by: Yao, Yuan, et al.
Published: (2024)
by: Yao, Yuan, et al.
Published: (2024)
LeechHijack: Covert Computational Resource Exploitation in Intelligent Agent Systems
by: Zhang, Yuanhe, et al.
Published: (2025)
by: Zhang, Yuanhe, et al.
Published: (2025)
MiniCPM: Unveiling the Potential of Small Language Models with Scalable Training Strategies
by: Hu, Shengding, et al.
Published: (2024)
by: Hu, Shengding, et al.
Published: (2024)
Multiple types of spin textures and robust valley physics in MP$_2$X$_6$
by: Liang, Li, et al.
Published: (2024)
by: Liang, Li, et al.
Published: (2024)
Accounting Reasoning in Large Language Models: Concepts, Evaluation, and Empirical Analysis
by: Zhou, Jie, et al.
Published: (2025)
by: Zhou, Jie, et al.
Published: (2025)
Pardon? Evaluating Conversational Repair in Large Audio-Language Models
by: Huang, Shuanghong, et al.
Published: (2026)
by: Huang, Shuanghong, et al.
Published: (2026)
Delicately Tailoring Morphology of Silver Phosphate Crystals with High‐Gravity
by: Qin Zhou, et al.
Published: (2025)
by: Qin Zhou, et al.
Published: (2025)
An Efficient Approach to the Design and Construction of Novel Energetic Materials: Co‐Crystallization and Chemical Reaction of Insensitive Explosive ICM‐102 in Acidic Solutions
by: Fei Wang, et al.
Published: (2024)
by: Fei Wang, et al.
Published: (2024)
Tuning Stability and Optoelectronic Properties of CsSnI 3 Via Rare‐Earth Doping
by: Haowei Wang, et al.
Published: (2026)
by: Haowei Wang, et al.
Published: (2026)
Detecting Performance Degradation under Data Shift in Pathology Vision-Language Model
by: Guan, Hao, et al.
Published: (2026)
by: Guan, Hao, et al.
Published: (2026)
TransFR: Transferable Federated Recommendation with Adapter Tuning on Pre-trained Language Models
by: Zhang, Honglei, et al.
Published: (2024)
by: Zhang, Honglei, et al.
Published: (2024)
Fabrication of Multifunctional Integrated Composites with Excellent Mechanical Synergistic Effects Producing from Aramid Honeycomb and Rigid Polyimide Foam
by: Jingjing Li, et al.
Published: (2025)
by: Jingjing Li, et al.
Published: (2025)
Provably Efficient Reinforcement Learning with Multinomial Logit Function Approximation
by: Li, Long-Fei, et al.
Published: (2024)
by: Li, Long-Fei, et al.
Published: (2024)
Text or Pixels? It Takes Half: On the Token Efficiency of Visual Text Inputs in Multimodal LLMs
by: Li, Yanhong, et al.
Published: (2025)
by: Li, Yanhong, et al.
Published: (2025)
Double Distillation Network for Multi-Agent Reinforcement Learning
by: Zhou, Yang, et al.
Published: (2025)
by: Zhou, Yang, et al.
Published: (2025)
FPGA-based Emulation and Device-Side Management for CXL-based Memory Tiering Systems
by: Chen, Yiqi, et al.
Published: (2025)
by: Chen, Yiqi, et al.
Published: (2025)
Percutaneous Tract Embolization Versus Conventional Drainage Following Percutaneous Transhepatic Cholangioscopy for Biliary Stones
by: Mengying Zhao, et al.
Published: (2025)
by: Mengying Zhao, et al.
Published: (2025)
Alpha‐Lipoic Acid Alleviates Isoflurane‐Induced Cognitive Dysfunction in Juvenile Mice by Activating Peroxisome Proliferator Activated Receptor Gamma Coactivator ‐1 Alpha
by: Xingkai Zhao, et al.
Published: (2025)
by: Xingkai Zhao, et al.
Published: (2025)
Staggered nonlinear spin generations in centrosymmetric altermagnets under electric current
by: Zhang, Jie, et al.
Published: (2025)
by: Zhang, Jie, et al.
Published: (2025)
Automatic Knowledge Graph Construction for Judicial Cases
by: Zhou, Jie, et al.
Published: (2024)
by: Zhou, Jie, et al.
Published: (2024)
Dense Cross-Scale Image Alignment With Fully Spatial Correlation and Just Noticeable Difference Guidance
by: You, Jinkun, et al.
Published: (2025)
by: You, Jinkun, et al.
Published: (2025)
Magnetically controllable nonlinear valley Hall effect in centrosymmetric ferromagnets
by: Fang, Ruijing, et al.
Published: (2025)
by: Fang, Ruijing, et al.
Published: (2025)
Modular Assembly of 2,3‐Dihydro‐4 H ‐benzothiazinones via a Multicomponent Cyclization Strategy
by: Guozheng Gu, et al.
Published: (2026)
by: Guozheng Gu, et al.
Published: (2026)
Domain Similarity-Perceived Label Assignment for Domain Generalized Underwater Object Detection
by: Li, Xisheng, et al.
Published: (2023)
by: Li, Xisheng, et al.
Published: (2023)
Adaptive Development of Soil Bacterial Communities to Ecological Processes Caused by Mining Subsidence
by: Yan Yu, et al.
Published: (2025)
by: Yan Yu, et al.
Published: (2025)
Video-based Sign Language Recognition without Temporal Segmentation
by: Huang, Jie, et al.
Published: (2018)
by: Huang, Jie, et al.
Published: (2018)
Forest2Seq: Revitalizing Order Prior for Sequential Indoor Scene Synthesis
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
Key Materials for Potassium‐Ion Batteries: Overcoming Challenges and Opening Up Horizons for Commercialization
by: Zhiwang Liu, et al.
Published: (2026)
by: Zhiwang Liu, et al.
Published: (2026)
Dehallu3D: Hallucination-Mitigated 3D Generation from Single Image via Cyclic View Consistency Refinement
by: Wang, Xiwen, et al.
Published: (2026)
by: Wang, Xiwen, et al.
Published: (2026)
MQADet: A Plug-and-Play Paradigm for Enhancing Open-Vocabulary Object Detection via Multimodal Question Answering
by: Li, Caixiong, et al.
Published: (2025)
by: Li, Caixiong, et al.
Published: (2025)
LARFT: Closing the Cognition-Action Gap for Length Instruction Following in Large Language Models
by: Zhang, Wei, et al.
Published: (2026)
by: Zhang, Wei, et al.
Published: (2026)
Similar Items
-
MiniCPM-SALA: Hybridizing Sparse and Linear Attention for Efficient Long-Context Modeling
by: MiniCPM Team, et al.
Published: (2026) -
InfLLM-V2: Dense-Sparse Switchable Attention for Seamless Short-to-Long Adaptation
by: Zhao, Weilin, et al.
Published: (2025) -
Effectively Enhancing Vision Language Large Models by Prompt Augmentation and Caption Utilization
by: Zhao, Minyi, et al.
Published: (2024) -
Robust Multimodal Large Language Models Against Modality Conflict
by: Zhang, Zongmeng, et al.
Published: (2025) -
BURN: Backdoor Unlearning via Adversarial Boundary Analysis
by: Su, Yanghao, et al.
Published: (2025)