Gespeichert in:
| Hauptverfasser: | Yan, Binwei, Fu, Yifei, Zhu, Mingjian, Chen, Hanting, Yuan, Mingxuan, Wang, Yunhe, Hu, Hailin |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2602.10874 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Transferable text data distillation by trajectory matching
von: Yao, Rong, et al.
Veröffentlicht: (2025)
von: Yao, Rong, et al.
Veröffentlicht: (2025)
Saliency-driven Dynamic Token Pruning for Large Language Models
von: Tao, Yao, et al.
Veröffentlicht: (2025)
von: Tao, Yao, et al.
Veröffentlicht: (2025)
MoRAgent: Parameter Efficient Agent Tuning with Mixture-of-Roles
von: Han, Jing, et al.
Veröffentlicht: (2025)
von: Han, Jing, et al.
Veröffentlicht: (2025)
DiJiang: Efficient Large Language Models through Compact Kernelization
von: Chen, Hanting, et al.
Veröffentlicht: (2024)
von: Chen, Hanting, et al.
Veröffentlicht: (2024)
Deferred Commitment Decoding for Diffusion Language Models
von: Shu, Yingte, et al.
Veröffentlicht: (2026)
von: Shu, Yingte, et al.
Veröffentlicht: (2026)
Pangu Embedded: An Efficient Dual-system LLM Reasoner with Metacognition
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
Unshackling Context Length: An Efficient Selective Attention Approach through Query-Key Compression
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
von: Wang, Haoyu, et al.
Veröffentlicht: (2025)
Benchmarking Machine Translation with Cultural Awareness
von: Yao, Binwei, et al.
Veröffentlicht: (2023)
von: Yao, Binwei, et al.
Veröffentlicht: (2023)
MemDLM: Memory-Enhanced DLM Training
von: Pei, Zehua, et al.
Veröffentlicht: (2026)
von: Pei, Zehua, et al.
Veröffentlicht: (2026)
SCOPE: Prompt Evolution for Enhancing Agent Effectiveness
von: Pei, Zehua, et al.
Veröffentlicht: (2025)
von: Pei, Zehua, et al.
Veröffentlicht: (2025)
Rethinking 1-bit Optimization Leveraging Pre-trained Large Language Models
von: Tu, Zhijun, et al.
Veröffentlicht: (2025)
von: Tu, Zhijun, et al.
Veröffentlicht: (2025)
Nexus: Higher-Order Attention Mechanisms in Transformers
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
von: Chen, Hanting, et al.
Veröffentlicht: (2025)
Multi-Perspective Attention Mechanism for Bias-Aware Sequential Recommendation
von: Fu, Mingjian, et al.
Veröffentlicht: (2025)
von: Fu, Mingjian, et al.
Veröffentlicht: (2025)
Omni-Dimensional Frequency Learner for General Time Series Analysis
von: Chen, Xianing, et al.
Veröffentlicht: (2024)
von: Chen, Xianing, et al.
Veröffentlicht: (2024)
DLLM Agent: See Farther, Run Faster
von: Zhen, Huiling, et al.
Veröffentlicht: (2026)
von: Zhen, Huiling, et al.
Veröffentlicht: (2026)
EAQuant: Enhancing Post-Training Quantization for MoE Models via Expert-Aware Optimization
von: Fu, Zhongqian, et al.
Veröffentlicht: (2025)
von: Fu, Zhongqian, et al.
Veröffentlicht: (2025)
Multiscale Positive-Unlabeled Detection of AI-Generated Texts
von: Tian, Yuchuan, et al.
Veröffentlicht: (2023)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2023)
Near-Policy: Accelerating On-Policy Distillation via Asynchronous Generation and Selective Packing
von: Rang, Miao, et al.
Veröffentlicht: (2026)
von: Rang, Miao, et al.
Veröffentlicht: (2026)
PanGu-$π$: Enhancing Language Model Architectures via Nonlinearity Compensation
von: Wang, Yunhe, et al.
Veröffentlicht: (2023)
von: Wang, Yunhe, et al.
Veröffentlicht: (2023)
DP-GTR: Differentially Private Prompt Protection via Group Text Rewriting
von: Li, Mingchen, et al.
Veröffentlicht: (2025)
von: Li, Mingchen, et al.
Veröffentlicht: (2025)
Multi-Granularity Semantic Revision for Large Language Model Distillation
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Liu, Xiaoyu, et al.
Veröffentlicht: (2024)
A Pluggable Multi-Task Learning Framework for Sentiment-Aware Financial Relation Extraction
von: Luo, Jinming, et al.
Veröffentlicht: (2025)
von: Luo, Jinming, et al.
Veröffentlicht: (2025)
Towards Efficient Agents: A Co-Design of Inference Architecture and System
von: Lin, Weizhe, et al.
Veröffentlicht: (2025)
von: Lin, Weizhe, et al.
Veröffentlicht: (2025)
A Survey on Transformer Compression
von: Tang, Yehui, et al.
Veröffentlicht: (2024)
von: Tang, Yehui, et al.
Veröffentlicht: (2024)
RDEx-MOP: Indicator-Guided Reconstructed Differential Evolution for Fixed-Budget Multiobjective Optimization
von: Tao, Sichen, et al.
Veröffentlicht: (2026)
von: Tao, Sichen, et al.
Veröffentlicht: (2026)
CFinBench: A Comprehensive Chinese Financial Benchmark for Large Language Models
von: Nie, Ying, et al.
Veröffentlicht: (2024)
von: Nie, Ying, et al.
Veröffentlicht: (2024)
C-Evolve: Consensus-based Evolution for Prompt Groups
von: Li, Tiancheng, et al.
Veröffentlicht: (2025)
von: Li, Tiancheng, et al.
Veröffentlicht: (2025)
Introducing MAPO: Momentum-Aided Gradient Descent Prompt Optimization
von: Cui, Anthony, et al.
Veröffentlicht: (2024)
von: Cui, Anthony, et al.
Veröffentlicht: (2024)
SITA: Learning Speaker-Invariant and Tone-Aware Speech Representations for Low-Resource Tonal Languages
von: Xu, Tianyi, et al.
Veröffentlicht: (2026)
von: Xu, Tianyi, et al.
Veröffentlicht: (2026)
CBQ: Cross-Block Quantization for Large Language Models
von: Ding, Xin, et al.
Veröffentlicht: (2023)
von: Ding, Xin, et al.
Veröffentlicht: (2023)
AI as a deliberative partner fosters intercultural empathy for Americans but fails for Latin American participants
von: Villanueva, Isabel, et al.
Veröffentlicht: (2025)
von: Villanueva, Isabel, et al.
Veröffentlicht: (2025)
Question-Aware Knowledge Graph Prompting for Enhancing Large Language Models
von: Liu, Haochen, et al.
Veröffentlicht: (2025)
von: Liu, Haochen, et al.
Veröffentlicht: (2025)
IAPT: Instruction-Aware Prompt Tuning for Large Language Models
von: Zhu, Wei, et al.
Veröffentlicht: (2024)
von: Zhu, Wei, et al.
Veröffentlicht: (2024)
Ace-Skill: Bootstrapping Multimodal Agents with Prioritized and Clustered Evolution
von: Xiong, Feng, et al.
Veröffentlicht: (2026)
von: Xiong, Feng, et al.
Veröffentlicht: (2026)
CAF-I: A Collaborative Multi-Agent Framework for Enhanced Irony Detection with Large Language Models
von: Liu, Ziqi., et al.
Veröffentlicht: (2025)
von: Liu, Ziqi., et al.
Veröffentlicht: (2025)
Beyond Prompt Content: Enhancing LLM Performance via Content-Format Integrated Prompt Optimization
von: Liu, Yuanye, et al.
Veröffentlicht: (2025)
von: Liu, Yuanye, et al.
Veröffentlicht: (2025)
Mixture of In-Context Experts Enhance LLMs' Long Context Awareness
von: Lin, Hongzhan, et al.
Veröffentlicht: (2024)
von: Lin, Hongzhan, et al.
Veröffentlicht: (2024)
SPAM: Spike-Aware Adam with Momentum Reset for Stable LLM Training
von: Huang, Tianjin, et al.
Veröffentlicht: (2025)
von: Huang, Tianjin, et al.
Veröffentlicht: (2025)
The Detection-Extraction Gap: Models Know the Answer Before They Can Say It
von: Wang, Hanyang, et al.
Veröffentlicht: (2026)
von: Wang, Hanyang, et al.
Veröffentlicht: (2026)
Towards Reliable and Empathetic Depression-Diagnosis-Oriented Chats
von: Lan, Kunyao, et al.
Veröffentlicht: (2024)
von: Lan, Kunyao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Transferable text data distillation by trajectory matching
von: Yao, Rong, et al.
Veröffentlicht: (2025) -
Saliency-driven Dynamic Token Pruning for Large Language Models
von: Tao, Yao, et al.
Veröffentlicht: (2025) -
MoRAgent: Parameter Efficient Agent Tuning with Mixture-of-Roles
von: Han, Jing, et al.
Veröffentlicht: (2025) -
DiJiang: Efficient Large Language Models through Compact Kernelization
von: Chen, Hanting, et al.
Veröffentlicht: (2024) -
Deferred Commitment Decoding for Diffusion Language Models
von: Shu, Yingte, et al.
Veröffentlicht: (2026)