MEND: Meta dEmonstratioN Distillation for Efficient and Effective In-Context Learning
Fuente:
arXiv
Guardado en:
| Autores principales: | Li, Yichuan, Ma, Xiyao, Lu, Sixing, Lee, Kyumin, Liu, Xiaohu, Guo, Chenlei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Empowering Large Language Models for Textual Data Augmentation
por: Li, Yichuan, et al.
Publicado: (2024)
por: Li, Yichuan, et al.
Publicado: (2024)
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
por: Liu, Zeyu Leo, et al.
Publicado: (2025)
por: Liu, Zeyu Leo, et al.
Publicado: (2025)
From Critique to Clarity: A Pathway to Faithful and Personalized Code Explanations with Large Language Models
por: Xu, Zexing, et al.
Publicado: (2024)
por: Xu, Zexing, et al.
Publicado: (2024)
STEPER: Step-wise Knowledge Distillation for Enhancing Reasoning Ability in Multi-Step Retrieval-Augmented Language Models
por: Lee, Kyumin, et al.
Publicado: (2025)
por: Lee, Kyumin, et al.
Publicado: (2025)
Reducing and Exploiting Data Augmentation Noise through Meta Reweighting Contrastive Learning for Text Classification
por: Mou, Guanyi, et al.
Publicado: (2024)
por: Mou, Guanyi, et al.
Publicado: (2024)
Multi-Stage Balanced Distillation: Addressing Long-Tail Challenges in Sequence-Level Knowledge Distillation
por: Zhou, Yuhang, et al.
Publicado: (2024)
por: Zhou, Yuhang, et al.
Publicado: (2024)
Piece of Table: A Divide-and-Conquer Approach for Selecting Subtables in Table Question Answering
por: Lee, Wonjin, et al.
Publicado: (2024)
por: Lee, Wonjin, et al.
Publicado: (2024)
Hybrid Linear Attention Done Right: Efficient Distillation and Effective Architectures for Extremely Long Contexts
por: Chen, Yingfa, et al.
Publicado: (2026)
por: Chen, Yingfa, et al.
Publicado: (2026)
Meta-Sel: Efficient Demonstration Selection for In-Context Learning via Supervised Meta-Learning
por: Wang, Xubin, et al.
Publicado: (2026)
por: Wang, Xubin, et al.
Publicado: (2026)
TSUBASA: Improving Long-Horizon Personalization via Evolving Memory and Self-Learning with Context Distillation
por: Zhang, Xinliang Frederick, et al.
Publicado: (2026)
por: Zhang, Xinliang Frederick, et al.
Publicado: (2026)
Quantity Convergence, Quality Divergence: Disentangling Fluency and Accuracy in L2 Mandarin Prosody
por: Shi, Yuqi, et al.
Publicado: (2026)
por: Shi, Yuqi, et al.
Publicado: (2026)
Beyond Perfect APIs: A Comprehensive Evaluation of LLM Agents Under Real-World API Complexity
por: Kim, Doyoung, et al.
Publicado: (2026)
por: Kim, Doyoung, et al.
Publicado: (2026)
END: Early Noise Dropping for Efficient and Effective Context Denoising
por: Jin, Hongye, et al.
Publicado: (2025)
por: Jin, Hongye, et al.
Publicado: (2025)
Meta In-Context Learning Makes Large Language Models Better Zero and Few-Shot Relation Extractors
por: Li, Guozheng, et al.
Publicado: (2024)
por: Li, Guozheng, et al.
Publicado: (2024)
LAWCAT: Efficient Distillation from Quadratic to Linear Attention with Convolution across Tokens for Long Context Modeling
por: Liu, Zeyu, et al.
Publicado: (2025)
por: Liu, Zeyu, et al.
Publicado: (2025)
From Deferral to Learning: Online In-Context Knowledge Distillation for LLM Cascades
por: Wu, Yu, et al.
Publicado: (2025)
por: Wu, Yu, et al.
Publicado: (2025)
Context Memorization for Efficient Long Context Generation
por: Okoshi, Yasuyuki, et al.
Publicado: (2026)
por: Okoshi, Yasuyuki, et al.
Publicado: (2026)
Evaluation of Coding Schemes for Transformer-based Gene Sequence Modeling
por: Gong, Chenlei, et al.
Publicado: (2025)
por: Gong, Chenlei, et al.
Publicado: (2025)
MetaSICL: Adapting Audiroty LLM via Meta Speech In-Context Learning
por: Zheng, Haolong, et al.
Publicado: (2026)
por: Zheng, Haolong, et al.
Publicado: (2026)
Distilling Monolingual and Crosslingual Word-in-Context Representations
por: Arase, Yuki, et al.
Publicado: (2024)
por: Arase, Yuki, et al.
Publicado: (2024)
Rapid Word Learning Through Meta In-Context Learning
por: Wang, Wentao, et al.
Publicado: (2025)
por: Wang, Wentao, et al.
Publicado: (2025)
EasyDistill: A Comprehensive Toolkit for Effective Knowledge Distillation of Large Language Models
por: Wang, Chengyu, et al.
Publicado: (2025)
por: Wang, Chengyu, et al.
Publicado: (2025)
Efficient Text Classification with Conformal In-Context Learning
por: Pantelidis, Ippokratis, et al.
Publicado: (2025)
por: Pantelidis, Ippokratis, et al.
Publicado: (2025)
EMSEdit: Efficient Multi-Step Meta-Learning-based Model Editing
por: Li, Xiaopeng, et al.
Publicado: (2025)
por: Li, Xiaopeng, et al.
Publicado: (2025)
Emojis Decoded: Leveraging ChatGPT for Enhanced Understanding in Social Media Communications
por: Zhou, Yuhang, et al.
Publicado: (2024)
por: Zhou, Yuhang, et al.
Publicado: (2024)
Context-level Language Modeling by Learning Predictive Context Embeddings
por: Dai, Beiya, et al.
Publicado: (2025)
por: Dai, Beiya, et al.
Publicado: (2025)
Distilling the Essence: Efficient Reasoning Distillation via Sequence Truncation
por: Chen, Wei-Rui, et al.
Publicado: (2025)
por: Chen, Wei-Rui, et al.
Publicado: (2025)
EntropyLong: Effective Long-Context Training via Predictive Uncertainty
por: Jia, Junlong, et al.
Publicado: (2025)
por: Jia, Junlong, et al.
Publicado: (2025)
WAND: Windowed Attention and Knowledge Distillation for Efficient Autoregressive Text-to-Speech Models
por: Lee, Hanna, et al.
Publicado: (2026)
por: Lee, Hanna, et al.
Publicado: (2026)
Timely Machine: Awareness of Time Makes Test-Time Scaling Agentic
por: Ma, Yichuan, et al.
Publicado: (2026)
por: Ma, Yichuan, et al.
Publicado: (2026)
Beyond Induction Heads: In-Context Meta Learning Induces Multi-Phase Circuit Emergence
por: Minegishi, Gouki, et al.
Publicado: (2025)
por: Minegishi, Gouki, et al.
Publicado: (2025)
OPSDL: On-Policy Self-Distillation for Long-Context Language Models
por: Zhang, Xinsen, et al.
Publicado: (2026)
por: Zhang, Xinsen, et al.
Publicado: (2026)
Whispering Context: Distilling Syntax and Semantics for Long Speech Transcripts
por: Altinok, Duygu
Publicado: (2025)
por: Altinok, Duygu
Publicado: (2025)
KV-Distill: Nearly Lossless Learnable Context Compression for LLMs
por: Chari, Vivek, et al.
Publicado: (2025)
por: Chari, Vivek, et al.
Publicado: (2025)
Effective Structured Prompting by Meta-Learning and Representative Verbalizer
por: Jiang, Weisen, et al.
Publicado: (2023)
por: Jiang, Weisen, et al.
Publicado: (2023)
Efficiently Distilling LLMs for Edge Applications
por: Kundu, Achintya, et al.
Publicado: (2024)
por: Kundu, Achintya, et al.
Publicado: (2024)
SkillFactory: Self-Distillation For Learning Cognitive Behaviors
por: Sprague, Zayne, et al.
Publicado: (2025)
por: Sprague, Zayne, et al.
Publicado: (2025)
Jailbreak and Guard Aligned Language Models with Only Few In-Context Demonstrations
por: Wei, Zeming, et al.
Publicado: (2023)
por: Wei, Zeming, et al.
Publicado: (2023)
Sentinel: Decoding Context Utilization via Attention Probing for Efficient LLM Context Compression
por: Zhang, Yong, et al.
Publicado: (2025)
por: Zhang, Yong, et al.
Publicado: (2025)
Efficient Technical Term Translation: A Knowledge Distillation Approach for Parenthetical Terminology Translation
por: Myung, Jiyoon, et al.
Publicado: (2024)
por: Myung, Jiyoon, et al.
Publicado: (2024)
Ejemplares similares
-
Empowering Large Language Models for Textual Data Augmentation
por: Li, Yichuan, et al.
Publicado: (2024) -
PropMEND: Hypernetworks for Knowledge Propagation in LLMs
por: Liu, Zeyu Leo, et al.
Publicado: (2025) -
From Critique to Clarity: A Pathway to Faithful and Personalized Code Explanations with Large Language Models
por: Xu, Zexing, et al.
Publicado: (2024) -
STEPER: Step-wise Knowledge Distillation for Enhancing Reasoning Ability in Multi-Step Retrieval-Augmented Language Models
por: Lee, Kyumin, et al.
Publicado: (2025) -
Reducing and Exploiting Data Augmentation Noise through Meta Reweighting Contrastive Learning for Text Classification
por: Mou, Guanyi, et al.
Publicado: (2024)