Class Distillation with Mahalanobis Contrast: An Efficient Training Paradigm for Pragmatic Language Understanding Tasks
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Chenlu, Lyu, Weimin, Banerjee, Ritwik |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Paying Attention to Deflections: Mining Pragmatic Nuances for Whataboutism Detection in Online Discourse
por: Phi, Khiem, et al.
Publicado: (2024)
por: Phi, Khiem, et al.
Publicado: (2024)
A Longitudinal, Multinational, and Multilingual Corpus of News Coverage of the Russo-Ukrainian War
por: Mohanty, Dikshya, et al.
Publicado: (2026)
por: Mohanty, Dikshya, et al.
Publicado: (2026)
PUB: A Pragmatics Understanding Benchmark for Assessing LLMs' Pragmatics Capabilities
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2024)
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2024)
Beyond Understanding: Evaluating the Pragmatic Gap in LLMs' Cultural Processing of Figurative Language
por: Attia, Mena, et al.
Publicado: (2025)
por: Attia, Mena, et al.
Publicado: (2025)
Understand the Implication: Learning to Think for Pragmatic Understanding
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2025)
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2025)
Understanding Understanding: A Pragmatic Framework Motivated by Large Language Models
por: Leyton-Brown, Kevin, et al.
Publicado: (2024)
por: Leyton-Brown, Kevin, et al.
Publicado: (2024)
Are Multimodal Large Language Models Pragmatically Competent Listeners in Simple Reference Resolution Tasks?
por: Junker, Simeon, et al.
Publicado: (2025)
por: Junker, Simeon, et al.
Publicado: (2025)
Towards Better Understanding of Contrastive Sentence Representation Learning: A Unified Paradigm for Gradient
por: Li, Mingxin, et al.
Publicado: (2024)
por: Li, Mingxin, et al.
Publicado: (2024)
Training Task Experts through Retrieval Based Distillation
por: Ge, Jiaxin, et al.
Publicado: (2024)
por: Ge, Jiaxin, et al.
Publicado: (2024)
Do Language Models Understand the Cognitive Tasks Given to Them? Investigations with the N-Back Paradigm
por: Hu, Xiaoyang, et al.
Publicado: (2024)
por: Hu, Xiaoyang, et al.
Publicado: (2024)
Pragmatic Theories Enhance Understanding of Implied Meanings in LLMs
por: Sato, Takuma, et al.
Publicado: (2025)
por: Sato, Takuma, et al.
Publicado: (2025)
The Pragmatic Mind of Machines: Tracing the Emergence of Pragmatic Competence in Large Language Models
por: Yu, Kefan, et al.
Publicado: (2025)
por: Yu, Kefan, et al.
Publicado: (2025)
BadCLM: Backdoor Attack in Clinical Language Models for Electronic Health Records
por: Lyu, Weimin, et al.
Publicado: (2024)
por: Lyu, Weimin, et al.
Publicado: (2024)
Developing a Pragmatic Benchmark for Assessing Korean Legal Language Understanding in Large Language Models
por: Kim, Yeeun, et al.
Publicado: (2024)
por: Kim, Yeeun, et al.
Publicado: (2024)
Distillation versus Contrastive Learning: How to Train Your Rerankers
por: Xu, Zhichao, et al.
Publicado: (2025)
por: Xu, Zhichao, et al.
Publicado: (2025)
NLoRA: Nyström-Initiated Low-Rank Adaptation for Large Language Models
por: Guo, Chenlu, et al.
Publicado: (2025)
por: Guo, Chenlu, et al.
Publicado: (2025)
VLCD: Vision-Language Contrastive Distillation for Accurate and Efficient Automatic Placenta Analysis
por: Mehta, Manas, et al.
Publicado: (2025)
por: Mehta, Manas, et al.
Publicado: (2025)
DistilQwen2.5: Industrial Practices of Training Distilled Open Lightweight Language Models
por: Wang, Chengyu, et al.
Publicado: (2025)
por: Wang, Chengyu, et al.
Publicado: (2025)
A Training-Free Regeneration Paradigm: Contrastive Reflection Memory Guided Self-Verification and Self-Improvement
por: Li, Yuran, et al.
Publicado: (2026)
por: Li, Yuran, et al.
Publicado: (2026)
Leveraging Large Language Models for Enhanced NLP Task Performance through Knowledge Distillation and Optimized Training Strategies
por: Huang, Yining, et al.
Publicado: (2024)
por: Huang, Yining, et al.
Publicado: (2024)
Sensivity of LLMs' Explanations to the Training Randomness:Context, Class & Task Dependencies
por: Loncour, Romain, et al.
Publicado: (2026)
por: Loncour, Romain, et al.
Publicado: (2026)
Contrastive Learning in Distilled Models
por: Lim, Valerie, et al.
Publicado: (2024)
por: Lim, Valerie, et al.
Publicado: (2024)
Pragmatic Competence Evaluation of Large Language Models for the Korean Language
por: Park, Dojun, et al.
Publicado: (2024)
por: Park, Dojun, et al.
Publicado: (2024)
QCRD: Quality-guided Contrastive Rationale Distillation for Large Language Models
por: Wang, Wei, et al.
Publicado: (2024)
por: Wang, Wei, et al.
Publicado: (2024)
Distilling Instruction-following Abilities of Large Language Models with Task-aware Curriculum Planning
por: Yue, Yuanhao, et al.
Publicado: (2024)
por: Yue, Yuanhao, et al.
Publicado: (2024)
A Dual-Task Paradigm to Investigate Sentence Comprehension Strategies in Language Models
por: Emura, Rei, et al.
Publicado: (2026)
por: Emura, Rei, et al.
Publicado: (2026)
Efficient End-to-End Visual Document Understanding with Rationale Distillation
por: Zhu, Wang, et al.
Publicado: (2023)
por: Zhu, Wang, et al.
Publicado: (2023)
Language Model Training Paradigms for Clinical Feature Embeddings
por: Hu, Yurong, et al.
Publicado: (2023)
por: Hu, Yurong, et al.
Publicado: (2023)
Distilling Fine-grained Sentiment Understanding from Large Language Models
por: Zhang, Yice, et al.
Publicado: (2024)
por: Zhang, Yice, et al.
Publicado: (2024)
A Learning Rate Path Switching Training Paradigm for Version Updates of Large Language Models
por: Wang, Zhihao, et al.
Publicado: (2024)
por: Wang, Zhihao, et al.
Publicado: (2024)
Unsupervised Out-of-Distribution Dialect Detection with Mahalanobis Distance
por: Das, Sourya Dipta, et al.
Publicado: (2023)
por: Das, Sourya Dipta, et al.
Publicado: (2023)
CHBench: A Chinese Dataset for Evaluating Health in Large Language Models
por: Guo, Chenlu, et al.
Publicado: (2024)
por: Guo, Chenlu, et al.
Publicado: (2024)
Evolutionary Contrastive Distillation for Language Model Alignment
por: Katz-Samuels, Julian, et al.
Publicado: (2024)
por: Katz-Samuels, Julian, et al.
Publicado: (2024)
Long-context Non-factoid Question Answering in Indic Languages
por: Mishra, Ritwik, et al.
Publicado: (2025)
por: Mishra, Ritwik, et al.
Publicado: (2025)
Task-Agnostic Detector for Insertion-Based Backdoor Attacks
por: Lyu, Weimin, et al.
Publicado: (2024)
por: Lyu, Weimin, et al.
Publicado: (2024)
AFD-SLU: Adaptive Feature Distillation for Spoken Language Understanding
por: Xie, Yan, et al.
Publicado: (2025)
por: Xie, Yan, et al.
Publicado: (2025)
Language Models are Bounded Pragmatic Speakers: Understanding RLHF from a Bayesian Cognitive Modeling Perspective
por: Nguyen, Khanh
Publicado: (2023)
por: Nguyen, Khanh
Publicado: (2023)
From Correction to Mastery: Reinforced Distillation of Large Language Model Agents
por: Lyu, Yuanjie, et al.
Publicado: (2025)
por: Lyu, Yuanjie, et al.
Publicado: (2025)
LittiChoQA: Literary Texts in Indic Languages Chosen for Question Answering
por: Khandelwal, Aarya, et al.
Publicado: (2026)
por: Khandelwal, Aarya, et al.
Publicado: (2026)
DDK: Distilling Domain Knowledge for Efficient Large Language Models
por: Liu, Jiaheng, et al.
Publicado: (2024)
por: Liu, Jiaheng, et al.
Publicado: (2024)
Ejemplares similares
-
Paying Attention to Deflections: Mining Pragmatic Nuances for Whataboutism Detection in Online Discourse
por: Phi, Khiem, et al.
Publicado: (2024) -
A Longitudinal, Multinational, and Multilingual Corpus of News Coverage of the Russo-Ukrainian War
por: Mohanty, Dikshya, et al.
Publicado: (2026) -
PUB: A Pragmatics Understanding Benchmark for Assessing LLMs' Pragmatics Capabilities
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2024) -
Beyond Understanding: Evaluating the Pragmatic Gap in LLMs' Cultural Processing of Figurative Language
por: Attia, Mena, et al.
Publicado: (2025) -
Understand the Implication: Learning to Think for Pragmatic Understanding
por: Sravanthi, Settaluri Lakshmi, et al.
Publicado: (2025)