Alignment Adapter to Improve the Performance of Compressed Deep Learning Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Rai, Rohit Raj, Dhaka, Abhishek, Awekar, Amit |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Compressed models are NOT miniature versions of large models
por: Rai, Rohit Raj, et al.
Publicado: (2024)
por: Rai, Rohit Raj, et al.
Publicado: (2024)
Compressed Models are NOT Trust-equivalent to Their Large Counterparts
por: Rai, Rohit Raj, et al.
Publicado: (2025)
por: Rai, Rohit Raj, et al.
Publicado: (2025)
Application Specific Compression of Deep Learning Models
por: Rai, Rohit Raj, et al.
Publicado: (2024)
por: Rai, Rohit Raj, et al.
Publicado: (2024)
Are Word Embedding Methods Stable and Should We Care About It?
por: Borah, Angana, et al.
Publicado: (2021)
por: Borah, Angana, et al.
Publicado: (2021)
Evaluating Chain-of-Thought Reasoning through Reusability and Verifiability
por: Aggarwal, Shashank, et al.
Publicado: (2026)
por: Aggarwal, Shashank, et al.
Publicado: (2026)
Effect of dimensionality change on the bias of word embeddings
por: Rai, Rohit Raj, et al.
Publicado: (2023)
por: Rai, Rohit Raj, et al.
Publicado: (2023)
Revisit and Outstrip Entity Alignment: A Perspective of Generative Models
por: Guo, Lingbing, et al.
Publicado: (2023)
por: Guo, Lingbing, et al.
Publicado: (2023)
ELMO: Efficiency via Low-precision and Peak Memory Optimization in Large Output Spaces
por: Zhang, Jinbin, et al.
Publicado: (2025)
por: Zhang, Jinbin, et al.
Publicado: (2025)
AT-RAG: An Adaptive RAG Model Enhancing Query Efficiency with Topic Filtering and Iterative Reasoning
por: Rezaei, Mohammad Reza, et al.
Publicado: (2024)
por: Rezaei, Mohammad Reza, et al.
Publicado: (2024)
Improving Neural Topic Models with Wasserstein Knowledge Distillation
por: Adhya, Suman, et al.
Publicado: (2023)
por: Adhya, Suman, et al.
Publicado: (2023)
LANE: Logic Alignment of Non-tuning Large Language Models and Online Recommendation Systems for Explainable Reason Generation
por: Zhao, Hongke, et al.
Publicado: (2024)
por: Zhao, Hongke, et al.
Publicado: (2024)
Scaling the Vocabulary of Non-autoregressive Models for Efficient Generative Retrieval
por: Valluri, Ravisri, et al.
Publicado: (2024)
por: Valluri, Ravisri, et al.
Publicado: (2024)
Refining Dimensions for Improving Clustering-based Cross-lingual Topic Models
por: Chang, Chia-Hsuan, et al.
Publicado: (2024)
por: Chang, Chia-Hsuan, et al.
Publicado: (2024)
Improving Legal Entity Recognition Using a Hybrid Transformer Model and Semantic Filtering Approach
por: Rajamanickam, Duraimurugan
Publicado: (2024)
por: Rajamanickam, Duraimurugan
Publicado: (2024)
Personas within Parameters: Fine-Tuning Small Language Models with Low-Rank Adapters to Mimic User Behaviors
por: Thakur, Himanshu, et al.
Publicado: (2025)
por: Thakur, Himanshu, et al.
Publicado: (2025)
Fine-Tuning Large Language Models and Evaluating Retrieval Methods for Improved Question Answering on Building Codes
por: Aqib, Mohammad, et al.
Publicado: (2025)
por: Aqib, Mohammad, et al.
Publicado: (2025)
Holistic Utility Preference Learning for Listwise Alignment
por: Zhou, Jiacong, et al.
Publicado: (2024)
por: Zhou, Jiacong, et al.
Publicado: (2024)
Federated Learning for ICD Classification with Lightweight Models and Pretrained Embeddings
por: Xu, Binbin, et al.
Publicado: (2025)
por: Xu, Binbin, et al.
Publicado: (2025)
Beyond Sequential Reranking: Reranker-Guided Search Improves Reasoning Intensive Retrieval
por: Xu, Haike, et al.
Publicado: (2025)
por: Xu, Haike, et al.
Publicado: (2025)
Open Deep Search: Democratizing Search with Open-source Reasoning Agents
por: Alzubi, Salaheddin, et al.
Publicado: (2025)
por: Alzubi, Salaheddin, et al.
Publicado: (2025)
A Language-Driven Framework for Improving Personalized Recommendations: Merging LLMs with Traditional Algorithms
por: Goldstein, Aaron, et al.
Publicado: (2025)
por: Goldstein, Aaron, et al.
Publicado: (2025)
IDGenRec: LLM-RecSys Alignment with Textual ID Learning
por: Tan, Juntao, et al.
Publicado: (2024)
por: Tan, Juntao, et al.
Publicado: (2024)
From Web Search towards Agentic Deep Research: Incentivizing Search with Reasoning Agents
por: Zhang, Weizhi, et al.
Publicado: (2025)
por: Zhang, Weizhi, et al.
Publicado: (2025)
Improving Word Translation via Two-Stage Contrastive Learning
por: Li, Yaoyiran, et al.
Publicado: (2022)
por: Li, Yaoyiran, et al.
Publicado: (2022)
Retrieval-Enhanced Machine Learning: Synthesis and Opportunities
por: Kim, To Eun, et al.
Publicado: (2024)
por: Kim, To Eun, et al.
Publicado: (2024)
SoftQE: Learned Representations of Queries Expanded by LLMs
por: Pimpalkhute, Varad, et al.
Publicado: (2024)
por: Pimpalkhute, Varad, et al.
Publicado: (2024)
Personalized Product Search Ranking: A Multi-Task Learning Approach with Tabular and Non-Tabular Data
por: Morishetti, Lalitesh, et al.
Publicado: (2025)
por: Morishetti, Lalitesh, et al.
Publicado: (2025)
Language Models As Semantic Indexers
por: Jin, Bowen, et al.
Publicado: (2023)
por: Jin, Bowen, et al.
Publicado: (2023)
SetCSE: Set Operations using Contrastive Learning of Sentence Embeddings
por: Liu, Kang
Publicado: (2024)
por: Liu, Kang
Publicado: (2024)
Exploring Contrastive Learning for Long-Tailed Multi-Label Text Classification
por: Audibert, Alexandre, et al.
Publicado: (2024)
por: Audibert, Alexandre, et al.
Publicado: (2024)
Medical Coding with Biomedical Transformer Ensembles and Zero/Few-shot Learning
por: Ziletti, Angelo, et al.
Publicado: (2022)
por: Ziletti, Angelo, et al.
Publicado: (2022)
mmBERT: A Modern Multilingual Encoder with Annealed Language Learning
por: Marone, Marc, et al.
Publicado: (2025)
por: Marone, Marc, et al.
Publicado: (2025)
LTR-ICD: A Learning-to-Rank Approach for Automatic ICD Coding
por: Mansoori, Mohammad, et al.
Publicado: (2025)
por: Mansoori, Mohammad, et al.
Publicado: (2025)
Familiarity-Aware Evidence Compression for Retrieval-Augmented Generation
por: Jung, Dongwon, et al.
Publicado: (2024)
por: Jung, Dongwon, et al.
Publicado: (2024)
Explainable Disentangled Representation Learning for Generalizable Authorship Attribution in the Era of Generative AI
por: Man, Hieu, et al.
Publicado: (2026)
por: Man, Hieu, et al.
Publicado: (2026)
Cluster-based Adaptive Retrieval: Dynamic Context Selection for RAG Applications
por: Xu, Yifan, et al.
Publicado: (2025)
por: Xu, Yifan, et al.
Publicado: (2025)
PaECTER: Patent-level Representation Learning using Citation-informed Transformers
por: Ghosh, Mainak, et al.
Publicado: (2024)
por: Ghosh, Mainak, et al.
Publicado: (2024)
The Atlas of In-Context Learning: How Attention Heads Shape In-Context Retrieval Augmentation
por: Kahardipraja, Patrick, et al.
Publicado: (2025)
por: Kahardipraja, Patrick, et al.
Publicado: (2025)
PEARL: Prototype-Enhanced Alignment for Label-Efficient Representation Learning with Deployment-Driven Insights from Digital Governance Communication Systems
por: Zhang, Ruiyu, et al.
Publicado: (2026)
por: Zhang, Ruiyu, et al.
Publicado: (2026)
An Integrated Data Processing Framework for Pretraining Foundation Models
por: Sun, Yiding, et al.
Publicado: (2024)
por: Sun, Yiding, et al.
Publicado: (2024)
Ejemplares similares
-
Compressed models are NOT miniature versions of large models
por: Rai, Rohit Raj, et al.
Publicado: (2024) -
Compressed Models are NOT Trust-equivalent to Their Large Counterparts
por: Rai, Rohit Raj, et al.
Publicado: (2025) -
Application Specific Compression of Deep Learning Models
por: Rai, Rohit Raj, et al.
Publicado: (2024) -
Are Word Embedding Methods Stable and Should We Care About It?
por: Borah, Angana, et al.
Publicado: (2021) -
Evaluating Chain-of-Thought Reasoning through Reusability and Verifiability
por: Aggarwal, Shashank, et al.
Publicado: (2026)