A Multi-Encoder Frozen-Decoder Approach for Fine-Tuning Large Language Models
Fuente:
arXiv
Saved in:
| Main Author: | Dhole, Kaustubh D. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020)
by: Banthia, Saumya, et al.
Published: (2020)
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
by: Sharma, Anika, et al.
Published: (2025)
by: Sharma, Anika, et al.
Published: (2025)
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
by: Gonzalez, Alberto Andres Valdes
Published: (2026)
RubricRAG: Towards Interpretable and Reliable LLM Evaluation via Domain Knowledge Retrieval for Rubric Generation
by: Dhole, Kaustubh D., et al.
Published: (2026)
by: Dhole, Kaustubh D., et al.
Published: (2026)
Applied Explainability for Large Language Models: A Comparative Study
by: Kancharla, Venkata Abhinandan
Published: (2026)
by: Kancharla, Venkata Abhinandan
Published: (2026)
Adaptive Activation Cancellation for Hallucination Mitigation in Large Language Models
by: Yocam, Eric, et al.
Published: (2026)
by: Yocam, Eric, et al.
Published: (2026)
A Lightweight Approach to Detection of AI-Generated Texts Using Stylometric Features
by: Aityan, Sergey K., et al.
Published: (2025)
by: Aityan, Sergey K., et al.
Published: (2025)
A Semantic Approach to Negation Detection and Word Disambiguation with Natural Language Processing
by: Okpala, Izunna, et al.
Published: (2023)
by: Okpala, Izunna, et al.
Published: (2023)
Word Importance Explains How Prompts Affect Language Model Outputs
by: Hackmann, Stefan, et al.
Published: (2024)
by: Hackmann, Stefan, et al.
Published: (2024)
Do LLMs Know What They Know? Measuring Metacognitive Efficiency with Signal Detection Theory
by: Cacioli, Jon-Paul
Published: (2026)
by: Cacioli, Jon-Paul
Published: (2026)
Computational Economics in Large Language Models: Exploring Model Behavior and Incentive Design under Resource Constraints
by: Reddy, Sandeep, et al.
Published: (2025)
by: Reddy, Sandeep, et al.
Published: (2025)
Efficient Strategy for Improving Large Language Model (LLM) Capabilities
by: Gutiérrez, Julián Camilo Velandia
Published: (2025)
by: Gutiérrez, Julián Camilo Velandia
Published: (2025)
When is dataset cartography ineffective? Using training dynamics does not improve robustness against Adversarial SQuAD
by: Mandal, Paul K.
Published: (2025)
by: Mandal, Paul K.
Published: (2025)
Large Language Models Are Not Strong Abstract Reasoners
by: Gendron, Gaël, et al.
Published: (2023)
by: Gendron, Gaël, et al.
Published: (2023)
To Retrieve or Not to Retrieve? Uncertainty Detection for Dynamic Retrieval Augmented Generation
by: Dhole, Kaustubh D.
Published: (2025)
by: Dhole, Kaustubh D.
Published: (2025)
Correspondence of high-dimensional emotion structures elicited by video clips between humans and Multimodal LLMs
by: Asanuma, Haruka, et al.
Published: (2025)
by: Asanuma, Haruka, et al.
Published: (2025)
Less is More: Learning Graph Tasks with Just LLMs
by: Shirai, Sola, et al.
Published: (2025)
by: Shirai, Sola, et al.
Published: (2025)
Accelerating Language Model Workflows with Prompt Choreography
by: Bai, TJ, et al.
Published: (2025)
by: Bai, TJ, et al.
Published: (2025)
Project Riley: Multimodal Multi-Agent LLM Collaboration with Emotional Reasoning and Voting
by: Ortigoso, Ana Rita, et al.
Published: (2025)
by: Ortigoso, Ana Rita, et al.
Published: (2025)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
by: Wen, Yuqiao, et al.
Published: (2025)
by: Wen, Yuqiao, et al.
Published: (2025)
MIMIC-SR-ICD11: A Dataset for Narrative-Based Diagnosis
by: Wu, Yuexin, et al.
Published: (2025)
by: Wu, Yuexin, et al.
Published: (2025)
DROID: Dual Representation for Out-of-Scope Intent Detection
by: Rashwan, Wael, et al.
Published: (2025)
by: Rashwan, Wael, et al.
Published: (2025)
The Rise of Verbal Tics in Large Language Models: A Systematic Analysis Across Frontier Models
by: Wu, Shuai, et al.
Published: (2026)
by: Wu, Shuai, et al.
Published: (2026)
ADALog: Adaptive Unsupervised Anomaly detection in Logs with Self-attention Masked Language Model
by: Pospieszny, Przemek, et al.
Published: (2025)
by: Pospieszny, Przemek, et al.
Published: (2025)
Listwise Direct Preference Optimization with Multi-Dimensional Preference Mixing
by: Sun, Yuhui, et al.
Published: (2025)
by: Sun, Yuhui, et al.
Published: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
by: Wen, Yuqiao, et al.
Published: (2024)
by: Wen, Yuqiao, et al.
Published: (2024)
Regularisation in neural networks: a survey and empirical analysis of approaches
by: Opperman, Christiaan P., et al.
Published: (2026)
by: Opperman, Christiaan P., et al.
Published: (2026)
Reading the Mood Behind Words: Integrating Prosody-Derived Emotional Context into Socially Responsive VR Agents
by: Jeong, SangYeop, et al.
Published: (2026)
by: Jeong, SangYeop, et al.
Published: (2026)
LLM Performance Predictors: Learning When to Escalate in Hybrid Human-AI Moderation Systems
by: Bachar, Or, et al.
Published: (2026)
by: Bachar, Or, et al.
Published: (2026)
Balancing the Scales: A Comprehensive Study on Tackling Class Imbalance in Binary Classification
by: Abdelhamid, Mohamed, et al.
Published: (2024)
by: Abdelhamid, Mohamed, et al.
Published: (2024)
Adaptive Riemannian Graph Neural Networks
by: Wang, Xudong, et al.
Published: (2025)
by: Wang, Xudong, et al.
Published: (2025)
Learnable Kernel Density Estimation for Graphs and Its Application to Graph-Level Anomaly Detection
by: Wang, Xudong, et al.
Published: (2025)
by: Wang, Xudong, et al.
Published: (2025)
Predicting When to Trust Vision-Language Models for Spatial Reasoning
by: Imran, Muhammad, et al.
Published: (2026)
by: Imran, Muhammad, et al.
Published: (2026)
TaylorShift: Shifting the Complexity of Self-Attention from Squared to Linear (and Back) using Taylor-Softmax
by: Nauen, Tobias Christian, et al.
Published: (2024)
by: Nauen, Tobias Christian, et al.
Published: (2024)
Efficient Hate Speech Detection: Evaluating 38 Models from Traditional Methods to Transformers
by: Abusaqer, Mahmoud, et al.
Published: (2025)
by: Abusaqer, Mahmoud, et al.
Published: (2025)
Deep Feed-Forward Neural Network for Bangla Isolated Speech Recognition
by: Bhadra, Dipayan, et al.
Published: (2025)
by: Bhadra, Dipayan, et al.
Published: (2025)
Measurement Risk in Supervised Financial NLP: Rubric and Metric Sensitivity on JF-ICR
by: Chang, Sidi, et al.
Published: (2026)
by: Chang, Sidi, et al.
Published: (2026)
Logits-Constrained Framework with RoBERTa for Ancient Chinese NER
by: Hua, Wenjie, et al.
Published: (2025)
by: Hua, Wenjie, et al.
Published: (2025)
Phase-Aware Wavelet-Based-Scattering Encoder-Decoder for Dense Predictions
by: Marrakchi, Ghassen, et al.
Published: (2026)
by: Marrakchi, Ghassen, et al.
Published: (2026)
Beyond Subtokens: A Rich Character Embedding for Low-resource and Morphologically Complex Languages
by: Schneider, Felix, et al.
Published: (2026)
by: Schneider, Felix, et al.
Published: (2026)
Similar Items
-
Classification of descriptions and summary using multiple passes of statistical and natural language toolkits
by: Banthia, Saumya, et al.
Published: (2020) -
Adversarially Probing Cross-Family Sound Symbolism in 27 Languages
by: Sharma, Anika, et al.
Published: (2025) -
Cost-Aware Model Selection for Text Classification: Multi-Objective Trade-offs Between Fine-Tuned Encoders and LLM Prompting in Production
by: Gonzalez, Alberto Andres Valdes
Published: (2026) -
RubricRAG: Towards Interpretable and Reliable LLM Evaluation via Domain Knowledge Retrieval for Rubric Generation
by: Dhole, Kaustubh D., et al.
Published: (2026) -
Applied Explainability for Large Language Models: A Comparative Study
by: Kancharla, Venkata Abhinandan
Published: (2026)