Towards Objective Fine-tuning: How LLMs' Prior Knowledge Causes Potential Poor Calibration?
Fuente:
arXiv
Guardado en:
| Autores principales: | Wang, Ziming, Shi, Zeyu, Zhou, Haoyi, Gao, Shiqi, Sun, Qingyun, Li, Jianxin |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Fine-Tuned LLMs Know They Don't Know: A Parameter-Efficient Approach to Recovering Honesty
por: Shi, Zeyu, et al.
Publicado: (2025)
por: Shi, Zeyu, et al.
Publicado: (2025)
MIKO: Multimodal Intention Knowledge Distillation from Large Language Models for Social-Media Commonsense Discovery
por: Lu, Feihong, et al.
Publicado: (2024)
por: Lu, Feihong, et al.
Publicado: (2024)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
por: Zhou, Huichi, et al.
Publicado: (2025)
por: Zhou, Huichi, et al.
Publicado: (2025)
Stage-wise Fine-tuning for Graph-to-Text Generation
por: Wang, Qingyun, et al.
Publicado: (2021)
por: Wang, Qingyun, et al.
Publicado: (2021)
Beyond Fine-tuning: Unleashing the Potential of Continuous Pretraining for Clinical LLMs
por: Christophe, Clément, et al.
Publicado: (2024)
por: Christophe, Clément, et al.
Publicado: (2024)
Double-Calibration: Towards Reliable LLMs via Calibrating Knowledge and Reasoning Confidence
por: Lu, Yuyin, et al.
Publicado: (2026)
por: Lu, Yuyin, et al.
Publicado: (2026)
Be Careful When Fine-tuning On Open-Source LLMs: Your Fine-tuning Data Could Be Secretly Stolen!
por: Zhang, Zhexin, et al.
Publicado: (2025)
por: Zhang, Zhexin, et al.
Publicado: (2025)
Explain Less, Understand More: Jargon Detection via Personalized Parameter-Efficient Fine-tuning
por: Wu, Bohao, et al.
Publicado: (2025)
por: Wu, Bohao, et al.
Publicado: (2025)
Beyond Human Preferences: Exploring Reinforcement Learning Trajectory Evaluation and Improvement through LLMs
por: Shen, Zichao, et al.
Publicado: (2024)
por: Shen, Zichao, et al.
Publicado: (2024)
Unlocking the Potentials of Retrieval-Augmented Generation for Diffusion Language Models
por: Yu, Chuanyue, et al.
Publicado: (2026)
por: Yu, Chuanyue, et al.
Publicado: (2026)
Towards Context-Robust LLMs: A Gated Representation Fine-tuning Approach
por: Zeng, Shenglai, et al.
Publicado: (2025)
por: Zeng, Shenglai, et al.
Publicado: (2025)
Disentangling Reasoning Tokens and Boilerplate Tokens For Language Model Fine-tuning
por: Ye, Ziang, et al.
Publicado: (2024)
por: Ye, Ziang, et al.
Publicado: (2024)
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models
por: Lyu, Yougang, et al.
Publicado: (2024)
por: Lyu, Yougang, et al.
Publicado: (2024)
Practical Membership Inference Attacks against Fine-tuned Large Language Models via Self-prompt Calibration
por: Fu, Wenjie, et al.
Publicado: (2023)
por: Fu, Wenjie, et al.
Publicado: (2023)
KaFT: Knowledge-aware Fine-tuning for Boosting LLMs' Domain-specific Question-Answering Performance
por: Zhong, Qihuang, et al.
Publicado: (2025)
por: Zhong, Qihuang, et al.
Publicado: (2025)
GSQ-Tuning: Group-Shared Exponents Integer in Fully Quantized Training for LLMs On-Device Fine-tuning
por: Zhou, Sifan, et al.
Publicado: (2025)
por: Zhou, Sifan, et al.
Publicado: (2025)
UltraLink: An Open-Source Knowledge-Enhanced Multilingual Supervised Fine-tuning Dataset
por: Wang, Haoyu, et al.
Publicado: (2024)
por: Wang, Haoyu, et al.
Publicado: (2024)
From Text to Trajectory: Exploring Complex Constraint Representation and Decomposition in Safe Reinforcement Learning
por: Dong, Pusen, et al.
Publicado: (2024)
por: Dong, Pusen, et al.
Publicado: (2024)
GRAVER: Generative Graph Vocabularies for Robust Graph Foundation Models Fine-tuning
por: Yuan, Haonan, et al.
Publicado: (2025)
por: Yuan, Haonan, et al.
Publicado: (2025)
Intent Mismatch Causes LLMs to Get Lost in Multi-Turn Conversation
por: Liu, Geng, et al.
Publicado: (2026)
por: Liu, Geng, et al.
Publicado: (2026)
KcMF: A Knowledge-compliant Framework for Schema and Entity Matching with Fine-tuning-free LLMs
por: Xu, Yongqin, et al.
Publicado: (2024)
por: Xu, Yongqin, et al.
Publicado: (2024)
Identity Lock: Locking API Fine-tuned LLMs With Identity-based Wake Words
por: Su, Hongyu, et al.
Publicado: (2025)
por: Su, Hongyu, et al.
Publicado: (2025)
How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition
por: Dong, Guanting, et al.
Publicado: (2023)
por: Dong, Guanting, et al.
Publicado: (2023)
Topic Modeling with Fine-tuning LLMs and Bag of Sentences
por: Schneider, Johannes
Publicado: (2024)
por: Schneider, Johannes
Publicado: (2024)
Fine-tuning and Utilization Methods of Domain-specific LLMs
por: Jeong, Cheonsu
Publicado: (2024)
por: Jeong, Cheonsu
Publicado: (2024)
medIKAL: Integrating Knowledge Graphs as Assistants of LLMs for Enhanced Clinical Diagnosis on EMRs
por: Jia, Mingyi, et al.
Publicado: (2024)
por: Jia, Mingyi, et al.
Publicado: (2024)
Precise Localization of Memories: A Fine-grained Neuron-level Knowledge Editing Technique for LLMs
por: Pan, Haowen, et al.
Publicado: (2025)
por: Pan, Haowen, et al.
Publicado: (2025)
ITERTL: An Iterative Framework for Fine-tuning LLMs for RTL Code Generation
por: Wu, Peiyang, et al.
Publicado: (2024)
por: Wu, Peiyang, et al.
Publicado: (2024)
RepCali: High Efficient Fine-tuning Via Representation Calibration in Latent Space for Pre-trained Language Models
por: Zhang, Fujun, et al.
Publicado: (2025)
por: Zhang, Fujun, et al.
Publicado: (2025)
Agent Fine-tuning through Distillation for Domain-specific LLMs in Microdomains
por: Xue, Yawen, et al.
Publicado: (2025)
por: Xue, Yawen, et al.
Publicado: (2025)
RoseLoRA: Row and Column-wise Sparse Low-rank Adaptation of Pre-trained Language Model for Knowledge Editing and Fine-tuning
por: Wang, Haoyu, et al.
Publicado: (2024)
por: Wang, Haoyu, et al.
Publicado: (2024)
Structure-aware Fine-tuning for Code Pre-trained Models
por: Wu, Jiayi, et al.
Publicado: (2024)
por: Wu, Jiayi, et al.
Publicado: (2024)
Aloe: A Family of Fine-tuned Open Healthcare LLMs
por: Gururajan, Ashwin Kumar, et al.
Publicado: (2024)
por: Gururajan, Ashwin Kumar, et al.
Publicado: (2024)
FineTuneBench: How well do commercial fine-tuning APIs infuse knowledge into LLMs?
por: Wu, Eric, et al.
Publicado: (2024)
por: Wu, Eric, et al.
Publicado: (2024)
Relational Knowledge Distillation Using Fine-tuned Function Vectors
por: Kang, Andrea, et al.
Publicado: (2026)
por: Kang, Andrea, et al.
Publicado: (2026)
UniARM: Towards a Unified Autoregressive Reward Model for Multi-Objective Test-Time Alignment
por: Xie, Hongyan, et al.
Publicado: (2026)
por: Xie, Hongyan, et al.
Publicado: (2026)
Fine-tuning Done Right in Model Editing
por: Yang, Wanli, et al.
Publicado: (2025)
por: Yang, Wanli, et al.
Publicado: (2025)
AgenticGEO: A Self-Evolving Agentic System for Generative Engine Optimization
por: Yuan, Jiaqi, et al.
Publicado: (2026)
por: Yuan, Jiaqi, et al.
Publicado: (2026)
Instructing the Architecture Search for Spatial-temporal Sequence Forecasting with LLM
por: Xue, Xin, et al.
Publicado: (2025)
por: Xue, Xin, et al.
Publicado: (2025)
Automatic Essay Multi-dimensional Scoring with Fine-tuning and Multiple Regression
por: Sun, Kun, et al.
Publicado: (2024)
por: Sun, Kun, et al.
Publicado: (2024)
Ejemplares similares
-
Fine-Tuned LLMs Know They Don't Know: A Parameter-Efficient Approach to Recovering Honesty
por: Shi, Zeyu, et al.
Publicado: (2025) -
MIKO: Multimodal Intention Knowledge Distillation from Large Language Models for Social-Media Commonsense Discovery
por: Lu, Feihong, et al.
Publicado: (2024) -
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
por: Zhou, Huichi, et al.
Publicado: (2025) -
Stage-wise Fine-tuning for Graph-to-Text Generation
por: Wang, Qingyun, et al.
Publicado: (2021) -
Beyond Fine-tuning: Unleashing the Potential of Continuous Pretraining for Clinical LLMs
por: Christophe, Clément, et al.
Publicado: (2024)