RobustFT: Robust Supervised Fine-tuning for Large Language Models under Noisy Response
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Luo, Junyu, Luo, Xiao, Ding, Kaize, Yuan, Jingyang, Xiao, Zhiping, Zhang, Ming |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Semi-supervised Fine-tuning for Large Language Models
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
Attention Bootstrapping for Multi-Modal Test-Time Adaptation
von: Zhao, Yusheng, et al.
Veröffentlicht: (2025)
von: Zhao, Yusheng, et al.
Veröffentlicht: (2025)
VRPO: Rethinking Value Modeling for Robust RL Training under Noisy Supervision
von: Zhu, Dingwei, et al.
Veröffentlicht: (2025)
von: Zhu, Dingwei, et al.
Veröffentlicht: (2025)
Multi-Party Supervised Fine-tuning of Language Models for Multi-Party Dialogue Generation
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024)
GALA: Graph Diffusion-based Alignment with Jigsaw for Source-free Domain Adaptation
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
How Abilities in Large Language Models are Affected by Supervised Fine-tuning Data Composition
von: Dong, Guanting, et al.
Veröffentlicht: (2023)
von: Dong, Guanting, et al.
Veröffentlicht: (2023)
Panacea: Mitigating Harmful Fine-tuning for Large Language Models via Post-fine-tuning Perturbation
von: Wang, Yibo, et al.
Veröffentlicht: (2025)
von: Wang, Yibo, et al.
Veröffentlicht: (2025)
Empowering Large Language Models for Textual Data Augmentation
von: Li, Yichuan, et al.
Veröffentlicht: (2024)
von: Li, Yichuan, et al.
Veröffentlicht: (2024)
Rank and Align: Towards Effective Source-free Graph Domain Adaptation
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
von: Luo, Junyu, et al.
Veröffentlicht: (2024)
On Active Privacy Auditing in Supervised Fine-tuning for White-Box Language Models
von: Sun, Qian, et al.
Veröffentlicht: (2024)
von: Sun, Qian, et al.
Veröffentlicht: (2024)
Hallucination Detection: Robustly Discerning Reliable Answers in Large Language Models
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
von: Chen, Yuyan, et al.
Veröffentlicht: (2024)
Federated Fine-tuning of Large Language Models under Heterogeneous Tasks and Client Resources
von: Bai, Jiamu, et al.
Veröffentlicht: (2024)
von: Bai, Jiamu, et al.
Veröffentlicht: (2024)
Improved Supervised Fine-Tuning for Large Language Models to Mitigate Catastrophic Forgetting
von: Ding, Fei, et al.
Veröffentlicht: (2025)
von: Ding, Fei, et al.
Veröffentlicht: (2025)
Entity Alignment with Noisy Annotations from Large Language Models
von: Chen, Shengyuan, et al.
Veröffentlicht: (2024)
von: Chen, Shengyuan, et al.
Veröffentlicht: (2024)
ShareLoRA: Parameter Efficient and Robust Large Language Model Fine-tuning via Shared Low-Rank Adaptation
von: Song, Yurun, et al.
Veröffentlicht: (2024)
von: Song, Yurun, et al.
Veröffentlicht: (2024)
JailBreakV: A Benchmark for Assessing the Robustness of MultiModal Large Language Models against Jailbreak Attacks
von: Luo, Weidi, et al.
Veröffentlicht: (2024)
von: Luo, Weidi, et al.
Veröffentlicht: (2024)
EchoDistill:Alignment Noisy-to-Clean Self-Distillation for Robust Audio LLMs
von: Lin, Liang, et al.
Veröffentlicht: (2026)
von: Lin, Liang, et al.
Veröffentlicht: (2026)
When Long Helps Short: How Context Length in Supervised Fine-tuning Affects Behavior of Large Language Models
von: Zheng, Yingming, et al.
Veröffentlicht: (2025)
von: Zheng, Yingming, et al.
Veröffentlicht: (2025)
Embracing Large Language Models in Traffic Flow Forecasting
von: Zhao, Yusheng, et al.
Veröffentlicht: (2024)
von: Zhao, Yusheng, et al.
Veröffentlicht: (2024)
Increasing the Robustness of the Fine-tuned Multilingual Machine-Generated Text Detectors
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
von: Macko, Dominik, et al.
Veröffentlicht: (2025)
Assessing Personalized AI Mentoring with Large Language Models in the Computing Field
von: Luo, Xiao, et al.
Veröffentlicht: (2024)
von: Luo, Xiao, et al.
Veröffentlicht: (2024)
ChatKBQA: A Generate-then-Retrieve Framework for Knowledge Base Question Answering with Fine-tuned Large Language Models
von: Luo, Haoran, et al.
Veröffentlicht: (2023)
von: Luo, Haoran, et al.
Veröffentlicht: (2023)
KnowTuning: Knowledge-aware Fine-tuning for Large Language Models
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
von: Lyu, Yougang, et al.
Veröffentlicht: (2024)
JudgeLM: Fine-tuned Large Language Models are Scalable Judges
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
von: Zhu, Lianghui, et al.
Veröffentlicht: (2023)
Dhati+: Fine-tuned Large Language Models for Arabic Subjectivity Evaluation
von: Bellaouar, Slimane, et al.
Veröffentlicht: (2025)
von: Bellaouar, Slimane, et al.
Veröffentlicht: (2025)
Compositional Subspace Representation Fine-tuning for Adaptive Large Language Models
von: Zhou, Andy
Veröffentlicht: (2025)
von: Zhou, Andy
Veröffentlicht: (2025)
AMANDA: Agentic Medical Knowledge Augmentation for Data-Efficient Medical Visual Question Answering
von: Wang, Ziqing, et al.
Veröffentlicht: (2025)
von: Wang, Ziqing, et al.
Veröffentlicht: (2025)
Impact of Fine-Tuning Methods on Memorization in Large Language Models
von: Hou, Jie, et al.
Veröffentlicht: (2025)
von: Hou, Jie, et al.
Veröffentlicht: (2025)
Plug-in and Fine-tuning: Bridging the Gap between Small Language Models and Large Language Models
von: Kim, Kyeonghyun, et al.
Veröffentlicht: (2025)
von: Kim, Kyeonghyun, et al.
Veröffentlicht: (2025)
Phased Instruction Fine-Tuning for Large Language Models
von: Pang, Wei, et al.
Veröffentlicht: (2024)
von: Pang, Wei, et al.
Veröffentlicht: (2024)
Sparse is Enough in Fine-tuning Pre-trained Large Language Models
von: Song, Weixi, et al.
Veröffentlicht: (2023)
von: Song, Weixi, et al.
Veröffentlicht: (2023)
Prompting or Fine-tuning? Exploring Large Language Models for Causal Graph Validation
von: Susanti, Yuni, et al.
Veröffentlicht: (2024)
von: Susanti, Yuni, et al.
Veröffentlicht: (2024)
AD-LLM: Benchmarking Large Language Models for Anomaly Detection
von: Yang, Tiankai, et al.
Veröffentlicht: (2024)
von: Yang, Tiankai, et al.
Veröffentlicht: (2024)
A Survey on Efficient Large Language Model Training: From Data-centric Perspectives
von: Luo, Junyu, et al.
Veröffentlicht: (2025)
von: Luo, Junyu, et al.
Veröffentlicht: (2025)
ITERTL: An Iterative Framework for Fine-tuning LLMs for RTL Code Generation
von: Wu, Peiyang, et al.
Veröffentlicht: (2024)
von: Wu, Peiyang, et al.
Veröffentlicht: (2024)
Advancing Single and Multi-task Text Classification through Large Language Model Fine-tuning
von: Zhao, Hang, et al.
Veröffentlicht: (2024)
von: Zhao, Hang, et al.
Veröffentlicht: (2024)
You Only Fine-tune Once: Many-Shot In-Context Fine-Tuning for Large Language Models
von: He, Wenchong, et al.
Veröffentlicht: (2025)
von: He, Wenchong, et al.
Veröffentlicht: (2025)
Robustness of Prompting: Enhancing Robustness of Large Language Models Against Prompting Attacks
von: Mu, Lin, et al.
Veröffentlicht: (2025)
von: Mu, Lin, et al.
Veröffentlicht: (2025)
Packing Analysis: Packing Is More Appropriate for Large Models or Datasets in Supervised Fine-tuning
von: Wang, Shuhe, et al.
Veröffentlicht: (2024)
von: Wang, Shuhe, et al.
Veröffentlicht: (2024)
Demystifying Instruction Mixing for Fine-tuning Large Language Models
von: Wang, Renxi, et al.
Veröffentlicht: (2023)
von: Wang, Renxi, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Semi-supervised Fine-tuning for Large Language Models
von: Luo, Junyu, et al.
Veröffentlicht: (2024) -
Attention Bootstrapping for Multi-Modal Test-Time Adaptation
von: Zhao, Yusheng, et al.
Veröffentlicht: (2025) -
VRPO: Rethinking Value Modeling for Robust RL Training under Noisy Supervision
von: Zhu, Dingwei, et al.
Veröffentlicht: (2025) -
Multi-Party Supervised Fine-tuning of Language Models for Multi-Party Dialogue Generation
von: Wang, Xiaoyu, et al.
Veröffentlicht: (2024) -
GALA: Graph Diffusion-based Alignment with Jigsaw for Source-free Domain Adaptation
von: Luo, Junyu, et al.
Veröffentlicht: (2024)