As Simple as Fine-tuning: LLM Alignment via Bidirectional Negative Feedback Loss
Fuente:
arXiv
Salvato in:
| Autori principali: | Mao, Xin, Li, Feng-Lin, Xu, Huimin, Zhang, Wei, Chen, Wang, Luu, Anh Tuan |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Don't Forget Your Reward Values: Language Model Alignment via Value-based Calibration
di: Mao, Xin, et al.
Pubblicazione: (2024)
di: Mao, Xin, et al.
Pubblicazione: (2024)
Full-Step-DPO: Self-Supervised Preference Optimization with Step-wise Rewards for Mathematical Reasoning
di: Xu, Huimin, et al.
Pubblicazione: (2025)
di: Xu, Huimin, et al.
Pubblicazione: (2025)
SynTQA: Synergistic Table-based Question Answering via Mixture of Text-to-SQL and E2E TQA
di: Zhang, Siyue, et al.
Pubblicazione: (2024)
di: Zhang, Siyue, et al.
Pubblicazione: (2024)
Expand BERT Representation with Visual Information via Grounded Language Learning with Multimodal Partial Alignment
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2023)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2023)
SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation
di: Xu, Huimin, et al.
Pubblicazione: (2025)
di: Xu, Huimin, et al.
Pubblicazione: (2025)
ChatKBQA: A Generate-then-Retrieve Framework for Knowledge Base Question Answering with Fine-tuned Large Language Models
di: Luo, Haoran, et al.
Pubblicazione: (2023)
di: Luo, Haoran, et al.
Pubblicazione: (2023)
Memento: Fine-tuning LLM Agents without Fine-tuning LLMs
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
di: Zhou, Huichi, et al.
Pubblicazione: (2025)
Fine-Mem: Fine-Grained Feedback Alignment for Long-Horizon Memory Management
di: Ma, Weitao, et al.
Pubblicazione: (2026)
di: Ma, Weitao, et al.
Pubblicazione: (2026)
Exploring the Potential of Large Language Models in Computational Argumentation
di: Chen, Guizhen, et al.
Pubblicazione: (2023)
di: Chen, Guizhen, et al.
Pubblicazione: (2023)
Why LLM Safety Guardrails Collapse After Fine-tuning: A Similarity Analysis Between Alignment and Fine-tuning Datasets
di: Hsiung, Lei, et al.
Pubblicazione: (2025)
di: Hsiung, Lei, et al.
Pubblicazione: (2025)
Unsupervised Hallucination Detection by Inspecting Reasoning Processes
di: Srey, Ponhvoan, et al.
Pubblicazione: (2025)
di: Srey, Ponhvoan, et al.
Pubblicazione: (2025)
PMSS: Pretrained Matrices Skeleton Selection for LLM Fine-tuning
di: Wang, Qibin, et al.
Pubblicazione: (2024)
di: Wang, Qibin, et al.
Pubblicazione: (2024)
Are LLMs Good Zero-Shot Fallacy Classifiers?
di: Pan, Fengjun, et al.
Pubblicazione: (2024)
di: Pan, Fengjun, et al.
Pubblicazione: (2024)
LLM-AutoDP: Automatic Data Processing via LLM Agents for Model Fine-tuning
di: Huang, Wei, et al.
Pubblicazione: (2026)
di: Huang, Wei, et al.
Pubblicazione: (2026)
UniBridge: A Unified Approach to Cross-Lingual Transfer Learning for Low-Resource Languages
di: Pham, Trinh, et al.
Pubblicazione: (2024)
di: Pham, Trinh, et al.
Pubblicazione: (2024)
Token-level Data Selection for Safe LLM Fine-tuning
di: Li, Yanping, et al.
Pubblicazione: (2026)
di: Li, Yanping, et al.
Pubblicazione: (2026)
A Survey on Neural Topic Models: Methods, Applications, and Challenges
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
di: Bhardwaj, Rishabh, et al.
Pubblicazione: (2024)
di: Bhardwaj, Rishabh, et al.
Pubblicazione: (2024)
On the Loss of Context-awareness in General Instruction Fine-tuning
di: Wang, Yihan, et al.
Pubblicazione: (2024)
di: Wang, Yihan, et al.
Pubblicazione: (2024)
BenchBench: Benchmarking Automated Benchmark Generation
di: Zheng, Yandan, et al.
Pubblicazione: (2026)
di: Zheng, Yandan, et al.
Pubblicazione: (2026)
Fine-tuning BERT with Bidirectional LSTM for Fine-grained Movie Reviews Sentiment Analysis
di: Nkhata, Gibson, et al.
Pubblicazione: (2025)
di: Nkhata, Gibson, et al.
Pubblicazione: (2025)
TREX: Automating LLM Fine-tuning via Agent-Driven Tree-based Exploration
di: Ma, Zerun, et al.
Pubblicazione: (2026)
di: Ma, Zerun, et al.
Pubblicazione: (2026)
RLSF: Fine-tuning LLMs via Symbolic Feedback
di: Jha, Piyush, et al.
Pubblicazione: (2024)
di: Jha, Piyush, et al.
Pubblicazione: (2024)
Enriching and Controlling Global Semantics for Text Summarization
di: Nguyen, Thong, et al.
Pubblicazione: (2021)
di: Nguyen, Thong, et al.
Pubblicazione: (2021)
Read as You See: Guiding Unimodal LLMs for Low-Resource Explainable Harmful Meme Detection
di: Pan, Fengjun, et al.
Pubblicazione: (2025)
di: Pan, Fengjun, et al.
Pubblicazione: (2025)
Small Updates, Big Doubts: Does Parameter-Efficient Fine-tuning Enhance Hallucination Detection ?
di: Hu, Xu, et al.
Pubblicazione: (2026)
di: Hu, Xu, et al.
Pubblicazione: (2026)
Towards the TopMost: A Topic Modeling System Toolkit
di: Wu, Xiaobao, et al.
Pubblicazione: (2023)
di: Wu, Xiaobao, et al.
Pubblicazione: (2023)
Mitigating Fine-tuning based Jailbreak Attack with Backdoor Enhanced Safety Alignment
di: Wang, Jiongxiao, et al.
Pubblicazione: (2024)
di: Wang, Jiongxiao, et al.
Pubblicazione: (2024)
Discrete Diffusion Language Model for Efficient Text Summarization
di: Dat, Do Huu, et al.
Pubblicazione: (2024)
di: Dat, Do Huu, et al.
Pubblicazione: (2024)
More Bias, Less Bias: BiasPrompting for Enhanced Multiple-Choice Question Answering
di: Vu, Duc Anh, et al.
Pubblicazione: (2025)
di: Vu, Duc Anh, et al.
Pubblicazione: (2025)
MRAG: A Modular Retrieval Framework for Time-Sensitive Question Answering
di: Siyue, Zhang, et al.
Pubblicazione: (2024)
di: Siyue, Zhang, et al.
Pubblicazione: (2024)
Is Translation All You Need? A Study on Solving Multilingual Tasks with Large Language Models
di: Liu, Chaoqun, et al.
Pubblicazione: (2024)
di: Liu, Chaoqun, et al.
Pubblicazione: (2024)
AKEW: Assessing Knowledge Editing in the Wild
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
MAIN: Mutual Alignment Is Necessary for instruction tuning
di: Yang, Fanyi, et al.
Pubblicazione: (2025)
di: Yang, Fanyi, et al.
Pubblicazione: (2025)
Utility-Diversity Aware Online Batch Selection for LLM Supervised Fine-tuning
di: Zou, Heming, et al.
Pubblicazione: (2025)
di: Zou, Heming, et al.
Pubblicazione: (2025)
Fine-tuning large language models for domain adaptation: Exploration of training strategies, scaling, model merging and synergistic capabilities
di: Lu, Wei, et al.
Pubblicazione: (2024)
di: Lu, Wei, et al.
Pubblicazione: (2024)
LAMPAT: Low-Rank Adaption for Multilingual Paraphrasing Using Adversarial Training
di: Le, Khoi M., et al.
Pubblicazione: (2024)
di: Le, Khoi M., et al.
Pubblicazione: (2024)
KDMCSE: Knowledge Distillation Multimodal Sentence Embeddings with Adaptive Angular margin Contrastive Learning
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2024)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2024)
Aspect-Based Summarization with Self-Aspect Retrieval Enhanced Generation
di: Feng, Yichao, et al.
Pubblicazione: (2025)
di: Feng, Yichao, et al.
Pubblicazione: (2025)
Learning From Failure: Integrating Negative Examples when Fine-tuning Large Language Models as Agents
di: Wang, Renxi, et al.
Pubblicazione: (2024)
di: Wang, Renxi, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Don't Forget Your Reward Values: Language Model Alignment via Value-based Calibration
di: Mao, Xin, et al.
Pubblicazione: (2024) -
Full-Step-DPO: Self-Supervised Preference Optimization with Step-wise Rewards for Mathematical Reasoning
di: Xu, Huimin, et al.
Pubblicazione: (2025) -
SynTQA: Synergistic Table-based Question Answering via Mixture of Text-to-SQL and E2E TQA
di: Zhang, Siyue, et al.
Pubblicazione: (2024) -
Expand BERT Representation with Visual Information via Grounded Language Learning with Multimodal Partial Alignment
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2023) -
SCOPE: Compress Mathematical Reasoning Steps for Efficient Automated Process Annotation
di: Xu, Huimin, et al.
Pubblicazione: (2025)