Distribution Preference Optimization: A Fine-grained Perspective for LLM Unlearning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Qin, Kai, Wu, Jiaqi, He, Jianxiang, Sun, Haoyuan, Zhao, Yifei, Liang, Bin, Chang, Yongzhe, Zhang, Tiantian, Liu, Houde |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025)
Generalizing Alignment Paradigm of Text-to-Image Generation with Preferences through $f$-divergence Minimization
von: Sun, Haoyuan, et al.
Veröffentlicht: (2024)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2024)
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025)
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
von: Fan, Chongyu, et al.
Veröffentlicht: (2024)
Graph Unlearning Meets Influence-aware Negative Preference Optimization
von: Chen, Qiang, et al.
Veröffentlicht: (2025)
von: Chen, Qiang, et al.
Veröffentlicht: (2025)
Optimizing Long-context LLM Serving via Fine-grained Sequence Parallelism
von: Li, Cong, et al.
Veröffentlicht: (2025)
von: Li, Cong, et al.
Veröffentlicht: (2025)
UACER: An Uncertainty-Adaptive Critic Ensemble Framework for Robust Adversarial Reinforcement Learning
von: Wu, Jiaxi, et al.
Veröffentlicht: (2025)
von: Wu, Jiaxi, et al.
Veröffentlicht: (2025)
Unlearning of Knowledge Graph Embedding via Preference Optimization
von: Liu, Jiajun, et al.
Veröffentlicht: (2025)
von: Liu, Jiajun, et al.
Veröffentlicht: (2025)
Reinforcement Learning Meets Masked Generative Models: Mask-GRPO for Text-to-Image Generation
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
FLRC: Fine-grained Low-Rank Compressor for Efficient LLM Inference
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
von: Lu, Yu-Chen, et al.
Veröffentlicht: (2025)
CoVis: A Collaborative Framework for Fine-grained Graphic Visual Understanding
von: Deng, Xiaoyu, et al.
Veröffentlicht: (2024)
von: Deng, Xiaoyu, et al.
Veröffentlicht: (2024)
BLUR: A Bi-Level Optimization Approach for LLM Unlearning
von: Reisizadeh, Hadi, et al.
Veröffentlicht: (2025)
von: Reisizadeh, Hadi, et al.
Veröffentlicht: (2025)
Fine-grained Video Dubbing Duration Alignment with Segment Supervised Preference Optimization
von: Cui, Chaoqun, et al.
Veröffentlicht: (2025)
von: Cui, Chaoqun, et al.
Veröffentlicht: (2025)
A Method on Searching Better Activation Functions
von: Sun, Haoyuan, et al.
Veröffentlicht: (2024)
von: Sun, Haoyuan, et al.
Veröffentlicht: (2024)
Rubrics as an Attack Surface: Stealthy Preference Drift in LLM Judges
von: Ding, Ruomeng, et al.
Veröffentlicht: (2026)
von: Ding, Ruomeng, et al.
Veröffentlicht: (2026)
Hierarchical Fine-grained Preference Optimization for Physically Plausible Video Generation
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
von: Chen, Harold Haodong, et al.
Veröffentlicht: (2025)
Fine-grained Preference Optimization Improves Zero-shot Text-to-Speech
von: Yao, Jixun, et al.
Veröffentlicht: (2025)
von: Yao, Jixun, et al.
Veröffentlicht: (2025)
Why Some Models Resist Unlearning: A Linear Stability Perspective
von: Chang, Wei-Kai, et al.
Veröffentlicht: (2026)
von: Chang, Wei-Kai, et al.
Veröffentlicht: (2026)
Beyond Hallucinations: Enhancing LVLMs through Hallucination-Aware Direct Preference Optimization
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2023)
von: Zhao, Zhiyuan, et al.
Veröffentlicht: (2023)
Understanding Forgetting in LLM Supervised Fine-Tuning and Preference Learning -- A Convex Optimization Perspective
von: Fernando, Heshan, et al.
Veröffentlicht: (2024)
von: Fernando, Heshan, et al.
Veröffentlicht: (2024)
EFUF: Efficient Fine-grained Unlearning Framework for Mitigating Hallucinations in Multimodal Large Language Models
von: Xing, Shangyu, et al.
Veröffentlicht: (2024)
von: Xing, Shangyu, et al.
Veröffentlicht: (2024)
UniFine: A Unified and Fine-grained Approach for Zero-shot Vision-Language Understanding
von: Wang, Zhecan, et al.
Veröffentlicht: (2023)
von: Wang, Zhecan, et al.
Veröffentlicht: (2023)
A General Framework to Enhance Fine-tuning-based LLM Unlearning
von: Ren, Jie, et al.
Veröffentlicht: (2025)
von: Ren, Jie, et al.
Veröffentlicht: (2025)
FGTR: Fine-Grained Multi-Table Retrieval via Hierarchical LLM Reasoning
von: Sun, Chaojie, et al.
Veröffentlicht: (2026)
von: Sun, Chaojie, et al.
Veröffentlicht: (2026)
Federated Unlearning: a Perspective of Stability and Fairness
von: Shao, Jiaqi, et al.
Veröffentlicht: (2024)
von: Shao, Jiaqi, et al.
Veröffentlicht: (2024)
On-Policy Optimization with Group Equivalent Preference for Multi-Programming Language Understanding
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
von: Wu, Haoyuan, et al.
Veröffentlicht: (2025)
Certified Minimax Unlearning with Generalization Rates and Deletion Capacity
von: Liu, Jiaqi, et al.
Veröffentlicht: (2023)
von: Liu, Jiaqi, et al.
Veröffentlicht: (2023)
Argus: Token Aware Distributed LLM Inference Optimization
von: Wu, Panlong, et al.
Veröffentlicht: (2025)
von: Wu, Panlong, et al.
Veröffentlicht: (2025)
D-FINE: Redefine Regression Task in DETRs as Fine-grained Distribution Refinement
von: Peng, Yansong, et al.
Veröffentlicht: (2024)
von: Peng, Yansong, et al.
Veröffentlicht: (2024)
EPO: Hierarchical LLM Agents with Environment Preference Optimization
von: Zhao, Qi, et al.
Veröffentlicht: (2024)
von: Zhao, Qi, et al.
Veröffentlicht: (2024)
Refine-n-Judge: Curating High-Quality Preference Chains for LLM-Fine-Tuning
von: Cayir, Derin, et al.
Veröffentlicht: (2025)
von: Cayir, Derin, et al.
Veröffentlicht: (2025)
Direct Preference Optimization for LLM-Enhanced Recommendation Systems
von: Sun, Chao, et al.
Veröffentlicht: (2024)
von: Sun, Chao, et al.
Veröffentlicht: (2024)
Putting on the Thinking Hats: A Survey on Chain of Thought Fine-tuning from the Perspective of Human Reasoning Mechanism
von: Chen, Xiaoshu, et al.
Veröffentlicht: (2025)
von: Chen, Xiaoshu, et al.
Veröffentlicht: (2025)
Eguard: Defending LLM Embeddings Against Inversion Attacks via Text Mutual Information Optimization
von: Liu, Tiantian, et al.
Veröffentlicht: (2024)
von: Liu, Tiantian, et al.
Veröffentlicht: (2024)
MEMO: Fine-grained Tensor Management For Ultra-long Context LLM Training
von: Zhao, Pinxue, et al.
Veröffentlicht: (2024)
von: Zhao, Pinxue, et al.
Veröffentlicht: (2024)
Fine-grained Analysis of Stability and Generalization for Stochastic Bilevel Optimization
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
von: Zhang, Xuelin, et al.
Veröffentlicht: (2026)
Principled RL for Flow Matching Emerges from the Chunk-level Policy Optimization
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
von: Luo, Yifu, et al.
Veröffentlicht: (2025)
LUME: LLM Unlearning with Multitask Evaluations
von: Ramakrishna, Anil, et al.
Veröffentlicht: (2025)
von: Ramakrishna, Anil, et al.
Veröffentlicht: (2025)
Unlock the Potential of Fine-grained LLM Serving via Dynamic Module Scaling
von: Wu, Jingfeng, et al.
Veröffentlicht: (2025)
von: Wu, Jingfeng, et al.
Veröffentlicht: (2025)
Morphology and Behavior Co-Optimization of Modular Satellites for Attitude Control
von: Wang, Yuxing, et al.
Veröffentlicht: (2024)
von: Wang, Yuxing, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Reinforcement Fine-Tuning Powers Reasoning Capability of Multimodal Large Language Models
von: Sun, Haoyuan, et al.
Veröffentlicht: (2025) -
Generalizing Alignment Paradigm of Text-to-Image Generation with Preferences through $f$-divergence Minimization
von: Sun, Haoyuan, et al.
Veröffentlicht: (2024) -
Unlearned but Not Forgotten: Data Extraction after Exact Unlearning in LLM
von: Wu, Xiaoyu, et al.
Veröffentlicht: (2025) -
Simplicity Prevails: Rethinking Negative Preference Optimization for LLM Unlearning
von: Fan, Chongyu, et al.
Veröffentlicht: (2024) -
Graph Unlearning Meets Influence-aware Negative Preference Optimization
von: Chen, Qiang, et al.
Veröffentlicht: (2025)