Plug-and-Play Training Framework for Preference Optimization
Fuente:
arXiv
Saved in:
| Main Authors: | Ma, Jingyuan, Li, Rui, Li, Zheng, Sha, Lei, Sui, Zhifang |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Be a Multitude to Itself: A Prompt Evolution Framework for Red Teaming
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
HauntAttack: When Attack Follows Reasoning as a Shadow
by: Ma, Jingyuan, et al.
Published: (2025)
by: Ma, Jingyuan, et al.
Published: (2025)
Harnessing the Plug-and-Play Controller by Prompting
by: Wang, Hao, et al.
Published: (2024)
by: Wang, Hao, et al.
Published: (2024)
Large Language Models Struggle with Unreasonability in Math Problems
by: Ma, Jingyuan, et al.
Published: (2024)
by: Ma, Jingyuan, et al.
Published: (2024)
How Far are LLMs from Being Our Digital Twins? A Benchmark for Persona-Based Behavior Chain Simulation
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
SelfBudgeter: Adaptive Token Allocation for Efficient LLM Reasoning
by: Li, Zheng, et al.
Published: (2025)
by: Li, Zheng, et al.
Published: (2025)
DSAS: A Universal Plug-and-Play Framework for Attention Optimization in Multi-Document Question Answering
by: Li, Jiakai, et al.
Published: (2025)
by: Li, Jiakai, et al.
Published: (2025)
ShieldLM: Empowering LLMs as Aligned, Customizable and Explainable Safety Detectors
by: Zhang, Zhexin, et al.
Published: (2024)
by: Zhang, Zhexin, et al.
Published: (2024)
Towards Harmonized Uncertainty Estimation for Large Language Models
by: Li, Rui, et al.
Published: (2025)
by: Li, Rui, et al.
Published: (2025)
Towards Better RL Training Data Utilization via Second-Order Rollout
by: Yang, Zhe, et al.
Published: (2026)
by: Yang, Zhe, et al.
Published: (2026)
Self-Boosting Large Language Models with Synthetic Preference Data
by: Dong, Qingxiu, et al.
Published: (2024)
by: Dong, Qingxiu, et al.
Published: (2024)
Oreo: A Plug-in Context Reconstructor to Enhance Retrieval-Augmented Generation
by: Li, Sha, et al.
Published: (2025)
by: Li, Sha, et al.
Published: (2025)
PCToolkit: A Unified Plug-and-Play Prompt Compression Toolkit of Large Language Models
by: Li, Jinyi, et al.
Published: (2024)
by: Li, Jinyi, et al.
Published: (2024)
CoLT: Reasoning with Chain of Latent Tool Calls
by: Zhu, Fangwei, et al.
Published: (2026)
by: Zhu, Fangwei, et al.
Published: (2026)
Be Your Own Red Teamer: Safety Alignment via Self-Play and Reflective Experience Replay
by: Wang, Hao, et al.
Published: (2026)
by: Wang, Hao, et al.
Published: (2026)
360-LLaMA-Factory: Plug & Play Sequence Parallelism for Long Post-Training
by: Zou, Haosheng, et al.
Published: (2025)
by: Zou, Haosheng, et al.
Published: (2025)
PAS: Data-Efficient Plug-and-Play Prompt Augmentation System
by: Zheng, Miao, et al.
Published: (2024)
by: Zheng, Miao, et al.
Published: (2024)
Reinforcement Pre-Training
by: Dong, Qingxiu, et al.
Published: (2025)
by: Dong, Qingxiu, et al.
Published: (2025)
A Survey on In-context Learning
by: Dong, Qingxiu, et al.
Published: (2022)
by: Dong, Qingxiu, et al.
Published: (2022)
Can Large Multimodal Models Uncover Deep Semantics Behind Images?
by: Yang, Yixin, et al.
Published: (2024)
by: Yang, Yixin, et al.
Published: (2024)
Towards Extreme Pruning of LLMs with Plug-and-Play Mixed Sparsity
by: Xu, Chi, et al.
Published: (2025)
by: Xu, Chi, et al.
Published: (2025)
Generation-Augmented Generation: A Plug-and-Play Framework for Private Knowledge Injection in Large Language Models
by: Li, Rongji, et al.
Published: (2026)
by: Li, Rongji, et al.
Published: (2026)
AgentMonitor: A Plug-and-Play Framework for Predictive and Secure Multi-Agent Systems
by: Chan, Chi-Min, et al.
Published: (2024)
by: Chan, Chi-Min, et al.
Published: (2024)
Self-Checker: Plug-and-Play Modules for Fact-Checking with Large Language Models
by: Li, Miaoran, et al.
Published: (2023)
by: Li, Miaoran, et al.
Published: (2023)
Reducing Hallucinations in Entity Abstract Summarization with Facts-Template Decomposition
by: Zhu, Fangwei, et al.
Published: (2024)
by: Zhu, Fangwei, et al.
Published: (2024)
Language Models Encode the Value of Numbers Linearly
by: Zhu, Fangwei, et al.
Published: (2024)
by: Zhu, Fangwei, et al.
Published: (2024)
CIP: A Plug-and-Play Causal Prompting Framework for Mitigating Hallucinations under Long-Context Noise
by: Ma, Qingsen, et al.
Published: (2025)
by: Ma, Qingsen, et al.
Published: (2025)
Decoupled Alignment for Robust Plug-and-Play Adaptation
by: Luo, Haozheng, et al.
Published: (2024)
by: Luo, Haozheng, et al.
Published: (2024)
A Knowledge Plug-and-Play Test Bed for Open-domain Dialogue Generation
by: Li, Xiangci, et al.
Published: (2024)
by: Li, Xiangci, et al.
Published: (2024)
Plug-and-Play Grounding of Reasoning in Multimodal Large Language Models
by: Chen, Jiaxing, et al.
Published: (2024)
by: Chen, Jiaxing, et al.
Published: (2024)
SCoRE: Benchmarking Long-Chain Reasoning in Commonsense Scenarios
by: Zhan, Weidong, et al.
Published: (2025)
by: Zhan, Weidong, et al.
Published: (2025)
PDR: A Plug-and-Play Positional Decay Framework for LLM Pre-training Data Detection
by: Liu, Jinhan, et al.
Published: (2026)
by: Liu, Jinhan, et al.
Published: (2026)
Revisiting Self-Play Preference Optimization: On the Role of Prompt Difficulty
by: Xiao, Yao, et al.
Published: (2025)
by: Xiao, Yao, et al.
Published: (2025)
Align With Purpose: Optimize Desired Properties in CTC Models with a General Plug-and-Play Framework
by: Segev, Eliya, et al.
Published: (2023)
by: Segev, Eliya, et al.
Published: (2023)
DISC: Plug-and-Play Decoding Intervention with Similarity of Characters for Chinese Spelling Check
by: Qiao, Ziheng, et al.
Published: (2024)
by: Qiao, Ziheng, et al.
Published: (2024)
Enhancing Reliability across Short and Long-Form QA via Reinforcement Learning
by: Wang, Yudong, et al.
Published: (2025)
by: Wang, Yudong, et al.
Published: (2025)
From the Least to the Most: Building a Plug-and-Play Visual Reasoner via Data Synthesis
by: Cheng, Chuanqi, et al.
Published: (2024)
by: Cheng, Chuanqi, et al.
Published: (2024)
Large Language Models Enhanced by Plug and Play Syntactic Knowledge for Aspect-based Sentiment Analysis
by: Tian, Yuanhe, et al.
Published: (2025)
by: Tian, Yuanhe, et al.
Published: (2025)
Exploring Activation Patterns of Parameters in Language Models
by: Wang, Yudong, et al.
Published: (2024)
by: Wang, Yudong, et al.
Published: (2024)
Chain-of-Thought Tokens are Computer Program Variables
by: Zhu, Fangwei, et al.
Published: (2025)
by: Zhu, Fangwei, et al.
Published: (2025)
Similar Items
-
Be a Multitude to Itself: A Prompt Evolution Framework for Red Teaming
by: Li, Rui, et al.
Published: (2025) -
HauntAttack: When Attack Follows Reasoning as a Shadow
by: Ma, Jingyuan, et al.
Published: (2025) -
Harnessing the Plug-and-Play Controller by Prompting
by: Wang, Hao, et al.
Published: (2024) -
Large Language Models Struggle with Unreasonability in Math Problems
by: Ma, Jingyuan, et al.
Published: (2024) -
How Far are LLMs from Being Our Digital Twins? A Benchmark for Persona-Based Behavior Chain Simulation
by: Li, Rui, et al.
Published: (2025)