Sem-DPO: Mitigating Semantic Inconsistency in Preference Optimization for Prompt Engineering
Fuente:
arXiv
Saved in:
| Main Authors: | Mohamed, Anas, Khan, Azal Ahmad, Wang, Xinran, Khan, Ahmad Faraz, Ge, Shuwen, Khan, Saman Bahzad, Ahmad, Ayaan, Anwar, Ali |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Personalized Federated Learning Techniques: Empirical Analysis
by: Khan, Azal Ahmad, et al.
Published: (2024)
by: Khan, Azal Ahmad, et al.
Published: (2024)
Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts
by: Ahmed, Ammar, et al.
Published: (2025)
by: Ahmed, Ammar, et al.
Published: (2025)
LADs: Leveraging LLMs for AI-Driven DevOps
by: Khan, Ahmad Faraz, et al.
Published: (2025)
by: Khan, Ahmad Faraz, et al.
Published: (2025)
Accelerating LLM Reasoning via Early Rejection with Partial Reward Modeling
by: Cheshmi, Seyyed Saeid, et al.
Published: (2025)
by: Cheshmi, Seyyed Saeid, et al.
Published: (2025)
IP-FL: Incentivized and Personalized Federated Learning
by: Khan, Ahmad Faraz, et al.
Published: (2023)
by: Khan, Ahmad Faraz, et al.
Published: (2023)
Faster Synchronous On-Policy RL via Straggler-Aware Group Sizing
by: Khan, Azal Ahmad, et al.
Published: (2026)
by: Khan, Azal Ahmad, et al.
Published: (2026)
You Don't Need Prompt Engineering Anymore: The Prompting Inversion
by: Khan, Imran
Published: (2025)
by: Khan, Imran
Published: (2025)
SABRE-FL: Selective and Accurate Backdoor Rejection for Federated Prompt Learning
by: Khan, Momin Ahmad, et al.
Published: (2025)
by: Khan, Momin Ahmad, et al.
Published: (2025)
FLStore: Efficient Federated Learning Storage for non-training workloads
by: Khan, Ahmad Faraz, et al.
Published: (2025)
by: Khan, Ahmad Faraz, et al.
Published: (2025)
Enhancing Few-Shot Transfer Learning with Optimized Multi-Task Prompt Tuning through Modular Prompt Composition
by: Pouramini, Ahmad, et al.
Published: (2024)
by: Pouramini, Ahmad, et al.
Published: (2024)
Safety Aware Task Planning via Large Language Models in Robotics
by: Khan, Azal Ahmad, et al.
Published: (2025)
by: Khan, Azal Ahmad, et al.
Published: (2025)
OPUS-VFL: Incentivizing Optimal Privacy-Utility Tradeoffs in Vertical Federated Learning
by: Madabushi, Sindhuja, et al.
Published: (2025)
by: Madabushi, Sindhuja, et al.
Published: (2025)
AlignCheck: a Semantic Open-Domain Metric for Factual Consistency Assessment
by: Aghaebrahimian, Ahmad
Published: (2025)
by: Aghaebrahimian, Ahmad
Published: (2025)
ArabicNumBench: Evaluating Arabic Number Reading in Large Language Models
by: Alhumud, Anas, et al.
Published: (2026)
by: Alhumud, Anas, et al.
Published: (2026)
Assessing the Impact of Code Changes on the Fault Localizability of Large Language Models
by: Haroon, Sabaat, et al.
Published: (2025)
by: Haroon, Sabaat, et al.
Published: (2025)
CLaC at SemEval-2025 Task 6: A Multi-Architecture Approach for Corporate Environmental Promise Verification
by: Turk, Nawar, et al.
Published: (2025)
by: Turk, Nawar, et al.
Published: (2025)
From Literal to Liberal: A Meta-Prompting Framework for Eliciting Human-Aligned Exception Handling in Large Language Models
by: Khan, Imran
Published: (2025)
by: Khan, Imran
Published: (2025)
CrossPT: Exploring Cross-Task Transferability through Multi-Task Prompt Tuning
by: Pouramini, Ahmad, et al.
Published: (2025)
by: Pouramini, Ahmad, et al.
Published: (2025)
Beyond Rebalancing: Benchmarking Binary Classifiers Under Class Imbalance Without Rebalancing Techniques
by: Nawaz, Ali, et al.
Published: (2025)
by: Nawaz, Ali, et al.
Published: (2025)
Learning from Contrastive Prompts: Automated Optimization and Adaptation
by: Li, Mingqi, et al.
Published: (2024)
by: Li, Mingqi, et al.
Published: (2024)
Stress Classification from ECG Signals Using Vision Transformer
by: Ahmad, Zeeshan, et al.
Published: (2026)
by: Ahmad, Zeeshan, et al.
Published: (2026)
AlphaDPO: Adaptive Reward Margin for Direct Preference Optimization
by: Wu, Junkang, et al.
Published: (2024)
by: Wu, Junkang, et al.
Published: (2024)
SEMFED: Semantic-Aware Resource-Efficient Federated Learning for Heterogeneous NLP Tasks
by: Hussain, Sajid, et al.
Published: (2025)
by: Hussain, Sajid, et al.
Published: (2025)
2D-DPO: Scaling Direct Preference Optimization with 2-Dimensional Supervision
by: Li, Shilong, et al.
Published: (2024)
by: Li, Shilong, et al.
Published: (2024)
Efficient IoT Intrusion Detection with an Improved Attention-Based CNN-BiLSTM Architecture
by: Naeem, Amna, et al.
Published: (2025)
by: Naeem, Amna, et al.
Published: (2025)
DPO Kernels: A Semantically-Aware, Kernel-Enhanced, and Divergence-Rich Paradigm for Direct Preference Optimization
by: Das, Amitava, et al.
Published: (2025)
by: Das, Amitava, et al.
Published: (2025)
Online DPO: Online Direct Preference Optimization with Fast-Slow Chasing
by: Qi, Biqing, et al.
Published: (2024)
by: Qi, Biqing, et al.
Published: (2024)
Multi-Preference Optimization: Generalizing DPO via Set-Level Contrasts
by: Gupta, Taneesh, et al.
Published: (2024)
by: Gupta, Taneesh, et al.
Published: (2024)
Prompting Implicit Discourse Relation Annotation
by: Yung, Frances, et al.
Published: (2024)
by: Yung, Frances, et al.
Published: (2024)
Rethinking DPO: The Role of Rejected Responses in Preference Misalignment
by: Cho, Jay Hyeon, et al.
Published: (2025)
by: Cho, Jay Hyeon, et al.
Published: (2025)
Promptception: How Sensitive Are Large Multimodal Models to Prompts?
by: Ismithdeen, Mohamed Insaf, et al.
Published: (2025)
by: Ismithdeen, Mohamed Insaf, et al.
Published: (2025)
HausaNLP at SemEval-2025 Task 3: Towards a Fine-Grained Model-Aware Hallucination Detection
by: Bala, Maryam, et al.
Published: (2025)
by: Bala, Maryam, et al.
Published: (2025)
Step-DPO: Step-wise Preference Optimization for Long-chain Reasoning of LLMs
by: Lai, Xin, et al.
Published: (2024)
by: Lai, Xin, et al.
Published: (2024)
Percentile Optimization in Wireless Networks- Part II: Beamforming for Cell-Edge Throughput Maximization
by: Khan, Ahmad Ali, et al.
Published: (2024)
by: Khan, Ahmad Ali, et al.
Published: (2024)
SP^2DPO: An LLM-assisted Semantic Per-Pair DPO Generalization
by: He, Chaoyue, et al.
Published: (2026)
by: He, Chaoyue, et al.
Published: (2026)
Evaluating Prompt Engineering Techniques for Accuracy and Confidence Elicitation in Medical LLMs
by: Naderi, Nariman, et al.
Published: (2025)
by: Naderi, Nariman, et al.
Published: (2025)
Mix- and MoE-DPO: A Variational Inference Approach to Direct Preference Optimization
by: Bohne, Jason, et al.
Published: (2025)
by: Bohne, Jason, et al.
Published: (2025)
Percentile Optimization in Wireless Networks- Part I: Power Control for Max-Min-Rate to Sum-Rate Maximization (and Everything in Between)
by: Khan, Ahmad Ali, et al.
Published: (2024)
by: Khan, Ahmad Ali, et al.
Published: (2024)
Towards cost-effective and resource-aware aggregation at Edge for Federated Learning
by: Khan, Ahmad Faraz, et al.
Published: (2022)
by: Khan, Ahmad Faraz, et al.
Published: (2022)
Sumotosima: A Framework and Dataset for Classifying and Summarizing Otoscopic Images
by: Khan, Eram Anwarul, et al.
Published: (2024)
by: Khan, Eram Anwarul, et al.
Published: (2024)
Similar Items
-
Personalized Federated Learning Techniques: Empirical Analysis
by: Khan, Azal Ahmad, et al.
Published: (2024) -
Retrieval-of-Thought: Efficient Reasoning via Reusing Thoughts
by: Ahmed, Ammar, et al.
Published: (2025) -
LADs: Leveraging LLMs for AI-Driven DevOps
by: Khan, Ahmad Faraz, et al.
Published: (2025) -
Accelerating LLM Reasoning via Early Rejection with Partial Reward Modeling
by: Cheshmi, Seyyed Saeid, et al.
Published: (2025) -
IP-FL: Incentivized and Personalized Federated Learning
by: Khan, Ahmad Faraz, et al.
Published: (2023)