DPO Unchained: Your Training Algorithm is Secretly Disentangled in Human Choice Theory
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Wenxuan, Zhang, Shujian, Magdalou, Brice, Lambert, John, Amid, Ehsan, Nock, Richard, Hard, Andrew |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Tempered Calculus for ML: Application to Hyperbolic Model Embedding
por: Nock, Richard, et al.
Publicado: (2024)
por: Nock, Richard, et al.
Publicado: (2024)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
por: Fadli, Samih
Publicado: (2025)
por: Fadli, Samih
Publicado: (2025)
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
por: Srishty, Sharmin Sultana, et al.
Publicado: (2026)
por: Srishty, Sharmin Sultana, et al.
Publicado: (2026)
Survey Transfer Learning: Recycling Data with Silicon Responses
por: Amini, Ali
Publicado: (2025)
por: Amini, Ali
Publicado: (2025)
TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors
por: Atad, Ido Andrew, et al.
Publicado: (2026)
por: Atad, Ido Andrew, et al.
Publicado: (2026)
Learned Relay Representations for Forward-Thinking Discrete Diffusion Models
por: Rozonoyer, Benjamin, et al.
Publicado: (2026)
por: Rozonoyer, Benjamin, et al.
Publicado: (2026)
PRISMA: Preference-Reinforced Self-Training Approach for Interpretable Emotionally Intelligent Negotiation Dialogues
por: Kajare, Prajwal Vijay, et al.
Publicado: (2026)
por: Kajare, Prajwal Vijay, et al.
Publicado: (2026)
Unveiling the Secret Recipe: A Guide For Supervised Fine-Tuning Small LLMs
por: Pareja, Aldo, et al.
Publicado: (2024)
por: Pareja, Aldo, et al.
Publicado: (2024)
Ask WhAI:Probing Belief Formation in Role-Primed LLM Agents
por: Moore, Keith, et al.
Publicado: (2025)
por: Moore, Keith, et al.
Publicado: (2025)
Your Pretrained Model Tells the Difficulty Itself: A Self-Adaptive Curriculum Learning Paradigm for Natural Language Understanding
por: Feng, Qi, et al.
Publicado: (2025)
por: Feng, Qi, et al.
Publicado: (2025)
Automated Bug Triaging using Instruction-Tuned Large Language Models
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
por: Kiashemshaki, Kiana, et al.
Publicado: (2025)
Merge-Bench: Resolve Merge Conflicts with Large Language Models
por: Schesch, Benedikt, et al.
Publicado: (2026)
por: Schesch, Benedikt, et al.
Publicado: (2026)
Truth as a Compression Artifact in Language Model Training
por: Krestnikov, Konstantin
Publicado: (2026)
por: Krestnikov, Konstantin
Publicado: (2026)
ACE: Exploring Activation Cosine Similarity and Variance for Accurate and Calibration-Efficient LLM Pruning
por: Mi, Zhendong, et al.
Publicado: (2025)
por: Mi, Zhendong, et al.
Publicado: (2025)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
por: Gwak, Jiho, et al.
Publicado: (2025)
por: Gwak, Jiho, et al.
Publicado: (2025)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
por: Liu, Zhongxin, et al.
Publicado: (2025)
por: Liu, Zhongxin, et al.
Publicado: (2025)
On the Influence of Discourse Relations in Persuasive Texts
por: Turk, Nawar, et al.
Publicado: (2025)
por: Turk, Nawar, et al.
Publicado: (2025)
Latent Instruction Representation Alignment: defending against jailbreaks, backdoors and undesired knowledge in LLMs
por: Easley, Eric, et al.
Publicado: (2026)
por: Easley, Eric, et al.
Publicado: (2026)
Calibrated Confidence Estimation for Tabular Question Answering
por: Voss, Lukas
Publicado: (2026)
por: Voss, Lukas
Publicado: (2026)
Automated CAD Modeling Sequence Generation from Text Descriptions via Transformer-Based Large Language Models
por: Liao, Jianxing, et al.
Publicado: (2025)
por: Liao, Jianxing, et al.
Publicado: (2025)
Scalable GPU-Accelerated Euler Characteristic Curves: Optimization and Differentiable Learning for PyTorch
por: Saxena, Udit
Publicado: (2025)
por: Saxena, Udit
Publicado: (2025)
JURY-RL: Votes Propose, Proofs Dispose for Label-Free RLVR
por: Chen, Xinjie, et al.
Publicado: (2026)
por: Chen, Xinjie, et al.
Publicado: (2026)
Descriptive Collision in Sparse Autoencoder Auto-Interpretability: When One Explanation Describes Many Features
por: McCann, Jordan F.
Publicado: (2026)
por: McCann, Jordan F.
Publicado: (2026)
Super Apriel: One Checkpoint, Many Speeds
por: Labs, SLAM, et al.
Publicado: (2026)
por: Labs, SLAM, et al.
Publicado: (2026)
Synergy over Discrepancy: A Partition-Based Approach to Multi-Domain LLM Fine-Tuning
por: Ye, Hua, et al.
Publicado: (2025)
por: Ye, Hua, et al.
Publicado: (2025)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
por: Han, Xudong, et al.
Publicado: (2025)
por: Han, Xudong, et al.
Publicado: (2025)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
por: Kugler, Kai
Publicado: (2025)
por: Kugler, Kai
Publicado: (2025)
Character-Level Transformer for Tajik-Persian Transliteration with a Parallel Lexical Corpus
por: Arabov, Mullosharaf K.
Publicado: (2026)
por: Arabov, Mullosharaf K.
Publicado: (2026)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
por: Resck, Lucas, et al.
Publicado: (2026)
por: Resck, Lucas, et al.
Publicado: (2026)
BitCal-TTS: Bit-Calibrated Test-Time Scaling for Quantized Reasoning Models
por: Patarlapalli, Sai Babu, et al.
Publicado: (2026)
por: Patarlapalli, Sai Babu, et al.
Publicado: (2026)
KerZOO: Kernel Function Informed Zeroth-Order Optimization for Accurate and Accelerated LLM Fine-Tuning
por: Mi, Zhendong, et al.
Publicado: (2025)
por: Mi, Zhendong, et al.
Publicado: (2025)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
por: Bertina, Abbas, et al.
Publicado: (2025)
por: Bertina, Abbas, et al.
Publicado: (2025)
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
por: Zhang, Bowen, et al.
Publicado: (2025)
por: Zhang, Bowen, et al.
Publicado: (2025)
Revisiting LRP: Positional Attribution as the Missing Ingredient for Transformer Explainability
por: Bakish, Yarden, et al.
Publicado: (2025)
por: Bakish, Yarden, et al.
Publicado: (2025)
QuAnTS: Question Answering on Time Series
por: Divo, Felix, et al.
Publicado: (2025)
por: Divo, Felix, et al.
Publicado: (2025)
EmoLoom-2B: Fast Base-Model Screening for Emotion Classification and VAD with Lexicon-Weak Supervision and KV-Off Evaluation
por: Li, Zilin, et al.
Publicado: (2026)
por: Li, Zilin, et al.
Publicado: (2026)
Why LoRA Resists Label Noise: A Theoretical Framework for Noise-Robust Parameter-Efficient Fine-Tuning
por: Steele, Brady
Publicado: (2026)
por: Steele, Brady
Publicado: (2026)
TRiMS: Real-Time Tracking of Minimal Sufficient Length for Efficient Reasoning via RL
por: Bian, Tingcheng, et al.
Publicado: (2026)
por: Bian, Tingcheng, et al.
Publicado: (2026)
Planning vs Reasoning: Ablations to Test Capabilities of LoRA layers
por: Redkar, Neel
Publicado: (2024)
por: Redkar, Neel
Publicado: (2024)
Evaluating the Systematic Reasoning Abilities of Large Language Models through Graph Coloring
por: Heyman, Alex, et al.
Publicado: (2025)
por: Heyman, Alex, et al.
Publicado: (2025)
Ejemplares similares
-
Tempered Calculus for ML: Application to Hyperbolic Model Embedding
por: Nock, Richard, et al.
Publicado: (2024) -
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
por: Fadli, Samih
Publicado: (2025) -
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
por: Srishty, Sharmin Sultana, et al.
Publicado: (2026) -
Survey Transfer Learning: Recycling Data with Silicon Responses
por: Amini, Ali
Publicado: (2025) -
TensorLens: End-to-End Transformer Analysis via High-Order Attention Tensors
por: Atad, Ido Andrew, et al.
Publicado: (2026)