Understanding the Logic of Direct Preference Alignment through Logic
Fuente:
arXiv
Saved in:
| Main Authors: | Richardson, Kyle, Srikumar, Vivek, Sabharwal, Ashish |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
by: Lin, Bill Yuchen, et al.
Published: (2025)
by: Lin, Bill Yuchen, et al.
Published: (2025)
A Logic for Expressing Log-Precision Transformers
by: Merrill, William, et al.
Published: (2022)
by: Merrill, William, et al.
Published: (2022)
Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation
by: Gupta, Ashim, et al.
Published: (2025)
by: Gupta, Ashim, et al.
Published: (2025)
Reinforcing Code Generation: Improving Text-to-SQL with Execution-Based Learning
by: Kulkarni, Atharv, et al.
Published: (2025)
by: Kulkarni, Atharv, et al.
Published: (2025)
InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis
by: Bentham, Oliver, et al.
Published: (2026)
by: Bentham, Oliver, et al.
Published: (2026)
The Expressive Power of Transformers with Chain of Thought
by: Merrill, William, et al.
Published: (2023)
by: Merrill, William, et al.
Published: (2023)
On the Reasoning Abilities of Masked Diffusion Language Models
by: Svete, Anej, et al.
Published: (2025)
by: Svete, Anej, et al.
Published: (2025)
Reason from Fallacy: Enhancing Large Language Models' Logical Reasoning through Logical Fallacy Understanding
by: Li, Yanda, et al.
Published: (2024)
by: Li, Yanda, et al.
Published: (2024)
Promptly Predicting Structures: The Return of Inference
by: Mehta, Maitrey, et al.
Published: (2024)
by: Mehta, Maitrey, et al.
Published: (2024)
The Illusion of State in State-Space Models
by: Merrill, William, et al.
Published: (2024)
by: Merrill, William, et al.
Published: (2024)
Hierarchical Alignment: Enforcing Hierarchical Instruction-Following in LLMs through Logical Consistency
by: Yang, Shu, et al.
Published: (2026)
by: Yang, Shu, et al.
Published: (2026)
Logical Phase Transitions: Understanding Collapse in LLM Logical Reasoning
by: Zhang, Xinglang, et al.
Published: (2026)
by: Zhang, Xinglang, et al.
Published: (2026)
Exact Expressive Power of Transformers with Padding
by: Merrill, William, et al.
Published: (2025)
by: Merrill, William, et al.
Published: (2025)
Distilling Formal Logic into Neural Spaces: A Kernel Alignment Approach for Signal Temporal Logic
by: Candussio, Sara, et al.
Published: (2026)
by: Candussio, Sara, et al.
Published: (2026)
Aligning with Logic: Measuring, Evaluating and Improving Logical Preference Consistency in Large Language Models
by: Liu, Yinhong, et al.
Published: (2024)
by: Liu, Yinhong, et al.
Published: (2024)
Answer, Assemble, Ace: Understanding How LMs Answer Multiple Choice Questions
by: Wiegreffe, Sarah, et al.
Published: (2024)
by: Wiegreffe, Sarah, et al.
Published: (2024)
The Role of Logic and Automata in Understanding Transformers
by: Lin, Anthony W., et al.
Published: (2025)
by: Lin, Anthony W., et al.
Published: (2025)
LLM-Symbolic Integration for Robust Temporal Tabular Reasoning
by: Kulkarni, Atharv, et al.
Published: (2025)
by: Kulkarni, Atharv, et al.
Published: (2025)
SUPER: Evaluating Agents on Setting Up and Executing Tasks from Research Repositories
by: Bogin, Ben, et al.
Published: (2024)
by: Bogin, Ben, et al.
Published: (2024)
Discovering a Shared Logical Subspace: Steering LLM Logical Reasoning via Alignment of Natural-Language and Symbolic Views
by: Fang, Feihao, et al.
Published: (2026)
by: Fang, Feihao, et al.
Published: (2026)
Language Model Planners do not Scale, but do Formalizers?
by: Jiang, Owen, et al.
Published: (2026)
by: Jiang, Owen, et al.
Published: (2026)
Leveraging Code to Improve In-context Learning for Semantic Parsing
by: Bogin, Ben, et al.
Published: (2023)
by: Bogin, Ben, et al.
Published: (2023)
In-Context Example Ordering Guided by Label Distributions
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
Found in Translation: Measuring Multilingual LLM Consistency as Simple as Translate then Evaluate
by: Gupta, Ashim, et al.
Published: (2025)
by: Gupta, Ashim, et al.
Published: (2025)
PolicyLR: A Logic Representation For Privacy Policies
by: Hooda, Ashish, et al.
Published: (2024)
by: Hooda, Ashish, et al.
Published: (2024)
RLearner-LLM: Balancing Logical Grounding and Fluency in Large Language Models via Hybrid Direct Preference Optimization
by: Bao, Qiming, et al.
Published: (2026)
by: Bao, Qiming, et al.
Published: (2026)
Revisiting Padded Transformer Expressivity: Which Architectural Choices Matter and Which Don't
by: Svete, Anej, et al.
Published: (2026)
by: Svete, Anej, et al.
Published: (2026)
Robust Preference Alignment via Directional Neighborhood Consensus
by: Mao, Ruochen, et al.
Published: (2025)
by: Mao, Ruochen, et al.
Published: (2025)
An Empirical Investigation of Matrix Factorization Methods for Pre-trained Transformers
by: Gupta, Ashim, et al.
Published: (2024)
by: Gupta, Ashim, et al.
Published: (2024)
Pronoun Logic
by: Bohrer, Rose, et al.
Published: (2024)
by: Bohrer, Rose, et al.
Published: (2024)
Logic of Montage
by: Takahashi, Hayami, et al.
Published: (2025)
by: Takahashi, Hayami, et al.
Published: (2025)
State Space Models are Strong Text Rerankers
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
Distillation versus Contrastive Learning: How to Train Your Rerankers
by: Xu, Zhichao, et al.
Published: (2025)
by: Xu, Zhichao, et al.
Published: (2025)
Making Implicit Premises Explicit in Logical Understanding of Enthymemes
by: Feng, Xuyao, et al.
Published: (2026)
by: Feng, Xuyao, et al.
Published: (2026)
Beyond Perplexity: Multi-dimensional Safety Evaluation of LLM Compression
by: Xu, Zhichao, et al.
Published: (2024)
by: Xu, Zhichao, et al.
Published: (2024)
Whispers of Doubt Amidst Echoes of Triumph in NLP Robustness
by: Gupta, Ashim, et al.
Published: (2023)
by: Gupta, Ashim, et al.
Published: (2023)
Defragmenting Language Models: An Interpretability-based Approach for Vocabulary Expansion
by: Mehta, Maitrey, et al.
Published: (2026)
by: Mehta, Maitrey, et al.
Published: (2026)
Boosting Logical Fallacy Reasoning in LLMs via Logical Structure Tree
by: Lei, Yuanyuan, et al.
Published: (2024)
by: Lei, Yuanyuan, et al.
Published: (2024)
Integrating Physician Diagnostic Logic into Large Language Models: Preference Learning from Process Feedback
by: Dou, Chengfeng, et al.
Published: (2024)
by: Dou, Chengfeng, et al.
Published: (2024)
Why Are Linear RNNs More Parallelizable?
by: Merrill, William, et al.
Published: (2026)
by: Merrill, William, et al.
Published: (2026)
Similar Items
-
ZebraLogic: On the Scaling Limits of LLMs for Logical Reasoning
by: Lin, Bill Yuchen, et al.
Published: (2025) -
A Logic for Expressing Log-Precision Transformers
by: Merrill, William, et al.
Published: (2022) -
Test-Time Scaling with Repeated Sampling Improves Multilingual Text Generation
by: Gupta, Ashim, et al.
Published: (2025) -
Reinforcing Code Generation: Improving Text-to-SQL with Execution-Based Learning
by: Kulkarni, Atharv, et al.
Published: (2025) -
InfiniteScienceGym: An Unbounded, Procedurally-Generated Benchmark for Scientific Analysis
by: Bentham, Oliver, et al.
Published: (2026)