The Sign Estimator: LLM Alignment in the Face of Choice Heterogeneity
Fuente:
arXiv
Saved in:
| Main Authors: | Aouad, Ali, Gadarri, Aymane El, Farias, Vivek F. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Value of Covariance Matching in Gaussian DDPMs and the Lanczos Sampler
by: Akhtar, Md Sahil, et al.
Published: (2026)
by: Akhtar, Md Sahil, et al.
Published: (2026)
High-dimensional Learning with Noisy Labels
by: Firdoussi, Aymane El, et al.
Published: (2024)
by: Firdoussi, Aymane El, et al.
Published: (2024)
$α$-LoRA: Effective Fine-Tuning via Base Model Rescaling
by: Firdoussi, Aymane El, et al.
Published: (2025)
by: Firdoussi, Aymane El, et al.
Published: (2025)
Self-Normalized Resets for Plasticity in Continual Learning
by: Farias, Vivek F., et al.
Published: (2024)
by: Farias, Vivek F., et al.
Published: (2024)
Beyond Pairwise: Empowering LLM Alignment With Ranked Choice Modeling
by: Tang, Yuxuan, et al.
Published: (2025)
by: Tang, Yuxuan, et al.
Published: (2025)
Direct Alignment with Heterogeneous Preferences
by: Shirali, Ali, et al.
Published: (2025)
by: Shirali, Ali, et al.
Published: (2025)
A foundation for exact binarized morphological neural networks
by: Aouad, Theodore, et al.
Published: (2024)
by: Aouad, Theodore, et al.
Published: (2024)
Feature Distillation is the Better Choice for Model-Heterogeneous Federated Learning
by: Li, Yichen, et al.
Published: (2025)
by: Li, Yichen, et al.
Published: (2025)
Open Problems in Differentiable Social Choice: Learning Mechanisms, Decisions, and Alignment
by: An, Zhiyu, et al.
Published: (2026)
by: An, Zhiyu, et al.
Published: (2026)
LLM Safety Alignment is Divergence Estimation in Disguise
by: Haldar, Rajdeep, et al.
Published: (2025)
by: Haldar, Rajdeep, et al.
Published: (2025)
Maximizing the Potential of Synthetic Data: Insights from Random Matrix Theory
by: Firdoussi, Aymane El, et al.
Published: (2024)
by: Firdoussi, Aymane El, et al.
Published: (2024)
Policy Optimization for Personalized Interventions in Behavioral Health
by: Baek, Jackie, et al.
Published: (2023)
by: Baek, Jackie, et al.
Published: (2023)
Tokenized Bandit for LLM Decoding and Alignment
by: Shin, Suho, et al.
Published: (2025)
by: Shin, Suho, et al.
Published: (2025)
Alignment Dynamics in LLM Fine-Tuning
by: Huang, Yuhan, et al.
Published: (2026)
by: Huang, Yuhan, et al.
Published: (2026)
Aggregation Alignment for Federated Learning with Mixture-of-Experts under Data Heterogeneity
by: Fang, Zihan, et al.
Published: (2026)
by: Fang, Zihan, et al.
Published: (2026)
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
by: Chen, Tiejin, et al.
Published: (2026)
by: Chen, Tiejin, et al.
Published: (2026)
Adversarial Preference Learning for Robust LLM Alignment
by: Wang, Yuanfu, et al.
Published: (2025)
by: Wang, Yuanfu, et al.
Published: (2025)
A Small Math Model: Recasting Strategy Choice Theory in an LLM-Inspired Architecture
by: Rahman, Roussel, et al.
Published: (2025)
by: Rahman, Roussel, et al.
Published: (2025)
Echo: Decoupling Inference and Training for Large-Scale RL Alignment on Heterogeneous Swarms
by: Xiao, Jie, et al.
Published: (2025)
by: Xiao, Jie, et al.
Published: (2025)
Rethinking LoRA for Data Heterogeneous Federated Learning: Subspace and State Alignment
by: Peng, Hongyi, et al.
Published: (2026)
by: Peng, Hongyi, et al.
Published: (2026)
ECLIPTICA -- A Framework for Switchable LLM Alignment via CITA - Contrastive Instruction-Tuned Alignment
by: Wanaskar, Kapil, et al.
Published: (2026)
by: Wanaskar, Kapil, et al.
Published: (2026)
PIPA: Preference Alignment as Prior-Informed Statistical Estimation
by: Li, Junbo, et al.
Published: (2025)
by: Li, Junbo, et al.
Published: (2025)
The Alignment Game: A Theory of Long-Horizon Alignment Through Recursive Curation
by: Falahati, Ali, et al.
Published: (2025)
by: Falahati, Ali, et al.
Published: (2025)
GRADE: Replacing Policy Gradients with Backpropagation for LLM Alignment
by: Nel, Lukas Abrie
Published: (2025)
by: Nel, Lukas Abrie
Published: (2025)
Training LLM Agents to Empower Humans
by: Ellis, Evan, et al.
Published: (2025)
by: Ellis, Evan, et al.
Published: (2025)
Adaptive Sparse Allocation with Mutual Choice & Feature Choice Sparse Autoencoders
by: Ayonrinde, Kola
Published: (2024)
by: Ayonrinde, Kola
Published: (2024)
Generalised Probabilistic Modelling and Improved Uncertainty Estimation in Comparative LLM-as-a-judge
by: Fathullah, Yassir, et al.
Published: (2025)
by: Fathullah, Yassir, et al.
Published: (2025)
Cost-Minimized Label-Flipping Poisoning Attack to LLM Alignment
by: Kusaka, Shigeki, et al.
Published: (2025)
by: Kusaka, Shigeki, et al.
Published: (2025)
Find A Winning Sign: Sign Is All We Need to Win the Lottery
by: Oh, Junghun, et al.
Published: (2025)
by: Oh, Junghun, et al.
Published: (2025)
Early Detection of Pancreatic Cancer Using Multimodal Learning on Electronic Health Records
by: Aouad, Mosbah, et al.
Published: (2025)
by: Aouad, Mosbah, et al.
Published: (2025)
Data-Driven Estimation of Heterogeneous Treatment Effects
by: Tran, Christopher, et al.
Published: (2023)
by: Tran, Christopher, et al.
Published: (2023)
Align-Pro: A Principled Approach to Prompt Optimization for LLM Alignment
by: Trivedi, Prashant, et al.
Published: (2025)
by: Trivedi, Prashant, et al.
Published: (2025)
Robust LLM Alignment via Distributionally Robust Direct Preference Optimization
by: Xu, Zaiyan, et al.
Published: (2025)
by: Xu, Zaiyan, et al.
Published: (2025)
Entropy-Guided Dynamic Tokens for Graph-LLM Alignment in Molecular Understanding
by: Jing, Zihao, et al.
Published: (2026)
by: Jing, Zihao, et al.
Published: (2026)
Intra-Fairness Dynamics: The Bias Spillover Effect in Targeted LLM Alignment
by: Paraschou, Eva, et al.
Published: (2026)
by: Paraschou, Eva, et al.
Published: (2026)
Safeguarding LLM Fine-tuning via Push-Pull Distributional Alignment
by: Wang, Haozhong, et al.
Published: (2026)
by: Wang, Haozhong, et al.
Published: (2026)
Style Outweighs Substance: Failure Modes of LLM Judges in Alignment Benchmarking
by: Feuer, Benjamin, et al.
Published: (2024)
by: Feuer, Benjamin, et al.
Published: (2024)
Moral Alignment for LLM Agents
by: Tennant, Elizaveta, et al.
Published: (2024)
by: Tennant, Elizaveta, et al.
Published: (2024)
Certified Signed Graph Unlearning
by: Zhao, Junpeng, et al.
Published: (2025)
by: Zhao, Junpeng, et al.
Published: (2025)
Toward Practical Entity Alignment Method Design: Insights from New Highly Heterogeneous Knowledge Graph Datasets
by: Jiang, Xuhui, et al.
Published: (2023)
by: Jiang, Xuhui, et al.
Published: (2023)
Similar Items
-
The Value of Covariance Matching in Gaussian DDPMs and the Lanczos Sampler
by: Akhtar, Md Sahil, et al.
Published: (2026) -
High-dimensional Learning with Noisy Labels
by: Firdoussi, Aymane El, et al.
Published: (2024) -
$α$-LoRA: Effective Fine-Tuning via Base Model Rescaling
by: Firdoussi, Aymane El, et al.
Published: (2025) -
Self-Normalized Resets for Plasticity in Continual Learning
by: Farias, Vivek F., et al.
Published: (2024) -
Beyond Pairwise: Empowering LLM Alignment With Ranked Choice Modeling
by: Tang, Yuxuan, et al.
Published: (2025)