Flipping Against All Odds: Reducing LLM Coin Flip Bias via Verbalized Rejection Sampling
Fuente:
arXiv
Saved in:
| Main Authors: | Xiao, Tim Z., Zenn, Johannes, Liu, Zhen, Liu, Weiyang, Bamler, Robert, Schölkopf, Bernhard |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Verbalized Machine Learning: Revisiting Machine Learning with Language Models
by: Xiao, Tim Z., et al.
Published: (2024)
by: Xiao, Tim Z., et al.
Published: (2024)
Differentiable Annealed Importance Sampling Minimizes The Symmetrized Kullback-Leibler Divergence Between Initial and Target Distribution
by: Zenn, Johannes, et al.
Published: (2024)
by: Zenn, Johannes, et al.
Published: (2024)
A Note on Generalization in Variational Autoencoders: How Effective Is Synthetic Data & Overparameterization?
by: Xiao, Tim Z., et al.
Published: (2023)
by: Xiao, Tim Z., et al.
Published: (2023)
Reparameterized LLM Training via Orthogonal Equivalence Transformation
by: Qiu, Zeju, et al.
Published: (2025)
by: Qiu, Zeju, et al.
Published: (2025)
A Coin Flip for Safety: LLM Judges Fail to Reliably Measure Adversarial Robustness
by: Schwinn, Leo, et al.
Published: (2026)
by: Schwinn, Leo, et al.
Published: (2026)
Your Finetuned Large Language Model is Already a Powerful Out-of-distribution Detector
by: Zhang, Andi, et al.
Published: (2024)
by: Zhang, Andi, et al.
Published: (2024)
A Compact Representation for Bayesian Neural Networks By Removing Permutation Symmetry
by: Xiao, Tim Z., et al.
Published: (2023)
by: Xiao, Tim Z., et al.
Published: (2023)
Enough Coin Flips Can Make LLMs Act Bayesian
by: Gupta, Ritwik, et al.
Published: (2025)
by: Gupta, Ritwik, et al.
Published: (2025)
Rethinking Easy-to-Hard: Limits of Curriculum Learning in Post-Training for Deductive Reasoning
by: Mordig, Maximilian, et al.
Published: (2026)
by: Mordig, Maximilian, et al.
Published: (2026)
Prompt Optimization Is a Coin Flip: Diagnosing When It Helps in Compound AI Systems
by: Zhang, Xing, et al.
Published: (2026)
by: Zhang, Xing, et al.
Published: (2026)
Balancing Molecular Information and Empirical Data in the Prediction of Physico-Chemical Properties
by: Zenn, Johannes, et al.
Published: (2024)
by: Zenn, Johannes, et al.
Published: (2024)
Orthogonal Finetuning Made Scalable
by: Qiu, Zeju, et al.
Published: (2025)
by: Qiu, Zeju, et al.
Published: (2025)
Against All Odds: Overcoming Typology, Script, and Language Confusion in Multilingual Embedding Inversion Attacks
by: Chen, Yiyi, et al.
Published: (2024)
by: Chen, Yiyi, et al.
Published: (2024)
Protocols for Quantum Weak Coin Flipping
by: Arora, Atul Singh, et al.
Published: (2024)
by: Arora, Atul Singh, et al.
Published: (2024)
Ready to Flip a Coin … Twice?
Published: (2025)
Published: (2025)
Cheat-Penalised Quantum Weak Coin-Flipping
by: Arora, Atul Singh, et al.
Published: (2025)
by: Arora, Atul Singh, et al.
Published: (2025)
AdvJudge-Zero: Binary Decision Flips in LLM-as-a-Judge via Adversarial Control Tokens
by: Li, Tung-Ling, et al.
Published: (2025)
by: Li, Tung-Ling, et al.
Published: (2025)
Coin-Flipping In The Brain: Statistical Learning with Neuronal Assemblies
by: Dabagia, Max, et al.
Published: (2024)
by: Dabagia, Max, et al.
Published: (2024)
Can Large Language Models Understand Symbolic Graphics Programs?
by: Qiu, Zeju, et al.
Published: (2024)
by: Qiu, Zeju, et al.
Published: (2024)
The Flip Side of RLHF: On-Policy Feedback for Reward Model Self-Supervised Improvement
by: Wang, Xiaobo, et al.
Published: (2026)
by: Wang, Xiaobo, et al.
Published: (2026)
FlipAttack: Jailbreak LLMs via Flipping
by: Liu, Yue, et al.
Published: (2024)
by: Liu, Yue, et al.
Published: (2024)
PEFT-Arena: Understanding Parameter-Efficient Finetuning from a Stability-Plasticity Perspective
by: Huang, Yangyi, et al.
Published: (2026)
by: Huang, Yangyi, et al.
Published: (2026)
Flipping the Dialogue: Training and Evaluating User Language Models
by: Naous, Tarek, et al.
Published: (2025)
by: Naous, Tarek, et al.
Published: (2025)
Exploring the Jungle of Bias: Political Bias Attribution in Language Models via Dependency Analysis
by: Jenny, David F., et al.
Published: (2023)
by: Jenny, David F., et al.
Published: (2023)
Improved Bounds for Coin Flipping, Leader Election, and Random Selection
by: Chattopadhyay, Eshan, et al.
Published: (2025)
by: Chattopadhyay, Eshan, et al.
Published: (2025)
Truth or Twist? Optimal Model Selection for Reliable Label Flipping Evaluation in LLM-based Counterfactuals
by: Wang, Qianli, et al.
Published: (2025)
by: Wang, Qianli, et al.
Published: (2025)
Detecting Localized Density Anomalies in Multivariate Data via Coin-Flip Statistics
by: Springer, Sebastian, et al.
Published: (2025)
by: Springer, Sebastian, et al.
Published: (2025)
Against All Odds: Refugees Coping in a Strange Land.
by: Mason, Elisa
Published: (1999)
by: Mason, Elisa
Published: (1999)
FlipGuard: Defending Preference Alignment against Update Regression with Constrained Optimization
by: Zhu, Mingye, et al.
Published: (2024)
by: Zhu, Mingye, et al.
Published: (2024)
Statistical Rejection Sampling Improves Preference Optimization
by: Liu, Tianqi, et al.
Published: (2023)
by: Liu, Tianqi, et al.
Published: (2023)
How Robust Are Router-LLMs? Analysis of the Fragility of LLM Routing Capabilities
by: Kassem, Aly M., et al.
Published: (2025)
by: Kassem, Aly M., et al.
Published: (2025)
A diverse Multilingual News Headlines Dataset from around the World
by: Leeb, Felix, et al.
Published: (2024)
by: Leeb, Felix, et al.
Published: (2024)
How Random is Random? Evaluating the Randomness and Humaness of LLMs' Coin Flips
by: Van Koevering, Katherine, et al.
Published: (2024)
by: Van Koevering, Katherine, et al.
Published: (2024)
Bottom-up Rebalancing Binary Search Trees by Flipping a Coin
by: Brodal, Gerth Stølting
Published: (2024)
by: Brodal, Gerth Stølting
Published: (2024)
Anticoagulation Monitoring During ECMO Support: Monitor or Flip a Coin?
by: Sasa Rajsic, et al.
Published: (2024)
by: Sasa Rajsic, et al.
Published: (2024)
What Are the Odds? Language Models Are Capable of Probabilistic Reasoning
by: Paruchuri, Akshay, et al.
Published: (2024)
by: Paruchuri, Akshay, et al.
Published: (2024)
Permute-and-Flip: An optimally stable and watermarkable decoder for LLMs
by: Zhao, Xuandong, et al.
Published: (2024)
by: Zhao, Xuandong, et al.
Published: (2024)
iFlip: Iterative Feedback-driven Counterfactual Example Refinement
by: Wang, Yilong, et al.
Published: (2026)
by: Wang, Yilong, et al.
Published: (2026)
Learning Analytics from Spoken Discussion Dialogs in Flipped Classroom
by: Su, Hang, et al.
Published: (2023)
by: Su, Hang, et al.
Published: (2023)
Reproducing HotFlip for Corpus Poisoning Attacks in Dense Retrieval
by: Li, Yongkang, et al.
Published: (2025)
by: Li, Yongkang, et al.
Published: (2025)
Similar Items
-
Verbalized Machine Learning: Revisiting Machine Learning with Language Models
by: Xiao, Tim Z., et al.
Published: (2024) -
Differentiable Annealed Importance Sampling Minimizes The Symmetrized Kullback-Leibler Divergence Between Initial and Target Distribution
by: Zenn, Johannes, et al.
Published: (2024) -
A Note on Generalization in Variational Autoencoders: How Effective Is Synthetic Data & Overparameterization?
by: Xiao, Tim Z., et al.
Published: (2023) -
Reparameterized LLM Training via Orthogonal Equivalence Transformation
by: Qiu, Zeju, et al.
Published: (2025) -
A Coin Flip for Safety: LLM Judges Fail to Reliably Measure Adversarial Robustness
by: Schwinn, Leo, et al.
Published: (2026)