On the Rejection Criterion for Proxy-based Test-time Alignment
Fuente:
arXiv
Saved in:
| Main Authors: | Hammal, Ayoub, Zweigenbaum, Pierre, Corro, Caio |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Kad: A Framework for Proxy-based Test-time Alignment with Knapsack Approximation Deferral
by: Hammal, Ayoub, et al.
Published: (2025)
by: Hammal, Ayoub, et al.
Published: (2025)
Suffix-Constrained Greedy Search Algorithms for Causal Language Models
by: Hammal, Ayoub, et al.
Published: (2026)
by: Hammal, Ayoub, et al.
Published: (2026)
Few-Shot Domain Adaptation for Named-Entity Recognition via Joint Constrained k-Means and Subspace Selection
by: Hammal, Ayoub, et al.
Published: (2024)
by: Hammal, Ayoub, et al.
Published: (2024)
A fast and sound tagging method for discontinuous named-entity recognition
by: Corro, Caio
Published: (2024)
by: Corro, Caio
Published: (2024)
Am I eligible? Natural Language Inference for Clinical Trial Patient Recruitment: the Patient's Point of View
by: Aguiar, Mathilde, et al.
Published: (2025)
by: Aguiar, Mathilde, et al.
Published: (2025)
SEME at SemEval-2024 Task 2: Comparing Masked and Generative Language Models on Natural Language Inference for Clinical Trials
by: Aguiar, Mathilde, et al.
Published: (2024)
by: Aguiar, Mathilde, et al.
Published: (2024)
Sparse Logistic Regression with High-order Features for Automatic Grammar Rule Extraction from Treebanks
by: Herrera, Santiago, et al.
Published: (2024)
by: Herrera, Santiago, et al.
Published: (2024)
A Study on Building Efficient Zero-Shot Relation Extraction Models
by: Thomas, Hugo, et al.
Published: (2026)
by: Thomas, Hugo, et al.
Published: (2026)
DocPolarBERT: A Pre-trained Model for Document Understanding with Relative Polar Coordinate Encoding of Layout Structures
by: Uthayasooriyar, Benno, et al.
Published: (2025)
by: Uthayasooriyar, Benno, et al.
Published: (2025)
Training LayoutLM from Scratch for Efficient Named-Entity Recognition in the Insurance Domain
by: Uthayasooriyar, Benno, et al.
Published: (2024)
by: Uthayasooriyar, Benno, et al.
Published: (2024)
Bregman Conditional Random Fields: Sequence Labeling with Parallelizable Inference Algorithms
by: Corro, Caio, et al.
Published: (2025)
by: Corro, Caio, et al.
Published: (2025)
Is Clinical Text Enough? A Multimodal Study on Mortality Prediction in Heart Failure Patients
by: Khettari, Oumaima El, et al.
Published: (2026)
by: Khettari, Oumaima El, et al.
Published: (2026)
Decoupled Proxy Alignment: Mitigating Language Prior Conflict for Multimodal Alignment in MLLM
by: Tan, Chenkun, et al.
Published: (2025)
by: Tan, Chenkun, et al.
Published: (2025)
PARHAF, a human-authored corpus of clinical reports for fictitious patients in French
by: Tannier, Xavier, et al.
Published: (2026)
by: Tannier, Xavier, et al.
Published: (2026)
EuroBERT: Scaling Multilingual Encoders for European Languages
by: Boizard, Nicolas, et al.
Published: (2025)
by: Boizard, Nicolas, et al.
Published: (2025)
Proxy-RLHF: Decoupling Generation and Alignment in Large Language Model with Proxy
by: Zhu, Yu, et al.
Published: (2024)
by: Zhu, Yu, et al.
Published: (2024)
Nested Named Entity Recognition as Single-Pass Sequence Labeling
by: Muñoz-Ortiz, Alberto, et al.
Published: (2025)
by: Muñoz-Ortiz, Alberto, et al.
Published: (2025)
Distributional Alignment as a Criterion for Designing Task Vectors in In-Context Learning
by: Kwon, Jihoon, et al.
Published: (2026)
by: Kwon, Jihoon, et al.
Published: (2026)
The Greatest Good Benchmark: Measuring LLMs' Alignment with Utilitarian Moral Dilemmas
by: Marraffini, Giovanni Franco Gabriel, et al.
Published: (2025)
by: Marraffini, Giovanni Franco Gabriel, et al.
Published: (2025)
Rethinking the Role of Proxy Rewards in Language Model Alignment
by: Kim, Sungdong, et al.
Published: (2024)
by: Kim, Sungdong, et al.
Published: (2024)
Omnimodal Dataset Distillation via High-order Proxy Alignment
by: Gao, Yuxuan, et al.
Published: (2026)
by: Gao, Yuxuan, et al.
Published: (2026)
A BERTology View of LLM Orchestrations: Token- and Layer-Selective Probes for Efficient Single-Pass Classification
by: Meyoyan, Gonzalo Ariel, et al.
Published: (2026)
by: Meyoyan, Gonzalo Ariel, et al.
Published: (2026)
Entropy Sentinel: Continuous LLM Accuracy Monitoring from Decoding Entropy Traces in STEM
by: Buffa, Pedro Memoli, et al.
Published: (2026)
by: Buffa, Pedro Memoli, et al.
Published: (2026)
Token-level Accept or Reject: A Micro Alignment Approach for Large Language Models
by: Zhang, Yang, et al.
Published: (2025)
by: Zhang, Yang, et al.
Published: (2025)
Reasoning over Boundaries: Enhancing Specification Alignment via Test-time Deliberation
by: Zhang, Haoran, et al.
Published: (2025)
by: Zhang, Haoran, et al.
Published: (2025)
SaulLM-7B: A pioneering Large Language Model for Law
by: Colombo, Pierre, et al.
Published: (2024)
by: Colombo, Pierre, et al.
Published: (2024)
$k$NNProxy: Efficient Training-Free Proxy Alignment for Black-Box Zero-Shot LLM-Generated Text Detection
by: Wong, Kahim, et al.
Published: (2026)
by: Wong, Kahim, et al.
Published: (2026)
ProxyLM: Predicting Language Model Performance on Multilingual Tasks via Proxy Models
by: Anugraha, David, et al.
Published: (2024)
by: Anugraha, David, et al.
Published: (2024)
GenARM: Reward Guided Generation with Autoregressive Reward Model for Test-time Alignment
by: Xu, Yuancheng, et al.
Published: (2024)
by: Xu, Yuancheng, et al.
Published: (2024)
Tuning Language Models by Proxy
by: Liu, Alisa, et al.
Published: (2024)
by: Liu, Alisa, et al.
Published: (2024)
ProxyThinker: Test-Time Guidance through Small Visual Reasoners
by: Xiao, Zilin, et al.
Published: (2025)
by: Xiao, Zilin, et al.
Published: (2025)
Autoregressive vs. Masked Diffusion Language Models: A Controlled Comparison
by: Vicentino, Caio
Published: (2026)
by: Vicentino, Caio
Published: (2026)
Reasons to Reject? Aligning Language Models with Judgments
by: Xu, Weiwen, et al.
Published: (2023)
by: Xu, Weiwen, et al.
Published: (2023)
Statistical Rejection Sampling Improves Preference Optimization
by: Liu, Tianqi, et al.
Published: (2023)
by: Liu, Tianqi, et al.
Published: (2023)
Safety Arithmetic: A Framework for Test-time Safety Alignment of Language Models by Steering Parameters and Activations
by: Hazra, Rima, et al.
Published: (2024)
by: Hazra, Rima, et al.
Published: (2024)
"Mirror" Language AI Models of Depression are Criterion-Contaminated
by: Li, Tong, et al.
Published: (2025)
by: Li, Tong, et al.
Published: (2025)
Fast Best-of-N Decoding via Speculative Rejection
by: Sun, Hanshi, et al.
Published: (2024)
by: Sun, Hanshi, et al.
Published: (2024)
Proxy Compression for Language Modeling
by: Zheng, Lin, et al.
Published: (2026)
by: Zheng, Lin, et al.
Published: (2026)
TARo: Token-level Adaptive Routing for LLM Test-time Alignment
by: Rai, Arushi, et al.
Published: (2026)
by: Rai, Arushi, et al.
Published: (2026)
Beyond Rejection Sampling: Trajectory Fusion for Scaling Mathematical Reasoning
by: Deng, Jie, et al.
Published: (2026)
by: Deng, Jie, et al.
Published: (2026)
Similar Items
-
Kad: A Framework for Proxy-based Test-time Alignment with Knapsack Approximation Deferral
by: Hammal, Ayoub, et al.
Published: (2025) -
Suffix-Constrained Greedy Search Algorithms for Causal Language Models
by: Hammal, Ayoub, et al.
Published: (2026) -
Few-Shot Domain Adaptation for Named-Entity Recognition via Joint Constrained k-Means and Subspace Selection
by: Hammal, Ayoub, et al.
Published: (2024) -
A fast and sound tagging method for discontinuous named-entity recognition
by: Corro, Caio
Published: (2024) -
Am I eligible? Natural Language Inference for Clinical Trial Patient Recruitment: the Patient's Point of View
by: Aguiar, Mathilde, et al.
Published: (2025)