Eliminating Inductive Bias in Reward Models with Information-Theoretic Guidance
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Zhuo, Cheng, Pengyu, Yu, Zhechao, Tong, Feifei, Gao, Anningzhe, Chang, Tsung-Hui, Wan, Xiang, Zhao, Erchao, Jiang, Xiaoxi, Jiang, Guanjun |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RLHF in an SFT Way: From Optimal Solution to Reward-Weighted Alignment
von: Du, Yuhao, et al.
Veröffentlicht: (2025)
von: Du, Yuhao, et al.
Veröffentlicht: (2025)
Atoxia: Red-teaming Large Language Models with Target Toxic Answers
von: Du, Yuhao, et al.
Veröffentlicht: (2024)
von: Du, Yuhao, et al.
Veröffentlicht: (2024)
CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR
von: Cui, Sijia, et al.
Veröffentlicht: (2026)
von: Cui, Sijia, et al.
Veröffentlicht: (2026)
Rationale Matters: Learning Transferable Rubrics via Proxy-Guided Critique for VLM Reward Models
von: Qiu, Weijie, et al.
Veröffentlicht: (2026)
von: Qiu, Weijie, et al.
Veröffentlicht: (2026)
MARCH: Multi-Agent Reinforced Self-Check for LLM Hallucination
von: Li, Zhuo, et al.
Veröffentlicht: (2026)
von: Li, Zhuo, et al.
Veröffentlicht: (2026)
Trace2Skill: Distill Trajectory-Local Lessons into Transferable Agent Skills
von: Ni, Jingwei, et al.
Veröffentlicht: (2026)
von: Ni, Jingwei, et al.
Veröffentlicht: (2026)
Grounding the Score: Explicit Visual Premise Verification for Reliable Vision-Language Process Reward Models
von: Wang, Junxin, et al.
Veröffentlicht: (2026)
von: Wang, Junxin, et al.
Veröffentlicht: (2026)
APLOT: Robust Reward Modeling via Adaptive Preference Learning with Optimal Transport
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
Theoretical Investigation on Inductive Bias of Isolation Forest
von: Zheng, Qin-Cheng, et al.
Veröffentlicht: (2025)
von: Zheng, Qin-Cheng, et al.
Veröffentlicht: (2025)
Self-Instructed Derived Prompt Generation Meets In-Context Learning: Unlocking New Potential of Black-Box LLMs
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
von: Li, Zhuo, et al.
Veröffentlicht: (2024)
ATP-Bench: Towards Agentic Tool Planning for MLLM Interleaved Generation
von: Liu, Yinuo, et al.
Veröffentlicht: (2026)
von: Liu, Yinuo, et al.
Veröffentlicht: (2026)
Writing-Zero: Bridge the Gap Between Non-verifiable Tasks and Verifiable Rewards
von: Jia, Ruipeng, et al.
Veröffentlicht: (2025)
von: Jia, Ruipeng, et al.
Veröffentlicht: (2025)
Bringing Stability to Diffusion: Decomposing and Reducing Variance of Training Masked Diffusion Models
von: Jia, Mengni, et al.
Veröffentlicht: (2025)
von: Jia, Mengni, et al.
Veröffentlicht: (2025)
Aligning Language Models Using Follow-up Likelihood as Reward Signal
von: Zhang, Chen, et al.
Veröffentlicht: (2024)
von: Zhang, Chen, et al.
Veröffentlicht: (2024)
Search Self-play: Pushing the Frontier of Agent Capability without Supervision
von: Lu, Hongliang, et al.
Veröffentlicht: (2025)
von: Lu, Hongliang, et al.
Veröffentlicht: (2025)
RoTHP: Rotary Position Embedding-based Transformer Hawkes Process
von: Gao, Anningzhe, et al.
Veröffentlicht: (2024)
von: Gao, Anningzhe, et al.
Veröffentlicht: (2024)
Add-One-In: Incremental Sample Selection for Large Language Models via a Choice-Based Greedy Paradigm
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
von: Li, Zhuo, et al.
Veröffentlicht: (2025)
Symplectic Inductive Bias for Data-Driven Target Reachability in Hamiltonian Systems
von: Ouyang, Zhuo, et al.
Veröffentlicht: (2026)
von: Ouyang, Zhuo, et al.
Veröffentlicht: (2026)
Perceptual Inductive Bias Is What You Need Before Contrastive Learning
von: Li, Tianqin, et al.
Veröffentlicht: (2025)
von: Li, Tianqin, et al.
Veröffentlicht: (2025)
Selective hydrodeoxygenation of lignin to 4‐ethylcyclohexanol catalyzed by Cu–Ni/MgCrOx spinel
von: Lixia Li, et al.
Veröffentlicht: (2025)
von: Lixia Li, et al.
Veröffentlicht: (2025)
LLMs Could Autonomously Learn Without External Supervision
von: Ji, Ke, et al.
Veröffentlicht: (2024)
von: Ji, Ke, et al.
Veröffentlicht: (2024)
Disentangling Granularity: An Implicit Inductive Bias in Factorized VAEs
von: Chen, Zihao, et al.
Veröffentlicht: (2025)
von: Chen, Zihao, et al.
Veröffentlicht: (2025)
InvDiff: Invariant Guidance for Bias Mitigation in Diffusion Models
von: Hou, Min, et al.
Veröffentlicht: (2024)
von: Hou, Min, et al.
Veröffentlicht: (2024)
Mamba Hawkes Process
von: Gao, Anningzhe, et al.
Veröffentlicht: (2024)
von: Gao, Anningzhe, et al.
Veröffentlicht: (2024)
OVM, Outcome-supervised Value Models for Planning in Mathematical Reasoning
von: Yu, Fei, et al.
Veröffentlicht: (2023)
von: Yu, Fei, et al.
Veröffentlicht: (2023)
Information Locality as an Inductive Bias for Neural Language Models
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
von: Someya, Taiga, et al.
Veröffentlicht: (2025)
Do We Always Need the Simplicity Bias? Looking for Optimal Inductive Biases in the Wild
von: Teney, Damien, et al.
Veröffentlicht: (2025)
von: Teney, Damien, et al.
Veröffentlicht: (2025)
From Pixels to Gigapixels: Bridging Local Inductive Bias and Long-Range Dependencies with Pixel-Mamba
von: Qiu, Zhongwei, et al.
Veröffentlicht: (2024)
von: Qiu, Zhongwei, et al.
Veröffentlicht: (2024)
EBaReT: Expert-guided Bag Reward Transformer for Auto Bidding
von: Li, Kaiyuan, et al.
Veröffentlicht: (2025)
von: Li, Kaiyuan, et al.
Veröffentlicht: (2025)
CoD, Towards an Interpretable Medical Agent using Chain of Diagnosis
von: Chen, Junying, et al.
Veröffentlicht: (2024)
von: Chen, Junying, et al.
Veröffentlicht: (2024)
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
von: Zhang, Feng, et al.
Veröffentlicht: (2026)
von: Zhang, Feng, et al.
Veröffentlicht: (2026)
SEGNO: Generalizing Equivariant Graph Neural Networks with Physical Inductive Biases
von: Liu, Yang, et al.
Veröffentlicht: (2023)
von: Liu, Yang, et al.
Veröffentlicht: (2023)
Unsupervised Mutual Learning of Discourse Parsing and Topic Segmentation in Dialogue
von: Xu, Jiahui, et al.
Veröffentlicht: (2024)
von: Xu, Jiahui, et al.
Veröffentlicht: (2024)
LLMs for Doctors: Leveraging Medical LLMs to Assist Doctors, Not Replace Them
von: Xie, Wenya, et al.
Veröffentlicht: (2024)
von: Xie, Wenya, et al.
Veröffentlicht: (2024)
Open Rubric System: Scaling Reinforcement Learning with Pairwise Adaptive Rubric
von: Jia, Ruipeng, et al.
Veröffentlicht: (2026)
von: Jia, Ruipeng, et al.
Veröffentlicht: (2026)
Dataset Difficulty and the Role of Inductive Bias
von: Kwok, Devin, et al.
Veröffentlicht: (2024)
von: Kwok, Devin, et al.
Veröffentlicht: (2024)
Interpolated-MLPs: Controllable Inductive Bias
von: Wu, Sean, et al.
Veröffentlicht: (2024)
von: Wu, Sean, et al.
Veröffentlicht: (2024)
Towards Exact Computation of Inductive Bias
von: Boopathy, Akhilan, et al.
Veröffentlicht: (2024)
von: Boopathy, Akhilan, et al.
Veröffentlicht: (2024)
An LLM-Agent-Based Framework for Age of Information Optimization in Heterogeneous Random Access Networks
von: Liu, Fang, et al.
Veröffentlicht: (2026)
von: Liu, Fang, et al.
Veröffentlicht: (2026)
Electromagnetic Property Sensing: A New Paradigm of Integrated Sensing and Communication
von: Jiang, Yuhua, et al.
Veröffentlicht: (2023)
von: Jiang, Yuhua, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
RLHF in an SFT Way: From Optimal Solution to Reward-Weighted Alignment
von: Du, Yuhao, et al.
Veröffentlicht: (2025) -
Atoxia: Red-teaming Large Language Models with Target Toxic Answers
von: Du, Yuhao, et al.
Veröffentlicht: (2024) -
CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR
von: Cui, Sijia, et al.
Veröffentlicht: (2026) -
Rationale Matters: Learning Transferable Rubrics via Proxy-Guided Critique for VLM Reward Models
von: Qiu, Weijie, et al.
Veröffentlicht: (2026) -
MARCH: Multi-Agent Reinforced Self-Check for LLM Hallucination
von: Li, Zhuo, et al.
Veröffentlicht: (2026)