Before the Model Learns the Bug:Fuzzing RLVR Verifiers
Fuente:
arXiv
Salvato in:
| Autore principale: | Ray, Jaideep |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Aletheia: What Makes RLVR For Code Verifiers Tick?
di: Venkatkrishna, Vatsal, et al.
Pubblicazione: (2026)
di: Venkatkrishna, Vatsal, et al.
Pubblicazione: (2026)
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
di: Helff, Lukas, et al.
Pubblicazione: (2026)
di: Helff, Lukas, et al.
Pubblicazione: (2026)
Backdoors in RLVR: Jailbreak Backdoors in LLMs From Verifiable Reward
di: Guo, Weiyang, et al.
Pubblicazione: (2026)
di: Guo, Weiyang, et al.
Pubblicazione: (2026)
Extending RLVR to Open-Ended Tasks via Verifiable Multiple-Choice Reformulation
di: Zhang, Mengyu, et al.
Pubblicazione: (2025)
di: Zhang, Mengyu, et al.
Pubblicazione: (2025)
RLPR: Extrapolating RLVR to General Domains without Verifiers
di: Yu, Tianyu, et al.
Pubblicazione: (2025)
di: Yu, Tianyu, et al.
Pubblicazione: (2025)
Causal Fuzzing for Verifying Machine Unlearning
di: Mazhar, Anna, et al.
Pubblicazione: (2025)
di: Mazhar, Anna, et al.
Pubblicazione: (2025)
PerMix-RLVR: Preserving Persona Expressivity under Verifiable-Reward Alignment
di: Oh, Jihwan, et al.
Pubblicazione: (2026)
di: Oh, Jihwan, et al.
Pubblicazione: (2026)
SWE-Synth: Synthesizing Verifiable Bug-Fix Data to Enable Large Language Models in Resolving Real-World Bugs
di: Pham, Minh V. T., et al.
Pubblicazione: (2025)
di: Pham, Minh V. T., et al.
Pubblicazione: (2025)
IRDS: Interpretable RLVR Data Selection via Verifier-Coupled Sparse Autoencoder Coverage
di: Li, Yuhan, et al.
Pubblicazione: (2026)
di: Li, Yuhan, et al.
Pubblicazione: (2026)
Text Before Vision: Staged Knowledge Injection Matters for Agentic RLVR in Ultra-High-Resolution Remote Sensing Understanding
di: Wang, Fengxiang, et al.
Pubblicazione: (2026)
di: Wang, Fengxiang, et al.
Pubblicazione: (2026)
RLVR-World: Training World Models with Reinforcement Learning
di: Wu, Jialong, et al.
Pubblicazione: (2025)
di: Wu, Jialong, et al.
Pubblicazione: (2025)
Learn More with Less: Uncertainty Consistency Guided Query Selection for RLVR
di: Yi, Hao, et al.
Pubblicazione: (2026)
di: Yi, Hao, et al.
Pubblicazione: (2026)
MEML-GRPO: Heterogeneous Multi-Expert Mutual Learning for RLVR Advancement
di: Jia, Weitao, et al.
Pubblicazione: (2025)
di: Jia, Weitao, et al.
Pubblicazione: (2025)
The Path Not Taken: RLVR Provably Learns Off the Principals
di: Zhu, Hanqing, et al.
Pubblicazione: (2025)
di: Zhu, Hanqing, et al.
Pubblicazione: (2025)
R1-Fuzz: Specializing Language Models for Textual Fuzzing via Reinforcement Learning
di: Lin, Jiayi, et al.
Pubblicazione: (2025)
di: Lin, Jiayi, et al.
Pubblicazione: (2025)
CLIPO: Contrastive Learning in Policy Optimization Generalizes RLVR
di: Cui, Sijia, et al.
Pubblicazione: (2026)
di: Cui, Sijia, et al.
Pubblicazione: (2026)
Does RLVR Extend Reasoning Boundaries? Investigating Capability Expansion in Vision-Language Models
di: Shen, Minghe, et al.
Pubblicazione: (2025)
di: Shen, Minghe, et al.
Pubblicazione: (2025)
FuzzingRL: Reinforcement Fuzz-Testing for Revealing VLM Failures
di: Xu, Jiajun, et al.
Pubblicazione: (2026)
di: Xu, Jiajun, et al.
Pubblicazione: (2026)
VIGIL: Defending LLM Agents Against Tool Stream Injection via Verify-Before-Commit
di: Lin, Junda, et al.
Pubblicazione: (2026)
di: Lin, Junda, et al.
Pubblicazione: (2026)
Verify Before You Commit: Towards Faithful Reasoning in LLM Agents via Self-Auditing
di: Yuan, Wenhao, et al.
Pubblicazione: (2026)
di: Yuan, Wenhao, et al.
Pubblicazione: (2026)
Automated Auditing of Hospital Discharge Summaries for Care Transitions
di: Dasula, Akshat, et al.
Pubblicazione: (2026)
di: Dasula, Akshat, et al.
Pubblicazione: (2026)
Quantifying Empirical Compute-Supervision Tradeoffs in RLVR
di: Mitsuhashi, Ryo, et al.
Pubblicazione: (2026)
di: Mitsuhashi, Ryo, et al.
Pubblicazione: (2026)
RLocator: Reinforcement Learning for Bug Localization
di: Chakraborty, Partha, et al.
Pubblicazione: (2023)
di: Chakraborty, Partha, et al.
Pubblicazione: (2023)
VL Norm: Rethink Loss Aggregation in RLVR
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
di: He, Zhiyuan, et al.
Pubblicazione: (2025)
Spurious Rewards: Rethinking Training Signals in RLVR
di: Shao, Rulin, et al.
Pubblicazione: (2025)
di: Shao, Rulin, et al.
Pubblicazione: (2025)
IntelliAsk: Learning to Ask High-Quality Research Questions via RLVR
di: Sharma, Karun, et al.
Pubblicazione: (2026)
di: Sharma, Karun, et al.
Pubblicazione: (2026)
Learning from Less: Measuring the Effectiveness of RLVR in Low Data and Compute Regimes
di: Bauer, Justin, et al.
Pubblicazione: (2026)
di: Bauer, Justin, et al.
Pubblicazione: (2026)
BugPilot: Complex Bug Generation for Efficient Learning of SWE Skills
di: Sonwane, Atharv, et al.
Pubblicazione: (2025)
di: Sonwane, Atharv, et al.
Pubblicazione: (2025)
Not Every Rubric Teaches Equally: Policy-Aware Rubric Rewards for RLVR
di: Tyagi, Utkarsh, et al.
Pubblicazione: (2026)
di: Tyagi, Utkarsh, et al.
Pubblicazione: (2026)
Probing RLVR training instability through the lens of objective-level hacking
di: Dong, Yiming, et al.
Pubblicazione: (2026)
di: Dong, Yiming, et al.
Pubblicazione: (2026)
Unlocking Exploration in RLVR: Uncertainty-aware Advantage Shaping for Deeper Reasoning
di: Xie, Can, et al.
Pubblicazione: (2025)
di: Xie, Can, et al.
Pubblicazione: (2025)
FuzzAug: Data Augmentation by Coverage-guided Fuzzing for Neural Test Generation
di: He, Yifeng, et al.
Pubblicazione: (2024)
di: He, Yifeng, et al.
Pubblicazione: (2024)
Enhancing In-Hospital Mortality Prediction Using Multi-Representational Learning with LLM-Generated Expert Summaries
di: Battula, Harshavardhan, et al.
Pubblicazione: (2024)
di: Battula, Harshavardhan, et al.
Pubblicazione: (2024)
Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration
di: Chen, Zhipeng, et al.
Pubblicazione: (2026)
di: Chen, Zhipeng, et al.
Pubblicazione: (2026)
FuzzWiz -- Fuzzing Framework for Efficient Hardware Coverage
di: Gadde, Deepak Narayan, et al.
Pubblicazione: (2024)
di: Gadde, Deepak Narayan, et al.
Pubblicazione: (2024)
Bug Analysis Towards Bug Resolution Time Prediction
di: Ozkan, Hasan Yagiz, et al.
Pubblicazione: (2024)
di: Ozkan, Hasan Yagiz, et al.
Pubblicazione: (2024)
On the Direction of RLVR Updates for LLM Reasoning: Identification and Exploitation
di: Huang, Kexin, et al.
Pubblicazione: (2026)
di: Huang, Kexin, et al.
Pubblicazione: (2026)
On the Implicit Reward Overfitting and the Low-rank Dynamics in RLVR
di: Ye, Hao, et al.
Pubblicazione: (2026)
di: Ye, Hao, et al.
Pubblicazione: (2026)
Rethinking Entropy Interventions in RLVR: An Entropy Change Perspective
di: Hao, Zhezheng, et al.
Pubblicazione: (2025)
di: Hao, Zhezheng, et al.
Pubblicazione: (2025)
Quantile Advantage Estimation: Stabilizing RLVR for LLM Reasoning
di: Wu, Junkang, et al.
Pubblicazione: (2025)
di: Wu, Junkang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Aletheia: What Makes RLVR For Code Verifiers Tick?
di: Venkatkrishna, Vatsal, et al.
Pubblicazione: (2026) -
LLMs Gaming Verifiers: RLVR can Lead to Reward Hacking
di: Helff, Lukas, et al.
Pubblicazione: (2026) -
Backdoors in RLVR: Jailbreak Backdoors in LLMs From Verifiable Reward
di: Guo, Weiyang, et al.
Pubblicazione: (2026) -
Extending RLVR to Open-Ended Tasks via Verifiable Multiple-Choice Reformulation
di: Zhang, Mengyu, et al.
Pubblicazione: (2025) -
RLPR: Extrapolating RLVR to General Domains without Verifiers
di: Yu, Tianyu, et al.
Pubblicazione: (2025)