ADHint: Adaptive Hints with Difficulty Priors for Reinforcement Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhang, Feng, Tan, Zezhong, Ma, Xinhong, Dong, Ziqiang, Leng, Xi, Zhao, Jianfei, Sun, Xin, Yang, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
di: Zhang, Feng, et al.
Pubblicazione: (2026)
di: Zhang, Feng, et al.
Pubblicazione: (2026)
Cross-Image Contrastive Decoding: Precise, Lossless Suppression of Language Priors in Large Vision-Language Models
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
Enhancing Diffusion-based Restoration Models via Difficulty-Adaptive Reinforcement Learning with IQA Reward
di: Xu, Xiaogang, et al.
Pubblicazione: (2025)
di: Xu, Xiaogang, et al.
Pubblicazione: (2025)
Estimating Noisy Class Posterior with Part-level Labels for Noisy Label Learning
di: Zhao, Rui, et al.
Pubblicazione: (2024)
di: Zhao, Rui, et al.
Pubblicazione: (2024)
GNNRL-Smoothing: A Prior-Free Reinforcement Learning Model for Mesh Smoothing
di: Wang, Zhichao, et al.
Pubblicazione: (2024)
di: Wang, Zhichao, et al.
Pubblicazione: (2024)
Mitigating Hallucination in Large Vision-Language Models through Aligning Attention Distribution to Information Flow
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
Cross-Layer Vision Smoothing: Enhancing Visual Understanding via Sustained Focus on Key Objects in Large Vision-Language Models
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
di: Zhao, Jianfei, et al.
Pubblicazione: (2025)
ViaRL: Adaptive Temporal Grounding via Visual Iterated Amplification Reinforcement Learning
di: Xu, Ziqiang, et al.
Pubblicazione: (2025)
di: Xu, Ziqiang, et al.
Pubblicazione: (2025)
Minerva-Ego: Spatiotemporal Hints for Egocentric Video Understanding
di: Nagrani, Arsha, et al.
Pubblicazione: (2026)
di: Nagrani, Arsha, et al.
Pubblicazione: (2026)
Rethinking Multi-domain Generalization with A General Learning Objective
di: Tan, Zhaorui, et al.
Pubblicazione: (2024)
di: Tan, Zhaorui, et al.
Pubblicazione: (2024)
Semi-off-Policy Reinforcement Learning for Vision-Language Slow-Thinking Reasoning
di: Shen, Junhao, et al.
Pubblicazione: (2025)
di: Shen, Junhao, et al.
Pubblicazione: (2025)
On the Difficulty of Learning a Meta-network for Training Data Selection
di: Du, Zilin, et al.
Pubblicazione: (2026)
di: Du, Zilin, et al.
Pubblicazione: (2026)
Towards Flash Thinking via Decoupled Advantage Policy Optimization
di: Tan, Zezhong, et al.
Pubblicazione: (2025)
di: Tan, Zezhong, et al.
Pubblicazione: (2025)
Self-Adaptive Gamma Context-Aware SSM-based Model for Metal Defect Detection
di: Sun, Sijin, et al.
Pubblicazione: (2025)
di: Sun, Sijin, et al.
Pubblicazione: (2025)
RL-Selector: Reinforcement Learning-Guided Data Selection via Redundancy Assessment
di: Yang, Suorong, et al.
Pubblicazione: (2025)
di: Yang, Suorong, et al.
Pubblicazione: (2025)
Enhancing Interpretability of AR-SSVEP-Based Motor Intention Recognition via CNN-BiLSTM and SHAP Analysis on EEG Data
di: Yang, Lin, et al.
Pubblicazione: (2025)
di: Yang, Lin, et al.
Pubblicazione: (2025)
TEACH: Text Encoding as Curriculum Hints for Scene Text Recognition
di: Yang, Xiahan, et al.
Pubblicazione: (2025)
di: Yang, Xiahan, et al.
Pubblicazione: (2025)
Learning to Rebalance Multi-Modal Optimization by Adaptively Masking Subnetworks
di: Yang, Yang, et al.
Pubblicazione: (2024)
di: Yang, Yang, et al.
Pubblicazione: (2024)
Learning to Adapt Frozen CLIP for Few-Shot Test-Time Domain Adaptation
di: Chi, Zhixiang, et al.
Pubblicazione: (2025)
di: Chi, Zhixiang, et al.
Pubblicazione: (2025)
Disentangled World Models: Learning to Transfer Semantic Knowledge from Distracting Videos for Reinforcement Learning
di: Wang, Qi, et al.
Pubblicazione: (2025)
di: Wang, Qi, et al.
Pubblicazione: (2025)
RelationMatch: Matching In-batch Relationships for Semi-supervised Learning
di: Zhang, Yifan, et al.
Pubblicazione: (2023)
di: Zhang, Yifan, et al.
Pubblicazione: (2023)
SpargeAttention2: Trainable Sparse Attention via Hybrid Top-k+Top-p Masking and Distillation Fine-Tuning
di: Zhang, Jintao, et al.
Pubblicazione: (2026)
di: Zhang, Jintao, et al.
Pubblicazione: (2026)
A Survey on Deep Clustering: From the Prior Perspective
di: Lu, Yiding, et al.
Pubblicazione: (2024)
di: Lu, Yiding, et al.
Pubblicazione: (2024)
Highlight Every Step: Knowledge Distillation via Collaborative Teaching
di: Zhao, Haoran, et al.
Pubblicazione: (2019)
di: Zhao, Haoran, et al.
Pubblicazione: (2019)
Neural Prior Estimation: Learning Class Priors from Latent Representations
di: Yavari, Masoud, et al.
Pubblicazione: (2026)
di: Yavari, Masoud, et al.
Pubblicazione: (2026)
Unlocking the Potential of Difficulty Prior in RL-based Multimodal Reasoning
di: Chen, Mingrui, et al.
Pubblicazione: (2025)
di: Chen, Mingrui, et al.
Pubblicazione: (2025)
Generalization in Online Reinforcement Learning for Mobile Agents
di: Gu, Li, et al.
Pubblicazione: (2026)
di: Gu, Li, et al.
Pubblicazione: (2026)
Closed-Loop Unsupervised Representation Disentanglement with $β$-VAE Distillation and Diffusion Probabilistic Feedback
di: Jin, Xin, et al.
Pubblicazione: (2024)
di: Jin, Xin, et al.
Pubblicazione: (2024)
DiffDesign: Controllable Diffusion with Meta Prior for Efficient Interior Design Generation
di: Yang, Yuxuan, et al.
Pubblicazione: (2024)
di: Yang, Yuxuan, et al.
Pubblicazione: (2024)
Consistency Diffusion Models for Single-Image 3D Reconstruction with Priors
di: Jiang, Chenru, et al.
Pubblicazione: (2025)
di: Jiang, Chenru, et al.
Pubblicazione: (2025)
Adaptive Decision Boundary for Few-Shot Class-Incremental Learning
di: Li, Linhao, et al.
Pubblicazione: (2025)
di: Li, Linhao, et al.
Pubblicazione: (2025)
Multimodal Adaptive Retrieval Augmented Generation through Internal Representation Learning
di: Du, Ruoshuang, et al.
Pubblicazione: (2026)
di: Du, Ruoshuang, et al.
Pubblicazione: (2026)
Taming Generative Diffusion Prior for Universal Blind Image Restoration
di: Tu, Siwei, et al.
Pubblicazione: (2024)
di: Tu, Siwei, et al.
Pubblicazione: (2024)
Uncertainty-aware Diffusion and Reinforcement Learning for Joint Plane Localization and Anomaly Diagnosis in 3D Ultrasound
di: Huang, Yuhao, et al.
Pubblicazione: (2025)
di: Huang, Yuhao, et al.
Pubblicazione: (2025)
XTransfer: Modality-Agnostic Few-Shot Model Transfer for Human Sensing at the Edge
di: Zhang, Yu, et al.
Pubblicazione: (2025)
di: Zhang, Yu, et al.
Pubblicazione: (2025)
Distill to Think, Foresee to Act: Cognitive-Physical Reinforcement Learning for Autonomous Driving
di: Wu, Yang, et al.
Pubblicazione: (2026)
di: Wu, Yang, et al.
Pubblicazione: (2026)
Efficient Reinforcement Learning Through Adaptively Pretrained Visual Encoder
di: Zhang, Yuhan, et al.
Pubblicazione: (2025)
di: Zhang, Yuhan, et al.
Pubblicazione: (2025)
Oscillation-Reduced MXFP4 Training for Vision Transformers
di: Chen, Yuxiang, et al.
Pubblicazione: (2025)
di: Chen, Yuxiang, et al.
Pubblicazione: (2025)
Cross-modal Active Complementary Learning with Self-refining Correspondence
di: Qin, Yang, et al.
Pubblicazione: (2023)
di: Qin, Yang, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Revisiting Reinforcement Learning with Verifiable Rewards from a Contrastive Perspective
di: Zhang, Feng, et al.
Pubblicazione: (2026) -
Cross-Image Contrastive Decoding: Precise, Lossless Suppression of Language Priors in Large Vision-Language Models
di: Zhao, Jianfei, et al.
Pubblicazione: (2025) -
Tell Model Where to Look: Mitigating Hallucinations in MLLMs by Vision-Guided Attention
di: Zhao, Jianfei, et al.
Pubblicazione: (2025) -
Enhancing Diffusion-based Restoration Models via Difficulty-Adaptive Reinforcement Learning with IQA Reward
di: Xu, Xiaogang, et al.
Pubblicazione: (2025) -
Estimating Noisy Class Posterior with Part-level Labels for Noisy Label Learning
di: Zhao, Rui, et al.
Pubblicazione: (2024)