PDCR: Perception-Decomposed Confidence Reward for Vision-Language Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Yoon, Hee Suk, Yoon, Eunseop, Hong, Ji Woo, Eom, SooHwan, Koo, Gwanhyeong, Hasegawa-Johnson, Mark, Dai, Qi, Luo, Chong, Yoo, Chang D. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PACR: Progressively Ascending Confidence Reward for LLM Reasoning
by: Yoon, Eunseop, et al.
Published: (2025)
by: Yoon, Eunseop, et al.
Published: (2025)
Decomposed On-Policy Distillation for Vision-Language Reasoning: Steering Gradients for Visual Grounding
by: Yoon, Hee Suk, et al.
Published: (2026)
by: Yoon, Hee Suk, et al.
Published: (2026)
High-Fidelity Text-to-Image Generation from Pre-Trained Vision-Language Models via Distribution-Conditioned Diffusion Decoding
by: Hong, Ji Woo, et al.
Published: (2026)
by: Hong, Ji Woo, et al.
Published: (2026)
AdaMER-CTC: Connectionist Temporal Classification with Adaptive Maximum Entropy Regularization for Automatic Speech Recognition
by: Eom, SooHwan, et al.
Published: (2024)
by: Eom, SooHwan, et al.
Published: (2024)
TLCR: Token-Level Continuous Reward for Fine-grained Reinforcement Learning from Human Feedback
by: Yoon, Eunseop, et al.
Published: (2024)
by: Yoon, Eunseop, et al.
Published: (2024)
ConfPO: Exploiting Policy Model Confidence for Critical Token Selection in Preference Optimization
by: Yoon, Hee Suk, et al.
Published: (2025)
by: Yoon, Hee Suk, et al.
Published: (2025)
SiamCTC: Learning Speech Representations through Monotonic Temporal Alignment
by: Eom, SooHwan, et al.
Published: (2026)
by: Eom, SooHwan, et al.
Published: (2026)
Can Video LLMs Refuse to Answer? Alignment for Answerability in Video Large Language Models
by: Yoon, Eunseop, et al.
Published: (2025)
by: Yoon, Eunseop, et al.
Published: (2025)
LI-TTA: Language Informed Test-Time Adaptation for Automatic Speech Recognition
by: Yoon, Eunseop, et al.
Published: (2024)
by: Yoon, Eunseop, et al.
Published: (2024)
C-TPT: Calibrated Test-Time Prompt Tuning for Vision-Language Models via Text Feature Dispersion
by: Yoon, Hee Suk, et al.
Published: (2024)
by: Yoon, Hee Suk, et al.
Published: (2024)
DNI: Dilutional Noise Initialization for Diffusion Video Editing
by: Yoon, Sunjae, et al.
Published: (2024)
by: Yoon, Sunjae, et al.
Published: (2024)
FlexiEdit: Frequency-Aware Latent Refinement for Enhanced Non-Rigid Editing
by: Koo, Gwanhyeong, et al.
Published: (2024)
by: Koo, Gwanhyeong, et al.
Published: (2024)
SimPSI: A Simple Strategy to Preserve Spectral Information in Time Series Data Augmentation
by: Ryu, Hyun, et al.
Published: (2023)
by: Ryu, Hyun, et al.
Published: (2023)
Selective Query-guided Debiasing for Video Corpus Moment Retrieval
by: Yoon, Sunjae, et al.
Published: (2022)
by: Yoon, Sunjae, et al.
Published: (2022)
Wavelet-Guided Acceleration of Text Inversion in Diffusion-Based Image Editing
by: Koo, Gwanhyeong, et al.
Published: (2024)
by: Koo, Gwanhyeong, et al.
Published: (2024)
HEAR: Hearing Enhanced Audio Response for Video-grounded Dialogue
by: Yoon, Sunjae, et al.
Published: (2023)
by: Yoon, Sunjae, et al.
Published: (2023)
FlowDrag: 3D-aware Drag-based Image Editing with Mesh-guided Deformation Vector Flow Fields
by: Koo, Gwanhyeong, et al.
Published: (2025)
by: Koo, Gwanhyeong, et al.
Published: (2025)
Occlusion-robust Stylization for Drawing-based 3D Animation
by: Yoon, Sunjae, et al.
Published: (2025)
by: Yoon, Sunjae, et al.
Published: (2025)
TPC: Test-time Procrustes Calibration for Diffusion-based Human Image Animation
by: Yoon, Sunjae, et al.
Published: (2024)
by: Yoon, Sunjae, et al.
Published: (2024)
FRAG: Frequency Adapting Group for Diffusion Video Editing
by: Yoon, Sunjae, et al.
Published: (2024)
by: Yoon, Sunjae, et al.
Published: (2024)
SCANet: Scene Complexity Aware Network for Weakly-Supervised Video Moment Retrieval
by: Yoon, Sunjae, et al.
Published: (2023)
by: Yoon, Sunjae, et al.
Published: (2023)
Zero-Shot Dual-Path Integration Framework for Open-Vocabulary 3D Instance Segmentation
by: Ton, Tri, et al.
Published: (2024)
by: Ton, Tri, et al.
Published: (2024)
ITA-MDT: Image-Timestep-Adaptive Masked Diffusion Transformer Framework for Image-Based Virtual Try-On
by: Hong, Ji Woo, et al.
Published: (2025)
by: Hong, Ji Woo, et al.
Published: (2025)
ESD: Expected Squared Difference as a Tuning-Free Trainable Calibration Measure
by: Yoon, Hee Suk, et al.
Published: (2023)
by: Yoon, Hee Suk, et al.
Published: (2023)
BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation
by: Yoon, Hee Suk, et al.
Published: (2024)
by: Yoon, Hee Suk, et al.
Published: (2024)
Versatile and Fast Location-Based Private Information Retrieval with Fully Homomorphic Encryption over the Torus
by: Yoo, Joon Soo, et al.
Published: (2025)
by: Yoo, Joon Soo, et al.
Published: (2025)
Programmable spectral shaping to improve the measurement precision of frequency comb mode-resolved spectral interferometric ranging
by: Jang, Yoon-Soo, et al.
Published: (2023)
by: Jang, Yoon-Soo, et al.
Published: (2023)
ThermoAct:Thermal-Aware Vision-Language-Action Models for Robotic Perception and Decision-Making
by: Son, Young-Chae, et al.
Published: (2026)
by: Son, Young-Chae, et al.
Published: (2026)
A Steady‐State Eulerian Smoothed Particle Hydrodynamics ( SPH ) Approach for Incompressible Flow and Heat Transfer Using the Semi‐Implicit Method for Pressure‐Linked Equations ( SIMPLE ) Algorithm
by: Tae Hwan Kim, et al.
Published: (2026)
by: Tae Hwan Kim, et al.
Published: (2026)
Bidirectional Biometric Authentication Using Transciphering and (T)FHE
by: Yoo, Joon Soo, et al.
Published: (2025)
by: Yoo, Joon Soo, et al.
Published: (2025)
Alix‐normalized exosomal programmed death‐ligand 1 analysis in urine enables precision monitoring of urothelial cancer
by: Hyun‐Kyung Woo, et al.
Published: (2024)
by: Hyun‐Kyung Woo, et al.
Published: (2024)
HuBERT-EE: Early Exiting HuBERT for Efficient Speech Recognition
by: Yoon, Ji Won, et al.
Published: (2022)
by: Yoon, Ji Won, et al.
Published: (2022)
Bad Seeing or Bad Thinking? Rewarding Perception for Vision-Language Reasoning
by: Wang, Haozhe, et al.
Published: (2026)
by: Wang, Haozhe, et al.
Published: (2026)
Highly Ordered Mesoporous Polymer‐Supported Phosphine as the Ligand for Organometallic Reaction: Suzuki‐Miyaura Cross‐Coupling of Aryl Chlorides at Room Temperature
by: Hwang Suk Kim, et al.
Published: (2024)
by: Hwang Suk Kim, et al.
Published: (2024)
Selective Perception for Robot: Task-Aware Attention in Multimodal VLA
by: Son, Young-Chae, et al.
Published: (2026)
by: Son, Young-Chae, et al.
Published: (2026)
“Off‐The‐Shelf” Bioartificial Liver Support System Using Cryopreserved Immobilized Hepatocyte Spheroids in a Porcine Acute Liver Failure Model
by: Ji‐Hyun Lee, et al.
Published: (2025)
by: Ji‐Hyun Lee, et al.
Published: (2025)
The study on the multiplicity dependence of ridge behavior in $pp$ collisions at $\sqrt{s}=13$ TeV at the LHC
by: Yoon, Jeongseok, et al.
Published: (2023)
by: Yoon, Jeongseok, et al.
Published: (2023)
Empowering Multimodal Respiratory Sound Classification with Counterfactual Adversarial Debiasing for Out-of-Distribution Robustness
by: Koo, Heejoon, et al.
Published: (2025)
by: Koo, Heejoon, et al.
Published: (2025)
Confidence-guided Refinement Reasoning for Zero-shot Question Answering
by: Jang, Youwon, et al.
Published: (2025)
by: Jang, Youwon, et al.
Published: (2025)
Approaching the quantum-limited precision in frequency-comb-based spectral interferometry for length measurements
by: Jang, Yoon-Soo, et al.
Published: (2025)
by: Jang, Yoon-Soo, et al.
Published: (2025)
Similar Items
-
PACR: Progressively Ascending Confidence Reward for LLM Reasoning
by: Yoon, Eunseop, et al.
Published: (2025) -
Decomposed On-Policy Distillation for Vision-Language Reasoning: Steering Gradients for Visual Grounding
by: Yoon, Hee Suk, et al.
Published: (2026) -
High-Fidelity Text-to-Image Generation from Pre-Trained Vision-Language Models via Distribution-Conditioned Diffusion Decoding
by: Hong, Ji Woo, et al.
Published: (2026) -
AdaMER-CTC: Connectionist Temporal Classification with Adaptive Maximum Entropy Regularization for Automatic Speech Recognition
by: Eom, SooHwan, et al.
Published: (2024) -
TLCR: Token-Level Continuous Reward for Fine-grained Reinforcement Learning from Human Feedback
by: Yoon, Eunseop, et al.
Published: (2024)