When In-Distribution Gains Fail: Evaluating Weak-to-Strong Reward Models under Preference Shift
Fuente:
arXiv
Salvato in:
| Autori principali: | Le, Khoi, Cao, Tri, Nguyen, Phong, Nguyen, Cong-Duy, Luu, Anh Tuan, Chunyan, Miao, Ng, See-Kiong, Nguyen, Thong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
di: Cao, Tri, et al.
Pubblicazione: (2026)
di: Cao, Tri, et al.
Pubblicazione: (2026)
Vision-and-Language Pretraining
di: Nguyen, Thong, et al.
Pubblicazione: (2022)
di: Nguyen, Thong, et al.
Pubblicazione: (2022)
READ: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling
di: Nguyen, Thong, et al.
Pubblicazione: (2023)
di: Nguyen, Thong, et al.
Pubblicazione: (2023)
DemaFormer: Damped Exponential Moving Average Transformer with Energy-Based Modeling for Temporal Language Grounding
di: Nguyen, Thong, et al.
Pubblicazione: (2023)
di: Nguyen, Thong, et al.
Pubblicazione: (2023)
Topic Modeling as Multi-Objective Contrastive Optimization
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
Temporal-Oriented Recipe for Transferring Large Vision-Language Model to Video Understanding
di: Nguyen, Thong, et al.
Pubblicazione: (2025)
di: Nguyen, Thong, et al.
Pubblicazione: (2025)
Motion-aware Contrastive Learning for Temporal Panoptic Scene Graph Generation
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
Encoding and Controlling Global Semantics for Long-form Video Question Answering
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
MAMA: Meta-optimized Angular Margin Contrastive Framework for Video-Language Representation Learning
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
Don't Read Everything: A Curvature-Conditioned Query for Linear Attention
di: Le, Dong, et al.
Pubblicazione: (2026)
di: Le, Dong, et al.
Pubblicazione: (2026)
Multi-Scale Contrastive Learning for Video Temporal Grounding
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
di: Nguyen, Thong Thanh, et al.
Pubblicazione: (2024)
KDMCSE: Knowledge Distillation Multimodal Sentence Embeddings with Adaptive Angular margin Contrastive Learning
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2024)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2024)
Expand BERT Representation with Visual Information via Grounded Language Learning with Multimodal Partial Alignment
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2023)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2023)
Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
More Bias, Less Bias: BiasPrompting for Enhanced Multiple-Choice Question Answering
di: Vu, Duc Anh, et al.
Pubblicazione: (2025)
di: Vu, Duc Anh, et al.
Pubblicazione: (2025)
Video-Language Understanding: A Survey from Model Architecture, Model Training, and Data Perspectives
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
di: Nguyen, Thong, et al.
Pubblicazione: (2024)
CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
di: Nguyen, Cong-Duy, et al.
Pubblicazione: (2025)
Adaptive Contrastive Learning on Multimodal Transformer for Review Helpfulness Predictions
di: Nguyen, Thong, et al.
Pubblicazione: (2022)
di: Nguyen, Thong, et al.
Pubblicazione: (2022)
A Survey on Neural Topic Models: Methods, Applications, and Challenges
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
Unlearning Backdoor Attacks for LLMs with Weak-to-Strong Knowledge Distillation
di: Zhao, Shuai, et al.
Pubblicazione: (2024)
di: Zhao, Shuai, et al.
Pubblicazione: (2024)
Gradient-Boosted Decision Tree for Listwise Context Model in Multimodal Review Helpfulness Prediction
di: Nguyen, Thong, et al.
Pubblicazione: (2023)
di: Nguyen, Thong, et al.
Pubblicazione: (2023)
On the Affinity, Rationality, and Diversity of Hierarchical Topic Modeling
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
Breaking PEFT Limitations: Leveraging Weak-to-Strong Knowledge Transfer for Backdoor Attacks in LLMs
di: Zhao, Shuai, et al.
Pubblicazione: (2024)
di: Zhao, Shuai, et al.
Pubblicazione: (2024)
Learning Uncertainty from Sequential Internal Dispersion in Large Language Models
di: Srey, Ponhvoan, et al.
Pubblicazione: (2026)
di: Srey, Ponhvoan, et al.
Pubblicazione: (2026)
Enriching and Controlling Global Semantics for Text Summarization
di: Nguyen, Thong, et al.
Pubblicazione: (2021)
di: Nguyen, Thong, et al.
Pubblicazione: (2021)
Semi-supervised 3D Semantic Scene Completion with 2D Vision Foundation Model Guidance
di: Pham, Duc-Hai, et al.
Pubblicazione: (2024)
di: Pham, Duc-Hai, et al.
Pubblicazione: (2024)
Cost-Adaptive Recourse Recommendation by Adaptive Preference Elicitation
di: Nguyen, Duy, et al.
Pubblicazione: (2024)
di: Nguyen, Duy, et al.
Pubblicazione: (2024)
Curriculum Demonstration Selection for In-Context Learning
di: Vu, Duc Anh, et al.
Pubblicazione: (2024)
di: Vu, Duc Anh, et al.
Pubblicazione: (2024)
CASUAL: Conditional Support Alignment for Domain Adaptation with Label Shift
di: Nguyen, Anh T, et al.
Pubblicazione: (2023)
di: Nguyen, Anh T, et al.
Pubblicazione: (2023)
Effects of magnetic field and structural parameters on multi-photon absorption spectra in Morse quantum wells with electron-phonon interactions
di: Vi, Tran Ky, et al.
Pubblicazione: (2024)
di: Vi, Tran Ky, et al.
Pubblicazione: (2024)
Evaluating the spatiotemporal variation of Ba River water quality in the agricultural and urban watershed in the highland of Vietnam
di: Tuan Anh Nguyen, et al.
Pubblicazione: (2024)
di: Tuan Anh Nguyen, et al.
Pubblicazione: (2024)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
di: Nguyen, Quang-Binh, et al.
Pubblicazione: (2025)
Handheld Thermal Devices Can Facilitate Population Monitoring of the Critically Endangered Delacour's Langur Trachypithecus delacouri in Difficult Terrains
di: Anh Tuan Nguyen, et al.
Pubblicazione: (2026)
di: Anh Tuan Nguyen, et al.
Pubblicazione: (2026)
Mercury: A Code Efficiency Benchmark for Code Large Language Models
di: Du, Mingzhe, et al.
Pubblicazione: (2024)
di: Du, Mingzhe, et al.
Pubblicazione: (2024)
Guiding VLM Agents with Process Rewards at Inference Time for GUI Navigation
di: Hu, Zhiyuan, et al.
Pubblicazione: (2025)
di: Hu, Zhiyuan, et al.
Pubblicazione: (2025)
Any3DIS: Class-Agnostic 3D Instance Segmentation by 2D Mask Tracking
di: Nguyen, Phuc, et al.
Pubblicazione: (2024)
di: Nguyen, Phuc, et al.
Pubblicazione: (2024)
Distributional Surgery for Language Model Activations
di: Nguyen, Bao, et al.
Pubblicazione: (2025)
di: Nguyen, Bao, et al.
Pubblicazione: (2025)
FurniMAS: Language-Guided Furniture Decoration using Multi-Agent System
di: Nguyen, Toan, et al.
Pubblicazione: (2025)
di: Nguyen, Toan, et al.
Pubblicazione: (2025)
Modeling Dynamic Topics in Chain-Free Fashion by Evolution-Tracking Contrastive Learning and Unassociated Word Exclusion
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
di: Wu, Xiaobao, et al.
Pubblicazione: (2024)
DiverseDream: Diverse Text-to-3D Synthesis with Augmented Text Embedding
di: Tran, Uy Dieu, et al.
Pubblicazione: (2023)
di: Tran, Uy Dieu, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
di: Cao, Tri, et al.
Pubblicazione: (2026) -
Vision-and-Language Pretraining
di: Nguyen, Thong, et al.
Pubblicazione: (2022) -
READ: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling
di: Nguyen, Thong, et al.
Pubblicazione: (2023) -
DemaFormer: Damped Exponential Moving Average Transformer with Energy-Based Modeling for Temporal Language Grounding
di: Nguyen, Thong, et al.
Pubblicazione: (2023) -
Topic Modeling as Multi-Objective Contrastive Optimization
di: Nguyen, Thong, et al.
Pubblicazione: (2024)