Improving Vision-language Models with Perception-centric Process Reward Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Min, Yingqian, Zhou, Kun, Li, Yifan, Wu, Yuhuan, Peng, Han, Du, Yifan, Zhao, Wayne Xin, Yang, Min, Wen, Ji-Rong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Revisiting the Necessity of Lengthy Chain-of-Thought in Vision-centric Reasoning Generalization
di: Du, Yifan, et al.
Pubblicazione: (2025)
di: Du, Yifan, et al.
Pubblicazione: (2025)
Beyond the Last Frame: Process-aware Evaluation for Generative Video Reasoning
di: Li, Yifan, et al.
Pubblicazione: (2025)
di: Li, Yifan, et al.
Pubblicazione: (2025)
Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models
di: Sun, Haoxiang, et al.
Pubblicazione: (2025)
di: Sun, Haoxiang, et al.
Pubblicazione: (2025)
Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models
di: Li, Yifan, et al.
Pubblicazione: (2024)
di: Li, Yifan, et al.
Pubblicazione: (2024)
Analyzing and Mitigating Object Hallucination: A Training Bias Perspective
di: Li, Yifan, et al.
Pubblicazione: (2025)
di: Li, Yifan, et al.
Pubblicazione: (2025)
R1-Searcher++: Incentivizing the Dynamic Knowledge Acquisition of LLMs via Reinforcement Learning
di: Song, Huatong, et al.
Pubblicazione: (2025)
di: Song, Huatong, et al.
Pubblicazione: (2025)
ICPC-Eval: Probing the Frontiers of LLM Reasoning with Competitive Programming Contests
di: Xu, Shiyi, et al.
Pubblicazione: (2025)
di: Xu, Shiyi, et al.
Pubblicazione: (2025)
Data-CUBE: Data Curriculum for Instruction-based Sentence Representation Learning
di: Min, Yingqian, et al.
Pubblicazione: (2024)
di: Min, Yingqian, et al.
Pubblicazione: (2024)
Do we Really Need Visual Instructions? Towards Visual Instruction-Free Fine-tuning for Large Vision-Language Models
di: Liu, Zikang, et al.
Pubblicazione: (2025)
di: Liu, Zikang, et al.
Pubblicazione: (2025)
Sticker-TTS: Learn to Utilize Historical Experience with a Sticker-driven Test-Time Scaling Framework
di: Chen, Jie, et al.
Pubblicazione: (2025)
di: Chen, Jie, et al.
Pubblicazione: (2025)
Towards Event-oriented Long Video Understanding
di: Du, Yifan, et al.
Pubblicazione: (2024)
di: Du, Yifan, et al.
Pubblicazione: (2024)
An Empirical Study on Eliciting and Improving R1-like Reasoning Models
di: Chen, Zhipeng, et al.
Pubblicazione: (2025)
di: Chen, Zhipeng, et al.
Pubblicazione: (2025)
ChainLM: Empowering Large Language Models with Improved Chain-of-Thought Prompting
di: Cheng, Xiaoxue, et al.
Pubblicazione: (2024)
di: Cheng, Xiaoxue, et al.
Pubblicazione: (2024)
Not Everything is All You Need: Toward Low-Redundant Optimization for Large Language Model Alignment
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
Improving Large Language Models via Fine-grained Reinforcement Learning with Minimum Editing Constraint
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
Unleashing Perception-Time Scaling to Multimodal Reasoning Models
di: Li, Yifan, et al.
Pubblicazione: (2025)
di: Li, Yifan, et al.
Pubblicazione: (2025)
ReasoningLM: Enabling Structural Subgraph Reasoning in Pre-trained Language Models for Question Answering over Knowledge Graph
di: Jiang, Jinhao, et al.
Pubblicazione: (2023)
di: Jiang, Jinhao, et al.
Pubblicazione: (2023)
A Survey of Large Language Models
di: Zhao, Wayne Xin, et al.
Pubblicazione: (2023)
di: Zhao, Wayne Xin, et al.
Pubblicazione: (2023)
Towards Long-horizon Agentic Multimodal Search
di: Du, Yifan, et al.
Pubblicazione: (2026)
di: Du, Yifan, et al.
Pubblicazione: (2026)
Exploring the Design Space of Visual Context Representation in Video MLLMs
di: Du, Yifan, et al.
Pubblicazione: (2024)
di: Du, Yifan, et al.
Pubblicazione: (2024)
R1-Searcher: Incentivizing the Search Capability in LLMs via Reinforcement Learning
di: Song, Huatong, et al.
Pubblicazione: (2025)
di: Song, Huatong, et al.
Pubblicazione: (2025)
DEVAL: A Framework for Evaluating and Improving the Derivation Capability of Large Language Models
di: Li, Yifan, et al.
Pubblicazione: (2025)
di: Li, Yifan, et al.
Pubblicazione: (2025)
What, Whether and How? Unveiling Process Reward Models for Thinking with Images Reasoning
di: Zhou, Yujin, et al.
Pubblicazione: (2026)
di: Zhou, Yujin, et al.
Pubblicazione: (2026)
What Makes for Good Visual Instructions? Synthesizing Complex Visual Reasoning Instructions for Visual Instruction Tuning
di: Du, Yifan, et al.
Pubblicazione: (2023)
di: Du, Yifan, et al.
Pubblicazione: (2023)
Improving Conversational Recommendation Systems via Counterfactual Data Simulation
di: Wang, Xiaolei, et al.
Pubblicazione: (2023)
di: Wang, Xiaolei, et al.
Pubblicazione: (2023)
Abstract 3D Perception for Spatial Intelligence in Vision-Language Models
di: Liu, Yifan, et al.
Pubblicazione: (2025)
di: Liu, Yifan, et al.
Pubblicazione: (2025)
Enhancing LLM Reasoning with Reward-guided Tree Search
di: Jiang, Jinhao, et al.
Pubblicazione: (2024)
di: Jiang, Jinhao, et al.
Pubblicazione: (2024)
Towards Effective Code-Integrated Reasoning
di: Bai, Fei, et al.
Pubblicazione: (2025)
di: Bai, Fei, et al.
Pubblicazione: (2025)
MolMetaLM: a Physicochemical Knowledge-Guided Molecular Meta Language Model
di: Wu, Yifan, et al.
Pubblicazione: (2024)
di: Wu, Yifan, et al.
Pubblicazione: (2024)
Low-rank Optimization Trajectories Modeling for LLM RLVR Acceleration
di: Chen, Zhipeng, et al.
Pubblicazione: (2026)
di: Chen, Zhipeng, et al.
Pubblicazione: (2026)
JiuZhang3.0: Efficiently Improving Mathematical Reasoning by Training Small Data Synthesis Models
di: Zhou, Kun, et al.
Pubblicazione: (2024)
di: Zhou, Kun, et al.
Pubblicazione: (2024)
Think More, Hallucinate Less: Mitigating Hallucinations via Dual Process of Fast and Slow Thinking
di: Cheng, Xiaoxue, et al.
Pubblicazione: (2025)
di: Cheng, Xiaoxue, et al.
Pubblicazione: (2025)
Adaptive Ability Decomposing for Unlocking Large Reasoning Model Effective Reinforcement Learning
di: Chen, Zhipeng, et al.
Pubblicazione: (2026)
di: Chen, Zhipeng, et al.
Pubblicazione: (2026)
Extracting and Combining Abilities For Building Multi-lingual Ability-enhanced Large Language Models
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
di: Chen, Zhipeng, et al.
Pubblicazione: (2024)
Search-Based Interaction For Conversation Recommendation via Generative Reward Model Based Simulated User
di: Wang, Xiaolei, et al.
Pubblicazione: (2025)
di: Wang, Xiaolei, et al.
Pubblicazione: (2025)
Virgo: A Preliminary Exploration on Reproducing o1-like MLLM
di: Du, Yifan, et al.
Pubblicazione: (2025)
di: Du, Yifan, et al.
Pubblicazione: (2025)
AVC-DPO: Aligned Video Captioning via Direct Preference Optimization
di: Tang, Jiyang, et al.
Pubblicazione: (2025)
di: Tang, Jiyang, et al.
Pubblicazione: (2025)
ViGoR: Improving Visual Grounding of Large Vision Language Models with Fine-Grained Reward Modeling
di: Yan, Siming, et al.
Pubblicazione: (2024)
di: Yan, Siming, et al.
Pubblicazione: (2024)
BAMBOO: A Comprehensive Benchmark for Evaluating Long Text Modeling Capacities of Large Language Models
di: Dong, Zican, et al.
Pubblicazione: (2023)
di: Dong, Zican, et al.
Pubblicazione: (2023)
Benchmarking GPT-5 for biomedical natural language processing
di: Hou, Yu, et al.
Pubblicazione: (2025)
di: Hou, Yu, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Revisiting the Necessity of Lengthy Chain-of-Thought in Vision-centric Reasoning Generalization
di: Du, Yifan, et al.
Pubblicazione: (2025) -
Beyond the Last Frame: Process-aware Evaluation for Generative Video Reasoning
di: Li, Yifan, et al.
Pubblicazione: (2025) -
Challenging the Boundaries of Reasoning: An Olympiad-Level Math Benchmark for Large Language Models
di: Sun, Haoxiang, et al.
Pubblicazione: (2025) -
Images are Achilles' Heel of Alignment: Exploiting Visual Vulnerabilities for Jailbreaking Multimodal Large Language Models
di: Li, Yifan, et al.
Pubblicazione: (2024) -
Analyzing and Mitigating Object Hallucination: A Training Bias Perspective
di: Li, Yifan, et al.
Pubblicazione: (2025)