StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhang, Xiangxiang, Wei, Jingxuan, Zhong, Donghong, Chen, Qi, Jia, Caijun, Tan, Cheng, Gu, Jinming, Qin, Xiaobo, Liu, Zhiping, Hu, Liang, Sun, Tong, Wu, Yuchen, Sun, Zewei, Lou, Chenwei, Zheng, Hua, Zhan, Tianyang, Wang, Changbao, Wu, Shuangzhi, Lin, Zefa, Guo, Chang, Yuan, Sihang, Chen, Riwei, Zhao, Shixiong, Zhang, Yingping, Wu, Gaowei, Yu, Bihui, Wu, Jiahui, Zhao, Zhehui, Liu, Qianqian, Tang, Ruofeng, Huang, Xingyue, Zhao, Bing, Zhang, Mengyang, Zhou, Youqiang |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
LLaVA-NeuMT: Selective Layer-Neuron Modulation for Efficient Multilingual Multimodal Translation
por: Wei, Jingxuan, et al.
Publicado: (2025)
por: Wei, Jingxuan, et al.
Publicado: (2025)
AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning
por: Lou, Chenwei, et al.
Publicado: (2025)
por: Lou, Chenwei, et al.
Publicado: (2025)
Estimation of Treatment Harm Rate via Partitioning
por: Liang, Wei, et al.
Publicado: (2025)
por: Liang, Wei, et al.
Publicado: (2025)
Nonparametric Bounds in Causal Mediation Analysis in the Presence of Unmeasured Confounding and Imperfect Compliance
por: Liang, Wei, et al.
Publicado: (2025)
por: Liang, Wei, et al.
Publicado: (2025)
AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
por: Yang, Dejie, et al.
Publicado: (2025)
por: Yang, Dejie, et al.
Publicado: (2025)
Structurally Modulated Formation of Cyanine J‐Aggregates with Sharp and Tunable Spectra for Multiplexed Optoacoustic and Fluorescence Bioimaging
por: Chaobang Zhang, et al.
Publicado: (2024)
por: Chaobang Zhang, et al.
Publicado: (2024)
Pseudo-Empirical Likelihood Methods for Causal Inference
por: Huang, Jingyue, et al.
Publicado: (2024)
por: Huang, Jingyue, et al.
Publicado: (2024)
Sample Empirical Likelihood Methods for Causal Inference
por: Huang, Jingyue, et al.
Publicado: (2024)
por: Huang, Jingyue, et al.
Publicado: (2024)
Ceftazidime‐Induced Agranulocytosis: A Case Report
por: Bingbin Dong, et al.
Publicado: (2025)
por: Bingbin Dong, et al.
Publicado: (2025)
ChartMind: A Comprehensive Benchmark for Complex Real-world Multimodal Chart Question Answering
por: Wei, Jingxuan, et al.
Publicado: (2025)
por: Wei, Jingxuan, et al.
Publicado: (2025)
Geoint-R1: Formalizing Multimodal Geometric Reasoning with Dynamic Auxiliary Constructions
por: Wei, Jingxuan, et al.
Publicado: (2025)
por: Wei, Jingxuan, et al.
Publicado: (2025)
Phospholipase A2 and Systemic‐Immune Inflammation Index as Early Predictors of Neurotoxicity Induced by Acute Glufosinate Ammonium Poisoning: A Population‐Based Case–Control Analysis
por: Xiang Xue, et al.
Publicado: (2026)
por: Xiang Xue, et al.
Publicado: (2026)
Formation of Ultra‐Narrowband SWIR J‐Aggregate Materials and Their Applications in Multispectral Optoacoustic Tomography Imaging
por: Chaobang Zhang, et al.
Publicado: (2026)
por: Chaobang Zhang, et al.
Publicado: (2026)
VRM: Knowledge Distillation via Virtual Relation Matching
por: Zhang, Weijia, et al.
Publicado: (2025)
por: Zhang, Weijia, et al.
Publicado: (2025)
Enhancing 1-Second 3D SELD Performance with Filter Bank Analysis and SCConv Integration in CST-Former
por: Zhang, Zhehui
Publicado: (2024)
por: Zhang, Zhehui
Publicado: (2024)
From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing
por: Wei, Jingxuan, et al.
Publicado: (2024)
por: Wei, Jingxuan, et al.
Publicado: (2024)
GenProve: Learning to Generate Text with Fine-Grained Provenance
por: Wei, Jingxuan, et al.
Publicado: (2026)
por: Wei, Jingxuan, et al.
Publicado: (2026)
Canvas-of-Thought: Grounding Reasoning via Mutable Structured States
por: Sun, Lingzhuang, et al.
Publicado: (2026)
por: Sun, Lingzhuang, et al.
Publicado: (2026)
Thinking with Drafting: Optical Decompression via Logical Reconstruction
por: Wei, Jingxuan, et al.
Publicado: (2026)
por: Wei, Jingxuan, et al.
Publicado: (2026)
How RL Unlocks the Aha Moment in Geometric Interleaved Reasoning
por: Zhang, Xiangxiang, et al.
Publicado: (2026)
por: Zhang, Xiangxiang, et al.
Publicado: (2026)
Maple: A Multi-agent System for Portable Deep Learning across Clusters
por: Wu, Molang, et al.
Publicado: (2025)
por: Wu, Molang, et al.
Publicado: (2025)
Transfer Learning Enhanced Single-choice Decision for Multi-choice Question Answering
por: Cui, Chenhao, et al.
Publicado: (2024)
por: Cui, Chenhao, et al.
Publicado: (2024)
Quantifying the retrieval uncertainties associated with systematic and random errors in multi‐satellite‐only precipitation estimates over the Chinese mainland
por: Zhehui Shen, et al.
Publicado: (2024)
por: Zhehui Shen, et al.
Publicado: (2024)
Therapeutic potential of dihydroartemisinin in mitigating radiation‐induced lung injury: Inhibition of ferroptosis through Nrf2/HO‐1 pathways in mice
por: Xin Ning, et al.
Publicado: (2024)
por: Xin Ning, et al.
Publicado: (2024)
GGBench: A Geometric Generative Reasoning Benchmark for Unified Multimodal Models
por: Wei, Jingxuan, et al.
Publicado: (2025)
por: Wei, Jingxuan, et al.
Publicado: (2025)
SoundScape: A Human-AI Co-Creation System Making Your Memories Heard
por: Zhong, Chongjun, et al.
Publicado: (2024)
por: Zhong, Chongjun, et al.
Publicado: (2024)
Improving participation equity in dialogic collaborative problem solving: A participatory visual learning analytical approach
por: Liru Hu, et al.
Publicado: (2024)
por: Liru Hu, et al.
Publicado: (2024)
Unintended Negative Impacts of Promotional Language in Patent Evaluation
por: Zhao, Bingkun, et al.
Publicado: (2026)
por: Zhao, Bingkun, et al.
Publicado: (2026)
ResearchPulse: Building Method-Experiment Chains through Multi-Document Scientific Inference
por: Chen, Qi, et al.
Publicado: (2025)
por: Chen, Qi, et al.
Publicado: (2025)
Tumeochrysa (Nineta) acuta Wu & Liu, 2024, sp. nov.
por: Wu, Jingyu, et al.
Publicado: (2024)
por: Wu, Jingyu, et al.
Publicado: (2024)
On the contribution of dwarf galaxies to reionization of the Universe
por: Wu, Zewei, et al.
Publicado: (2024)
por: Wu, Zewei, et al.
Publicado: (2024)
Micronumerical Simulation of Compression Failure and Size Effect in Lightweight Aggregate Concrete
por: Nannan Sun, et al.
Publicado: (2026)
por: Nannan Sun, et al.
Publicado: (2026)
Chemical Fuel‐Driven Stiffening of Transient Hydrogels via Vitrifiable Phase Separation
por: Ying Zhao, et al.
Publicado: (2025)
por: Ying Zhao, et al.
Publicado: (2025)
A Survey on Data-Centric Recommender Systems
por: Lai, Riwei, et al.
Publicado: (2024)
por: Lai, Riwei, et al.
Publicado: (2024)
Statistical Inference with Nonignorable Non-Probability Survey Samples
por: Liu, Yang, et al.
Publicado: (2024)
por: Liu, Yang, et al.
Publicado: (2024)
RL-Struct: A Lightweight Reinforcement Learning Framework for Reliable Structured Output in LLMs
por: Hu, Ruike, et al.
Publicado: (2025)
por: Hu, Ruike, et al.
Publicado: (2025)
Octopus: Alleviating Hallucination via Dynamic Contrastive Decoding
por: Suo, Wei, et al.
Publicado: (2025)
por: Suo, Wei, et al.
Publicado: (2025)
Diagnosing and Mitigating Semantic Inconsistencies in Wikidata's Classification Hierarchy
por: Zhao, Shixiong, et al.
Publicado: (2025)
por: Zhao, Shixiong, et al.
Publicado: (2025)
Integrating Ecosystem Resilience to Revisit Land Degradation Neutrality in China
por: Chenwei Zhang, et al.
Publicado: (2026)
por: Chenwei Zhang, et al.
Publicado: (2026)
Head Anchor Enhanced Detection and Association for Crowded Pedestrian Tracking
por: Wu, Zewei, et al.
Publicado: (2025)
por: Wu, Zewei, et al.
Publicado: (2025)
Ejemplares similares
-
LLaVA-NeuMT: Selective Layer-Neuron Modulation for Efficient Multilingual Multimodal Translation
por: Wei, Jingxuan, et al.
Publicado: (2025) -
AdaCoT: Pareto-Optimal Adaptive Chain-of-Thought Triggering via Reinforcement Learning
por: Lou, Chenwei, et al.
Publicado: (2025) -
Estimation of Treatment Harm Rate via Partitioning
por: Liang, Wei, et al.
Publicado: (2025) -
Nonparametric Bounds in Causal Mediation Analysis in the Presence of Unmeasured Confounding and Imperfect Compliance
por: Liang, Wei, et al.
Publicado: (2025) -
AR-VRM: Imitating Human Motions for Visual Robot Manipulation with Analogical Reasoning
por: Yang, Dejie, et al.
Publicado: (2025)