SeePhys Pro: Diagnosing Modality Transfer and Blind-Training Effects in Multimodal RLVR for Physics Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xiang, Kun, Zhang, Terry Jingchen, Liu, Zirong, Zhou, Bokai, Tang, Yueling, Yu, Junjie, Lu, Jiacong, Huang, Shangrui, Li, Heng, Zhang, Likui, Liu, Kunkun, Zhang, Changzheng, Fang, Yangle, Guo, Boqiang, Zhen, Hui-Ling, Tu, Dandan, Huang, Yinya, Liang, Xiaodan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SeePhys: Does Seeing Help Thinking? -- Benchmarking Vision-Based Physics Reasoning
von: Xiang, Kun, et al.
Veröffentlicht: (2025)
von: Xiang, Kun, et al.
Veröffentlicht: (2025)
Multimodal Reasoning for Science: Technical Report and 1st Place Solution to the ICML 2025 SeePhys Challenge
von: Liang, Hao, et al.
Veröffentlicht: (2025)
von: Liang, Hao, et al.
Veröffentlicht: (2025)
Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI
von: Xiang, Kun, et al.
Veröffentlicht: (2025)
von: Xiang, Kun, et al.
Veröffentlicht: (2025)
Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
ProPhy: Progressive Physical Alignment for Dynamic World Simulation
von: Wang, Zijun, et al.
Veröffentlicht: (2025)
von: Wang, Zijun, et al.
Veröffentlicht: (2025)
AtomThink: Multimodal Slow Thinking with Atomic Step Reasoning
von: Xiang, Kun, et al.
Veröffentlicht: (2024)
von: Xiang, Kun, et al.
Veröffentlicht: (2024)
Climate risk and green total factor productivity in agriculture: The moderating role of climate policy uncertainty
von: Miao Wang, et al.
Veröffentlicht: (2024)
von: Miao Wang, et al.
Veröffentlicht: (2024)
Educators' Perceptions of Large Language Models as Tutors: Comparing Human and AI Tutors in a Blind Text-only Setting
von: Chowdhury, Sankalan Pal, et al.
Veröffentlicht: (2025)
von: Chowdhury, Sankalan Pal, et al.
Veröffentlicht: (2025)
ATG: Benchmarking Automated Theorem Generation for Generative Language Models
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
TreeRPO: Tree Relative Policy Optimization
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
Recycling Failures: Salvaging Exploration in RLVR via Fine-Grained Off-Policy Guidance
von: Ren, Yanwei, et al.
Veröffentlicht: (2026)
von: Ren, Yanwei, et al.
Veröffentlicht: (2026)
Robust H2 and H∞ Filter Design for Polytopic Systems by Dilating Matrices
von: Yuefeng Cui, et al.
Veröffentlicht: (2025)
von: Yuefeng Cui, et al.
Veröffentlicht: (2025)
Test of Time: Rethinking Temporal Signal of Benchmark Contamination
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)
Stargazer: A Scalable Model-Fitting Benchmark Environment for AI Agents under Astrophysical Constraints
von: Liu, Xinge, et al.
Veröffentlicht: (2026)
von: Liu, Xinge, et al.
Veröffentlicht: (2026)
CLOMO: Counterfactual Logical Modification with Large Language Models
von: Huang, Yinya, et al.
Veröffentlicht: (2023)
von: Huang, Yinya, et al.
Veröffentlicht: (2023)
Exploring Multi-Temperature Strategies for Token- and Rollout-Level Control in RLVR
von: Zhuang, Haomin, et al.
Veröffentlicht: (2025)
von: Zhuang, Haomin, et al.
Veröffentlicht: (2025)
Endo-G$^{2}$T: Geometry-Guided & Temporally Aware Time-Embedded 4DGS For Endoscopic Scenes
von: Liu, Yangle, et al.
Veröffentlicht: (2025)
von: Liu, Yangle, et al.
Veröffentlicht: (2025)
Quantile Advantage Estimation: Stabilizing RLVR for LLM Reasoning
von: Wu, Junkang, et al.
Veröffentlicht: (2025)
von: Wu, Junkang, et al.
Veröffentlicht: (2025)
EAST: Environment-Aware Stylized Transition Along the Reality-Virtuality Continuum
von: Zhang, Xiaohan, et al.
Veröffentlicht: (2025)
von: Zhang, Xiaohan, et al.
Veröffentlicht: (2025)
Multifractal analysis of the convergence exponents for the digits in $d$-decaying Gauss like dynamical systems
von: Song, Kunkun, et al.
Veröffentlicht: (2024)
von: Song, Kunkun, et al.
Veröffentlicht: (2024)
PhysGame: Uncovering Physical Commonsense Violations in Gameplay Videos
von: Cao, Meng, et al.
Veröffentlicht: (2024)
von: Cao, Meng, et al.
Veröffentlicht: (2024)
FVEL: Interactive Formal Verification Environment with Large Language Models via Theorem Proving
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
von: Lin, Xiaohan, et al.
Veröffentlicht: (2024)
Reducing Differential Item Functioning via Process Data
von: Chen, Ling, et al.
Veröffentlicht: (2025)
von: Chen, Ling, et al.
Veröffentlicht: (2025)
An event‐triggered zonotopic Gaussian state estimator for discrete‐time stochastic multiplicative systems with zonotopic set‐membership and stochastic uncertainties
von: Ye Chen, et al.
Veröffentlicht: (2024)
von: Ye Chen, et al.
Veröffentlicht: (2024)
ANCoEF: Asynchronous Neuromorphic Algorithm/Hardware Co-Exploration Framework with a Fully Asynchronous Simulator
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
ArcPro: Architectural Programs for Structured 3D Abstraction of Sparse Points
von: Huang, Qirui, et al.
Veröffentlicht: (2025)
von: Huang, Qirui, et al.
Veröffentlicht: (2025)
ORMind: A Cognitive-Inspired End-to-End Reasoning Framework for Operations Research
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2025)
von: Wang, Zhiyuan, et al.
Veröffentlicht: (2025)
AlignedCoT: Prompting Large Language Models via Native-Speaking Demonstrations
von: Yang, Zhicheng, et al.
Veröffentlicht: (2023)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2023)
Critique to Verify: Accurate and Honest Test-Time Scaling with RL-Trained Verifiers
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025)
SEED: A Structural Encoder for Embedding-Driven Decoding in Time Series Prediction with LLMs
von: Li, Fengze, et al.
Veröffentlicht: (2025)
von: Li, Fengze, et al.
Veröffentlicht: (2025)
Orchestrating Tokens and Sequences: Dynamic Hybrid Policy Optimization for RLVR
von: Min, Zijun, et al.
Veröffentlicht: (2026)
von: Min, Zijun, et al.
Veröffentlicht: (2026)
Does LLM Alignment Really Need Diversity? An Empirical Study of Adapting RLVR Methods for Moral Reasoning
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2026)
von: Zhang, Zhaowei, et al.
Veröffentlicht: (2026)
Open-Medical-R1: How to Choose Data for RLVR Training at Medicine Domain
von: Qiu, Zhongxi, et al.
Veröffentlicht: (2025)
von: Qiu, Zhongxi, et al.
Veröffentlicht: (2025)
The Suture‐In‐Needle, Closed‐Loop Technique for Repositioning a Dislocated Akreos Adapt Intraocular Lens
von: Jingjing Zhang, et al.
Veröffentlicht: (2025)
von: Jingjing Zhang, et al.
Veröffentlicht: (2025)
MUSTARD: Mastering Uniform Synthesis of Theorem and Proof Data
von: Huang, Yinya, et al.
Veröffentlicht: (2024)
von: Huang, Yinya, et al.
Veröffentlicht: (2024)
The Impact of Economic Growth Targets on Environmental Pollution: A Study From Chinese Cities
von: Yicheng Zhou, et al.
Veröffentlicht: (2024)
von: Yicheng Zhou, et al.
Veröffentlicht: (2024)
GT-HarmBench: Benchmarking AI Safety Risks Through the Lens of Game Theory
von: Cobben, Pepijn, et al.
Veröffentlicht: (2026)
von: Cobben, Pepijn, et al.
Veröffentlicht: (2026)
MEL: Efficient Multi-Task Evolutionary Learning for High-Dimensional Feature Selection
von: Wang, Xubin, et al.
Veröffentlicht: (2024)
von: Wang, Xubin, et al.
Veröffentlicht: (2024)
LvHcS52, a Litopenaeus vannamei hemocyanin-derived peptide, restricts WSSV infection by promoting phagocytosis and activating the STAT signaling pathway.
von: Zhan, Shixiong, et al.
Veröffentlicht: (2025)
von: Zhan, Shixiong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SeePhys: Does Seeing Help Thinking? -- Benchmarking Vision-Based Physics Reasoning
von: Xiang, Kun, et al.
Veröffentlicht: (2025) -
Multimodal Reasoning for Science: Technical Report and 1st Place Solution to the ICML 2025 SeePhys Challenge
von: Liang, Hao, et al.
Veröffentlicht: (2025) -
Aligning Perception, Reasoning, Modeling and Interaction: A Survey on Physical AI
von: Xiang, Kun, et al.
Veröffentlicht: (2025) -
Depth-Breadth Synergy in RLVR: Unlocking LLM Reasoning Gains with Adaptive Exploration
von: Yang, Zhicheng, et al.
Veröffentlicht: (2025) -
Lean Meets Theoretical Computer Science: Scalable Synthesis of Theorem Proving Challenges in Formal-Informal Pairs
von: Zhang, Terry Jingchen, et al.
Veröffentlicht: (2025)