Physics-R1: An Audited Olympiad Corpus and Recipe for Visual Physics Reasoning
Fuente:
arXiv
Guardado en:
| Autor principal: | Yang, Shan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Learning the meanings of function words from grounded language using a visual question answering model
por: Portelance, Eva, et al.
Publicado: (2023)
por: Portelance, Eva, et al.
Publicado: (2023)
Reframing linguistic bootstrapping as joint inference using visually-grounded grammar induction models
por: Portelance, Eva, et al.
Publicado: (2024)
por: Portelance, Eva, et al.
Publicado: (2024)
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
por: Dai, Song, et al.
Publicado: (2025)
por: Dai, Song, et al.
Publicado: (2025)
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
Survey Transfer Learning: Recycling Data with Silicon Responses
por: Amini, Ali
Publicado: (2025)
por: Amini, Ali
Publicado: (2025)
Waking Up Blind: Cold-Start Optimization of Supervision-Free Agentic Trajectories for Grounded Visual Perception
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
por: Bajpai, Ashutosh, et al.
Publicado: (2026)
Game-RL: Synthesizing Multimodal Verifiable Game Data to Boost VLMs' General Reasoning
por: Tong, Jingqi, et al.
Publicado: (2025)
por: Tong, Jingqi, et al.
Publicado: (2025)
A Roadmap for Multilingual, Multimodal Domain Independent Deception Detection
por: Boumber, Dainis, et al.
Publicado: (2024)
por: Boumber, Dainis, et al.
Publicado: (2024)
TensLoRA: Tensor Alternatives for Low-Rank Adaptation
por: Marmoret, Axel, et al.
Publicado: (2025)
por: Marmoret, Axel, et al.
Publicado: (2025)
Character-Level Transformer for Tajik-Persian Transliteration with a Parallel Lexical Corpus
por: Arabov, Mullosharaf K.
Publicado: (2026)
por: Arabov, Mullosharaf K.
Publicado: (2026)
ICG: Improving Cover Image Generation via MLLM-based Prompting and Personalized Preference Alignment
por: Bian, Zhipeng, et al.
Publicado: (2026)
por: Bian, Zhipeng, et al.
Publicado: (2026)
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
por: Masrourisaadat, Nila, et al.
Publicado: (2024)
por: Masrourisaadat, Nila, et al.
Publicado: (2024)
QoSGMAA: A Robust Multi-Order Graph Attention and Adversarial Framework for Sparse QoS Prediction
por: Du, Guanchen, et al.
Publicado: (2025)
por: Du, Guanchen, et al.
Publicado: (2025)
Unpacking Failure Modes of Generative Policies: Runtime Monitoring of Consistency and Progress
por: Agia, Christopher, et al.
Publicado: (2024)
por: Agia, Christopher, et al.
Publicado: (2024)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
por: Shahin, Nada, et al.
Publicado: (2025)
por: Shahin, Nada, et al.
Publicado: (2025)
SALLIE: Safeguarding Against Latent Language & Image Exploits
por: Azov, Guy, et al.
Publicado: (2026)
por: Azov, Guy, et al.
Publicado: (2026)
Spatially-Aware Speaker for Vision-and-Language Navigation Instruction Generation
por: Gopinathan, Muraleekrishna, et al.
Publicado: (2024)
por: Gopinathan, Muraleekrishna, et al.
Publicado: (2024)
D-COT: Disciplined Chain-of-Thought Learning for Efficient Reasoning in Small Language Models
por: Ubukata, Shunsuke
Publicado: (2026)
por: Ubukata, Shunsuke
Publicado: (2026)
Induce, Align, Predict: Zero-Shot Stance Detection via Cognitive Inductive Reasoning
por: Zhang, Bowen, et al.
Publicado: (2025)
por: Zhang, Bowen, et al.
Publicado: (2025)
TRiMS: Real-Time Tracking of Minimal Sufficient Length for Efficient Reasoning via RL
por: Bian, Tingcheng, et al.
Publicado: (2026)
por: Bian, Tingcheng, et al.
Publicado: (2026)
WildRoadBench: A Wild Aerial Road-Damage Grounding Benchmark for Vision-Language Models and Autonomous Agents
por: Liu, Bingnan, et al.
Publicado: (2026)
por: Liu, Bingnan, et al.
Publicado: (2026)
Unpacking Hateful Memes: Presupposed Context and False Claims
por: Cai, Weibin, et al.
Publicado: (2025)
por: Cai, Weibin, et al.
Publicado: (2025)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
por: Fadli, Samih
Publicado: (2025)
por: Fadli, Samih
Publicado: (2025)
Perception-Consistency Multimodal Large Language Models Reasoning via Caption-Regularized Policy Optimization
por: Tu, Songjun, et al.
Publicado: (2025)
por: Tu, Songjun, et al.
Publicado: (2025)
From Benchmarking to Reasoning: A Dual-Aspect, Large-Scale Evaluation of LLMs on Vietnamese Legal Text
por: Le, Van-Truong
Publicado: (2026)
por: Le, Van-Truong
Publicado: (2026)
Aligning by Misaligning: Boundary-aware Curriculum Learning for Multimodal Alignment
por: Ye, Hua, et al.
Publicado: (2025)
por: Ye, Hua, et al.
Publicado: (2025)
Text-to-Events: Synthetic Event Camera Streams from Conditional Text Input
por: Ott, Joachim, et al.
Publicado: (2024)
por: Ott, Joachim, et al.
Publicado: (2024)
PhysNote: Self-Knowledge Notes for Evolvable Physical Reasoning in Vision-Language Model
por: Zhang, Sinin, et al.
Publicado: (2026)
por: Zhang, Sinin, et al.
Publicado: (2026)
Towards Alignment-Centric Paradigm: A Survey of Instruction Tuning in Large Language Models
por: Han, Xudong, et al.
Publicado: (2025)
por: Han, Xudong, et al.
Publicado: (2025)
What Makes a Good Story and How Can We Measure It? A Comprehensive Survey of Story Evaluation
por: Yang, Dingyi, et al.
Publicado: (2024)
por: Yang, Dingyi, et al.
Publicado: (2024)
Enhancing Spatial Reasoning in Vision-Language Models via Chain-of-Thought Prompting and Reinforcement Learning
por: Ji, Binbin, et al.
Publicado: (2025)
por: Ji, Binbin, et al.
Publicado: (2025)
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation
por: Shahin, Nada, et al.
Publicado: (2025)
por: Shahin, Nada, et al.
Publicado: (2025)
Calibrated Confidence Estimation for Tabular Question Answering
por: Voss, Lukas
Publicado: (2026)
por: Voss, Lukas
Publicado: (2026)
Mitigating Cross-Lingual Cultural Inconsistencies in LLMs via Consensus-Driven Preference Optimisation
por: Resck, Lucas, et al.
Publicado: (2026)
por: Resck, Lucas, et al.
Publicado: (2026)
EmoLoom-2B: Fast Base-Model Screening for Emotion Classification and VAD with Lexicon-Weak Supervision and KV-Off Evaluation
por: Li, Zilin, et al.
Publicado: (2026)
por: Li, Zilin, et al.
Publicado: (2026)
Layer-Aware Embedding Fusion for LLMs in Text Classifications
por: Gwak, Jiho, et al.
Publicado: (2025)
por: Gwak, Jiho, et al.
Publicado: (2025)
KSHSeek: Data-Driven Approaches to Mitigating and Detecting Knowledge-Shortcut Hallucinations in Generative Models
por: Liu, Zhongxin, et al.
Publicado: (2025)
por: Liu, Zhongxin, et al.
Publicado: (2025)
On the Influence of Discourse Relations in Persuasive Texts
por: Turk, Nawar, et al.
Publicado: (2025)
por: Turk, Nawar, et al.
Publicado: (2025)
Emergent Lexical Semantics in Neural Language Models: Testing Martin's Law on LLM-Generated Text
por: Kugler, Kai
Publicado: (2025)
por: Kugler, Kai
Publicado: (2025)
Bridging the Gap: An Intermediate Language for Enhanced and Cost-Effective Grapheme-to-Phoneme Conversion with Homographs with Multiple Pronunciations Disambiguation
por: Bertina, Abbas, et al.
Publicado: (2025)
por: Bertina, Abbas, et al.
Publicado: (2025)
Ejemplares similares
-
Learning the meanings of function words from grounded language using a visual question answering model
por: Portelance, Eva, et al.
Publicado: (2023) -
Reframing linguistic bootstrapping as joint inference using visually-grounded grammar induction models
por: Portelance, Eva, et al.
Publicado: (2024) -
PhysicsArena: The First Multimodal Physics Reasoning Benchmark Exploring Variable, Process, and Solution Dimensions
por: Dai, Song, et al.
Publicado: (2025) -
SpatialMath: Spatial Comprehension-Infused Symbolic Reasoning for Mathematical Problem-Solving
por: Bajpai, Ashutosh, et al.
Publicado: (2026) -
Survey Transfer Learning: Recycling Data with Silicon Responses
por: Amini, Ali
Publicado: (2025)