Solving Physics Olympiad via Reinforcement Learning on Physics Simulators
Fuente:
arXiv
Saved in:
| Main Authors: | Prabhudesai, Mihir, Satpathy, Aryan, Li, Yangmin, Qin, Zheyang, Bhardwaj, Nikash, Zadeh, Amir, Li, Chuan, Fragkiadaki, Katerina, Pathak, Deepak |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Iterative Refinement Improves Compositional Image Generation
by: Jaiswal, Shantanu, et al.
Published: (2026)
by: Jaiswal, Shantanu, et al.
Published: (2026)
Video Diffusion Alignment via Reward Gradients
by: Prabhudesai, Mihir, et al.
Published: (2024)
by: Prabhudesai, Mihir, et al.
Published: (2024)
Diffusion Beats Autoregressive in Data-Constrained Settings
by: Prabhudesai, Mihir, et al.
Published: (2025)
by: Prabhudesai, Mihir, et al.
Published: (2025)
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
by: Prabhudesai, Mihir, et al.
Published: (2023)
by: Prabhudesai, Mihir, et al.
Published: (2023)
Unified Multimodal Discrete Diffusion
by: Swerdlow, Alexander, et al.
Published: (2025)
by: Swerdlow, Alexander, et al.
Published: (2025)
Self-Questioning Language Models
by: Chen, Lili, et al.
Published: (2025)
by: Chen, Lili, et al.
Published: (2025)
Maximizing Confidence Alone Improves Reasoning
by: Prabhudesai, Mihir, et al.
Published: (2025)
by: Prabhudesai, Mihir, et al.
Published: (2025)
Learning to Assist: Physics-Grounded Human-Human Control via Multi-Agent Reinforcement Learning
by: Shibata, Yuto, et al.
Published: (2026)
by: Shibata, Yuto, et al.
Published: (2026)
RoboTAG: End-to-end Robot Configuration Estimation via Topological Alignment Graph
by: Liu, Yifan, et al.
Published: (2025)
by: Liu, Yifan, et al.
Published: (2025)
Reinforcement Learning for Collision-free Flight Exploiting Deep Collision Encoding
by: Kulkarni, Mihir, et al.
Published: (2024)
by: Kulkarni, Mihir, et al.
Published: (2024)
3D Diffuser Actor: Policy Diffusion with 3D Scene Representations
by: Ke, Tsung-Wei, et al.
Published: (2024)
by: Ke, Tsung-Wei, et al.
Published: (2024)
Dex4D: Task-Agnostic Point Track Policy for Sim-to-Real Dexterous Manipulation
by: Kuang, Yuxuan, et al.
Published: (2026)
by: Kuang, Yuxuan, et al.
Published: (2026)
Reinforcement Learning for Active Perception in Autonomous Navigation
by: Malczyk, Grzegorz, et al.
Published: (2026)
by: Malczyk, Grzegorz, et al.
Published: (2026)
Semantically-driven Deep Reinforcement Learning for Inspection Path Planning
by: Malczyk, Grzegorz, et al.
Published: (2025)
by: Malczyk, Grzegorz, et al.
Published: (2025)
RoboGen: Towards Unleashing Infinite Data for Automated Robot Learning via Generative Simulation
by: Wang, Yufei, et al.
Published: (2023)
by: Wang, Yufei, et al.
Published: (2023)
3D FlowMatch Actor: Unified 3D Policy for Single- and Dual-Arm Manipulation
by: Gkanatsios, Nikolaos, et al.
Published: (2025)
by: Gkanatsios, Nikolaos, et al.
Published: (2025)
LocoFormer: Generalist Locomotion via Long-context Adaptation
by: Liu, Min, et al.
Published: (2025)
by: Liu, Min, et al.
Published: (2025)
Tractable Joint Prediction and Planning over Discrete Behavior Modes for Urban Driving
by: Villaflor, Adam, et al.
Published: (2024)
by: Villaflor, Adam, et al.
Published: (2024)
Can LLMs Lie? Investigation beyond Hallucination
by: Huan, Haoran, et al.
Published: (2025)
by: Huan, Haoran, et al.
Published: (2025)
Efficient Knowledge Transfer for Jump-Starting Control Policy Learning of Multirotors through Physics-Aware Neural Architectures
by: Rehberg, Welf, et al.
Published: (2026)
by: Rehberg, Welf, et al.
Published: (2026)
Aerial Gym Simulator: A Framework for Highly Parallelized Simulation of Aerial Robots
by: Kulkarni, Mihir, et al.
Published: (2025)
by: Kulkarni, Mihir, et al.
Published: (2025)
Deep Generative Models in Robotics: A Survey on Learning from Multimodal Demonstrations
by: Urain, Julen, et al.
Published: (2024)
by: Urain, Julen, et al.
Published: (2024)
ParkingWorld: End-to-End Autonomous Parking Reinforcement Learning from Corrective Experience in 3DGS Simulation
by: Yu, Zhengcheng, et al.
Published: (2026)
by: Yu, Zhengcheng, et al.
Published: (2026)
Dynamic Non-Prehensile Object Transport via Model-Predictive Reinforcement Learning
by: Jawale, Neel, et al.
Published: (2024)
by: Jawale, Neel, et al.
Published: (2024)
roto 2.0: The Robot Tactile Olympiad
by: Miller, Elle, et al.
Published: (2026)
by: Miller, Elle, et al.
Published: (2026)
AimBot: A Simple Auxiliary Visual Cue to Enhance Spatial Awareness of Visuomotor Policies
by: Dai, Yinpei, et al.
Published: (2025)
by: Dai, Yinpei, et al.
Published: (2025)
Path Planning on Multi-level Point Cloud with a Weighted Traversability Graph
by: Tang, Yujie, et al.
Published: (2025)
by: Tang, Yujie, et al.
Published: (2025)
Generative Multi-Robot Motion Planning via Diffusion Modeling with Multi-Agent Reinforcement Learning Guidance
by: Lee, Suk Ki, et al.
Published: (2026)
by: Lee, Suk Ki, et al.
Published: (2026)
MagBotSim: Physics-Based Simulation and Reinforcement Learning Environments for Magnetic Robotics
by: Bergmann, Lara, et al.
Published: (2025)
by: Bergmann, Lara, et al.
Published: (2025)
Physics-informed Imitative Reinforcement Learning for Real-world Driving
by: Zhou, Hang, et al.
Published: (2024)
by: Zhou, Hang, et al.
Published: (2024)
Adaptive Gain Scheduling using Reinforcement Learning for Quadcopter Control
by: Timmerman, Mike, et al.
Published: (2024)
by: Timmerman, Mike, et al.
Published: (2024)
Universal Dexterous Functional Grasping via Demonstration-Editing Reinforcement Learning
by: Mao, Chuan, et al.
Published: (2025)
by: Mao, Chuan, et al.
Published: (2025)
Reinforcement Learning for Solving Robotic Reaching Tasks in the Neurorobotics Platform
by: Szep, Marton, et al.
Published: (2022)
by: Szep, Marton, et al.
Published: (2022)
MRIC: Model-Based Reinforcement-Imitation Learning with Mixture-of-Codebooks for Autonomous Driving Simulation
by: He, Baotian, et al.
Published: (2024)
by: He, Baotian, et al.
Published: (2024)
Parallel Reinforcement Learning Simulation for Visual Quadrotor Navigation
by: Saunders, Jack, et al.
Published: (2022)
by: Saunders, Jack, et al.
Published: (2022)
Data-Efficient Task Generalization via Probabilistic Model-based Meta Reinforcement Learning
by: Bhardwaj, Arjun, et al.
Published: (2023)
by: Bhardwaj, Arjun, et al.
Published: (2023)
A Survey: Learning Embodied Intelligence from Physical Simulators and World Models
by: Long, Xiaoxiao, et al.
Published: (2025)
by: Long, Xiaoxiao, et al.
Published: (2025)
Solving Robotics Tasks with Prior Demonstration via Exploration-Efficient Deep Reinforcement Learning
by: Shen, Chengyandan, et al.
Published: (2025)
by: Shen, Chengyandan, et al.
Published: (2025)
Generative Simulation for Policy Learning in Physical Human-Robot Interaction
by: Wang, Junxiang, et al.
Published: (2026)
by: Wang, Junxiang, et al.
Published: (2026)
SimLauncher: Launching Sample-Efficient Real-world Robotic Reinforcement Learning via Simulation Pre-training
by: Wu, Mingdong, et al.
Published: (2025)
by: Wu, Mingdong, et al.
Published: (2025)
Similar Items
-
Iterative Refinement Improves Compositional Image Generation
by: Jaiswal, Shantanu, et al.
Published: (2026) -
Video Diffusion Alignment via Reward Gradients
by: Prabhudesai, Mihir, et al.
Published: (2024) -
Diffusion Beats Autoregressive in Data-Constrained Settings
by: Prabhudesai, Mihir, et al.
Published: (2025) -
Aligning Text-to-Image Diffusion Models with Reward Backpropagation
by: Prabhudesai, Mihir, et al.
Published: (2023) -
Unified Multimodal Discrete Diffusion
by: Swerdlow, Alexander, et al.
Published: (2025)