OGPO: Sample Efficient Full-Finetuning of Generative Control Policies
Fuente:
arXiv
Salvato in:
| Autori principali: | Patil, Sarvesh, Nakamoto, Mitsuhiko, Agarwal, Manan, Saxena, Shashwat, Zhang, Jesse, Anantharaman, Giri, Winston, Cleah, Pan, Chaoyi, Chen, Douglas, Huang, Nai-Chieh, Temel, Zeynep, Kroemer, Oliver, Levine, Sergey, Gupta, Abhishek, Da, Hongkai, Shah, Paarth, Simchowitz, Max |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
From Fold to Function: Simulation-Driven Design of Origami Mechanisms
di: Han, Tianhui, et al.
Pubblicazione: (2025)
di: Han, Tianhui, et al.
Pubblicazione: (2025)
PuffyBot: An Untethered Shape Morphing Robot for Multi-environment Locomotion
di: Singh, Shashwat, et al.
Pubblicazione: (2025)
di: Singh, Shashwat, et al.
Pubblicazione: (2025)
Much Ado About Noising: Dispelling the Myths of Generative Robotic Control
di: Pan, Chaoyi, et al.
Pubblicazione: (2025)
di: Pan, Chaoyi, et al.
Pubblicazione: (2025)
Tilde: Teleoperation for Dexterous In-Hand Manipulation Learning with a DeltaHand
di: Si, Zilin, et al.
Pubblicazione: (2024)
di: Si, Zilin, et al.
Pubblicazione: (2024)
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
di: Nakamoto, Mitsuhiko, et al.
Pubblicazione: (2024)
di: Nakamoto, Mitsuhiko, et al.
Pubblicazione: (2024)
A multi-modal tactile fingertip design for robotic hands to enhance dexterous manipulation
di: Xu, Zhuowei, et al.
Pubblicazione: (2025)
di: Xu, Zhuowei, et al.
Pubblicazione: (2025)
TerraSkipper: A Centimeter-Scale Robot for Multi-Terrain Skipping and Crawling
di: Singh, Shashwat, et al.
Pubblicazione: (2026)
di: Singh, Shashwat, et al.
Pubblicazione: (2026)
Diamond Maps: Efficient Reward Alignment via Stochastic Flow Maps
di: Holderrieth, Peter, et al.
Pubblicazione: (2026)
di: Holderrieth, Peter, et al.
Pubblicazione: (2026)
MResT: Multi-Resolution Sensing for Real-Time Control with Vision-Language Models
di: Saxena, Saumya, et al.
Pubblicazione: (2024)
di: Saxena, Saumya, et al.
Pubblicazione: (2024)
Reconfigurable Robot Control Using Flexible Coupling Mechanisms
di: Yi, Sha, et al.
Pubblicazione: (2023)
di: Yi, Sha, et al.
Pubblicazione: (2023)
Steering Your Diffusion Policy with Latent Space Reinforcement Learning
di: Wagenmaker, Andrew, et al.
Pubblicazione: (2025)
di: Wagenmaker, Andrew, et al.
Pubblicazione: (2025)
Using Non-Expert Data to Robustify Imitation Learning via Offline Reinforcement Learning
di: Huang, Kevin, et al.
Pubblicazione: (2025)
di: Huang, Kevin, et al.
Pubblicazione: (2025)
Action Chunking and Exploratory Data Collection Yield Exponential Improvements in Behavior Cloning for Continuous Control
di: Zhang, Thomas T., et al.
Pubblicazione: (2025)
di: Zhang, Thomas T., et al.
Pubblicazione: (2025)
Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning
di: Nakamoto, Mitsuhiko, et al.
Pubblicazione: (2023)
di: Nakamoto, Mitsuhiko, et al.
Pubblicazione: (2023)
Comparison of CO‐OP and goal‐directed training on occupational performance and functional status in children with cerebral palsy: Three‐armed randomised trial
di: Zeynep Kolit, et al.
Pubblicazione: (2025)
di: Zeynep Kolit, et al.
Pubblicazione: (2025)
Hodoscope: Unsupervised Monitoring for AI Misbehaviors
di: Zhong, Ziqian, et al.
Pubblicazione: (2026)
di: Zhong, Ziqian, et al.
Pubblicazione: (2026)
Predicting Emergent Capabilities by Finetuning
di: Snell, Charlie, et al.
Pubblicazione: (2024)
di: Snell, Charlie, et al.
Pubblicazione: (2024)
CAPITAL BATTLES , HABITUS AND SPACE : Urban Segregation in Bursa, Türkiye
di: Kemal Temel
Pubblicazione: (2026)
di: Kemal Temel
Pubblicazione: (2026)
The Impact of Cooperation on Firms' Innovation Propensity in Emerging Economies
di: Serdal Temel
Pubblicazione: (2013)
di: Serdal Temel
Pubblicazione: (2013)
Scalable Inference for Bayesian Multinomial Logistic-Normal Dynamic Linear Models
di: Saxena, Manan, et al.
Pubblicazione: (2024)
di: Saxena, Manan, et al.
Pubblicazione: (2024)
Against Public‐Facing Religious Bio‐Restrictionism
di: Muralidharan Anantharaman
Pubblicazione: (2025)
di: Muralidharan Anantharaman
Pubblicazione: (2025)
Non-Euclidean elasticity for rods and almost isometric embeddings of geodesic tubes
di: Kroemer, Milan, et al.
Pubblicazione: (2025)
di: Kroemer, Milan, et al.
Pubblicazione: (2025)
ETHER: Efficient Finetuning of Large-Scale Models with Hyperplane Reflections
di: Bini, Massimo, et al.
Pubblicazione: (2024)
di: Bini, Massimo, et al.
Pubblicazione: (2024)
SQUIDs for detection of potential dark matter candidates
di: Sivakumar, Siddarth, et al.
Pubblicazione: (2024)
di: Sivakumar, Siddarth, et al.
Pubblicazione: (2024)
Momentum Streams for Optimizer-Inspired Transformers
di: Gai, Jingchu, et al.
Pubblicazione: (2026)
di: Gai, Jingchu, et al.
Pubblicazione: (2026)
SparseContrast: Dynamic Sparse Attention for Efficient and Accurate Contrastive Learning in Medical Imaging
di: Prasad, Paarth, et al.
Pubblicazione: (2026)
di: Prasad, Paarth, et al.
Pubblicazione: (2026)
Topology-Constrained Quantized nnUNet for Efficient and Anatomically Accurate 3D Tooth Segmentation
di: Prasad, Paarth, et al.
Pubblicazione: (2026)
di: Prasad, Paarth, et al.
Pubblicazione: (2026)
Rational terms of UV origin to all loop orders
di: Duhr, Claude, et al.
Pubblicazione: (2023)
di: Duhr, Claude, et al.
Pubblicazione: (2023)
Unfamiliar Finetuning Examples Control How Language Models Hallucinate
di: Kang, Katie, et al.
Pubblicazione: (2024)
di: Kang, Katie, et al.
Pubblicazione: (2024)
Leveraging Simulation-Based Model Preconditions for Fast Action Parameter Optimization with Multiple Models
di: Seker, M. Yunus, et al.
Pubblicazione: (2024)
di: Seker, M. Yunus, et al.
Pubblicazione: (2024)
A Long Way From Home: A Rare Case of Cutaneous Metastasis to the Scalp of Hepatocellular Carcinoma
di: Evan Eggiman, et al.
Pubblicazione: (2024)
di: Evan Eggiman, et al.
Pubblicazione: (2024)
Groupoid exactness and the weak containment problem
di: Anantharaman-Delaroche, Claire
Pubblicazione: (2016)
di: Anantharaman-Delaroche, Claire
Pubblicazione: (2016)
SkinGrip: An Adaptive Soft Robotic Manipulator with Capacitive Sensing for Whole-Limb Bed Bathing Assistance
di: Liu, Fukang, et al.
Pubblicazione: (2024)
di: Liu, Fukang, et al.
Pubblicazione: (2024)
Cross‐cultural adaptation and psychometric evaluation of the Turkish version of the Smombie Scale for Adolescents
di: Ebru Sönmez Sari, et al.
Pubblicazione: (2024)
di: Ebru Sönmez Sari, et al.
Pubblicazione: (2024)
SAM Fewshot Finetuning for Anatomical Segmentation in Medical Images
di: Xie, Weiyi, et al.
Pubblicazione: (2024)
di: Xie, Weiyi, et al.
Pubblicazione: (2024)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
di: Wagenmaker, Andrew, et al.
Pubblicazione: (2025)
di: Wagenmaker, Andrew, et al.
Pubblicazione: (2025)
Numerical analysis of debonding behavior of FRP sheets bonded to concrete focusing on mortar skin
di: Mitsuhiko Ozaki, et al.
Pubblicazione: (2025)
di: Mitsuhiko Ozaki, et al.
Pubblicazione: (2025)
Rational Design Strategies for Stimuli‐Responsive DNAzymes Using Modified and Artificial Nucleotides
di: Yusuke Takezawa, et al.
Pubblicazione: (2026)
di: Yusuke Takezawa, et al.
Pubblicazione: (2026)
Inducing Robustness in a 2 Dimensional Direct Preference Optimization Paradigm
di: Shashidhar, Sarvesh, et al.
Pubblicazione: (2025)
di: Shashidhar, Sarvesh, et al.
Pubblicazione: (2025)
The Pitfalls of Imitation Learning when Actions are Continuous
di: Simchowitz, Max, et al.
Pubblicazione: (2025)
di: Simchowitz, Max, et al.
Pubblicazione: (2025)
Documenti analoghi
-
From Fold to Function: Simulation-Driven Design of Origami Mechanisms
di: Han, Tianhui, et al.
Pubblicazione: (2025) -
PuffyBot: An Untethered Shape Morphing Robot for Multi-environment Locomotion
di: Singh, Shashwat, et al.
Pubblicazione: (2025) -
Much Ado About Noising: Dispelling the Myths of Generative Robotic Control
di: Pan, Chaoyi, et al.
Pubblicazione: (2025) -
Tilde: Teleoperation for Dexterous In-Hand Manipulation Learning with a DeltaHand
di: Si, Zilin, et al.
Pubblicazione: (2024) -
Steering Your Generalists: Improving Robotic Foundation Models via Value Guidance
di: Nakamoto, Mitsuhiko, et al.
Pubblicazione: (2024)