Universal Neural Functionals
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhou, Allan, Finn, Chelsea, Harrison, James |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
TQL: Scaling Q-Functions with Transformers by Preventing Attention Collapse
por: Dong, Perry, et al.
Publicado: (2026)
por: Dong, Perry, et al.
Publicado: (2026)
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
por: Hsu, Sheryl, et al.
Publicado: (2024)
por: Hsu, Sheryl, et al.
Publicado: (2024)
Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning
por: Xie, Johnathan, et al.
Publicado: (2024)
por: Xie, Johnathan, et al.
Publicado: (2024)
EXPO: Stable Reinforcement Learning with Expressive Policies
por: Dong, Perry, et al.
Publicado: (2025)
por: Dong, Perry, et al.
Publicado: (2025)
FASTER: Value-Guided Sampling for Fast RL
por: Dong, Perry, et al.
Publicado: (2026)
por: Dong, Perry, et al.
Publicado: (2026)
Reinforcement Learning via Implicit Imitation Guidance
por: Dong, Perry, et al.
Publicado: (2025)
por: Dong, Perry, et al.
Publicado: (2025)
MemER: Scaling Up Memory for Robot Control via Experience Retrieval
por: Sridhar, Ajay, et al.
Publicado: (2025)
por: Sridhar, Ajay, et al.
Publicado: (2025)
Learning Long-Context Diffusion Policies via Past-Token Prediction
por: Torne, Marcel, et al.
Publicado: (2025)
por: Torne, Marcel, et al.
Publicado: (2025)
Value Flows
por: Dong, Perry, et al.
Publicado: (2025)
por: Dong, Perry, et al.
Publicado: (2025)
Curating Demonstrations using Online Experience
por: Chen, Annie S., et al.
Publicado: (2025)
por: Chen, Annie S., et al.
Publicado: (2025)
Efficient Data Collection for Robotic Manipulation via Compositional Generalization
por: Gao, Jensen, et al.
Publicado: (2024)
por: Gao, Jensen, et al.
Publicado: (2024)
Posterior Behavioral Cloning: Pretraining BC Policies for Efficient RL Finetuning
por: Wagenmaker, Andrew, et al.
Publicado: (2025)
por: Wagenmaker, Andrew, et al.
Publicado: (2025)
Polychromic Objectives for Reinforcement Learning
por: Hamid, Jubayer Ibn, et al.
Publicado: (2025)
por: Hamid, Jubayer Ibn, et al.
Publicado: (2025)
Conservative Prediction via Data-Driven Confidence Minimization
por: Choi, Caroline, et al.
Publicado: (2023)
por: Choi, Caroline, et al.
Publicado: (2023)
Affordance-Guided Reinforcement Learning via Visual Prompting
por: Lee, Olivia Y., et al.
Publicado: (2024)
por: Lee, Olivia Y., et al.
Publicado: (2024)
Agent Q: Advanced Reasoning and Learning for Autonomous AI Agents
por: Putta, Pranav, et al.
Publicado: (2024)
por: Putta, Pranav, et al.
Publicado: (2024)
Just Enough Thinking: Efficient Reasoning with Adaptive Length Penalties Reinforcement Learning
por: Xiang, Violet, et al.
Publicado: (2025)
por: Xiang, Violet, et al.
Publicado: (2025)
Clarify: Improving Model Robustness With Natural Language Corrections
por: Lee, Yoonho, et al.
Publicado: (2024)
por: Lee, Yoonho, et al.
Publicado: (2024)
Calibrating Language Models with Adaptive Temperature Scaling
por: Xie, Johnathan, et al.
Publicado: (2024)
por: Xie, Johnathan, et al.
Publicado: (2024)
Contrastive Preference Learning: Learning from Human Feedback without RL
por: Hejna, Joey, et al.
Publicado: (2023)
por: Hejna, Joey, et al.
Publicado: (2023)
Fine-Tuning Vision-Language-Action Models: Optimizing Speed and Success
por: Kim, Moo Jin, et al.
Publicado: (2025)
por: Kim, Moo Jin, et al.
Publicado: (2025)
MOTO: Offline Pre-training to Online Fine-tuning for Model-based Robot Learning
por: Rafailov, Rafael, et al.
Publicado: (2024)
por: Rafailov, Rafael, et al.
Publicado: (2024)
Bidirectional Decoding: Improving Action Chunking via Guided Test-Time Sampling
por: Liu, Yuejiang, et al.
Publicado: (2024)
por: Liu, Yuejiang, et al.
Publicado: (2024)
A Critical Evaluation of AI Feedback for Aligning Large Language Models
por: Sharma, Archit, et al.
Publicado: (2024)
por: Sharma, Archit, et al.
Publicado: (2024)
Policy Agnostic RL: Offline RL and Online RL Fine-Tuning of Any Class and Backbone
por: Mark, Max Sobol, et al.
Publicado: (2024)
por: Mark, Max Sobol, et al.
Publicado: (2024)
Cal-QL: Calibrated Offline RL Pre-Training for Efficient Online Fine-Tuning
por: Nakamoto, Mitsuhiko, et al.
Publicado: (2023)
por: Nakamoto, Mitsuhiko, et al.
Publicado: (2023)
Direct Preference Optimization: Your Language Model is Secretly a Reward Model
por: Rafailov, Rafael, et al.
Publicado: (2023)
por: Rafailov, Rafael, et al.
Publicado: (2023)
Deriving Neural Scaling Laws from the statistics of natural language
por: Cagnetta, Francesco, et al.
Publicado: (2026)
por: Cagnetta, Francesco, et al.
Publicado: (2026)
Target-Aligned Reinforcement Learning
por: Pleiss, Leonard S., et al.
Publicado: (2026)
por: Pleiss, Leonard S., et al.
Publicado: (2026)
Mobile ALOHA: Learning Bimanual Mobile Manipulation with Low-Cost Whole-Body Teleoperation
por: Fu, Zipeng, et al.
Publicado: (2024)
por: Fu, Zipeng, et al.
Publicado: (2024)
Neural Green's Functions
por: Yoo, Seungwoo, et al.
Publicado: (2025)
por: Yoo, Seungwoo, et al.
Publicado: (2025)
RLVF: Learning from Verbal Feedback without Overgeneralization
por: Stephan, Moritz, et al.
Publicado: (2024)
por: Stephan, Moritz, et al.
Publicado: (2024)
RLAD: Training LLMs to Discover Abstractions for Solving Reasoning Problems
por: Qu, Yuxiao, et al.
Publicado: (2025)
por: Qu, Yuxiao, et al.
Publicado: (2025)
Universal Value-Function Uncertainties
por: Zanger, Moritz A., et al.
Publicado: (2025)
por: Zanger, Moritz A., et al.
Publicado: (2025)
Yell At Your Robot: Improving On-the-Fly from Language Corrections
por: Shi, Lucy Xiaoyang, et al.
Publicado: (2024)
por: Shi, Lucy Xiaoyang, et al.
Publicado: (2024)
World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry
por: Liu, Yuejiang, et al.
Publicado: (2026)
por: Liu, Yuejiang, et al.
Publicado: (2026)
Towards Universal Neural Likelihood Inference
por: Brahmavar, Shreyas Bhat, et al.
Publicado: (2025)
por: Brahmavar, Shreyas Bhat, et al.
Publicado: (2025)
Helpful DoggyBot: Open-World Object Fetching using Legged Robots and Vision-Language Models
por: Wu, Qi, et al.
Publicado: (2024)
por: Wu, Qi, et al.
Publicado: (2024)
Neural Functions for Learning Periodic Signal
por: Cho, Woojin, et al.
Publicado: (2025)
por: Cho, Woojin, et al.
Publicado: (2025)
Scaling Laws for Reward Model Overoptimization in Direct Alignment Algorithms
por: Rafailov, Rafael, et al.
Publicado: (2024)
por: Rafailov, Rafael, et al.
Publicado: (2024)
Ejemplares similares
-
TQL: Scaling Q-Functions with Transformers by Preventing Attention Collapse
por: Dong, Perry, et al.
Publicado: (2026) -
Grounding by Trying: LLMs with Reinforcement Learning-Enhanced Retrieval
por: Hsu, Sheryl, et al.
Publicado: (2024) -
Self-Guided Masked Autoencoders for Domain-Agnostic Self-Supervised Learning
por: Xie, Johnathan, et al.
Publicado: (2024) -
EXPO: Stable Reinforcement Learning with Expressive Policies
por: Dong, Perry, et al.
Publicado: (2025) -
FASTER: Value-Guided Sampling for Fast RL
por: Dong, Perry, et al.
Publicado: (2026)