$\pi2\text{vec}$: Policy Representations with Successor Features
Fuente:
arXiv
Saved in:
| Main Authors: | Scarpellini, Gianluca, Konyushkova, Ksenia, Fantacci, Claudio, Paine, Tom Le, Chen, Yutian, Denil, Misha |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vision-Language Model Dialog Games for Self-Improvement
by: Konyushkova, Ksenia, et al.
Published: (2025)
by: Konyushkova, Ksenia, et al.
Published: (2025)
Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following
by: Myers, Vivek, et al.
Published: (2025)
by: Myers, Vivek, et al.
Published: (2025)
Rapid Object Annotation
by: Denil, Misha
Published: (2024)
by: Denil, Misha
Published: (2024)
GATS: Gather-Attend-Scatter
by: Zolna, Konrad, et al.
Published: (2024)
by: Zolna, Konrad, et al.
Published: (2024)
Full-Gradient Successor Feature Representations
by: Shrirao, Ritish, et al.
Published: (2026)
by: Shrirao, Ritish, et al.
Published: (2026)
Fast Feature Field ($\text{F}^3$): A Predictive Representation of Events
by: Das, Richeek, et al.
Published: (2025)
by: Das, Richeek, et al.
Published: (2025)
Focusing Robot Open-Ended Reinforcement Learning Through Users' Purposes
by: Cartoni, Emilio, et al.
Published: (2025)
by: Cartoni, Emilio, et al.
Published: (2025)
Decoding High-Dimensional Finger Motion from EMG Using Riemannian Features and RNNs
by: Colot, Martin, et al.
Published: (2026)
by: Colot, Martin, et al.
Published: (2026)
SPLASH! Sample-efficient Preference-based inverse reinforcement learning for Long-horizon Adversarial tasks from Suboptimal Hierarchical demonstrations
by: Crowley, Peter, et al.
Published: (2025)
by: Crowley, Peter, et al.
Published: (2025)
AirIO: Learning Inertial Odometry with Enhanced IMU Feature Observability
by: Qiu, Yuheng, et al.
Published: (2025)
by: Qiu, Yuheng, et al.
Published: (2025)
Learning Safe Autonomous Driving Policies Using Predictive Safety Representations
by: Keswani, Mahesh, et al.
Published: (2025)
by: Keswani, Mahesh, et al.
Published: (2025)
Foundation Policies with Hilbert Representations
by: Park, Seohong, et al.
Published: (2024)
by: Park, Seohong, et al.
Published: (2024)
Distributional Successor Features Enable Zero-Shot Policy Optimization
by: Zhu, Chuning, et al.
Published: (2024)
by: Zhu, Chuning, et al.
Published: (2024)
Language-Conditioned Representations and Mixture-of-Experts Policy for Robust Multi-Task Robotic Manipulation
by: Zhang, Xiucheng, et al.
Published: (2025)
by: Zhang, Xiucheng, et al.
Published: (2025)
What is the relation between Slow Feature Analysis and the Successor Representation?
by: Seabrook, Eddie, et al.
Published: (2024)
by: Seabrook, Eddie, et al.
Published: (2024)
Ethics2vec: aligning automatic agents and human preferences
by: Bontempi, Gianluca
Published: (2025)
by: Bontempi, Gianluca
Published: (2025)
Diffusion Policy through Conditional Proximal Policy Optimization
by: Liu, Ben, et al.
Published: (2026)
by: Liu, Ben, et al.
Published: (2026)
Time-Unified Diffusion Policy with Action Discrimination for Robotic Manipulation
by: Niu, Ye, et al.
Published: (2025)
by: Niu, Ye, et al.
Published: (2025)
PolicyFlow: Policy Optimization with Continuous Normalizing Flow in Reinforcement Learning
by: Yang, Shunpeng, et al.
Published: (2026)
by: Yang, Shunpeng, et al.
Published: (2026)
Practice Makes Perfect: Planning to Learn Skill Parameter Policies
by: Kumar, Nishanth, et al.
Published: (2024)
by: Kumar, Nishanth, et al.
Published: (2024)
DynamicDTA: Drug-Target Binding Affinity Prediction Using Dynamic Descriptors and Graph Representation
by: Luo, Dan, et al.
Published: (2025)
by: Luo, Dan, et al.
Published: (2025)
Diffusion Policy Policy Optimization
by: Ren, Allen Z., et al.
Published: (2024)
by: Ren, Allen Z., et al.
Published: (2024)
State-wise Constrained Policy Optimization
by: Zhao, Weiye, et al.
Published: (2023)
by: Zhao, Weiye, et al.
Published: (2023)
OmniGuide: Universal Guidance Fields for Enhancing Generalist Robot Policies
by: Song, Yunzhou, et al.
Published: (2026)
by: Song, Yunzhou, et al.
Published: (2026)
State Chrono Representation for Enhancing Generalization in Reinforcement Learning
by: Chen, Jianda, et al.
Published: (2024)
by: Chen, Jianda, et al.
Published: (2024)
Which Features are Best for Successor Features?
by: Ollivier, Yann
Published: (2025)
by: Ollivier, Yann
Published: (2025)
Dichotomous Diffusion Policy Optimization
by: Liang, Ruiming, et al.
Published: (2025)
by: Liang, Ruiming, et al.
Published: (2025)
Learning Successor Features the Simple Way
by: Chua, Raymond, et al.
Published: (2024)
by: Chua, Raymond, et al.
Published: (2024)
FDPP: Fine-tune Diffusion Policy with Human Preference
by: Chen, Yuxin, et al.
Published: (2025)
by: Chen, Yuxin, et al.
Published: (2025)
Learning Getting-Up Policies for Real-World Humanoid Robots
by: He, Xialin, et al.
Published: (2025)
by: He, Xialin, et al.
Published: (2025)
Hierarchical Successor Representation for Robust Transfer
by: Yu, Changmin, et al.
Published: (2026)
by: Yu, Changmin, et al.
Published: (2026)
Latent Policy Steering through One-Step Flow Policies
by: Im, Hokyun, et al.
Published: (2026)
by: Im, Hokyun, et al.
Published: (2026)
Residual Off-Policy RL for Finetuning Behavior Cloning Policies
by: Ankile, Lars, et al.
Published: (2025)
by: Ankile, Lars, et al.
Published: (2025)
PEEK: Guiding and Minimal Image Representations for Zero-Shot Generalization of Robot Manipulation Policies
by: Zhang, Jesse, et al.
Published: (2025)
by: Zhang, Jesse, et al.
Published: (2025)
Learning Visual Feature-Based World Models via Residual Latent Action
by: Zhang, Xinyu, et al.
Published: (2026)
by: Zhang, Xinyu, et al.
Published: (2026)
Equivariant Diffusion Policy
by: Wang, Dian, et al.
Published: (2024)
by: Wang, Dian, et al.
Published: (2024)
BiKC: Keypose-Conditioned Consistency Policy for Bimanual Robotic Manipulation
by: Yu, Dongjie, et al.
Published: (2024)
by: Yu, Dongjie, et al.
Published: (2024)
Benchmarking Smoothness and Reducing High-Frequency Oscillations in Continuous Control Policies
by: Christmann, Guilherme, et al.
Published: (2024)
by: Christmann, Guilherme, et al.
Published: (2024)
DreamPolicy: A Unified World-model Policy for Scalable Humanoid Locomotion
by: Fan, Yahao, et al.
Published: (2025)
by: Fan, Yahao, et al.
Published: (2025)
Sparse Diffusion Policy: A Sparse, Reusable, and Flexible Policy for Robot Learning
by: Wang, Yixiao, et al.
Published: (2024)
by: Wang, Yixiao, et al.
Published: (2024)
Similar Items
-
Vision-Language Model Dialog Games for Self-Improvement
by: Konyushkova, Ksenia, et al.
Published: (2025) -
Temporal Representation Alignment: Successor Features Enable Emergent Compositionality in Robot Instruction Following
by: Myers, Vivek, et al.
Published: (2025) -
Rapid Object Annotation
by: Denil, Misha
Published: (2024) -
GATS: Gather-Attend-Scatter
by: Zolna, Konrad, et al.
Published: (2024) -
Full-Gradient Successor Feature Representations
by: Shrirao, Ritish, et al.
Published: (2026)