Leveraging Human Feedback for Semantically-Relevant Skill Discovery
Fuente:
arXiv
Guardado en:
| Autores principales: | Hussonnois, Maxence, Karimpanal, Thommen George, Rana, Santu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
por: Hussonnois, Maxence, et al.
Publicado: (2025)
por: Hussonnois, Maxence, et al.
Publicado: (2025)
MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2025)
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2025)
Dynamic Policy Fusion for User Alignment Without Re-Interaction
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2024)
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2024)
ASPECT:Analogical Semantic Policy Execution via Language Conditioned Transfer
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2026)
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2026)
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)
ECoDe: A Sample-Efficient Method for Co-Design of Robotic Agents
por: Nagiredla, Kishan R., et al.
Publicado: (2023)
por: Nagiredla, Kishan R., et al.
Publicado: (2023)
TRUST: Test-time Resource Utilization for Superior Trustworthiness
por: Harikumar, Haripriya, et al.
Publicado: (2025)
por: Harikumar, Haripriya, et al.
Publicado: (2025)
Novel Kernel Models and Exact Representor Theory for Neural Networks Beyond the Over-Parameterized Regime
por: Shilton, Alistair, et al.
Publicado: (2024)
por: Shilton, Alistair, et al.
Publicado: (2024)
Automated Skill Discovery for Language Agents through Exploration and Iterative Feedback
por: Yang, Yongjin, et al.
Publicado: (2025)
por: Yang, Yongjin, et al.
Publicado: (2025)
Improving Multilingual Language Models by Aligning Representations through Steering
por: Mahmoud, Omar, et al.
Publicado: (2025)
por: Mahmoud, Omar, et al.
Publicado: (2025)
Reference Grounded Skill Discovery
por: Rho, Seungeun, et al.
Publicado: (2025)
por: Rho, Seungeun, et al.
Publicado: (2025)
Agentic Skill Discovery
por: Zhao, Xufeng, et al.
Publicado: (2024)
por: Zhao, Xufeng, et al.
Publicado: (2024)
Variational Offline Multi-agent Skill Discovery
por: Chen, Jiayu, et al.
Publicado: (2024)
por: Chen, Jiayu, et al.
Publicado: (2024)
From Tabula Rasa to Emergent Abilities: Discovering Robot Skills via Real-World Unsupervised Quality-Diversity
por: Grillotti, Luca, et al.
Publicado: (2025)
por: Grillotti, Luca, et al.
Publicado: (2025)
CAX: Cellular Automata Accelerated in JAX
por: Faldor, Maxence, et al.
Publicado: (2024)
por: Faldor, Maxence, et al.
Publicado: (2024)
Efficient Skill Discovery via Regret-Aware Optimization
por: Zhang, He, et al.
Publicado: (2025)
por: Zhang, He, et al.
Publicado: (2025)
SUSD: Structured Unsupervised Skill Discovery through State Factorization
por: Hosseini, Seyed Mohammad Hadi, et al.
Publicado: (2026)
por: Hosseini, Seyed Mohammad Hadi, et al.
Publicado: (2026)
Prompt Optimization with Human Feedback
por: Lin, Xiaoqiang, et al.
Publicado: (2024)
por: Lin, Xiaoqiang, et al.
Publicado: (2024)
Leveraging Skills from Unlabeled Prior Data for Efficient Online Exploration
por: Wilcoxson, Max, et al.
Publicado: (2024)
por: Wilcoxson, Max, et al.
Publicado: (2024)
The Unintended Trade-off of AI Alignment:Balancing Hallucination Mitigation and Safety in LLMs
por: Mahmoud, Omar, et al.
Publicado: (2025)
por: Mahmoud, Omar, et al.
Publicado: (2025)
Unsupervised Skill Discovery as Exploration for Learning Agile Locomotion
por: Rho, Seungeun, et al.
Publicado: (2025)
por: Rho, Seungeun, et al.
Publicado: (2025)
Online Feedback Efficient Active Target Discovery in Partially Observable Environments
por: Sarkar, Anindya, et al.
Publicado: (2025)
por: Sarkar, Anindya, et al.
Publicado: (2025)
Language Guided Skill Discovery
por: Rho, Seungeun, et al.
Publicado: (2024)
por: Rho, Seungeun, et al.
Publicado: (2024)
Neuro-Symbolic Skill Discovery for Conditional Multi-Level Planning
por: Aktas, Hakan, et al.
Publicado: (2024)
por: Aktas, Hakan, et al.
Publicado: (2024)
Understanding the Learning Dynamics of Alignment with Human Feedback
por: Im, Shawn, et al.
Publicado: (2024)
por: Im, Shawn, et al.
Publicado: (2024)
Evolutionary Task Discovery: Advancing Reasoning Frontiers via Skill Composition and Complexity Scaling
por: Ye, Liqin, et al.
Publicado: (2026)
por: Ye, Liqin, et al.
Publicado: (2026)
From Attribution Maps to Human-Understandable Explanations through Concept Relevance Propagation
por: Achtibat, Reduan, et al.
Publicado: (2022)
por: Achtibat, Reduan, et al.
Publicado: (2022)
Unsupervised Skill Discovery for Robotic Manipulation through Automatic Task Generation
por: Jansonnie, Paul, et al.
Publicado: (2024)
por: Jansonnie, Paul, et al.
Publicado: (2024)
ComSD: Balancing Behavioral Quality and Diversity in Unsupervised Skill Discovery
por: Liu, Xin, et al.
Publicado: (2023)
por: Liu, Xin, et al.
Publicado: (2023)
RLAF: Reinforcement Learning from Automaton Feedback
por: Alinejad, Mahyar, et al.
Publicado: (2025)
por: Alinejad, Mahyar, et al.
Publicado: (2025)
Can LLMs Leverage Observational Data? Towards Data-Driven Causal Discovery with LLMs
por: Susanti, Yuni, et al.
Publicado: (2025)
por: Susanti, Yuni, et al.
Publicado: (2025)
Adaptive Preference Scaling for Reinforcement Learning with Human Feedback
por: Hong, Ilgee, et al.
Publicado: (2024)
por: Hong, Ilgee, et al.
Publicado: (2024)
Corruption Robust Offline Reinforcement Learning with Human Feedback
por: Mandal, Debmalya, et al.
Publicado: (2024)
por: Mandal, Debmalya, et al.
Publicado: (2024)
Focused Skill Discovery: Learning to Control Specific State Variables while Minimizing Side Effects
por: Carr, Jonathan Colaço, et al.
Publicado: (2025)
por: Carr, Jonathan Colaço, et al.
Publicado: (2025)
Option Discovery Using LLM-guided Semantic Hierarchical Reinforcement Learning
por: Shek, Chak Lam, et al.
Publicado: (2025)
por: Shek, Chak Lam, et al.
Publicado: (2025)
Zero-Shot LLMs in Human-in-the-Loop RL: Replacing Human Feedback for Reward Shaping
por: Nazir, Mohammad Saif, et al.
Publicado: (2025)
por: Nazir, Mohammad Saif, et al.
Publicado: (2025)
MaestroMotif: Skill Design from Artificial Intelligence Feedback
por: Klissarov, Martin, et al.
Publicado: (2024)
por: Klissarov, Martin, et al.
Publicado: (2024)
Revolutionizing Biomarker Discovery: Leveraging Generative AI for Bio-Knowledge-Embedded Continuous Space Exploration
por: Ying, Wangyang, et al.
Publicado: (2024)
por: Ying, Wangyang, et al.
Publicado: (2024)
CognitionNet: A Collaborative Neural Network for Play Style Discovery in Online Skill Gaming Platform
por: Talwadker, Rukma, et al.
Publicado: (2025)
por: Talwadker, Rukma, et al.
Publicado: (2025)
Leveraging Parameter Space Symmetries for Reasoning Skill Transfer in LLMs
por: Horoi, Stefan, et al.
Publicado: (2025)
por: Horoi, Stefan, et al.
Publicado: (2025)
Ejemplares similares
-
Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
por: Hussonnois, Maxence, et al.
Publicado: (2025) -
MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2025) -
Dynamic Policy Fusion for User Alignment Without Re-Interaction
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2024) -
ASPECT:Analogical Semantic Policy Execution via Language Conditioned Transfer
por: Palattuparambil, Ajsal Shereef, et al.
Publicado: (2026) -
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer
por: Nguyen, Minh Hoang, et al.
Publicado: (2025)