Dynamic Policy Fusion for User Alignment Without Re-Interaction
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Palattuparambil, Ajsal Shereef, Karimpanal, Thommen George, Rana, Santu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer
von: Palattuparambil, Ajsal Shereef, et al.
Veröffentlicht: (2025)
von: Palattuparambil, Ajsal Shereef, et al.
Veröffentlicht: (2025)
ASPECT:Analogical Semantic Policy Execution via Language Conditioned Transfer
von: Palattuparambil, Ajsal Shereef, et al.
Veröffentlicht: (2026)
von: Palattuparambil, Ajsal Shereef, et al.
Veröffentlicht: (2026)
Leveraging Human Feedback for Semantically-Relevant Skill Discovery
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2026)
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2026)
Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2025)
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2025)
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)
ECoDe: A Sample-Efficient Method for Co-Design of Robotic Agents
von: Nagiredla, Kishan R., et al.
Veröffentlicht: (2023)
von: Nagiredla, Kishan R., et al.
Veröffentlicht: (2023)
TRUST: Test-time Resource Utilization for Superior Trustworthiness
von: Harikumar, Haripriya, et al.
Veröffentlicht: (2025)
von: Harikumar, Haripriya, et al.
Veröffentlicht: (2025)
Novel Kernel Models and Exact Representor Theory for Neural Networks Beyond the Over-Parameterized Regime
von: Shilton, Alistair, et al.
Veröffentlicht: (2024)
von: Shilton, Alistair, et al.
Veröffentlicht: (2024)
VUSA: Virtually Upscaled Systolic Array Architecture to Exploit Unstructured Sparsity in AI Acceleration
von: Helal, Shereef, et al.
Veröffentlicht: (2025)
von: Helal, Shereef, et al.
Veröffentlicht: (2025)
The Unintended Trade-off of AI Alignment:Balancing Hallucination Mitigation and Safety in LLMs
von: Mahmoud, Omar, et al.
Veröffentlicht: (2025)
von: Mahmoud, Omar, et al.
Veröffentlicht: (2025)
Improving Multilingual Language Models by Aligning Representations through Steering
von: Mahmoud, Omar, et al.
Veröffentlicht: (2025)
von: Mahmoud, Omar, et al.
Veröffentlicht: (2025)
From Demonstrations to Rewards: Alignment Without Explicit Human Preferences
von: Zeng, Siliang, et al.
Veröffentlicht: (2025)
von: Zeng, Siliang, et al.
Veröffentlicht: (2025)
Prototypical Self-Explainable Models Without Re-training
von: Gautam, Srishti, et al.
Veröffentlicht: (2023)
von: Gautam, Srishti, et al.
Veröffentlicht: (2023)
Ranking Joint Policies in Dynamic Games using Evolutionary Dynamics
von: Koliou, Natalia, et al.
Veröffentlicht: (2025)
von: Koliou, Natalia, et al.
Veröffentlicht: (2025)
IMLE Policy: Fast and Sample Efficient Visuomotor Policy Learning via Implicit Maximum Likelihood Estimation
von: Rana, Krishan, et al.
Veröffentlicht: (2025)
von: Rana, Krishan, et al.
Veröffentlicht: (2025)
MVCL-DAF++: Enhancing Multimodal Intent Recognition via Prototype-Aware Contrastive Alignment and Coarse-to-Fine Dynamic Attention Fusion
von: Huang, Haofeng, et al.
Veröffentlicht: (2025)
von: Huang, Haofeng, et al.
Veröffentlicht: (2025)
GRADE: Replacing Policy Gradients with Backpropagation for LLM Alignment
von: Nel, Lukas Abrie
Veröffentlicht: (2025)
von: Nel, Lukas Abrie
Veröffentlicht: (2025)
PSPO*: An Effective Process-supervised Policy Optimization for Reasoning Alignment
von: Li, Jiawei, et al.
Veröffentlicht: (2024)
von: Li, Jiawei, et al.
Veröffentlicht: (2024)
ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL
von: He, Zelin, et al.
Veröffentlicht: (2026)
von: He, Zelin, et al.
Veröffentlicht: (2026)
Initializing Services in Interactive ML Systems for Diverse Users
von: Bose, Avinandan, et al.
Veröffentlicht: (2023)
von: Bose, Avinandan, et al.
Veröffentlicht: (2023)
Alignment Dynamics in LLM Fine-Tuning
von: Huang, Yuhan, et al.
Veröffentlicht: (2026)
von: Huang, Yuhan, et al.
Veröffentlicht: (2026)
Semantic-aware Wasserstein Policy Regularization for Large Language Model Alignment
von: Na, Byeonghu, et al.
Veröffentlicht: (2026)
von: Na, Byeonghu, et al.
Veröffentlicht: (2026)
OPERA: Automatic Offline Policy Evaluation with Re-weighted Aggregates of Multiple Estimators
von: Nie, Allen, et al.
Veröffentlicht: (2024)
von: Nie, Allen, et al.
Veröffentlicht: (2024)
LoRe: Adaptive Interaction-Evaluation Routing with Per-Step Interaction Budgets for Iterative Graph Solvers
von: Li, Jintao, et al.
Veröffentlicht: (2026)
von: Li, Jintao, et al.
Veröffentlicht: (2026)
Understanding the Learning Dynamics of Alignment with Human Feedback
von: Im, Shawn, et al.
Veröffentlicht: (2024)
von: Im, Shawn, et al.
Veröffentlicht: (2024)
Few-shot Steerable Alignment: Adapting Rewards and LLM Policies with Neural Processes
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2024)
von: Kobalczyk, Katarzyna, et al.
Veröffentlicht: (2024)
Beyond Alignment: Expanding Reasoning Capacity via Manifold-Reshaping Policy Optimization
von: Wang, Dayu, et al.
Veröffentlicht: (2026)
von: Wang, Dayu, et al.
Veröffentlicht: (2026)
Meta-Aligner: Bidirectional Preference-Policy Optimization for Multi-Objective LLMs Alignment
von: Xu, Wenzhe, et al.
Veröffentlicht: (2026)
von: Xu, Wenzhe, et al.
Veröffentlicht: (2026)
TRACE: Distilling Where It Matters via Token-Routed Self On-Policy Alignment
von: Wang, Jiaxuan, et al.
Veröffentlicht: (2026)
von: Wang, Jiaxuan, et al.
Veröffentlicht: (2026)
Probabilistic Token Alignment for Large Language Model Fusion
von: Zeng, Runjia, et al.
Veröffentlicht: (2025)
von: Zeng, Runjia, et al.
Veröffentlicht: (2025)
RePO: Replay-Enhanced Policy Optimization
von: Li, Siheng, et al.
Veröffentlicht: (2025)
von: Li, Siheng, et al.
Veröffentlicht: (2025)
Personalized Alignment Revisited: The Necessity and Sufficiency of User Diversity
von: Kang, Enoch Hyunwook
Veröffentlicht: (2026)
von: Kang, Enoch Hyunwook
Veröffentlicht: (2026)
UserBench: An Interactive Gym Environment for User-Centric Agents
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
von: Qian, Cheng, et al.
Veröffentlicht: (2025)
Dynamic Search for Inference-Time Alignment in Diffusion Models
von: Li, Xiner, et al.
Veröffentlicht: (2025)
von: Li, Xiner, et al.
Veröffentlicht: (2025)
Reflective Preference Optimization (RPO): Enhancing On-Policy Alignment via Hint-Guided Reflection
von: Zhao, Zihui, et al.
Veröffentlicht: (2025)
von: Zhao, Zihui, et al.
Veröffentlicht: (2025)
Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement Learning
von: Ma, Hao, et al.
Veröffentlicht: (2025)
von: Ma, Hao, et al.
Veröffentlicht: (2025)
RLZero: Direct Policy Inference from Language Without In-Domain Supervision
von: Sikchi, Harshit, et al.
Veröffentlicht: (2024)
von: Sikchi, Harshit, et al.
Veröffentlicht: (2024)
COLLAGE: Adaptive Fusion-based Retrieval for Augmented Policy Learning
von: Kumar, Sateesh, et al.
Veröffentlicht: (2025)
von: Kumar, Sateesh, et al.
Veröffentlicht: (2025)
Explaining AI Without Code: A User Study on Explainable AI
von: Abarca, Natalia, et al.
Veröffentlicht: (2025)
von: Abarca, Natalia, et al.
Veröffentlicht: (2025)
TractRLFusion: A GPT-Based Multi-Critic Policy Fusion Framework for Fiber Tractography
von: Joshi, Ankita, et al.
Veröffentlicht: (2026)
von: Joshi, Ankita, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MAGIK: Mapping to Analogous Goals via Imagination-enabled Knowledge Transfer
von: Palattuparambil, Ajsal Shereef, et al.
Veröffentlicht: (2025) -
ASPECT:Analogical Semantic Policy Execution via Language Conditioned Transfer
von: Palattuparambil, Ajsal Shereef, et al.
Veröffentlicht: (2026) -
Leveraging Human Feedback for Semantically-Relevant Skill Discovery
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2026) -
Human-Aligned Skill Discovery: Balancing Behaviour Exploration and Alignment
von: Hussonnois, Maxence, et al.
Veröffentlicht: (2025) -
Beyond the Known: Decision Making with Counterfactual Reasoning Decision Transformer
von: Nguyen, Minh Hoang, et al.
Veröffentlicht: (2025)