FLoRA: Sample-Efficient Preference-based RL via Low-Rank Style Adaptation of Reward Functions
Fuente:
arXiv
Salvato in:
| Autori principali: | Marta, Daniel, Holk, Simon, Vasco, Miguel, Lundell, Jens, Homberger, Timon, Busch, Finn, Andersson, Olov, Kragic, Danica, Leite, Iolanda |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
DIV-Nav: Open-Vocabulary Spatial Relationships for Multi-Object Navigation
di: Ortega-Peimbert, Jesús, et al.
Pubblicazione: (2025)
di: Ortega-Peimbert, Jesús, et al.
Pubblicazione: (2025)
One Map to Find Them All: Real-time Open-Vocabulary Mapping for Zero-shot Multi-Object Navigation
di: Busch, Finn Lukas, et al.
Pubblicazione: (2024)
di: Busch, Finn Lukas, et al.
Pubblicazione: (2024)
FUS3DMaps: Scalable and Accurate Open-Vocabulary Semantic Mapping by 3D Fusion of Voxel- and Instance-Level Layers
di: Homberger, Timon, et al.
Pubblicazione: (2026)
di: Homberger, Timon, et al.
Pubblicazione: (2026)
FLoRA: Federated Fine-Tuning Large Language Models with Heterogeneous Low-Rank Adaptations
di: Wang, Ziyao, et al.
Pubblicazione: (2024)
di: Wang, Ziyao, et al.
Pubblicazione: (2024)
PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning
di: Holk, Simon, et al.
Pubblicazione: (2024)
di: Holk, Simon, et al.
Pubblicazione: (2024)
R-FLoRA: Residual-Statistic-Gated Low-Rank Adaptation for Single-Image Face Morphing Attack Detection
di: Ramachandra, Raghavendra
Pubblicazione: (2026)
di: Ramachandra, Raghavendra
Pubblicazione: (2026)
FLoRA: Enhancing Vision-Language Models with Parameter-Efficient Federated Learning
di: Nguyen, Duy Phuong, et al.
Pubblicazione: (2024)
di: Nguyen, Duy Phuong, et al.
Pubblicazione: (2024)
S$^2$-Diffusion: Generalizing from Instance-level to Category-level Skills in Robot Manipulation
di: Yang, Quantao, et al.
Pubblicazione: (2025)
di: Yang, Quantao, et al.
Pubblicazione: (2025)
Humans Coexist, So Must Embodied Artificial Agents
di: Kuehn, Hannah, et al.
Pubblicazione: (2025)
di: Kuehn, Hannah, et al.
Pubblicazione: (2025)
The FLoRA Engine: Using Analytics to Measure and Facilitate Learners' own Regulation Activities
di: Li, Xinyu, et al.
Pubblicazione: (2024)
di: Li, Xinyu, et al.
Pubblicazione: (2024)
qa-FLoRA: Data-free query-adaptive Fusion of LoRAs for LLMs
di: Shukla, Shreya, et al.
Pubblicazione: (2025)
di: Shukla, Shreya, et al.
Pubblicazione: (2025)
DexDiffuser: Generating Dexterous Grasps with Diffusion Models
di: Weng, Zehang, et al.
Pubblicazione: (2024)
di: Weng, Zehang, et al.
Pubblicazione: (2024)
Pushing Everything Everywhere All At Once: Probabilistic Prehensile Pushing
di: Perugini, Patrizio, et al.
Pubblicazione: (2025)
di: Perugini, Patrizio, et al.
Pubblicazione: (2025)
CAPGrasp: An $\mathbb{R}^3\times \text{SO(2)-equivariant}$ Continuous Approach-Constrained Generative Grasp Sampler
di: Weng, Zehang, et al.
Pubblicazione: (2023)
di: Weng, Zehang, et al.
Pubblicazione: (2023)
FLoRA: An Advanced AI-Powered Engine to Facilitate Hybrid Human-AI Regulated Learning
di: Li, Xinyu, et al.
Pubblicazione: (2025)
di: Li, Xinyu, et al.
Pubblicazione: (2025)
Geometry of Uncertainty: Learning Metric Spaces for Multimodal State Estimation in RL
di: Reichlin, Alfredo, et al.
Pubblicazione: (2026)
di: Reichlin, Alfredo, et al.
Pubblicazione: (2026)
The Impact of VR and 2D Interfaces on Human Feedback in Preference-Based Robot Learning
di: de Heuvel, Jorge, et al.
Pubblicazione: (2025)
di: de Heuvel, Jorge, et al.
Pubblicazione: (2025)
Reduced-order Control and Geometric Structure of Learned Lagrangian Latent Dynamics
di: Friedl, Katharina, et al.
Pubblicazione: (2026)
di: Friedl, Katharina, et al.
Pubblicazione: (2026)
A Riemannian Framework for Learning Reduced-order Lagrangian Dynamics
di: Friedl, Katharina, et al.
Pubblicazione: (2024)
di: Friedl, Katharina, et al.
Pubblicazione: (2024)
FLoRA: Fused forward-backward adapters for parameter efficient fine-tuning and reducing inference-time latencies of LLMs
di: Gowda, Dhananjaya, et al.
Pubblicazione: (2025)
di: Gowda, Dhananjaya, et al.
Pubblicazione: (2025)
Walking on the Fiber: A Simple Geometric Approximation for Bayesian Neural Networks
di: Reichlin, Alfredo, et al.
Pubblicazione: (2025)
di: Reichlin, Alfredo, et al.
Pubblicazione: (2025)
FLoRA: Fusion-Latent for Optical Reconstruction and Flood Area Segmentation via Cross-Modal Multi-Task Distillation Network
di: Talreja, Jagrati, et al.
Pubblicazione: (2026)
di: Talreja, Jagrati, et al.
Pubblicazione: (2026)
CompSLAM: Complementary Hierarchical Multi-Modal Localization and Mapping for Robot Autonomy in Underground Environments
di: Khattak, Shehryar, et al.
Pubblicazione: (2025)
di: Khattak, Shehryar, et al.
Pubblicazione: (2025)
Grasping a Handful: Sequential Multi-Object Dexterous Grasp Generation
di: Lu, Haofei, et al.
Pubblicazione: (2025)
di: Lu, Haofei, et al.
Pubblicazione: (2025)
The 1st InterAI Workshop: Interactive AI for Human-centered Robotics
di: Zhang, Yuchong, et al.
Pubblicazione: (2024)
di: Zhang, Yuchong, et al.
Pubblicazione: (2024)
Goal-Conditioned Reinforcement Learning from Sub-Optimal Data on Metric Spaces
di: Reichlin, Alfredo, et al.
Pubblicazione: (2024)
di: Reichlin, Alfredo, et al.
Pubblicazione: (2024)
Will You Participate? Exploring the Potential of Robotics Competitions on Human-centric Topics
di: Zhang, Yuchong, et al.
Pubblicazione: (2024)
di: Zhang, Yuchong, et al.
Pubblicazione: (2024)
zFLoRA: Zero-Latency Fused Low-Rank Adapters
di: Gowda, Dhananjaya, et al.
Pubblicazione: (2025)
di: Gowda, Dhananjaya, et al.
Pubblicazione: (2025)
Preference Aligned Visuomotor Diffusion Policies for Deformable Object Manipulation
di: Moletta, Marco, et al.
Pubblicazione: (2026)
di: Moletta, Marco, et al.
Pubblicazione: (2026)
Cloth-Splatting: 3D Cloth State Estimation from RGB Supervision
di: Longhini, Alberta, et al.
Pubblicazione: (2025)
di: Longhini, Alberta, et al.
Pubblicazione: (2025)
FLAME: A Federated Learning Benchmark for Robotic Manipulation
di: Betran, Santiago Bou, et al.
Pubblicazione: (2025)
di: Betran, Santiago Bou, et al.
Pubblicazione: (2025)
Can Transformers Smell Like Humans?
di: Taleb, Farzaneh, et al.
Pubblicazione: (2024)
di: Taleb, Farzaneh, et al.
Pubblicazione: (2024)
Reducing Variance in Meta-Learning via Laplace Approximation for Regression Tasks
di: Reichlin, Alfredo, et al.
Pubblicazione: (2024)
di: Reichlin, Alfredo, et al.
Pubblicazione: (2024)
StyleQoRA: Quality-Aware Low-Rank Adaptation for Few-Shot Multi-Style Editing
di: Cao, Cong, et al.
Pubblicazione: (2025)
di: Cao, Cong, et al.
Pubblicazione: (2025)
FLoE: Fisher-Based Layer Selection for Efficient Sparse Adaptation of Low-Rank Experts
di: Wang, Xinyi, et al.
Pubblicazione: (2025)
di: Wang, Xinyi, et al.
Pubblicazione: (2025)
Real-Time Iteration Scheme for Diffusion Policy
di: Duan, Yufei, et al.
Pubblicazione: (2025)
di: Duan, Yufei, et al.
Pubblicazione: (2025)
Reframing Human-Robot Interaction Through Extended Reality: Unlocking Safer, Smarter, and More Empathic Interactions with Virtual Robots and Foundation Models
di: Zhang, Yuchong, et al.
Pubblicazione: (2025)
di: Zhang, Yuchong, et al.
Pubblicazione: (2025)
Vision Beyond Boundaries: An Initial Design Space of Domain-specific Large Vision Models in Human-robot Interaction
di: Zhang, Yuchong, et al.
Pubblicazione: (2024)
di: Zhang, Yuchong, et al.
Pubblicazione: (2024)
Raising Body Ownership in End-to-End Visuomotor Policy Learning via Robot-Centric Pooling
di: Zhuang, Zheyu, et al.
Pubblicazione: (2024)
di: Zhuang, Zheyu, et al.
Pubblicazione: (2024)
Learning to Localize Reference Trajectories in Image-Space for Visual Navigation
di: Busch, Finn Lukas, et al.
Pubblicazione: (2026)
di: Busch, Finn Lukas, et al.
Pubblicazione: (2026)
Documenti analoghi
-
DIV-Nav: Open-Vocabulary Spatial Relationships for Multi-Object Navigation
di: Ortega-Peimbert, Jesús, et al.
Pubblicazione: (2025) -
One Map to Find Them All: Real-time Open-Vocabulary Mapping for Zero-shot Multi-Object Navigation
di: Busch, Finn Lukas, et al.
Pubblicazione: (2024) -
FUS3DMaps: Scalable and Accurate Open-Vocabulary Semantic Mapping by 3D Fusion of Voxel- and Instance-Level Layers
di: Homberger, Timon, et al.
Pubblicazione: (2026) -
FLoRA: Federated Fine-Tuning Large Language Models with Heterogeneous Low-Rank Adaptations
di: Wang, Ziyao, et al.
Pubblicazione: (2024) -
PREDILECT: Preferences Delineated with Zero-Shot Language-based Reasoning in Reinforcement Learning
di: Holk, Simon, et al.
Pubblicazione: (2024)