Mapping representations in Reinforcement Learning via Semantic Alignment for Zero-Shot Stitching
Fuente:
arXiv
Saved in:
| Main Authors: | Ricciardi, Antonio Pio, Maiorca, Valentino, Moschella, Luca, Marin, Riccardo, Rodolà, Emanuele |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
R3L: Relative Representations for Reinforcement Learning
by: Ricciardi, Antonio Pio, et al.
Published: (2024)
by: Ricciardi, Antonio Pio, et al.
Published: (2024)
Deep Hedging Under Non-Convexity: Limitations and a Case for AlphaZero
by: Maggiolo, Matteo, et al.
Published: (2025)
by: Maggiolo, Matteo, et al.
Published: (2025)
The Inhibitor: ReLU and Addition-Based Attention for Efficient Transformers under Fully Homomorphic Encryption on the Torus
by: Brännvall, Rickard, et al.
Published: (2023)
by: Brännvall, Rickard, et al.
Published: (2023)
Semantic Reward Collapse and the Preservation of Epistemic Integrity in Adaptive AI Systems
by: Parris, William
Published: (2026)
by: Parris, William
Published: (2026)
MAcPNN: Mutual Assisted Learning on Data Streams with Temporal Dependence
by: Giannini, Federico, et al.
Published: (2026)
by: Giannini, Federico, et al.
Published: (2026)
From Features to Graphs: Exploring Graph Structures and Pairwise Interactions via GNNs
by: Yamchote, Phaphontee, et al.
Published: (2025)
by: Yamchote, Phaphontee, et al.
Published: (2025)
Step-Aware Residual-Guided Diffusion for EEG Spatial Super-Resolution
by: Liu, Hongjun, et al.
Published: (2025)
by: Liu, Hongjun, et al.
Published: (2025)
SafeAgent: Safeguarding LLM Agents via an Automated Risk Simulator
by: Zhou, Xueyang, et al.
Published: (2025)
by: Zhou, Xueyang, et al.
Published: (2025)
IntSeqBERT: Learning Arithmetic Structure in OEIS via Modulo-Spectrum Embeddings
by: Nakasho, Kazuhisa
Published: (2026)
by: Nakasho, Kazuhisa
Published: (2026)
FeNeC: Enhancing Continual Learning via Feature Clustering with Neighbor- or Logit-Based Classification
by: Książek, Kamil, et al.
Published: (2025)
by: Książek, Kamil, et al.
Published: (2025)
cPNN: Continuous Progressive Neural Networks for Evolving Streaming Time Series
by: Giannini, Federico, et al.
Published: (2026)
by: Giannini, Federico, et al.
Published: (2026)
Model Capability Dominates: Inference-Time Optimization Lessons from AIMO 3
by: Nitarach, Natapong
Published: (2026)
by: Nitarach, Natapong
Published: (2026)
Stage-wise Dynamics of Classifier-Free Guidance in Diffusion Models
by: Jin, Cheng, et al.
Published: (2025)
by: Jin, Cheng, et al.
Published: (2025)
Drift-Resilient TabPFN: In-Context Learning Temporal Distribution Shifts on Tabular Data
by: Helli, Kai, et al.
Published: (2024)
by: Helli, Kai, et al.
Published: (2024)
X-Factor: Quality Is a Dataset-Intrinsic Property
by: Couch, Josiah, et al.
Published: (2025)
by: Couch, Josiah, et al.
Published: (2025)
A Constraint-Preserving Neural Network Approach for Solving Mean-Field Games Equilibrium
by: Liu, Jinwei, et al.
Published: (2025)
by: Liu, Jinwei, et al.
Published: (2025)
The Domain Mixed Unit: A New Neural Arithmetic Layer
by: Curry, Paul
Published: (2025)
by: Curry, Paul
Published: (2025)
Deep learning four decades of human migration
by: Gaskin, Thomas, et al.
Published: (2025)
by: Gaskin, Thomas, et al.
Published: (2025)
Your contrastive learning problem is secretly a distribution alignment problem
by: Chen, Zihao, et al.
Published: (2025)
by: Chen, Zihao, et al.
Published: (2025)
Developing Explainable Machine Learning Model using Augmented Concept Activation Vector
by: Hassanpour, Reza, et al.
Published: (2024)
by: Hassanpour, Reza, et al.
Published: (2024)
Don't Look Back in Anger: MAGIC Net for Streaming Continual Learning with Temporal Dependence
by: Giannini, Federico, et al.
Published: (2026)
by: Giannini, Federico, et al.
Published: (2026)
NOTAI.AI: Explainable Detection of Machine-Generated Text via Curvature and Feature Attribution
by: Breneur, Oleksandr Marchenko, et al.
Published: (2026)
by: Breneur, Oleksandr Marchenko, et al.
Published: (2026)
Upside Down Reinforcement Learning with Policy Generators
by: Di Ventura, Jacopo, et al.
Published: (2025)
by: Di Ventura, Jacopo, et al.
Published: (2025)
Correction and Corruption: A Two-Rate View of Error Flow in LLM Protocols
by: Reitich, Fernando
Published: (2026)
by: Reitich, Fernando
Published: (2026)
Distinguished In Uniform: Self Attention Vs. Virtual Nodes
by: Rosenbluth, Eran, et al.
Published: (2024)
by: Rosenbluth, Eran, et al.
Published: (2024)
Evaluating Model-Agnostic Meta-Learning on MetaWorld ML10 Benchmark: Fast Adaptation in Robotic Manipulation Tasks
by: Atamuradov, Sanjar
Published: (2025)
by: Atamuradov, Sanjar
Published: (2025)
Just In Time Transformers
by: Benali, Ahmed Ala Eddine, et al.
Published: (2024)
by: Benali, Ahmed Ala Eddine, et al.
Published: (2024)
torchsom: The Reference PyTorch Library for Self-Organizing Maps
by: Berthier, Louis, et al.
Published: (2025)
by: Berthier, Louis, et al.
Published: (2025)
A Practical Guide to Streaming Continual Learning
by: Cossu, Andrea, et al.
Published: (2026)
by: Cossu, Andrea, et al.
Published: (2026)
A Comparison Between Decision Transformers and Traditional Offline Reinforcement Learning Algorithms
by: Caunhye, Ali Murtaza, et al.
Published: (2025)
by: Caunhye, Ali Murtaza, et al.
Published: (2025)
Scalpel-SAM: A Semi-Supervised Paradigm for Adapting SAM to Infrared Small Object Detection
by: Liu, Zihan, et al.
Published: (2025)
by: Liu, Zihan, et al.
Published: (2025)
Formulation and Therapeutic Assessment of a Zinc Oxide, Silver, and Cerium Oxide Enriched Ointment for Accelerated Wound Healing in Aged Models
by: Yousaf, Iqra, et al.
Published: (2025)
by: Yousaf, Iqra, et al.
Published: (2025)
Agentic Discovery of Neural Architectures: AIRA-Compose and AIRA-Design
by: Pepe, Alberto, et al.
Published: (2026)
by: Pepe, Alberto, et al.
Published: (2026)
When Does Content-Based Routing Work? Representation Requirements for Selective Attention in Hybrid Sequence Models
by: Basu, Abhinaba
Published: (2026)
by: Basu, Abhinaba
Published: (2026)
An Improved Adaptive PID Optimizer with Enhanced Convergence and Stability for Deep Learning
by: Saini, Saurabh, et al.
Published: (2026)
by: Saini, Saurabh, et al.
Published: (2026)
Rethinking the Multilingual Reasoning Gap with Layer Swap
by: Lasbordes, Maxence, et al.
Published: (2026)
by: Lasbordes, Maxence, et al.
Published: (2026)
Territory Paint Wars: Diagnosing and Mitigating Failure Modes in Competitive Multi-Agent PPO
by: Singh, Diyansha
Published: (2026)
by: Singh, Diyansha
Published: (2026)
Feature Selection Based on Reinforcement Learning and Hazard State Classification for Magnetic Adhesion Wall-Climbing Robots
by: Ma, Zhen, et al.
Published: (2025)
by: Ma, Zhen, et al.
Published: (2025)
Enhancing Predictive Accuracy in Tennis: Integrating Fuzzy Logic and CV-GRNN for Dynamic Match Outcome and Player Momentum Analysis
by: Li, Kechen, et al.
Published: (2025)
by: Li, Kechen, et al.
Published: (2025)
Neural Concept Verifier: Scaling Prover-Verifier Games via Concept Encodings
by: Turan, Berkant, et al.
Published: (2025)
by: Turan, Berkant, et al.
Published: (2025)
Similar Items
-
R3L: Relative Representations for Reinforcement Learning
by: Ricciardi, Antonio Pio, et al.
Published: (2024) -
Deep Hedging Under Non-Convexity: Limitations and a Case for AlphaZero
by: Maggiolo, Matteo, et al.
Published: (2025) -
The Inhibitor: ReLU and Addition-Based Attention for Efficient Transformers under Fully Homomorphic Encryption on the Torus
by: Brännvall, Rickard, et al.
Published: (2023) -
Semantic Reward Collapse and the Preservation of Epistemic Integrity in Adaptive AI Systems
by: Parris, William
Published: (2026) -
MAcPNN: Mutual Assisted Learning on Data Streams with Temporal Dependence
by: Giannini, Federico, et al.
Published: (2026)