Auxiliary Reward Generation with Transition Distance Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Siyuan, Han, Shijie, Zhao, Yingnan, Liang, By, Liu, Peng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SemiReward: A General Reward Model for Semi-supervised Learning
by: Li, Siyuan, et al.
Published: (2023)
by: Li, Siyuan, et al.
Published: (2023)
When Distance Distracts: Representation Distance Bias in BT-Loss for Reward Models
by: Xie, Tong, et al.
Published: (2025)
by: Xie, Tong, et al.
Published: (2025)
Richer Representations for Neural Algorithmic Reasoning via Auxiliary Reconstruction
by: Huang, Jiafu, et al.
Published: (2026)
by: Huang, Jiafu, et al.
Published: (2026)
Why and How Auxiliary Tasks Improve JEPA Representations
by: Yu, Jiacan, et al.
Published: (2025)
by: Yu, Jiacan, et al.
Published: (2025)
Representation Learning of Auxiliary Concepts for Improved Student Modeling and Exercise Recommendation
by: Badran, Yahya, et al.
Published: (2025)
by: Badran, Yahya, et al.
Published: (2025)
Personalized Subgraph Federated Learning with Differentiable Auxiliary Projections
by: Zhuo, Wei, et al.
Published: (2025)
by: Zhuo, Wei, et al.
Published: (2025)
TimeRewarder: Learning Dense Reward from Passive Videos via Frame-wise Temporal Distance
by: Liu, Yuyang, et al.
Published: (2025)
by: Liu, Yuyang, et al.
Published: (2025)
R^3: Replay, Reflection, and Ranking Rewards for LLM Reinforcement Learning
by: Jiang, Zhizheng, et al.
Published: (2026)
by: Jiang, Zhizheng, et al.
Published: (2026)
Debiasing Reward Models by Representation Learning with Guarantees
by: Ng, Ignavier, et al.
Published: (2025)
by: Ng, Ignavier, et al.
Published: (2025)
CAML: Collaborative Auxiliary Modality Learning for Multi-Agent Systems
by: Liu, Rui, et al.
Published: (2025)
by: Liu, Rui, et al.
Published: (2025)
Flow Matching with Arbitrary Auxiliary Paths
by: Peng, Xin, et al.
Published: (2026)
by: Peng, Xin, et al.
Published: (2026)
Humanoid-inspired Causal Representation Learning for Domain Generalization
by: Tao, Ze, et al.
Published: (2025)
by: Tao, Ze, et al.
Published: (2025)
Forming Auxiliary High-confident Instance-level Loss to Promote Learning from Label Proportions
by: Ma, Tianhao, et al.
Published: (2024)
by: Ma, Tianhao, et al.
Published: (2024)
Discovering the Representation Bottleneck of Graph Neural Networks
by: Wu, Fang, et al.
Published: (2022)
by: Wu, Fang, et al.
Published: (2022)
Disentangled Generative Graph Representation Learning
by: Hu, Xinyue, et al.
Published: (2024)
by: Hu, Xinyue, et al.
Published: (2024)
Unveiling the Significance of Toddler-Inspired Reward Transition in Goal-Oriented Reinforcement Learning
by: Park, Junseok, et al.
Published: (2024)
by: Park, Junseok, et al.
Published: (2024)
Notes on the Reward Representation of Posterior Updates
by: Ortega, Pedro A.
Published: (2026)
by: Ortega, Pedro A.
Published: (2026)
Efficient Generative Model Training via Embedded Representation Warmup
by: Liu, Deyuan, et al.
Published: (2025)
by: Liu, Deyuan, et al.
Published: (2025)
Decoupled Split Learning via Auxiliary Loss
by: Zihad, Anower, et al.
Published: (2026)
by: Zihad, Anower, et al.
Published: (2026)
GenURL: A General Framework for Unsupervised Representation Learning
by: Li, Siyuan, et al.
Published: (2021)
by: Li, Siyuan, et al.
Published: (2021)
Identifiable Exchangeable Mechanisms for Causal Structure and Representation Learning
by: Reizinger, Patrik, et al.
Published: (2024)
by: Reizinger, Patrik, et al.
Published: (2024)
Vision-Based Generic Potential Function for Policy Alignment in Multi-Agent Reinforcement Learning
by: Ma, Hao, et al.
Published: (2025)
by: Ma, Hao, et al.
Published: (2025)
Logit Distance Bounds Representational Similarity
by: Nielsen, Beatrix M. G., et al.
Published: (2026)
by: Nielsen, Beatrix M. G., et al.
Published: (2026)
Temporal Representations for Exploration: Learning Complex Exploratory Behavior without Extrinsic Rewards
by: Mohamed, Faisal, et al.
Published: (2026)
by: Mohamed, Faisal, et al.
Published: (2026)
VRAIL: Vectorized Reward-based Attribution for Interpretable Learning
by: Kim, Jina, et al.
Published: (2025)
by: Kim, Jina, et al.
Published: (2025)
ProtoGCD: Unified and Unbiased Prototype Learning for Generalized Category Discovery
by: Ma, Shijie, et al.
Published: (2025)
by: Ma, Shijie, et al.
Published: (2025)
Improving Continual Learning Performance and Efficiency with Auxiliary Classifiers
by: Szatkowski, Filip, et al.
Published: (2024)
by: Szatkowski, Filip, et al.
Published: (2024)
Grid and Road Expressions Are Complementary for Trajectory Representation Learning
by: Zhou, Silin, et al.
Published: (2024)
by: Zhou, Silin, et al.
Published: (2024)
VCformer: Variable Correlation Transformer with Inherent Lagged Correlation for Multivariate Time Series Forecasting
by: Yang, Yingnan, et al.
Published: (2024)
by: Yang, Yingnan, et al.
Published: (2024)
Skill Learning via Policy Diversity Yields Identifiable Representations for Reinforcement Learning
by: Reizinger, Patrik, et al.
Published: (2025)
by: Reizinger, Patrik, et al.
Published: (2025)
Multimodal Representation Learning Conditioned on Semantic Relations
by: Qiao, Yang, et al.
Published: (2025)
by: Qiao, Yang, et al.
Published: (2025)
DyMRL: Dynamic Multispace Representation Learning for Multimodal Event Forecasting in Knowledge Graph
by: Zhao, Feng, et al.
Published: (2026)
by: Zhao, Feng, et al.
Published: (2026)
Reward Learning From Preference With Ties
by: Liu, Jinsong, et al.
Published: (2024)
by: Liu, Jinsong, et al.
Published: (2024)
GFRIEND: Generative Few-shot Reward Inference through EfficieNt DPO
by: Zhao, Yiyang, et al.
Published: (2025)
by: Zhao, Yiyang, et al.
Published: (2025)
Can We Really Learn One Representation to Optimize All Rewards?
by: Zheng, Chongyi, et al.
Published: (2026)
by: Zheng, Chongyi, et al.
Published: (2026)
Exploring Pass-Rate Reward in Reinforcement Learning for Code Generation
by: Li, Xin-Ye, et al.
Published: (2026)
by: Li, Xin-Ye, et al.
Published: (2026)
RDesign: Hierarchical Data-efficient Representation Learning for Tertiary Structure-based RNA Design
by: Tan, Cheng, et al.
Published: (2023)
by: Tan, Cheng, et al.
Published: (2023)
Learn-to-Distance: Distance Learning for Detecting LLM-Generated Text
by: Zhou, Hongyi, et al.
Published: (2026)
by: Zhou, Hongyi, et al.
Published: (2026)
Robust Model-Based Reinforcement Learning with an Adversarial Auxiliary Model
by: Herremans, Siemen, et al.
Published: (2024)
by: Herremans, Siemen, et al.
Published: (2024)
RationalRewards: Reasoning Rewards Scale Visual Generation Both Training and Test Time
by: Wang, Haozhe, et al.
Published: (2026)
by: Wang, Haozhe, et al.
Published: (2026)
Similar Items
-
SemiReward: A General Reward Model for Semi-supervised Learning
by: Li, Siyuan, et al.
Published: (2023) -
When Distance Distracts: Representation Distance Bias in BT-Loss for Reward Models
by: Xie, Tong, et al.
Published: (2025) -
Richer Representations for Neural Algorithmic Reasoning via Auxiliary Reconstruction
by: Huang, Jiafu, et al.
Published: (2026) -
Why and How Auxiliary Tasks Improve JEPA Representations
by: Yu, Jiacan, et al.
Published: (2025) -
Representation Learning of Auxiliary Concepts for Improved Student Modeling and Exercise Recommendation
by: Badran, Yahya, et al.
Published: (2025)