Transductive Reward Inference on Graph
Fuente:
arXiv
Saved in:
| Main Authors: | Qu, Bohao, Cao, Xiaofeng, Guo, Qing, Chang, Yi, Tsang, Ivor W., Zhang, Chengqi |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Policy Dispersion in Non-Markovian Environment
by: Qu, Bohao, et al.
Published: (2023)
by: Qu, Bohao, et al.
Published: (2023)
Graph Transductive Defense: a Two-Stage Defense for Graph Membership Inference Attacks
by: Niu, Peizhi, et al.
Published: (2024)
by: Niu, Peizhi, et al.
Published: (2024)
Distributional Multi-objective Black-box Optimization for Diffusion-model Inference-time Multi-Target Generation
by: Tan, Kim Yong, et al.
Published: (2025)
by: Tan, Kim Yong, et al.
Published: (2025)
Imitation from Diverse Behaviors: Wasserstein Quality Diversity Imitation Learning with Single-Step Archive Exploration
by: Yu, Xingrui, et al.
Published: (2024)
by: Yu, Xingrui, et al.
Published: (2024)
Diversifying Policy Behaviors with Extrinsic Behavioral Curiosity
by: Wan, Zhenglin, et al.
Published: (2024)
by: Wan, Zhenglin, et al.
Published: (2024)
Efficient Reinforcement Learning in Probabilistic Reward Machines
by: Lin, Xiaofeng, et al.
Published: (2024)
by: Lin, Xiaofeng, et al.
Published: (2024)
Beyond-Expert Performance with Limited Demonstrations: Efficient Imitation Learning with Double Exploration
by: Zhao, Heyang, et al.
Published: (2025)
by: Zhao, Heyang, et al.
Published: (2025)
Analytical Survey of Learning with Low-Resource Data: From Analysis to Investigation
by: Cao, Xiaofeng, et al.
Published: (2025)
by: Cao, Xiaofeng, et al.
Published: (2025)
Flow-Direct: Feedback-Efficient and Reusable Guidance for Flow Models via Non-Parametric Guidance Field
by: Tan, Kim Yong, et al.
Published: (2026)
by: Tan, Kim Yong, et al.
Published: (2026)
Fast Direct: Query-Efficient Online Black-box Guidance for Diffusion-model Target Generation
by: Tan, Kim Yong, et al.
Published: (2025)
by: Tan, Kim Yong, et al.
Published: (2025)
Sharpness-Aware Black-Box Optimization
by: Ye, Feiyang, et al.
Published: (2024)
by: Ye, Feiyang, et al.
Published: (2024)
FZOO: Fast Zeroth-Order Optimizer for Fine-Tuning Large Language Models towards Adam-Scale Speed
by: Dang, Sizhe, et al.
Published: (2025)
by: Dang, Sizhe, et al.
Published: (2025)
Concept Matching with Agent for Out-of-Distribution Detection
by: Lee, Yuxiao, et al.
Published: (2024)
by: Lee, Yuxiao, et al.
Published: (2024)
Transductive Active Learning: Theory and Applications
by: Hübotter, Jonas, et al.
Published: (2024)
by: Hübotter, Jonas, et al.
Published: (2024)
Beyond the Aggregation Dilemma: Prior-Retaining Decoupled Learning for Multimodal Graphs
by: Yan, Hao, et al.
Published: (2026)
by: Yan, Hao, et al.
Published: (2026)
Olaf-World: Orienting Latent Actions for Video World Modeling
by: Jiang, Yuxin, et al.
Published: (2026)
by: Jiang, Yuxin, et al.
Published: (2026)
Biologically Plausible Brain Graph Transformer
by: Peng, Ciyuan, et al.
Published: (2025)
by: Peng, Ciyuan, et al.
Published: (2025)
Mitigating Mismatch within Reference-based Preference Optimization
by: Yuan, Suqin, et al.
Published: (2026)
by: Yuan, Suqin, et al.
Published: (2026)
ChaosNexus: A Foundation Model for ODE-based Chaotic System Forecasting with Hierarchical Multi-scale Awareness
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
FANFOLD: Graph Normalizing Flows-driven Asymmetric Network for Unsupervised Graph-Level Anomaly Detection
by: Cao, Rui, et al.
Published: (2024)
by: Cao, Rui, et al.
Published: (2024)
CodeScaler: Scaling Code LLM Training and Test-Time Inference via Reward Models
by: Zhu, Xiao, et al.
Published: (2026)
by: Zhu, Xiao, et al.
Published: (2026)
HyperGraphX: Graph Transductive Learning with Hyperdimensional Computing and Message Passing
by: Cong, Guojing, et al.
Published: (2025)
by: Cong, Guojing, et al.
Published: (2025)
Dual-Balancing for Multi-Task Learning
by: Lin, Baijiong, et al.
Published: (2023)
by: Lin, Baijiong, et al.
Published: (2023)
Uncover and Unlearn Nuisances: Agnostic Fully Test-Time Adaptation
by: Srey, Ponhvoan, et al.
Published: (2025)
by: Srey, Ponhvoan, et al.
Published: (2025)
Transduction is All You Need for Structured Data Workflows
by: Gliozzo, Alfio, et al.
Published: (2025)
by: Gliozzo, Alfio, et al.
Published: (2025)
Latent Reward: LLM-Empowered Credit Assignment in Episodic Reinforcement Learning
by: Qu, Yun, et al.
Published: (2024)
by: Qu, Yun, et al.
Published: (2024)
BLAST: Block-Level Adaptive Structured Matrices for Efficient Deep Neural Network Inference
by: Lee, Changwoo, et al.
Published: (2024)
by: Lee, Changwoo, et al.
Published: (2024)
Value of Information and Reward Specification in Active Inference and POMDPs
by: Wei, Ran
Published: (2024)
by: Wei, Ran
Published: (2024)
Transductive Confidence Machine and its application to Medical Data Sets
by: Lindsay, David
Published: (2024)
by: Lindsay, David
Published: (2024)
Mamba Integrated with Physics Principles Masters Long-term Chaotic System Forecasting
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
Reward Centering
by: Naik, Abhishek, et al.
Published: (2024)
by: Naik, Abhishek, et al.
Published: (2024)
Explain Less, Understand More: Jargon Detection via Personalized Parameter-Efficient Fine-tuning
by: Wu, Bohao, et al.
Published: (2025)
by: Wu, Bohao, et al.
Published: (2025)
MemReward: Graph-Based Experience Memory for LLM Reward Prediction with Limited Labels
by: Luo, Tianyang, et al.
Published: (2026)
by: Luo, Tianyang, et al.
Published: (2026)
Agentics 2.0: Logical Transduction Algebra for Agentic Data Workflows
by: Gliozzo, Alfio Massimiliano, et al.
Published: (2026)
by: Gliozzo, Alfio Massimiliano, et al.
Published: (2026)
CAESAR: Enhancing Federated RL in Heterogeneous MDPs through Convergence-Aware Sampling with Screening
by: Mak, Hei Yi, et al.
Published: (2024)
by: Mak, Hei Yi, et al.
Published: (2024)
Combining Induction and Transduction for Abstract Reasoning
by: Li, Wen-Ding, et al.
Published: (2024)
by: Li, Wen-Ding, et al.
Published: (2024)
Reward Shaping for Inference-Time Alignment: A Stackelberg Game Perspective
by: Wang, Haichuan, et al.
Published: (2026)
by: Wang, Haichuan, et al.
Published: (2026)
Effective Illicit Account Detection on Large Cryptocurrency MultiGraphs
by: Ding, Zhihao, et al.
Published: (2023)
by: Ding, Zhihao, et al.
Published: (2023)
Safety Modulation: Enhancing Safety in Reinforcement Learning through Cost-Modulated Rewards
by: Zhang, Hanping, et al.
Published: (2025)
by: Zhang, Hanping, et al.
Published: (2025)
Compositional Conservatism: A Transductive Approach in Offline Reinforcement Learning
by: Song, Yeda, et al.
Published: (2024)
by: Song, Yeda, et al.
Published: (2024)
Similar Items
-
Policy Dispersion in Non-Markovian Environment
by: Qu, Bohao, et al.
Published: (2023) -
Graph Transductive Defense: a Two-Stage Defense for Graph Membership Inference Attacks
by: Niu, Peizhi, et al.
Published: (2024) -
Distributional Multi-objective Black-box Optimization for Diffusion-model Inference-time Multi-Target Generation
by: Tan, Kim Yong, et al.
Published: (2025) -
Imitation from Diverse Behaviors: Wasserstein Quality Diversity Imitation Learning with Single-Step Archive Exploration
by: Yu, Xingrui, et al.
Published: (2024) -
Diversifying Policy Behaviors with Extrinsic Behavioral Curiosity
by: Wan, Zhenglin, et al.
Published: (2024)