SPAR: Support-Preserving Action Rectification
Fuente:
arXiv
Saved in:
| Main Authors: | Zhao, Jiaxin, Pan, Weihang, Liang, Xun, Lin, Binbin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs
by: Zhang, Yuxiang, et al.
Published: (2025)
by: Zhang, Yuxiang, et al.
Published: (2025)
Geometric Manifold Rectification for Imbalanced Learning
by: Wang, Xubin, et al.
Published: (2026)
by: Wang, Xubin, et al.
Published: (2026)
FedSDR: Federated Self-Distillation with Rectification
by: Ren, Ziheng, et al.
Published: (2026)
by: Ren, Ziheng, et al.
Published: (2026)
Mode-Dependent Rectification for Stable PPO Training
by: Mohamad, Mohamad, et al.
Published: (2026)
by: Mohamad, Mohamad, et al.
Published: (2026)
Preserve Support, Not Correspondence: Dynamic Routing for Offline Reinforcement Learning
by: Mu, Zhancun, et al.
Published: (2026)
by: Mu, Zhancun, et al.
Published: (2026)
GFT: From Imitation to Reward Fine-Tuning with Unbiased Group Advantages and Dynamic Coefficient Rectification
by: Gan, Wangjie, et al.
Published: (2026)
by: Gan, Wangjie, et al.
Published: (2026)
Efficient Inference Using Large Language Models with Limited Human Data: Fine-Tuning then Rectification
by: Wang, Lei, et al.
Published: (2025)
by: Wang, Lei, et al.
Published: (2025)
Annealing Self-Distillation Rectification Improves Adversarial Training
by: Wu, Yu-Yu, et al.
Published: (2023)
by: Wu, Yu-Yu, et al.
Published: (2023)
Efficient Rectification of Neuro-Symbolic Reasoning Inconsistencies by Abductive Reflection
by: Hu, Wen-Chao, et al.
Published: (2024)
by: Hu, Wen-Chao, et al.
Published: (2024)
FORLER: Federated Offline Reinforcement Learning with Q-Ensemble and Actor Rectification
by: Qiao, Nan, et al.
Published: (2026)
by: Qiao, Nan, et al.
Published: (2026)
A Rectification-Based Approach for Distilling Boosted Trees into Decision Trees
by: Audemard, Gilles, et al.
Published: (2025)
by: Audemard, Gilles, et al.
Published: (2025)
Can Past Experience Accelerate LLM Reasoning?
by: Pan, Bo, et al.
Published: (2025)
by: Pan, Bo, et al.
Published: (2025)
Reason in Chains, Learn in Trees: Self-Rectification and Grafting for Multi-turn Agent Policy Optimization
by: Li, Yu, et al.
Published: (2026)
by: Li, Yu, et al.
Published: (2026)
Decentralized Structural-RNN for Robot Crowd Navigation with Deep Reinforcement Learning
by: Liu, Shuijing, et al.
Published: (2020)
by: Liu, Shuijing, et al.
Published: (2020)
Preservation of Feature Stability in Machine Learning Under Data Uncertainty for Decision Support in Critical Domains
by: Capała, Karol, et al.
Published: (2024)
by: Capała, Karol, et al.
Published: (2024)
Action-Adaptive Continual Learning: Enabling Policy Generalization under Dynamic Action Spaces
by: Pan, Chaofan, et al.
Published: (2025)
by: Pan, Chaofan, et al.
Published: (2025)
Early Quantization Shrinks Codebook: A Simple Fix for Diversity-Preserving Tokenization
by: Zhao, Wenhao, et al.
Published: (2026)
by: Zhao, Wenhao, et al.
Published: (2026)
Predictive AI Can Support Human Learning while Preserving Error Diversity
by: He, Vivianna Fang, et al.
Published: (2025)
by: He, Vivianna Fang, et al.
Published: (2025)
NoWag: A Unified Framework for Shape Preserving Compression of Large Language Models
by: Liu, Lawrence, et al.
Published: (2025)
by: Liu, Lawrence, et al.
Published: (2025)
Preserving Temporal Dynamics in Time Series Generation
by: Lin, Ci, et al.
Published: (2026)
by: Lin, Ci, et al.
Published: (2026)
Implicit Neural Differential Model for Spatiotemporal Dynamics
by: Akhare, Deepak, et al.
Published: (2025)
by: Akhare, Deepak, et al.
Published: (2025)
Optimizing Privacy-Preserving Primitives to Support LLM-Scale Applications
by: Jandali, Yaman, et al.
Published: (2025)
by: Jandali, Yaman, et al.
Published: (2025)
From $\log π$ to $π$: Taming Divergence in Soft Clipping via Bilateral Decoupled Decay of Probability Gradient Weight
by: Fu, Xiaoliang, et al.
Published: (2026)
by: Fu, Xiaoliang, et al.
Published: (2026)
Provably Efficient Action-Manipulation Attack Against Continuous Reinforcement Learning
by: Luo, Zhi, et al.
Published: (2024)
by: Luo, Zhi, et al.
Published: (2024)
PDNNet: PDN-Aware GNN-CNN Heterogeneous Network for Dynamic IR Drop Prediction
by: Zhao, Yuxiang, et al.
Published: (2024)
by: Zhao, Yuxiang, et al.
Published: (2024)
Ever: Mitigating Hallucination in Large Language Models through Real-Time Verification and Rectification
by: Kang, Haoqiang, et al.
Published: (2023)
by: Kang, Haoqiang, et al.
Published: (2023)
Privacy Preserving Reinforcement Learning with One-Sided Feedback
by: Cong, Lin William, et al.
Published: (2026)
by: Cong, Lin William, et al.
Published: (2026)
Adjusting the Output of Decision Transformer with Action Gradient
by: Lin, Rui, et al.
Published: (2025)
by: Lin, Rui, et al.
Published: (2025)
MASPO: Unifying Gradient Utilization, Probability Mass, and Signal Reliability for Robust and Sample-Efficient LLM Reasoning
by: Fu, Xiaoliang, et al.
Published: (2026)
by: Fu, Xiaoliang, et al.
Published: (2026)
Internalizing LLM Reasoning via Discovery and Replay of Latent Actions
by: Shi, Zhenning, et al.
Published: (2026)
by: Shi, Zhenning, et al.
Published: (2026)
Preferred-Action-Optimized Diffusion Policies for Offline Reinforcement Learning
by: Zhang, Tianle, et al.
Published: (2024)
by: Zhang, Tianle, et al.
Published: (2024)
Geometry Preserving Loss Functions Promote Improved Adaptation of Blackbox Generative Model
by: Mitra, Sinjini, et al.
Published: (2026)
by: Mitra, Sinjini, et al.
Published: (2026)
FLAIN: Mitigating Backdoor Attacks in Federated Learning via Flipping Weight Updates of Low-Activation Input Neurons
by: Ding, Binbin, et al.
Published: (2024)
by: Ding, Binbin, et al.
Published: (2024)
Mitigating Reward Over-Optimization in RLHF via Behavior-Supported Regularization
by: Dai, Juntao, et al.
Published: (2025)
by: Dai, Juntao, et al.
Published: (2025)
Do Latent-CoT Models Think Step-by-Step? A Mechanistic Study on Sequential Reasoning Tasks
by: Liang, Jia, et al.
Published: (2026)
by: Liang, Jia, et al.
Published: (2026)
Off-OAB: Off-Policy Policy Gradient Method with Optimal Action-Dependent Baseline
by: Meng, Wenjia, et al.
Published: (2024)
by: Meng, Wenjia, et al.
Published: (2024)
An Advantage-based Optimization Method for Reinforcement Learning in Large Action Space
by: Lin, Hai, et al.
Published: (2024)
by: Lin, Hai, et al.
Published: (2024)
Attribution-Guided Model Rectification of Unreliable Neural Network Behaviors
by: Yang, Peiyu, et al.
Published: (2026)
by: Yang, Peiyu, et al.
Published: (2026)
The Heterophilic Snowflake Hypothesis: Training and Empowering GNNs for Heterophilic Graphs
by: Wang, Kun, et al.
Published: (2024)
by: Wang, Kun, et al.
Published: (2024)
Dive into Waves: Morlet Spectral Transformer for Cross-Subject Emotion Decoding from EEG
by: Qing, Jiaxin, et al.
Published: (2026)
by: Qing, Jiaxin, et al.
Published: (2026)
Similar Items
-
TokenSqueeze: Performance-Preserving Compression for Reasoning LLMs
by: Zhang, Yuxiang, et al.
Published: (2025) -
Geometric Manifold Rectification for Imbalanced Learning
by: Wang, Xubin, et al.
Published: (2026) -
FedSDR: Federated Self-Distillation with Rectification
by: Ren, Ziheng, et al.
Published: (2026) -
Mode-Dependent Rectification for Stable PPO Training
by: Mohamad, Mohamad, et al.
Published: (2026) -
Preserve Support, Not Correspondence: Dynamic Routing for Offline Reinforcement Learning
by: Mu, Zhancun, et al.
Published: (2026)