Swift-Sarsa: Fast and Robust Linear Control
Fuente:
arXiv
Saved in:
| Main Authors: | Javed, Khurram, Sutton, Richard S. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Step-size Optimization for Continual Learning
by: Degris, Thomas, et al.
Published: (2024)
by: Degris, Thomas, et al.
Published: (2024)
SwiftKV: Fast Prefill-Optimized Inference with Knowledge-Preserving Model Transformation
by: Qiao, Aurick, et al.
Published: (2024)
by: Qiao, Aurick, et al.
Published: (2024)
Can LLMs Reconcile Knowledge Conflicts in Counterfactual Reasoning
by: Yamin, Khurram, et al.
Published: (2025)
by: Yamin, Khurram, et al.
Published: (2025)
SwiftF0: Fast and Accurate Monophonic Pitch Detection
by: Nieradzik, Lars
Published: (2025)
by: Nieradzik, Lars
Published: (2025)
Swift Sampler: Efficient Learning of Sampler by 10 Parameters
by: Yao, Jiawei, et al.
Published: (2024)
by: Yao, Jiawei, et al.
Published: (2024)
RIFT: A Scalable Methodology for LLM Accelerator Fault Assessment using Reinforcement Learning
by: Khalil, Khurram, et al.
Published: (2025)
by: Khalil, Khurram, et al.
Published: (2025)
Reward Centering
by: Naik, Abhishek, et al.
Published: (2024)
by: Naik, Abhishek, et al.
Published: (2024)
SwiftPrune: Hessian-Free Weight Pruning for Large Language Models
by: Kang, Yuhan, et al.
Published: (2025)
by: Kang, Yuhan, et al.
Published: (2025)
Lipschitz-aware Linearity Grafting for Certified Robustness
by: Han, Yongjin, et al.
Published: (2025)
by: Han, Yongjin, et al.
Published: (2025)
Calibrating LLM Judges: Linear Probes for Fast and Reliable Uncertainty Estimation
by: Radharapu, Bhaktipriya, et al.
Published: (2025)
by: Radharapu, Bhaktipriya, et al.
Published: (2025)
Ignition Phase : Standard Training for Fast Adversarial Robustness
by: Yu-Hang, Wang, et al.
Published: (2025)
by: Yu-Hang, Wang, et al.
Published: (2025)
Fast and Interpretable Mixed-Integer Linear Program Solving by Learning Model Reduction
by: Li, Yixuan, et al.
Published: (2024)
by: Li, Yixuan, et al.
Published: (2024)
Kernel Banzhaf: A Fast and Robust Estimator for Banzhaf Values
by: Liu, Yurong, et al.
Published: (2024)
by: Liu, Yurong, et al.
Published: (2024)
Swift Hydra: Self-Reinforcing Generative Framework for Anomaly Detection with Multiple Mamba Models
by: Do, Nguyen, et al.
Published: (2025)
by: Do, Nguyen, et al.
Published: (2025)
Uncertainty-aware Human Mobility Modeling and Anomaly Detection
by: Wen, Haomin, et al.
Published: (2024)
by: Wen, Haomin, et al.
Published: (2024)
Explainable AI-Guided Efficient Approximate DNN Generation for Multi-Pod Systolic Arrays
by: Siddique, Ayesha, et al.
Published: (2025)
by: Siddique, Ayesha, et al.
Published: (2025)
MetaOptimize: A Framework for Optimizing Step Sizes and Other Meta-parameters
by: Sharifnassab, Arsalan, et al.
Published: (2024)
by: Sharifnassab, Arsalan, et al.
Published: (2024)
Fast and Efficient Gossip Algorithms for Robust and Non-smooth Decentralized Learning
by: van Elst, Anna, et al.
Published: (2026)
by: van Elst, Anna, et al.
Published: (2026)
Fast and Robust Contextual Node Representation Learning over Dynamic Graphs
by: Guo, Xingzhi, et al.
Published: (2024)
by: Guo, Xingzhi, et al.
Published: (2024)
SFR-GNN: Simple and Fast Robust GNNs against Structural Attacks
by: Ai, Xing, et al.
Published: (2024)
by: Ai, Xing, et al.
Published: (2024)
EPSILON: Adaptive Fault Mitigation in Approximate Deep Neural Network using Statistical Signatures
by: Khalil, Khurram, et al.
Published: (2025)
by: Khalil, Khurram, et al.
Published: (2025)
Fast and Robust Likelihood-Guided Diffusion Posterior Sampling with Amortized Variational Inference
by: Zheng, Léon, et al.
Published: (2026)
by: Zheng, Léon, et al.
Published: (2026)
Policy Regularized Distributionally Robust Markov Decision Processes with Linear Function Approximation
by: Gu, Jingwen, et al.
Published: (2025)
by: Gu, Jingwen, et al.
Published: (2025)
FastTD3: Simple, Fast, and Capable Reinforcement Learning for Humanoid Control
by: Seo, Younggyo, et al.
Published: (2025)
by: Seo, Younggyo, et al.
Published: (2025)
Linear Mixture Distributionally Robust Markov Decision Processes
by: Liu, Zhishuai, et al.
Published: (2025)
by: Liu, Zhishuai, et al.
Published: (2025)
Auxiliary task discovery through generate-and-test
by: Rafiee, Banafsheh, et al.
Published: (2022)
by: Rafiee, Banafsheh, et al.
Published: (2022)
Intentional Updates for Streaming Reinforcement Learning
by: Sharifnassab, Arsalan, et al.
Published: (2026)
by: Sharifnassab, Arsalan, et al.
Published: (2026)
Fast and Robust: Task Sampling with Posterior and Diversity Synergies for Adaptive Decision-Makers in Randomized Environments
by: Qu, Yun, et al.
Published: (2025)
by: Qu, Yun, et al.
Published: (2025)
Robust Offline Reinforcement Learning with Linearly Structured f-Divergence Regularization
by: Tang, Cheng, et al.
Published: (2024)
by: Tang, Cheng, et al.
Published: (2024)
FastDSAC: Unlocking the Potential of Maximum Entropy RL in High-Dimensional Humanoid Control
by: Xue, Jun, et al.
Published: (2026)
by: Xue, Jun, et al.
Published: (2026)
Enhancing Autonomous Online Intrusion Detection for IoT with Balanced Learning, Reliable Pseudo-Labels, and Lightweight Architectures
by: Afzaal, Hanzala, et al.
Published: (2026)
by: Afzaal, Hanzala, et al.
Published: (2026)
SwiftRL: Towards Efficient Reinforcement Learning on Real Processing-In-Memory Systems
by: Gogineni, Kailash, et al.
Published: (2024)
by: Gogineni, Kailash, et al.
Published: (2024)
Robust Multi-Objective Controlled Decoding of Large Language Models
by: Son, Seongho, et al.
Published: (2025)
by: Son, Seongho, et al.
Published: (2025)
CSRA: Controlled Spectral Residual Augmentation for Robust Sepsis Prediction
by: Guo, Honglin, et al.
Published: (2026)
by: Guo, Honglin, et al.
Published: (2026)
On the Robustness of Transformers against Context Hijacking for Linear Classification
by: Li, Tianle, et al.
Published: (2025)
by: Li, Tianle, et al.
Published: (2025)
Robust Non-Linear Correlations via Polynomial Regression
by: Giuliani, Luca, et al.
Published: (2025)
by: Giuliani, Luca, et al.
Published: (2025)
Linearizing Models for Efficient yet Robust Private Inference
by: Sarkar, Sreetama, et al.
Published: (2024)
by: Sarkar, Sreetama, et al.
Published: (2024)
Enhancing Robustness of Federated Learning via Server Learning
by: Mai, Van Sy, et al.
Published: (2026)
by: Mai, Van Sy, et al.
Published: (2026)
Contextual Linear Bandits under Noisy Features: Towards Bayesian Oracles
by: Kim, Jung-hun, et al.
Published: (2017)
by: Kim, Jung-hun, et al.
Published: (2017)
FastGAS: Fast Graph-based Annotation Selection for In-Context Learning
by: Chen, Zihan, et al.
Published: (2024)
by: Chen, Zihan, et al.
Published: (2024)
Similar Items
-
Step-size Optimization for Continual Learning
by: Degris, Thomas, et al.
Published: (2024) -
SwiftKV: Fast Prefill-Optimized Inference with Knowledge-Preserving Model Transformation
by: Qiao, Aurick, et al.
Published: (2024) -
Can LLMs Reconcile Knowledge Conflicts in Counterfactual Reasoning
by: Yamin, Khurram, et al.
Published: (2025) -
SwiftF0: Fast and Accurate Monophonic Pitch Detection
by: Nieradzik, Lars
Published: (2025) -
Swift Sampler: Efficient Learning of Sampler by 10 Parameters
by: Yao, Jiawei, et al.
Published: (2024)