Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
Fuente:
arXiv
Saved in:
| Main Authors: | Wigmore, Jerrod, Shrader, Brooke, Modiano, Eytan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
by: Hu, Yuelin, et al.
Published: (2026)
by: Hu, Yuelin, et al.
Published: (2026)
A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report
by: Wigmore, Jerrod, et al.
Published: (2025)
by: Wigmore, Jerrod, et al.
Published: (2025)
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
by: Zeytuncu, Yunus E.
Published: (2026)
by: Zeytuncu, Yunus E.
Published: (2026)
Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
by: Calvo, Anherutowa
Published: (2026)
by: Calvo, Anherutowa
Published: (2026)
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
by: Pratyush, Spandan
Published: (2026)
by: Pratyush, Spandan
Published: (2026)
Time-to-Injury Forecasting in Elite Female Football: A DeepHit Survival Approach
by: Catterall, Victoria, et al.
Published: (2026)
by: Catterall, Victoria, et al.
Published: (2026)
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
by: Kashyap, Ankit
Published: (2025)
by: Kashyap, Ankit
Published: (2025)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
by: Xu, Zhe
Published: (2026)
by: Xu, Zhe
Published: (2026)
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
by: Beck, Florentin, et al.
Published: (2025)
by: Beck, Florentin, et al.
Published: (2025)
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
by: Arbuzov, Mikhail L., et al.
Published: (2026)
by: Arbuzov, Mikhail L., et al.
Published: (2026)
Explaining Bayesian Optimization by Shapley Values Facilitates Human-AI Collaboration
by: Rodemann, Julian, et al.
Published: (2024)
by: Rodemann, Julian, et al.
Published: (2024)
CoupleEvo: Evolving Heuristics for Coupled Optimization Problems Using Large Language Models
by: Bömer, Thomas, et al.
Published: (2026)
by: Bömer, Thomas, et al.
Published: (2026)
Learning Nonlinear Regime Transitions via Semi-Parametric State-Space Models
by: Hiremath, Prakul Sunil
Published: (2026)
by: Hiremath, Prakul Sunil
Published: (2026)
Unified Unbiased Variance Estimation for Maximum Mean Discrepancy: Robust Finite-Sample Performance with Imbalanced Data and Exact Acceleration under Null and Alternative Hypotheses
by: Zhong, Shijie, et al.
Published: (2026)
by: Zhong, Shijie, et al.
Published: (2026)
Interactive LLM-assisted Curriculum Learning for Multi-Task Evolutionary Policy Search
by: Sakallioglu, Berfin, et al.
Published: (2026)
by: Sakallioglu, Berfin, et al.
Published: (2026)
DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
A Prescriptive Framework for Determining Optimal Days for Short-Term Traffic Counts
by: Mukwaya, Arthur, et al.
Published: (2025)
by: Mukwaya, Arthur, et al.
Published: (2025)
Vision-Guided Iterative Refinement for Frontend Code Generation
by: Sansford, Hannah, et al.
Published: (2026)
by: Sansford, Hannah, et al.
Published: (2026)
Evolving Programmatic Skill Networks
by: Shi, Haochen, et al.
Published: (2026)
by: Shi, Haochen, et al.
Published: (2026)
Optimizing Job Allocation using Reinforcement Learning with Graph Neural Networks
by: Quaedvlieg, Lars C. P. M.
Published: (2025)
by: Quaedvlieg, Lars C. P. M.
Published: (2025)
Enhanced and Efficient Reasoning in Large Learning Models
by: Valiant, Leslie G.
Published: (2026)
by: Valiant, Leslie G.
Published: (2026)
When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth
by: Guo, Dongxin, et al.
Published: (2026)
by: Guo, Dongxin, et al.
Published: (2026)
DRO-InstructZero: Distributionally Robust Prompt Optimization for Large Language Models
by: Li, Yangyang
Published: (2025)
by: Li, Yangyang
Published: (2025)
Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation
by: Chen, Mu-Chi, et al.
Published: (2026)
by: Chen, Mu-Chi, et al.
Published: (2026)
Understanding and Tackling Over-Dilution in Graph Neural Networks
by: Lee, Junhyun, et al.
Published: (2025)
by: Lee, Junhyun, et al.
Published: (2025)
On the approximation capability of GNNs in node classification/regression tasks
by: D'Inverno, Giuseppe Alessio, et al.
Published: (2021)
by: D'Inverno, Giuseppe Alessio, et al.
Published: (2021)
Graph Coloring for Multi-Task Learning
by: Patapati, Santosh
Published: (2025)
by: Patapati, Santosh
Published: (2025)
Alignment Verifiability in Large Language Models: Normative Indistinguishability under Behavioral Evaluation
by: Santos-Grueiro, Igor
Published: (2026)
by: Santos-Grueiro, Igor
Published: (2026)
Diffusion-MPC in Discrete Domains: Feasibility Constraints, Horizon Effects, and Critic Alignment: Case study with Tetris
by: Wang, Haochuan Kevin
Published: (2026)
by: Wang, Haochuan Kevin
Published: (2026)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
by: Lauffer, Niklas, et al.
Published: (2025)
by: Lauffer, Niklas, et al.
Published: (2025)
CBR -- Boosting Adaptive Classification By Retrieval of Encrypted Network Traffic with Out-of-distribution
by: Lukach, Amir, et al.
Published: (2024)
by: Lukach, Amir, et al.
Published: (2024)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
by: Tiwari, Dhruv
Published: (2025)
by: Tiwari, Dhruv
Published: (2025)
Scattered Forest Search: Smarter Code Space Exploration with LLMs
by: Light, Jonathan, et al.
Published: (2024)
by: Light, Jonathan, et al.
Published: (2024)
Causality-Driven Neural Network Repair: Challenges and Opportunities
by: Vares, Fatemeh, et al.
Published: (2025)
by: Vares, Fatemeh, et al.
Published: (2025)
RoboGrind: Intuitive and Interactive Surface Treatment with Industrial Robots
by: Alt, Benjamin, et al.
Published: (2024)
by: Alt, Benjamin, et al.
Published: (2024)
Chain of Unit-Physics: A Primitive-Centric Approach to Scientific Code Synthesis
by: Sharma, Vansh, et al.
Published: (2025)
by: Sharma, Vansh, et al.
Published: (2025)
SCOPE: Selective Conformal Optimized Pairwise LLM Judging
by: Badshah, Sher, et al.
Published: (2026)
by: Badshah, Sher, et al.
Published: (2026)
Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs
by: Kaiser, Daniel, et al.
Published: (2026)
by: Kaiser, Daniel, et al.
Published: (2026)
Graph Neural Networks are Heuristics
by: Min, Yimeng, et al.
Published: (2026)
by: Min, Yimeng, et al.
Published: (2026)
A Multidisciplinary Approach to Telegram Data Analysis
by: Varbanov, Velizar, et al.
Published: (2024)
by: Varbanov, Velizar, et al.
Published: (2024)
Similar Items
-
Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
by: Hu, Yuelin, et al.
Published: (2026) -
A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report
by: Wigmore, Jerrod, et al.
Published: (2025) -
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
by: Zeytuncu, Yunus E.
Published: (2026) -
Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
by: Calvo, Anherutowa
Published: (2026) -
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
by: Pratyush, Spandan
Published: (2026)