Intervention-Assisted Policy Gradient Methods for Online Stochastic Queuing Network Optimization: Technical Report
Fuente:
arXiv
Guardado en:
| Autores principales: | Wigmore, Jerrod, Shrader, Brooke, Modiano, Eytan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
por: Hu, Yuelin, et al.
Publicado: (2026)
por: Hu, Yuelin, et al.
Publicado: (2026)
A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report
por: Wigmore, Jerrod, et al.
Publicado: (2025)
por: Wigmore, Jerrod, et al.
Publicado: (2025)
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
por: Zeytuncu, Yunus E.
Publicado: (2026)
por: Zeytuncu, Yunus E.
Publicado: (2026)
Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
por: Calvo, Anherutowa
Publicado: (2026)
por: Calvo, Anherutowa
Publicado: (2026)
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
por: Pratyush, Spandan
Publicado: (2026)
por: Pratyush, Spandan
Publicado: (2026)
Time-to-Injury Forecasting in Elite Female Football: A DeepHit Survival Approach
por: Catterall, Victoria, et al.
Publicado: (2026)
por: Catterall, Victoria, et al.
Publicado: (2026)
Recurrent Memory-Augmented Transformers with Chunked Attention for Long-Context Language Modeling
por: Kashyap, Ankit
Publicado: (2025)
por: Kashyap, Ankit
Publicado: (2025)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
por: Xu, Zhe
Publicado: (2026)
por: Xu, Zhe
Publicado: (2026)
TRIM: Achieving Extreme Sparsity with Targeted Row-wise Iterative Metric-driven Pruning
por: Beck, Florentin, et al.
Publicado: (2025)
por: Beck, Florentin, et al.
Publicado: (2025)
The Architecture of Errors: From Universal Impossibility to Patch-Local LLM Reliability
por: Arbuzov, Mikhail L., et al.
Publicado: (2026)
por: Arbuzov, Mikhail L., et al.
Publicado: (2026)
Explaining Bayesian Optimization by Shapley Values Facilitates Human-AI Collaboration
por: Rodemann, Julian, et al.
Publicado: (2024)
por: Rodemann, Julian, et al.
Publicado: (2024)
CoupleEvo: Evolving Heuristics for Coupled Optimization Problems Using Large Language Models
por: Bömer, Thomas, et al.
Publicado: (2026)
por: Bömer, Thomas, et al.
Publicado: (2026)
Learning Nonlinear Regime Transitions via Semi-Parametric State-Space Models
por: Hiremath, Prakul Sunil
Publicado: (2026)
por: Hiremath, Prakul Sunil
Publicado: (2026)
Unified Unbiased Variance Estimation for Maximum Mean Discrepancy: Robust Finite-Sample Performance with Imbalanced Data and Exact Acceleration under Null and Alternative Hypotheses
por: Zhong, Shijie, et al.
Publicado: (2026)
por: Zhong, Shijie, et al.
Publicado: (2026)
Interactive LLM-assisted Curriculum Learning for Multi-Task Evolutionary Policy Search
por: Sakallioglu, Berfin, et al.
Publicado: (2026)
por: Sakallioglu, Berfin, et al.
Publicado: (2026)
DELTA: Variational Disentangled Learning for Privacy-Preserving Data Reprogramming
por: Malarkkan, Arun Vignesh, et al.
Publicado: (2025)
por: Malarkkan, Arun Vignesh, et al.
Publicado: (2025)
A Prescriptive Framework for Determining Optimal Days for Short-Term Traffic Counts
por: Mukwaya, Arthur, et al.
Publicado: (2025)
por: Mukwaya, Arthur, et al.
Publicado: (2025)
Vision-Guided Iterative Refinement for Frontend Code Generation
por: Sansford, Hannah, et al.
Publicado: (2026)
por: Sansford, Hannah, et al.
Publicado: (2026)
Evolving Programmatic Skill Networks
por: Shi, Haochen, et al.
Publicado: (2026)
por: Shi, Haochen, et al.
Publicado: (2026)
Optimizing Job Allocation using Reinforcement Learning with Graph Neural Networks
por: Quaedvlieg, Lars C. P. M.
Publicado: (2025)
por: Quaedvlieg, Lars C. P. M.
Publicado: (2025)
Enhanced and Efficient Reasoning in Large Learning Models
por: Valiant, Leslie G.
Publicado: (2026)
por: Valiant, Leslie G.
Publicado: (2026)
When Do Early-Exit Networks Generalize? A PAC-Bayesian Theory of Adaptive Depth
por: Guo, Dongxin, et al.
Publicado: (2026)
por: Guo, Dongxin, et al.
Publicado: (2026)
DRO-InstructZero: Distributionally Robust Prompt Optimization for Large Language Models
por: Li, Yangyang
Publicado: (2025)
por: Li, Yangyang
Publicado: (2025)
Exploring LLM-based Verilog Code Generation with Data-Efficient Fine-Tuning and Testbench Automation
por: Chen, Mu-Chi, et al.
Publicado: (2026)
por: Chen, Mu-Chi, et al.
Publicado: (2026)
Understanding and Tackling Over-Dilution in Graph Neural Networks
por: Lee, Junhyun, et al.
Publicado: (2025)
por: Lee, Junhyun, et al.
Publicado: (2025)
On the approximation capability of GNNs in node classification/regression tasks
por: D'Inverno, Giuseppe Alessio, et al.
Publicado: (2021)
por: D'Inverno, Giuseppe Alessio, et al.
Publicado: (2021)
Graph Coloring for Multi-Task Learning
por: Patapati, Santosh
Publicado: (2025)
por: Patapati, Santosh
Publicado: (2025)
Alignment Verifiability in Large Language Models: Normative Indistinguishability under Behavioral Evaluation
por: Santos-Grueiro, Igor
Publicado: (2026)
por: Santos-Grueiro, Igor
Publicado: (2026)
Diffusion-MPC in Discrete Domains: Feasibility Constraints, Horizon Effects, and Critic Alignment: Case study with Tetris
por: Wang, Haochuan Kevin
Publicado: (2026)
por: Wang, Haochuan Kevin
Publicado: (2026)
Robust and Diverse Multi-Agent Learning via Rational Policy Gradient
por: Lauffer, Niklas, et al.
Publicado: (2025)
por: Lauffer, Niklas, et al.
Publicado: (2025)
CBR -- Boosting Adaptive Classification By Retrieval of Encrypted Network Traffic with Out-of-distribution
por: Lukach, Amir, et al.
Publicado: (2024)
por: Lukach, Amir, et al.
Publicado: (2024)
Emotion-Inspired Learning Signals (EILS): A Homeostatic Framework for Adaptive Autonomous Agents
por: Tiwari, Dhruv
Publicado: (2025)
por: Tiwari, Dhruv
Publicado: (2025)
Scattered Forest Search: Smarter Code Space Exploration with LLMs
por: Light, Jonathan, et al.
Publicado: (2024)
por: Light, Jonathan, et al.
Publicado: (2024)
Causality-Driven Neural Network Repair: Challenges and Opportunities
por: Vares, Fatemeh, et al.
Publicado: (2025)
por: Vares, Fatemeh, et al.
Publicado: (2025)
RoboGrind: Intuitive and Interactive Surface Treatment with Industrial Robots
por: Alt, Benjamin, et al.
Publicado: (2024)
por: Alt, Benjamin, et al.
Publicado: (2024)
Chain of Unit-Physics: A Primitive-Centric Approach to Scientific Code Synthesis
por: Sharma, Vansh, et al.
Publicado: (2025)
por: Sharma, Vansh, et al.
Publicado: (2025)
SCOPE: Selective Conformal Optimized Pairwise LLM Judging
por: Badshah, Sher, et al.
Publicado: (2026)
por: Badshah, Sher, et al.
Publicado: (2026)
Beyond Accuracy: Decomposing the Reasoning Efficiency of LLMs
por: Kaiser, Daniel, et al.
Publicado: (2026)
por: Kaiser, Daniel, et al.
Publicado: (2026)
Graph Neural Networks are Heuristics
por: Min, Yimeng, et al.
Publicado: (2026)
por: Min, Yimeng, et al.
Publicado: (2026)
A Multidisciplinary Approach to Telegram Data Analysis
por: Varbanov, Velizar, et al.
Publicado: (2024)
por: Varbanov, Velizar, et al.
Publicado: (2024)
Ejemplares similares
-
Hidden Failure Modes of Gradient Modification under Adam in Continual Learning, and Adaptive Decoupled Moment Routing as a Repair
por: Hu, Yuelin, et al.
Publicado: (2026) -
A Novel Switch-Type Policy Network for Resource Allocation Problems: Technical Report
por: Wigmore, Jerrod, et al.
Publicado: (2025) -
4OPS: Structural Difficulty Modeling in Integer Arithmetic Puzzles
por: Zeytuncu, Yunus E.
Publicado: (2026) -
Predictable Gradient Manifolds in Deep Learning: Temporal Path-Length and Intrinsic Rank as a Complexity Regime
por: Calvo, Anherutowa
Publicado: (2026) -
Grammatically-Guided Sparse Attention for Efficient and Interpretable Transformers
por: Pratyush, Spandan
Publicado: (2026)