Saved in:
| Main Authors: | Akkerman, Fabian, Knofius, Nils, van der Heijden, Matthieu, Mes, Martijn |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2410.03887 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Learning Dynamic Selection and Pricing of Out-of-Home Deliveries
by: Akkerman, Fabian, et al.
Published: (2023)
by: Akkerman, Fabian, et al.
Published: (2023)
Dynamic Neighborhood Construction for Structured Large Discrete Action Spaces
by: Akkerman, Fabian, et al.
Published: (2023)
by: Akkerman, Fabian, et al.
Published: (2023)
Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces
by: Hoppe, Heiko, et al.
Published: (2026)
by: Hoppe, Heiko, et al.
Published: (2026)
Deep Reinforcement Learning for Solving Management Problems: Towards A Large Management Mode
by: Jiang, Jinyang, et al.
Published: (2024)
by: Jiang, Jinyang, et al.
Published: (2024)
Green AI in Action: Strategic Model Selection for Ensembles in Production
by: Nijkamp, Nienke, et al.
Published: (2024)
by: Nijkamp, Nienke, et al.
Published: (2024)
Revealing Interpretable Failure Modes of VLMs
by: Chaudhary, Isha, et al.
Published: (2026)
by: Chaudhary, Isha, et al.
Published: (2026)
Degradation Modeling and Prognostic Analysis Under Unknown Failure Modes
by: Fu, Ying, et al.
Published: (2024)
by: Fu, Ying, et al.
Published: (2024)
SED2AM: Solving Multi-Trip Time-Dependent Vehicle Routing Problem using Deep Reinforcement Learning
by: Mozhdehi, Arash, et al.
Published: (2025)
by: Mozhdehi, Arash, et al.
Published: (2025)
Mode-Dependent Rectification for Stable PPO Training
by: Mohamad, Mohamad, et al.
Published: (2026)
by: Mohamad, Mohamad, et al.
Published: (2026)
Graph of Thoughts: Solving Elaborate Problems with Large Language Models
by: Besta, Maciej, et al.
Published: (2023)
by: Besta, Maciej, et al.
Published: (2023)
Style Outweighs Substance: Failure Modes of LLM Judges in Alignment Benchmarking
by: Feuer, Benjamin, et al.
Published: (2024)
by: Feuer, Benjamin, et al.
Published: (2024)
MemFail: Stress-Testing Failure Modes of LLM Memory Systems
by: Garg, Ishir, et al.
Published: (2026)
by: Garg, Ishir, et al.
Published: (2026)
Matching Problems to Solutions: An Explainable Way of Solving Machine Learning Problems
by: Saleh, Lokman, et al.
Published: (2024)
by: Saleh, Lokman, et al.
Published: (2024)
The $\mathbf{Y}$-Combinator for LLMs: Solving Long-Context Rot with $λ$-Calculus
by: Roy, Amartya, et al.
Published: (2026)
by: Roy, Amartya, et al.
Published: (2026)
Unsupervised Learning for Solving the Travelling Salesman Problem
by: Min, Yimeng, et al.
Published: (2023)
by: Min, Yimeng, et al.
Published: (2023)
Diagnosing Failure Modes of Shared-State Collaboration in Resource-Constrained Visual Agents
by: Zhou, Yunpeng
Published: (2026)
by: Zhou, Yunpeng
Published: (2026)
Failure Modes in Multi-Hop QA: The Weakest Link Effect and the Recognition Bottleneck
by: Zhang, Meiru, et al.
Published: (2026)
by: Zhang, Meiru, et al.
Published: (2026)
Extreme Low-Bit Inference in Reasoning Models: Failure Modes and Targeted Recovery
by: Alimaskina, Ekaterina, et al.
Published: (2026)
by: Alimaskina, Ekaterina, et al.
Published: (2026)
Conditional Equivalence of DPO and RLHF: Implicit Assumption, Failure Modes, and Provable Alignment
by: Yang, Zhiqin, et al.
Published: (2026)
by: Yang, Zhiqin, et al.
Published: (2026)
Rate optimal learning of equilibria from data
by: Freihaut, Till, et al.
Published: (2025)
by: Freihaut, Till, et al.
Published: (2025)
Machine Learning Predictions for Traffic Equilibria in Road Renovation Scheduling
by: Bosch, Robbert, et al.
Published: (2025)
by: Bosch, Robbert, et al.
Published: (2025)
Metacognitive Capabilities of LLMs: An Exploration in Mathematical Problem Solving
by: Didolkar, Aniket, et al.
Published: (2024)
by: Didolkar, Aniket, et al.
Published: (2024)
Generalization in LLM Problem Solving: The Case of the Shortest Path
by: Tong, Yao, et al.
Published: (2026)
by: Tong, Yao, et al.
Published: (2026)
Solving Diffusion Inverse Problems with Restart Posterior Sampling
by: Ahmed, Bilal, et al.
Published: (2025)
by: Ahmed, Bilal, et al.
Published: (2025)
Query-Dependent Prompt Evaluation and Optimization with Offline Inverse RL
by: Sun, Hao, et al.
Published: (2023)
by: Sun, Hao, et al.
Published: (2023)
TRACE: Reconstruction-Based Anomaly Detection in Ensemble and Time-Dependent Simulations
by: Gadirov, Hamid, et al.
Published: (2026)
by: Gadirov, Hamid, et al.
Published: (2026)
Unleashing the Denoising Capability of Diffusion Prior for Solving Inverse Problems
by: Zhang, Jiawei, et al.
Published: (2024)
by: Zhang, Jiawei, et al.
Published: (2024)
Policy Split: Incentivizing Dual-Mode Exploration in LLM Reinforcement with Dual-Mode Entropy Regularization
by: Yao, Jiashu, et al.
Published: (2026)
by: Yao, Jiashu, et al.
Published: (2026)
Principled Data Augmentation for Learning to Solve Quadratic Programming Problems
by: Qian, Chendi, et al.
Published: (2025)
by: Qian, Chendi, et al.
Published: (2025)
Adversarial Generative Flow Network for Solving Vehicle Routing Problems
by: Zhang, Ni, et al.
Published: (2025)
by: Zhang, Ni, et al.
Published: (2025)
Learning to Solve Orienteering Problem with Time Windows and Variable Profits
by: Gao, Songqun, et al.
Published: (2026)
by: Gao, Songqun, et al.
Published: (2026)
Defending Against Unforeseen Failure Modes with Latent Adversarial Training
by: Casper, Stephen, et al.
Published: (2024)
by: Casper, Stephen, et al.
Published: (2024)
Smaug: Fixing Failure Modes of Preference Optimisation with DPO-Positive
by: Pal, Arka, et al.
Published: (2024)
by: Pal, Arka, et al.
Published: (2024)
Revisiting On-Policy Distillation: Empirical Failure Modes and Simple Fixes
by: Fu, Yuqian, et al.
Published: (2026)
by: Fu, Yuqian, et al.
Published: (2026)
Chain of Simulation: A Dual-Mode Reasoning Framework for Large Language Models with Dynamic Problem Routing
by: Sheikhi, Saeid
Published: (2026)
by: Sheikhi, Saeid
Published: (2026)
Towards Learning Foundation Models for Heuristic Functions to Solve Pathfinding Problems
by: Khandelwal, Vedant, et al.
Published: (2024)
by: Khandelwal, Vedant, et al.
Published: (2024)
Solving Probabilistic Verification Problems of Neural Networks using Branch and Bound
by: Boetius, David, et al.
Published: (2024)
by: Boetius, David, et al.
Published: (2024)
PeersimGym: An Environment for Solving the Task Offloading Problem with Reinforcement Learning
by: Metelo, Frederico, et al.
Published: (2024)
by: Metelo, Frederico, et al.
Published: (2024)
Learning Solution-Aware Transformers for Efficiently Solving Quadratic Assignment Problem
by: Tan, Zhentao, et al.
Published: (2024)
by: Tan, Zhentao, et al.
Published: (2024)
LLMOPT: Learning to Define and Solve General Optimization Problems from Scratch
by: Jiang, Caigao, et al.
Published: (2024)
by: Jiang, Caigao, et al.
Published: (2024)
Similar Items
-
Learning Dynamic Selection and Pricing of Out-of-Home Deliveries
by: Akkerman, Fabian, et al.
Published: (2023) -
Dynamic Neighborhood Construction for Structured Large Discrete Action Spaces
by: Akkerman, Fabian, et al.
Published: (2023) -
Breaking the Grid: Distance-Guided Reinforcement Learning in Large Discrete Action Spaces
by: Hoppe, Heiko, et al.
Published: (2026) -
Deep Reinforcement Learning for Solving Management Problems: Towards A Large Management Mode
by: Jiang, Jinyang, et al.
Published: (2024) -
Green AI in Action: Strategic Model Selection for Ensembles in Production
by: Nijkamp, Nienke, et al.
Published: (2024)