Model-Agnostic Solutions for Deep Reinforcement Learning in Non-Ergodic Contexts
Fuente:
arXiv
Saved in:
| Main Authors: | Verbruggen, Bert, Vanhoyweghen, Arne, Ginis, Vincent |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026)
by: Baumann, Dominik, et al.
Published: (2026)
Lexical Hints of Accuracy in LLM Reasoning Chains
by: Vanhoyweghen, Arne, et al.
Published: (2025)
by: Vanhoyweghen, Arne, et al.
Published: (2025)
Metro 3 in Brussels under uncertainty: scenario-based public transport accessibility analysis
by: Verbeken, Brecht, et al.
Published: (2025)
by: Verbeken, Brecht, et al.
Published: (2025)
The Effectiveness of Curvature-Based Rewiring and the Role of Hyperparameters in GNNs Revisited
by: Tori, Floriano, et al.
Published: (2024)
by: Tori, Floriano, et al.
Published: (2024)
The Relationship Between Reasoning and Performance in Large Language Models -- o3 (mini) Thinks Harder, Not Longer
by: Ballon, Marthe, et al.
Published: (2025)
by: Ballon, Marthe, et al.
Published: (2025)
Probing the Trajectories of Reasoning Traces in Large Language Models
by: Ballon, Marthe, et al.
Published: (2026)
by: Ballon, Marthe, et al.
Published: (2026)
Continual Deep Reinforcement Learning with Task-Agnostic Policy Distillation
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
by: Hafez, Muhammad Burhan, et al.
Published: (2024)
Estimating problem difficulty without ground truth using Large Language Model comparisons
by: Ballon, Marthe, et al.
Published: (2025)
by: Ballon, Marthe, et al.
Published: (2025)
Flexible Counterfactual Explanations with Generative Models
by: Hellemans, Stig, et al.
Published: (2025)
by: Hellemans, Stig, et al.
Published: (2025)
Model-Agnostic Zeroth-Order Policy Optimization for Meta-Learning of Ergodic Linear Quadratic Regulators
by: Pan, Yunian, et al.
Published: (2024)
by: Pan, Yunian, et al.
Published: (2024)
Structurally Human, Semantically Biased: Detecting LLM-Generated References with Embeddings and GNNs
by: Mobini, Melika, et al.
Published: (2026)
by: Mobini, Melika, et al.
Published: (2026)
Benchmarks Saturate When The Model Gets Smarter Than The Judge
by: Ballon, Marthe, et al.
Published: (2026)
by: Ballon, Marthe, et al.
Published: (2026)
Constrained Meta Agnostic Reinforcement Learning
by: Daaboul, Karam, et al.
Published: (2024)
by: Daaboul, Karam, et al.
Published: (2024)
How Deep Do Large Language Models Internalize Scientific Literature and Citation Practices?
by: Algaba, Andres, et al.
Published: (2025)
by: Algaba, Andres, et al.
Published: (2025)
Agnostic Reinforcement Learning: Foundations and Algorithms
by: Li, Gene
Published: (2025)
by: Li, Gene
Published: (2025)
Progressive Safeguards for Safe and Model-Agnostic Reinforcement Learning
by: Omi, Nabil, et al.
Published: (2024)
by: Omi, Nabil, et al.
Published: (2024)
In-Context Reinforcement Learning through Bayesian Fusion of Context and Value Prior
by: Berkes, Anaïs, et al.
Published: (2026)
by: Berkes, Anaïs, et al.
Published: (2026)
The Role of Environment Access in Agnostic Reinforcement Learning
by: Krishnamurthy, Akshay, et al.
Published: (2025)
by: Krishnamurthy, Akshay, et al.
Published: (2025)
Variable-Agnostic Causal Exploration for Reinforcement Learning
by: Nguyen, Minh Hoang, et al.
Published: (2024)
by: Nguyen, Minh Hoang, et al.
Published: (2024)
No $D_{\text{train}}$: Model-Agnostic Counterfactual Explanations Using Reinforcement Learning
by: Sun, Xiangyu, et al.
Published: (2024)
by: Sun, Xiangyu, et al.
Published: (2024)
Anomaly Detection in Power Grids via Context-Agnostic Learning
by: Park, SangWoo, et al.
Published: (2024)
by: Park, SangWoo, et al.
Published: (2024)
Statistical Context Detection for Deep Lifelong Reinforcement Learning
by: Dick, Jeffery, et al.
Published: (2024)
by: Dick, Jeffery, et al.
Published: (2024)
Large Language Models Reflect Human Citation Patterns with a Heightened Citation Bias
by: Algaba, Andres, et al.
Published: (2024)
by: Algaba, Andres, et al.
Published: (2024)
Deep Transfer $Q$-Learning for Offline Non-Stationary Reinforcement Learning
by: Chai, Jinhang, et al.
Published: (2025)
by: Chai, Jinhang, et al.
Published: (2025)
From Lab to Reality: A Practical Evaluation of Deep Learning Models and LLMs for Vulnerability Detection
by: Lu, Chaomeng, et al.
Published: (2025)
by: Lu, Chaomeng, et al.
Published: (2025)
Task-Agnostic Contrastive Pretraining for Relational Deep Learning
by: Peleška, Jakub, et al.
Published: (2025)
by: Peleška, Jakub, et al.
Published: (2025)
Early Evidence of Vibe-Proving with Consumer LLMs: A Case Study on Spectral Region Characterization with ChatGPT-5.2 (Thinking)
by: Verbeken, Brecht, et al.
Published: (2026)
by: Verbeken, Brecht, et al.
Published: (2026)
Probing Graph Neural Network Activation Patterns Through Graph Topology
by: Tori, Floriano, et al.
Published: (2026)
by: Tori, Floriano, et al.
Published: (2026)
In-Context Function Learning in Large Language Models
by: Akata, Elif, et al.
Published: (2026)
by: Akata, Elif, et al.
Published: (2026)
Batch-ICL: Effective, Efficient, and Order-Agnostic In-Context Learning
by: Zhang, Kaiyi, et al.
Published: (2024)
by: Zhang, Kaiyi, et al.
Published: (2024)
MAML-en-LLM: Model Agnostic Meta-Training of LLMs for Improved In-Context Learning
by: Sinha, Sanchit, et al.
Published: (2024)
by: Sinha, Sanchit, et al.
Published: (2024)
Online Reinforcement Learning in Non-Stationary Context-Driven Environments
by: Hamadanian, Pouya, et al.
Published: (2023)
by: Hamadanian, Pouya, et al.
Published: (2023)
Human-in-the-Loop LLM Grading for Handwritten Mathematics Assessments
by: Vanhoyweghen, Arne, et al.
Published: (2026)
by: Vanhoyweghen, Arne, et al.
Published: (2026)
In-Context Deep Learning via Transformer Models
by: Wu, Weimin, et al.
Published: (2024)
by: Wu, Weimin, et al.
Published: (2024)
Deep Reinforcement Learning based Triggering Function for Early Classifiers of Time Series
by: Renault, Aurélien, et al.
Published: (2025)
by: Renault, Aurélien, et al.
Published: (2025)
DLBacktrace: A Model Agnostic Explainability for any Deep Learning Models
by: Sankarapu, Vinay Kumar, et al.
Published: (2024)
by: Sankarapu, Vinay Kumar, et al.
Published: (2024)
DeepOSets: Non-Autoregressive In-Context Learning with Permutation-Invariance Inductive Bias
by: Chiu, Shao-Ting, et al.
Published: (2024)
by: Chiu, Shao-Ting, et al.
Published: (2024)
Context Bootstrapped Reinforcement Learning
by: Agashe, Saaket, et al.
Published: (2026)
by: Agashe, Saaket, et al.
Published: (2026)
Safe In-Context Reinforcement Learning
by: Moeini, Amir, et al.
Published: (2025)
by: Moeini, Amir, et al.
Published: (2025)
Ergodic Risk Measures: Towards a Risk-Aware Foundation for Continual Reinforcement Learning
by: Rojas, Juan Sebastian, et al.
Published: (2025)
by: Rojas, Juan Sebastian, et al.
Published: (2025)
Similar Items
-
Ergodicity in reinforcement learning
by: Baumann, Dominik, et al.
Published: (2026) -
Lexical Hints of Accuracy in LLM Reasoning Chains
by: Vanhoyweghen, Arne, et al.
Published: (2025) -
Metro 3 in Brussels under uncertainty: scenario-based public transport accessibility analysis
by: Verbeken, Brecht, et al.
Published: (2025) -
The Effectiveness of Curvature-Based Rewiring and the Role of Hyperparameters in GNNs Revisited
by: Tori, Floriano, et al.
Published: (2024) -
The Relationship Between Reasoning and Performance in Large Language Models -- o3 (mini) Thinks Harder, Not Longer
by: Ballon, Marthe, et al.
Published: (2025)