Saved in:
| Main Authors: | Uppalapati, Khartik, Abdulkareem, Shakeel, Yimenicioglu, Bora |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2510.06267 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
by: Yousaf, Iqra
Published: (2024)
by: Yousaf, Iqra
Published: (2024)
Stable and Privacy-Preserving Synthetic Educational Data with Empirical Marginals: A Copula-Based Approach
by: Ramos, Gabriel Diaz, et al.
Published: (2026)
by: Ramos, Gabriel Diaz, et al.
Published: (2026)
LeanProgress: Guiding Search for Neural Theorem Proving via Proof Progress Prediction
by: George, Robert Joseph, et al.
Published: (2025)
by: George, Robert Joseph, et al.
Published: (2025)
APC-GNN++: An Adaptive Patient-Centric GNN with Context-Aware Attention and Mini-Graph Explainability for Diabetes Classification
by: Berkani, Khaled
Published: (2025)
by: Berkani, Khaled
Published: (2025)
Tiny QA Benchmark++: Ultra-Lightweight, Synthetic Multilingual Dataset Generation & Smoke-Tests for Continuous LLM Evaluation
by: Koc, Vincent
Published: (2025)
by: Koc, Vincent
Published: (2025)
SAGE: Scale-Aware Gradual Evolution for Continual Knowledge Graph Embedding
by: Li, Yifei, et al.
Published: (2025)
by: Li, Yifei, et al.
Published: (2025)
Sketch Decompositions for Classical Planning via Deep Reinforcement Learning
by: Aichmüller, Michael, et al.
Published: (2024)
by: Aichmüller, Michael, et al.
Published: (2024)
Zero-Shot Context Generalization in Reinforcement Learning from Few Training Contexts
by: Chapman, James, et al.
Published: (2025)
by: Chapman, James, et al.
Published: (2025)
From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning
by: Tomashevskiy, Timofey
Published: (2026)
by: Tomashevskiy, Timofey
Published: (2026)
FlowRL: Flow-Augmented Few-Shot Reinforcement Learning for Semi-Structured Sensor Data
by: Pivezhandi, Mohammad, et al.
Published: (2024)
by: Pivezhandi, Mohammad, et al.
Published: (2024)
A Parallel Hybrid Action Space Reinforcement Learning Model for Real-world Adaptive Traffic Signal Control
by: Wang, Yuxuan, et al.
Published: (2025)
by: Wang, Yuxuan, et al.
Published: (2025)
Predicting and improving test-time scaling laws via reward tail-guided search
by: Li, Muheng, et al.
Published: (2026)
by: Li, Muheng, et al.
Published: (2026)
NeSIG: A Neuro-Symbolic Method for Learning to Generate Planning Problems
by: Núñez-Molina, Carlos, et al.
Published: (2023)
by: Núñez-Molina, Carlos, et al.
Published: (2023)
Constrained Auto-Bidding via Generative Response Modeling
by: Yang, Eunseok, et al.
Published: (2026)
by: Yang, Eunseok, et al.
Published: (2026)
Fractional Policy Gradients: Reinforcement Learning with Long-Term Memory
by: Pawar, Urvi, et al.
Published: (2025)
by: Pawar, Urvi, et al.
Published: (2025)
Learning to Select Goals in Automated Planning with Deep-Q Learning
by: Núñez-Molina, Carlos, et al.
Published: (2024)
by: Núñez-Molina, Carlos, et al.
Published: (2024)
Selective Progress-Aware Querying for Human-in-the-Loop Reinforcement Learning
by: Muraleedharan, Anujith, et al.
Published: (2025)
by: Muraleedharan, Anujith, et al.
Published: (2025)
SCULPT: Constraint-Guided Pruned MCTS that Carves Efficient Paths for Mathematical Reasoning
by: Fang, Qitong, et al.
Published: (2026)
by: Fang, Qitong, et al.
Published: (2026)
ScaleMAP: Preserving Local Density and Neighborhood Structure in Low-Dimensional Embeddings
by: Poorna, Rajas, et al.
Published: (2026)
by: Poorna, Rajas, et al.
Published: (2026)
Factored Diffusion Policies:Compositionally Generalized Robot Control with a Single Score Network
by: Mitra, Sayan, et al.
Published: (2026)
by: Mitra, Sayan, et al.
Published: (2026)
Evo-DKD: Dual-Knowledge Decoding for Autonomous Ontology Evolution in Large Language Models
by: Raman, Vishal, et al.
Published: (2025)
by: Raman, Vishal, et al.
Published: (2025)
From Demonstrations to Safe Deployment: Path-Consistent Safety Filtering for Diffusion Policies
by: Römer, Ralf, et al.
Published: (2025)
by: Römer, Ralf, et al.
Published: (2025)
The Final-Stage Bottleneck: A Systematic Dissection of the R-Learner for Network Causal Inference
by: Sairam, S, et al.
Published: (2025)
by: Sairam, S, et al.
Published: (2025)
From Next Token Prediction to (STRIPS) World Models
by: Núñez-Molina, Carlos, et al.
Published: (2025)
by: Núñez-Molina, Carlos, et al.
Published: (2025)
Autopilot-Preserving Residual Q-Learning with HJB-Inspired Finite-Action Risk Filtering for Fixed-Wing UAV Command Supervision
by: Iscan, Mehmet, et al.
Published: (2026)
by: Iscan, Mehmet, et al.
Published: (2026)
PIRS: Physics-Informed Reward Shaping for SAC-Based Building Energy Management
by: Zaregarizi, Shadmehr, et al.
Published: (2026)
by: Zaregarizi, Shadmehr, et al.
Published: (2026)
Are We Winning the Wrong Game? Revisiting Evaluation Practices for Long-Term Time Series Forecasting
by: Phungtua-eng, Thanapol, et al.
Published: (2026)
by: Phungtua-eng, Thanapol, et al.
Published: (2026)
GoldenStart: Q-Guided Priors and Entropy Control for Distilling Flow Policies
by: Zhang, He, et al.
Published: (2026)
by: Zhang, He, et al.
Published: (2026)
StepScorer: Accelerating Reinforcement Learning with Step-wise Scoring and Psychological Regret Modeling
by: Xu, Zhe
Published: (2026)
by: Xu, Zhe
Published: (2026)
Improving Industrial Injection Molding Processes with Explainable AI for Quality Classification
by: Rottenwalter, Georg, et al.
Published: (2025)
by: Rottenwalter, Georg, et al.
Published: (2025)
Novel Approaches to Artificial Intelligence Development Based on the Nearest Neighbor Method
by: Priezzhev, I. I., et al.
Published: (2025)
by: Priezzhev, I. I., et al.
Published: (2025)
Advancements in synthetic data extraction for industrial injection molding
by: Rottenwalter, Georg, et al.
Published: (2025)
by: Rottenwalter, Georg, et al.
Published: (2025)
Predicting Future Actions of Reinforcement Learning Agents
by: Chung, Stephen, et al.
Published: (2024)
by: Chung, Stephen, et al.
Published: (2024)
Machine Learning Based Path Planning for Improved Rover Navigation (Pre-Print Version)
by: Abcouwer, Neil, et al.
Published: (2020)
by: Abcouwer, Neil, et al.
Published: (2020)
Foundational Requirements for Artificial General Intelligence: A Falsifiable Framework Based on Signal Prediction
by: Šprogar, Matej
Published: (2025)
by: Šprogar, Matej
Published: (2025)
Deep Reinforcement Learning for Day-to-day Dynamic Tolling in Tradable Credit Schemes
by: Wu, Xiaoyi, et al.
Published: (2025)
by: Wu, Xiaoyi, et al.
Published: (2025)
Bilinear Mamba-Koopman Neural MPC for Varying Dynamics
by: Pagi, Matan, et al.
Published: (2026)
by: Pagi, Matan, et al.
Published: (2026)
Differentiable Symbolic Planning: A Neural Architecture for Constraint Reasoning with Learned Feasibility
by: Oruganti, Venkatakrishna Reddy
Published: (2026)
by: Oruganti, Venkatakrishna Reddy
Published: (2026)
From Theory to Practice with RAVEN-UCB: Addressing Non-Stationarity in Multi-Armed Bandits through Variance Adaptation
by: Fang, Junyi, et al.
Published: (2025)
by: Fang, Junyi, et al.
Published: (2025)
Embedded Safety-Aligned Intelligence via Differentiable Internal Alignment Embeddings
by: Rathva, Harsh, et al.
Published: (2025)
by: Rathva, Harsh, et al.
Published: (2025)
Similar Items
-
AI and Machine Learning Approaches for Predicting Nanoparticles Toxicity The Critical Role of Physiochemical Properties
by: Yousaf, Iqra
Published: (2024) -
Stable and Privacy-Preserving Synthetic Educational Data with Empirical Marginals: A Copula-Based Approach
by: Ramos, Gabriel Diaz, et al.
Published: (2026) -
LeanProgress: Guiding Search for Neural Theorem Proving via Proof Progress Prediction
by: George, Robert Joseph, et al.
Published: (2025) -
APC-GNN++: An Adaptive Patient-Centric GNN with Context-Aware Attention and Mini-Graph Explainability for Diabetes Classification
by: Berkani, Khaled
Published: (2025) -
Tiny QA Benchmark++: Ultra-Lightweight, Synthetic Multilingual Dataset Generation & Smoke-Tests for Continuous LLM Evaluation
by: Koc, Vincent
Published: (2025)