RadDQN: a Deep Q Learning-based Architecture for Finding Time-efficient Minimum Radiation Exposure Pathway
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Sadhu, Biswajit, Sadhu, Trijit, Anand, S. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Interpolation-Driven Machine Learning Approaches for Plume Shine Dose Estimation: A Comparison of XGBoost, Random Forest, and TabNet
von: Sadhu, Biswajit, et al.
Veröffentlicht: (2026)
von: Sadhu, Biswajit, et al.
Veröffentlicht: (2026)
Improving Audio Event Recognition with Consistency Regularization
von: Sadhu, Shanmuka, et al.
Veröffentlicht: (2025)
von: Sadhu, Shanmuka, et al.
Veröffentlicht: (2025)
Hierarchical Pedagogical Oversight: A Multi-Agent Adversarial Framework for Reliable AI Tutoring
von: Sadhu, Saisab, et al.
Veröffentlicht: (2025)
von: Sadhu, Saisab, et al.
Veröffentlicht: (2025)
KANQAS: Kolmogorov-Arnold Network for Quantum Architecture Search
von: Kundu, Akash, et al.
Veröffentlicht: (2024)
von: Kundu, Akash, et al.
Veröffentlicht: (2024)
$β$-DQN: Improving Deep Q-Learning By Evolving the Behavior
von: Zhang, Hongming, et al.
Veröffentlicht: (2025)
von: Zhang, Hongming, et al.
Veröffentlicht: (2025)
Task Decoding based on Eye Movements using Synthetic Data Augmentation
von: Sadhu, Shanmuka, et al.
Veröffentlicht: (2025)
von: Sadhu, Shanmuka, et al.
Veröffentlicht: (2025)
Weaker LLMs' Opinions Also Matter: Mixture of Opinions Enhances LLM's Mathematical Reasoning
von: Chen, Yanan, et al.
Veröffentlicht: (2025)
von: Chen, Yanan, et al.
Veröffentlicht: (2025)
ToMCAT: Theory-of-Mind for Cooperative Agents in Teams via Multiagent Diffusion Policies
von: Sequeira, Pedro, et al.
Veröffentlicht: (2025)
von: Sequeira, Pedro, et al.
Veröffentlicht: (2025)
Beyond ReLU: Chebyshev-DQN for Enhanced Deep Q-Networks
von: Yazdannik, Saman, et al.
Veröffentlicht: (2025)
von: Yazdannik, Saman, et al.
Veröffentlicht: (2025)
Athena: Safe Autonomous Agents with Verbal Contrastive Learning
von: Sadhu, Tanmana, et al.
Veröffentlicht: (2024)
von: Sadhu, Tanmana, et al.
Veröffentlicht: (2024)
Can We Rely on LLM Agents to Draft Long-Horizon Plans? Let's Take TravelPlanner as an Example
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
von: Chen, Yanan, et al.
Veröffentlicht: (2024)
Evaluation of Reinforcement Learning for Autonomous Penetration Testing using A3C, Q-learning and DQN
von: Becker, Norman, et al.
Veröffentlicht: (2024)
von: Becker, Norman, et al.
Veröffentlicht: (2024)
Modified Double DQN: addressing stability
von: Halat, Shervin, et al.
Veröffentlicht: (2021)
von: Halat, Shervin, et al.
Veröffentlicht: (2021)
A Controlled Study of Double DQN and Dueling DQN Under Cross-Environment Transfer
von: Nasir, Azkaa, et al.
Veröffentlicht: (2026)
von: Nasir, Azkaa, et al.
Veröffentlicht: (2026)
On the vanishing of Twisted negative K-theory and homotopy invariance
von: Sadhu, Vivek
Veröffentlicht: (2024)
von: Sadhu, Vivek
Veröffentlicht: (2024)
Value of Information-Enhanced Exploration in Bootstrapped DQN
von: Plataniotis, Stergios, et al.
Veröffentlicht: (2025)
von: Plataniotis, Stergios, et al.
Veröffentlicht: (2025)
Analysis of Bluffing by DQN and CFR in Leduc Hold'em Poker
von: Zaciragic, Tarik, et al.
Veröffentlicht: (2025)
von: Zaciragic, Tarik, et al.
Veröffentlicht: (2025)
ComTraQ-MPC: Meta-Trained DQN-MPC Integration for Trajectory Tracking with Limited Active Localization Updates
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2024)
von: Puthumanaillam, Gokul, et al.
Veröffentlicht: (2024)
RadHiera: Semantic Hierarchical Reinforcement Learning for Medical Report Generation
von: Du, Bodong, et al.
Veröffentlicht: (2025)
von: Du, Bodong, et al.
Veröffentlicht: (2025)
Improving the Diversity of Bootstrapped DQN by Replacing Priors With Noise
von: Meng, Li, et al.
Veröffentlicht: (2022)
von: Meng, Li, et al.
Veröffentlicht: (2022)
Hybrid DQN-TD3 Reinforcement Learning for Autonomous Navigation in Dynamic Environments
von: He, Xiaoyi, et al.
Veröffentlicht: (2025)
von: He, Xiaoyi, et al.
Veröffentlicht: (2025)
RadOnc-GPT: An Autonomous LLM Agent for Real-Time Patient Outcomes Labeling at Scale
von: Holmes, Jason, et al.
Veröffentlicht: (2025)
von: Holmes, Jason, et al.
Veröffentlicht: (2025)
SymDQN: Symbolic Knowledge and Reasoning in Neural Network-based Reinforcement Learning
von: Amador, Ivo, et al.
Veröffentlicht: (2025)
von: Amador, Ivo, et al.
Veröffentlicht: (2025)
DQNC2S: DQN-based Cross-stream Crisis event Summarizer
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
von: Cambrin, Daniele Rege, et al.
Veröffentlicht: (2024)
Finding Minimum-Cost Explanations for Predictions made by Tree Ensembles
von: Törnblom, John, et al.
Veröffentlicht: (2023)
von: Törnblom, John, et al.
Veröffentlicht: (2023)
Le développement rural dans l'Inde
von: Sadhu Singh Dhami
Veröffentlicht: (1954)
von: Sadhu Singh Dhami
Veröffentlicht: (1954)
Rural development in India
von: Sadhu Singh Dhami
Veröffentlicht: (1954)
von: Sadhu Singh Dhami
Veröffentlicht: (1954)
El fomento rural en la India
von: Sadhu Singh Dhami
Veröffentlicht: (1954)
von: Sadhu Singh Dhami
Veröffentlicht: (1954)
OpenRad: a Curated Repository of Open-access AI models for Radiology
von: Vrettos, Konstantinos, et al.
Veröffentlicht: (2026)
von: Vrettos, Konstantinos, et al.
Veröffentlicht: (2026)
DQN Performance with Epsilon Greedy Policies and Prioritized Experience Replay
von: Perkins, Daniel, et al.
Veröffentlicht: (2025)
von: Perkins, Daniel, et al.
Veröffentlicht: (2025)
Finding Nontrivial Minimum Fixed Points in Discrete Dynamical Systems
von: Qiu, Zirou, et al.
Veröffentlicht: (2023)
von: Qiu, Zirou, et al.
Veröffentlicht: (2023)
Finding the DeepDream for Time Series: Activation Maximization for Univariate Time Series
von: Schlegel, Udo, et al.
Veröffentlicht: (2024)
von: Schlegel, Udo, et al.
Veröffentlicht: (2024)
Toward Physics-Aware Deep Learning Architectures for LiDAR Intensity Simulation
von: Anand, Vivek, et al.
Veröffentlicht: (2024)
von: Anand, Vivek, et al.
Veröffentlicht: (2024)
Ensemble Elastic DQN: A novel multi-step ensemble approach to address overestimation in deep value-based reinforcement learning
von: Ly, Adrian, et al.
Veröffentlicht: (2025)
von: Ly, Adrian, et al.
Veröffentlicht: (2025)
Causal Deep Q Network
von: Khelifi, Elouanes, et al.
Veröffentlicht: (2025)
von: Khelifi, Elouanes, et al.
Veröffentlicht: (2025)
Any-Class Presence Likelihood for Robust Multi-Label Classification with Abundant Negative Data
von: Tissera, Dumindu, et al.
Veröffentlicht: (2025)
von: Tissera, Dumindu, et al.
Veröffentlicht: (2025)
Finding Clustering Algorithms in the Transformer Architecture
von: Clarkson, Kenneth L., et al.
Veröffentlicht: (2025)
von: Clarkson, Kenneth L., et al.
Veröffentlicht: (2025)
Sign-Separated Finite-Time Error Analysis of Q-Learning
von: Lee, Donghwan
Veröffentlicht: (2026)
von: Lee, Donghwan
Veröffentlicht: (2026)
Time After Time: Deep-Q Effect Estimation for Interventions on When and What to do
von: Wald, Yoav, et al.
Veröffentlicht: (2025)
von: Wald, Yoav, et al.
Veröffentlicht: (2025)
Early prediction of onset of sepsis in Clinical Setting
von: Mohammad, Fahim, et al.
Veröffentlicht: (2024)
von: Mohammad, Fahim, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Interpolation-Driven Machine Learning Approaches for Plume Shine Dose Estimation: A Comparison of XGBoost, Random Forest, and TabNet
von: Sadhu, Biswajit, et al.
Veröffentlicht: (2026) -
Improving Audio Event Recognition with Consistency Regularization
von: Sadhu, Shanmuka, et al.
Veröffentlicht: (2025) -
Hierarchical Pedagogical Oversight: A Multi-Agent Adversarial Framework for Reliable AI Tutoring
von: Sadhu, Saisab, et al.
Veröffentlicht: (2025) -
KANQAS: Kolmogorov-Arnold Network for Quantum Architecture Search
von: Kundu, Akash, et al.
Veröffentlicht: (2024) -
$β$-DQN: Improving Deep Q-Learning By Evolving the Behavior
von: Zhang, Hongming, et al.
Veröffentlicht: (2025)