Can We Optimize Deep RL Policy Weights as Trajectory Modeling?
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Tang, Hongyao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SpikeRL: A Scalable and Energy-efficient Framework for Deep Spiking Reinforcement Learning
von: Tahmid, Tokey, et al.
Veröffentlicht: (2025)
von: Tahmid, Tokey, et al.
Veröffentlicht: (2025)
t-DGR: A Trajectory-Based Deep Generative Replay Method for Continual Learning in Decision Making
von: Yue, William, et al.
Veröffentlicht: (2024)
von: Yue, William, et al.
Veröffentlicht: (2024)
A Deep Dive into Effects of Structural Bias on CMA-ES Performance along Affine Trajectories
von: van Stein, Niki, et al.
Veröffentlicht: (2024)
von: van Stein, Niki, et al.
Veröffentlicht: (2024)
AM-PPO: (Advantage) Alpha-Modulation with Proximal Policy Optimization
von: Sane, Soham
Veröffentlicht: (2025)
von: Sane, Soham
Veröffentlicht: (2025)
Position: Leverage Foundational Models for Black-Box Optimization
von: Song, Xingyou, et al.
Veröffentlicht: (2024)
von: Song, Xingyou, et al.
Veröffentlicht: (2024)
Weight Sparsity Complements Activity Sparsity in Neuromorphic Language Models
von: Mukherji, Rishav, et al.
Veröffentlicht: (2024)
von: Mukherji, Rishav, et al.
Veröffentlicht: (2024)
DGPO: RL-Steered Graph Diffusion for Neural Architecture Generation
von: Liuliakov, Aleksei, et al.
Veröffentlicht: (2026)
von: Liuliakov, Aleksei, et al.
Veröffentlicht: (2026)
A Low Latency Adaptive Coding Spiking Framework for Deep Reinforcement Learning
von: Qin, Lang, et al.
Veröffentlicht: (2022)
von: Qin, Lang, et al.
Veröffentlicht: (2022)
Surrogate Benchmarks for Model Merging Optimization
von: Akizuki, Rio, et al.
Veröffentlicht: (2025)
von: Akizuki, Rio, et al.
Veröffentlicht: (2025)
Offline Model-Based Optimization by Learning to Rank
von: Tan, Rong-Xi, et al.
Veröffentlicht: (2024)
von: Tan, Rong-Xi, et al.
Veröffentlicht: (2024)
REACT: Revealing Evolutionary Action Consequence Trajectories for Interpretable Reinforcement Learning
von: Altmann, Philipp, et al.
Veröffentlicht: (2024)
von: Altmann, Philipp, et al.
Veröffentlicht: (2024)
Set-based Neural Network Encoding Without Weight Tying
von: Andreis, Bruno, et al.
Veröffentlicht: (2023)
von: Andreis, Bruno, et al.
Veröffentlicht: (2023)
Weight Decay Regimes in Grokking Transformers: Cheap Online Diagnostics
von: Verma, Lucky
Veröffentlicht: (2026)
von: Verma, Lucky
Veröffentlicht: (2026)
EOE: Evolutionary Optimization of Experts for Training Language Models
von: Chen, Yingshi
Veröffentlicht: (2025)
von: Chen, Yingshi
Veröffentlicht: (2025)
Evolved Sample Weights for Bias Mitigation: Effectiveness Depends on the Fairness Objective
von: Saini, Anil K., et al.
Veröffentlicht: (2025)
von: Saini, Anil K., et al.
Veröffentlicht: (2025)
Large Language Models As Evolution Strategies
von: Lange, Robert Tjarko, et al.
Veröffentlicht: (2024)
von: Lange, Robert Tjarko, et al.
Veröffentlicht: (2024)
Sensitivity-Aware Mixed-Precision Quantization and Width Optimization of Deep Neural Networks Through Cluster-Based Tree-Structured Parzen Estimation
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2023)
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2023)
Evaluating the Robustness of Deep-Learning Algorithm-Selection Models by Evolving Adversarial Instances
von: Hart, Emma, et al.
Veröffentlicht: (2024)
von: Hart, Emma, et al.
Veröffentlicht: (2024)
Recursive Dynamics in Fast-Weights Homeostatic Reentry Networks: Toward Reflective Intelligence
von: Chae, B. G.
Veröffentlicht: (2025)
von: Chae, B. G.
Veröffentlicht: (2025)
Beyond Single-Model Optimization: Preserving Plasticity in Continual Reinforcement Learning
von: Lillo, Lute, et al.
Veröffentlicht: (2026)
von: Lillo, Lute, et al.
Veröffentlicht: (2026)
Scalable Event-by-event Processing of Neuromorphic Sensory Signals With Deep State-Space Models
von: Schöne, Mark, et al.
Veröffentlicht: (2024)
von: Schöne, Mark, et al.
Veröffentlicht: (2024)
What Planning Problems Can A Relational Neural Network Solve?
von: Mao, Jiayuan, et al.
Veröffentlicht: (2023)
von: Mao, Jiayuan, et al.
Veröffentlicht: (2023)
Agent-GWO: Collaborative Agents for Dynamic Prompt Optimization in Large Language Models
von: Wang, Xudong, et al.
Veröffentlicht: (2026)
von: Wang, Xudong, et al.
Veröffentlicht: (2026)
A Firefly Algorithm for Mixed-Variable Optimization Based on Hybrid Distance Modeling
von: Bechir, Ousmane Tom, et al.
Veröffentlicht: (2026)
von: Bechir, Ousmane Tom, et al.
Veröffentlicht: (2026)
On-line Policy Improvement using Monte-Carlo Search
von: Tesauro, Gerald, et al.
Veröffentlicht: (2025)
von: Tesauro, Gerald, et al.
Veröffentlicht: (2025)
Discovering Effective Policies for Land-Use Planning with Neuroevolution
von: Young, Daniel, et al.
Veröffentlicht: (2023)
von: Young, Daniel, et al.
Veröffentlicht: (2023)
Towards Universal Offline Black-Box Optimization via Learning Language Model Embeddings
von: Tan, Rong-Xi, et al.
Veröffentlicht: (2025)
von: Tan, Rong-Xi, et al.
Veröffentlicht: (2025)
Lossless Compression of Neural Network Components: Weights, Checkpoints, and K/V Caches in Low-Precision Formats
von: Heilper, Anat, et al.
Veröffentlicht: (2025)
von: Heilper, Anat, et al.
Veröffentlicht: (2025)
Deep Oscillatory Neural Network
von: Rohan, Nurani Rajagopal, et al.
Veröffentlicht: (2024)
von: Rohan, Nurani Rajagopal, et al.
Veröffentlicht: (2024)
Deep Neural Regression Collapse
von: Rangamani, Akshay, et al.
Veröffentlicht: (2026)
von: Rangamani, Akshay, et al.
Veröffentlicht: (2026)
Optimizing PM2.5 Forecasting Accuracy with Hybrid Meta-Heuristic and Machine Learning Models
von: Ghafariasl, Parviz, et al.
Veröffentlicht: (2024)
von: Ghafariasl, Parviz, et al.
Veröffentlicht: (2024)
Towards White Box Deep Learning
von: Satkiewicz, Maciej
Veröffentlicht: (2024)
von: Satkiewicz, Maciej
Veröffentlicht: (2024)
Automated Deep Learning for Load Forecasting
von: Keisler, Julie, et al.
Veröffentlicht: (2024)
von: Keisler, Julie, et al.
Veröffentlicht: (2024)
Evolutionary Developmental Biology Can Serve as the Conceptual Foundation for a New Design Paradigm in Artificial Intelligence
von: Erden, Zeki Doruk, et al.
Veröffentlicht: (2025)
von: Erden, Zeki Doruk, et al.
Veröffentlicht: (2025)
Interpretable Fine-Gray Deep Survival Model for Competing Risks: Predicting Post-Discharge Foot Complications for Diabetic Patients in Ontario
von: Ramachandram, Dhanesh, et al.
Veröffentlicht: (2025)
von: Ramachandram, Dhanesh, et al.
Veröffentlicht: (2025)
Robust Lagrangian and Adversarial Policy Gradient for Robust Constrained Markov Decision Processes
von: Bossens, David M.
Veröffentlicht: (2023)
von: Bossens, David M.
Veröffentlicht: (2023)
Large Language Model as Meta-Surrogate for Data-Driven Many-Task Optimization: A Proof-of-Principle Study
von: Zhang, Xian-Rong, et al.
Veröffentlicht: (2025)
von: Zhang, Xian-Rong, et al.
Veröffentlicht: (2025)
Vector Policy Optimization: Training for Diversity Improves Test-Time Search
von: Bahlous-Boldi, Ryan, et al.
Veröffentlicht: (2026)
von: Bahlous-Boldi, Ryan, et al.
Veröffentlicht: (2026)
EvoJail: Evolutionary Diverse Jailbreak Prompt Generation for Large Language Models
von: Tang, Rui, et al.
Veröffentlicht: (2026)
von: Tang, Rui, et al.
Veröffentlicht: (2026)
Deep Reinforcement Learning with Spiking Q-learning
von: Chen, Ding, et al.
Veröffentlicht: (2022)
von: Chen, Ding, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
SpikeRL: A Scalable and Energy-efficient Framework for Deep Spiking Reinforcement Learning
von: Tahmid, Tokey, et al.
Veröffentlicht: (2025) -
t-DGR: A Trajectory-Based Deep Generative Replay Method for Continual Learning in Decision Making
von: Yue, William, et al.
Veröffentlicht: (2024) -
A Deep Dive into Effects of Structural Bias on CMA-ES Performance along Affine Trajectories
von: van Stein, Niki, et al.
Veröffentlicht: (2024) -
AM-PPO: (Advantage) Alpha-Modulation with Proximal Policy Optimization
von: Sane, Soham
Veröffentlicht: (2025) -
Position: Leverage Foundational Models for Black-Box Optimization
von: Song, Xingyou, et al.
Veröffentlicht: (2024)