FM-EAC: Feature Model-based Enhanced Actor-Critic for Multi-Task Control in Dynamic Environments
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Quanxi, Mao, Wencan, Tsukada, Manabu, Lui, John C. S., Ji, Yusheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
EIA-SEC: Improved Actor-Critic Framework for Multi-UAV Collaborative Control in Smart Agriculture
von: Zhou, Quanxi, et al.
Veröffentlicht: (2025)
von: Zhou, Quanxi, et al.
Veröffentlicht: (2025)
Trajectory Planning for UAV-Based Smart Farming Using Imitation-Based Triple Deep Q-Learning
von: Mao, Wencan, et al.
Veröffentlicht: (2025)
von: Mao, Wencan, et al.
Veröffentlicht: (2025)
Controlling Large Language Model-based Agents for Large-Scale Decision-Making: An Actor-Critic Approach
von: Zhang, Bin, et al.
Veröffentlicht: (2023)
von: Zhang, Bin, et al.
Veröffentlicht: (2023)
Large Language Models for Human-like Autonomous Driving: A Survey
von: Li, Yun, et al.
Veröffentlicht: (2024)
von: Li, Yun, et al.
Veröffentlicht: (2024)
K-Score: Kalman Filter as a Principled Alternative to Reward Normalization in Reinforcement Learning
von: Xia, Zixuan, et al.
Veröffentlicht: (2026)
von: Xia, Zixuan, et al.
Veröffentlicht: (2026)
Where Do You Go? Pedestrian Trajectory Prediction using Scene Features
von: Rezaei, Mohammad Ali, et al.
Veröffentlicht: (2025)
von: Rezaei, Mohammad Ali, et al.
Veröffentlicht: (2025)
Graph Attention-based Decentralized Actor-Critic for Dual-Objective Control of Multi-UAV Swarms
von: Peng, Haoran, et al.
Veröffentlicht: (2025)
von: Peng, Haoran, et al.
Veröffentlicht: (2025)
Multi-Agent Actor-Critic with Harmonic Annealing Pruning for Dynamic Spectrum Access Systems
von: Stamatelis, George, et al.
Veröffentlicht: (2025)
von: Stamatelis, George, et al.
Veröffentlicht: (2025)
Enhancing Decision-Making of Large Language Models via Actor-Critic
von: Dong, Heng, et al.
Veröffentlicht: (2025)
von: Dong, Heng, et al.
Veröffentlicht: (2025)
Multi-Robot Multi-Queue Control via Exhaustive Assignment Actor-Critic Learning
von: Merati, Mohammad, et al.
Veröffentlicht: (2026)
von: Merati, Mohammad, et al.
Veröffentlicht: (2026)
Large Language Model-Enhanced Multi-Armed Bandits
von: Sun, Jiahang, et al.
Veröffentlicht: (2025)
von: Sun, Jiahang, et al.
Veröffentlicht: (2025)
Enhancing Multi-Agent Collaboration with Attention-Based Actor-Critic Policies
von: Belinchon, Hugo Garrido-Lestache, et al.
Veröffentlicht: (2025)
von: Belinchon, Hugo Garrido-Lestache, et al.
Veröffentlicht: (2025)
Scalable Neighborhood-Based Multi-Agent Actor-Critic
von: Goppelsroeder, Tim, et al.
Veröffentlicht: (2026)
von: Goppelsroeder, Tim, et al.
Veröffentlicht: (2026)
Asymmetric Actor-Critic for Multi-turn LLM Agents
von: Jiang, Shuli, et al.
Veröffentlicht: (2026)
von: Jiang, Shuli, et al.
Veröffentlicht: (2026)
Revisiting Discrete Soft Actor-Critic
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
von: Zhou, Haibin, et al.
Veröffentlicht: (2022)
Broad Critic Deep Actor Reinforcement Learning for Continuous Control
von: Thalagala, Shiron, et al.
Veröffentlicht: (2024)
von: Thalagala, Shiron, et al.
Veröffentlicht: (2024)
Multi-Agent Actor-Critics in Autonomous Cyber Defense
von: Wang, Mingjun, et al.
Veröffentlicht: (2024)
von: Wang, Mingjun, et al.
Veröffentlicht: (2024)
MSACL: Multi-Step Actor-Critic Learning with Lyapunov Certificates for Exponentially Stabilizing Control
von: Zhang, Yongwei, et al.
Veröffentlicht: (2025)
von: Zhang, Yongwei, et al.
Veröffentlicht: (2025)
Actor-Critic based Online Data Mixing For Language Model Pre-Training
von: Ma, Jing, et al.
Veröffentlicht: (2025)
von: Ma, Jing, et al.
Veröffentlicht: (2025)
AC4MPC: Actor-Critic Reinforcement Learning for Nonlinear Model Predictive Control
von: Reiter, Rudolf, et al.
Veröffentlicht: (2024)
von: Reiter, Rudolf, et al.
Veröffentlicht: (2024)
Causal Scene Narration with Runtime Safety Supervision for Vision-Language-Action Driving
von: Li, Yun, et al.
Veröffentlicht: (2026)
von: Li, Yun, et al.
Veröffentlicht: (2026)
An Open-Source Modular Benchmark for Diffusion-Based Motion Planning in Closed-Loop Autonomous Driving
von: Li, Yun, et al.
Veröffentlicht: (2026)
von: Li, Yun, et al.
Veröffentlicht: (2026)
MapFM: Foundation Model-Driven HD Mapping with Multi-Task Contextual Learning
von: Ivanov, Leonid, et al.
Veröffentlicht: (2025)
von: Ivanov, Leonid, et al.
Veröffentlicht: (2025)
CA-AC-MPC: CUDA-Accelerated Actor-Critic Model Predictive Control
von: Buo, Antoonio, et al.
Veröffentlicht: (2026)
von: Buo, Antoonio, et al.
Veröffentlicht: (2026)
Online Efficient Safety-Critical Control for Mobile Robots in Unknown Dynamic Multi-Obstacle Environments
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
von: Zhang, Yu, et al.
Veröffentlicht: (2024)
Quality-Diversity Actor-Critic: Learning High-Performing and Diverse Behaviors via Value and Successor Features Critics
von: Grillotti, Luca, et al.
Veröffentlicht: (2024)
von: Grillotti, Luca, et al.
Veröffentlicht: (2024)
Functional Critics Are Essential for Actor-Critic: From Off-Policy Stability to Efficient Exploration
von: Bai, Qinxun, et al.
Veröffentlicht: (2025)
von: Bai, Qinxun, et al.
Veröffentlicht: (2025)
Multi-Agent Actor-Critic Generative AI for Query Resolution and Analysis
von: Rahman, Mohammad Wali Ur, et al.
Veröffentlicht: (2025)
von: Rahman, Mohammad Wali Ur, et al.
Veröffentlicht: (2025)
ACC-Collab: An Actor-Critic Approach to Multi-Agent LLM Collaboration
von: Estornell, Andrew, et al.
Veröffentlicht: (2024)
von: Estornell, Andrew, et al.
Veröffentlicht: (2024)
GTDE: Grouped Training with Decentralized Execution for Multi-agent Actor-Critic
von: Li, Mengxian, et al.
Veröffentlicht: (2024)
von: Li, Mengxian, et al.
Veröffentlicht: (2024)
EAC-MoE: Expert-Selection Aware Compressor for Mixture-of-Experts Large Language Models
von: Chen, Yuanteng, et al.
Veröffentlicht: (2025)
von: Chen, Yuanteng, et al.
Veröffentlicht: (2025)
Value Improved Actor Critic Algorithms
von: Oren, Yaniv, et al.
Veröffentlicht: (2024)
von: Oren, Yaniv, et al.
Veröffentlicht: (2024)
Diffusion Actor-Critic with Entropy Regulator
von: Wang, Yinuo, et al.
Veröffentlicht: (2024)
von: Wang, Yinuo, et al.
Veröffentlicht: (2024)
Average-Reward Soft Actor-Critic
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
von: Adamczyk, Jacob, et al.
Veröffentlicht: (2025)
GraphFM: Graph Factorization Machines for Feature Interaction Modeling
von: Wu, Shu, et al.
Veröffentlicht: (2021)
von: Wu, Shu, et al.
Veröffentlicht: (2021)
Monte Carlo Beam Search for Actor-Critic Reinforcement Learning in Continuous Control
von: Alzorgan, Hazim, et al.
Veröffentlicht: (2025)
von: Alzorgan, Hazim, et al.
Veröffentlicht: (2025)
Adviser-Actor-Critic: Eliminating Steady-State Error in Reinforcement Learning Control
von: Chen, Donghe, et al.
Veröffentlicht: (2025)
von: Chen, Donghe, et al.
Veröffentlicht: (2025)
Learning Decentralized LLM Collaboration with Multi-Agent Actor Critic
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
von: Liu, Shuo, et al.
Veröffentlicht: (2026)
Trust, Don't Trust, or Flip: Robust Preference-Based Reinforcement Learning with Multi-Expert Feedback
von: Hosseini, Seyed Amir, et al.
Veröffentlicht: (2026)
von: Hosseini, Seyed Amir, et al.
Veröffentlicht: (2026)
SToFM: a Multi-scale Foundation Model for Spatial Transcriptomics
von: Zhao, Suyuan, et al.
Veröffentlicht: (2025)
von: Zhao, Suyuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
EIA-SEC: Improved Actor-Critic Framework for Multi-UAV Collaborative Control in Smart Agriculture
von: Zhou, Quanxi, et al.
Veröffentlicht: (2025) -
Trajectory Planning for UAV-Based Smart Farming Using Imitation-Based Triple Deep Q-Learning
von: Mao, Wencan, et al.
Veröffentlicht: (2025) -
Controlling Large Language Model-based Agents for Large-Scale Decision-Making: An Actor-Critic Approach
von: Zhang, Bin, et al.
Veröffentlicht: (2023) -
Large Language Models for Human-like Autonomous Driving: A Survey
von: Li, Yun, et al.
Veröffentlicht: (2024) -
K-Score: Kalman Filter as a Principled Alternative to Reward Normalization in Reinforcement Learning
von: Xia, Zixuan, et al.
Veröffentlicht: (2026)