Multi-agent transformer-accelerated RL for satisfaction of STL specifications
Fuente:
arXiv
Saved in:
| Main Authors: | Forsberg, Albin Larsson, Nikou, Alexandros, Feljan, Aneta Vulgarakis, Tumova, Jana |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
A Survey on the Integration of Generative AI for Critical Thinking in Mobile Networks
by: Karapantelakis, Athanasios, et al.
Published: (2024)
by: Karapantelakis, Athanasios, et al.
Published: (2024)
Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines
by: Levina, Kristina, et al.
Published: (2025)
by: Levina, Kristina, et al.
Published: (2025)
Numeric Reward Machines
by: Levina, Kristina, et al.
Published: (2024)
by: Levina, Kristina, et al.
Published: (2024)
Reinforcement Learning with Reward Machines for Sleep Control in Mobile Networks
by: Levina, Kristina, et al.
Published: (2026)
by: Levina, Kristina, et al.
Published: (2026)
Scale When Needed: Adaptive Neuron-level Mixed Precision Quantization Aware Training
by: Varshney, Ayush K., et al.
Published: (2026)
by: Varshney, Ayush K., et al.
Published: (2026)
TTT: A Temporal Refinement Heuristic for Tenuously Tractable Discrete Time Reachability Problems
by: Sidrane, Chelsea, et al.
Published: (2024)
by: Sidrane, Chelsea, et al.
Published: (2024)
BURNS: Backward Underapproximate Reachability for Neural-Feedback-Loop Systems
by: Sidrane, Chelsea, et al.
Published: (2025)
by: Sidrane, Chelsea, et al.
Published: (2025)
When to restart? Exploring escalating restarts on convergence
by: Varshney, Ayush K., et al.
Published: (2026)
by: Varshney, Ayush K., et al.
Published: (2026)
Robust STL Control Synthesis under Maximal Disturbance Sets
by: Verhagen, Joris, et al.
Published: (2024)
by: Verhagen, Joris, et al.
Published: (2024)
Cooperative Multi-agent RL with Communication Constraints
by: Xiong, Nuoya, et al.
Published: (2026)
by: Xiong, Nuoya, et al.
Published: (2026)
Contrastive and Transfer Learning for Effective Audio Fingerprinting through a Real-World Evaluation Protocol
by: Nikou, Christos, et al.
Published: (2025)
by: Nikou, Christos, et al.
Published: (2025)
Validating Generalist Robots with Situation Calculus and STL Falsification
by: Li, Changwen, et al.
Published: (2026)
by: Li, Changwen, et al.
Published: (2026)
MUA-RL: Multi-turn User-interacting Agent Reinforcement Learning for agentic tool use
by: Zhao, Weikang, et al.
Published: (2025)
by: Zhao, Weikang, et al.
Published: (2025)
Using Large Language Models to Understand Telecom Standards
by: Karapantelakis, Athanasios, et al.
Published: (2024)
by: Karapantelakis, Athanasios, et al.
Published: (2024)
Designing Practical Models for Isolated Word Visual Speech Recognition
by: Panagos, Iason Ioannis, et al.
Published: (2025)
by: Panagos, Iason Ioannis, et al.
Published: (2025)
Research on Metro Transportation Flow Prediction Based on the STL-GRU Combined Model
by: Zhou, Zijie, et al.
Published: (2025)
by: Zhou, Zijie, et al.
Published: (2025)
Towards Socially and Morally Aware RL agent: Reward Design With LLM
by: Wang, Zhaoyue
Published: (2024)
by: Wang, Zhaoyue
Published: (2024)
Lightweight Operations for Visual Speech Recognition
by: Panagos, Iason Ioannis, et al.
Published: (2025)
by: Panagos, Iason Ioannis, et al.
Published: (2025)
SkyRL-Agent: Efficient RL Training for Multi-turn LLM Agent
by: Cao, Shiyi, et al.
Published: (2025)
by: Cao, Shiyi, et al.
Published: (2025)
Traffic flow forecasting, STL decomposition, Hybrid model, LSTM, ARIMA, XGBoost, Intelligent transportation systems
by: Yuan, Fujiang, et al.
Published: (2025)
by: Yuan, Fujiang, et al.
Published: (2025)
ProRL Agent: Rollout-as-a-Service for RL Training of Multi-Turn LLM Agents
by: Zhang, Hao, et al.
Published: (2026)
by: Zhang, Hao, et al.
Published: (2026)
Grounding Natural Language for Multi-agent Decision-Making with Multi-agentic LLMs
by: Huh, Dom, et al.
Published: (2025)
by: Huh, Dom, et al.
Published: (2025)
Realistic pedestrian-driver interaction modelling using multi-agent RL with human perceptual-motor constraints
by: Wang, Yueyang, et al.
Published: (2025)
by: Wang, Yueyang, et al.
Published: (2025)
Sampling-based Pareto Optimization for Chance-constrained Monotone Submodular Problems
by: Yan, Xiankun, et al.
Published: (2024)
by: Yan, Xiankun, et al.
Published: (2024)
Beyond Static Instruction: A Multi-agent AI Framework for Adaptive Augmented Reality Robot Training
by: Leins, Nicolas, et al.
Published: (2026)
by: Leins, Nicolas, et al.
Published: (2026)
ReasonSTL: Bridging Natural Language and Signal Temporal Logic via Tool-Augmented Process-Rewarded Learning
by: Ye, Bowen, et al.
Published: (2026)
by: Ye, Bowen, et al.
Published: (2026)
SAC-GLAM: Improving Online RL for LLM agents with Soft Actor-Critic and Hindsight Relabeling
by: Gaven, Loris, et al.
Published: (2024)
by: Gaven, Loris, et al.
Published: (2024)
AgentRL: Scaling Agentic Reinforcement Learning with a Multi-Turn, Multi-Task Framework
by: Zhang, Hanchen, et al.
Published: (2025)
by: Zhang, Hanchen, et al.
Published: (2025)
Layered LA-MAPF: a decomposition of large agent MAPF instance to accelerate solving without compromising solvability
by: Yao, Zhuo
Published: (2024)
by: Yao, Zhuo
Published: (2024)
Runtime Analysis of Evolutionary Diversity Optimization on the Multi-objective (LeadingOnes, TrailingZeros) Problem
by: Antipov, Denis, et al.
Published: (2024)
by: Antipov, Denis, et al.
Published: (2024)
Synthetic SQL Column Descriptions and Their Impact on Text-to-SQL Performance
by: Wretblad, Niklas, et al.
Published: (2024)
by: Wretblad, Niklas, et al.
Published: (2024)
Single-agent vs. Multi-agents for Automated Video Analysis of On-Screen Collaborative Learning Behaviors
by: Peng, Likai, et al.
Published: (2026)
by: Peng, Likai, et al.
Published: (2026)
Consolidation via Policy Information Regularization in Deep RL for Multi-Agent Games
by: Malloy, Tailia, et al.
Published: (2020)
by: Malloy, Tailia, et al.
Published: (2020)
ZipRL: Adaptive Multi-Turn Context Compression with Hindsight Response Replay
by: Hu, Zhexin, et al.
Published: (2026)
by: Hu, Zhexin, et al.
Published: (2026)
Performing Arts as a Limit to the Digital: Theatrical Space, Embodied Co‑Presence, and the Politics of Ephemeral Action
by: Schismenos, Alexandros
Published: (2026)
by: Schismenos, Alexandros
Published: (2026)
Bi-Objective Evolutionary Optimization for Large-Scale Open Pit Mine Scheduling Problem under Uncertainty with Chance Constraints
by: Pathiranage, Ishara Hewa, et al.
Published: (2025)
by: Pathiranage, Ishara Hewa, et al.
Published: (2025)
On the Use of Evolutionary Optimization for the Dynamic Chance Constrained Open-Pit Mine Scheduling Problem
by: Pathiranage, Ishara Hewa, et al.
Published: (2026)
by: Pathiranage, Ishara Hewa, et al.
Published: (2026)
Multi-agent Multi-armed Bandits with Stochastic Sharable Arm Capacities
by: Xie, Hong, et al.
Published: (2024)
by: Xie, Hong, et al.
Published: (2024)
Design Conductor 2.0: An agent builds a TurboQuant inference accelerator in 80 hours
by: The Verkor Team, et al.
Published: (2026)
by: The Verkor Team, et al.
Published: (2026)
Offline Multi-task Transfer RL with Representational Penalization
by: Bose, Avinandan, et al.
Published: (2024)
by: Bose, Avinandan, et al.
Published: (2024)
Similar Items
-
A Survey on the Integration of Generative AI for Critical Thinking in Mobile Networks
by: Karapantelakis, Athanasios, et al.
Published: (2024) -
Reinforcement Learning for Long-Horizon Unordered Tasks: From Boolean to Coupled Reward Machines
by: Levina, Kristina, et al.
Published: (2025) -
Numeric Reward Machines
by: Levina, Kristina, et al.
Published: (2024) -
Reinforcement Learning with Reward Machines for Sleep Control in Mobile Networks
by: Levina, Kristina, et al.
Published: (2026) -
Scale When Needed: Adaptive Neuron-level Mixed Precision Quantization Aware Training
by: Varshney, Ayush K., et al.
Published: (2026)