Learning Embeddings for Sequential Tasks Using Population of Agents
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mahajan, Mridul, Tzannetos, Georgios, Radanovic, Goran, Singla, Adish |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Proximal Curriculum with Task Correlations for Deep Reinforcement Learning
von: Tzannetos, Georgios, et al.
Veröffentlicht: (2024)
von: Tzannetos, Georgios, et al.
Veröffentlicht: (2024)
Neural Task Synthesis for Visual Programming
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2023)
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2023)
Reward Model Learning vs. Direct Policy Optimization: A Comparative Analysis of Learning from Human Preferences
von: Nika, Andi, et al.
Veröffentlicht: (2024)
von: Nika, Andi, et al.
Veröffentlicht: (2024)
Corruption Robust Offline Reinforcement Learning with Human Feedback
von: Mandal, Debmalya, et al.
Veröffentlicht: (2024)
von: Mandal, Debmalya, et al.
Veröffentlicht: (2024)
Curriculum Design for Trajectory-Constrained Agent: Compressing Chain-of-Thought Tokens in LLMs
von: Tzannetos, Georgios, et al.
Veröffentlicht: (2025)
von: Tzannetos, Georgios, et al.
Veröffentlicht: (2025)
Reward Design for Justifiable Sequential Decision-Making
von: Sukovic, Aleksa, et al.
Veröffentlicht: (2024)
von: Sukovic, Aleksa, et al.
Veröffentlicht: (2024)
Benchmarking the Robustness of Agentic Systems to Adversarially-Induced Harms
von: Nöther, Jonathan, et al.
Veröffentlicht: (2025)
von: Nöther, Jonathan, et al.
Veröffentlicht: (2025)
Text-Diffusion Red-Teaming of Large Language Models: Unveiling Harmful Behaviors with Proximity Constraints
von: Nöther, Jonathan, et al.
Veröffentlicht: (2025)
von: Nöther, Jonathan, et al.
Veröffentlicht: (2025)
MaMa: A Game-Theoretic Approach for Designing Safe Agentic Systems
von: Nöther, Jonathan, et al.
Veröffentlicht: (2026)
von: Nöther, Jonathan, et al.
Veröffentlicht: (2026)
Towards Generalizable Agents in Text-Based Educational Environments: A Study of Integrating RL with LLMs
von: Radmehr, Bahar, et al.
Veröffentlicht: (2024)
von: Radmehr, Bahar, et al.
Veröffentlicht: (2024)
Corruption-robust Offline Multi-agent Reinforcement Learning From Human Feedback
von: Nika, Andi, et al.
Veröffentlicht: (2026)
von: Nika, Andi, et al.
Veröffentlicht: (2026)
Corruption-Robust Offline Two-Player Zero-Sum Markov Games
von: Nika, Andi, et al.
Veröffentlicht: (2024)
von: Nika, Andi, et al.
Veröffentlicht: (2024)
Hints-In-Browser: Benchmarking Language Models for Programming Feedback Generation
von: Kotalwar, Nachiket, et al.
Veröffentlicht: (2024)
von: Kotalwar, Nachiket, et al.
Veröffentlicht: (2024)
Distributionally Robust Reinforcement Learning with Human Feedback
von: Mandal, Debmalya, et al.
Veröffentlicht: (2025)
von: Mandal, Debmalya, et al.
Veröffentlicht: (2025)
PolyNet: Learning Diverse Solution Strategies for Neural Combinatorial Optimization
von: Hottung, André, et al.
Veröffentlicht: (2024)
von: Hottung, André, et al.
Veröffentlicht: (2024)
Policy Teaching via Data Poisoning in Learning from Human Preferences
von: Nika, Andi, et al.
Veröffentlicht: (2025)
von: Nika, Andi, et al.
Veröffentlicht: (2025)
Reinforcement Learning for Durable Algorithmic Recourse
von: Ceccon, Marina, et al.
Veröffentlicht: (2025)
von: Ceccon, Marina, et al.
Veröffentlicht: (2025)
Prompt Optimization Across Multiple Agents for Representing Diverse Human Populations
von: Nguyen, Manh Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Manh Hung, et al.
Veröffentlicht: (2025)
Counterfactual Effect Decomposition in Multi-Agent Sequential Decision Making
von: Triantafyllou, Stelios, et al.
Veröffentlicht: (2024)
von: Triantafyllou, Stelios, et al.
Veröffentlicht: (2024)
Representation Learning for Sequential Volumetric Design Tasks
von: Alam, Md Ferdous, et al.
Veröffentlicht: (2023)
von: Alam, Md Ferdous, et al.
Veröffentlicht: (2023)
Adaptive Opponent Policy Detection in Multi-Agent MDPs: Real-Time Strategy Switch Identification Using Running Error Estimation
von: Mridul, Mohidul Haque, et al.
Veröffentlicht: (2024)
von: Mridul, Mohidul Haque, et al.
Veröffentlicht: (2024)
Pathology-Aware Multi-View Contrastive Learning for Patient-Independent ECG Reconstruction
von: Youssef, Youssef, et al.
Veröffentlicht: (2026)
von: Youssef, Youssef, et al.
Veröffentlicht: (2026)
Learning Hidden Markov Models Using Conditional Samples
von: Kakade, Sham M., et al.
Veröffentlicht: (2023)
von: Kakade, Sham M., et al.
Veröffentlicht: (2023)
Exposing Limitations of Language Model Agents in Sequential-Task Compositions on the Web
von: Furuta, Hiroki, et al.
Veröffentlicht: (2023)
von: Furuta, Hiroki, et al.
Veröffentlicht: (2023)
Joint Optimization of Multi-Objective Reinforcement Learning with Policy Gradient Based Algorithm
von: Bai, Qinbo, et al.
Veröffentlicht: (2021)
von: Bai, Qinbo, et al.
Veröffentlicht: (2021)
Translating the Rashomon Effect to Sequential Decision-Making Tasks
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
von: Gross, Dennis, et al.
Veröffentlicht: (2025)
Sequential Stochastic Combinatorial Optimization Using Hierarchal Reinforcement Learning
von: Feng, Xinsong, et al.
Veröffentlicht: (2025)
von: Feng, Xinsong, et al.
Veröffentlicht: (2025)
Beyond Loss Guidance: Using PDE Residuals as Spectral Attention in Diffusion Neural Operators
von: Sawhney, Medha, et al.
Veröffentlicht: (2025)
von: Sawhney, Medha, et al.
Veröffentlicht: (2025)
Offline Multi-Agent Reinforcement Learning via In-Sample Sequential Policy Optimization
von: Liu, Zongkai, et al.
Veröffentlicht: (2024)
von: Liu, Zongkai, et al.
Veröffentlicht: (2024)
Agent-Specific Effects: A Causal Effect Propagation Analysis in Multi-Agent MDPs
von: Triantafyllou, Stelios, et al.
Veröffentlicht: (2023)
von: Triantafyllou, Stelios, et al.
Veröffentlicht: (2023)
Divergent-Convergent Thinking in Large Language Models for Creative Problem Generation
von: Nguyen, Manh Hung, et al.
Veröffentlicht: (2025)
von: Nguyen, Manh Hung, et al.
Veröffentlicht: (2025)
Benchmarking Generative Models on Computational Thinking Tests in Elementary Visual Programming
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2024)
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2024)
When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks?
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2026)
von: Hatgis-Kessell, Stephane, et al.
Veröffentlicht: (2026)
Performative Reinforcement Learning with Linear Markov Decision Process
von: Mandal, Debmalya, et al.
Veröffentlicht: (2024)
von: Mandal, Debmalya, et al.
Veröffentlicht: (2024)
Mirage: Model-Agnostic Graph Distillation for Graph Classification
von: Gupta, Mridul, et al.
Veröffentlicht: (2023)
von: Gupta, Mridul, et al.
Veröffentlicht: (2023)
Informativeness of Reward Functions in Reinforcement Learning
von: Devidze, Rati, et al.
Veröffentlicht: (2024)
von: Devidze, Rati, et al.
Veröffentlicht: (2024)
Iterative Amortized Inference: Unifying In-Context Learning and Learned Optimizers
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
von: Mittal, Sarthak, et al.
Veröffentlicht: (2025)
Task2vec Readiness: Diagnostics for Federated Learning from Pre-Training Embeddings
von: Mafuz, Cristiano, et al.
Veröffentlicht: (2026)
von: Mafuz, Cristiano, et al.
Veröffentlicht: (2026)
Integrating Sequential and Relational Modeling for User Events: Datasets and Prediction Tasks
von: Fathony, Rizal, et al.
Veröffentlicht: (2025)
von: Fathony, Rizal, et al.
Veröffentlicht: (2025)
On Corruption-Robustness in Performative Reinforcement Learning
von: Pollatos, Vasilis, et al.
Veröffentlicht: (2025)
von: Pollatos, Vasilis, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Proximal Curriculum with Task Correlations for Deep Reinforcement Learning
von: Tzannetos, Georgios, et al.
Veröffentlicht: (2024) -
Neural Task Synthesis for Visual Programming
von: Pădurean, Victor-Alexandru, et al.
Veröffentlicht: (2023) -
Reward Model Learning vs. Direct Policy Optimization: A Comparative Analysis of Learning from Human Preferences
von: Nika, Andi, et al.
Veröffentlicht: (2024) -
Corruption Robust Offline Reinforcement Learning with Human Feedback
von: Mandal, Debmalya, et al.
Veröffentlicht: (2024) -
Curriculum Design for Trajectory-Constrained Agent: Compressing Chain-of-Thought Tokens in LLMs
von: Tzannetos, Georgios, et al.
Veröffentlicht: (2025)