pAI/MSc: ML Theory Research with Humans on the Loop
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Abdelmoneum, Mahmoud, Beneventano, Pierfrancesco, Poggio, Tomaso |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Incorporating Human Flexibility through Reward Preferences in Human-AI Teaming
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2023)
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2023)
Moving Out: Physically-grounded Human-AI Collaboration
von: Kang, Xuhui, et al.
Veröffentlicht: (2025)
von: Kang, Xuhui, et al.
Veröffentlicht: (2025)
Too Sharp, Too Sure: When Calibration Follows Curvature
von: Morosini, Alessandro, et al.
Veröffentlicht: (2026)
von: Morosini, Alessandro, et al.
Veröffentlicht: (2026)
LoopBench: Discovering Emergent Symmetry Breaking Strategies with LLM Swarms
von: Parsaee, Ali, et al.
Veröffentlicht: (2025)
von: Parsaee, Ali, et al.
Veröffentlicht: (2025)
PaperOrchestra: A Multi-Agent Framework for Automated AI Research Paper Writing
von: Song, Yiwen, et al.
Veröffentlicht: (2026)
von: Song, Yiwen, et al.
Veröffentlicht: (2026)
Disaster Management in the Era of Agentic AI Systems: A Vision for Collective Human-Machine Intelligence for Augmented Resilience
von: Li, Bo, et al.
Veröffentlicht: (2025)
von: Li, Bo, et al.
Veröffentlicht: (2025)
Machine Theory of Mind for Autonomous Cyber-Defence
von: Swaby, Luke, et al.
Veröffentlicht: (2024)
von: Swaby, Luke, et al.
Veröffentlicht: (2024)
AutoML-Agent: A Multi-Agent LLM Framework for Full-Pipeline AutoML
von: Trirat, Patara, et al.
Veröffentlicht: (2024)
von: Trirat, Patara, et al.
Veröffentlicht: (2024)
Probably Approximately Consensus: On the Learning Theory of Finding Common Ground
von: Blair, Carter, et al.
Veröffentlicht: (2026)
von: Blair, Carter, et al.
Veröffentlicht: (2026)
Learning to Cooperate with Humans using Generative Agents
von: Liang, Yancheng, et al.
Veröffentlicht: (2024)
von: Liang, Yancheng, et al.
Veröffentlicht: (2024)
ToMCAT: Theory-of-Mind for Cooperative Agents in Teams via Multiagent Diffusion Policies
von: Sequeira, Pedro, et al.
Veröffentlicht: (2025)
von: Sequeira, Pedro, et al.
Veröffentlicht: (2025)
How Neural Networks Learn the Support is an Implicit Regularization Effect of SGD
von: Beneventano, Pierfrancesco, et al.
Veröffentlicht: (2024)
von: Beneventano, Pierfrancesco, et al.
Veröffentlicht: (2024)
Cognify: Supercharging Gen-AI Workflows With Hierarchical Autotuning
von: He, Zijian, et al.
Veröffentlicht: (2025)
von: He, Zijian, et al.
Veröffentlicht: (2025)
Aligning Compound AI Systems via System-level DPO
von: Wang, Xiangwen, et al.
Veröffentlicht: (2025)
von: Wang, Xiangwen, et al.
Veröffentlicht: (2025)
Hybrid Agentic AI and Multi-Agent Systems in Smart Manufacturing
von: Farahani, Mojtaba A., et al.
Veröffentlicht: (2025)
von: Farahani, Mojtaba A., et al.
Veröffentlicht: (2025)
Bounded Coupled AI Learning Dynamics in Tri-Hierarchical Drone Swarms
von: Bychkov, Oleksii
Veröffentlicht: (2026)
von: Bychkov, Oleksii
Veröffentlicht: (2026)
A Comprehensive Review of AI Agents: Transforming Possibilities in Technology and Beyond
von: Qu, Xiaodong, et al.
Veröffentlicht: (2025)
von: Qu, Xiaodong, et al.
Veröffentlicht: (2025)
Generative Multi-Agent Collaboration in Embodied AI: A Systematic Review
von: Wu, Di, et al.
Veröffentlicht: (2025)
von: Wu, Di, et al.
Veröffentlicht: (2025)
Generative AI Against Poaching: Latent Composite Flow Matching for Wildlife Conservation
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
von: Kong, Lingkai, et al.
Veröffentlicht: (2025)
Learning Individual Intrinsic Reward in Multi-Agent Reinforcement Learning via Incorporating Generalized Human Expertise
von: Wu, Xuefei, et al.
Veröffentlicht: (2025)
von: Wu, Xuefei, et al.
Veröffentlicht: (2025)
Towards Understanding, Analyzing, and Optimizing Agentic AI Execution: A CPU-Centric Perspective
von: Raj, Ritik, et al.
Veröffentlicht: (2025)
von: Raj, Ritik, et al.
Veröffentlicht: (2025)
PLAICraft: Large-Scale Time-Aligned Vision-Speech-Action Dataset for Embodied AI
von: He, Yingchen, et al.
Veröffentlicht: (2025)
von: He, Yingchen, et al.
Veröffentlicht: (2025)
M3HF: Multi-agent Reinforcement Learning from Multi-phase Human Feedback of Mixed Quality
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
von: Wang, Ziyan, et al.
Veröffentlicht: (2025)
Talk Freely, Execute Strictly: Schema-Gated Agentic AI for Flexible and Reproducible Scientific Workflows
von: Strickland, Joel, et al.
Veröffentlicht: (2026)
von: Strickland, Joel, et al.
Veröffentlicht: (2026)
EngiAI: A Multi-Agent Framework and Benchmark Suite for LLM-Driven Engineering Design
von: Molinari, Gioele, et al.
Veröffentlicht: (2026)
von: Molinari, Gioele, et al.
Veröffentlicht: (2026)
EdgeAgentX: A Novel Framework for Agentic AI at the Edge in Military Communication Networks
von: Ray, Abir
Veröffentlicht: (2025)
von: Ray, Abir
Veröffentlicht: (2025)
A Hierarchical Hybrid AI Approach: Integrating Deep Reinforcement Learning and Scripted Agents in Combat Simulations
von: Black, Scotty, et al.
Veröffentlicht: (2025)
von: Black, Scotty, et al.
Veröffentlicht: (2025)
HDDLGym: A Tool for Studying Multi-Agent Hierarchical Problems Defined in HDDL with OpenAI Gym
von: La, Ngoc, et al.
Veröffentlicht: (2025)
von: La, Ngoc, et al.
Veröffentlicht: (2025)
Proactive Agent Research Environment: Simulating Active Users to Evaluate Proactive Assistants
von: Nathani, Deepak, et al.
Veröffentlicht: (2026)
von: Nathani, Deepak, et al.
Veröffentlicht: (2026)
Can Vibe Coding Beat Graduate CS Students? An LLM vs. Human Coding Tournament on Market-driven Strategic Planning
von: Danassis, Panayiotis, et al.
Veröffentlicht: (2025)
von: Danassis, Panayiotis, et al.
Veröffentlicht: (2025)
Analyzing Operator States and the Impact of AI-Enhanced Decision Support in Control Rooms: A Human-in-the-Loop Specialized Reinforcement Learning Framework for Intervention Strategies
von: Abbas, Ammar N., et al.
Veröffentlicht: (2024)
von: Abbas, Ammar N., et al.
Veröffentlicht: (2024)
ThinkTank: A Framework for Generalizing Domain-Specific AI Agent Systems into Universal Collaborative Intelligence Platforms
von: Surabhi, Praneet Sai Madhu, et al.
Veröffentlicht: (2025)
von: Surabhi, Praneet Sai Madhu, et al.
Veröffentlicht: (2025)
AIvril: AI-Driven RTL Generation With Verification In-The-Loop
von: Islam, Mubashir ul, et al.
Veröffentlicht: (2024)
von: Islam, Mubashir ul, et al.
Veröffentlicht: (2024)
Topological Structure Learning Should Be A Research Priority for LLM-Based Multi-Agent Systems
von: Yang, Jiaxi, et al.
Veröffentlicht: (2025)
von: Yang, Jiaxi, et al.
Veröffentlicht: (2025)
Momentum Further Constrains Sharpness at the Edge of Stochastic Stability
von: Andreyev, Arseniy, et al.
Veröffentlicht: (2026)
von: Andreyev, Arseniy, et al.
Veröffentlicht: (2026)
A Data-Driven Discretized CS:GO Simulation Environment to Facilitate Strategic Multi-Agent Planning Research
von: Wang, Yunzhe, et al.
Veröffentlicht: (2025)
von: Wang, Yunzhe, et al.
Veröffentlicht: (2025)
Towards Reliable ML Feature Engineering via Planning in Constrained-Topology of LLM Agents
von: Thakur, Himanshu, et al.
Veröffentlicht: (2026)
von: Thakur, Himanshu, et al.
Veröffentlicht: (2026)
Reliability and Effectiveness of Autonomous AI Agents in Supply Chain Management
von: Long, Carol Xuan, et al.
Veröffentlicht: (2026)
von: Long, Carol Xuan, et al.
Veröffentlicht: (2026)
HealthFlow: A Self-Evolving AI Agent with Meta Planning for Autonomous Healthcare Research
von: Zhu, Yinghao, et al.
Veröffentlicht: (2025)
von: Zhu, Yinghao, et al.
Veröffentlicht: (2025)
InteRACT: Transformer Models for Human Intent Prediction Conditioned on Robot Actions
von: Kedia, Kushal, et al.
Veröffentlicht: (2023)
von: Kedia, Kushal, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Incorporating Human Flexibility through Reward Preferences in Human-AI Teaming
von: Bhambri, Siddhant, et al.
Veröffentlicht: (2023) -
Moving Out: Physically-grounded Human-AI Collaboration
von: Kang, Xuhui, et al.
Veröffentlicht: (2025) -
Too Sharp, Too Sure: When Calibration Follows Curvature
von: Morosini, Alessandro, et al.
Veröffentlicht: (2026) -
LoopBench: Discovering Emergent Symmetry Breaking Strategies with LLM Swarms
von: Parsaee, Ali, et al.
Veröffentlicht: (2025) -
PaperOrchestra: A Multi-Agent Framework for Automated AI Research Paper Writing
von: Song, Yiwen, et al.
Veröffentlicht: (2026)