Allen: Rethinking MAS Design through Step-Level Policy Autonomy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhou, Qiangong, Wang, Zhiting, Yao, Mingyou, Liu, Zongyang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Det-SAM2:Technical Report on the Self-Prompting Segmentation Framework Based on Segment Anything Model 2
von: Wang, Zhiting, et al.
Veröffentlicht: (2024)
von: Wang, Zhiting, et al.
Veröffentlicht: (2024)
VideoChat-M1: Collaborative Policy Planning for Video Understanding via Multi-Agent Reinforcement Learning
von: Chen, Boyu, et al.
Veröffentlicht: (2025)
von: Chen, Boyu, et al.
Veröffentlicht: (2025)
HiLight: Technical Report on the Motern AI Video Language Model
von: Wang, Zhiting, et al.
Veröffentlicht: (2024)
von: Wang, Zhiting, et al.
Veröffentlicht: (2024)
AgentCVR: Active Multi-Agent Cross-Video Reasoning via Script-Simulated Reinforcement Learning
von: Qiu, Yilun, et al.
Veröffentlicht: (2026)
von: Qiu, Yilun, et al.
Veröffentlicht: (2026)
PreGSU-A Generalized Traffic Scene Understanding Model for Autonomous Driving based on Pre-trained Graph Attention Network
von: Wang, Yuning, et al.
Veröffentlicht: (2024)
von: Wang, Yuning, et al.
Veröffentlicht: (2024)
Sentinel: Embodied Cooperative Spatial Reasoning and Planning
von: Lin, Xiangye, et al.
Veröffentlicht: (2026)
von: Lin, Xiangye, et al.
Veröffentlicht: (2026)
AdaptFly: Prompt-Guided Adaptation of Foundation Models for Low-Altitude UAV Networks
von: Chen, Jiao, et al.
Veröffentlicht: (2025)
von: Chen, Jiao, et al.
Veröffentlicht: (2025)
What Makes Good Collaborative Views? Contrastive Mutual Information Maximization for Multi-Agent Perception
von: Su, Wanfang, et al.
Veröffentlicht: (2024)
von: Su, Wanfang, et al.
Veröffentlicht: (2024)
V2X-DGPE: Addressing Domain Gaps and Pose Errors for Robust Collaborative 3D Object Detection
von: Wang, Sichao, et al.
Veröffentlicht: (2025)
von: Wang, Sichao, et al.
Veröffentlicht: (2025)
A Multi-Agent Perception-Action Alliance for Efficient Long Video Reasoning
von: Xu, Yichang, et al.
Veröffentlicht: (2026)
von: Xu, Yichang, et al.
Veröffentlicht: (2026)
End-to-End Autonomous Driving through V2X Cooperation
von: Yu, Haibao, et al.
Veröffentlicht: (2024)
von: Yu, Haibao, et al.
Veröffentlicht: (2024)
MAG-3D: Multi-Agent Grounded Reasoning for 3D Understanding
von: Zheng, Henry, et al.
Veröffentlicht: (2026)
von: Zheng, Henry, et al.
Veröffentlicht: (2026)
AstroVLM: Expert Multi-agent Collaborative Reasoning for Astronomical Imaging Quality Diagnosis
von: Han, Yaohui, et al.
Veröffentlicht: (2026)
von: Han, Yaohui, et al.
Veröffentlicht: (2026)
Enhancing CLIP Robustness via Cross-Modality Alignment
von: Zhu, Xingyu, et al.
Veröffentlicht: (2025)
von: Zhu, Xingyu, et al.
Veröffentlicht: (2025)
AniMaker: Multi-Agent Animated Storytelling with MCTS-Driven Clip Generation
von: Shi, Haoyuan, et al.
Veröffentlicht: (2025)
von: Shi, Haoyuan, et al.
Veröffentlicht: (2025)
Multi-Agent Amodal Completion: Direct Synthesis with Fine-Grained Semantic Guidance
von: Fan, Hongxing, et al.
Veröffentlicht: (2025)
von: Fan, Hongxing, et al.
Veröffentlicht: (2025)
Unified End-to-End V2X Cooperative Autonomous Driving
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
von: Li, Zhiwei, et al.
Veröffentlicht: (2024)
HiMemFormer: Hierarchical Memory-Aware Transformer for Multi-Agent Action Anticipation
von: Wang, Zirui, et al.
Veröffentlicht: (2024)
von: Wang, Zirui, et al.
Veröffentlicht: (2024)
CAMON: Cooperative Agents for Multi-Object Navigation with LLM-based Conversations
von: Wu, Pengying, et al.
Veröffentlicht: (2024)
von: Wu, Pengying, et al.
Veröffentlicht: (2024)
FetalAgents: A Multi-Agent System for Fetal Ultrasound Image and Video Analysis
von: Hu, Xiaotian, et al.
Veröffentlicht: (2026)
von: Hu, Xiaotian, et al.
Veröffentlicht: (2026)
Towards Reliable Fetal Ultrasound Interpretation with Multi-Agent Collaboration
von: Hu, Xiaotian, et al.
Veröffentlicht: (2026)
von: Hu, Xiaotian, et al.
Veröffentlicht: (2026)
Visual Multi-Agent System: Mitigating Hallucination Snowballing via Visual Flow
von: Yu, Xinlei, et al.
Veröffentlicht: (2025)
von: Yu, Xinlei, et al.
Veröffentlicht: (2025)
Fast2comm:Collaborative perception combined with prior knowledge
von: Zhang, Zhengbin, et al.
Veröffentlicht: (2025)
von: Zhang, Zhengbin, et al.
Veröffentlicht: (2025)
Hollywood Town: Long-Video Generation via Cross-Modal Multi-Agent Orchestration
von: Wei, Zheng, et al.
Veröffentlicht: (2025)
von: Wei, Zheng, et al.
Veröffentlicht: (2025)
Med-GRIM: Enhanced Zero-Shot Medical VQA using prompt-embedded Multimodal Graph RAG
von: Madavan, Rakesh Raj, et al.
Veröffentlicht: (2025)
von: Madavan, Rakesh Raj, et al.
Veröffentlicht: (2025)
A Case Study of Counting the Number of Unique Users in Linear and Non-Linear Trails -- A Multi-Agent System Approach
von: Rahman, Tanvir
Veröffentlicht: (2025)
von: Rahman, Tanvir
Veröffentlicht: (2025)
TraF-Align: Trajectory-aware Feature Alignment for Asynchronous Multi-agent Perception
von: Song, Zhiying, et al.
Veröffentlicht: (2025)
von: Song, Zhiying, et al.
Veröffentlicht: (2025)
Chain-of-Anomaly Thoughts with Large Vision-Language Models
von: Domingos, Pedro, et al.
Veröffentlicht: (2025)
von: Domingos, Pedro, et al.
Veröffentlicht: (2025)
VideoMultiAgents: A Multi-Agent Framework for Video Question Answering
von: Kugo, Noriyuki, et al.
Veröffentlicht: (2025)
von: Kugo, Noriyuki, et al.
Veröffentlicht: (2025)
Active Scout: Multi-Target Tracking Using Neural Radiance Fields in Dense Urban Environments
von: Hsu, Christopher D., et al.
Veröffentlicht: (2024)
von: Hsu, Christopher D., et al.
Veröffentlicht: (2024)
LogiStory: A Logic-Aware Framework for Multi-Image Story Visualization
von: Meng, Chutian, et al.
Veröffentlicht: (2026)
von: Meng, Chutian, et al.
Veröffentlicht: (2026)
Visual Sensor Pose Optimisation Using Visibility Models for Smart Cities
von: Arnold, Eduardo, et al.
Veröffentlicht: (2021)
von: Arnold, Eduardo, et al.
Veröffentlicht: (2021)
Cascading multi-agent anomaly detection in surveillance systems via vision-language models and embedding-based classification
von: Rehman, Tayyab, et al.
Veröffentlicht: (2026)
von: Rehman, Tayyab, et al.
Veröffentlicht: (2026)
FootBots: A Transformer-based Architecture for Motion Prediction in Soccer
von: Capellera, Guillem, et al.
Veröffentlicht: (2024)
von: Capellera, Guillem, et al.
Veröffentlicht: (2024)
ProCrit: Self-Elicited Multi-Perspective Reasoning with Critic-Guided Revision for Multimodal Sarcasm Detection
von: Xu, Yingjia, et al.
Veröffentlicht: (2026)
von: Xu, Yingjia, et al.
Veröffentlicht: (2026)
Learning Collective Dynamics of Multi-Agent Systems using Event-based Vision
von: Lee, Minah, et al.
Veröffentlicht: (2024)
von: Lee, Minah, et al.
Veröffentlicht: (2024)
TranSPORTmer: A Holistic Approach to Trajectory Understanding in Multi-Agent Sports
von: Capellera, Guillem, et al.
Veröffentlicht: (2024)
von: Capellera, Guillem, et al.
Veröffentlicht: (2024)
ReCCur: A Recursive Corner-Case Curation Framework for Robust Vision-Language Understanding in Open and Edge Scenarios
von: Wei, Yihan, et al.
Veröffentlicht: (2026)
von: Wei, Yihan, et al.
Veröffentlicht: (2026)
A Stochastic Geo-spatiotemporal Bipartite Network to Optimize GCOOS Sensor Placement Strategies
von: Holmberg, Ted Edward, et al.
Veröffentlicht: (2024)
von: Holmberg, Ted Edward, et al.
Veröffentlicht: (2024)
SPAgent: Adaptive Task Decomposition and Model Selection for General Video Generation and Editing
von: Tu, Rong-Cheng, et al.
Veröffentlicht: (2024)
von: Tu, Rong-Cheng, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Det-SAM2:Technical Report on the Self-Prompting Segmentation Framework Based on Segment Anything Model 2
von: Wang, Zhiting, et al.
Veröffentlicht: (2024) -
VideoChat-M1: Collaborative Policy Planning for Video Understanding via Multi-Agent Reinforcement Learning
von: Chen, Boyu, et al.
Veröffentlicht: (2025) -
HiLight: Technical Report on the Motern AI Video Language Model
von: Wang, Zhiting, et al.
Veröffentlicht: (2024) -
AgentCVR: Active Multi-Agent Cross-Video Reasoning via Script-Simulated Reinforcement Learning
von: Qiu, Yilun, et al.
Veröffentlicht: (2026) -
PreGSU-A Generalized Traffic Scene Understanding Model for Autonomous Driving based on Pre-trained Graph Attention Network
von: Wang, Yuning, et al.
Veröffentlicht: (2024)