AgentCollabBench: Diagnosing When Good Agents Make Bad Collaborators
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Mazumder, Aritra, Dipta, Shubhashis Roy, Lia, Nusrat Jahan, Khan, Tanzila, Hossain, Kainat Raisa, Shri, Nehaa, Debsarkar, Shubhrangshu, Tasnim, Humayra, Shawon, Gour Gupal Talukder, Mitra, Debjoty, Rani, Sumaiya Ahmed, Anik, Al Jami Islam, Khan, Al Nafeu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Machine Learning Derived Blood Input for Dynamic PET Images of Rat Heart
von: Debsarkar, Shubhrangshu, et al.
Veröffentlicht: (2025)
von: Debsarkar, Shubhrangshu, et al.
Veröffentlicht: (2025)
Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability
von: Lia, Nusrat Jahan, et al.
Veröffentlicht: (2026)
von: Lia, Nusrat Jahan, et al.
Veröffentlicht: (2026)
TRIAGE: Evaluating Prospective Metacognitive Control in LLMs under Resource Constraints
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026)
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026)
Read Between the Lines: A Benchmark for Uncovering Political Bias in Bangla News Articles
von: Lia, Nusrat Jahan, et al.
Veröffentlicht: (2025)
von: Lia, Nusrat Jahan, et al.
Veröffentlicht: (2025)
†DAGGER: Distractor-Aware Graph Generation for Executable Reasoning in Math Problems
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026)
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026)
BanglaLlama: LLaMA for Bangla Language
von: Zehady, Abdullah Khan, et al.
Veröffentlicht: (2024)
von: Zehady, Abdullah Khan, et al.
Veröffentlicht: (2024)
Omni-Modal Dissonance Benchmark: Systematically Breaking Modality Consensus to Probe Robustness and Calibrated Abstention
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026)
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026)
FedMentor: Domain-Aware Differential Privacy for Heterogeneous Federated LLMs in Mental Health
von: Sarwar, Nobin, et al.
Veröffentlicht: (2025)
von: Sarwar, Nobin, et al.
Veröffentlicht: (2025)
If We May De-Presuppose: Robustly Verifying Claims through Presupposition-Free Question Decomposition
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
UMBCLU at SemEval-2024 Task 1A and 1C: Semantic Textual Relatedness with and without machine translation
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2024)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2024)
HU at SemEval-2024 Task 8A: Can Contrastive Learning Learn Embeddings to Detect Machine-Generated Text?
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2024)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2024)
Q2E: Query-to-Event Decomposition for Zero-Shot Multilingual Text-to-Video Retrieval
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
BanglaTalk: Towards Real-Time Speech Assistance for Bengali Regional Dialects
von: Hasan, Jakir, et al.
Veröffentlicht: (2025)
von: Hasan, Jakir, et al.
Veröffentlicht: (2025)
PromptGuard at BLP-2025 Task 1: A Few-Shot Classification Framework Using Majority Voting and Keyword Similarity for Bengali Hate Speech Detection
von: Hossan, Rakib, et al.
Veröffentlicht: (2025)
von: Hossan, Rakib, et al.
Veröffentlicht: (2025)
DecomposeRL: Learning to Ask Useful, Informative, and Diverse Questions for Semi-Supervised, Traceable Claim Verification
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
GanitLLM: Difficulty-Aware Bengali Mathematical Reasoning through Curriculum-GRPO
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
VC-Inspector: Advancing Reference-free Evaluation of Video Captions with Factual Analysis
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2025)
PA3: Policy-Aware Agent Alignment through Chain-of-Thought
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
von: Dipta, Shubhashis Roy, et al.
Veröffentlicht: (2026)
Effect of patient‐centered self‐management intervention on glycemic control, self‐efficacy, and self‐care behaviors in South Asian adults with type 2 diabetes mellitus: A multicenter randomized controlled trial
von: Kainat Asmat, et al.
Veröffentlicht: (2024)
von: Kainat Asmat, et al.
Veröffentlicht: (2024)
Learning How to Use Tools, Not Just When: Pattern-Aware Tool-Integrated Reasoning
von: Xu, Ningning, et al.
Veröffentlicht: (2025)
von: Xu, Ningning, et al.
Veröffentlicht: (2025)
Deficiencies of women's participation in climate governance and sustainable development challenges in Bangladesh
von: Jeba Humayra, et al.
Veröffentlicht: (2024)
von: Jeba Humayra, et al.
Veröffentlicht: (2024)
Empirical Analysis of Nature-Inspired Algorithms for Autism Spectrum Disorder Detection Using 3D Video Dataset
von: Panchal, Aneesh, et al.
Veröffentlicht: (2025)
von: Panchal, Aneesh, et al.
Veröffentlicht: (2025)
Collab: Controlled Decoding using Mixture of Agents for LLM Alignment
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2025)
von: Chakraborty, Souradip, et al.
Veröffentlicht: (2025)
AgentCollab: A Self-Evaluation-Driven Collaboration Paradigm for Efficient LLM Agents
von: Gao, Wenbo, et al.
Veröffentlicht: (2026)
von: Gao, Wenbo, et al.
Veröffentlicht: (2026)
The Efficacy of Clove Seed Aqueous Extract as an Anesthetic Agent on the Stress Response of Mystus cavasius During Transportation
von: Nafisa Khatun, et al.
Veröffentlicht: (2026)
von: Nafisa Khatun, et al.
Veröffentlicht: (2026)
Los h litos de la intimidad / Abd Ar-Rahman Al Jami; Traducción, prólogo y notas, Jordi Quingles
von: Jami, Abd Ar-Rahman Al
von: Jami, Abd Ar-Rahman Al
Quantum Energy Teleportation across Multi-Qubit Systems using W-State Entanglement
von: Khan, Alif Elham, et al.
Veröffentlicht: (2025)
von: Khan, Alif Elham, et al.
Veröffentlicht: (2025)
ACC-Collab: An Actor-Critic Approach to Multi-Agent LLM Collaboration
von: Estornell, Andrew, et al.
Veröffentlicht: (2024)
von: Estornell, Andrew, et al.
Veröffentlicht: (2024)
Collab-Overcooked: Benchmarking and Evaluating Large Language Models as Collaborative Agents
von: Sun, Haochen, et al.
Veröffentlicht: (2025)
von: Sun, Haochen, et al.
Veröffentlicht: (2025)
A PROPOSED TRAINING MODEL THAT ADDRESSES THE NEEDS OF PARENTS OF CHILDREN WITH AUTISM SPECTRUM DISORDER IN MUSCAT GOVERNORATE, SULTANATE OF OMAN
von: Al_Aghbari, Sumaiya Yousuf Saif
Veröffentlicht: (2025)
von: Al_Aghbari, Sumaiya Yousuf Saif
Veröffentlicht: (2025)
Cornerstones or Stumbling Blocks? Deciphering the Rock Tokens in On-Policy Distillation
von: Jiang, Yuxuan, et al.
Veröffentlicht: (2026)
von: Jiang, Yuxuan, et al.
Veröffentlicht: (2026)
CollabEval: Enhancing LLM-as-a-Judge via Multi-Agent Collaboration
von: Qian, Yiyue, et al.
Veröffentlicht: (2026)
von: Qian, Yiyue, et al.
Veröffentlicht: (2026)
ChatCollab: Exploring Collaboration Between Humans and AI Agents in Software Teams
von: Klieger, Benjamin, et al.
Veröffentlicht: (2024)
von: Klieger, Benjamin, et al.
Veröffentlicht: (2024)
BanglaIPA: Towards Robust Text-to-IPA Transcription with Contextual Rewriting in Bengali
von: Hasan, Jakir, et al.
Veröffentlicht: (2026)
von: Hasan, Jakir, et al.
Veröffentlicht: (2026)
Three-dimensional imaging of biological cells using surface plasmon coupled emission
von: Mazumder, Anik, et al.
Veröffentlicht: (2025)
von: Mazumder, Anik, et al.
Veröffentlicht: (2025)
EVALUATION OF NURSING INTERVENTIONS TO PREVENT PRESSURE ULCERS IN CHRONICALLY ILL PATIENTS IN A TERTIARY HEALTHCARE SETTING IN DISTRICT MARDAN
von: Haya Khan, Rida Parvez, Iqra Bibi, Yusra Shanzay Khan, Tanzila Nawaz, Muhammad Sulaiman
Veröffentlicht: (2026)
von: Haya Khan, Rida Parvez, Iqra Bibi, Yusra Shanzay Khan, Tanzila Nawaz, Muhammad Sulaiman
Veröffentlicht: (2026)
Shadow loss: Memory-linear deep metric learning for efficient training
von: Khan, Alif Elham, et al.
Veröffentlicht: (2023)
von: Khan, Alif Elham, et al.
Veröffentlicht: (2023)
Cuffless Blood Pressure Prediction from Speech Sentences using Deep Learning Methods
von: Kainat
Veröffentlicht: (2025)
von: Kainat
Veröffentlicht: (2025)
Operational-umbral approach to bivariate degenerate Hermite polynomials and their partial orthogonality
von: Raza, Nusrat, et al.
Veröffentlicht: (2025)
von: Raza, Nusrat, et al.
Veröffentlicht: (2025)
Comprehensive DFT Study on the Structural, Electronic, Optical, Mechanical, and Thermodynamic Behavior of Lead‐Free AZnX 3 (A = Al, Ag; X = Cl, Br) Perovskites for Optoelectronic Applications
von: Mushfique Azad Takin, et al.
Veröffentlicht: (2025)
von: Mushfique Azad Takin, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Machine Learning Derived Blood Input for Dynamic PET Images of Rat Heart
von: Debsarkar, Shubhrangshu, et al.
Veröffentlicht: (2025) -
Cross-Lingual Sentiment Misalignment: Auditing Multilingual Language Models for Inversion Risk, Dialectal Representation, and Affective Stability
von: Lia, Nusrat Jahan, et al.
Veröffentlicht: (2026) -
TRIAGE: Evaluating Prospective Metacognitive Control in LLMs under Resource Constraints
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026) -
Read Between the Lines: A Benchmark for Uncovering Political Bias in Bangla News Articles
von: Lia, Nusrat Jahan, et al.
Veröffentlicht: (2025) -
†DAGGER: Distractor-Aware Graph Generation for Executable Reasoning in Math Problems
von: Nazi, Zabir Al, et al.
Veröffentlicht: (2026)