MegaChat: A Synthetic Persian Q&A Dataset for High-Quality Sales Chatbot Evaluation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Rahmani, Mahdi, Saffari, AmirHossein, Rahmani, Reyhane |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Information-Driven Fault Detection and Identification for Multi-Agent Spacecraft Systems: Collaborative On-Orbit Inspection Mission
von: Gupta, Akshita, et al.
Veröffentlicht: (2025)
von: Gupta, Akshita, et al.
Veröffentlicht: (2025)
MiRAGE: A Multiagent Framework for Generating Multimodal Multihop Question-Answer Dataset for RAG Evaluation
von: Sahu, Chandan Kumar, et al.
Veröffentlicht: (2026)
von: Sahu, Chandan Kumar, et al.
Veröffentlicht: (2026)
ReasonMed: A 370K Multi-Agent Generated Dataset for Advancing Medical Reasoning
von: Sun, Yu, et al.
Veröffentlicht: (2025)
von: Sun, Yu, et al.
Veröffentlicht: (2025)
Agent-Centric Projection of Prompting Techniques and Implications for Synthetic Training Data for Large Language Models
von: Dhamani, Dhruv, et al.
Veröffentlicht: (2025)
von: Dhamani, Dhruv, et al.
Veröffentlicht: (2025)
AgentArch: A Comprehensive Benchmark to Evaluate Agent Architectures in Enterprise
von: Bogavelli, Tara, et al.
Veröffentlicht: (2025)
von: Bogavelli, Tara, et al.
Veröffentlicht: (2025)
WorkBench: a Benchmark Dataset for Agents in a Realistic Workplace Setting
von: Styles, Olly, et al.
Veröffentlicht: (2024)
von: Styles, Olly, et al.
Veröffentlicht: (2024)
MIND-Skill: Quality-Guaranteed Skill Generation via Multi-Agent Induction and Deduction
von: Li, Yixuan, et al.
Veröffentlicht: (2026)
von: Li, Yixuan, et al.
Veröffentlicht: (2026)
RPA-Check: A Multi-Stage Automated Framework for Evaluating Dynamic LLM-based Role-Playing Agents
von: Rosati, Riccardo, et al.
Veröffentlicht: (2026)
von: Rosati, Riccardo, et al.
Veröffentlicht: (2026)
MARCO: Multi-Agent Real-time Chat Orchestration
von: Shrimal, Anubhav, et al.
Veröffentlicht: (2024)
von: Shrimal, Anubhav, et al.
Veröffentlicht: (2024)
Toward Real-World Chinese Psychological Support Dialogues: CPsDD Dataset and a Co-Evolving Multi-Agent System
von: Shi, Yuanchen, et al.
Veröffentlicht: (2025)
von: Shi, Yuanchen, et al.
Veröffentlicht: (2025)
Adaptive Monitoring and Real-World Evaluation of Agentic AI Systems
von: Shukla, Manish
Veröffentlicht: (2025)
von: Shukla, Manish
Veröffentlicht: (2025)
ClinicBot: A Guideline-Grounded Clinical Chatbot with Prioritized Evidence RAG and Verifiable Citations
von: Nananukul, Navapat, et al.
Veröffentlicht: (2026)
von: Nananukul, Navapat, et al.
Veröffentlicht: (2026)
Collab-Overcooked: Benchmarking and Evaluating Large Language Models as Collaborative Agents
von: Sun, Haochen, et al.
Veröffentlicht: (2025)
von: Sun, Haochen, et al.
Veröffentlicht: (2025)
ATOD: An Evaluation Framework and Benchmark for Agentic Task-Oriented Dialogue Systems
von: Zhang, Yifei, et al.
Veröffentlicht: (2026)
von: Zhang, Yifei, et al.
Veröffentlicht: (2026)
Prompt Stability Matters: Evaluating and Optimizing Auto-Generated Prompt in General-Purpose Systems
von: Chen, Ke, et al.
Veröffentlicht: (2025)
von: Chen, Ke, et al.
Veröffentlicht: (2025)
Incorporating Error Level Noise Embedding for Improving LLM-Assisted Robustness in Persian Speech Recognition
von: Rahmani, Zahra, et al.
Veröffentlicht: (2025)
von: Rahmani, Zahra, et al.
Veröffentlicht: (2025)
Stated Preference for Interaction and Continued Engagement (SPICE): Evaluating an LLM's Willingness to Re-engage in Conversation
von: Rost, Thomas Manuel, et al.
Veröffentlicht: (2025)
von: Rost, Thomas Manuel, et al.
Veröffentlicht: (2025)
Advancing Agentic Systems: Dynamic Task Decomposition, Tool Integration and Evaluation using Novel Metrics and Dataset
von: Gabriel, Adrian Garret, et al.
Veröffentlicht: (2024)
von: Gabriel, Adrian Garret, et al.
Veröffentlicht: (2024)
High-order Interactions Modeling for Interpretable Multi-Agent Q-Learning
von: Xu, Qinyu, et al.
Veröffentlicht: (2025)
von: Xu, Qinyu, et al.
Veröffentlicht: (2025)
ChatMotion: A Multimodal Multi-Agent for Human Motion Analysis
von: Li, Lei, et al.
Veröffentlicht: (2025)
von: Li, Lei, et al.
Veröffentlicht: (2025)
A Multi-Memory Segment System for Generating High-Quality Long-Term Memory Content in Agents
von: Zhang, Gaoke, et al.
Veröffentlicht: (2025)
von: Zhang, Gaoke, et al.
Veröffentlicht: (2025)
A Comprehensive Survey of Self-Evolving AI Agents: A New Paradigm Bridging Foundation Models and Lifelong Agentic Systems
von: Fang, Jinyuan, et al.
Veröffentlicht: (2025)
von: Fang, Jinyuan, et al.
Veröffentlicht: (2025)
Governed Memory: A Production Architecture for Multi-Agent Workflows
von: Taheri, Hamed
Veröffentlicht: (2026)
von: Taheri, Hamed
Veröffentlicht: (2026)
ColorAgent: Building A Robust, Personalized, and Interactive OS Agent
von: Li, Ning, et al.
Veröffentlicht: (2025)
von: Li, Ning, et al.
Veröffentlicht: (2025)
A Text-to-Game Engine for UGC-Based Role-Playing Games
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
von: Zhang, Lei, et al.
Veröffentlicht: (2024)
LLM-as-RNN: A Recurrent Language Model for Memory Updates and Sequence Prediction
von: Lu, Yuxing, et al.
Veröffentlicht: (2026)
von: Lu, Yuxing, et al.
Veröffentlicht: (2026)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
von: Liu, Zijun, et al.
Veröffentlicht: (2023)
von: Liu, Zijun, et al.
Veröffentlicht: (2023)
LinguistAgent: A Reflective Multi-Model Platform for Automated Linguistic Annotation
von: Li, Bingru
Veröffentlicht: (2026)
von: Li, Bingru
Veröffentlicht: (2026)
DeepMEL: A Multi-Agent Collaboration Framework for Multimodal Entity Linking
von: Wang, Fang, et al.
Veröffentlicht: (2025)
von: Wang, Fang, et al.
Veröffentlicht: (2025)
MASLab: A Unified and Comprehensive Codebase for LLM-based Multi-Agent Systems
von: Ye, Rui, et al.
Veröffentlicht: (2025)
von: Ye, Rui, et al.
Veröffentlicht: (2025)
SciAgent: A Unified Multi-Agent System for Generalistic Scientific Reasoning
von: Li, Xuchen, et al.
Veröffentlicht: (2025)
von: Li, Xuchen, et al.
Veröffentlicht: (2025)
Large Language Model based Multi-Agents: A Survey of Progress and Challenges
von: Guo, Taicheng, et al.
Veröffentlicht: (2024)
von: Guo, Taicheng, et al.
Veröffentlicht: (2024)
Rethinking the Reliability of Multi-agent System: A Perspective from Byzantine Fault Tolerance
von: Zheng, Lifan, et al.
Veröffentlicht: (2025)
von: Zheng, Lifan, et al.
Veröffentlicht: (2025)
SMoA: Improving Multi-agent Large Language Models with Sparse Mixture-of-Agents
von: Li, Dawei, et al.
Veröffentlicht: (2024)
von: Li, Dawei, et al.
Veröffentlicht: (2024)
Towards Ethical Multi-Agent Systems of Large Language Models: A Mechanistic Interpretability Perspective
von: Lee, Jae Hee, et al.
Veröffentlicht: (2025)
von: Lee, Jae Hee, et al.
Veröffentlicht: (2025)
ConfAgents: A Conformal-Guided Multi-Agent Framework for Cost-Efficient Medical Diagnosis
von: Zhao, Huiya, et al.
Veröffentlicht: (2025)
von: Zhao, Huiya, et al.
Veröffentlicht: (2025)
FinRAG-12B: A Production-Validated Recipe for Grounded Question Answering in Banking
von: Katerenchuk, Denys, et al.
Veröffentlicht: (2026)
von: Katerenchuk, Denys, et al.
Veröffentlicht: (2026)
MARBLE: A Multi-Agent Rule-Based LLM Reasoning Engine for Accident Severity Prediction
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
von: Qasim, Kaleem Ullah, et al.
Veröffentlicht: (2025)
G2CP: A Graph-Grounded Communication Protocol for Verifiable and Efficient Multi-Agent Reasoning
von: Khaled, Karim Ben, et al.
Veröffentlicht: (2026)
von: Khaled, Karim Ben, et al.
Veröffentlicht: (2026)
UnityMAS-O: A General RL Optimization Framework for LLM-Based Multi-Agent Systems
von: Chen, Yiqun, et al.
Veröffentlicht: (2026)
von: Chen, Yiqun, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Information-Driven Fault Detection and Identification for Multi-Agent Spacecraft Systems: Collaborative On-Orbit Inspection Mission
von: Gupta, Akshita, et al.
Veröffentlicht: (2025) -
MiRAGE: A Multiagent Framework for Generating Multimodal Multihop Question-Answer Dataset for RAG Evaluation
von: Sahu, Chandan Kumar, et al.
Veröffentlicht: (2026) -
ReasonMed: A 370K Multi-Agent Generated Dataset for Advancing Medical Reasoning
von: Sun, Yu, et al.
Veröffentlicht: (2025) -
Agent-Centric Projection of Prompting Techniques and Implications for Synthetic Training Data for Large Language Models
von: Dhamani, Dhruv, et al.
Veröffentlicht: (2025) -
AgentArch: A Comprehensive Benchmark to Evaluate Agent Architectures in Enterprise
von: Bogavelli, Tara, et al.
Veröffentlicht: (2025)