Kaleidoscopic Teaming in Multi Agent Simulations
Fuente:
arXiv
Guardado en:
| Autores principales: | Mehrabi, Ninareh, Kumarage, Tharindu, Chang, Kai-Wei, Galstyan, Aram, Gupta, Rahul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Towards Safety Reasoning in LLMs: AI-agentic Deliberation for Policy-embedded CoT Data Creation
por: Kumarage, Tharindu, et al.
Publicado: (2025)
por: Kumarage, Tharindu, et al.
Publicado: (2025)
FLIRT: Feedback Loop In-context Red Teaming
por: Mehrabi, Ninareh, et al.
Publicado: (2023)
por: Mehrabi, Ninareh, et al.
Publicado: (2023)
ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System
por: Liang, Jiacheng, et al.
Publicado: (2026)
por: Liang, Jiacheng, et al.
Publicado: (2026)
Data Advisor: Dynamic Data Curation for Safety Alignment of Large Language Models
por: Wang, Fei, et al.
Publicado: (2024)
por: Wang, Fei, et al.
Publicado: (2024)
K-Edit: Language Model Editing with Contextual Knowledge Awareness
por: Markowitz, Elan, et al.
Publicado: (2025)
por: Markowitz, Elan, et al.
Publicado: (2025)
Tree-of-Traversals: A Zero-Shot Reasoning Algorithm for Augmenting Black-box Language Models with Knowledge Graphs
por: Markowitz, Elan, et al.
Publicado: (2024)
por: Markowitz, Elan, et al.
Publicado: (2024)
Strategize Globally, Adapt Locally: A Multi-Turn Red Teaming Agent with Dual-Level Learning
por: Chen, Si, et al.
Publicado: (2025)
por: Chen, Si, et al.
Publicado: (2025)
Tokenization Matters: Navigating Data-Scarce Tokenization for Gender Inclusive Language Technologies
por: Ovalle, Anaelia, et al.
Publicado: (2023)
por: Ovalle, Anaelia, et al.
Publicado: (2023)
Emergent Strategic Reasoning Risks in AI: A Taxonomy-Driven Evaluation Framework
por: Kumarage, Tharindu, et al.
Publicado: (2026)
por: Kumarage, Tharindu, et al.
Publicado: (2026)
SWAN: Semantic Watermarking with Abstract Meaning Representation
por: Ye, Ziping, et al.
Publicado: (2026)
por: Ye, Ziping, et al.
Publicado: (2026)
FERRET: Framework for Expansion Reliant Red Teaming
por: Mehrabi, Ninareh, et al.
Publicado: (2026)
por: Mehrabi, Ninareh, et al.
Publicado: (2026)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
por: Kumarage, Tharindu, et al.
Publicado: (2024)
por: Kumarage, Tharindu, et al.
Publicado: (2024)
DiCoRe: Enhancing Zero-shot Event Detection via Divergent-Convergent LLM Reasoning
por: Parekh, Tanmay, et al.
Publicado: (2025)
por: Parekh, Tanmay, et al.
Publicado: (2025)
On the steerability of large language models toward data-driven personas
por: Li, Junyi, et al.
Publicado: (2023)
por: Li, Junyi, et al.
Publicado: (2023)
Diagnosing Memorization in Chain-of-Thought Reasoning, One Token at a Time
por: Li, Huihan, et al.
Publicado: (2025)
por: Li, Huihan, et al.
Publicado: (2025)
Co-Evolving Agents: Learning from Failures as Hard Negatives
por: Jung, Yeonsung, et al.
Publicado: (2025)
por: Jung, Yeonsung, et al.
Publicado: (2025)
Attribute Controlled Fine-tuning for Large Language Models: A Case Study on Detoxification
por: Meng, Tao, et al.
Publicado: (2024)
por: Meng, Tao, et al.
Publicado: (2024)
Adaptive Video Understanding Agent: Enhancing efficiency with dynamic frame sampling and feedback-driven reasoning
por: Jeoung, Sullam, et al.
Publicado: (2024)
por: Jeoung, Sullam, et al.
Publicado: (2024)
RedditESS: A Mental Health Social Support Interaction Dataset -- Understanding Effective Social Support to Refine AI-Driven Support Tools
por: Alghamdi, Zeyad, et al.
Publicado: (2025)
por: Alghamdi, Zeyad, et al.
Publicado: (2025)
IdleSpec: Exploiting Idle Time via Speculative Planning for LLM Agents
por: Choi, Daewon, et al.
Publicado: (2026)
por: Choi, Daewon, et al.
Publicado: (2026)
Asking Back: Interaction-Layer Antidistillation Watermarks
por: Yang, Guang, et al.
Publicado: (2026)
por: Yang, Guang, et al.
Publicado: (2026)
Generative Kaleidoscopic Networks
por: Shrivastava, Harsh
Publicado: (2024)
por: Shrivastava, Harsh
Publicado: (2024)
A Survey of AI-generated Text Forensic Systems: Detection, Attribution, and Characterization
por: Kumarage, Tharindu, et al.
Publicado: (2024)
por: Kumarage, Tharindu, et al.
Publicado: (2024)
KG-LLM-Bench: A Scalable Benchmark for Evaluating LLM Reasoning on Textualized Knowledge Graphs
por: Markowitz, Elan, et al.
Publicado: (2025)
por: Markowitz, Elan, et al.
Publicado: (2025)
MultiLS: A Multi-task Lexical Simplification Framework
por: North, Kai, et al.
Publicado: (2024)
por: North, Kai, et al.
Publicado: (2024)
Graph Based Deep Reinforcement Learning Aided by Transformers for Multi-Agent Cooperation
por: Elrod, Michael, et al.
Publicado: (2025)
por: Elrod, Michael, et al.
Publicado: (2025)
MDTeamGPT: A Self-Evolving LLM-based Multi-Agent Framework for Multi-Disciplinary Team Medical Consultation
por: Chen, Kai, et al.
Publicado: (2025)
por: Chen, Kai, et al.
Publicado: (2025)
Kaleidoscope: Learnable Masks for Heterogeneous Multi-agent Reinforcement Learning
por: Li, Xinran, et al.
Publicado: (2024)
por: Li, Xinran, et al.
Publicado: (2024)
SeRA: Self-Reviewing and Alignment of Large Language Models using Implicit Reward Margins
por: Ko, Jongwoo, et al.
Publicado: (2024)
por: Ko, Jongwoo, et al.
Publicado: (2024)
Ontology-Aware RAG for Improved Question-Answering in Cybersecurity Education
por: Zhao, Chengshuai, et al.
Publicado: (2024)
por: Zhao, Chengshuai, et al.
Publicado: (2024)
ExComm: Exploration-Stage Communication for Error-Resilient Agentic Test-Time Scaling
por: Song, Woomin, et al.
Publicado: (2026)
por: Song, Woomin, et al.
Publicado: (2026)
X-Teaming: Multi-Turn Jailbreaks and Defenses with Adaptive Multi-Agents
por: Rahman, Salman, et al.
Publicado: (2025)
por: Rahman, Salman, et al.
Publicado: (2025)
Prompt Perturbation Consistency Learning for Robust Language Models
por: Qiang, Yao, et al.
Publicado: (2024)
por: Qiang, Yao, et al.
Publicado: (2024)
DeepForgeSeal: Latent Space-Driven Semi-Fragile Watermarking for Deepfake Detection Using Multi-Agent Adversarial Reinforcement Learning
por: Fernando, Tharindu, et al.
Publicado: (2025)
por: Fernando, Tharindu, et al.
Publicado: (2025)
ALEXSIS-PT: A New Resource for Portuguese Lexical Simplification
por: North, Kai, et al.
Publicado: (2022)
por: North, Kai, et al.
Publicado: (2022)
Kaleidoscope Gallery: Exploring Ethics and Generative AI Through Art
por: Issak, Alayt, et al.
Publicado: (2025)
por: Issak, Alayt, et al.
Publicado: (2025)
Claw AI Lab: An Autonomous Multi-Agent Research Team
por: Wu, Fan, et al.
Publicado: (2026)
por: Wu, Fan, et al.
Publicado: (2026)
Multi-Agent Teams Hold Experts Back
por: Pappu, Aneesh, et al.
Publicado: (2026)
por: Pappu, Aneesh, et al.
Publicado: (2026)
CyberBOT: Towards Reliable Cybersecurity Education via Ontology-Grounded Retrieval Augmented Generation
por: Zhao, Chengshuai, et al.
Publicado: (2025)
por: Zhao, Chengshuai, et al.
Publicado: (2025)
Before Humans Join the Team: Diagnosing Coordination Failures in Healthcare Robot Team Simulation
por: Bai, Yuanchen, et al.
Publicado: (2025)
por: Bai, Yuanchen, et al.
Publicado: (2025)
Ejemplares similares
-
Towards Safety Reasoning in LLMs: AI-agentic Deliberation for Policy-embedded CoT Data Creation
por: Kumarage, Tharindu, et al.
Publicado: (2025) -
FLIRT: Feedback Loop In-context Red Teaming
por: Mehrabi, Ninareh, et al.
Publicado: (2023) -
ARES: Adaptive Red-Teaming and End-to-End Repair of Policy-Reward System
por: Liang, Jiacheng, et al.
Publicado: (2026) -
Data Advisor: Dynamic Data Curation for Safety Alignment of Large Language Models
por: Wang, Fei, et al.
Publicado: (2024) -
K-Edit: Language Model Editing with Contextual Knowledge Awareness
por: Markowitz, Elan, et al.
Publicado: (2025)