CORTEX: Collaborative LLM Agents for High-Stakes Alert Triage
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Bowen, Tay, Yuan Shen, Liu, Howard, Pan, Jinhao, Luo, Kun, Zhu, Ziwei, Jordan, Chris |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Bias Association Discovery Framework for Open-Ended LLM Generations
by: Pan, Jinhao, et al.
Published: (2025)
by: Pan, Jinhao, et al.
Published: (2025)
Learning to Explain: Prototype-Based Surrogate Models for LLM Classification
by: Wei, Bowen, et al.
Published: (2025)
by: Wei, Bowen, et al.
Published: (2025)
What's Not Said Still Hurts: A Description-Based Evaluation Framework for Measuring Social Bias in LLMs
by: Pan, Jinhao, et al.
Published: (2025)
by: Pan, Jinhao, et al.
Published: (2025)
Advancing Interpretability in Text Classification through Prototype Learning
by: Wei, Bowen, et al.
Published: (2024)
by: Wei, Bowen, et al.
Published: (2024)
Guideline-Grounded Evidence Accumulation for High-Stakes Agent Verification
by: Zhang, Yichi, et al.
Published: (2026)
by: Zhang, Yichi, et al.
Published: (2026)
Talent or Luck? Evaluating Attribution Bias in Large Language Models
by: Raj, Chahat, et al.
Published: (2025)
by: Raj, Chahat, et al.
Published: (2025)
KG-Agent: An Efficient Autonomous Agent Framework for Complex Reasoning over Knowledge Graph
by: Jiang, Jinhao, et al.
Published: (2024)
by: Jiang, Jinhao, et al.
Published: (2024)
Implicit Behavioral Alignment of Language Agents in High-Stakes Crowd Simulations
by: Wang, Yunzhe, et al.
Published: (2025)
by: Wang, Yunzhe, et al.
Published: (2025)
Domaino1s: Guiding LLM Reasoning for Explainable Answers in High-Stakes Domains
by: Chu, Xu, et al.
Published: (2025)
by: Chu, Xu, et al.
Published: (2025)
High-Stakes Personalization: Rethinking LLM Customization for Individual Investor Decision-Making
by: Sawant, Yash Ganpat
Published: (2026)
by: Sawant, Yash Ganpat
Published: (2026)
AgentDropout: Dynamic Agent Elimination for Token-Efficient and High-Performance LLM-Based Multi-Agent Collaboration
by: Wang, Zhexuan, et al.
Published: (2025)
by: Wang, Zhexuan, et al.
Published: (2025)
Spoiler Alert: Narrative Forecasting as a Metric for Tension in LLM Storytelling
by: Sui, Peiqi, et al.
Published: (2026)
by: Sui, Peiqi, et al.
Published: (2026)
MedAide: Information Fusion and Anatomy of Medical Intents via LLM-based Agent Collaboration
by: Yang, Dingkang, et al.
Published: (2024)
by: Yang, Dingkang, et al.
Published: (2024)
Mitigating Hallucination in Large Vision-Language Models via Adaptive Attention Calibration
by: Fazli, Mehrdad, et al.
Published: (2025)
by: Fazli, Mehrdad, et al.
Published: (2025)
VIGNETTE: Socially Grounded Bias Evaluation for Vision-Language Models
by: Raj, Chahat, et al.
Published: (2025)
by: Raj, Chahat, et al.
Published: (2025)
Overhearing LLM Agents: A Survey, Taxonomy, and Roadmap
by: Zhu, Andrew, et al.
Published: (2025)
by: Zhu, Andrew, et al.
Published: (2025)
A Logical-Rule Autoencoder for Interpretable Recommendations
by: Pan, Jinhao, et al.
Published: (2026)
by: Pan, Jinhao, et al.
Published: (2026)
Beyond English and Evasion: A Human-Annotated Multi-Domain Benchmark for High-Stakes LLM Safety Evaluation in Chinese
by: Zaghouani, Wajdi, et al.
Published: (2026)
by: Zaghouani, Wajdi, et al.
Published: (2026)
Cochain: Balancing Insufficient and Excessive Collaboration in LLM Agent Workflows
by: Zhao, Jiaxing, et al.
Published: (2025)
by: Zhao, Jiaxing, et al.
Published: (2025)
Dental-TriageBench: Benchmarking Multimodal Reasoning for Hierarchical Dental Triage
by: He, Ziyi, et al.
Published: (2026)
by: He, Ziyi, et al.
Published: (2026)
AgentRouter: A Knowledge-Graph-Guided LLM Router for Collaborative Multi-Agent Question Answering
by: Zhang, Zheyuan, et al.
Published: (2025)
by: Zhang, Zheyuan, et al.
Published: (2025)
WISE-Flow: Workflow-Induced Structured Experience for Self-Evolving Conversational Service Agents
by: Zhou, Yuqing, et al.
Published: (2026)
by: Zhou, Yuqing, et al.
Published: (2026)
PRBench: Large-Scale Expert Rubrics for Evaluating High-Stakes Professional Reasoning
by: Akyürek, Afra Feyza, et al.
Published: (2025)
by: Akyürek, Afra Feyza, et al.
Published: (2025)
Triaging Threats to Specialized Guardrails
by: Mo, Wenjie Jacky, et al.
Published: (2026)
by: Mo, Wenjie Jacky, et al.
Published: (2026)
AgentCollab: A Self-Evaluation-Driven Collaboration Paradigm for Efficient LLM Agents
by: Gao, Wenbo, et al.
Published: (2026)
by: Gao, Wenbo, et al.
Published: (2026)
PPAI: Enabling Personalized LLM Agent Interoperability for Collaborative Edge Intelligence
by: Wang, Zile, et al.
Published: (2026)
by: Wang, Zile, et al.
Published: (2026)
LLMs Struggle to Reject False Presuppositions when Misinformation Stakes are High
by: Sieker, Judith, et al.
Published: (2025)
by: Sieker, Judith, et al.
Published: (2025)
Evaluating Differentially Private Synthetic Data Generation in High-Stakes Domains
by: Ramesh, Krithika, et al.
Published: (2024)
by: Ramesh, Krithika, et al.
Published: (2024)
Rationale-Augmented Retrieval with Constrained LLM Re-Ranking for Task Discovery
by: Wei, Bowen
Published: (2025)
by: Wei, Bowen
Published: (2025)
SafetyFlow: An Agent-Flow System for Automated LLM Safety Benchmarking
by: Zhu, Xiangyang, et al.
Published: (2025)
by: Zhu, Xiangyang, et al.
Published: (2025)
TriageSim: A Conversational Emergency Triage Simulation Framework from Structured Electronic Health Records
by: Srirag, Dipankar, et al.
Published: (2026)
by: Srirag, Dipankar, et al.
Published: (2026)
A LLM Benchmark based on the Minecraft Builder Dialog Agent Task
by: Madge, Chris, et al.
Published: (2024)
by: Madge, Chris, et al.
Published: (2024)
Beyond Semantic Understanding: Preserving Collaborative Frequency Components in LLM-based Recommendation
by: Wang, Minhao, et al.
Published: (2025)
by: Wang, Minhao, et al.
Published: (2025)
LLM-Collaboration on Automatic Science Journalism for the General Audience
by: Jiang, Gongyao, et al.
Published: (2024)
by: Jiang, Gongyao, et al.
Published: (2024)
SRAP-Agent: Simulating and Optimizing Scarce Resource Allocation Policy with LLM-based Agent
by: Ji, Jiarui, et al.
Published: (2024)
by: Ji, Jiarui, et al.
Published: (2024)
A Dynamic LLM-Powered Agent Network for Task-Oriented Agent Collaboration
by: Liu, Zijun, et al.
Published: (2023)
by: Liu, Zijun, et al.
Published: (2023)
Internal Representation, Not Clinical Knowledge: Where Apparent LLM Triage Failures Originate
by: Navarro, David Fraile, et al.
Published: (2026)
by: Navarro, David Fraile, et al.
Published: (2026)
When One LLM Drools, Multi-LLM Collaboration Rules
by: Feng, Shangbin, et al.
Published: (2025)
by: Feng, Shangbin, et al.
Published: (2025)
ACC-Collab: An Actor-Critic Approach to Multi-Agent LLM Collaboration
by: Estornell, Andrew, et al.
Published: (2024)
by: Estornell, Andrew, et al.
Published: (2024)
SynthTextEval: Synthetic Text Data Generation and Evaluation for High-Stakes Domains
by: Ramesh, Krithika, et al.
Published: (2025)
by: Ramesh, Krithika, et al.
Published: (2025)
Similar Items
-
Bias Association Discovery Framework for Open-Ended LLM Generations
by: Pan, Jinhao, et al.
Published: (2025) -
Learning to Explain: Prototype-Based Surrogate Models for LLM Classification
by: Wei, Bowen, et al.
Published: (2025) -
What's Not Said Still Hurts: A Description-Based Evaluation Framework for Measuring Social Bias in LLMs
by: Pan, Jinhao, et al.
Published: (2025) -
Advancing Interpretability in Text Classification through Prototype Learning
by: Wei, Bowen, et al.
Published: (2024) -
Guideline-Grounded Evidence Accumulation for High-Stakes Agent Verification
by: Zhang, Yichi, et al.
Published: (2026)