OpenHospital: A Thing-in-itself Arena for Evolving and Benchmarking LLM-based Collective Intelligence
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Peigen, Ding, Rui, Mao, Yuren, Jiang, Ziyan, Ye, Yuxiang, Gao, Yunjun, Zhang, Ying, Sun, Renjie, Lai, Longbin, Qian, Zhengping |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
G-Boost: Boosting Private SLMs with General LLMs
by: Fan, Yijiang, et al.
Published: (2025)
by: Fan, Yijiang, et al.
Published: (2025)
Agent-Kernel: A MicroKernel Multi-Agent System Framework for Adaptive Social Simulation Powered by LLMs
by: Mao, Yuren, et al.
Published: (2025)
by: Mao, Yuren, et al.
Published: (2025)
scAgent: Universal Single-Cell Annotation via a LLM Agent
by: Mao, Yuren, et al.
Published: (2025)
by: Mao, Yuren, et al.
Published: (2025)
SpecDB: LLM-Generated Customized Databases via Feature-Oriented Decomposition
by: Lou, Yunkai, et al.
Published: (2026)
by: Lou, Yunkai, et al.
Published: (2026)
Snoopy: Effective and Efficient Semantic Join Discovery via Proxy Columns
by: Guo, Yuxiang, et al.
Published: (2025)
by: Guo, Yuxiang, et al.
Published: (2025)
CT-Agent: A Multimodal-LLM Agent for 3D CT Radiology Question Answering
by: Mao, Yuren, et al.
Published: (2025)
by: Mao, Yuren, et al.
Published: (2025)
Birdie: Natural Language-Driven Table Discovery Using Differentiable Search Index
by: Guo, Yuxiang, et al.
Published: (2025)
by: Guo, Yuxiang, et al.
Published: (2025)
FIT-RAG: Black-Box RAG with Factual Information and Token Reduction
by: Mao, Yuren, et al.
Published: (2024)
by: Mao, Yuren, et al.
Published: (2024)
Text-to-Pipeline: Bridging Natural Language and Data Preparation Pipelines
by: Ge, Yuhang, et al.
Published: (2025)
by: Ge, Yuhang, et al.
Published: (2025)
SQL-Factory: A Multi-Agent Framework for High-Quality and Large-Scale SQL Generation
by: Li, Jiahui, et al.
Published: (2025)
by: Li, Jiahui, et al.
Published: (2025)
LASER: A Data-Centric Method for Low-Cost and Efficient SQL Rewriting based on SQL-GRPO
by: Li, Jiahui, et al.
Published: (2026)
by: Li, Jiahui, et al.
Published: (2026)
SparDL: Distributed Deep Learning Training with Efficient Sparse Communication
by: Zhao, Minjun, et al.
Published: (2023)
by: Zhao, Minjun, et al.
Published: (2023)
A Hypergraph-Based Framework for Exploratory Business Intelligence
by: Lou, Yunkai, et al.
Published: (2026)
by: Lou, Yunkai, et al.
Published: (2026)
ClawArena: Benchmarking AI Agents in Evolving Information Environments
by: Ji, Haonian, et al.
Published: (2026)
by: Ji, Haonian, et al.
Published: (2026)
OpenDataArena: A Fair and Open Arena for Benchmarking Post-Training Dataset Value
by: Cai, Mengzhang, et al.
Published: (2025)
by: Cai, Mengzhang, et al.
Published: (2025)
Open-LLM-Leaderboard: From Multi-choice to Open-style Questions for LLMs Evaluation, Benchmark, and Arena
by: Myrzakhan, Aidar, et al.
Published: (2024)
by: Myrzakhan, Aidar, et al.
Published: (2024)
CodeArena: A Collective Evaluation Platform for LLM Code Generation
by: Du, Mingzhe, et al.
Published: (2025)
by: Du, Mingzhe, et al.
Published: (2025)
A Survey on LoRA of Large Language Models
by: Mao, Yuren, et al.
Published: (2024)
by: Mao, Yuren, et al.
Published: (2024)
DAgent: A Relational Database-Driven Data Analysis Report Generation Agent
by: Xu, Wenyi, et al.
Published: (2025)
by: Xu, Wenyi, et al.
Published: (2025)
Liberalism against itself
by: Samuel Moyn
Published: (2024)
by: Samuel Moyn
Published: (2024)
Oblique Subduction Model Script (Japan Sea Openning Paper) - Underworld2
by: Luo, Peigen, et al.
Published: (2026)
by: Luo, Peigen, et al.
Published: (2026)
DTBench: A Synthetic Benchmark for Document-to-Table Extraction
by: Guo, Yuxiang, et al.
Published: (2026)
by: Guo, Yuxiang, et al.
Published: (2026)
F-invariant in cluster algebras
by: Cao, Peigen
Published: (2023)
by: Cao, Peigen
Published: (2023)
Modules determined by their Newton polytopes
by: Cao, Peigen
Published: (2025)
by: Cao, Peigen
Published: (2025)
F-invariant and E-invariant
by: Cao, Peigen
Published: (2025)
by: Cao, Peigen
Published: (2025)
Newton polytopes in cluster algebras and $τ$-tilting theory
by: Cao, Peigen
Published: (2026)
by: Cao, Peigen
Published: (2026)
Moreover. America smiles at itself
Published: (1996)
Published: (1996)
Clean, safe and it drives itself
by: Varios
Published: (2013)
by: Varios
Published: (2013)
Can We Predict Before Executing Machine Learning Agents?
by: Zheng, Jingsheng, et al.
Published: (2026)
by: Zheng, Jingsheng, et al.
Published: (2026)
FinSQL: Model-Agnostic LLMs-based Text-to-SQL Framework for Financial Analysis
by: Zhang, Chao, et al.
Published: (2024)
by: Zhang, Chao, et al.
Published: (2024)
Self-Disguise Attack: Induce the LLM to disguise itself for AIGT detection evasion
by: Zhou, Yinghan, et al.
Published: (2025)
by: Zhou, Yinghan, et al.
Published: (2025)
LLM-as-a-Prophet: Understanding Predictive Intelligence with Prophet Arena
by: Yang, Qingchuan, et al.
Published: (2025)
by: Yang, Qingchuan, et al.
Published: (2025)
RouterArena: An Open Platform for Comprehensive Comparison of LLM Routers
by: Lu, Yifan, et al.
Published: (2025)
by: Lu, Yifan, et al.
Published: (2025)
MedBookVQA: A Systematic and Comprehensive Medical Benchmark Derived from Open-Access Book
by: Yip, Sau Lai, et al.
Published: (2025)
by: Yip, Sau Lai, et al.
Published: (2025)
SkillClaw: Let Skills Evolve Collectively with Agentic Evolver
by: Ma, Ziyu, et al.
Published: (2026)
by: Ma, Ziyu, et al.
Published: (2026)
Benchmarking LLM Faithfulness in RAG with Evolving Leaderboards
by: Tamber, Manveer Singh, et al.
Published: (2025)
by: Tamber, Manveer Singh, et al.
Published: (2025)
FedAIoT: A Federated Learning Benchmark for Artificial Intelligence of Things
by: Alam, Samiul, et al.
Published: (2023)
by: Alam, Samiul, et al.
Published: (2023)
The new Avengers fear itself #3 (2018)
Published: (2012)
Published: (2012)
Science and technology. Nature rarely repeats itself
Published: (1997)
Published: (1997)
Can the growing human population feed itself?
Published: (1994)
Published: (1994)
Similar Items
-
G-Boost: Boosting Private SLMs with General LLMs
by: Fan, Yijiang, et al.
Published: (2025) -
Agent-Kernel: A MicroKernel Multi-Agent System Framework for Adaptive Social Simulation Powered by LLMs
by: Mao, Yuren, et al.
Published: (2025) -
scAgent: Universal Single-Cell Annotation via a LLM Agent
by: Mao, Yuren, et al.
Published: (2025) -
SpecDB: LLM-Generated Customized Databases via Feature-Oriented Decomposition
by: Lou, Yunkai, et al.
Published: (2026) -
Snoopy: Effective and Efficient Semantic Join Discovery via Proxy Columns
by: Guo, Yuxiang, et al.
Published: (2025)