Salvato in:
| Autori principali: | Zhou, Hanlin, Chan, Huah Yong |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2602.01797 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Runtime Burden Allocation for Structured LLM Routing in Agentic Expert Systems: A Full-Factorial Cross-Backend Methodology
di: Hanlin, Zhou, et al.
Pubblicazione: (2026)
di: Hanlin, Zhou, et al.
Pubblicazione: (2026)
ADEMA: A Knowledge-State Orchestration Architecture for Long-Horizon Knowledge Synthesis with LLMAgents
di: Hanlin, Zhou, et al.
Pubblicazione: (2026)
di: Hanlin, Zhou, et al.
Pubblicazione: (2026)
Automatic Adjustment of HPA Parameters and Attack Prevention in Kubernetes Using Random Forests
di: Zhou, Hanlin, et al.
Pubblicazione: (2026)
di: Zhou, Hanlin, et al.
Pubblicazione: (2026)
Position: agentic AI orchestration should be Bayes-consistent
di: Papamarkou, Theodore, et al.
Pubblicazione: (2026)
di: Papamarkou, Theodore, et al.
Pubblicazione: (2026)
Dynamic fairness-aware recommendation through multi-agent social choice
di: Aird, Amanda, et al.
Pubblicazione: (2023)
di: Aird, Amanda, et al.
Pubblicazione: (2023)
MARS: toward more efficient multi-agent collaboration for LLM reasoning
di: Wang, Xiao, et al.
Pubblicazione: (2025)
di: Wang, Xiao, et al.
Pubblicazione: (2025)
Evidence-based diagnostic reasoning with multi-agent copilot for human pathology
di: Weishaupt, Luca L., et al.
Pubblicazione: (2025)
di: Weishaupt, Luca L., et al.
Pubblicazione: (2025)
TrajOnco: a multi-agent framework for temporal reasoning over longitudinal EHR for multi-cancer early detection
di: Zeng, Sihang, et al.
Pubblicazione: (2026)
di: Zeng, Sihang, et al.
Pubblicazione: (2026)
Improving monotonic optimization in heterogeneous multi-agent reinforcement learning with optimal marginal deterministic policy gradient
di: Yu, Xiaoyang, et al.
Pubblicazione: (2025)
di: Yu, Xiaoyang, et al.
Pubblicazione: (2025)
EMA Policy Gradient: Taming Reinforcement Learning for LLMs with EMA Anchor and Top-k KL
di: Zhang, Lunjun, et al.
Pubblicazione: (2026)
di: Zhang, Lunjun, et al.
Pubblicazione: (2026)
Learning for routing: A guided review of recent developments and future directions
di: Zhou, Fangting, et al.
Pubblicazione: (2025)
di: Zhou, Fangting, et al.
Pubblicazione: (2025)
AIonopedia: an LLM agent orchestrating multimodal learning for ionic liquid discovery
di: Yin, Yuqi, et al.
Pubblicazione: (2025)
di: Yin, Yuqi, et al.
Pubblicazione: (2025)
Adaptive routing protocols for determining optimal paths in AI multi-agent systems: a priority- and learning-enhanced approach
di: Panayotov, Theodor, et al.
Pubblicazione: (2025)
di: Panayotov, Theodor, et al.
Pubblicazione: (2025)
SIKeD: Self-guided Iterative Knowledge Distillation for mathematical reasoning
di: Adarsh, Shivam, et al.
Pubblicazione: (2024)
di: Adarsh, Shivam, et al.
Pubblicazione: (2024)
Automated legal reasoning with discretion to act using s(LAW)
di: Arias, Joaquín, et al.
Pubblicazione: (2024)
di: Arias, Joaquín, et al.
Pubblicazione: (2024)
WiseMind: a knowledge-guided multi-agent framework for accurate and empathetic psychiatric diagnosis
di: Wu, Yuqi, et al.
Pubblicazione: (2025)
di: Wu, Yuqi, et al.
Pubblicazione: (2025)
Biomedical reasoning in action: Multi-agent System for Auditable Biomedical Evidence Synthesis
di: Wysocki, Oskar, et al.
Pubblicazione: (2025)
di: Wysocki, Oskar, et al.
Pubblicazione: (2025)
Learning to reason about rare diseases through retrieval-augmented agents
di: Kim, Ha Young, et al.
Pubblicazione: (2025)
di: Kim, Ha Young, et al.
Pubblicazione: (2025)
SciAgents: Automating scientific discovery through multi-agent intelligent graph reasoning
di: Ghafarollahi, Alireza, et al.
Pubblicazione: (2024)
di: Ghafarollahi, Alireza, et al.
Pubblicazione: (2024)
Reinforcement Learning in hyperbolic space for multi-step reasoning
di: Xu, Tao, et al.
Pubblicazione: (2025)
di: Xu, Tao, et al.
Pubblicazione: (2025)
DecompSR: A dataset for decomposed analyses of compositional multihop spatial reasoning
di: McPheat, Lachlan, et al.
Pubblicazione: (2025)
di: McPheat, Lachlan, et al.
Pubblicazione: (2025)
Retrieval-augmented reasoning with lean language models
di: Chan, Ryan Sze-Yin, et al.
Pubblicazione: (2025)
di: Chan, Ryan Sze-Yin, et al.
Pubblicazione: (2025)
EMA Without the Lag: Bias-Corrected Iterate Averaging Schemes
di: Block, Adam, et al.
Pubblicazione: (2025)
di: Block, Adam, et al.
Pubblicazione: (2025)
Is continuous CoT better suited for multi-lingual reasoning?
di: Bashir, Ali Hamza, et al.
Pubblicazione: (2026)
di: Bashir, Ali Hamza, et al.
Pubblicazione: (2026)
Pushing the Limits of Low-Bit Optimizers: A Focus on EMA Dynamics
di: Xu, Cong, et al.
Pubblicazione: (2025)
di: Xu, Cong, et al.
Pubblicazione: (2025)
Incorporating uncertainty quantification into travel mode choice modeling: a Bayesian neural network (BNN) approach and an uncertainty-guided active survey framework
di: Zheng, Shuwen, et al.
Pubblicazione: (2024)
di: Zheng, Shuwen, et al.
Pubblicazione: (2024)
From LLM-anation to LLM-orchestrator: Coordinating Small Models for Data Labeling
di: Lu, Yao, et al.
Pubblicazione: (2025)
di: Lu, Yao, et al.
Pubblicazione: (2025)
The cognitive companion: a lightweight parallel monitoring architecture for detecting and recovering from reasoning degradation in LLM agents
di: Khan, Rafflesia, et al.
Pubblicazione: (2026)
di: Khan, Rafflesia, et al.
Pubblicazione: (2026)
Configurable multi-agent framework for scalable and realistic testing of llm-based agents
di: Wang, Sai, et al.
Pubblicazione: (2025)
di: Wang, Sai, et al.
Pubblicazione: (2025)
Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
di: Sclar, Melanie, et al.
Pubblicazione: (2024)
di: Sclar, Melanie, et al.
Pubblicazione: (2024)
Performance of AI agents based on reasoning language models on ALD process optimization tasks
di: Yanguas-Gil, Angel
Pubblicazione: (2026)
di: Yanguas-Gil, Angel
Pubblicazione: (2026)
A Kubernetes custom scheduler based on reinforcement learning for compute-intensive pods
di: Zhou, Hanlin, et al.
Pubblicazione: (2026)
di: Zhou, Hanlin, et al.
Pubblicazione: (2026)
Performance Comparison of IBN orchestration using LLM and SLMs
di: Phone, Wai Lwin, et al.
Pubblicazione: (2026)
di: Phone, Wai Lwin, et al.
Pubblicazione: (2026)
MetaOpenFOAM: an LLM-based multi-agent framework for CFD
di: Chen, Yuxuan, et al.
Pubblicazione: (2024)
di: Chen, Yuxuan, et al.
Pubblicazione: (2024)
Causal vs. Anticausal merging of predictors
di: Mejia, Sergio Hernan Garrido, et al.
Pubblicazione: (2025)
di: Mejia, Sergio Hernan Garrido, et al.
Pubblicazione: (2025)
Reasonably reasoning AI agents can avoid game-theoretic failures in zero-shot, provably
di: Kang, Enoch Hyunwook
Pubblicazione: (2026)
di: Kang, Enoch Hyunwook
Pubblicazione: (2026)
Adaptive parameter sharing for multi-agent reinforcement learning
di: Li, Dapeng, et al.
Pubblicazione: (2023)
di: Li, Dapeng, et al.
Pubblicazione: (2023)
Automated stereotactic radiosurgery planning using a human-in-the-loop reasoning large language model agent
di: Nusrat, Humza, et al.
Pubblicazione: (2025)
di: Nusrat, Humza, et al.
Pubblicazione: (2025)
PathReasoning: A multimodal reasoning agent for query-based ROI navigation on whole-slide images
di: Zhang, Kunpeng, et al.
Pubblicazione: (2025)
di: Zhang, Kunpeng, et al.
Pubblicazione: (2025)
Using multi-agent architecture to mitigate the risk of LLM hallucinations
di: Amer, Abd Elrahman, et al.
Pubblicazione: (2025)
di: Amer, Abd Elrahman, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Runtime Burden Allocation for Structured LLM Routing in Agentic Expert Systems: A Full-Factorial Cross-Backend Methodology
di: Hanlin, Zhou, et al.
Pubblicazione: (2026) -
ADEMA: A Knowledge-State Orchestration Architecture for Long-Horizon Knowledge Synthesis with LLMAgents
di: Hanlin, Zhou, et al.
Pubblicazione: (2026) -
Automatic Adjustment of HPA Parameters and Attack Prevention in Kubernetes Using Random Forests
di: Zhou, Hanlin, et al.
Pubblicazione: (2026) -
Position: agentic AI orchestration should be Bayes-consistent
di: Papamarkou, Theodore, et al.
Pubblicazione: (2026) -
Dynamic fairness-aware recommendation through multi-agent social choice
di: Aird, Amanda, et al.
Pubblicazione: (2023)