ALIGN: Aligned Delegation with Performance Guarantees for Multi-Agent LLM Reasoning
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhu, Tong, Chen, Baiting, Zhou, Jin, Zhou, Hua, Sankararaman, Sriram, Dai, Xiaowu |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Prediction-Powered Conditional Inference
por: Sui, Yang, et al.
Publicado: (2026)
por: Sui, Yang, et al.
Publicado: (2026)
Incentivizing Truthful Language Models via Peer Elicitation Games
por: Chen, Baiting, et al.
Publicado: (2025)
por: Chen, Baiting, et al.
Publicado: (2025)
Detecting LLM-Generated Text with Performance Guarantees
por: Zhou, Hongyi, et al.
Publicado: (2026)
por: Zhou, Hongyi, et al.
Publicado: (2026)
Learn then Decide: A Learning Approach for Designing Data Marketplaces
por: Gao, Yingqi, et al.
Publicado: (2025)
por: Gao, Yingqi, et al.
Publicado: (2025)
Common-agency Games for Multi-Objective Test-Time Alignment
por: Chen, Baiting, et al.
Publicado: (2026)
por: Chen, Baiting, et al.
Publicado: (2026)
dotears: Scalable, consistent DAG estimation using observational and interventional data
por: Xue, Albert, et al.
Publicado: (2023)
por: Xue, Albert, et al.
Publicado: (2023)
Uncertainty-Aware Multimodal Learning via Conformal Shapley Intervals
por: Chandy, Mathew, et al.
Publicado: (2026)
por: Chandy, Mathew, et al.
Publicado: (2026)
CACTI: Leveraging Copy Masking and Contextual Information to Improve Tabular Data Imputation
por: Gorla, Aditya, et al.
Publicado: (2025)
por: Gorla, Aditya, et al.
Publicado: (2025)
A Robust Multi-Item Auction Design with Statistical Learning
por: Han, Jiale, et al.
Publicado: (2023)
por: Han, Jiale, et al.
Publicado: (2023)
Prior-Aligned Meta-RL: Thompson Sampling with Learned Priors and Guarantees in Finite-Horizon MDPs
por: Zhou, Runlin, et al.
Publicado: (2025)
por: Zhou, Runlin, et al.
Publicado: (2025)
Variance Reduction via Resampling and Experience Replay
por: Han, Jiale, et al.
Publicado: (2025)
por: Han, Jiale, et al.
Publicado: (2025)
Performative Risk Control: Calibrating Models for Reliable Deployment under Performativity
por: Li, Victor, et al.
Publicado: (2025)
por: Li, Victor, et al.
Publicado: (2025)
Probabilistic Soundness Guarantees in LLM Reasoning Chains
por: You, Weiqiu, et al.
Publicado: (2025)
por: You, Weiqiu, et al.
Publicado: (2025)
Post-Regularization Confidence Bands for Ordinary Differential Equations
por: Dai, Xiaowu, et al.
Publicado: (2021)
por: Dai, Xiaowu, et al.
Publicado: (2021)
Calibrate-Then-Delegate: Safety Monitoring with Risk and Budget Guarantees via Model Cascades
por: Pona, Edoardo, et al.
Publicado: (2026)
por: Pona, Edoardo, et al.
Publicado: (2026)
Certainty in Uncertainty: Reasoning over Uncertain Knowledge Graphs with Statistical Guarantees
por: Zhu, Yuqicheng, et al.
Publicado: (2025)
por: Zhu, Yuqicheng, et al.
Publicado: (2025)
Multi-Objective Alignment of Language Models for Personalized Psychotherapy
por: Beikzadeh, Mehrab, et al.
Publicado: (2026)
por: Beikzadeh, Mehrab, et al.
Publicado: (2026)
SWEET-RL: Training Multi-Turn LLM Agents on Collaborative Reasoning Tasks
por: Zhou, Yifei, et al.
Publicado: (2025)
por: Zhou, Yifei, et al.
Publicado: (2025)
Conformal Prediction: A Data Perspective
por: Zhou, Xiaofan, et al.
Publicado: (2024)
por: Zhou, Xiaofan, et al.
Publicado: (2024)
Efficient Evaluation of LLM Performance with Statistical Guarantees
por: Wu, Skyler, et al.
Publicado: (2026)
por: Wu, Skyler, et al.
Publicado: (2026)
Online Auction Design Using Distribution-Free Uncertainty Quantification with Applications to E-Commerce
por: Han, Jiale, et al.
Publicado: (2024)
por: Han, Jiale, et al.
Publicado: (2024)
ALIGN: Adversarial Learning for Generalizable Speech Neuroprosthesis
por: Zhang, Zhanqi, et al.
Publicado: (2026)
por: Zhang, Zhanqi, et al.
Publicado: (2026)
Kernelized Advantage Estimation: From Nonparametric Statistics to LLM Reasoning
por: Gong, Shijin, et al.
Publicado: (2026)
por: Gong, Shijin, et al.
Publicado: (2026)
Contextual Bandits with Packing and Covering Constraints: A Modular Lagrangian Approach via Regression
por: Slivkins, Aleksandrs, et al.
Publicado: (2022)
por: Slivkins, Aleksandrs, et al.
Publicado: (2022)
Fairness-aware kidney exchange and kidney paired donation
por: Zhang, Mingrui, et al.
Publicado: (2025)
por: Zhang, Mingrui, et al.
Publicado: (2025)
A Data Envelopment Analysis Approach for Assessing Fairness in Resource Allocation: Application to Kidney Exchange Programs
por: Kaazempur-Mofrad, Ali, et al.
Publicado: (2024)
por: Kaazempur-Mofrad, Ali, et al.
Publicado: (2024)
The Gossiping Insert-Eliminate Algorithm for Multi-Agent Bandits
por: Chawla, Ronshee, et al.
Publicado: (2020)
por: Chawla, Ronshee, et al.
Publicado: (2020)
AutoMCU: Feasibility-First MCU Neural Network Customization via LLM-based Multi-Agent Systems
por: Dai, Penglin, et al.
Publicado: (2026)
por: Dai, Penglin, et al.
Publicado: (2026)
GR-Agent: Adaptive Graph Reasoning Agent under Incomplete Knowledge
por: Zhou, Dongzhuoran, et al.
Publicado: (2025)
por: Zhou, Dongzhuoran, et al.
Publicado: (2025)
Efficient Rectification of Neuro-Symbolic Reasoning Inconsistencies by Abductive Reflection
por: Hu, Wen-Chao, et al.
Publicado: (2024)
por: Hu, Wen-Chao, et al.
Publicado: (2024)
Probabilistic Hash Embeddings for Online Learning of Categorical Features
por: Li, Aodong, et al.
Publicado: (2025)
por: Li, Aodong, et al.
Publicado: (2025)
SAT: Sequential Agent Tuning for Coordinator Free Plug and Play Multi-LLM Training with Monotonic Improvement Guarantees
por: Xie, Yi, et al.
Publicado: (2026)
por: Xie, Yi, et al.
Publicado: (2026)
From Debate to Equilibrium: Belief-Driven Multi-Agent LLM Reasoning via Bayesian Nash Equilibrium
por: Yi, Xie, et al.
Publicado: (2025)
por: Yi, Xie, et al.
Publicado: (2025)
AdaDetectGPT: Adaptive Detection of LLM-Generated Text with Statistical Guarantees
por: Zhou, Hongyi, et al.
Publicado: (2025)
por: Zhou, Hongyi, et al.
Publicado: (2025)
On the Provable Performance Guarantee of Efficient Reasoning Models
por: Zeng, Hao, et al.
Publicado: (2025)
por: Zeng, Hao, et al.
Publicado: (2025)
Learning With Multi-Group Guarantees For Clusterable Subpopulations
por: Dai, Jessica, et al.
Publicado: (2024)
por: Dai, Jessica, et al.
Publicado: (2024)
Generalization Guarantees of Gradient Descent for Multi-Layer Neural Networks
por: Wang, Puyu, et al.
Publicado: (2023)
por: Wang, Puyu, et al.
Publicado: (2023)
MALT: Improving Reasoning with Multi-Agent LLM Training
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
por: Motwani, Sumeet Ramesh, et al.
Publicado: (2024)
Towards Trustworthy Multimodal Moderation via Policy-Aligned Reasoning and Hierarchical Labeling
por: Li, Anqi, et al.
Publicado: (2025)
por: Li, Anqi, et al.
Publicado: (2025)
Simulation-Based Benchmarking of Reinforcement Learning Agents for Personalized Retail Promotions
por: Xia, Yu, et al.
Publicado: (2024)
por: Xia, Yu, et al.
Publicado: (2024)
Ejemplares similares
-
Prediction-Powered Conditional Inference
por: Sui, Yang, et al.
Publicado: (2026) -
Incentivizing Truthful Language Models via Peer Elicitation Games
por: Chen, Baiting, et al.
Publicado: (2025) -
Detecting LLM-Generated Text with Performance Guarantees
por: Zhou, Hongyi, et al.
Publicado: (2026) -
Learn then Decide: A Learning Approach for Designing Data Marketplaces
por: Gao, Yingqi, et al.
Publicado: (2025) -
Common-agency Games for Multi-Objective Test-Time Alignment
por: Chen, Baiting, et al.
Publicado: (2026)