Every Response Counts: Quantifying Uncertainty of LLM-based Multi-Agent Systems through Tensor Decomposition
Fuente:
arXiv
Guardado en:
| Autores principales: | Chen, Tiejin, Yao, Huaiyuan, Chen, Jia, Papalexakis, Evangelos E., Wei, Hua |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Uncertainty Quantification of Large Language Models through Multi-Dimensional Responses
por: Chen, Tiejin, et al.
Publicado: (2025)
por: Chen, Tiejin, et al.
Publicado: (2025)
Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders
por: Patel, Het, et al.
Publicado: (2026)
por: Patel, Het, et al.
Publicado: (2026)
Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks
por: Patel, Het, et al.
Publicado: (2025)
por: Patel, Het, et al.
Publicado: (2025)
LLM Uncertainty Quantification through Directional Entailment Graph and Claim Level Response Augmentation
por: Da, Longchao, et al.
Publicado: (2024)
por: Da, Longchao, et al.
Publicado: (2024)
Instructional Agents: Reducing Teaching Faculty Workload through Multi-Agent Instructional Design
por: Yao, Huaiyuan, et al.
Publicado: (2025)
por: Yao, Huaiyuan, et al.
Publicado: (2025)
SamBaTen: Sampling-based Batch Incremental Tensor Decomposition
por: Gujral, Ekta, et al.
Publicado: (2017)
por: Gujral, Ekta, et al.
Publicado: (2017)
GPT-generated Text Detection: Benchmark Dataset and Tensor-based Detection Method
por: Qazi, Zubair, et al.
Publicado: (2024)
por: Qazi, Zubair, et al.
Publicado: (2024)
Position: Uncertainty Quantification in LLMs is Just Unsupervised Clustering
por: Chen, Tiejin, et al.
Publicado: (2026)
por: Chen, Tiejin, et al.
Publicado: (2026)
CoCoTen: Detecting Adversarial Inputs to Large Language Models through Latent Space Features of Contextual Co-occurrence Tensors
por: Kadali, Sri Durga Sai Sowmya, et al.
Publicado: (2025)
por: Kadali, Sri Durga Sai Sowmya, et al.
Publicado: (2025)
Beyond ROUGE: N-Gram Subspace Features for LLM Hallucination Detection
por: Li, Jerry, et al.
Publicado: (2025)
por: Li, Jerry, et al.
Publicado: (2025)
From Rubrics to Reliable Scores: Evidence-Grounded Text Evaluation with LLM Judges
por: Hong, Yihan, et al.
Publicado: (2026)
por: Hong, Yihan, et al.
Publicado: (2026)
Zer0-Jack: A Memory-efficient Gradient-based Jailbreaking Method for Black-box Multi-modal Large Language Models
por: Chen, Tiejin, et al.
Publicado: (2024)
por: Chen, Tiejin, et al.
Publicado: (2024)
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
por: Chen, Tiejin, et al.
Publicado: (2026)
por: Chen, Tiejin, et al.
Publicado: (2026)
Uncertainty Regularized Evidential Regression
por: Ye, Kai, et al.
Publicado: (2024)
por: Ye, Kai, et al.
Publicado: (2024)
MIRIX: Multi-Agent Memory System for LLM-Based Agents
por: Wang, Yu, et al.
Publicado: (2025)
por: Wang, Yu, et al.
Publicado: (2025)
Uncertainty Quantification and Confidence Calibration in Large Language Models: A Survey
por: Liu, Xiaoou, et al.
Publicado: (2025)
por: Liu, Xiaoou, et al.
Publicado: (2025)
Make Every Draft Count: Hidden State based Speculative Decoding
por: Chen, Yuetao, et al.
Publicado: (2026)
por: Chen, Yuetao, et al.
Publicado: (2026)
Every FLOP Counts: Scaling a 300B Mixture-of-Experts LING LLM without Premium GPUs
por: Ling Team, et al.
Publicado: (2025)
por: Ling Team, et al.
Publicado: (2025)
Tensor Completion for Surrogate Modeling of Material Property Prediction
por: Pakala, Shaan, et al.
Publicado: (2025)
por: Pakala, Shaan, et al.
Publicado: (2025)
Creativity in LLM-based Multi-Agent Systems: A Survey
por: Lin, Yi-Cheng, et al.
Publicado: (2025)
por: Lin, Yi-Cheng, et al.
Publicado: (2025)
Make Every Penny Count: Difficulty-Adaptive Self-Consistency for Cost-Efficient Reasoning
por: Wang, Xinglin, et al.
Publicado: (2024)
por: Wang, Xinglin, et al.
Publicado: (2024)
Multi-View Spectral Clustering for Graphs with Multiple View Structures
por: Tsitsikas, Yorgos, et al.
Publicado: (2025)
por: Tsitsikas, Yorgos, et al.
Publicado: (2025)
LangMARL: Natural Language Multi-Agent Reinforcement Learning
por: Yao, Huaiyuan, et al.
Publicado: (2026)
por: Yao, Huaiyuan, et al.
Publicado: (2026)
Optima: Optimizing Effectiveness and Efficiency for LLM-Based Multi-Agent System
por: Chen, Weize, et al.
Publicado: (2024)
por: Chen, Weize, et al.
Publicado: (2024)
Every Token Counts: Generalizing 16M Ultra-Long Context in Large Language Models
por: Hu, Xiang, et al.
Publicado: (2025)
por: Hu, Xiang, et al.
Publicado: (2025)
Every Step Counts: Decoding Trajectories as Authorship Fingerprints of dLLMs
por: Li, Qi, et al.
Publicado: (2025)
por: Li, Qi, et al.
Publicado: (2025)
MASLab: A Unified and Comprehensive Codebase for LLM-based Multi-Agent Systems
por: Ye, Rui, et al.
Publicado: (2025)
por: Ye, Rui, et al.
Publicado: (2025)
Efficiently Generating Multidimensional Calorimeter Data with Tensor Decomposition Parameterization
por: Goulart, Paimon, et al.
Publicado: (2025)
por: Goulart, Paimon, et al.
Publicado: (2025)
Watch Every Step! LLM Agent Learning via Iterative Step-Level Process Refinement
por: Xiong, Weimin, et al.
Publicado: (2024)
por: Xiong, Weimin, et al.
Publicado: (2024)
TRAWL: Tensor Reduced and Approximated Weights for Large Language Models
por: Luo, Yiran, et al.
Publicado: (2024)
por: Luo, Yiran, et al.
Publicado: (2024)
Transforming Behavioral Neuroscience Discovery with In-Context Learning and AI-Enhanced Tensor Methods
por: Goulart, Paimon, et al.
Publicado: (2026)
por: Goulart, Paimon, et al.
Publicado: (2026)
Researchy Questions: A Dataset of Multi-Perspective, Decompositional Questions for LLM Web Agents
por: Rosset, Corby, et al.
Publicado: (2024)
por: Rosset, Corby, et al.
Publicado: (2024)
MemAgent: Reshaping Long-Context LLM with Multi-Conv RL-based Memory Agent
por: Yu, Hongli, et al.
Publicado: (2025)
por: Yu, Hongli, et al.
Publicado: (2025)
AgentGym-RL: Training LLM Agents for Long-Horizon Decision Making through Multi-Turn Reinforcement Learning
por: Xi, Zhiheng, et al.
Publicado: (2025)
por: Xi, Zhiheng, et al.
Publicado: (2025)
Understanding LLM Performance Degradation in Multi-Instance Processing: The Roles of Instance Count and Context Length
por: Chen, Jingxuan, et al.
Publicado: (2026)
por: Chen, Jingxuan, et al.
Publicado: (2026)
Every Character Counts: From Vulnerability to Defense in Phishing Detection
por: Chiper, Maria, et al.
Publicado: (2025)
por: Chiper, Maria, et al.
Publicado: (2025)
Data Compressibility Quantifies LLM Memorization
por: Huang, Yizhan, et al.
Publicado: (2025)
por: Huang, Yizhan, et al.
Publicado: (2025)
When Every Token Counts: Optimal Segmentation for Low-Resource Language Models
por: Raj, Bharath, et al.
Publicado: (2024)
por: Raj, Bharath, et al.
Publicado: (2024)
Modeling LLM Agent Reviewer Dynamics in Elo-Ranked Review System
por: Huang, Hsiang-Wei, et al.
Publicado: (2026)
por: Huang, Hsiang-Wei, et al.
Publicado: (2026)
Structured Uncertainty guided Clarification for LLM Agents
por: Suri, Manan, et al.
Publicado: (2025)
por: Suri, Manan, et al.
Publicado: (2025)
Ejemplares similares
-
Uncertainty Quantification of Large Language Models through Multi-Dimensional Responses
por: Chen, Tiejin, et al.
Publicado: (2025) -
Are LLM Uncertainty and Correctness Encoded by the Same Features? A Functional Dissociation via Sparse Autoencoders
por: Patel, Het, et al.
Publicado: (2026) -
Robust Vision-Language Models via Tensor Decomposition: A Defense Against Adversarial Attacks
por: Patel, Het, et al.
Publicado: (2025) -
LLM Uncertainty Quantification through Directional Entailment Graph and Claim Level Response Augmentation
por: Da, Longchao, et al.
Publicado: (2024) -
Instructional Agents: Reducing Teaching Faculty Workload through Multi-Agent Instructional Design
por: Yao, Huaiyuan, et al.
Publicado: (2025)