Multi-LLM Adaptive Conformal Inference for Reliable LLM Responses
Fuente:
arXiv
Guardado en:
| Autores principales: | Noh, Kangjun, Lee, Seongchan, Kim, Ilmun, Song, Kyungwoo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
General Frameworks for Conditional Two-Sample Testing
por: Lee, Seongchan, et al.
Publicado: (2024)
por: Lee, Seongchan, et al.
Publicado: (2024)
Uncertainty-driven Embedding Convolution
por: Lim, Sungjun, et al.
Publicado: (2025)
por: Lim, Sungjun, et al.
Publicado: (2025)
Spurious Correlation-Aware Embedding Regularization for Worst-Group Robustness
por: Park, Subeen, et al.
Publicado: (2025)
por: Park, Subeen, et al.
Publicado: (2025)
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
por: Chen, Tiejin, et al.
Publicado: (2026)
por: Chen, Tiejin, et al.
Publicado: (2026)
Margin-Adaptive Confidence Ranking for Reliable LLM Judgement
por: Jin, Gaojie, et al.
Publicado: (2026)
por: Jin, Gaojie, et al.
Publicado: (2026)
Reliable Self-Harm Risk Screening via Adaptive Multi-Agent LLM Systems
por: Karnam, Meghana, et al.
Publicado: (2026)
por: Karnam, Meghana, et al.
Publicado: (2026)
GDFlow: Anomaly Detection with NCDE-based Normalizing Flow for Advanced Driver Assistance System
por: Lee, Kangjun, et al.
Publicado: (2024)
por: Lee, Kangjun, et al.
Publicado: (2024)
RAP: Runtime Adaptive Pruning for LLM Inference
por: Liu, Huanrong, et al.
Publicado: (2025)
por: Liu, Huanrong, et al.
Publicado: (2025)
Diagnosing LLM Judge Reliability: Conformal Prediction Sets and Transitivity Violations
por: Gupta, Manan, et al.
Publicado: (2026)
por: Gupta, Manan, et al.
Publicado: (2026)
MIDUS: Memory-Infused Depth Up-Scaling
por: Kim, Taero, et al.
Publicado: (2025)
por: Kim, Taero, et al.
Publicado: (2025)
Semi-Supervised Preference Optimization with Limited Feedback
por: Lee, Seonggyun, et al.
Publicado: (2025)
por: Lee, Seonggyun, et al.
Publicado: (2025)
Conformal Constrained Policy Optimization for Cost-Effective LLM Agents
por: Si, Wenwen, et al.
Publicado: (2025)
por: Si, Wenwen, et al.
Publicado: (2025)
AdaBlock-dLLM: Semantic-Aware Diffusion LLM Inference via Adaptive Block Size
por: Lu, Guanxi, et al.
Publicado: (2025)
por: Lu, Guanxi, et al.
Publicado: (2025)
Eigen-Value: Efficient Domain-Robust Data Valuation via Eigenvalue-Based Approach
por: Choi, Youngjun, et al.
Publicado: (2025)
por: Choi, Youngjun, et al.
Publicado: (2025)
Perturb-and-Compare Approach for Detecting Out-of-Distribution Samples in Constrained Access Environments
por: Lee, Heeyoung, et al.
Publicado: (2024)
por: Lee, Heeyoung, et al.
Publicado: (2024)
LBC: Language-Based-Classifier for Out-Of-Variable Generalization
por: Noh, Kangjun, et al.
Publicado: (2024)
por: Noh, Kangjun, et al.
Publicado: (2024)
Towards Reliable LLM Evaluation: Correcting the Winner's Curse in Adaptive Benchmarking
por: Xu, Yang, et al.
Publicado: (2026)
por: Xu, Yang, et al.
Publicado: (2026)
CSAttention: Centroid-Scoring Attention for Accelerating LLM Inference
por: Song, Chuxu, et al.
Publicado: (2026)
por: Song, Chuxu, et al.
Publicado: (2026)
Sufficient Invariant Learning for Distribution Shift
por: Kim, Taero, et al.
Publicado: (2022)
por: Kim, Taero, et al.
Publicado: (2022)
RAMP: Reinforcement Adaptive Mixed Precision Quantization for Efficient On Device LLM Inference
por: Gautam, Arpit Singh, et al.
Publicado: (2026)
por: Gautam, Arpit Singh, et al.
Publicado: (2026)
CATS: Cascaded Adaptive Tree Speculation for Memory-Limited LLM Inference Acceleration
por: Han, Yuning, et al.
Publicado: (2026)
por: Han, Yuning, et al.
Publicado: (2026)
A Technical Exploration of Causal Inference with Hybrid LLM Synthetic Data
por: Kim, Dana, et al.
Publicado: (2025)
por: Kim, Dana, et al.
Publicado: (2025)
FastMTP: Accelerating LLM Inference with Enhanced Multi-Token Prediction
por: Cai, Yuxuan, et al.
Publicado: (2025)
por: Cai, Yuxuan, et al.
Publicado: (2025)
TIDE: Temporal Incremental Draft Engine for Self-Improving LLM Inference
por: Park, Jiyoung, et al.
Publicado: (2026)
por: Park, Jiyoung, et al.
Publicado: (2026)
Reliability Auditing for Downstream LLM tasks in Psychiatry: LLM-Generated Hospitalization Risk Scores
por: Panda, Shevya, et al.
Publicado: (2026)
por: Panda, Shevya, et al.
Publicado: (2026)
eFedLLM: Efficient LLM Inference Based on Federated Learning
por: Ding, Shengwen, et al.
Publicado: (2024)
por: Ding, Shengwen, et al.
Publicado: (2024)
WebLLM: A High-Performance In-Browser LLM Inference Engine
por: Ruan, Charlie F., et al.
Publicado: (2024)
por: Ruan, Charlie F., et al.
Publicado: (2024)
Adaptively Robust LLM Inference Optimization under Prediction Uncertainty
por: Chen, Zixi, et al.
Publicado: (2025)
por: Chen, Zixi, et al.
Publicado: (2025)
Multi-Task GRPO: Reliable LLM Reasoning Across Tasks
por: Ramesh, Shyam Sundhar, et al.
Publicado: (2026)
por: Ramesh, Shyam Sundhar, et al.
Publicado: (2026)
CARE: Confounder-Aware Aggregation for Reliable LLM Evaluation
por: Zhao, Jitian, et al.
Publicado: (2026)
por: Zhao, Jitian, et al.
Publicado: (2026)
Energy-Efficient Wireless LLM Inference via Uncertainty and Importance-Aware Speculative Decoding
por: Park, Jihoon, et al.
Publicado: (2025)
por: Park, Jihoon, et al.
Publicado: (2025)
Adaptive Layer Selection for Layer-Wise Token Pruning in LLM Inference
por: Taniguchi, Rei, et al.
Publicado: (2026)
por: Taniguchi, Rei, et al.
Publicado: (2026)
Structure-Adaptive Conformal Inference for Large-Scale Out-of-Distribution Testing
por: Sun, Rongyi, et al.
Publicado: (2026)
por: Sun, Rongyi, et al.
Publicado: (2026)
Interactive Critique-Revision Training for Reliable Structured LLM Generation
por: Yu, Fei Xu, et al.
Publicado: (2026)
por: Yu, Fei Xu, et al.
Publicado: (2026)
SDQ: Sparse Decomposed Quantization for LLM Inference
por: Jeong, Geonhwa, et al.
Publicado: (2024)
por: Jeong, Geonhwa, et al.
Publicado: (2024)
Recursive Speculative Decoding: Accelerating LLM Inference via Sampling Without Replacement
por: Jeon, Wonseok, et al.
Publicado: (2024)
por: Jeon, Wonseok, et al.
Publicado: (2024)
SUN: Shared Use of Next-token Prediction for Efficient Multi-LLM Disaggregated Serving
por: Woo, Sunghyeon, et al.
Publicado: (2026)
por: Woo, Sunghyeon, et al.
Publicado: (2026)
Dynamic and Adaptive Feature Generation with LLM
por: Zhang, Xinhao, et al.
Publicado: (2024)
por: Zhang, Xinhao, et al.
Publicado: (2024)
Training-free LLM Verification via Recycling Few-shot Examples
por: Lee, Dongseok, et al.
Publicado: (2025)
por: Lee, Dongseok, et al.
Publicado: (2025)
Adaptive Inference-Time Scaling via Cyclic Diffusion Search
por: Lee, Gyubin, et al.
Publicado: (2025)
por: Lee, Gyubin, et al.
Publicado: (2025)
Ejemplares similares
-
General Frameworks for Conditional Two-Sample Testing
por: Lee, Seongchan, et al.
Publicado: (2024) -
Uncertainty-driven Embedding Convolution
por: Lim, Sungjun, et al.
Publicado: (2025) -
Spurious Correlation-Aware Embedding Regularization for Worst-Group Robustness
por: Park, Subeen, et al.
Publicado: (2025) -
Conformal Feedback Alignment: Quantifying Answer-Level Reliability for Robust LLM Alignment
por: Chen, Tiejin, et al.
Publicado: (2026) -
Margin-Adaptive Confidence Ranking for Reliable LLM Judgement
por: Jin, Gaojie, et al.
Publicado: (2026)