Efficient MAP Estimation of LLM Judgment Performance with Prior Transfer
Fuente:
arXiv
Saved in:
| Main Authors: | Qu, Huaizhi, Choi, Inyoung, Tan, Zhen, Wang, Song, Yun, Sukwon, Long, Qi, Siddiqui, Faizan, Lee, Kwonjoon, Chen, Tianlong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multi-Agent Debate for LLM Judges with Adaptive Stability Detection
by: Hu, Tianyu, et al.
Published: (2025)
by: Hu, Tianyu, et al.
Published: (2025)
I2MoE: Interpretable Multimodal Interaction-aware Mixture-of-Experts
by: Xin, Jiayi, et al.
Published: (2025)
by: Xin, Jiayi, et al.
Published: (2025)
Flex-MoE: Modeling Arbitrary Modality Combination via the Flexible Mixture-of-Experts
by: Yun, Sukwon, et al.
Published: (2024)
by: Yun, Sukwon, et al.
Published: (2024)
$\textit{Agents Under Siege}$: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025)
DOGe: Defensive Output Generation for LLM Protection Against Knowledge Distillation
by: Li, Pingzhi, et al.
Published: (2025)
by: Li, Pingzhi, et al.
Published: (2025)
GEM: 3D Gaussian Splatting for Efficient and Accurate Cryo-EM Reconstruction
by: Qu, Huaizhi, et al.
Published: (2025)
by: Qu, Huaizhi, et al.
Published: (2025)
Mew: Multiplexed Immunofluorescence Image Analysis through an Efficient Multiplex Network
by: Yun, Sukwon, et al.
Published: (2024)
by: Yun, Sukwon, et al.
Published: (2024)
Task-Aware Resolution Optimization for Visual Large Language Models
by: Luo, Weiqing, et al.
Published: (2025)
by: Luo, Weiqing, et al.
Published: (2025)
Spatial Coordinates as a Cell Language: A Multi-Sentence Framework for Imaging Mass Cytometry Analysis
by: Chen, Chi-Jane, et al.
Published: (2025)
by: Chen, Chi-Jane, et al.
Published: (2025)
M2D2M: Multi-Motion Generation from Text with Discrete Diffusion Models
by: Chi, Seunggeun, et al.
Published: (2024)
by: Chi, Seunggeun, et al.
Published: (2024)
Graph-of-Agents: A Graph-based Framework for Multi-Agent LLM Collaboration
by: Yun, Sukwon, et al.
Published: (2026)
by: Yun, Sukwon, et al.
Published: (2026)
PETS: A Principled Framework Towards Optimal Trajectory Allocation for Efficient Test-Time Self-Consistency
by: Liu, Zhangyi, et al.
Published: (2026)
by: Liu, Zhangyi, et al.
Published: (2026)
BrainMAP: Learning Multiple Activation Pathways in Brain Networks
by: Wang, Song, et al.
Published: (2024)
by: Wang, Song, et al.
Published: (2024)
Oldie but Goodie: Re-illuminating Label Propagation on Graphs with Partially Observed Features
by: Yun, Sukwon, et al.
Published: (2025)
by: Yun, Sukwon, et al.
Published: (2025)
Skill-Based Mixture-of-Experts: Adaptive Routing for Heterogeneous Reasoning via Inferred Skills
by: Chen, Justin Chih-Yao, et al.
Published: (2025)
by: Chen, Justin Chih-Yao, et al.
Published: (2025)
Harnessing Your DRAM and SSD for Sustainable and Accessible LLM Inference with Mixed-Precision and Multi-level Caching
by: Peng, Jie, et al.
Published: (2024)
by: Peng, Jie, et al.
Published: (2024)
PortLLM: Personalizing Evolving Large Language Models with Training-Free and Portable Model Patches
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2024)
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2024)
MAP: Multi-user Personalization with Collaborative LLM-powered Agents
by: Lee, Christine, et al.
Published: (2025)
by: Lee, Christine, et al.
Published: (2025)
Metacognitive Self-Correction for Multi-Agent System via Prototype-Guided Next-Execution Reconstruction
by: Shen, Xu, et al.
Published: (2025)
by: Shen, Xu, et al.
Published: (2025)
Tuning-Free Accountable Intervention for LLM Deployment -- A Metacognitive Approach
by: Tan, Zhen, et al.
Published: (2024)
by: Tan, Zhen, et al.
Published: (2024)
Understanding the Role of Hallucination in Reinforcement Post-Training of Multimodal Reasoning Models
by: Zhang, Gengwei, et al.
Published: (2026)
by: Zhang, Gengwei, et al.
Published: (2026)
EditCast3D: Single-Frame-Guided 3D Editing with Video Propagation and View Selection
by: Qu, Huaizhi, et al.
Published: (2025)
by: Qu, Huaizhi, et al.
Published: (2025)
Cut the Crap: An Economical Communication Pipeline for LLM-based Multi-Agent Systems
by: Zhang, Guibin, et al.
Published: (2024)
by: Zhang, Guibin, et al.
Published: (2024)
Advanced Materials for CO2 Capture: A Critical Review of Emerging Adsorbents and Technologies
by: Muhammad Faizan, et al.
Published: (2025)
by: Muhammad Faizan, et al.
Published: (2025)
FAST-EQA: Efficient Embodied Question Answering with Global and Local Region Relevancy
by: Zhang, Haochen, et al.
Published: (2026)
by: Zhang, Haochen, et al.
Published: (2026)
T-MAP: Red-Teaming LLM Agents with Trajectory-aware Evolutionary Search
by: Lee, Hyomin, et al.
Published: (2026)
by: Lee, Hyomin, et al.
Published: (2026)
EQA-RM: A Generative Embodied Reward Model with Test-time Scaling
by: Chen, Yuhang, et al.
Published: (2025)
by: Chen, Yuhang, et al.
Published: (2025)
Beyond Redundancy: Diverse and Specialized Multi-Expert Sparse Autoencoder
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
SPATIA: Multimodal Generation and Prediction of Spatial Cell Phenotypes
by: Kong, Zhenglun, et al.
Published: (2025)
by: Kong, Zhenglun, et al.
Published: (2025)
Association of GLP ‐1 Receptor Agonists With Hepatic Decompensation in the All of Us Research Program
by: Inyoung Hwang, et al.
Published: (2026)
by: Inyoung Hwang, et al.
Published: (2026)
Noisy MRI Reconstruction via MAP Estimation with an Implicit Deep-Denoiser Prior
by: Janjušević, Nikola, et al.
Published: (2025)
by: Janjušević, Nikola, et al.
Published: (2025)
Who's Your Judge? On the Detectability of LLM-Generated Judgments
by: Li, Dawei, et al.
Published: (2025)
by: Li, Dawei, et al.
Published: (2025)
Spread Preference Annotation: Direct Preference Judgment for Efficient LLM Alignment
by: Kim, Dongyoung, et al.
Published: (2024)
by: Kim, Dongyoung, et al.
Published: (2024)
The Quest for Efficient Reasoning: A Data-Centric Benchmark to CoT Distillation
by: Zhang, Ruichen, et al.
Published: (2025)
by: Zhang, Ruichen, et al.
Published: (2025)
Measuring Real-World Prompt Injection Attacks in LLM-based Resume Screening
by: Zhang, Mohan, et al.
Published: (2026)
by: Zhang, Mohan, et al.
Published: (2026)
DALK: Dynamic Co-Augmentation of LLMs and KG to answer Alzheimer's Disease Questions with Scientific Literature
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
GRNFormer: A Biologically-Guided Framework for Integrating Gene Regulatory Networks into RNA Foundation Models
by: Qiu, Mufan, et al.
Published: (2025)
by: Qiu, Mufan, et al.
Published: (2025)
Improving LLM-as-a-Judge Inference with the Judgment Distribution
by: Wang, Victor, et al.
Published: (2025)
by: Wang, Victor, et al.
Published: (2025)
From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
Pinched self-dual Weyl curvature on Einstein four-manifolds
by: Kim, Inyoung
Published: (2025)
by: Kim, Inyoung
Published: (2025)
Similar Items
-
Multi-Agent Debate for LLM Judges with Adaptive Stability Detection
by: Hu, Tianyu, et al.
Published: (2025) -
I2MoE: Interpretable Multimodal Interaction-aware Mixture-of-Experts
by: Xin, Jiayi, et al.
Published: (2025) -
Flex-MoE: Modeling Arbitrary Modality Combination via the Flexible Mixture-of-Experts
by: Yun, Sukwon, et al.
Published: (2024) -
$\textit{Agents Under Siege}$: Breaking Pragmatic Multi-Agent LLM Systems with Optimized Prompt Attacks
by: Khan, Rana Muhammad Shahroz, et al.
Published: (2025) -
DOGe: Defensive Output Generation for LLM Protection Against Knowledge Distillation
by: Li, Pingzhi, et al.
Published: (2025)