From Generation to Judgment: Opportunities and Challenges of LLM-as-a-judge
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Dawei, Jiang, Bohan, Huang, Liangjie, Beigi, Alimohammad, Zhao, Chengshuai, Tan, Zhen, Bhattacharjee, Amrita, Jiang, Yuxuan, Chen, Canyu, Wu, Tianhao, Shu, Kai, Cheng, Lu, Liu, Huan |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
by: Beigi, Alimohammad, et al.
Published: (2024)
by: Beigi, Alimohammad, et al.
Published: (2024)
Model Attribution in LLM-Generated Disinformation: A Domain Generalization Approach with Supervised Contrastive Learning
by: Beigi, Alimohammad, et al.
Published: (2024)
by: Beigi, Alimohammad, et al.
Published: (2024)
Large Language Models for Data Annotation and Synthesis: A Survey
by: Tan, Zhen, et al.
Published: (2024)
by: Tan, Zhen, et al.
Published: (2024)
Who's Your Judge? On the Detectability of LLM-Generated Judgments
by: Li, Dawei, et al.
Published: (2025)
by: Li, Dawei, et al.
Published: (2025)
CAMO: Causality-Guided Adversarial Multimodal Domain Generalization for Crisis Classification
by: Ma, Pingchuan, et al.
Published: (2025)
by: Ma, Pingchuan, et al.
Published: (2025)
Are Today's LLMs Ready to Explain Well-Being Concepts?
by: Jiang, Bohan, et al.
Published: (2025)
by: Jiang, Bohan, et al.
Published: (2025)
Catching Chameleons: Detecting Evolving Disinformation Generated using Large Language Models
by: Jiang, Bohan, et al.
Published: (2024)
by: Jiang, Bohan, et al.
Published: (2024)
Preference Leakage: A Contamination Problem in LLM-as-a-judge
by: Li, Dawei, et al.
Published: (2025)
by: Li, Dawei, et al.
Published: (2025)
Is Chain-of-Thought Reasoning of LLMs a Mirage? A Data Distribution Lens
by: Zhao, Chengshuai, et al.
Published: (2025)
by: Zhao, Chengshuai, et al.
Published: (2025)
An Interventional Approach to Real-Time Disaster Assessment via Causal Attribution
by: Vishnubhatla, Saketh, et al.
Published: (2025)
by: Vishnubhatla, Saketh, et al.
Published: (2025)
Can LLM-Generated Misinformation Be Detected?
by: Chen, Canyu, et al.
Published: (2023)
by: Chen, Canyu, et al.
Published: (2023)
Beyond Accuracy: The Role of Calibration in Self-Improving Large Language Models
by: Huang, Liangjie, et al.
Published: (2025)
by: Huang, Liangjie, et al.
Published: (2025)
To See is Not to Learn: Protecting Multimodal Data from Unauthorized Fine-Tuning of Large Vision-Language Model
by: Zhao, Chengshuai, et al.
Published: (2026)
by: Zhao, Chengshuai, et al.
Published: (2026)
Media Bias Matters: Understanding the Impact of Politically Biased News on Vaccine Attitudes in Social Media
by: Jiang, Bohan, et al.
Published: (2024)
by: Jiang, Bohan, et al.
Published: (2024)
Sentiment and Social Signals in the Climate Crisis: A Survey on Analyzing Social Media Responses to Extreme Weather Events
by: Shaeri, Pouya, et al.
Published: (2025)
by: Shaeri, Pouya, et al.
Published: (2025)
Curiosity-Driven LLM-as-a-judge for Personalized Creative Judgment
by: Kumar, Vanya Bannihatti, et al.
Published: (2025)
by: Kumar, Vanya Bannihatti, et al.
Published: (2025)
Harnessing Artificial Intelligence to Combat Online Hate: Exploring the Challenges and Opportunities of Large Language Models in Hate Speech Detection
by: Kumarage, Tharindu, et al.
Published: (2024)
by: Kumarage, Tharindu, et al.
Published: (2024)
Authorship Attribution in the Era of LLMs: Problems, Methodologies, and Challenges
by: Huang, Baixiang, et al.
Published: (2024)
by: Huang, Baixiang, et al.
Published: (2024)
From Calculation to Adjudication: Examining LLM judges on Mathematical Reasoning Tasks
by: Stephan, Andreas, et al.
Published: (2024)
by: Stephan, Andreas, et al.
Published: (2024)
Towards LLM-guided Causal Explainability for Black-box Text Classifiers
by: Bhattacharjee, Amrita, et al.
Published: (2023)
by: Bhattacharjee, Amrita, et al.
Published: (2023)
Zero-shot LLM-guided Counterfactual Generation: A Case Study on NLP Model Evaluation
by: Bhattacharjee, Amrita, et al.
Published: (2024)
by: Bhattacharjee, Amrita, et al.
Published: (2024)
Tri-Accel: Curvature-Aware Precision-Adaptive and Memory-Elastic Optimization for Efficient GPU Usage
by: Sheibanian, Mohsen, et al.
Published: (2025)
by: Sheibanian, Mohsen, et al.
Published: (2025)
Fediverse Sharing: Cross-Platform Interaction Dynamics between Threads and Mastodon Users
by: Jeong, Ujun, et al.
Published: (2025)
by: Jeong, Ujun, et al.
Published: (2025)
ResumeFlow: An LLM-facilitated Pipeline for Personalized Resume Generation and Refinement
by: Zinjad, Saurabh Bhausaheb, et al.
Published: (2024)
by: Zinjad, Saurabh Bhausaheb, et al.
Published: (2024)
Do LLMs Understand Ambiguity in Text? A Case Study in Open-world Question Answering
by: Keluskar, Aryan, et al.
Published: (2024)
by: Keluskar, Aryan, et al.
Published: (2024)
Can Large Language Models Identify Authorship?
by: Huang, Baixiang, et al.
Published: (2024)
by: Huang, Baixiang, et al.
Published: (2024)
BlueTempNet: A Temporal Multi-network Dataset of Social Interactions in Bluesky Social
by: Jeong, Ujun, et al.
Published: (2024)
by: Jeong, Ujun, et al.
Published: (2024)
Assessing the Impact of Conspiracy Theories Using Large Language Models
by: Jiang, Bohan, et al.
Published: (2024)
by: Jiang, Bohan, et al.
Published: (2024)
Exploring Large Language Models for Feature Selection: A Data-centric Perspective
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
Towards Interpretable Hate Speech Detection using Large Language Model-extracted Rationales
by: Nirmal, Ayushi, et al.
Published: (2024)
by: Nirmal, Ayushi, et al.
Published: (2024)
Adversarial Text Purification: A Large Language Model Approach for Defense
by: Moraffah, Raha, et al.
Published: (2024)
by: Moraffah, Raha, et al.
Published: (2024)
EAGLE: A Domain Generalization Framework for AI-generated Text Detection
by: Bhattacharjee, Amrita, et al.
Published: (2024)
by: Bhattacharjee, Amrita, et al.
Published: (2024)
Probing to Refine: Reinforcement Distillation of LLMs via Explanatory Inversion
by: Tan, Zhen, et al.
Published: (2026)
by: Tan, Zhen, et al.
Published: (2026)
Science Opportunities of Wet Extreme Mass-Ratio Inspirals
by: Lyu, Zhenwei, et al.
Published: (2024)
by: Lyu, Zhenwei, et al.
Published: (2024)
SCALE: Towards Collaborative Content Analysis in Social Science with Large Language Model Agents and Human Intervention
by: Zhao, Chengshuai, et al.
Published: (2025)
by: Zhao, Chengshuai, et al.
Published: (2025)
MetaGAD: Meta Representation Adaptation for Few-Shot Graph Anomaly Detection
by: Xu, Xiongxiao, et al.
Published: (2023)
by: Xu, Xiongxiao, et al.
Published: (2023)
DRP: Distilled Reasoning Pruning with Skill-aware Step Decomposition for Efficient Large Reasoning Models
by: Jiang, Yuxuan, et al.
Published: (2025)
by: Jiang, Yuxuan, et al.
Published: (2025)
Benchmarking LLMs' Judgments with No Gold Standard
by: Xu, Shengwei, et al.
Published: (2024)
by: Xu, Shengwei, et al.
Published: (2024)
Assessing On-the-Ground Disaster Impact Using Online Data Sources
by: Vishnubhatla, Saketh, et al.
Published: (2025)
by: Vishnubhatla, Saketh, et al.
Published: (2025)
Contextualization Distillation from Large Language Model for Knowledge Graph Completion
by: Li, Dawei, et al.
Published: (2024)
by: Li, Dawei, et al.
Published: (2024)
Similar Items
-
Can LLMs Improve Multimodal Fact-Checking by Asking Relevant Questions?
by: Beigi, Alimohammad, et al.
Published: (2024) -
Model Attribution in LLM-Generated Disinformation: A Domain Generalization Approach with Supervised Contrastive Learning
by: Beigi, Alimohammad, et al.
Published: (2024) -
Large Language Models for Data Annotation and Synthesis: A Survey
by: Tan, Zhen, et al.
Published: (2024) -
Who's Your Judge? On the Detectability of LLM-Generated Judgments
by: Li, Dawei, et al.
Published: (2025) -
CAMO: Causality-Guided Adversarial Multimodal Domain Generalization for Crisis Classification
by: Ma, Pingchuan, et al.
Published: (2025)