Who is a Better Matchmaker? Human vs. Algorithmic Judge Assignment in a High-Stakes Startup Competition
Fuente:
arXiv
Saved in:
| Main Authors: | Xi, Sarina, Pi, Orelia, Zhang, Miaomiao, Xiong, Becca, Lane, Jacqueline Ng, Shah, Nihar B. |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
FLAWS: A Benchmark for Error Identification and Localization in Scientific Papers
by: Xi, Sarina, et al.
Published: (2025)
by: Xi, Sarina, et al.
Published: (2025)
Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AI
by: Liu, Houjiang, et al.
Published: (2023)
by: Liu, Houjiang, et al.
Published: (2023)
Sima AIunty: Caste Audit in LLM-Driven Matchmaking
by: Naik, Atharva, et al.
Published: (2026)
by: Naik, Atharva, et al.
Published: (2026)
Who Judges the Judge? Evaluating LLM-as-a-Judge for French Medical open-ended QA
by: Belmadani, Ikram, et al.
Published: (2026)
by: Belmadani, Ikram, et al.
Published: (2026)
MR. Judge: Multimodal Reasoner as a Judge
by: Pi, Renjie, et al.
Published: (2025)
by: Pi, Renjie, et al.
Published: (2025)
AI vs. Human Judgment of Content Moderation: LLM-as-a-Judge and Ethics-Based Response Refusals
by: Pasch, Stefan
Published: (2025)
by: Pasch, Stefan
Published: (2025)
Personality-Enhanced Social Recommendations in SAMI: Exploring the Role of Personality Detection in Matchmaking
by: Harbison, Brittany, et al.
Published: (2025)
by: Harbison, Brittany, et al.
Published: (2025)
A Principled Approach to Randomized Selection under Uncertainty: Applications to Peer Review and Grant Funding
by: Goldberg, Alexander, et al.
Published: (2025)
by: Goldberg, Alexander, et al.
Published: (2025)
Human vs. Generative AI in Content Creation Competition: Symbiosis or Conflict?
by: Yao, Fan, et al.
Published: (2024)
by: Yao, Fan, et al.
Published: (2024)
A Gold Standard Dataset for the Reviewer Assignment Problem
by: Stelmakh, Ivan, et al.
Published: (2023)
by: Stelmakh, Ivan, et al.
Published: (2023)
Who Wrote This? Identifying Machine vs Human-Generated Text in Hausa
by: Sani, Babangida, et al.
Published: (2025)
by: Sani, Babangida, et al.
Published: (2025)
Text-to-Image Diffusion Models are Great Sketch-Photo Matchmakers
by: Koley, Subhadeep, et al.
Published: (2024)
by: Koley, Subhadeep, et al.
Published: (2024)
Raising the Stakes: Assessing the Influence of Stakes on User Reliance Behavior in Human-AI Decision-Making
by: Johnson, David S.
Published: (2025)
by: Johnson, David S.
Published: (2025)
Responsible LLM Deployment for High-Stake Decisions by Decentralized Technologies and Human-AI Interactions
by: Sachan, Swati, et al.
Published: (2025)
by: Sachan, Swati, et al.
Published: (2025)
CLASH: Evaluating Language Models on Judging High-Stakes Dilemmas from Multiple Perspectives
by: Lee, Ayoung, et al.
Published: (2025)
by: Lee, Ayoung, et al.
Published: (2025)
Who Controls the Conversation? User Perspectives On Generative AI (LLM) System Prompts
by: Neumann, Anna, et al.
Published: (2026)
by: Neumann, Anna, et al.
Published: (2026)
Understanding the Sampling Algorithm for Watt Spectrum
by: Miao, Jilang, et al.
Published: (2024)
by: Miao, Jilang, et al.
Published: (2024)
Technical Report: Competition Solution For BetterMixture
by: Zhao, Shuaijiang, et al.
Published: (2024)
by: Zhao, Shuaijiang, et al.
Published: (2024)
MLLM-as-a-Judge for Image Safety without Human Labeling
by: Wang, Zhenting, et al.
Published: (2024)
by: Wang, Zhenting, et al.
Published: (2024)
Shapley Value-based Approach for Redistributing Revenue of Matchmaking of Private Transactions in Blockchains
by: Rasheed, et al.
Published: (2025)
by: Rasheed, et al.
Published: (2025)
The Judge Who Never Admits: Hidden Shortcuts in LLM-based Evaluation
by: Marioriyad, Arash, et al.
Published: (2026)
by: Marioriyad, Arash, et al.
Published: (2026)
Who Gets the Mic? Investigating Gender Bias in the Speaker Assignment of a Speech-LLM
by: Puhach, Dariia, et al.
Published: (2025)
by: Puhach, Dariia, et al.
Published: (2025)
Algorithmic Advice as a Strategic Signal on Competitive Markets
by: Rebholz, Tobias R., et al.
Published: (2025)
by: Rebholz, Tobias R., et al.
Published: (2025)
Self-Judge: Selective Instruction Following with Alignment Self-Evaluation
by: Ye, Hai, et al.
Published: (2024)
by: Ye, Hai, et al.
Published: (2024)
Judging the Judges: Human Validation of Multi-LLM Evaluation for High-Quality K--12 Science Instructional Materials
by: He, Peng, et al.
Published: (2026)
by: He, Peng, et al.
Published: (2026)
Google or ChatGPT: Who is the Better Helper for University Students
by: Zhang, Mengmeng, et al.
Published: (2024)
by: Zhang, Mengmeng, et al.
Published: (2024)
Smooth Partial Lotteries for Stable Randomized Selection
by: Goldberg, Alexander, et al.
Published: (2026)
by: Goldberg, Alexander, et al.
Published: (2026)
Interactive AI and Human Behavior: Challenges and Pathways for AI Governance
by: Pi, Yulu, et al.
Published: (2025)
by: Pi, Yulu, et al.
Published: (2025)
"Who Am I, and Who Else Is Here?" Behavioral Differentiation Without Role Assignment in Multi-Agent LLM Systems
by: Kandoussi, Houssam EL
Published: (2026)
by: Kandoussi, Houssam EL
Published: (2026)
DIALECTIC: A Multi-Agent System for Startup Evaluation
by: Bae, Jae Yoon, et al.
Published: (2026)
by: Bae, Jae Yoon, et al.
Published: (2026)
Do Before You Judge: Self-Reference as a Pathway to Better LLM Evaluation
by: Lin, Wei-Hsiang, et al.
Published: (2025)
by: Lin, Wei-Hsiang, et al.
Published: (2025)
Introducing a novel Location-Assignment Algorithm for Activity-Based Transport Models: CARLA
by: Petre, Felix, et al.
Published: (2025)
by: Petre, Felix, et al.
Published: (2025)
Judge Anything: MLLM as a Judge Across Any Modality
by: Pu, Shu, et al.
Published: (2025)
by: Pu, Shu, et al.
Published: (2025)
Human-Centered Design Recommendations for LLM-as-a-Judge
by: Pan, Qian, et al.
Published: (2024)
by: Pan, Qian, et al.
Published: (2024)
Homogeneous Algorithms Can Reduce Competition in Personalized Pricing
by: Jo, Nathanael, et al.
Published: (2025)
by: Jo, Nathanael, et al.
Published: (2025)
Better Language Model-Based Judging Reward Modeling through Scaling Comprehension Boundaries
by: Ning, Meiling, et al.
Published: (2025)
by: Ning, Meiling, et al.
Published: (2025)
Stable Marriage: Loyalty vs. Competition
by: Ronen, Amit, et al.
Published: (2025)
by: Ronen, Amit, et al.
Published: (2025)
How Consistent Are Humans When Grading Programming Assignments?
by: Messer, Marcus, et al.
Published: (2024)
by: Messer, Marcus, et al.
Published: (2024)
Can LLM be a Personalized Judge?
by: Dong, Yijiang River, et al.
Published: (2024)
by: Dong, Yijiang River, et al.
Published: (2024)
Towards Algorithmic Fidelity: Mental Health Representation across Demographics in Synthetic vs. Human-generated Data
by: Mori, Shinka, et al.
Published: (2024)
by: Mori, Shinka, et al.
Published: (2024)
Similar Items
-
FLAWS: A Benchmark for Error Identification and Localization in Scientific Papers
by: Xi, Sarina, et al.
Published: (2025) -
Human-centered NLP Fact-checking: Co-Designing with Fact-checkers using Matchmaking for AI
by: Liu, Houjiang, et al.
Published: (2023) -
Sima AIunty: Caste Audit in LLM-Driven Matchmaking
by: Naik, Atharva, et al.
Published: (2026) -
Who Judges the Judge? Evaluating LLM-as-a-Judge for French Medical open-ended QA
by: Belmadani, Ikram, et al.
Published: (2026) -
MR. Judge: Multimodal Reasoner as a Judge
by: Pi, Renjie, et al.
Published: (2025)