LLMs and the Madness of Crowds
Fuente:
arXiv
Salvato in:
| Autore principale: | Bradley, William F. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Enhancing LLM Evaluations: The Garbling Trick
di: Bradley, William F.
Pubblicazione: (2024)
di: Bradley, William F.
Pubblicazione: (2024)
Voices in a Crowd: Searching for Clusters of Unique Perspectives
di: Vitsakis, Nikolas, et al.
Pubblicazione: (2024)
di: Vitsakis, Nikolas, et al.
Pubblicazione: (2024)
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
di: Nakkiran, Preetum, et al.
Pubblicazione: (2025)
di: Nakkiran, Preetum, et al.
Pubblicazione: (2025)
Wisdom of the Silicon Crowd: LLM Ensemble Prediction Capabilities Rival Human Crowd Accuracy
di: Schoenegger, Philipp, et al.
Pubblicazione: (2024)
di: Schoenegger, Philipp, et al.
Pubblicazione: (2024)
Reward Modeling with Ordinal Feedback: Wisdom of the Crowd
di: Liu, Shang, et al.
Pubblicazione: (2024)
di: Liu, Shang, et al.
Pubblicazione: (2024)
Human-LLM Hybrid Text Answer Aggregation for Crowd Annotations
di: Li, Jiyi
Pubblicazione: (2024)
di: Li, Jiyi
Pubblicazione: (2024)
MixEval: Deriving Wisdom of the Crowd from LLM Benchmark Mixtures
di: Ni, Jinjie, et al.
Pubblicazione: (2024)
di: Ni, Jinjie, et al.
Pubblicazione: (2024)
Wisdom of the Crowds in Forecasting: Forecast Summarization for Supporting Future Event Prediction
di: Saha, Anisha, et al.
Pubblicazione: (2025)
di: Saha, Anisha, et al.
Pubblicazione: (2025)
Encoder-Free Knowledge-Graph Reasoning with LLMs via Hyperdimensional Path Retrieval
di: Liu, Yezi, et al.
Pubblicazione: (2025)
di: Liu, Yezi, et al.
Pubblicazione: (2025)
Humor in AI: Massive Scale Crowd-Sourced Preferences and Benchmarks for Cartoon Captioning
di: Zhang, Jifan, et al.
Pubblicazione: (2024)
di: Zhang, Jifan, et al.
Pubblicazione: (2024)
CLAA: Cross-Layer Attention Aggregation for Accelerating LLM Prefill
di: McDanel, Bradley, et al.
Pubblicazione: (2026)
di: McDanel, Bradley, et al.
Pubblicazione: (2026)
LLMs Encode Their Failures: Predicting Success from Pre-Generation Activations
di: Lugoloobi, William, et al.
Pubblicazione: (2026)
di: Lugoloobi, William, et al.
Pubblicazione: (2026)
CLEAR: Revealing How Noise and Ambiguity Degrade Reliability in LLMs for Medicine
di: Guo, Kevin H., et al.
Pubblicazione: (2026)
di: Guo, Kevin H., et al.
Pubblicazione: (2026)
Learning to Trust the Crowd: A Multi-Model Consensus Reasoning Engine for Large Language Models
di: Kallem, Pranav
Pubblicazione: (2026)
di: Kallem, Pranav
Pubblicazione: (2026)
Scaling Efficient LLMs
di: Kausik, B. N.
Pubblicazione: (2024)
di: Kausik, B. N.
Pubblicazione: (2024)
Do LLMs Understand Your Translations? Evaluating Paragraph-level MT with Question Answering
di: Fernandes, Patrick, et al.
Pubblicazione: (2025)
di: Fernandes, Patrick, et al.
Pubblicazione: (2025)
The Illusion of Stochasticity in LLMs
di: Gu, Xiangming, et al.
Pubblicazione: (2026)
di: Gu, Xiangming, et al.
Pubblicazione: (2026)
Do LLMs Recognize Your Preferences? Evaluating Personalized Preference Following in LLMs
di: Zhao, Siyan, et al.
Pubblicazione: (2025)
di: Zhao, Siyan, et al.
Pubblicazione: (2025)
RLPF: Reinforcement Learning from Prediction Feedback for User Summarization with LLMs
di: Wu, Jiaxing, et al.
Pubblicazione: (2024)
di: Wu, Jiaxing, et al.
Pubblicazione: (2024)
Multicalibration for Confidence Scoring in LLMs
di: Detommaso, Gianluca, et al.
Pubblicazione: (2024)
di: Detommaso, Gianluca, et al.
Pubblicazione: (2024)
Refusal in LLMs is an Affine Function
di: Marshall, Thomas, et al.
Pubblicazione: (2024)
di: Marshall, Thomas, et al.
Pubblicazione: (2024)
Detecting Conceptual Abstraction in LLMs
di: Regneri, Michaela, et al.
Pubblicazione: (2024)
di: Regneri, Michaela, et al.
Pubblicazione: (2024)
Can LLMs subtract numbers?
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
di: Jobanputra, Mayank, et al.
Pubblicazione: (2025)
On Pruning State-Space LLMs
di: Ghattas, Tamer, et al.
Pubblicazione: (2025)
di: Ghattas, Tamer, et al.
Pubblicazione: (2025)
Emergent Response Planning in LLMs
di: Dong, Zhichen, et al.
Pubblicazione: (2025)
di: Dong, Zhichen, et al.
Pubblicazione: (2025)
Leverage Unlearning to Sanitize LLMs
di: Boutet, Antoine, et al.
Pubblicazione: (2025)
di: Boutet, Antoine, et al.
Pubblicazione: (2025)
Leveraging the true depth of LLMs
di: González, Ramón Calvo, et al.
Pubblicazione: (2025)
di: González, Ramón Calvo, et al.
Pubblicazione: (2025)
Confidence-Credibility Aware Weighted Ensembles of Small LLMs Outperform Large LLMs in Emotion Detection
di: Elgabry, Menna, et al.
Pubblicazione: (2025)
di: Elgabry, Menna, et al.
Pubblicazione: (2025)
AMUSD: Asynchronous Multi-Device Speculative Decoding for LLM Acceleration
di: McDanel, Bradley
Pubblicazione: (2024)
di: McDanel, Bradley
Pubblicazione: (2024)
Low-Bit Quantization Favors Undertrained LLMs: Scaling Laws for Quantized LLMs with 100T Training Tokens
di: Ouyang, Xu, et al.
Pubblicazione: (2024)
di: Ouyang, Xu, et al.
Pubblicazione: (2024)
How Data Inter-connectivity Shapes LLMs Unlearning: A Structural Unlearning Perspective
di: Qiu, Xinchi, et al.
Pubblicazione: (2024)
di: Qiu, Xinchi, et al.
Pubblicazione: (2024)
Benchmarking LLMs' Judgments with No Gold Standard
di: Xu, Shengwei, et al.
Pubblicazione: (2024)
di: Xu, Shengwei, et al.
Pubblicazione: (2024)
Jailbreaking LLMs with Arabic Transliteration and Arabizi
di: Ghanim, Mansour Al, et al.
Pubblicazione: (2024)
di: Ghanim, Mansour Al, et al.
Pubblicazione: (2024)
Toward a Theory of Tokenization in LLMs
di: Rajaraman, Nived, et al.
Pubblicazione: (2024)
di: Rajaraman, Nived, et al.
Pubblicazione: (2024)
Data Science with LLMs and Interpretable Models
di: Bordt, Sebastian, et al.
Pubblicazione: (2024)
di: Bordt, Sebastian, et al.
Pubblicazione: (2024)
Aligning (Medical) LLMs for (Counterfactual) Fairness
di: Poulain, Raphael, et al.
Pubblicazione: (2024)
di: Poulain, Raphael, et al.
Pubblicazione: (2024)
AgreeMate: Teaching LLMs to Haggle
di: Chatterjee, Ainesh, et al.
Pubblicazione: (2024)
di: Chatterjee, Ainesh, et al.
Pubblicazione: (2024)
Reasoning Boosts Opinion Alignment in LLMs
di: Berdoz, Frédéric, et al.
Pubblicazione: (2026)
di: Berdoz, Frédéric, et al.
Pubblicazione: (2026)
On the Calibration of Multilingual Question Answering LLMs
di: Yang, Yahan, et al.
Pubblicazione: (2023)
di: Yang, Yahan, et al.
Pubblicazione: (2023)
PEARL: Towards Permutation-Resilient LLMs
di: Chen, Liang, et al.
Pubblicazione: (2025)
di: Chen, Liang, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Enhancing LLM Evaluations: The Garbling Trick
di: Bradley, William F.
Pubblicazione: (2024) -
Voices in a Crowd: Searching for Clusters of Unique Perspectives
di: Vitsakis, Nikolas, et al.
Pubblicazione: (2024) -
Trained on Tokens, Calibrated on Concepts: The Emergence of Semantic Calibration in LLMs
di: Nakkiran, Preetum, et al.
Pubblicazione: (2025) -
Wisdom of the Silicon Crowd: LLM Ensemble Prediction Capabilities Rival Human Crowd Accuracy
di: Schoenegger, Philipp, et al.
Pubblicazione: (2024) -
Reward Modeling with Ordinal Feedback: Wisdom of the Crowd
di: Liu, Shang, et al.
Pubblicazione: (2024)