PPLqa: An Unsupervised Information-Theoretic Quality Metric for Comparing Generative Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Friedland, Gerald, Huang, Xin, Cui, Yueying, Kapoor, Vishaal, Khetan, Ashish, Das, Sanjiv |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
BoAT v2 -- A Web-Based Dependency Annotation Tool with Focus on Agglutinative Languages
by: Akkurt, Salih Furkan, et al.
Published: (2022)
by: Akkurt, Salih Furkan, et al.
Published: (2022)
PrivSpike: Employing Homomorphic Encryption for Private Inference of Deep Spiking Neural Networks
by: Njungle, Nges Brian, et al.
Published: (2025)
by: Njungle, Nges Brian, et al.
Published: (2025)
Curated endoscopic retrograde cholangiopancreatography images dataset
by: Andrade, Alda João, et al.
Published: (2026)
by: Andrade, Alda João, et al.
Published: (2026)
Data Readiness for AI: A 360-Degree Survey
by: Hiniduma, Kaveen, et al.
Published: (2024)
by: Hiniduma, Kaveen, et al.
Published: (2024)
Fin-Fact: A Benchmark Dataset for Multimodal Financial Fact Checking and Explanation Generation
by: Rangapur, Aman, et al.
Published: (2023)
by: Rangapur, Aman, et al.
Published: (2023)
A Super-Learner with Large Language Models for Medical Emergency Advising
by: Aityan, Sergey K., et al.
Published: (2025)
by: Aityan, Sergey K., et al.
Published: (2025)
A Labeled Array Distance Metric for Measuring Image Segmentation Quality
by: Berijanian, Maryam, et al.
Published: (2024)
by: Berijanian, Maryam, et al.
Published: (2024)
Improving the Precision of CNNs for Magnetic Resonance Spectral Modeling
by: LaMaster, John, et al.
Published: (2024)
by: LaMaster, John, et al.
Published: (2024)
The Battle of LLMs: A Comparative Study in Conversational QA Tasks
by: Rangapur, Aryan, et al.
Published: (2024)
by: Rangapur, Aryan, et al.
Published: (2024)
Context-Aware Clustering using Large Language Models
by: Tipirneni, Sindhu, et al.
Published: (2024)
by: Tipirneni, Sindhu, et al.
Published: (2024)
Defining and Quantifying Creative Behavior in Popular Image Generators
by: Ramaswamy, Aditi, et al.
Published: (2025)
by: Ramaswamy, Aditi, et al.
Published: (2025)
Knowledge Distillation for Large Language Models
by: La Torre, Alejandro Paredes, et al.
Published: (2026)
by: La Torre, Alejandro Paredes, et al.
Published: (2026)
StatCounter: A Longitudinal Study of a Portable Scholarly Metric Display
by: Oppenlaender, Jonas
Published: (2026)
by: Oppenlaender, Jonas
Published: (2026)
Hebbian Memory-Augmented Recurrent Networks: Engram Neurons in Deep Learning
by: Szelogowski, Daniel
Published: (2025)
by: Szelogowski, Daniel
Published: (2025)
Engram Memory Encoding and Retrieval: A Neurocomputational Perspective
by: Szelogowski, Daniel
Published: (2025)
by: Szelogowski, Daniel
Published: (2025)
Morpheme Induction for Emergent Language
by: Boldt, Brendon, et al.
Published: (2025)
by: Boldt, Brendon, et al.
Published: (2025)
EBBS: An Ensemble with Bi-Level Beam Search for Zero-Shot Machine Translation
by: Wen, Yuqiao, et al.
Published: (2024)
by: Wen, Yuqiao, et al.
Published: (2024)
ELCC: the Emergent Language Corpus Collection
by: Boldt, Brendon, et al.
Published: (2024)
by: Boldt, Brendon, et al.
Published: (2024)
Searching for the Most Human-like Emergent Language
by: Boldt, Brendon, et al.
Published: (2025)
by: Boldt, Brendon, et al.
Published: (2025)
XferBench: a Data-Driven Benchmark for Emergent Language
by: Boldt, Brendon, et al.
Published: (2024)
by: Boldt, Brendon, et al.
Published: (2024)
Spatially Disaggregated Energy Consumption and Emissions in End-use Sectors for Germany and Spain
by: Patil, Shruthi, et al.
Published: (2025)
by: Patil, Shruthi, et al.
Published: (2025)
Measuring the originality of intellectual property assets based on machine learning outputs
by: Ragot, Sébastien
Published: (2020)
by: Ragot, Sébastien
Published: (2020)
Towards Statistically Significant Taxonomy Aware Co-location Pattern Detection
by: Ghosh, Subhankar, et al.
Published: (2024)
by: Ghosh, Subhankar, et al.
Published: (2024)
QPMeL - Quantum-Aware Classically-Trained Embeddings via Projective Metric Learning
by: Sharma, Vinayak, et al.
Published: (2023)
by: Sharma, Vinayak, et al.
Published: (2023)
The Economic Implications of Large Language Model Selection on Earnings and Return on Investment: A Decision Theoretic Model
by: Xexéo, Geraldo, et al.
Published: (2024)
by: Xexéo, Geraldo, et al.
Published: (2024)
Information-Theoretic Quality Metric of Low-Dimensional Embeddings
by: Gutiérrez-Bernal, Sebastián, et al.
Published: (2025)
by: Gutiérrez-Bernal, Sebastián, et al.
Published: (2025)
Knowing Isn't Understanding: Re-grounding Generative Proactivity with Epistemic and Behavioral Insight
by: Kaur, Kirandeep, et al.
Published: (2026)
by: Kaur, Kirandeep, et al.
Published: (2026)
iRoCo: Intuitive Robot Control From Anywhere Using a Smartwatch
by: Weigend, Fabian C, et al.
Published: (2024)
by: Weigend, Fabian C, et al.
Published: (2024)
Anytime, Anywhere: Human Arm Pose from Smartwatch Data for Ubiquitous Robot Control and Teleoperation
by: Weigend, Fabian C, et al.
Published: (2023)
by: Weigend, Fabian C, et al.
Published: (2023)
Brain Organoid Computing -- an Overview
by: Talavera, Yannic, et al.
Published: (2025)
by: Talavera, Yannic, et al.
Published: (2025)
Language Detection for Transliterated Content
by: S, Selva Kumar, et al.
Published: (2024)
by: S, Selva Kumar, et al.
Published: (2024)
Exploring Model Invariance with Discrete Search for Ultra-Low-Bit Quantization
by: Wen, Yuqiao, et al.
Published: (2025)
by: Wen, Yuqiao, et al.
Published: (2025)
The QCET Taxonomy of Standard Quality Criterion Names and Definitions for the Evaluation of NLP Systems
by: Belz, Anya, et al.
Published: (2025)
by: Belz, Anya, et al.
Published: (2025)
BACE: LLM-based Code Generation through Bayesian Anchored Co-Evolution of Code and Test Populations
by: Silva, Kaushitha, et al.
Published: (2026)
by: Silva, Kaushitha, et al.
Published: (2026)
Automated, Unsupervised, and Auto-parameterized Inference of Data Patterns and Anomaly Detection
by: Qin, Qiaolin, et al.
Published: (2024)
by: Qin, Qiaolin, et al.
Published: (2024)
Reducing False Discoveries in Statistically-Significant Regional-Colocation Mining: A Summary of Results
by: Ghosh, Subhankar, et al.
Published: (2024)
by: Ghosh, Subhankar, et al.
Published: (2024)
Comparative Performance Analysis of Transformer-Based Pre-Trained Models for Detecting Keratoconus Disease
by: Ahmed, Nayeem, et al.
Published: (2024)
by: Ahmed, Nayeem, et al.
Published: (2024)
A Framework for Collaborating a Large Language Model Tool in Brainstorming for Triggering Creative Thoughts
by: Chang, Hung-Fu, et al.
Published: (2024)
by: Chang, Hung-Fu, et al.
Published: (2024)
On the Opportunities of Large Language Models for Programming Process Data
by: Edwards, John, et al.
Published: (2024)
by: Edwards, John, et al.
Published: (2024)
An Exploration of Default Images in Text-to-Image Generation
by: Simonen, Hannu, et al.
Published: (2025)
by: Simonen, Hannu, et al.
Published: (2025)
Similar Items
-
BoAT v2 -- A Web-Based Dependency Annotation Tool with Focus on Agglutinative Languages
by: Akkurt, Salih Furkan, et al.
Published: (2022) -
PrivSpike: Employing Homomorphic Encryption for Private Inference of Deep Spiking Neural Networks
by: Njungle, Nges Brian, et al.
Published: (2025) -
Curated endoscopic retrograde cholangiopancreatography images dataset
by: Andrade, Alda João, et al.
Published: (2026) -
Data Readiness for AI: A 360-Degree Survey
by: Hiniduma, Kaveen, et al.
Published: (2024) -
Fin-Fact: A Benchmark Dataset for Multimodal Financial Fact Checking and Explanation Generation
by: Rangapur, Aman, et al.
Published: (2023)