The signal is the ceiling: Measurement limits of LLM-predicted experience ratings from open-ended survey text
Fuente:
arXiv
Salvato in:
| Autori principali: | Hong, Andrew, Potteiger, Jason, Zapata, Luis E. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
LLM Predictive Scoring and Validation: Inferring Experience Ratings from Unstructured Text
di: Potteiger, Jason, et al.
Pubblicazione: (2026)
di: Potteiger, Jason, et al.
Pubblicazione: (2026)
SAGEval: The frontiers of Satisfactory Agent based NLG Evaluation for reference-free open-ended text
di: Ghosh, Reshmi, et al.
Pubblicazione: (2024)
di: Ghosh, Reshmi, et al.
Pubblicazione: (2024)
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
di: Macko, Dominik, et al.
Pubblicazione: (2025)
di: Macko, Dominik, et al.
Pubblicazione: (2025)
Synthetically generated text for supervised text analysis
di: Halterman, Andrew
Pubblicazione: (2023)
di: Halterman, Andrew
Pubblicazione: (2023)
End-to-end Speech Recognition with similar length speech and text
di: Fan, Peng, et al.
Pubblicazione: (2025)
di: Fan, Peng, et al.
Pubblicazione: (2025)
Who Judges the Judge? Evaluating LLM-as-a-Judge for French Medical open-ended QA
di: Belmadani, Ikram, et al.
Pubblicazione: (2026)
di: Belmadani, Ikram, et al.
Pubblicazione: (2026)
Prompting open-source and commercial language models for grammatical error correction of English learner text
di: Davis, Christopher, et al.
Pubblicazione: (2024)
di: Davis, Christopher, et al.
Pubblicazione: (2024)
How do we measure privacy in text? A survey of text anonymization metrics
di: Ren, Yaxuan, et al.
Pubblicazione: (2025)
di: Ren, Yaxuan, et al.
Pubblicazione: (2025)
Writing as a testbed for open ended agents
di: Gooding, Sian, et al.
Pubblicazione: (2025)
di: Gooding, Sian, et al.
Pubblicazione: (2025)
Leveraging the power of transformers for guilt detection in text
di: Meque, Abdul Gafar Manuel, et al.
Pubblicazione: (2024)
di: Meque, Abdul Gafar Manuel, et al.
Pubblicazione: (2024)
End-to-end solution for linked open data query logs analytics
di: Lanasri, Dihia
Pubblicazione: (2024)
di: Lanasri, Dihia
Pubblicazione: (2024)
On the effectiveness of LLMs for automatic grading of open-ended questions in Spanish
di: Capdehourat, Germán, et al.
Pubblicazione: (2025)
di: Capdehourat, Germán, et al.
Pubblicazione: (2025)
BeanCounter: A low-toxicity, large-scale, and open dataset of business-oriented text
di: Wang, Siyan, et al.
Pubblicazione: (2024)
di: Wang, Siyan, et al.
Pubblicazione: (2024)
Leaving the barn door open for Clever Hans: Simple features predict LLM benchmark answers
di: Pacchiardi, Lorenzo, et al.
Pubblicazione: (2024)
di: Pacchiardi, Lorenzo, et al.
Pubblicazione: (2024)
Semantic-Augmented Latent Topic Modeling with LLM-in-the-Loop
di: Hong, Mengze, et al.
Pubblicazione: (2025)
di: Hong, Mengze, et al.
Pubblicazione: (2025)
Counterfactual LLM-based Framework for Measuring Rhetorical Style
di: Qiu, Jingyi, et al.
Pubblicazione: (2025)
di: Qiu, Jingyi, et al.
Pubblicazione: (2025)
Comparison of pipeline, sequence-to-sequence, and GPT models for end-to-end relation extraction: experiments with the rare disease use-case
di: Gupta, Shashank, et al.
Pubblicazione: (2023)
di: Gupta, Shashank, et al.
Pubblicazione: (2023)
Fietje: An open, efficient LLM for Dutch
di: Vanroy, Bram
Pubblicazione: (2024)
di: Vanroy, Bram
Pubblicazione: (2024)
Text-only domain adaptation for end-to-end ASR using integrated text-to-mel-spectrogram generator
di: Bataev, Vladimir, et al.
Pubblicazione: (2023)
di: Bataev, Vladimir, et al.
Pubblicazione: (2023)
Benchmark of stylistic variation in LLM-generated texts
di: Milička, Jiří, et al.
Pubblicazione: (2025)
di: Milička, Jiří, et al.
Pubblicazione: (2025)
Generalizable prediction of academic performance from short texts on social media
di: Smirnov, Ivan
Pubblicazione: (2019)
di: Smirnov, Ivan
Pubblicazione: (2019)
In your own words: computationally identifying interpretable themes in free-text survey data
di: Wang, Jenny S, et al.
Pubblicazione: (2026)
di: Wang, Jenny S, et al.
Pubblicazione: (2026)
Few shot chain-of-thought driven reasoning to prompt LLMs for open ended medical question answering
di: Nachane, Saeel Sandeep, et al.
Pubblicazione: (2024)
di: Nachane, Saeel Sandeep, et al.
Pubblicazione: (2024)
A unified front-end framework for English text-to-speech synthesis
di: Ying, Zelin, et al.
Pubblicazione: (2023)
di: Ying, Zelin, et al.
Pubblicazione: (2023)
LLM-based feature generation from text for interpretable machine learning
di: Balek, Vojtěch, et al.
Pubblicazione: (2024)
di: Balek, Vojtěch, et al.
Pubblicazione: (2024)
Private prediction for large-scale synthetic text generation
di: Amin, Kareem, et al.
Pubblicazione: (2024)
di: Amin, Kareem, et al.
Pubblicazione: (2024)
MathTutorBench: A Benchmark for Measuring Open-ended Pedagogical Capabilities of LLM Tutors
di: Macina, Jakub, et al.
Pubblicazione: (2025)
di: Macina, Jakub, et al.
Pubblicazione: (2025)
Scalable and consistent few-shot classification of survey responses using text embeddings
di: Mjaaland, Jonas Timmann, et al.
Pubblicazione: (2025)
di: Mjaaland, Jonas Timmann, et al.
Pubblicazione: (2025)
Forma mentis networks predict creativity ratings of short texts via interpretable artificial intelligence in human and GPT-simulated raters
di: Haim, Edith, et al.
Pubblicazione: (2024)
di: Haim, Edith, et al.
Pubblicazione: (2024)
Researchers waste 80% of LLM annotation costs by classifying one text at a time
di: Pipal, Christian, et al.
Pubblicazione: (2026)
di: Pipal, Christian, et al.
Pubblicazione: (2026)
Comparing LLM-generated and human-authored news text using formal syntactic theory
di: Zamaraeva, Olga, et al.
Pubblicazione: (2025)
di: Zamaraeva, Olga, et al.
Pubblicazione: (2025)
Perception of Knowledge Boundary for Large Language Models through Semi-open-ended Question Answering
di: Wen, Zhihua, et al.
Pubblicazione: (2024)
di: Wen, Zhihua, et al.
Pubblicazione: (2024)
Prompto: An open source library for asynchronous querying of LLM endpoints
di: Chan, Ryan Sze-Yin, et al.
Pubblicazione: (2024)
di: Chan, Ryan Sze-Yin, et al.
Pubblicazione: (2024)
From text to multimodal: a survey of adversarial example generation in question answering systems
di: Yigit, Gulsum, et al.
Pubblicazione: (2023)
di: Yigit, Gulsum, et al.
Pubblicazione: (2023)
Effects of diversity incentives on sample diversity and downstream model performance in LLM-based text augmentation
di: Cegin, Jan, et al.
Pubblicazione: (2024)
di: Cegin, Jan, et al.
Pubblicazione: (2024)
A Judge-free LLM Open-ended Generation Benchmark Based on the Distributional Hypothesis
di: Imajo, Kentaro, et al.
Pubblicazione: (2025)
di: Imajo, Kentaro, et al.
Pubblicazione: (2025)
Dial-In LLM: Human-Aligned LLM-in-the-loop Intent Clustering for Customer Service Dialogues
di: Hong, Mengze, et al.
Pubblicazione: (2024)
di: Hong, Mengze, et al.
Pubblicazione: (2024)
Sentence Embeddings as an intermediate target in end-to-end summarisation
di: Zembrzuski, Maciej, et al.
Pubblicazione: (2025)
di: Zembrzuski, Maciej, et al.
Pubblicazione: (2025)
CheckEval: A reliable LLM-as-a-Judge framework for evaluating text generation using checklists
di: Lee, Yukyung, et al.
Pubblicazione: (2024)
di: Lee, Yukyung, et al.
Pubblicazione: (2024)
Large Language Model (LLM) AI text generation detection based on transformer deep learning algorithm
di: Mo, Yuhong, et al.
Pubblicazione: (2024)
di: Mo, Yuhong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
LLM Predictive Scoring and Validation: Inferring Experience Ratings from Unstructured Text
di: Potteiger, Jason, et al.
Pubblicazione: (2026) -
SAGEval: The frontiers of Satisfactory Agent based NLG Evaluation for reference-free open-ended text
di: Ghosh, Reshmi, et al.
Pubblicazione: (2024) -
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
di: Macko, Dominik, et al.
Pubblicazione: (2025) -
Synthetically generated text for supervised text analysis
di: Halterman, Andrew
Pubblicazione: (2023) -
End-to-end Speech Recognition with similar length speech and text
di: Fan, Peng, et al.
Pubblicazione: (2025)