Automated Quality Assessment for LLM-Based Complex Qualitative Coding: A Confidence-Diversity Framework
Fuente:
arXiv
Guardado en:
| Autores principales: | Zhao, Zhilong, Liu, Yindi |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Confidence-Diversity Framework for Calibrating AI Judgement in Accessible Qualitative Coding Tasks
por: Zhao, Zhilong, et al.
Publicado: (2025)
por: Zhao, Zhilong, et al.
Publicado: (2025)
A Novel Kuhnian Ontology for Epistemic Classification of STM Scholarly Articles
por: Saqr, Khalid M.
Publicado: (2020)
por: Saqr, Khalid M.
Publicado: (2020)
Perception-Aware Bias Detection for Query Suggestions
por: Haak, Fabian, et al.
Publicado: (2026)
por: Haak, Fabian, et al.
Publicado: (2026)
Why is it so hard to find a job now? Enter Ghost Jobs
por: Ng, Hunter
Publicado: (2024)
por: Ng, Hunter
Publicado: (2024)
MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition
por: Abtahi, Farhad, et al.
Publicado: (2026)
por: Abtahi, Farhad, et al.
Publicado: (2026)
AI Psychometrics: Evaluating the Psychological Reasoning of Large Language Models with Psychometric Validities
por: Li, Yibai, et al.
Publicado: (2026)
por: Li, Yibai, et al.
Publicado: (2026)
Pairwise Comparison for Bias Identification and Quantification
por: Haak, Fabian, et al.
Publicado: (2025)
por: Haak, Fabian, et al.
Publicado: (2025)
Claim Automation using Large Language Model
por: Mo, Zhengda, et al.
Publicado: (2026)
por: Mo, Zhengda, et al.
Publicado: (2026)
The Ethics Engine: A Modular Pipeline for Accessible Psychometric Assessment of Large Language Models
por: Van Clief, Jake, et al.
Publicado: (2025)
por: Van Clief, Jake, et al.
Publicado: (2025)
LLM-based IR-system for Bank Supervisors
por: Aarab, Ilias
Publicado: (2025)
por: Aarab, Ilias
Publicado: (2025)
Narrative Fingerprints: Multi-Scale Author Identification via Novelty Curve Dynamics
por: Zimmerman, Fred, et al.
Publicado: (2026)
por: Zimmerman, Fred, et al.
Publicado: (2026)
Mapping the Web of Science, a large-scale graph and text-based dataset with LLM embeddings
por: Kunt, Tim, et al.
Publicado: (2026)
por: Kunt, Tim, et al.
Publicado: (2026)
Towards Explainable Automated Data Quality Enhancement without Domain Knowledge
por: Sarr, Djibril
Publicado: (2024)
por: Sarr, Djibril
Publicado: (2024)
Argument Quality Annotation and Gender Bias Detection in Financial Communication through Large Language Models
por: Alhamzeh, Alaa, et al.
Publicado: (2025)
por: Alhamzeh, Alaa, et al.
Publicado: (2025)
The Association of Transformer-based Sentiment Analysis with Symptom Distress and Deterioration in Routine Psychotherapy Care
por: Faust, Douglas K., et al.
Publicado: (2026)
por: Faust, Douglas K., et al.
Publicado: (2026)
A Structural Text-Based Scaling Model for Analyzing Political Discourse
por: Vávra, Jan, et al.
Publicado: (2024)
por: Vávra, Jan, et al.
Publicado: (2024)
Moral Semantics Survive Machine Translation: Cross-Lingual Evidence from Moral Foundations Corpora
por: Skorski, Maciej
Publicado: (2026)
por: Skorski, Maciej
Publicado: (2026)
Sentiment and Volatility in Financial Markets: A Review of BERT and GARCH Applications during Geopolitical Crises
por: Mino, Domenica, et al.
Publicado: (2025)
por: Mino, Domenica, et al.
Publicado: (2025)
Artificial intelligence and downscaling global climate model future projections
por: Benestad, Rasmus E.
Publicado: (2026)
por: Benestad, Rasmus E.
Publicado: (2026)
Beyond Human Judgment: A Bayesian Evaluation of LLMs' Moral Values Understanding
por: Skorski, Maciej, et al.
Publicado: (2025)
por: Skorski, Maciej, et al.
Publicado: (2025)
2025 Southeast Asia Eleven Nations Influence Index Report
por: Meng, Wei
Publicado: (2025)
por: Meng, Wei
Publicado: (2025)
Automating Feedback Analysis in Surgical Training: Detection, Categorization, and Assessment
por: Nasriddinov, Firdavs, et al.
Publicado: (2024)
por: Nasriddinov, Firdavs, et al.
Publicado: (2024)
ValueBlindBench: Agreement-Gated Stress Testing of LLM-Judged Investment Rationales Before Returns Are Observable
por: Chang, Sidi, et al.
Publicado: (2026)
por: Chang, Sidi, et al.
Publicado: (2026)
Conformal Prediction Sets for Next-Token Prediction in Large Language Models: Balancing Coverage Guarantees with Set Efficiency
por: Kotla, Yoshith Roy, et al.
Publicado: (2025)
por: Kotla, Yoshith Roy, et al.
Publicado: (2025)
Computational Social Linguistics for Telugu Cultural Preservation: Novel Algorithms for Chandassu Metrical Pattern Recognition
por: Pavan, Boddu Sri, et al.
Publicado: (2025)
por: Pavan, Boddu Sri, et al.
Publicado: (2025)
Can Large Language Models Revolutionize Survey Research? Experiments with Disaster Preparedness Responses
por: Wang, Yan, et al.
Publicado: (2026)
por: Wang, Yan, et al.
Publicado: (2026)
Taskmaster Deconstructed: A Quantitative Look at Tension, Volatility, and Viewer Ratings
por: Silver, David H.
Publicado: (2025)
por: Silver, David H.
Publicado: (2025)
AETAS: Analysis of Evolving Temporal Affect and Semantics for Legal History
por: Wang, Qizhi
Publicado: (2025)
por: Wang, Qizhi
Publicado: (2025)
Semantic Search over 9 Million Mathematical Theorems
por: Alexander, Luke, et al.
Publicado: (2026)
por: Alexander, Luke, et al.
Publicado: (2026)
Modeling Political Discourse with Sentence-BERT and BERTopic
por: Mendonca, Margarida, et al.
Publicado: (2025)
por: Mendonca, Margarida, et al.
Publicado: (2025)
A Method for Quantifying Human Risk and a Blueprint for LLM Integration
por: Canale, Giuseppe
Publicado: (2025)
por: Canale, Giuseppe
Publicado: (2025)
Scalable AI-Driven Analytics for User Engagement and Stance Detection on Social Media
por: Seneviratne, Thammitage Piyumi Wathsala, et al.
Publicado: (2026)
por: Seneviratne, Thammitage Piyumi Wathsala, et al.
Publicado: (2026)
Whisper-LM: Improving ASR Models with Language Models for Low-Resource Languages
por: de Zuazo, Xabier, et al.
Publicado: (2025)
por: de Zuazo, Xabier, et al.
Publicado: (2025)
An Information-Theoretic Approach for Detecting Edits in AI-Generated Text
por: Kashtan, Idan, et al.
Publicado: (2023)
por: Kashtan, Idan, et al.
Publicado: (2023)
LLM-supported document separation for printed reviews from zbMATH Open
por: Pluzhnikov, Ivan, et al.
Publicado: (2026)
por: Pluzhnikov, Ivan, et al.
Publicado: (2026)
Modeling and Visualization Reasoning for Stakeholders in Education and Industry Integration Systems: Research on Structured Synthetic Dialogue Data Generation Based on NIST Standards
por: Meng, Wei
Publicado: (2025)
por: Meng, Wei
Publicado: (2025)
Supercharging Federated Intelligence Retrieval
por: Stripelis, Dimitris, et al.
Publicado: (2026)
por: Stripelis, Dimitris, et al.
Publicado: (2026)
The Syntactic Acceptability Dataset (Preview): A Resource for Machine Learning and Linguistic Analysis of English
por: Juzek, Tom S
Publicado: (2025)
por: Juzek, Tom S
Publicado: (2025)
Tokens with Meaning: A Hybrid Tokenization Approach for Turkish
por: Bayram, M. Ali, et al.
Publicado: (2025)
por: Bayram, M. Ali, et al.
Publicado: (2025)
Tokenization Standards for Linguistic Integrity: Turkish as a Benchmark
por: Bayram, M. Ali, et al.
Publicado: (2025)
por: Bayram, M. Ali, et al.
Publicado: (2025)
Ejemplares similares
-
A Confidence-Diversity Framework for Calibrating AI Judgement in Accessible Qualitative Coding Tasks
por: Zhao, Zhilong, et al.
Publicado: (2025) -
A Novel Kuhnian Ontology for Epistemic Classification of STM Scholarly Articles
por: Saqr, Khalid M.
Publicado: (2020) -
Perception-Aware Bias Detection for Query Suggestions
por: Haak, Fabian, et al.
Publicado: (2026) -
Why is it so hard to find a job now? Enter Ghost Jobs
por: Ng, Hunter
Publicado: (2024) -
MEDLEY-BENCH: Scale Buys Evaluation but Not Control in AI Metacognition
por: Abtahi, Farhad, et al.
Publicado: (2026)