Self-reflection in Automated Qualitative Coding: Improving Text Annotation through Secondary LLM Critique
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Dunivin, Zackary Okun, Noori, Mobina, Frey, Seth, Atkinson, Curtis |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Scalable Qualitative Coding with LLMs: Chain-of-Thought Reasoning Matches Human Performance in Some Hermeneutic Tasks
par: Dunivin, Zackary Okun
Publié: (2024)
par: Dunivin, Zackary Okun
Publié: (2024)
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
par: Fu, Lingyue, et autres
Publié: (2025)
par: Fu, Lingyue, et autres
Publié: (2025)
BanglaForge: LLM Collaboration with Self-Refinement for Bangla Code Generation
par: Dihan, Mahir Labib, et autres
Publié: (2025)
par: Dihan, Mahir Labib, et autres
Publié: (2025)
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios
par: Huang, Shiting, et autres
Publié: (2025)
par: Huang, Shiting, et autres
Publié: (2025)
NaviQAte: Functionality-Guided Web Application Navigation
par: Shahbandeh, Mobina, et autres
Publié: (2024)
par: Shahbandeh, Mobina, et autres
Publié: (2024)
LLM-as-a-Judge for Reference-less Automatic Code Validation and Refinement for Natural Language to Bash in IT Automation
par: Vo, Ngoc Phuoc An, et autres
Publié: (2025)
par: Vo, Ngoc Phuoc An, et autres
Publié: (2025)
DevEval: A Manually-Annotated Code Generation Benchmark Aligned with Real-World Code Repositories
par: Li, Jia, et autres
Publié: (2024)
par: Li, Jia, et autres
Publié: (2024)
ProbeLLM: Automating Principled Diagnosis of LLM Failures
par: Huang, Yue, et autres
Publié: (2026)
par: Huang, Yue, et autres
Publié: (2026)
Evaluating and Achieving Controllable Code Completion in Code LLM
par: Zhang, Jiajun, et autres
Publié: (2026)
par: Zhang, Jiajun, et autres
Publié: (2026)
Code Fingerprints: Disentangled Attribution of LLM-Generated Code
par: Guo, Jiaxun, et autres
Publié: (2026)
par: Guo, Jiaxun, et autres
Publié: (2026)
SEW: Self-Evolving Agentic Workflows for Automated Code Generation
par: Liu, Siwei, et autres
Publié: (2025)
par: Liu, Siwei, et autres
Publié: (2025)
Improving Code Localization with Repository Memory
par: Wang, Boshi, et autres
Publié: (2025)
par: Wang, Boshi, et autres
Publié: (2025)
Functional Consistency of LLM Code Embeddings: A Self-Evolving Data Synthesis Framework for Benchmarking
par: Li, Zhuohao, et autres
Publié: (2025)
par: Li, Zhuohao, et autres
Publié: (2025)
PerfCodeGen: Improving Performance of LLM Generated Code with Execution Feedback
par: Peng, Yun, et autres
Publié: (2024)
par: Peng, Yun, et autres
Publié: (2024)
Comparing Developer and LLM Biases in Code Evaluation
par: Mittal, Aditya, et autres
Publié: (2026)
par: Mittal, Aditya, et autres
Publié: (2026)
EffiSkill: Agent Skill Based Automated Code Efficiency Optimization
par: Wang, Zimu, et autres
Publié: (2026)
par: Wang, Zimu, et autres
Publié: (2026)
Enhanced Automated Code Vulnerability Repair using Large Language Models
par: de-Fitero-Dominguez, David, et autres
Publié: (2024)
par: de-Fitero-Dominguez, David, et autres
Publié: (2024)
CYCLE: Learning to Self-Refine the Code Generation
par: Ding, Yangruibo, et autres
Publié: (2024)
par: Ding, Yangruibo, et autres
Publié: (2024)
From Critique to Clarity: A Pathway to Faithful and Personalized Code Explanations with Large Language Models
par: Xu, Zexing, et autres
Publié: (2024)
par: Xu, Zexing, et autres
Publié: (2024)
Don't Judge Code by Its Cover: Exploring Biases in LLM Judges for Code Evaluation
par: Moon, Jiwon, et autres
Publié: (2025)
par: Moon, Jiwon, et autres
Publié: (2025)
LLMSniffer: Detecting LLM-Generated Code via GraphCodeBERT and Supervised Contrastive Learning
par: Dihan, Mahir Labib, et autres
Publié: (2026)
par: Dihan, Mahir Labib, et autres
Publié: (2026)
Measuring LLM Code Generation Stability via Structural Entropy
par: Song, Yewei, et autres
Publié: (2025)
par: Song, Yewei, et autres
Publié: (2025)
Showing LLM-Generated Code Selectively Based on Confidence of LLMs
par: Li, Jia, et autres
Publié: (2024)
par: Li, Jia, et autres
Publié: (2024)
SWE-Pruner: Self-Adaptive Context Pruning for Coding Agents
par: Wang, Yuhang, et autres
Publié: (2026)
par: Wang, Yuhang, et autres
Publié: (2026)
Leveraging Print Debugging to Improve Code Generation in Large Language Models
par: Hu, Xueyu, et autres
Publié: (2024)
par: Hu, Xueyu, et autres
Publié: (2024)
GrowthHacker: Automated Off-Policy Evaluation Optimization Using Code-Modifying LLM Agents
par: Wu, Jie JW, et autres
Publié: (2025)
par: Wu, Jie JW, et autres
Publié: (2025)
ProjectEval: A Benchmark for Programming Agents Automated Evaluation on Project-Level Code Generation
par: Liu, Kaiyuan, et autres
Publié: (2025)
par: Liu, Kaiyuan, et autres
Publié: (2025)
MATCH: Task-Driven Code Evaluation through Contrastive Learning
par: Ghoummaid, Marah, et autres
Publié: (2025)
par: Ghoummaid, Marah, et autres
Publié: (2025)
EffiLearner: Enhancing Efficiency of Generated Code via Self-Optimization
par: Huang, Dong, et autres
Publié: (2024)
par: Huang, Dong, et autres
Publié: (2024)
Alibaba LingmaAgent: Improving Automated Issue Resolution via Comprehensive Repository Exploration
par: Ma, Yingwei, et autres
Publié: (2024)
par: Ma, Yingwei, et autres
Publié: (2024)
SelfCodeAlign: Self-Alignment for Code Generation
par: Wei, Yuxiang, et autres
Publié: (2024)
par: Wei, Yuxiang, et autres
Publié: (2024)
StepCoder: Improve Code Generation with Reinforcement Learning from Compiler Feedback
par: Dou, Shihan, et autres
Publié: (2024)
par: Dou, Shihan, et autres
Publié: (2024)
ArtifactsBench: Bridging the Visual-Interactive Gap in LLM Code Generation Evaluation
par: Zhang, Chenchen, et autres
Publié: (2025)
par: Zhang, Chenchen, et autres
Publié: (2025)
UICoder: Finetuning Large Language Models to Generate User Interface Code through Automated Feedback
par: Wu, Jason, et autres
Publié: (2024)
par: Wu, Jason, et autres
Publié: (2024)
Improving Small Language Models for Code Generation with Reinforcement Learning from Verification Feedback
par: Skopin, Egor, et autres
Publié: (2026)
par: Skopin, Egor, et autres
Publié: (2026)
Show and Tell: Prompt Strategies for Style Control in Multi-Turn LLM Code Generation
par: Bohr, Jeremiah
Publié: (2025)
par: Bohr, Jeremiah
Publié: (2025)
Generating Equivalent Representations of Code By A Self-Reflection Approach
par: Li, Jia, et autres
Publié: (2024)
par: Li, Jia, et autres
Publié: (2024)
LLM Agents Improve Semantic Code Search
par: Jain, Sarthak, et autres
Publié: (2024)
par: Jain, Sarthak, et autres
Publié: (2024)
To Diff or Not to Diff? Structure-Aware and Adaptive Output Formats for Efficient LLM-based Code Editing
par: Cheng, Wei, et autres
Publié: (2026)
par: Cheng, Wei, et autres
Publié: (2026)
SRLCG: Self-Rectified Large-Scale Code Generation with Multidimensional Chain-of-Thought and Dynamic Backtracking
par: Ma, Hongru, et autres
Publié: (2025)
par: Ma, Hongru, et autres
Publié: (2025)
Documents similaires
-
Scalable Qualitative Coding with LLMs: Chain-of-Thought Reasoning Matches Human Performance in Some Hermeneutic Tasks
par: Dunivin, Zackary Okun
Publié: (2024) -
Automatically Benchmarking LLM Code Agents through Agent-Driven Annotation and Evaluation
par: Fu, Lingyue, et autres
Publié: (2025) -
BanglaForge: LLM Collaboration with Self-Refinement for Bangla Code Generation
par: Dihan, Mahir Labib, et autres
Publié: (2025) -
CRITICTOOL: Evaluating Self-Critique Capabilities of Large Language Models in Tool-Calling Error Scenarios
par: Huang, Shiting, et autres
Publié: (2025) -
NaviQAte: Functionality-Guided Web Application Navigation
par: Shahbandeh, Mobina, et autres
Publié: (2024)