Can Frontier LLMs Replace Annotators in Biomedical Text Mining? Analyzing Challenges and Exploring Solutions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhao, Yichong, Goto, Susumu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Are LLMs Ready to Replace Bangla Annotators?
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2026)
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2026)
Can LLMs Understand Unvoiced Speech? Exploring EMG-to-Text Conversion with LLMs
von: Mohapatra, Payal, et al.
Veröffentlicht: (2025)
von: Mohapatra, Payal, et al.
Veröffentlicht: (2025)
Automated Text Mining of Experimental Methodologies from Biomedical Literature
von: Guo, Ziqing
Veröffentlicht: (2024)
von: Guo, Ziqing
Veröffentlicht: (2024)
Applying Text Mining to Analyze Human Question Asking in Creativity Research
von: Wróblewska, Anna, et al.
Veröffentlicht: (2025)
von: Wróblewska, Anna, et al.
Veröffentlicht: (2025)
Can Large Language Models Replace Data Scientists in Biomedical Research?
von: Wang, Zifeng, et al.
Veröffentlicht: (2024)
von: Wang, Zifeng, et al.
Veröffentlicht: (2024)
BioPars: A Pretrained Biomedical Large Language Model for Persian Biomedical Text Mining
von: Merzah, Baqer M., et al.
Veröffentlicht: (2025)
von: Merzah, Baqer M., et al.
Veröffentlicht: (2025)
The Alternative Annotator Test for LLM-as-a-Judge: How to Statistically Justify Replacing Human Annotators with LLMs
von: Calderon, Nitay, et al.
Veröffentlicht: (2025)
von: Calderon, Nitay, et al.
Veröffentlicht: (2025)
Can Open-Source LLMs Compete with Commercial Models? Exploring the Few-Shot Performance of Current GPT Models in Biomedical Tasks
von: Ateia, Samy, et al.
Veröffentlicht: (2024)
von: Ateia, Samy, et al.
Veröffentlicht: (2024)
Analyzing Semantic Change through Lexical Replacements
von: Periti, Francesco, et al.
Veröffentlicht: (2024)
von: Periti, Francesco, et al.
Veröffentlicht: (2024)
Can We Edit LLMs for Long-Tail Biomedical Knowledge?
von: Yi, Xinhao, et al.
Veröffentlicht: (2025)
von: Yi, Xinhao, et al.
Veröffentlicht: (2025)
A Retrieval-Augmented Knowledge Mining Method with Deep Thinking LLMs for Biomedical Research and Clinical Support
von: Feng, Yichun, et al.
Veröffentlicht: (2025)
von: Feng, Yichun, et al.
Veröffentlicht: (2025)
Exploring Consciousness in LLMs: A Systematic Survey of Theories, Implementations, and Frontier Risks
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
von: Chen, Sirui, et al.
Veröffentlicht: (2025)
Rethinking Hate Speech Detection on Social Media: Can LLMs Replace Traditional Models?
von: Singh, Daman Deep, et al.
Veröffentlicht: (2025)
von: Singh, Daman Deep, et al.
Veröffentlicht: (2025)
PubMedCausal: A Span-Level Annotated Corpus for Causal Relation Extraction in Biomedical Text
von: Kunle-John, Ifeoluwa, et al.
Veröffentlicht: (2026)
von: Kunle-John, Ifeoluwa, et al.
Veröffentlicht: (2026)
From Text to Emotion: Unveiling the Emotion Annotation Capabilities of LLMs
von: Niu, Minxue, et al.
Veröffentlicht: (2024)
von: Niu, Minxue, et al.
Veröffentlicht: (2024)
Short-Path Prompting in LLMs: Analyzing Reasoning Instability and Solutions for Robust Performance
von: Tang, Zuoli, et al.
Veröffentlicht: (2025)
von: Tang, Zuoli, et al.
Veröffentlicht: (2025)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
von: Chatrath, Veronica, et al.
Veröffentlicht: (2024)
von: Chatrath, Veronica, et al.
Veröffentlicht: (2024)
SelectLLM: Can LLMs Select Important Instructions to Annotate?
von: Parkar, Ritik Sachin, et al.
Veröffentlicht: (2024)
von: Parkar, Ritik Sachin, et al.
Veröffentlicht: (2024)
Efficient Solutions For An Intriguing Failure of LLMs: Long Context Window Does Not Mean LLMs Can Analyze Long Sequences Flawlessly
von: Hosseini, Peyman, et al.
Veröffentlicht: (2024)
von: Hosseini, Peyman, et al.
Veröffentlicht: (2024)
Beyond Traditional Benchmarks: Analyzing Behaviors of Open LLMs on Data-to-Text Generation
von: Kasner, Zdeněk, et al.
Veröffentlicht: (2024)
von: Kasner, Zdeněk, et al.
Veröffentlicht: (2024)
Analyzing Dataset Annotation Quality Management in the Wild
von: Klie, Jan-Christoph, et al.
Veröffentlicht: (2023)
von: Klie, Jan-Christoph, et al.
Veröffentlicht: (2023)
MultiChallenge: A Realistic Multi-Turn Conversation Evaluation Benchmark Challenging to Frontier LLMs
von: Sirdeshmukh, Ved, et al.
Veröffentlicht: (2025)
von: Sirdeshmukh, Ved, et al.
Veröffentlicht: (2025)
Who and What? Using Linguistic Features and Annotator Characteristics to Analyze Annotation Variation
von: Maurer, Maximilian, et al.
Veröffentlicht: (2026)
von: Maurer, Maximilian, et al.
Veröffentlicht: (2026)
Can Vision Replace Text in Working Memory? Evidence from Spatial n-Back in Vision-Language Models
von: Liang, Sichu, et al.
Veröffentlicht: (2026)
von: Liang, Sichu, et al.
Veröffentlicht: (2026)
Vocabulary Transfer for Biomedical Texts: Add Tokens if You Can Not Add Data
von: Singh, Priyanka, et al.
Veröffentlicht: (2022)
von: Singh, Priyanka, et al.
Veröffentlicht: (2022)
Exploring Traffic Crash Narratives in Jordan Using Text Mining Analytics
von: Jaradat, Shadi, et al.
Veröffentlicht: (2024)
von: Jaradat, Shadi, et al.
Veröffentlicht: (2024)
Unveiling Cultural Blind Spots: Analyzing the Limitations of mLLMs in Procedural Text Comprehension
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
von: Yari, Amir Hossein, et al.
Veröffentlicht: (2025)
Streamlining Biomedical Research with Specialized LLMs
von: Chen, Linqing, et al.
Veröffentlicht: (2025)
von: Chen, Linqing, et al.
Veröffentlicht: (2025)
Exploring Cross-lingual Latent Transplantation: Mutual Opportunities and Open Challenges
von: Ye, Yangfan, et al.
Veröffentlicht: (2024)
von: Ye, Yangfan, et al.
Veröffentlicht: (2024)
From Fallback to Frontline: When Can LLMs be Superior Annotators of Human Perspectives?
von: Amin, Hasan, et al.
Veröffentlicht: (2026)
von: Amin, Hasan, et al.
Veröffentlicht: (2026)
LLMs for Doctors: Leveraging Medical LLMs to Assist Doctors, Not Replace Them
von: Xie, Wenya, et al.
Veröffentlicht: (2024)
von: Xie, Wenya, et al.
Veröffentlicht: (2024)
Argument Mining as a Text-to-Text Generation Task
von: Kawarada, Masayuki, et al.
Veröffentlicht: (2026)
von: Kawarada, Masayuki, et al.
Veröffentlicht: (2026)
Can Generic LLMs Help Analyze Child-adult Interactions Involving Children with Autism in Clinical Observation?
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
von: Feng, Tiantian, et al.
Veröffentlicht: (2024)
Behavior-Equivalent Token: Single-Token Replacement for Long Prompts in LLMs
von: Dong, Jiancheng, et al.
Veröffentlicht: (2025)
von: Dong, Jiancheng, et al.
Veröffentlicht: (2025)
Do LLMs Surpass Encoders for Biomedical NER?
von: Obeidat, Motasem S, et al.
Veröffentlicht: (2025)
von: Obeidat, Motasem S, et al.
Veröffentlicht: (2025)
On-the-fly Definition Augmentation of LLMs for Biomedical NER
von: Munnangi, Monica, et al.
Veröffentlicht: (2024)
von: Munnangi, Monica, et al.
Veröffentlicht: (2024)
Can Agent Conquer Web? Exploring the Frontiers of ChatGPT Atlas Agent in Web Games
von: Zhang, Jingran, et al.
Veröffentlicht: (2025)
von: Zhang, Jingran, et al.
Veröffentlicht: (2025)
Open-Source LLMs for Text Annotation: A Practical Guide for Model Setting and Fine-Tuning
von: Alizadeh, Meysam, et al.
Veröffentlicht: (2023)
von: Alizadeh, Meysam, et al.
Veröffentlicht: (2023)
Can LLMs Help Uncover Insights about LLMs? A Large-Scale, Evolving Literature Analysis of Frontier LLMs
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
von: Park, Jungsoo, et al.
Veröffentlicht: (2025)
Coal Mining Question Answering with LLMs
von: Rivera, Antonio Carlos, et al.
Veröffentlicht: (2024)
von: Rivera, Antonio Carlos, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Are LLMs Ready to Replace Bangla Annotators?
von: Hasan, Md. Najib, et al.
Veröffentlicht: (2026) -
Can LLMs Understand Unvoiced Speech? Exploring EMG-to-Text Conversion with LLMs
von: Mohapatra, Payal, et al.
Veröffentlicht: (2025) -
Automated Text Mining of Experimental Methodologies from Biomedical Literature
von: Guo, Ziqing
Veröffentlicht: (2024) -
Applying Text Mining to Analyze Human Question Asking in Creativity Research
von: Wróblewska, Anna, et al.
Veröffentlicht: (2025) -
Can Large Language Models Replace Data Scientists in Biomedical Research?
von: Wang, Zifeng, et al.
Veröffentlicht: (2024)