Hands-On Tutorial: Labeling with LLM and Human-in-the-Loop
Fuente:
arXiv
Saved in:
| Main Authors: | Artemova, Ekaterina, Tsvigun, Akim, Schlechtweg, Dominik, Fedorova, Natalia, Chernyshev, Konstantin, Tilga, Sergei, Obmoroshev, Boris |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Voices of Freelance Professional Writers on AI: Limitations, Expectations, and Fears
by: Ivanova, Anastasiia, et al.
Published: (2025)
by: Ivanova, Anastasiia, et al.
Published: (2025)
Tendem: A Hybrid AI+Human Platform
by: Chernyshev, Konstantin, et al.
Published: (2026)
by: Chernyshev, Konstantin, et al.
Published: (2026)
U-MATH: A University-Level Benchmark for Evaluating Mathematical Skills in LLMs
by: Chernyshev, Konstantin, et al.
Published: (2024)
by: Chernyshev, Konstantin, et al.
Published: (2024)
Beemo: Benchmark of Expert-edited Machine-generated Outputs
by: Artemova, Ekaterina, et al.
Published: (2024)
by: Artemova, Ekaterina, et al.
Published: (2024)
JEEM: Vision-Language Understanding in Four Arabic Dialects
by: Kadaoui, Karima, et al.
Published: (2025)
by: Kadaoui, Karima, et al.
Published: (2025)
Enriching Word Usage Graphs with Cluster Definitions
by: Fedorova, Mariia, et al.
Published: (2024)
by: Fedorova, Mariia, et al.
Published: (2024)
XL-DURel: Finetuning Sentence Transformers for Ordinal Word-in-Context Classification
by: Yadav, Sachin, et al.
Published: (2025)
by: Yadav, Sachin, et al.
Published: (2025)
A Head to Predict and a Head to Question: Pre-trained Uncertainty Quantification Heads for Hallucination Detection in LLM Outputs
by: Shelmanov, Artem, et al.
Published: (2025)
by: Shelmanov, Artem, et al.
Published: (2025)
Towards Automating Text Annotation: A Case Study on Semantic Proximity Annotation using GPT-4
by: Yadav, Sachin, et al.
Published: (2024)
by: Yadav, Sachin, et al.
Published: (2024)
Presence or Absence: Are Unknown Word Usages in Dictionaries?
by: Ma, Xianghe, et al.
Published: (2024)
by: Ma, Xianghe, et al.
Published: (2024)
Exploring the Robustness of Task-oriented Dialogue Systems for Colloquial German Varieties
by: Artemova, Ekaterina, et al.
Published: (2024)
by: Artemova, Ekaterina, et al.
Published: (2024)
The LSCD Benchmark: a Testbed for Diachronic Word Meaning Tasks
by: Schlechtweg, Dominik, et al.
Published: (2024)
by: Schlechtweg, Dominik, et al.
Published: (2024)
Detection of Non-recorded Word Senses in English and Swedish
by: Lautenschlager, Jonathan, et al.
Published: (2024)
by: Lautenschlager, Jonathan, et al.
Published: (2024)
LUNA: A Framework for Language Understanding and Naturalness Assessment
by: Saidov, Marat, et al.
Published: (2024)
by: Saidov, Marat, et al.
Published: (2024)
RuBia: A Russian Language Bias Detection Dataset
by: Grigoreva, Veronika, et al.
Published: (2024)
by: Grigoreva, Veronika, et al.
Published: (2024)
AIpom at SemEval-2024 Task 8: Detecting AI-produced Outputs in M4
by: Shirnin, Alexander, et al.
Published: (2024)
by: Shirnin, Alexander, et al.
Published: (2024)
Papilusion at DAGPap24: Paper or Illusion? Detecting AI-generated Scientific Papers
by: Andreev, Nikita, et al.
Published: (2024)
by: Andreev, Nikita, et al.
Published: (2024)
REPA: Russian Error Types Annotation for Evaluating Text Generation and Judgment Capabilities
by: Pugachev, Alexander, et al.
Published: (2025)
by: Pugachev, Alexander, et al.
Published: (2025)
RuBLiMP: Russian Benchmark of Linguistic Minimal Pairs
by: Taktasheva, Ekaterina, et al.
Published: (2024)
by: Taktasheva, Ekaterina, et al.
Published: (2024)
Donkii: Can Annotation Error Detection Methods Find Errors in Instruction-Tuning Datasets?
by: Weber-Genzel, Leon, et al.
Published: (2023)
by: Weber-Genzel, Leon, et al.
Published: (2023)
NER-Luxury: Named entity recognition for the fashion and luxury domain
by: Mousterou, Akim
Published: (2024)
by: Mousterou, Akim
Published: (2024)
DWUG: A large Resource of Diachronic Word Usage Graphs in Four Languages
by: Schlechtweg, Dominik, et al.
Published: (2021)
by: Schlechtweg, Dominik, et al.
Published: (2021)
Is Human-Like Text Liked by Humans? Multilingual Human Detection and Preference Against AI
by: Wang, Yuxia, et al.
Published: (2025)
by: Wang, Yuxia, et al.
Published: (2025)
LLM-DetectAIve: a Tool for Fine-Grained Machine-Generated Text Detection
by: Abassy, Mervat, et al.
Published: (2024)
by: Abassy, Mervat, et al.
Published: (2024)
ATGen: A Framework for Active Text Generation
by: Tsvigun, Akim, et al.
Published: (2025)
by: Tsvigun, Akim, et al.
Published: (2025)
SemEval-2026 Task 4: Narrative Story Similarity and Narrative Representation Learning
by: Hatzel, Hans Ole, et al.
Published: (2026)
by: Hatzel, Hans Ole, et al.
Published: (2026)
GenAI Content Detection Task 1: English and Multilingual Machine-Generated Text Detection: AI vs. Human
by: Wang, Yuxia, et al.
Published: (2025)
by: Wang, Yuxia, et al.
Published: (2025)
LLM-based Automated Grading with Human-in-the-Loop
by: Chu, Yucheng, et al.
Published: (2025)
by: Chu, Yucheng, et al.
Published: (2025)
Sebastian, Basti, Wastl?! Recognizing Named Entities in Bavarian Dialectal Data
by: Peng, Siyao, et al.
Published: (2024)
by: Peng, Siyao, et al.
Published: (2024)
Low-Resource, High-Impact: Building Corpora for Inclusive Language Technologies
by: Artemova, Ekaterina, et al.
Published: (2025)
by: Artemova, Ekaterina, et al.
Published: (2025)
Benchmarking Uncertainty Quantification Methods for Large Language Models with LM-Polygraph
by: Vashurin, Roman, et al.
Published: (2024)
by: Vashurin, Roman, et al.
Published: (2024)
Exploring Empty Spaces: Human-in-the-Loop Data Augmentation
by: Yeh, Catherine, et al.
Published: (2024)
by: Yeh, Catherine, et al.
Published: (2024)
Tutorial Proposal: Speculative Decoding for Efficient LLM Inference
by: Xia, Heming, et al.
Published: (2025)
by: Xia, Heming, et al.
Published: (2025)
The DURel Annotation Tool: Human and Computational Measurement of Semantic Proximity, Sense Clusters and Semantic Change
by: Schlechtweg, Dominik, et al.
Published: (2023)
by: Schlechtweg, Dominik, et al.
Published: (2023)
Towards A Human-in-the-Loop LLM Approach to Collaborative Discourse Analysis
by: Cohn, Clayton, et al.
Published: (2024)
by: Cohn, Clayton, et al.
Published: (2024)
Label-Looping: Highly Efficient Decoding for Transducers
by: Bataev, Vladimir, et al.
Published: (2024)
by: Bataev, Vladimir, et al.
Published: (2024)
Facilitating large language model Russian adaptation with Learned Embedding Propagation
by: Tikhomirov, Mikhail, et al.
Published: (2024)
by: Tikhomirov, Mikhail, et al.
Published: (2024)
A Fully Automated Pipeline for Conversational Discourse Annotation: Tree Scheme Generation and Labeling with Large Language Models
by: Petukhova, Kseniia, et al.
Published: (2025)
by: Petukhova, Kseniia, et al.
Published: (2025)
No Free Labels: Limitations of LLM-as-a-Judge Without Human Grounding
by: Krumdick, Michael, et al.
Published: (2025)
by: Krumdick, Michael, et al.
Published: (2025)
Learning to Predict Usage Options of Product Reviews with LLM-Generated Labels
by: Kohlenberg, Leo, et al.
Published: (2024)
by: Kohlenberg, Leo, et al.
Published: (2024)
Similar Items
-
Voices of Freelance Professional Writers on AI: Limitations, Expectations, and Fears
by: Ivanova, Anastasiia, et al.
Published: (2025) -
Tendem: A Hybrid AI+Human Platform
by: Chernyshev, Konstantin, et al.
Published: (2026) -
U-MATH: A University-Level Benchmark for Evaluating Mathematical Skills in LLMs
by: Chernyshev, Konstantin, et al.
Published: (2024) -
Beemo: Benchmark of Expert-edited Machine-generated Outputs
by: Artemova, Ekaterina, et al.
Published: (2024) -
JEEM: Vision-Language Understanding in Four Arabic Dialects
by: Kadaoui, Karima, et al.
Published: (2025)