Taec: a Manually annotated text dataset for trait and phenotype extraction and entity linking in wheat breeding literature
Fuente:
arXiv
Saved in:
| Main Authors: | Nédellec, Claire, Sauvion, Clara, Bossy, Robert, Borovikova, Mariya, Deléger, Louise |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
ESNERA: Empirical and semantic named entity alignment for named entity dataset merging
by: Zhang, Xiaobo, et al.
Published: (2025)
by: Zhang, Xiaobo, et al.
Published: (2025)
Algorithms For Automatic Accentuation And Transcription Of Russian Texts In Speech Recognition Systems
by: Iakovenko, Olga, et al.
Published: (2024)
by: Iakovenko, Olga, et al.
Published: (2024)
Are generative AI text annotations systematically biased?
by: Stolwijk, Sjoerd B., et al.
Published: (2025)
by: Stolwijk, Sjoerd B., et al.
Published: (2025)
MaterioMiner -- An ontology-based text mining dataset for extraction of process-structure-property entities
by: Durmaz, Ali Riza, et al.
Published: (2024)
by: Durmaz, Ali Riza, et al.
Published: (2024)
Large language models struggle with ethnographic text annotation
by: Goodall, Leonardo S., et al.
Published: (2026)
by: Goodall, Leonardo S., et al.
Published: (2026)
FRACCO: A gold-standard annotated corpus of oncological entities with ICD-O-3.1 normalisation
by: Pignat, Johann, et al.
Published: (2025)
by: Pignat, Johann, et al.
Published: (2025)
Multimodal large language model for wheat breeding: a new exploration of smart breeding
by: Yang, Guofeng, et al.
Published: (2024)
by: Yang, Guofeng, et al.
Published: (2024)
Noise reduction in BERT NER models for clinical entity extraction
by: Jiwani, Kuldeep, et al.
Published: (2026)
by: Jiwani, Kuldeep, et al.
Published: (2026)
The impact of fine tuning in LLaMA on hallucinations for named entity extraction in legal documentation
by: Vargas, Francisco, et al.
Published: (2025)
by: Vargas, Francisco, et al.
Published: (2025)
Generalized knowledge-enhanced framework for biomedical entity and relation extraction
by: Nguyen, Minh, et al.
Published: (2024)
by: Nguyen, Minh, et al.
Published: (2024)
Tgea: An error-annotated dataset and benchmark tasks for text generation from pretrained language models
by: He, Jie, et al.
Published: (2025)
by: He, Jie, et al.
Published: (2025)
Fully automatic extraction of morphological traits from the Web: utopia or reality?
by: Marcos, Diego, et al.
Published: (2024)
by: Marcos, Diego, et al.
Published: (2024)
LLM_annotate: A Python package for annotating and analyzing fiction characters
by: Rosenbusch, Hannes
Published: (2025)
by: Rosenbusch, Hannes
Published: (2025)
Mitigating Unintended Memorization with LoRA in Federated Learning for LLMs
by: Bossy, Thierry, et al.
Published: (2025)
by: Bossy, Thierry, et al.
Published: (2025)
Causality extraction from medical text using Large Language Models (LLMs)
by: Gopalakrishnan, Seethalakshmi, et al.
Published: (2024)
by: Gopalakrishnan, Seethalakshmi, et al.
Published: (2024)
Can Risk-taking AI-Assistants suitably represent entities
by: Mazyaki, Ali, et al.
Published: (2025)
by: Mazyaki, Ali, et al.
Published: (2025)
Med-gte-hybrid: A contextual embedding transformer model for extracting actionable information from clinical texts
by: Kumar, Aditya, et al.
Published: (2025)
by: Kumar, Aditya, et al.
Published: (2025)
Leveraging large language models for efficient representation learning for entity resolution
by: Xu, Xiaowei, et al.
Published: (2024)
by: Xu, Xiaowei, et al.
Published: (2024)
ks-lit-3m: A 3.1 million word kashmiri text dataset for large language model pretraining
by: Malik, Haq Nawaz
Published: (2026)
by: Malik, Haq Nawaz
Published: (2026)
Scalable multilingual PII annotation for responsible AI in LLMs
by: Meena, Bharti, et al.
Published: (2025)
by: Meena, Bharti, et al.
Published: (2025)
ESG-FTSE: A corpus of news articles with ESG relevance labels and use cases
by: Pavlova, Mariya, et al.
Published: (2024)
by: Pavlova, Mariya, et al.
Published: (2024)
Improving Semantic Understanding in Speech Language Models via Brain-tuning
by: Moussa, Omer, et al.
Published: (2024)
by: Moussa, Omer, et al.
Published: (2024)
ADMEDTAGGER: an annotation framework for distillation of expert knowledge for the Polish medical language
by: Górski, Franciszek, et al.
Published: (2025)
by: Górski, Franciszek, et al.
Published: (2025)
CARMA: Comprehensive Automatically-annotated Reddit Mental Health Dataset for Arabic
by: Mankarious, Saad, et al.
Published: (2025)
by: Mankarious, Saad, et al.
Published: (2025)
Reducing annotator bias by belief elicitation
by: Jakobsen, Terne Sasha Thorn, et al.
Published: (2024)
by: Jakobsen, Terne Sasha Thorn, et al.
Published: (2024)
Beyond speculation: Measuring the growing presence of LLM-generated texts in multilingual disinformation
by: Macko, Dominik, et al.
Published: (2025)
by: Macko, Dominik, et al.
Published: (2025)
Symbol-based entity marker highlighting for enhanced text mining in materials science with generative AI
by: Lee, Junhyeong, et al.
Published: (2025)
by: Lee, Junhyeong, et al.
Published: (2025)
Abusive text transformation using LLMs
by: Chandra, Rohitash, et al.
Published: (2025)
by: Chandra, Rohitash, et al.
Published: (2025)
dafny-annotator: AI-Assisted Verification of Dafny Programs
by: Poesia, Gabriel, et al.
Published: (2024)
by: Poesia, Gabriel, et al.
Published: (2024)
Retrieval augmented generation based dynamic prompting for few-shot biomedical named entity recognition using large language models
by: Ge, Yao, et al.
Published: (2025)
by: Ge, Yao, et al.
Published: (2025)
Benchmark of stylistic variation in LLM-generated texts
by: Milička, Jiří, et al.
Published: (2025)
by: Milička, Jiří, et al.
Published: (2025)
PAGE: Prompt Augmentation for text Generation Enhancement
by: Pacchiotti, Mauro Jose, et al.
Published: (2025)
by: Pacchiotti, Mauro Jose, et al.
Published: (2025)
Serialized EHR make for good text representations
by: Chou, Zhirong, et al.
Published: (2025)
by: Chou, Zhirong, et al.
Published: (2025)
AutoManual: Constructing Instruction Manuals by LLM Agents via Interactive Environmental Learning
by: Chen, Minghao, et al.
Published: (2024)
by: Chen, Minghao, et al.
Published: (2024)
Is 'Hope' a person or an idea? A pilot benchmark for NER: comparing traditional NLP tools and large language models on ambiguous entities
by: Latifi, Payam
Published: (2025)
by: Latifi, Payam
Published: (2025)
Assessing the potential of LLM-assisted annotation for corpus-based pragmatics and discourse analysis: The case of apology
by: Yu, Danni, et al.
Published: (2023)
by: Yu, Danni, et al.
Published: (2023)
ReMeREC: Relation-aware and Multi-entity Referring Expression Comprehension
by: Hu, Yizhi, et al.
Published: (2025)
by: Hu, Yizhi, et al.
Published: (2025)
Beyond checkmate: exploring the creative chokepoints in AI text
by: Tripto, Nafis Irtiza, et al.
Published: (2025)
by: Tripto, Nafis Irtiza, et al.
Published: (2025)
Can professional translators identify machine-generated text?
by: Farrell, Michael
Published: (2026)
by: Farrell, Michael
Published: (2026)
Facilitating phenotyping from clinical texts: the medkit library
by: Neuraz, Antoine, et al.
Published: (2024)
by: Neuraz, Antoine, et al.
Published: (2024)
Similar Items
-
ESNERA: Empirical and semantic named entity alignment for named entity dataset merging
by: Zhang, Xiaobo, et al.
Published: (2025) -
Algorithms For Automatic Accentuation And Transcription Of Russian Texts In Speech Recognition Systems
by: Iakovenko, Olga, et al.
Published: (2024) -
Are generative AI text annotations systematically biased?
by: Stolwijk, Sjoerd B., et al.
Published: (2025) -
MaterioMiner -- An ontology-based text mining dataset for extraction of process-structure-property entities
by: Durmaz, Ali Riza, et al.
Published: (2024) -
Large language models struggle with ethnographic text annotation
by: Goodall, Leonardo S., et al.
Published: (2026)