Ground Truth Generation for Multilingual Historical NLP using LLMs
Fuente:
arXiv
Guardado en:
| Autores principales: | Gladstone, Clovis, Fang, Zhao, Stewart, Spencer Dean |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
A Comparative Analysis of Word Segmentation, Part-of-Speech Tagging, and Named Entity Recognition for Historical Chinese Sources, 1900-1950
por: Fang, Zhao, et al.
Publicado: (2025)
por: Fang, Zhao, et al.
Publicado: (2025)
CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks
por: Lin, Peiqin, et al.
Publicado: (2026)
por: Lin, Peiqin, et al.
Publicado: (2026)
Tokenization and Representation Biases in Multilingual Models on Dialectal NLP Tasks
por: Kanjirangat, Vani, et al.
Publicado: (2025)
por: Kanjirangat, Vani, et al.
Publicado: (2025)
Collective Reasoning Among LLMs: A Framework for Answer Validation Without Ground Truth
por: Davoudi, Seyed Pouyan Mousavi, et al.
Publicado: (2025)
por: Davoudi, Seyed Pouyan Mousavi, et al.
Publicado: (2025)
Testing the Limits of Truth Directions in LLMs
por: Poulis, Angelos, et al.
Publicado: (2026)
por: Poulis, Angelos, et al.
Publicado: (2026)
AggTruth: Contextual Hallucination Detection using Aggregated Attention Scores in LLMs
por: Matys, Piotr, et al.
Publicado: (2025)
por: Matys, Piotr, et al.
Publicado: (2025)
How Context Shapes Truth: Geometric Transformations of Statement-level Truth Representations in LLMs
por: Adarsh, Shivam, et al.
Publicado: (2026)
por: Adarsh, Shivam, et al.
Publicado: (2026)
Human-Centric NLP or AI-Centric Illusion?: A Critical Investigation
por: Spencer, Piyapath T
Publicado: (2024)
por: Spencer, Piyapath T
Publicado: (2024)
MultiNRC: A Challenging and Native Multilingual Reasoning Evaluation Benchmark for LLMs
por: Fabbri, Alexander R., et al.
Publicado: (2025)
por: Fabbri, Alexander R., et al.
Publicado: (2025)
Truth is Universal: Robust Detection of Lies in LLMs
por: Bürger, Lennart, et al.
Publicado: (2024)
por: Bürger, Lennart, et al.
Publicado: (2024)
Tower+: Bridging Generality and Translation Specialization in Multilingual LLMs
por: Rei, Ricardo, et al.
Publicado: (2025)
por: Rei, Ricardo, et al.
Publicado: (2025)
On the Evaluation Practices in Multilingual NLP: Can Machine Translation Offer an Alternative to Human Translations?
por: Choenni, Rochelle, et al.
Publicado: (2024)
por: Choenni, Rochelle, et al.
Publicado: (2024)
GhanaNLP Parallel Corpora: Comprehensive Multilingual Resources for Low-Resource Ghanaian Languages
por: Gyamfi, Lawrence Adu, et al.
Publicado: (2026)
por: Gyamfi, Lawrence Adu, et al.
Publicado: (2026)
Advancing NLP Security by Leveraging LLMs as Adversarial Engines
por: Srinivasan, Sudarshan, et al.
Publicado: (2024)
por: Srinivasan, Sudarshan, et al.
Publicado: (2024)
Autonomous Evaluation of LLMs for Truth Maintenance and Reasoning Tasks
por: Karia, Rushang, et al.
Publicado: (2024)
por: Karia, Rushang, et al.
Publicado: (2024)
TruthRL: Incentivizing Truthful LLMs via Reinforcement Learning
por: Wei, Zhepei, et al.
Publicado: (2025)
por: Wei, Zhepei, et al.
Publicado: (2025)
Comparative Performance of Advanced NLP Models and LLMs in Multilingual Geo-Entity Detection
por: Kopanov, Kalin
Publicado: (2024)
por: Kopanov, Kalin
Publicado: (2024)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
por: Calderon, Nitay, et al.
Publicado: (2024)
por: Calderon, Nitay, et al.
Publicado: (2024)
Graphing the Truth: Structured Visualizations for Automated Hallucination Detection in LLMs
por: Agrawal, Tanmay
Publicado: (2025)
por: Agrawal, Tanmay
Publicado: (2025)
Reasoning Isn't Enough: Examining Truth-Bias and Sycophancy in LLMs
por: Barkett, Emilio, et al.
Publicado: (2025)
por: Barkett, Emilio, et al.
Publicado: (2025)
Debating with More Persuasive LLMs Leads to More Truthful Answers
por: Khan, Akbir, et al.
Publicado: (2024)
por: Khan, Akbir, et al.
Publicado: (2024)
Fact or Fiction? Can LLMs be Reliable Annotators for Political Truths?
por: Chatrath, Veronica, et al.
Publicado: (2024)
por: Chatrath, Veronica, et al.
Publicado: (2024)
Ground-Truth Subgraphs for Better Training and Evaluation of Knowledge Graph Augmented LLMs
por: Cattaneo, Alberto, et al.
Publicado: (2025)
por: Cattaneo, Alberto, et al.
Publicado: (2025)
Multilingual LLMs Are Not Multilingual Thinkers: Evidence from Hindi Analogy Evaluation
por: Gupta, Ashray, et al.
Publicado: (2025)
por: Gupta, Ashray, et al.
Publicado: (2025)
Evaluating Deduplication Techniques for Economic Research Paper Titles with a Focus on Semantic Similarity using NLP and LLMs
por: You, Doohee, et al.
Publicado: (2024)
por: You, Doohee, et al.
Publicado: (2024)
From Transformers to LLMs: A Systematic Survey of Efficiency Considerations in NLP
por: Ansar, Wazib, et al.
Publicado: (2024)
por: Ansar, Wazib, et al.
Publicado: (2024)
The Consensus Trap: Dissecting Subjectivity and the "Ground Truth" Illusion in Data Annotation
por: Munir, Sheza, et al.
Publicado: (2026)
por: Munir, Sheza, et al.
Publicado: (2026)
Single Ground Truth Is Not Enough: Adding Flexibility to Aspect-Based Sentiment Analysis Evaluation
por: Yang, Soyoung, et al.
Publicado: (2024)
por: Yang, Soyoung, et al.
Publicado: (2024)
Ranking Large Language Models without Ground Truth
por: Dhurandhar, Amit, et al.
Publicado: (2024)
por: Dhurandhar, Amit, et al.
Publicado: (2024)
The Riddle of Reflection: Evaluating Reasoning and Self-Awareness in Multilingual LLMs using Indian Riddles
por: M, Abhinav P, et al.
Publicado: (2025)
por: M, Abhinav P, et al.
Publicado: (2025)
Beyond Catalogue Counts: the Dataset Visibility Asymmetry in Low-Resource Multilingual NLP
por: Tan, Zhiyin, et al.
Publicado: (2026)
por: Tan, Zhiyin, et al.
Publicado: (2026)
A Framework to Assess Multilingual Vulnerabilities of LLMs
por: Tang, Likai, et al.
Publicado: (2025)
por: Tang, Likai, et al.
Publicado: (2025)
Intertwining CP and NLP: The Generation of Unreasonably Constrained Sentences
por: Bonlarron, Alexandre, et al.
Publicado: (2024)
por: Bonlarron, Alexandre, et al.
Publicado: (2024)
Machine-Assisted Grading of Nationwide School-Leaving Essay Exams with LLMs and Statistical NLP
por: Karjus, Andres, et al.
Publicado: (2026)
por: Karjus, Andres, et al.
Publicado: (2026)
TruthStance: An Annotated Dataset of Conversations on Truth Social
por: Ameen, Fathima, et al.
Publicado: (2026)
por: Ameen, Fathima, et al.
Publicado: (2026)
MASE: Interpretable NLP Models via Model-Agnostic Saliency Estimation
por: Yang, Zhou, et al.
Publicado: (2025)
por: Yang, Zhou, et al.
Publicado: (2025)
Evaluation of Multilingual LLMs Personalized Text Generation Capabilities Targeting Groups and Social-Media Platforms
por: Macko, Dominik
Publicado: (2026)
por: Macko, Dominik
Publicado: (2026)
Lessons Without Borders? Evaluating Cultural Alignment of LLMs Using Multilingual Story Moral Generation
por: Wu, Sophie, et al.
Publicado: (2026)
por: Wu, Sophie, et al.
Publicado: (2026)
Multilingual Prompt Engineering in Large Language Models: A Survey Across NLP Tasks
por: Vatsal, Shubham, et al.
Publicado: (2025)
por: Vatsal, Shubham, et al.
Publicado: (2025)
CorefInst: Leveraging LLMs for Multilingual Coreference Resolution
por: Arslan, Tuğba Pamay, et al.
Publicado: (2025)
por: Arslan, Tuğba Pamay, et al.
Publicado: (2025)
Ejemplares similares
-
A Comparative Analysis of Word Segmentation, Part-of-Speech Tagging, and Named Entity Recognition for Historical Chinese Sources, 1900-1950
por: Fang, Zhao, et al.
Publicado: (2025) -
CulturALL: Benchmarking Multilingual and Multicultural Competence of LLMs on Grounded Tasks
por: Lin, Peiqin, et al.
Publicado: (2026) -
Tokenization and Representation Biases in Multilingual Models on Dialectal NLP Tasks
por: Kanjirangat, Vani, et al.
Publicado: (2025) -
Collective Reasoning Among LLMs: A Framework for Answer Validation Without Ground Truth
por: Davoudi, Seyed Pouyan Mousavi, et al.
Publicado: (2025) -
Testing the Limits of Truth Directions in LLMs
por: Poulis, Angelos, et al.
Publicado: (2026)