Evaluating Deduplication Techniques for Economic Research Paper Titles with a Focus on Semantic Similarity using NLP and LLMs
Fuente:
arXiv
Saved in:
| Main Authors: | You, Doohee, Fraiberger, S |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
CSS: Contrastive Semantic Similarity for Uncertainty Quantification of LLMs
by: Ao, Shuang, et al.
Published: (2024)
by: Ao, Shuang, et al.
Published: (2024)
Ground Truth Generation for Multilingual Historical NLP using LLMs
by: Gladstone, Clovis, et al.
Published: (2025)
by: Gladstone, Clovis, et al.
Published: (2025)
Trust & Safety of LLMs and LLMs in Trust & Safety
by: You, Doohee, et al.
Published: (2024)
by: You, Doohee, et al.
Published: (2024)
TexIm FAST: Text-to-Image Representation for Semantic Similarity Evaluation using Transformers
by: Ansar, Wazib, et al.
Published: (2024)
by: Ansar, Wazib, et al.
Published: (2024)
Advancing NLP Security by Leveraging LLMs as Adversarial Engines
by: Srinivasan, Sudarshan, et al.
Published: (2024)
by: Srinivasan, Sudarshan, et al.
Published: (2024)
Rethinking Word Similarity: Semantic Similarity through Classification Confusion
by: Zhou, Kaitlyn, et al.
Published: (2025)
by: Zhou, Kaitlyn, et al.
Published: (2025)
TofuEval: Evaluating Hallucinations of LLMs on Topic-Focused Dialogue Summarization
by: Tang, Liyan, et al.
Published: (2024)
by: Tang, Liyan, et al.
Published: (2024)
Efficient Title Reranker for Fast and Improved Knowledge-Intense NLP
by: Chen, Ziyi, et al.
Published: (2023)
by: Chen, Ziyi, et al.
Published: (2023)
Decrypting Cryptic Crosswords: Semantically Complex Wordplay Puzzles as a Target for NLP
by: Rozner, Josh, et al.
Published: (2021)
by: Rozner, Josh, et al.
Published: (2021)
On Behalf of the Stakeholders: Trends in NLP Model Interpretability in the Era of LLMs
by: Calderon, Nitay, et al.
Published: (2024)
by: Calderon, Nitay, et al.
Published: (2024)
Supporting Artifact Evaluation with LLMs: A Study with Published Security Research Papers
by: Heye, David, et al.
Published: (2026)
by: Heye, David, et al.
Published: (2026)
A Comprehensive Framework for Semantic Similarity Analysis of Human and AI-Generated Text Using Transformer Architectures and Ensemble Techniques
by: Gao, Lifu, et al.
Published: (2025)
by: Gao, Lifu, et al.
Published: (2025)
Linguistically Conditioned Semantic Textual Similarity
by: Tu, Jingxuan, et al.
Published: (2024)
by: Tu, Jingxuan, et al.
Published: (2024)
Evaluation Metrics for Text Data Augmentation in NLP
by: Amadeus, Marcellus, et al.
Published: (2024)
by: Amadeus, Marcellus, et al.
Published: (2024)
Select, Label, Evaluate: Active Testing in NLP
by: Purificato, Antonio, et al.
Published: (2026)
by: Purificato, Antonio, et al.
Published: (2026)
Speaking of Language: Reflections on Metalanguage Research in NLP
by: Schneider, Nathan, et al.
Published: (2026)
by: Schneider, Nathan, et al.
Published: (2026)
From Transformers to LLMs: A Systematic Survey of Efficiency Considerations in NLP
by: Ansar, Wazib, et al.
Published: (2024)
by: Ansar, Wazib, et al.
Published: (2024)
Automated Detection of Clinical Entities in Lung and Breast Cancer Reports Using NLP Techniques
by: Moreno-Casanova, J., et al.
Published: (2025)
by: Moreno-Casanova, J., et al.
Published: (2025)
KurdSTS: The Kurdish Semantic Textual Similarity
by: Abdullah, Abdulhady Abas, et al.
Published: (2025)
by: Abdullah, Abdulhady Abas, et al.
Published: (2025)
PaperAsk: A Benchmark for Reliability Evaluation of LLMs in Paper Search and Reading
by: Wu, Yutao, et al.
Published: (2025)
by: Wu, Yutao, et al.
Published: (2025)
PaperBench: Evaluating AI's Ability to Replicate AI Research
by: Starace, Giulio, et al.
Published: (2025)
by: Starace, Giulio, et al.
Published: (2025)
BaichuanSEED: Sharing the Potential of ExtensivE Data Collection and Deduplication by Introducing a Competitive Large Language Model Baseline
by: Dong, Guosheng, et al.
Published: (2024)
by: Dong, Guosheng, et al.
Published: (2024)
The Pitfalls of Publishing in the Age of LLMs: Strange and Surprising Adventures with a High-Impact NLP Journal
by: Verma, Rakesh M., et al.
Published: (2024)
by: Verma, Rakesh M., et al.
Published: (2024)
Enhancing Steganographic Text Extraction: Evaluating the Impact of NLP Models on Accuracy and Semantic Coherence
by: Li, Mingyang, et al.
Published: (2024)
by: Li, Mingyang, et al.
Published: (2024)
Estimating Text Similarity based on Semantic Concept Embeddings
by: der Brück, Tim vor, et al.
Published: (2024)
by: der Brück, Tim vor, et al.
Published: (2024)
Surviving the Unseen: Predictive Defense for Novel Multi-Turn Multimodal Attacks
by: You, Doohee
Published: (2026)
by: You, Doohee
Published: (2026)
Kernel Language Entropy: Fine-grained Uncertainty Quantification for LLMs from Semantic Similarities
by: Nikitin, Alexander, et al.
Published: (2024)
by: Nikitin, Alexander, et al.
Published: (2024)
Towards Explainable Job Title Matching: Leveraging Semantic Textual Relatedness and Knowledge Graphs
by: Zadykian, Vadim, et al.
Published: (2025)
by: Zadykian, Vadim, et al.
Published: (2025)
Machine-Assisted Grading of Nationwide School-Leaving Essay Exams with LLMs and Statistical NLP
by: Karjus, Andres, et al.
Published: (2026)
by: Karjus, Andres, et al.
Published: (2026)
Leveraging KV Similarity for Online Structured Pruning in LLMs
by: Lee, Jungmin, et al.
Published: (2025)
by: Lee, Jungmin, et al.
Published: (2025)
Should We Fine-Tune or RAG? Evaluating Different Techniques to Adapt LLMs for Dialogue
by: Alghisi, Simone, et al.
Published: (2024)
by: Alghisi, Simone, et al.
Published: (2024)
Automatic Design of Semantic Similarity Ensembles Using Grammatical Evolution
by: Martinez-Gil, Jorge
Published: (2023)
by: Martinez-Gil, Jorge
Published: (2023)
How Small Transformation Expose the Weakness of Semantic Similarity Measures
by: Nikiema, Serge Lionel, et al.
Published: (2025)
by: Nikiema, Serge Lionel, et al.
Published: (2025)
Magnitude Matters: a Superior Class of Similarity Metrics for Holistic Semantic Understanding
by: Parupudi, V. S. Raghu
Published: (2025)
by: Parupudi, V. S. Raghu
Published: (2025)
Multi-LLM Thematic Analysis with Dual Reliability Metrics: Combining Cohen's Kappa and Semantic Similarity for Qualitative Research Validation
by: Jain, Nilesh, et al.
Published: (2025)
by: Jain, Nilesh, et al.
Published: (2025)
A Comprehensive Evaluation framework of Alignment Techniques for LLMs
by: Azmat, Muneeza, et al.
Published: (2025)
by: Azmat, Muneeza, et al.
Published: (2025)
Memory-Driven Role-Playing: Evaluation and Enhancement of Persona Knowledge Utilization in LLMs
by: Wang, Kai, et al.
Published: (2026)
by: Wang, Kai, et al.
Published: (2026)
A Survey on Prompting Techniques in LLMs
by: Bhandari, Prabin
Published: (2023)
by: Bhandari, Prabin
Published: (2023)
SCORE: A Semantic Evaluation Framework for Generative Document Parsing
by: Li, Renyu, et al.
Published: (2025)
by: Li, Renyu, et al.
Published: (2025)
ERAS: Evaluating the Robustness of Chinese NLP Models to Morphological Garden Path Errors
by: Li, Qinchan, et al.
Published: (2024)
by: Li, Qinchan, et al.
Published: (2024)
Similar Items
-
CSS: Contrastive Semantic Similarity for Uncertainty Quantification of LLMs
by: Ao, Shuang, et al.
Published: (2024) -
Ground Truth Generation for Multilingual Historical NLP using LLMs
by: Gladstone, Clovis, et al.
Published: (2025) -
Trust & Safety of LLMs and LLMs in Trust & Safety
by: You, Doohee, et al.
Published: (2024) -
TexIm FAST: Text-to-Image Representation for Semantic Similarity Evaluation using Transformers
by: Ansar, Wazib, et al.
Published: (2024) -
Advancing NLP Security by Leveraging LLMs as Adversarial Engines
by: Srinivasan, Sudarshan, et al.
Published: (2024)