Semantic similarity estimation for domain specific data using BERT and other techniques
Fuente:
arXiv
Gespeichert in:
| 1. Verfasser: | Prashanth, R. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Leveraging text data for causal inference using electronic health records
von: Mozer, Reagan, et al.
Veröffentlicht: (2023)
von: Mozer, Reagan, et al.
Veröffentlicht: (2023)
A Latent Dirichlet Allocation (LDA) Semantic Text Analytics Approach to Explore Topical Features in Charity Crowdfunding Campaigns
von: Muzumdar, Prathamesh, et al.
Veröffentlicht: (2024)
von: Muzumdar, Prathamesh, et al.
Veröffentlicht: (2024)
SigBERT: Combining Narrative Medical Reports and Rough Path Signature Theory for Survival Risk Estimation in Oncology
von: Minchella, Paul, et al.
Veröffentlicht: (2025)
von: Minchella, Paul, et al.
Veröffentlicht: (2025)
Judging It, Washing It: Scoring and Greenwashing Corporate Climate Disclosures using Large Language Models
von: Chuang, Marianne, et al.
Veröffentlicht: (2025)
von: Chuang, Marianne, et al.
Veröffentlicht: (2025)
The Proxy Presumption: From Semantic Embeddings to Valid Social Measures
von: Li, Baishi, et al.
Veröffentlicht: (2026)
von: Li, Baishi, et al.
Veröffentlicht: (2026)
TransitGPT: A Generative AI-based framework for interacting with GTFS data using Large Language Models
von: Devunuri, Saipraneeth, et al.
Veröffentlicht: (2024)
von: Devunuri, Saipraneeth, et al.
Veröffentlicht: (2024)
Sentiment Informed Sentence BERT-Ensemble Algorithm for Depression Detection
von: Ogunleye, Bayode, et al.
Veröffentlicht: (2024)
von: Ogunleye, Bayode, et al.
Veröffentlicht: (2024)
Semiotic Reconstruction of Destination Expectation Constructs An LLM-Driven Computational Paradigm for Social Media Tourism Analytics
von: Lan, Haotian, et al.
Veröffentlicht: (2025)
von: Lan, Haotian, et al.
Veröffentlicht: (2025)
LAVA: Language Model Assisted Verbal Autopsy for Cause-of-Death Determination
von: Chen, Yiqun T., et al.
Veröffentlicht: (2025)
von: Chen, Yiqun T., et al.
Veröffentlicht: (2025)
Large Language Models for Full-Text Methods Assessment: A Case Study on Mediation Analysis
von: Zhang, Wenqing, et al.
Veröffentlicht: (2025)
von: Zhang, Wenqing, et al.
Veröffentlicht: (2025)
Enhancing Systematic Reviews with Large Language Models: Using GPT-4 and Kimi
von: Kaptur, Dandan Chen, et al.
Veröffentlicht: (2025)
von: Kaptur, Dandan Chen, et al.
Veröffentlicht: (2025)
Does a Large Language Model Really Speak in Human-Like Language?
von: Park, Mose, et al.
Veröffentlicht: (2025)
von: Park, Mose, et al.
Veröffentlicht: (2025)
Gender Inequality in English Textbooks Around the World: an NLP Approach
von: Liu, Tairan
Veröffentlicht: (2025)
von: Liu, Tairan
Veröffentlicht: (2025)
Statistical Multicriteria Evaluation of LLM-Generated Text
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2025)
von: Arias, Esteban Garces, et al.
Veröffentlicht: (2025)
Extracting Emotion Phrases from Tweets using BART
von: Rezapour, Mahdi
Veröffentlicht: (2024)
von: Rezapour, Mahdi
Veröffentlicht: (2024)
Language Markers of Emotion Flexibility Predict Depression and Anxiety Treatment Outcomes
von: Brindle, Benjamin, et al.
Veröffentlicht: (2026)
von: Brindle, Benjamin, et al.
Veröffentlicht: (2026)
Agent Q-Mix: Selecting the Right Action for LLM Multi-Agent Systems through Reinforcement Learning
von: Jiang, Eric Hanchen, et al.
Veröffentlicht: (2026)
von: Jiang, Eric Hanchen, et al.
Veröffentlicht: (2026)
Adding Error Bars to Evals: A Statistical Approach to Language Model Evaluations
von: Miller, Evan
Veröffentlicht: (2024)
von: Miller, Evan
Veröffentlicht: (2024)
Bias in Language Models: Beyond Trick Tests and Toward RUTEd Evaluation
von: Lum, Kristian, et al.
Veröffentlicht: (2024)
von: Lum, Kristian, et al.
Veröffentlicht: (2024)
Sampling the Swadesh List to Identify Similar Languages with Tree Spaces
von: Ordway, Garett, et al.
Veröffentlicht: (2024)
von: Ordway, Garett, et al.
Veröffentlicht: (2024)
Auditing the Use of Language Models to Guide Hiring Decisions
von: Gaebler, Johann D., et al.
Veröffentlicht: (2024)
von: Gaebler, Johann D., et al.
Veröffentlicht: (2024)
Exploring the Comprehension of ChatGPT in Traditional Chinese Medicine Knowledge
von: Yizhen, Li, et al.
Veröffentlicht: (2024)
von: Yizhen, Li, et al.
Veröffentlicht: (2024)
Mind the Unseen Mass: Unmasking LLM Hallucinations via Soft-Hybrid Alphabet Estimation
von: Pan, Hongxing, et al.
Veröffentlicht: (2026)
von: Pan, Hongxing, et al.
Veröffentlicht: (2026)
Still no evidence for an effect of the proportion of non-native speakers on language complexity -- A response to Kauhanen, Einhaus & Walkden (2023)
von: Koplenig, Alexander
Veröffentlicht: (2023)
von: Koplenig, Alexander
Veröffentlicht: (2023)
The Multi-Range Theory of Translation Quality Measurement: MQM scoring models and Statistical Quality Control
von: Lommel, Arle, et al.
Veröffentlicht: (2024)
von: Lommel, Arle, et al.
Veröffentlicht: (2024)
Personalized Prediction of Perceived Message Effectiveness Using Large Language Model Based Digital Twins
von: Han, Jasmin, et al.
Veröffentlicht: (2026)
von: Han, Jasmin, et al.
Veröffentlicht: (2026)
Language Hierarchization Provides the Optimal Solution to Human Working Memory Limits
von: Chen, Luyao, et al.
Veröffentlicht: (2026)
von: Chen, Luyao, et al.
Veröffentlicht: (2026)
Less than one percent of words would be affected by gender-inclusive language in German press texts
von: Müller-Spitzer, Carolin, et al.
Veröffentlicht: (2024)
von: Müller-Spitzer, Carolin, et al.
Veröffentlicht: (2024)
Probing Minimalist Phase Structure in LLMs: What Universal Dependencies Cannot Represent
von: Chen, Yuanhao, et al.
Veröffentlicht: (2026)
von: Chen, Yuanhao, et al.
Veröffentlicht: (2026)
Emotion Detection with Transformers: A Comparative Study
von: Rezapour, Mahdi
Veröffentlicht: (2024)
von: Rezapour, Mahdi
Veröffentlicht: (2024)
A Novel Metric for Measuring the Robustness of Large Language Models in Non-adversarial Scenarios
von: Ackerman, Samuel, et al.
Veröffentlicht: (2024)
von: Ackerman, Samuel, et al.
Veröffentlicht: (2024)
Accurate early detection of Parkinson's disease from SPECT imaging through Convolutional Neural Networks
von: Prashanth, R.
Veröffentlicht: (2024)
von: Prashanth, R.
Veröffentlicht: (2024)
Documents Are People and Words Are Items: A Psychometric Approach to Textual Data with Contextual Embeddings
von: Chen, Jinsong
Veröffentlicht: (2025)
von: Chen, Jinsong
Veröffentlicht: (2025)
Constructing the Truth: Text Mining and Linguistic Networks in Public Hearings of Case 03 of the Special Jurisdiction for Peace (JEP)
von: Sosa, Juan, et al.
Veröffentlicht: (2025)
von: Sosa, Juan, et al.
Veröffentlicht: (2025)
Systematic Evaluation of Uncertainty Estimation Methods in Large Language Models
von: Hobelsberger, Christian, et al.
Veröffentlicht: (2025)
von: Hobelsberger, Christian, et al.
Veröffentlicht: (2025)
Exploring the Potential Role of Generative AI in the TRAPD Procedure for Survey Translation
von: Metheney, Erica Ann, et al.
Veröffentlicht: (2024)
von: Metheney, Erica Ann, et al.
Veröffentlicht: (2024)
Improving Probabilistic Models in Text Classification via Active Learning
von: Bosley, Mitchell, et al.
Veröffentlicht: (2022)
von: Bosley, Mitchell, et al.
Veröffentlicht: (2022)
The Human Flourishing Geographic Index: A County-Level Dataset for the United States, 2013--2023
von: Iacus, Stefano M., et al.
Veröffentlicht: (2025)
von: Iacus, Stefano M., et al.
Veröffentlicht: (2025)
The Cambridge Law Corpus: A Dataset for Legal AI Research
von: Östling, Andreas, et al.
Veröffentlicht: (2023)
von: Östling, Andreas, et al.
Veröffentlicht: (2023)
Assessing biomedical knowledge robustness in large language models by query-efficient sampling attacks
von: Xian, R. Patrick, et al.
Veröffentlicht: (2024)
von: Xian, R. Patrick, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Leveraging text data for causal inference using electronic health records
von: Mozer, Reagan, et al.
Veröffentlicht: (2023) -
A Latent Dirichlet Allocation (LDA) Semantic Text Analytics Approach to Explore Topical Features in Charity Crowdfunding Campaigns
von: Muzumdar, Prathamesh, et al.
Veröffentlicht: (2024) -
SigBERT: Combining Narrative Medical Reports and Rough Path Signature Theory for Survival Risk Estimation in Oncology
von: Minchella, Paul, et al.
Veröffentlicht: (2025) -
Judging It, Washing It: Scoring and Greenwashing Corporate Climate Disclosures using Large Language Models
von: Chuang, Marianne, et al.
Veröffentlicht: (2025) -
The Proxy Presumption: From Semantic Embeddings to Valid Social Measures
von: Li, Baishi, et al.
Veröffentlicht: (2026)