Extracting domain-specific terms using contextual word embeddings
Fuente:
arXiv
Saved in:
| Main Authors: | Repar, Andraž, Lavrač, Nada, Pollak, Senja |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Evaluating and explaining training strategies for zero-shot cross-lingual news sentiment analysis
by: Andrenšek, Luka, et al.
Published: (2024)
by: Andrenšek, Luka, et al.
Published: (2024)
From Symbolic to Neural and Back: Exploring Knowledge Graph-Large Language Model Synergies
by: Škrlj, Blaž, et al.
Published: (2025)
by: Škrlj, Blaž, et al.
Published: (2025)
Incremental Graph Construction Enables Robust Spectral Clustering of Texts
by: Pranjić, Marko, et al.
Published: (2026)
by: Pranjić, Marko, et al.
Published: (2026)
Multi-Task Learning for Features Extraction in Financial Annual Reports
by: Montariol, Syrielle, et al.
Published: (2024)
by: Montariol, Syrielle, et al.
Published: (2024)
A Computational Framework to Identify Self-Aspects in Text
by: Caporusso, Jaya, et al.
Published: (2025)
by: Caporusso, Jaya, et al.
Published: (2025)
Transformer verbatim in-context retrieval across time and scale
by: Armeni, Kristijan, et al.
Published: (2024)
by: Armeni, Kristijan, et al.
Published: (2024)
Semantics or spelling? Probing contextual word embeddings with orthographic noise
by: Matthews, Jacob A., et al.
Published: (2024)
by: Matthews, Jacob A., et al.
Published: (2024)
FuDoBa: Fusing Document and Knowledge Graph-based Representations with Bayesian Optimisation
by: Koloski, Boshko, et al.
Published: (2025)
by: Koloski, Boshko, et al.
Published: (2025)
A graph-based analysis of semantic types and coercion in contextualized word embeddings
by: Chen, Long, et al.
Published: (2026)
by: Chen, Long, et al.
Published: (2026)
The "Right" Discourse on Migration: Analysing Migration-Related Tweets in Right and Far-Right Political Movements
by: Chatterjee, Nishan, et al.
Published: (2025)
by: Chatterjee, Nishan, et al.
Published: (2025)
AutoML-guided Fusion of Entity and LLM-based Representations for Document Classification
by: Koloski, Boshko, et al.
Published: (2024)
by: Koloski, Boshko, et al.
Published: (2024)
Tracking Semantic Change in Slovene: A Novel Dataset and Optimal Transport-Based Distance
by: Pranjić, Marko, et al.
Published: (2024)
by: Pranjić, Marko, et al.
Published: (2024)
Measuring Catastrophic Forgetting in Cross-Lingual Transfer Paradigms: Exploring Tuning Strategies
by: Koloski, Boshko, et al.
Published: (2023)
by: Koloski, Boshko, et al.
Published: (2023)
Recent Advances and Future Directions in Literature-Based Discovery
by: Kastrin, Andrej, et al.
Published: (2025)
by: Kastrin, Andrej, et al.
Published: (2025)
SEKE: Specialised Experts for Keyword Extraction
by: Martinc, Matej, et al.
Published: (2024)
by: Martinc, Matej, et al.
Published: (2024)
A Computational Analysis of the Dehumanisation of Migrants from Syria and Ukraine in Slovene News Media
by: Caporusso, Jaya, et al.
Published: (2024)
by: Caporusso, Jaya, et al.
Published: (2024)
Environmental, Social and Governance Sentiment Analysis on Slovene News: A Novel Dataset and Models
by: Dodig, Paula, et al.
Published: (2026)
by: Dodig, Paula, et al.
Published: (2026)
Does mBERT understand Romansh? Evaluating word embeddings using word alignment
by: Dolev, Eyal Liron
Published: (2023)
by: Dolev, Eyal Liron
Published: (2023)
Make Literature-Based Discovery Great Again through Reproducible Pipelines
by: Cestnik, Bojan, et al.
Published: (2025)
by: Cestnik, Bojan, et al.
Published: (2025)
Effect of dimensionality change on the bias of word embeddings
by: Rai, Rohit Raj, et al.
Published: (2023)
by: Rai, Rohit Raj, et al.
Published: (2023)
Multilingual Cognitive Impairment Detection in the Era of Foundation Models
by: Hoogland, Damar, et al.
Published: (2026)
by: Hoogland, Damar, et al.
Published: (2026)
Semantic Properties of cosine based bias scores for word embeddings
by: Schröder, Sarah, et al.
Published: (2024)
by: Schröder, Sarah, et al.
Published: (2024)
The SAME score: Improved cosine based bias score for word embeddings
by: Schröder, Sarah, et al.
Published: (2022)
by: Schröder, Sarah, et al.
Published: (2022)
Predicting drug-gene relations via analogy tasks with word embeddings
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
by: Yamagiwa, Hiroaki, et al.
Published: (2024)
Explaining word embeddings with perfect fidelity: Case study in research impact prediction
by: Dvorackova, Lucie, et al.
Published: (2024)
by: Dvorackova, Lucie, et al.
Published: (2024)
Multilingual acoustic word embeddings for zero-resource languages
by: Jacobs, Christiaan
Published: (2024)
by: Jacobs, Christiaan
Published: (2024)
From communities to interpretable network and word embedding: an unified approach
by: Prouteau, Thibault, et al.
Published: (2024)
by: Prouteau, Thibault, et al.
Published: (2024)
CoastTerm: a Corpus for Multidisciplinary Term Extraction in Coastal Scientific Literature
by: Delaunay, Julien, et al.
Published: (2024)
by: Delaunay, Julien, et al.
Published: (2024)
A new kid on the block: Distributional semantics predicts the word-specific tone signatures of monosyllabic words in conversational Taiwan Mandarin
by: Jin, Xiaoyun, et al.
Published: (2025)
by: Jin, Xiaoyun, et al.
Published: (2025)
When can isotropy help adapt LLMs' next word prediction to numerical domains?
by: Shelim, Rashed, et al.
Published: (2025)
by: Shelim, Rashed, et al.
Published: (2025)
Semantic similarity estimation for domain specific data using BERT and other techniques
by: Prashanth, R.
Published: (2025)
by: Prashanth, R.
Published: (2025)
Topic-Conversation Relevance (TCR) Dataset and Benchmarks
by: Fan, Yaran, et al.
Published: (2024)
by: Fan, Yaran, et al.
Published: (2024)
Med-gte-hybrid: A contextual embedding transformer model for extracting actionable information from clinical texts
by: Kumar, Aditya, et al.
Published: (2025)
by: Kumar, Aditya, et al.
Published: (2025)
Are we describing the same sound? An analysis of word embedding spaces of expressive piano performance
by: Peter, Silvan David, et al.
Published: (2023)
by: Peter, Silvan David, et al.
Published: (2023)
Targeted control of fast prototyping through domain-specific interface
by: Shi, Yu-Zhe, et al.
Published: (2025)
by: Shi, Yu-Zhe, et al.
Published: (2025)
Injecting Wiktionary to improve token-level contextual representations using contrastive learning
by: Mosolova, Anna, et al.
Published: (2024)
by: Mosolova, Anna, et al.
Published: (2024)
What is a word?
by: Murphy, Elliot
Published: (2024)
by: Murphy, Elliot
Published: (2024)
Introduction of a novel word embedding approach based on technology labels extracted from patent data
by: Standke, Mark, et al.
Published: (2021)
by: Standke, Mark, et al.
Published: (2021)
Toward domain-specific machine translation and quality estimation systems
by: Sharami, Javad Pourmostafa Roshan
Published: (2026)
by: Sharami, Javad Pourmostafa Roshan
Published: (2026)
FAQ-Gen: An automated system to generate domain-specific FAQs to aid content comprehension
by: Kale, Sahil, et al.
Published: (2024)
by: Kale, Sahil, et al.
Published: (2024)
Similar Items
-
Evaluating and explaining training strategies for zero-shot cross-lingual news sentiment analysis
by: Andrenšek, Luka, et al.
Published: (2024) -
From Symbolic to Neural and Back: Exploring Knowledge Graph-Large Language Model Synergies
by: Škrlj, Blaž, et al.
Published: (2025) -
Incremental Graph Construction Enables Robust Spectral Clustering of Texts
by: Pranjić, Marko, et al.
Published: (2026) -
Multi-Task Learning for Features Extraction in Financial Annual Reports
by: Montariol, Syrielle, et al.
Published: (2024) -
A Computational Framework to Identify Self-Aspects in Text
by: Caporusso, Jaya, et al.
Published: (2025)