Spoken Word2Vec: Learning Skipgram Embeddings from Speech
Fuente:
arXiv
Saved in:
| Main Authors: | Sayeed, Mohammad Amaan, Aldarmaki, Hanan |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SparQLe: Speech Queries to Text Translation Through LLMs
by: Djanibekov, Amirbek, et al.
Published: (2025)
by: Djanibekov, Amirbek, et al.
Published: (2025)
Linear Semantic Segmentation for Low-Resource Spoken Dialects
by: Chirkunov, Kirill, et al.
Published: (2026)
by: Chirkunov, Kirill, et al.
Published: (2026)
STTATTS: Unified Speech-To-Text And Text-To-Speech Model
by: Toyin, Hawau Olamide, et al.
Published: (2024)
by: Toyin, Hawau Olamide, et al.
Published: (2024)
Personal Attribute Leakage in Federated Speech Models
by: Al-Ali, Hamdan, et al.
Published: (2025)
by: Al-Ali, Hamdan, et al.
Published: (2025)
Are LLMs Good Text Diacritizers? An Arabic and Yoruba Case Study
by: Toyin, Hawau Olamide, et al.
Published: (2025)
by: Toyin, Hawau Olamide, et al.
Published: (2025)
Morphemes Without Borders: Evaluating Root-Pattern Morphology in Arabic Tokenizers and LLMs
by: Alakeel, Yara, et al.
Published: (2026)
by: Alakeel, Yara, et al.
Published: (2026)
ArVoice: A Multi-Speaker Dataset for Arabic Speech Synthesis
by: Toyin, Hawau Olamide, et al.
Published: (2025)
by: Toyin, Hawau Olamide, et al.
Published: (2025)
Mixat: A Data Set of Bilingual Emirati-English Speech
by: Ali, Maryam Al, et al.
Published: (2024)
by: Ali, Maryam Al, et al.
Published: (2024)
SpokenWOZ: A Large-Scale Speech-Text Benchmark for Spoken Task-Oriented Dialogue Agents
by: Si, Shuzheng, et al.
Published: (2023)
by: Si, Shuzheng, et al.
Published: (2023)
Causal2Vec: Improving Decoder-only LLMs as Embedding Models through a Contextual Token
by: Lin, Ailiang, et al.
Published: (2025)
by: Lin, Ailiang, et al.
Published: (2025)
Joint Learning of Context and Feedback Embeddings in Spoken Dialogue
by: Qian, Livia, et al.
Published: (2024)
by: Qian, Livia, et al.
Published: (2024)
Automatic Restoration of Diacritics for Speech Data Sets
by: Shatnawi, Sara, et al.
Published: (2023)
by: Shatnawi, Sara, et al.
Published: (2023)
Optimal Transport Regularization for Speech Text Alignment in Spoken Language Models
by: Xu, Wenze, et al.
Published: (2025)
by: Xu, Wenze, et al.
Published: (2025)
A Comparative Analysis of Bilingual and Trilingual Wav2Vec Models for Automatic Speech Recognition in Multilingual Oral History Archives
by: Lehečka, Jan, et al.
Published: (2024)
by: Lehečka, Jan, et al.
Published: (2024)
VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks
by: Jiang, Ziyan, et al.
Published: (2024)
by: Jiang, Ziyan, et al.
Published: (2024)
Chem42: a Family of chemical Language Models for Target-aware Ligand Generation
by: Singh, Aahan, et al.
Published: (2025)
by: Singh, Aahan, et al.
Published: (2025)
Prot42: a Novel Family of Protein Language Models for Target-aware Protein Binder Generation
by: Sayeed, Mohammad Amaan, et al.
Published: (2025)
by: Sayeed, Mohammad Amaan, et al.
Published: (2025)
JEEM: Vision-Language Understanding in Four Arabic Dialects
by: Kadaoui, Karima, et al.
Published: (2025)
by: Kadaoui, Karima, et al.
Published: (2025)
E2Vec: Feature Embedding with Temporal Information for Analyzing Student Actions in E-Book Systems
by: Miyazaki, Yuma, et al.
Published: (2024)
by: Miyazaki, Yuma, et al.
Published: (2024)
TriNER: A Series of Named Entity Recognition Models For Hindi, Bengali & Marathi
by: Dhamaskar, Mohammed Amaan, et al.
Published: (2025)
by: Dhamaskar, Mohammed Amaan, et al.
Published: (2025)
LLM2Vec: Large Language Models Are Secretly Powerful Text Encoders
by: BehnamGhader, Parishad, et al.
Published: (2024)
by: BehnamGhader, Parishad, et al.
Published: (2024)
A Comparative Analysis of Static Word Embeddings for Hungarian
by: Gedeon, Máté
Published: (2025)
by: Gedeon, Máté
Published: (2025)
A Comprehensive Analysis of Static Word Embeddings for Turkish
by: Sarıtaş, Karahan, et al.
Published: (2024)
by: Sarıtaş, Karahan, et al.
Published: (2024)
Clinical Annotations for Automatic Stuttering Severity Assessment
by: Valente, Ana Rita, et al.
Published: (2025)
by: Valente, Ana Rita, et al.
Published: (2025)
Exploring ASR-Based Wav2Vec2 for Automated Speech Disorder Assessment: Insights and Analysis
by: Nguyen, Tuan, et al.
Published: (2024)
by: Nguyen, Tuan, et al.
Published: (2024)
CWTM: Leveraging Contextualized Word Embeddings from BERT for Neural Topic Modeling
by: Fang, Zheng, et al.
Published: (2023)
by: Fang, Zheng, et al.
Published: (2023)
Categorical Classification of Book Summaries Using Word Embedding Techniques
by: Keskin, Kerem, et al.
Published: (2025)
by: Keskin, Kerem, et al.
Published: (2025)
DiffuSpeech: Silent Thought, Spoken Answer via Unified Speech-Text Diffusion
by: Lou, Yuxuan, et al.
Published: (2026)
by: Lou, Yuxuan, et al.
Published: (2026)
Unrequited Emotions: Investigating the Gaps in Motivation and Practice in Speech Emotion Recognition Research
by: Wong, Taryn, et al.
Published: (2026)
by: Wong, Taryn, et al.
Published: (2026)
Word Embeddings Are Steers for Language Models
by: Han, Chi, et al.
Published: (2023)
by: Han, Chi, et al.
Published: (2023)
HC$^2$L: Hybrid and Cooperative Contrastive Learning for Cross-lingual Spoken Language Understanding
by: Xing, Bowen, et al.
Published: (2024)
by: Xing, Bowen, et al.
Published: (2024)
Spoken Grammar Assessment Using LLM
by: Kopparapu, Sunil Kumar, et al.
Published: (2024)
by: Kopparapu, Sunil Kumar, et al.
Published: (2024)
WESR: Scaling and Evaluating Word-level Event-Speech Recognition
by: Yang, Chenchen, et al.
Published: (2026)
by: Yang, Chenchen, et al.
Published: (2026)
Contrastive and Consistency Learning for Neural Noisy-Channel Model in Spoken Language Understanding
by: Kim, Suyoung, et al.
Published: (2024)
by: Kim, Suyoung, et al.
Published: (2024)
Static Word Embeddings for Sentence Semantic Representation
by: Wada, Takashi, et al.
Published: (2025)
by: Wada, Takashi, et al.
Published: (2025)
Interpretable Topic Extraction and Word Embedding Learning using row-stochastic DEDICOM
by: Hillebrand, Lars, et al.
Published: (2025)
by: Hillebrand, Lars, et al.
Published: (2025)
Demographic Attributes Prediction from Speech Using WavLM Embeddings
by: Yang, Yuchen, et al.
Published: (2025)
by: Yang, Yuchen, et al.
Published: (2025)
Survey of Pseudonymization, Abstractive Summarization & Spell Checker for Hindi and Marathi
by: Ransing, Rasika, et al.
Published: (2024)
by: Ransing, Rasika, et al.
Published: (2024)
Small LLMs Do Not Learn a Generalizable Theory of Mind via Reinforcement Learning
by: Sarangi, Sneheel, et al.
Published: (2025)
by: Sarangi, Sneheel, et al.
Published: (2025)
MoSECroT: Model Stitching with Static Word Embeddings for Crosslingual Zero-shot Transfer
by: Ye, Haotian, et al.
Published: (2024)
by: Ye, Haotian, et al.
Published: (2024)
Similar Items
-
SparQLe: Speech Queries to Text Translation Through LLMs
by: Djanibekov, Amirbek, et al.
Published: (2025) -
Linear Semantic Segmentation for Low-Resource Spoken Dialects
by: Chirkunov, Kirill, et al.
Published: (2026) -
STTATTS: Unified Speech-To-Text And Text-To-Speech Model
by: Toyin, Hawau Olamide, et al.
Published: (2024) -
Personal Attribute Leakage in Federated Speech Models
by: Al-Ali, Hamdan, et al.
Published: (2025) -
Are LLMs Good Text Diacritizers? An Arabic and Yoruba Case Study
by: Toyin, Hawau Olamide, et al.
Published: (2025)