SENSE models: an open source solution for multilingual and multimodal semantic-based tasks
Fuente:
arXiv
Saved in:
| Main Authors: | Mdhaffar, Salima, Elleuch, Haroun, Chellaf, Chaimae, Nguyen, Ha, Estève, Yannick |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Using Multimodal and Language-Agnostic Sentence Embeddings for Abstractive Summarization
by: Chellaf, Chaimae, et al.
Published: (2026)
by: Chellaf, Chaimae, et al.
Published: (2026)
ADI-20: Arabic Dialect Identification dataset and models
by: Elleuch, Haroun, et al.
Published: (2025)
by: Elleuch, Haroun, et al.
Published: (2025)
SLURP-TN : Resource for Tunisian Dialect Spoken Language Understanding
by: Elleuch, Haroun, et al.
Published: (2026)
by: Elleuch, Haroun, et al.
Published: (2026)
TEDxTN: A Three-way Speech Translation Corpus for Code-Switched Tunisian Arabic - English
by: Bougares, Fethi, et al.
Published: (2025)
by: Bougares, Fethi, et al.
Published: (2025)
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
by: Mdhaffar, Salima, et al.
Published: (2024)
by: Mdhaffar, Salima, et al.
Published: (2024)
ELYADATA & LIA at NADI 2025: ASR and ADI Subtasks
by: Elleuch, Haroun, et al.
Published: (2025)
by: Elleuch, Haroun, et al.
Published: (2025)
Ara-Best-RQ: Multi Dialectal Arabic SSL
by: Elleuch, Haroun, et al.
Published: (2026)
by: Elleuch, Haroun, et al.
Published: (2026)
Learning Multiple Utterance-Level Attribute Representations with a Unified Speech Encoder
by: Bouziane, Maryem, et al.
Published: (2026)
by: Bouziane, Maryem, et al.
Published: (2026)
A dual task learning approach to fine-tune a multilingual semantic speech encoder for Spoken Language Understanding
by: Laperrière, Gaëlle, et al.
Published: (2024)
by: Laperrière, Gaëlle, et al.
Published: (2024)
CV-18 NER: Augmented Common Voice for Named Entity Recognition from Arabic Speech
by: Saidi, Youssef, et al.
Published: (2026)
by: Saidi, Youssef, et al.
Published: (2026)
Sonos Voice Control Bias Assessment Dataset: A Methodology for Demographic Bias Assessment in Voice Assistants
by: Sekkat, Chloé, et al.
Published: (2024)
by: Sekkat, Chloé, et al.
Published: (2024)
In-domain SSL pre-training and streaming ASR
by: Duret, Jarod, et al.
Published: (2025)
by: Duret, Jarod, et al.
Published: (2025)
OpenStaxQA: A multilingual dataset based on open-source college textbooks
by: Gupta, Pranav
Published: (2025)
by: Gupta, Pranav
Published: (2025)
Semantic enrichment towards efficient speech representations
by: Laperrière, Gaëlle, et al.
Published: (2023)
by: Laperrière, Gaëlle, et al.
Published: (2023)
DHPLT: large-scale multilingual diachronic corpora and word representations for semantic change modelling
by: Fedorova, Mariia, et al.
Published: (2026)
by: Fedorova, Mariia, et al.
Published: (2026)
WorldMedQA-V: a multilingual, multimodal medical examination dataset for multimodal language models evaluation
by: Matos, João, et al.
Published: (2024)
by: Matos, João, et al.
Published: (2024)
Disentangling concept semantics via multilingual averaging in Sparse Autoencoders
by: O'Reilly, Cliff, et al.
Published: (2025)
by: O'Reilly, Cliff, et al.
Published: (2025)
An Ultra-Low Latency, End-to-End Streaming Speech Synthesis Architecture via Block-Wise Generation and Depth-Wise Codec Decoding
by: Su, Tianhui, et al.
Published: (2026)
by: Su, Tianhui, et al.
Published: (2026)
A multimodal multiplex of the mental lexicon for multilingual individuals
by: Huynh, Maria, et al.
Published: (2025)
by: Huynh, Maria, et al.
Published: (2025)
FLiP: Towards understanding and interpreting multimodal multilingual sentence embeddings
by: Kesiraju, Santosh, et al.
Published: (2026)
by: Kesiraju, Santosh, et al.
Published: (2026)
Neuron Specialization: Leveraging intrinsic task modularity for multilingual machine translation
by: Tan, Shaomu, et al.
Published: (2024)
by: Tan, Shaomu, et al.
Published: (2024)
Cross-linguistic disagreement as a conflict of semantic alignment norms in multilingual AI~Linguistic Diversity as a Problem for Philosophy, Cognitive Science, and AI~
by: Mizumoto, Masaharu, et al.
Published: (2025)
by: Mizumoto, Masaharu, et al.
Published: (2025)
The correlation between nativelike selection and prototypicality: a multilingual onomasiological case study using semantic embedding
by: Zhang, Huasheng
Published: (2024)
by: Zhang, Huasheng
Published: (2024)
Automatic register identification for the open web using multilingual deep learning
by: Henriksson, Erik, et al.
Published: (2024)
by: Henriksson, Erik, et al.
Published: (2024)
Morphosyntactic probing of multilingual BERT models
by: Acs, Judit, et al.
Published: (2023)
by: Acs, Judit, et al.
Published: (2023)
Investigating Low-Cost LLM Annotation for~Spoken Dialogue Understanding Datasets
by: Druart, Lucas, et al.
Published: (2024)
by: Druart, Lucas, et al.
Published: (2024)
Open Implementation and Study of BEST-RQ for Speech Processing
by: Whetten, Ryan, et al.
Published: (2024)
by: Whetten, Ryan, et al.
Published: (2024)
Analyzing Speech Unit Selection for Textless Speech-to-Speech Translation
by: Duret, Jarod, et al.
Published: (2024)
by: Duret, Jarod, et al.
Published: (2024)
LeBenchmark 2.0: a Standardized, Replicable and Enhanced Framework for Self-supervised Representations of French Speech
by: Parcollet, Titouan, et al.
Published: (2023)
by: Parcollet, Titouan, et al.
Published: (2023)
Is one brick enough to break the wall of spoken dialogue state tracking?
by: Druart, Lucas, et al.
Published: (2023)
by: Druart, Lucas, et al.
Published: (2023)
Not all ANIMALs are equal: metaphorical framing through source domains and semantic frames
by: Otmakhova, Yulia, et al.
Published: (2026)
by: Otmakhova, Yulia, et al.
Published: (2026)
EuroGEST: Investigating gender stereotypes in multilingual language models
by: Rowe, Jacqueline, et al.
Published: (2025)
by: Rowe, Jacqueline, et al.
Published: (2025)
A study of Vietnamese readability assessing through semantic and statistical features
by: Le, Hung Tuan, et al.
Published: (2024)
by: Le, Hung Tuan, et al.
Published: (2024)
Enhancing Entity Aware Machine Translation with Multi-task Learning
by: Trieu, An, et al.
Published: (2025)
by: Trieu, An, et al.
Published: (2025)
Large language models in materials science and the need for open-source approaches
by: Yang, Fengxu, et al.
Published: (2025)
by: Yang, Fengxu, et al.
Published: (2025)
A semantic embedding space based on large language models for modelling human beliefs
by: Lee, Byunghwee, et al.
Published: (2024)
by: Lee, Byunghwee, et al.
Published: (2024)
Prompting open-source and commercial language models for grammatical error correction of English learner text
by: Davis, Christopher, et al.
Published: (2024)
by: Davis, Christopher, et al.
Published: (2024)
Effective vocabulary expanding of multilingual language models for extremely low-resource languages
by: Zheng, Jianyu
Published: (2026)
by: Zheng, Jianyu
Published: (2026)
Team LA at SCIDOCA shared task 2025: Citation Discovery via relation-based zero-shot retrieval
by: An, Trieu, et al.
Published: (2025)
by: An, Trieu, et al.
Published: (2025)
SENSE: Semantic Embedding Navigation with Soft-gated Evaluation for Retrieval-based Speculative Decoding
by: Chen, Shaowen, et al.
Published: (2026)
by: Chen, Shaowen, et al.
Published: (2026)
Similar Items
-
Using Multimodal and Language-Agnostic Sentence Embeddings for Abstractive Summarization
by: Chellaf, Chaimae, et al.
Published: (2026) -
ADI-20: Arabic Dialect Identification dataset and models
by: Elleuch, Haroun, et al.
Published: (2025) -
SLURP-TN : Resource for Tunisian Dialect Spoken Language Understanding
by: Elleuch, Haroun, et al.
Published: (2026) -
TEDxTN: A Three-way Speech Translation Corpus for Code-Switched Tunisian Arabic - English
by: Bougares, Fethi, et al.
Published: (2025) -
Performance Analysis of Speech Encoders for Low-Resource SLU and ASR in Tunisian Dialect
by: Mdhaffar, Salima, et al.
Published: (2024)