EMS: Efficient and Effective Massively Multilingual Sentence Embedding Learning
Fuente:
arXiv
Salvato in:
| Autori principali: | Mao, Zhuoyuan, Chu, Chenhui, Kurohashi, Sadao |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2022
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Language Lives in Sparse Dimensions: Toward Interpretable and Efficient Multilingual Control for Large Language Models
di: Zhong, Chengzhi, et al.
Pubblicazione: (2025)
di: Zhong, Chengzhi, et al.
Pubblicazione: (2025)
Beyond English-Centric LLMs: What Language Do Multilingual Language Models Think in?
di: Zhong, Chengzhi, et al.
Pubblicazione: (2024)
di: Zhong, Chengzhi, et al.
Pubblicazione: (2024)
Assessing Agentic Large Language Models in Multilingual National Bias
di: Liu, Qianying, et al.
Pubblicazione: (2025)
di: Liu, Qianying, et al.
Pubblicazione: (2025)
MELD-ST: An Emotion-aware Speech Translation Dataset
di: Chen, Sirou, et al.
Pubblicazione: (2024)
di: Chen, Sirou, et al.
Pubblicazione: (2024)
Rapidly Developing High-quality Instruction Data and Evaluation Benchmark for Large Language Models with Minimal Human Effort: A Case Study on Japanese
di: Sun, Yikun, et al.
Pubblicazione: (2024)
di: Sun, Yikun, et al.
Pubblicazione: (2024)
A Joint Neural Baseline for Concept, Assertion, and Relation Extraction from Clinical Text
di: Cheng, Fei, et al.
Pubblicazione: (2026)
di: Cheng, Fei, et al.
Pubblicazione: (2026)
RecMind: Japanese Movie Recommendation Dialogue with Seeker's Internal State
di: Kodama, Takashi, et al.
Pubblicazione: (2024)
di: Kodama, Takashi, et al.
Pubblicazione: (2024)
Reformulating Domain Adaptation of Large Language Models as Adapt-Retrieve-Revise: A Case Study on Chinese Legal Domain
di: wan, Zhen, et al.
Pubblicazione: (2023)
di: wan, Zhen, et al.
Pubblicazione: (2023)
Omnilingual SONAR: Cross-Lingual and Cross-Modal Sentence Embeddings Bridging Massively Multilingual Text and Speech
di: Omnilingual SONAR Team, et al.
Pubblicazione: (2026)
di: Omnilingual SONAR Team, et al.
Pubblicazione: (2026)
Adaptative Bilingual Aligning Using Multilingual Sentence Embedding
di: Kraif, Olivier
Pubblicazione: (2024)
di: Kraif, Olivier
Pubblicazione: (2024)
Rethinking Meeting Effectiveness: A Benchmark and Framework for Temporal Fine-grained Automatic Meeting Effectiveness Evaluation
di: Li, Yihang, et al.
Pubblicazione: (2026)
di: Li, Yihang, et al.
Pubblicazione: (2026)
Causal Tree Extraction from Medical Case Reports: A Novel Task for Experts-like Text Comprehension
di: Yahata, Sakiko, et al.
Pubblicazione: (2025)
di: Yahata, Sakiko, et al.
Pubblicazione: (2025)
Multilingual Sentence-T5: Scalable Sentence Encoders for Multilingual Applications
di: Yano, Chihiro, et al.
Pubblicazione: (2024)
di: Yano, Chihiro, et al.
Pubblicazione: (2024)
AcTED: Automatic Acquisition of Typical Event Duration for Semi-supervised Temporal Commonsense QA
di: Virgo, Felix, et al.
Pubblicazione: (2024)
di: Virgo, Felix, et al.
Pubblicazione: (2024)
Reading Comprehension using Entity-based Memory Network
di: Wang, Xun, et al.
Pubblicazione: (2016)
di: Wang, Xun, et al.
Pubblicazione: (2016)
Cross-lingual Embedding Clustering for Hierarchical Softmax in Low-Resource Multilingual Speech Recognition
di: Yang, Zhengdong, et al.
Pubblicazione: (2025)
di: Yang, Zhengdong, et al.
Pubblicazione: (2025)
Datasets for Multilingual Answer Sentence Selection
di: Gabburo, Matteo, et al.
Pubblicazione: (2024)
di: Gabburo, Matteo, et al.
Pubblicazione: (2024)
SONAR-SLT: Multilingual Sign Language Translation via Language-Agnostic Sentence Embedding Supervision
di: Hamidullah, Yasser, et al.
Pubblicazione: (2025)
di: Hamidullah, Yasser, et al.
Pubblicazione: (2025)
Vec2Summ: Text Summarization via Probabilistic Sentence Embeddings
di: Li, Mao, et al.
Pubblicazione: (2025)
di: Li, Mao, et al.
Pubblicazione: (2025)
Reasoning Depth and Environment Complexity: A Controlled Study of RLVR Data Allocation across Logical Reasoning Tasks
di: Zhu, Yihua, et al.
Pubblicazione: (2026)
di: Zhu, Yihua, et al.
Pubblicazione: (2026)
J-CRe3: A Japanese Conversation Dataset for Real-world Reference Resolution
di: Ueda, Nobuhiro, et al.
Pubblicazione: (2024)
di: Ueda, Nobuhiro, et al.
Pubblicazione: (2024)
Memorization, Emergence, and Explaining Reversal Failures: A Controlled Study of Relational Semantics in LLMs
di: Zhu, Yihua, et al.
Pubblicazione: (2026)
di: Zhu, Yihua, et al.
Pubblicazione: (2026)
MMTEB: Massive Multilingual Text Embedding Benchmark
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2025)
di: Enevoldsen, Kenneth, et al.
Pubblicazione: (2025)
Domain Adaptation for Japanese Sentence Embeddings with Contrastive Learning based on Synthetic Sentence Generation
di: Chen, Zihao, et al.
Pubblicazione: (2025)
di: Chen, Zihao, et al.
Pubblicazione: (2025)
SpeechIQ: Speech-Agentic Intelligence Quotient Across Cognitive Levels in Voice Understanding by Large Language Models
di: Wan, Zhen, et al.
Pubblicazione: (2025)
di: Wan, Zhen, et al.
Pubblicazione: (2025)
Combining Discrete Wavelet and Cosine Transforms for Efficient Sentence Embedding
di: Salama, Rana, et al.
Pubblicazione: (2025)
di: Salama, Rana, et al.
Pubblicazione: (2025)
Space Decomposition for Sentence Embedding
di: Ponwitayarat, Wuttikorn, et al.
Pubblicazione: (2024)
di: Ponwitayarat, Wuttikorn, et al.
Pubblicazione: (2024)
Text Simplification with Sentence Embeddings
di: Shardlow, Matthew
Pubblicazione: (2025)
di: Shardlow, Matthew
Pubblicazione: (2025)
Tuning LLMs with Contrastive Alignment Instructions for Machine Translation in Unseen, Low-resource Languages
di: Mao, Zhuoyuan, et al.
Pubblicazione: (2024)
di: Mao, Zhuoyuan, et al.
Pubblicazione: (2024)
BIS Reasoning 1.0: The First Large-Scale Japanese Benchmark for Belief-Inconsistent Syllogistic Reasoning
di: Nguyen, Ha-Thanh, et al.
Pubblicazione: (2025)
di: Nguyen, Ha-Thanh, et al.
Pubblicazione: (2025)
FOCUS: Effective Embedding Initialization for Monolingual Specialization of Multilingual Models
di: Dobler, Konstantin, et al.
Pubblicazione: (2023)
di: Dobler, Konstantin, et al.
Pubblicazione: (2023)
Set-Theoretic Compositionality of Sentence Embeddings
di: Bansal, Naman, et al.
Pubblicazione: (2025)
di: Bansal, Naman, et al.
Pubblicazione: (2025)
Sentence Representations via Gaussian Embedding
di: Yoda, Shohei, et al.
Pubblicazione: (2023)
di: Yoda, Shohei, et al.
Pubblicazione: (2023)
Improving Multimodal Contrastive Learning of Sentence Embeddings with Object-Phrase Alignment
di: Zhao, Kaiyan, et al.
Pubblicazione: (2025)
di: Zhao, Kaiyan, et al.
Pubblicazione: (2025)
Refining Sentence Embedding Model through Ranking Sentences Generation with Large Language Models
di: He, Liyang, et al.
Pubblicazione: (2025)
di: He, Liyang, et al.
Pubblicazione: (2025)
Do LLMs Align Human Values Regarding Social Biases? Judging and Explaining Social Biases with LLMs
di: Liu, Yang, et al.
Pubblicazione: (2025)
di: Liu, Yang, et al.
Pubblicazione: (2025)
Understanding the Prompt Sensitivity
di: Liu, Yang, et al.
Pubblicazione: (2026)
di: Liu, Yang, et al.
Pubblicazione: (2026)
Efficient and Flexible Topic Modeling using Pretrained Embeddings and Bag of Sentences
di: Schneider, Johannes
Pubblicazione: (2023)
di: Schneider, Johannes
Pubblicazione: (2023)
TransAug: Translate as Augmentation for Sentence Embeddings
di: Wang, Jue
Pubblicazione: (2021)
di: Wang, Jue
Pubblicazione: (2021)
Linear Cross-Lingual Mapping of Sentence Embeddings
di: Vasilyev, Oleg, et al.
Pubblicazione: (2023)
di: Vasilyev, Oleg, et al.
Pubblicazione: (2023)
Documenti analoghi
-
Language Lives in Sparse Dimensions: Toward Interpretable and Efficient Multilingual Control for Large Language Models
di: Zhong, Chengzhi, et al.
Pubblicazione: (2025) -
Beyond English-Centric LLMs: What Language Do Multilingual Language Models Think in?
di: Zhong, Chengzhi, et al.
Pubblicazione: (2024) -
Assessing Agentic Large Language Models in Multilingual National Bias
di: Liu, Qianying, et al.
Pubblicazione: (2025) -
MELD-ST: An Emotion-aware Speech Translation Dataset
di: Chen, Sirou, et al.
Pubblicazione: (2024) -
Rapidly Developing High-quality Instruction Data and Evaluation Benchmark for Large Language Models with Minimal Human Effort: A Case Study on Japanese
di: Sun, Yikun, et al.
Pubblicazione: (2024)