jina-embeddings-v5-text: Task-Targeted Embedding Distillation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Akram, Mohammad Kalim, Sturua, Saba, Havriushenko, Nastia, Herreros, Quentin, Günther, Michael, Werk, Maximilian, Xiao, Han |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
jina-embeddings-v5-omni: Geometry-preserving Embeddings via Locked Aligned Towers
von: Hönicke, Florian, et al.
Veröffentlicht: (2026)
von: Hönicke, Florian, et al.
Veröffentlicht: (2026)
jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval
von: Günther, Michael, et al.
Veröffentlicht: (2025)
von: Günther, Michael, et al.
Veröffentlicht: (2025)
jina-embeddings-v3: Multilingual Embeddings With Task LoRA
von: Sturua, Saba, et al.
Veröffentlicht: (2024)
von: Sturua, Saba, et al.
Veröffentlicht: (2024)
jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
Jina Embeddings 2: 8192-Token General-Purpose Text Embeddings for Long Documents
von: Günther, Michael, et al.
Veröffentlicht: (2023)
von: Günther, Michael, et al.
Veröffentlicht: (2023)
Jina-ColBERT-v2: A General-Purpose Multilingual Late Interaction Retriever
von: Jha, Rohan, et al.
Veröffentlicht: (2024)
von: Jha, Rohan, et al.
Veröffentlicht: (2024)
Multi-Task Contrastive Learning for 8192-Token Bilingual Text Embeddings
von: Mohr, Isabelle, et al.
Veröffentlicht: (2024)
von: Mohr, Isabelle, et al.
Veröffentlicht: (2024)
Efficient Code Embeddings from Code Generation Models
von: Kryvosheieva, Daria, et al.
Veröffentlicht: (2025)
von: Kryvosheieva, Daria, et al.
Veröffentlicht: (2025)
Jina CLIP: Your CLIP Model Is Also Your Text Retriever
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024)
jina-reranker-v3: Last but Not Late Interaction for Listwise Document Reranking
von: Wang, Feng, et al.
Veröffentlicht: (2025)
von: Wang, Feng, et al.
Veröffentlicht: (2025)
jina-vlm: Small Multilingual Vision Language Model
von: Koukounas, Andreas, et al.
Veröffentlicht: (2025)
von: Koukounas, Andreas, et al.
Veröffentlicht: (2025)
Detecting text level intellectual influence with knowledge graph embeddings
von: Li, Lucian, et al.
Veröffentlicht: (2024)
von: Li, Lucian, et al.
Veröffentlicht: (2024)
Late Chunking: Contextual Chunk Embeddings Using Long-Context Embedding Models
von: Günther, Michael, et al.
Veröffentlicht: (2024)
von: Günther, Michael, et al.
Veröffentlicht: (2024)
SpeechMapper: Speech-to-text Embedding Projector for LLMs
von: Mohapatra, Biswesh, et al.
Veröffentlicht: (2026)
von: Mohapatra, Biswesh, et al.
Veröffentlicht: (2026)
Scalable and consistent few-shot classification of survey responses using text embeddings
von: Mjaaland, Jonas Timmann, et al.
Veröffentlicht: (2025)
von: Mjaaland, Jonas Timmann, et al.
Veröffentlicht: (2025)
Cropping outperforms dropout as an augmentation strategy for self-supervised training of text embeddings
von: González-Márquez, Rita, et al.
Veröffentlicht: (2025)
von: González-Márquez, Rita, et al.
Veröffentlicht: (2025)
Embedding Inversion via Conditional Masked Diffusion Language Models
von: Xiao, Han
Veröffentlicht: (2026)
von: Xiao, Han
Veröffentlicht: (2026)
Targeted Distillation for Sentiment Analysis
von: Zhang, Yice, et al.
Veröffentlicht: (2025)
von: Zhang, Yice, et al.
Veröffentlicht: (2025)
BanglaEmbed: Efficient Sentence Embedding Models for a Low-Resource Language Using Cross-Lingual Distillation Techniques
von: Kabir, Muhammad Rafsan, et al.
Veröffentlicht: (2024)
von: Kabir, Muhammad Rafsan, et al.
Veröffentlicht: (2024)
Conan-embedding: General Text Embedding with More and Better Negative Samples
von: Li, Shiyu, et al.
Veröffentlicht: (2024)
von: Li, Shiyu, et al.
Veröffentlicht: (2024)
Analyzing Political Bias in LLMs via Target-Oriented Sentiment Classification
von: Elbouanani, Akram, et al.
Veröffentlicht: (2025)
von: Elbouanani, Akram, et al.
Veröffentlicht: (2025)
Benchmarking pre-trained text embedding models in aligning built asset information
von: Shahinmoghadam, Mehrzad, et al.
Veröffentlicht: (2024)
von: Shahinmoghadam, Mehrzad, et al.
Veröffentlicht: (2024)
Multi-Sense Embeddings for Language Models and Knowledge Distillation
von: Wang, Qitong, et al.
Veröffentlicht: (2025)
von: Wang, Qitong, et al.
Veröffentlicht: (2025)
Recent advances in text embedding: A Comprehensive Review of Top-Performing Methods on the MTEB Benchmark
von: Cao, Hongliu
Veröffentlicht: (2024)
von: Cao, Hongliu
Veröffentlicht: (2024)
Conan-Embedding-v2: Training an LLM from Scratch for Text Embeddings
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
von: Li, Shiyu, et al.
Veröffentlicht: (2025)
Med-gte-hybrid: A contextual embedding transformer model for extracting actionable information from clinical texts
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
von: Kumar, Aditya, et al.
Veröffentlicht: (2025)
Test-Time Compute for Frozen Embedding Models through Agentic Program Search
von: Xiao, Han
Veröffentlicht: (2026)
von: Xiao, Han
Veröffentlicht: (2026)
ATEB: Evaluating and Improving Advanced NLP Tasks for Text Embedding Models
von: Han, Simeng, et al.
Veröffentlicht: (2025)
von: Han, Simeng, et al.
Veröffentlicht: (2025)
Adaptation of Embedding Models to Financial Filings via LLM Distillation
von: Brenner, Eliot, et al.
Veröffentlicht: (2025)
von: Brenner, Eliot, et al.
Veröffentlicht: (2025)
M3-Embedding: Multi-Linguality, Multi-Functionality, Multi-Granularity Text Embeddings Through Self-Knowledge Distillation
von: Chen, Jianlv, et al.
Veröffentlicht: (2024)
von: Chen, Jianlv, et al.
Veröffentlicht: (2024)
Training Task Experts through Retrieval Based Distillation
von: Ge, Jiaxin, et al.
Veröffentlicht: (2024)
von: Ge, Jiaxin, et al.
Veröffentlicht: (2024)
MoSE: Hierarchical Self-Distillation Enhances Early Layer Embeddings
von: Gurioli, Andrea, et al.
Veröffentlicht: (2025)
von: Gurioli, Andrea, et al.
Veröffentlicht: (2025)
Compass-Embedding v4: Robust Contrastive Learning for Multilingual E-commerce Embeddings
von: Ueareeworakul, Pakorn, et al.
Veröffentlicht: (2025)
von: Ueareeworakul, Pakorn, et al.
Veröffentlicht: (2025)
PWESuite: Phonetic Word Embeddings and Tasks They Facilitate
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
von: Zouhar, Vilém, et al.
Veröffentlicht: (2023)
TurkEmbed: Turkish Embedding Model on NLI & STS Tasks
von: Ezerceli, Özay, et al.
Veröffentlicht: (2025)
von: Ezerceli, Özay, et al.
Veröffentlicht: (2025)
ETT-CKGE: Efficient Task-driven Tokens for Continual Knowledge Graph Embedding
von: Zhu, Lijing, et al.
Veröffentlicht: (2025)
von: Zhu, Lijing, et al.
Veröffentlicht: (2025)
Synthetically generated text for supervised text analysis
von: Halterman, Andrew
Veröffentlicht: (2023)
von: Halterman, Andrew
Veröffentlicht: (2023)
CapsF: Capsule Fusion for Extracting psychiatric stressors for suicide from twitter
von: Dadgostarnia, Mohammad Ali, et al.
Veröffentlicht: (2024)
von: Dadgostarnia, Mohammad Ali, et al.
Veröffentlicht: (2024)
Dynamic Embeddings with Task-Oriented prompting
von: Balloccu, Allmin, et al.
Veröffentlicht: (2024)
von: Balloccu, Allmin, et al.
Veröffentlicht: (2024)
IT5: Text-to-text Pretraining for Italian Language Understanding and Generation
von: Sarti, Gabriele, et al.
Veröffentlicht: (2022)
von: Sarti, Gabriele, et al.
Veröffentlicht: (2022)
Ähnliche Einträge
-
jina-embeddings-v5-omni: Geometry-preserving Embeddings via Locked Aligned Towers
von: Hönicke, Florian, et al.
Veröffentlicht: (2026) -
jina-embeddings-v4: Universal Embeddings for Multimodal Multilingual Retrieval
von: Günther, Michael, et al.
Veröffentlicht: (2025) -
jina-embeddings-v3: Multilingual Embeddings With Task LoRA
von: Sturua, Saba, et al.
Veröffentlicht: (2024) -
jina-clip-v2: Multilingual Multimodal Embeddings for Text and Images
von: Koukounas, Andreas, et al.
Veröffentlicht: (2024) -
Jina Embeddings 2: 8192-Token General-Purpose Text Embeddings for Long Documents
von: Günther, Michael, et al.
Veröffentlicht: (2023)