Breaking the Batch Barrier (B3) of Contrastive Learning via Smart Batch Mining
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Thirukovalluru, Raghuveer, Meng, Rui, Liu, Ye, K, Karthikeyan, Su, Mingyi, Nie, Ping, Yavuz, Semih, Zhou, Yingbo, Chen, Wenhu, Dhingra, Bhuwan |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2024)
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2024)
Atomic Self-Consistency for Better Long Form Generations
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2024)
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2024)
InData: Towards Secure Multi-Step, Tool-Based Data Analysis
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
Document-as-Image Representations Fall Short for Scientific Retrieval
von: Khalighinejad, Ghazal, et al.
Veröffentlicht: (2026)
von: Khalighinejad, Ghazal, et al.
Veröffentlicht: (2026)
Atomic Consistency Preference Optimization for Long-Form Question Answering
von: Chen, Jingfeng, et al.
Veröffentlicht: (2025)
von: Chen, Jingfeng, et al.
Veröffentlicht: (2025)
Calibrating Long-form Generations from Large Language Models
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
Text-Guided Semantic Image Encoder
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2025)
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2025)
Additive Large Language Models for Semi-Structured Text
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
ClinStructor: AI-Powered Structuring of Unstructured Clinical Texts
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
von: K, Karthikeyan, et al.
Veröffentlicht: (2025)
VLM2Vec: Training Vision-Language Models for Massive Multimodal Embedding Tasks
von: Jiang, Ziyan, et al.
Veröffentlicht: (2024)
von: Jiang, Ziyan, et al.
Veröffentlicht: (2024)
Parameter-Efficient Detoxification with Contrastive Decoding
von: Niu, Tong, et al.
Veröffentlicht: (2024)
von: Niu, Tong, et al.
Veröffentlicht: (2024)
Retrofitting Small Multilingual Models for Retrieval: Matching 7B Performance with 300M Parameters
von: Tu, Lifu, et al.
Veröffentlicht: (2025)
von: Tu, Lifu, et al.
Veröffentlicht: (2025)
LDDR: Linear-DPP-Based Dynamic-Resolution Frame Sampling for Video MLLMs
von: Chen, Jingfeng, et al.
Veröffentlicht: (2026)
von: Chen, Jingfeng, et al.
Veröffentlicht: (2026)
Investigating Factuality in Long-Form Text Generation: The Roles of Self-Known and Self-Unknown
von: Tu, Lifu, et al.
Veröffentlicht: (2024)
von: Tu, Lifu, et al.
Veröffentlicht: (2024)
HPE:Answering Complex Questions over Text by Hybrid Question Parsing and Execution
von: Liu, Ye, et al.
Veröffentlicht: (2023)
von: Liu, Ye, et al.
Veröffentlicht: (2023)
Traffic Light or Light Traffic? Investigating Phrasal Semantics in Large Language Models
von: Meng, Rui, et al.
Veröffentlicht: (2024)
von: Meng, Rui, et al.
Veröffentlicht: (2024)
Breaking the Memory Barrier: Near Infinite Batch Size Scaling for Contrastive Loss
von: Cheng, Zesen, et al.
Veröffentlicht: (2024)
von: Cheng, Zesen, et al.
Veröffentlicht: (2024)
VLM2Vec-V2: Advancing Multimodal Embedding for Videos, Images, and Visual Documents
von: Meng, Rui, et al.
Veröffentlicht: (2025)
von: Meng, Rui, et al.
Veröffentlicht: (2025)
CodeXEmbed: A Generalist Embedding Model Family for Multiligual and Multi-task Code Retrieval
von: Liu, Ye, et al.
Veröffentlicht: (2024)
von: Liu, Ye, et al.
Veröffentlicht: (2024)
Modeling Uncertainty and Using Post-fusion as Fallback Improves Retrieval Augmented Generation with LLMs
von: Liu, Ye, et al.
Veröffentlicht: (2023)
von: Liu, Ye, et al.
Veröffentlicht: (2023)
Unlocking Anticipatory Text Generation: A Constrained Approach for Large Language Models Decoding
von: Tu, Lifu, et al.
Veröffentlicht: (2023)
von: Tu, Lifu, et al.
Veröffentlicht: (2023)
RVPO: Risk-Sensitive Alignment via Variance Regularization
von: Montero, Ivan, et al.
Veröffentlicht: (2026)
von: Montero, Ivan, et al.
Veröffentlicht: (2026)
JudgeRank: Leveraging Large Language Models for Reasoning-Intensive Reranking
von: Niu, Tong, et al.
Veröffentlicht: (2024)
von: Niu, Tong, et al.
Veröffentlicht: (2024)
Improving LLM Reasoning through Scaling Inference Computation with Collaborative Verification
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
von: Liang, Zhenwen, et al.
Veröffentlicht: (2024)
Breaking the $\log(1/Δ_2)$ Barrier: Better Batched Best Arm Identification with Adaptive Grids
von: Jin, Tianyuan, et al.
Veröffentlicht: (2025)
von: Jin, Tianyuan, et al.
Veröffentlicht: (2025)
AntBatchInfer: Elastic Batch Inference in the Kubernetes Cluster
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
von: Li, Siyuan, et al.
Veröffentlicht: (2024)
AugTriever: Unsupervised Dense Retrieval and Domain Adaptation by Scalable Data Augmentation
von: Meng, Rui, et al.
Veröffentlicht: (2022)
von: Meng, Rui, et al.
Veröffentlicht: (2022)
MineDraft: A Framework for Batch Parallel Speculative Decoding
von: Tang, Zhenwei, et al.
Veröffentlicht: (2026)
von: Tang, Zhenwei, et al.
Veröffentlicht: (2026)
Real-time Factuality Assessment from Adversarial Feedback
von: Chen, Sanxing, et al.
Veröffentlicht: (2024)
von: Chen, Sanxing, et al.
Veröffentlicht: (2024)
ChatShop: Interactive Information Seeking with Language Agents
von: Chen, Sanxing, et al.
Veröffentlicht: (2024)
von: Chen, Sanxing, et al.
Veröffentlicht: (2024)
Fuzzy Speculative Decoding for a Tunable Accuracy-Runtime Tradeoff
von: Holsman, Maximilian, et al.
Veröffentlicht: (2025)
von: Holsman, Maximilian, et al.
Veröffentlicht: (2025)
BADM: Batch ADMM for Deep Learning
von: Wang, Ouya, et al.
Veröffentlicht: (2024)
von: Wang, Ouya, et al.
Veröffentlicht: (2024)
Batch Transformer: Look for Attention in Batch
von: Her, Myung Beom, et al.
Veröffentlicht: (2024)
von: Her, Myung Beom, et al.
Veröffentlicht: (2024)
Coding Agents are Effective Long-Context Processors
von: Cao, Weili, et al.
Veröffentlicht: (2026)
von: Cao, Weili, et al.
Veröffentlicht: (2026)
BatchLLM: Optimizing Large Batched LLM Inference with Global Prefix Sharing and Throughput-oriented Token Batching
von: Zheng, Zhen, et al.
Veröffentlicht: (2024)
von: Zheng, Zhen, et al.
Veröffentlicht: (2024)
Online Linear Programming with Batching
von: Xu, Haoran, et al.
Veröffentlicht: (2024)
von: Xu, Haoran, et al.
Veröffentlicht: (2024)
SMDP-Based Dynamic Batching for Improving Responsiveness and Energy Efficiency of Batch Services
von: Xu, Yaodan, et al.
Veröffentlicht: (2025)
von: Xu, Yaodan, et al.
Veröffentlicht: (2025)
BucketServe: Bucket-Based Dynamic Batching for Smart and Efficient LLM Inference Serving
von: Zheng, Wanyi, et al.
Veröffentlicht: (2025)
von: Zheng, Wanyi, et al.
Veröffentlicht: (2025)
Vision2Code: A Multi-Domain Benchmark for Evaluating Image-to-Code Generation
von: Periasami, Ajay Vikram, et al.
Veröffentlicht: (2026)
von: Periasami, Ajay Vikram, et al.
Veröffentlicht: (2026)
Hierarchical Multi-Label Classification of Online Vaccine Concerns
von: Zhu, Chloe Qinyu, et al.
Veröffentlicht: (2024)
von: Zhu, Chloe Qinyu, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
GenEOL: Harnessing the Generative Power of LLMs for Training-Free Sentence Embeddings
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2024) -
Atomic Self-Consistency for Better Long Form Generations
von: Thirukovalluru, Raghuveer, et al.
Veröffentlicht: (2024) -
InData: Towards Secure Multi-Step, Tool-Based Data Analysis
von: K, Karthikeyan, et al.
Veröffentlicht: (2025) -
Document-as-Image Representations Fall Short for Scientific Retrieval
von: Khalighinejad, Ghazal, et al.
Veröffentlicht: (2026) -
Atomic Consistency Preference Optimization for Long-Form Question Answering
von: Chen, Jingfeng, et al.
Veröffentlicht: (2025)