MuCo: Multi-turn Contrastive Learning for Multimodal Embedding Model
Fuente:
arXiv
Salvato in:
| Autori principali: | Gu, Geonmo, Heo, Byeongho, Yu, Jaemyung, Hwang, Jaehui, Kim, Taekyung, Lee, Sangmin, Jun, HeeJae, Kang, Yoohoon, Yun, Sangdoo, Han, Dongyoon |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CompoDiff: Versatile Composed Image Retrieval With Latent Diffusion
di: Gu, Geonmo, et al.
Pubblicazione: (2023)
di: Gu, Geonmo, et al.
Pubblicazione: (2023)
Language-only Efficient Training of Zero-shot Composed Image Retrieval
di: Gu, Geonmo, et al.
Pubblicazione: (2023)
di: Gu, Geonmo, et al.
Pubblicazione: (2023)
Oops, Wait: Token-Level Signals as a Lens into LLM Reasoning
di: Hwang, Jaehui, et al.
Pubblicazione: (2026)
di: Hwang, Jaehui, et al.
Pubblicazione: (2026)
Learning to See What You Need: Gaze Attention for Multimodal Large Language Models
di: Song, Junha, et al.
Pubblicazione: (2026)
di: Song, Junha, et al.
Pubblicazione: (2026)
Masking meets Supervision: A Strong Learning Alliance
di: Heo, Byeongho, et al.
Pubblicazione: (2023)
di: Heo, Byeongho, et al.
Pubblicazione: (2023)
Rotary Position Embedding for Vision Transformer
di: Heo, Byeongho, et al.
Pubblicazione: (2024)
di: Heo, Byeongho, et al.
Pubblicazione: (2024)
Token Bottleneck: One Token to Remember Dynamics
di: Kim, Taekyung, et al.
Pubblicazione: (2025)
di: Kim, Taekyung, et al.
Pubblicazione: (2025)
Morphing Tokens Draw Strong Masked Image Models
di: Kim, Taekyung, et al.
Pubblicazione: (2023)
di: Kim, Taekyung, et al.
Pubblicazione: (2023)
RL makes MLLMs see better than SFT
di: Song, Junha, et al.
Pubblicazione: (2025)
di: Song, Junha, et al.
Pubblicazione: (2025)
Match me if you can: Semi-Supervised Semantic Correspondence Learning with Unpaired Images
di: Kim, Jiwon, et al.
Pubblicazione: (2023)
di: Kim, Jiwon, et al.
Pubblicazione: (2023)
Learning with Unmasked Tokens Drives Stronger Vision Learners
di: Kim, Taekyung, et al.
Pubblicazione: (2023)
di: Kim, Taekyung, et al.
Pubblicazione: (2023)
What Defines Good Reasoning in LLMs? Dissecting Reasoning Steps with Multi-Aspect Evaluation
di: Do, Heejin, et al.
Pubblicazione: (2025)
di: Do, Heejin, et al.
Pubblicazione: (2025)
Similarity of Neural Architectures using Adversarial Attack Transferability
di: Hwang, Jaehui, et al.
Pubblicazione: (2022)
di: Hwang, Jaehui, et al.
Pubblicazione: (2022)
Exploring Conditions for Diffusion models in Robotic Control
di: Shin, Heeseong, et al.
Pubblicazione: (2025)
di: Shin, Heeseong, et al.
Pubblicazione: (2025)
MuCo-KGC: Multi-Context-Aware Knowledge Graph Completion
di: Gul, Haji, et al.
Pubblicazione: (2025)
di: Gul, Haji, et al.
Pubblicazione: (2025)
MuCo: Publishing Microdata with Privacy Preservation through Mutual Cover
di: Li, Boyu, et al.
Pubblicazione: (2020)
di: Li, Boyu, et al.
Pubblicazione: (2020)
VisualScratchpad: Inference-time Visual Concepts Analysis in Vision Language Models
di: Lim, Hyesu, et al.
Pubblicazione: (2026)
di: Lim, Hyesu, et al.
Pubblicazione: (2026)
DenseNets Reloaded: Paradigm Shift Beyond ResNets and ViTs
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
di: Kim, Donghyun, et al.
Pubblicazione: (2024)
HYPE: Hyperbolic Entailment Filtering for Underspecified Images and Texts
di: Kim, Wonjae, et al.
Pubblicazione: (2024)
di: Kim, Wonjae, et al.
Pubblicazione: (2024)
SetCSE: Set Operations using Contrastive Learning of Sentence Embeddings
di: Liu, Kang
Pubblicazione: (2024)
di: Liu, Kang
Pubblicazione: (2024)
MuRAR: A Simple and Effective Multimodal Retrieval and Answer Refinement Framework for Multimodal Question Answering
di: Zhu, Zhengyuan, et al.
Pubblicazione: (2024)
di: Zhu, Zhengyuan, et al.
Pubblicazione: (2024)
SARDINE: A Simulator for Automated Recommendation in Dynamic and Interactive Environments
di: Deffayet, Romain, et al.
Pubblicazione: (2023)
di: Deffayet, Romain, et al.
Pubblicazione: (2023)
C-SEO Bench: Does Conversational SEO Work?
di: Puerto, Haritz, et al.
Pubblicazione: (2025)
di: Puerto, Haritz, et al.
Pubblicazione: (2025)
Imagine All The Relevance: Scenario-Profiled Indexing with Knowledge Expansion for Dense Retrieval
di: Lee, Sangam, et al.
Pubblicazione: (2025)
di: Lee, Sangam, et al.
Pubblicazione: (2025)
Is Contrastive Learning Necessary? A Study of Data Augmentation vs Contrastive Learning in Sequential Recommendation
di: Zhou, Peilin, et al.
Pubblicazione: (2024)
di: Zhou, Peilin, et al.
Pubblicazione: (2024)
Personalized Parameter-Efficient Fine-Tuning of Foundation Models for Multimodal Recommendation
di: Kim, Sunwoo, et al.
Pubblicazione: (2026)
di: Kim, Sunwoo, et al.
Pubblicazione: (2026)
Unifying Multimodal Retrieval via Document Screenshot Embedding
di: Ma, Xueguang, et al.
Pubblicazione: (2024)
di: Ma, Xueguang, et al.
Pubblicazione: (2024)
Equivariant Contrastive Learning for Sequential Recommendation
di: Zhou, Peilin, et al.
Pubblicazione: (2022)
di: Zhou, Peilin, et al.
Pubblicazione: (2022)
DNNs May Determine Major Properties of Their Outputs Early, with Timing Possibly Driven by Bias
di: Park, Song, et al.
Pubblicazione: (2025)
di: Park, Song, et al.
Pubblicazione: (2025)
Text Embeddings by Weakly-Supervised Contrastive Pre-training
di: Wang, Liang, et al.
Pubblicazione: (2022)
di: Wang, Liang, et al.
Pubblicazione: (2022)
BIPCL: Bilateral Intent-Enhanced Sequential Recommendation via Embedding Perturbation Contrastive Learning
di: Zhang, Shanfan, et al.
Pubblicazione: (2026)
di: Zhang, Shanfan, et al.
Pubblicazione: (2026)
LEXA: Legal Case Retrieval via Graph Contrastive Learning with Contextualised LLM Embeddings
di: Tang, Yanran, et al.
Pubblicazione: (2024)
di: Tang, Yanran, et al.
Pubblicazione: (2024)
CoST: Contrastive Quantization based Semantic Tokenization for Generative Recommendation
di: Zhu, Jieming, et al.
Pubblicazione: (2024)
di: Zhu, Jieming, et al.
Pubblicazione: (2024)
RECOR: Reasoning-focused Multi-turn Conversational Retrieval Benchmark
di: Ali, Mohammed, et al.
Pubblicazione: (2026)
di: Ali, Mohammed, et al.
Pubblicazione: (2026)
Sequences as Nodes for Contrastive Multimodal Graph Recommendation
di: Sahyouni, Bucher, et al.
Pubblicazione: (2026)
di: Sahyouni, Bucher, et al.
Pubblicazione: (2026)
Vela: Scalable Embeddings with Voice Large Language Models for Multimodal Retrieval
di: Hu, Ruofan, et al.
Pubblicazione: (2025)
di: Hu, Ruofan, et al.
Pubblicazione: (2025)
CoRECT: A Framework for Evaluating Embedding Compression Techniques at Scale
di: Caspari, L., et al.
Pubblicazione: (2025)
di: Caspari, L., et al.
Pubblicazione: (2025)
CoLLM: Integrating Collaborative Embeddings into Large Language Models for Recommendation
di: Zhang, Yang, et al.
Pubblicazione: (2023)
di: Zhang, Yang, et al.
Pubblicazione: (2023)
Unveiling Contrastive Learning's Capability of Neighborhood Aggregation for Collaborative Filtering
di: Zhang, Yu, et al.
Pubblicazione: (2025)
di: Zhang, Yu, et al.
Pubblicazione: (2025)
NeuroCLIP: Brain-Inspired Prompt Tuning for EEG-to-Image Multimodal Contrastive Learning
di: Wang, Jiyuan, et al.
Pubblicazione: (2025)
di: Wang, Jiyuan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CompoDiff: Versatile Composed Image Retrieval With Latent Diffusion
di: Gu, Geonmo, et al.
Pubblicazione: (2023) -
Language-only Efficient Training of Zero-shot Composed Image Retrieval
di: Gu, Geonmo, et al.
Pubblicazione: (2023) -
Oops, Wait: Token-Level Signals as a Lens into LLM Reasoning
di: Hwang, Jaehui, et al.
Pubblicazione: (2026) -
Learning to See What You Need: Gaze Attention for Multimodal Large Language Models
di: Song, Junha, et al.
Pubblicazione: (2026) -
Masking meets Supervision: A Strong Learning Alliance
di: Heo, Byeongho, et al.
Pubblicazione: (2023)