Enhancing Multimodal Entity Linking with Jaccard Distance-based Conditional Contrastive Learning and Contextual Visual Augmentation
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Nguyen, Cong-Duy, Wu, Xiaobao, Nguyen, Thong, Zhao, Shuai, Le, Khoi, Nguyen, Viet-Anh, Yichao, Feng, Luu, Anh Tuan |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
KDMCSE: Knowledge Distillation Multimodal Sentence Embeddings with Adaptive Angular margin Contrastive Learning
par: Nguyen, Cong-Duy, et autres
Publié: (2024)
par: Nguyen, Cong-Duy, et autres
Publié: (2024)
CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base
par: Nguyen, Cong-Duy, et autres
Publié: (2025)
par: Nguyen, Cong-Duy, et autres
Publié: (2025)
More Bias, Less Bias: BiasPrompting for Enhanced Multiple-Choice Question Answering
par: Vu, Duc Anh, et autres
Publié: (2025)
par: Vu, Duc Anh, et autres
Publié: (2025)
Adaptive Contrastive Learning on Multimodal Transformer for Review Helpfulness Predictions
par: Nguyen, Thong, et autres
Publié: (2022)
par: Nguyen, Thong, et autres
Publié: (2022)
Don't Read Everything: A Curvature-Conditioned Query for Linear Attention
par: Le, Dong, et autres
Publié: (2026)
par: Le, Dong, et autres
Publié: (2026)
Expand BERT Representation with Visual Information via Grounded Language Learning with Multimodal Partial Alignment
par: Nguyen, Cong-Duy, et autres
Publié: (2023)
par: Nguyen, Cong-Duy, et autres
Publié: (2023)
Topic Modeling as Multi-Objective Contrastive Optimization
par: Nguyen, Thong, et autres
Publié: (2024)
par: Nguyen, Thong, et autres
Publié: (2024)
A Survey on Neural Topic Models: Methods, Applications, and Challenges
par: Wu, Xiaobao, et autres
Publié: (2024)
par: Wu, Xiaobao, et autres
Publié: (2024)
On the Affinity, Rationality, and Diversity of Hierarchical Topic Modeling
par: Wu, Xiaobao, et autres
Publié: (2024)
par: Wu, Xiaobao, et autres
Publié: (2024)
Vision-and-Language Pretraining
par: Nguyen, Thong, et autres
Publié: (2022)
par: Nguyen, Thong, et autres
Publié: (2022)
Motion-aware Contrastive Learning for Temporal Panoptic Scene Graph Generation
par: Nguyen, Thong Thanh, et autres
Publié: (2024)
par: Nguyen, Thong Thanh, et autres
Publié: (2024)
MAMA: Meta-optimized Angular Margin Contrastive Framework for Video-Language Representation Learning
par: Nguyen, Thong, et autres
Publié: (2024)
par: Nguyen, Thong, et autres
Publié: (2024)
Gradient-Boosted Decision Tree for Listwise Context Model in Multimodal Review Helpfulness Prediction
par: Nguyen, Thong, et autres
Publié: (2023)
par: Nguyen, Thong, et autres
Publié: (2023)
READ: Recurrent Adapter with Partial Video-Language Alignment for Parameter-Efficient Transfer Learning in Low-Resource Video-Language Modeling
par: Nguyen, Thong, et autres
Publié: (2023)
par: Nguyen, Thong, et autres
Publié: (2023)
Multi-Scale Contrastive Learning for Video Temporal Grounding
par: Nguyen, Thong Thanh, et autres
Publié: (2024)
par: Nguyen, Thong Thanh, et autres
Publié: (2024)
DemaFormer: Damped Exponential Moving Average Transformer with Energy-Based Modeling for Temporal Language Grounding
par: Nguyen, Thong, et autres
Publié: (2023)
par: Nguyen, Thong, et autres
Publié: (2023)
Modeling Dynamic Topics in Chain-Free Fashion by Evolution-Tracking Contrastive Learning and Unassociated Word Exclusion
par: Wu, Xiaobao, et autres
Publié: (2024)
par: Wu, Xiaobao, et autres
Publié: (2024)
Learning Uncertainty from Sequential Internal Dispersion in Large Language Models
par: Srey, Ponhvoan, et autres
Publié: (2026)
par: Srey, Ponhvoan, et autres
Publié: (2026)
Encoding and Controlling Global Semantics for Long-form Video Question Answering
par: Nguyen, Thong Thanh, et autres
Publié: (2024)
par: Nguyen, Thong Thanh, et autres
Publié: (2024)
When In-Distribution Gains Fail: Evaluating Weak-to-Strong Reward Models under Preference Shift
par: Le, Khoi, et autres
Publié: (2026)
par: Le, Khoi, et autres
Publié: (2026)
Unlearning Backdoor Attacks for LLMs with Weak-to-Strong Knowledge Distillation
par: Zhao, Shuai, et autres
Publié: (2024)
par: Zhao, Shuai, et autres
Publié: (2024)
Curriculum Demonstration Selection for In-Context Learning
par: Vu, Duc Anh, et autres
Publié: (2024)
par: Vu, Duc Anh, et autres
Publié: (2024)
Provably data-driven projection method for quadratic programming
par: Nguyen, Anh Tuan, et autres
Publié: (2025)
par: Nguyen, Anh Tuan, et autres
Publié: (2025)
Temporal-Oriented Recipe for Transferring Large Vision-Language Model to Video Understanding
par: Nguyen, Thong, et autres
Publié: (2025)
par: Nguyen, Thong, et autres
Publié: (2025)
Provably Data-driven Multiple Hyper-parameter Tuning with Structured Loss Function
par: Le, Tung Quoc, et autres
Publié: (2026)
par: Le, Tung Quoc, et autres
Publié: (2026)
Provably Data-driven Lagrangian Relaxation for Mixed Integer Linear Programming
par: Le, Tung Quoc, et autres
Publié: (2026)
par: Le, Tung Quoc, et autres
Publié: (2026)
CSD-VAR: Content-Style Decomposition in Visual Autoregressive Models
par: Nguyen, Quang-Binh, et autres
Publié: (2025)
par: Nguyen, Quang-Binh, et autres
Publié: (2025)
Cost-Adaptive Recourse Recommendation by Adaptive Preference Elicitation
par: Nguyen, Duy, et autres
Publié: (2024)
par: Nguyen, Duy, et autres
Publié: (2024)
FASTopic: Pretrained Transformer is a Fast, Adaptive, Stable, and Transferable Topic Model
par: Wu, Xiaobao, et autres
Publié: (2024)
par: Wu, Xiaobao, et autres
Publié: (2024)
InfoCTM: A Mutual Information Maximization Perspective of Cross-Lingual Topic Modeling
par: Wu, Xiaobao, et autres
Publié: (2023)
par: Wu, Xiaobao, et autres
Publié: (2023)
Multimodal Contrastive Representation Learning in Augmented Biomedical Knowledge Graphs
par: Dang, Tien, et autres
Publié: (2025)
par: Dang, Tien, et autres
Publié: (2025)
Enriching and Controlling Global Semantics for Text Summarization
par: Nguyen, Thong, et autres
Publié: (2021)
par: Nguyen, Thong, et autres
Publié: (2021)
Tracking the Truth: Object-Centric Spatio-Temporal Monitoring for Video Large Language Models
par: Cao, Tri, et autres
Publié: (2026)
par: Cao, Tri, et autres
Publié: (2026)
Aspect-Based Summarization with Self-Aspect Retrieval Enhanced Generation
par: Feng, Yichao, et autres
Publié: (2025)
par: Feng, Yichao, et autres
Publié: (2025)
CASUAL: Conditional Support Alignment for Domain Adaptation with Label Shift
par: Nguyen, Anh T, et autres
Publié: (2023)
par: Nguyen, Anh T, et autres
Publié: (2023)
Distributional Surgery for Language Model Activations
par: Nguyen, Bao, et autres
Publié: (2025)
par: Nguyen, Bao, et autres
Publié: (2025)
Effects of magnetic field and structural parameters on multi-photon absorption spectra in Morse quantum wells with electron-phonon interactions
par: Vi, Tran Ky, et autres
Publié: (2024)
par: Vi, Tran Ky, et autres
Publié: (2024)
Torsion subgroups of Whitehead groups of graded division algebras
par: Khanh, Huynh Viet, et autres
Publié: (2024)
par: Khanh, Huynh Viet, et autres
Publié: (2024)
Towards Reliable Truth-Aligned Uncertainty Estimation in Large Language Models
par: Srey, Ponhvoan, et autres
Publié: (2026)
par: Srey, Ponhvoan, et autres
Publié: (2026)
A Comparative Analysis of Contextual Representation Flow in State-Space and Transformer Architectures
par: Hoang, Nhat M., et autres
Publié: (2025)
par: Hoang, Nhat M., et autres
Publié: (2025)
Documents similaires
-
KDMCSE: Knowledge Distillation Multimodal Sentence Embeddings with Adaptive Angular margin Contrastive Learning
par: Nguyen, Cong-Duy, et autres
Publié: (2024) -
CutPaste&Find: Efficient Multimodal Hallucination Detector with Visual-aid Knowledge Base
par: Nguyen, Cong-Duy, et autres
Publié: (2025) -
More Bias, Less Bias: BiasPrompting for Enhanced Multiple-Choice Question Answering
par: Vu, Duc Anh, et autres
Publié: (2025) -
Adaptive Contrastive Learning on Multimodal Transformer for Review Helpfulness Predictions
par: Nguyen, Thong, et autres
Publié: (2022) -
Don't Read Everything: A Curvature-Conditioned Query for Linear Attention
par: Le, Dong, et autres
Publié: (2026)