TSVC:Tripartite Learning with Semantic Variation Consistency for Robust Image-Text Retrieval
Fuente:
arXiv
Saved in:
| Main Authors: | Lyu, Shuai, Tian, Zijing, Ou, Zhonghong, Zhu, Yifan, Zhang, Xiao, Ha, Qiankun, Luo, Haoran, Song, Meina |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL
by: Lyu, Shuai, et al.
Published: (2025)
by: Lyu, Shuai, et al.
Published: (2025)
Structures Meet Semantics: Multimodal Fusion via Graph Contrastive Learning
by: Sun, Jiangfeng, et al.
Published: (2025)
by: Sun, Jiangfeng, et al.
Published: (2025)
OmniRAG-Agent: Agentic Omnimodal Reasoning for Low-Resource Long Audio-Video Question Answering
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
Text-Video Retrieval with Global-Local Semantic Consistent Learning
by: Zhang, Haonan, et al.
Published: (2024)
by: Zhang, Haonan, et al.
Published: (2024)
CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning
by: Feng, Yu, et al.
Published: (2024)
by: Feng, Yu, et al.
Published: (2024)
ERGeoBench:A Comprehensive Benchmark for Embodied Reasoning and Geo-localization in Multimodal Large Language Models
by: Xue, Kaiwen, et al.
Published: (2026)
by: Xue, Kaiwen, et al.
Published: (2026)
Evaluating Semantic Variation in Text-to-Image Synthesis: A Causal Perspective
by: Zhu, Xiangru, et al.
Published: (2024)
by: Zhu, Xiangru, et al.
Published: (2024)
Leveraging Text-Driven Semantic Variation for Robust OOD Segmentation
by: Song, Seungheon, et al.
Published: (2025)
by: Song, Seungheon, et al.
Published: (2025)
FGU3R: Fine-Grained Fusion via Unified 3D Representation for Multimodal 3D Object Detection
by: Zhang, Guoxin, et al.
Published: (2025)
by: Zhang, Guoxin, et al.
Published: (2025)
RevGNN: Negative Sampling Enhanced Contrastive Graph Learning for Academic Reviewer Recommendation
by: Liao, Weibin, et al.
Published: (2024)
by: Liao, Weibin, et al.
Published: (2024)
Consistency and Discrepancy-Based Contrastive Tripartite Graph Learning for Recommendations
by: Guo, Linxin, et al.
Published: (2024)
by: Guo, Linxin, et al.
Published: (2024)
LungCURE: Benchmarking Multimodal Real-World Clinical Reasoning for Precision Lung Cancer Diagnosis and Treatment
by: Hao, Fangyu, et al.
Published: (2026)
by: Hao, Fangyu, et al.
Published: (2026)
Boosting Text-to-Chart Retrieval through Training with Synthesized Semantic Insights
by: Wu, Yifan, et al.
Published: (2025)
by: Wu, Yifan, et al.
Published: (2025)
HyperGraphRAG: Retrieval-Augmented Generation via Hypergraph-Structured Knowledge Representation
by: Luo, Haoran, et al.
Published: (2025)
by: Luo, Haoran, et al.
Published: (2025)
VarDiU: A Variational Diffusive Upper Bound for One-Step Diffusion Distillation
by: Wang, Leyang, et al.
Published: (2025)
by: Wang, Leyang, et al.
Published: (2025)
HiFlow: Hierarchical Feedback-Driven Optimization for Constrained Long-Form Text Generation
by: Zhu, Yifan, et al.
Published: (2026)
by: Zhu, Yifan, et al.
Published: (2026)
ChatKBQA: A Generate-then-Retrieve Framework for Knowledge Base Question Answering with Fine-tuned Large Language Models
by: Luo, Haoran, et al.
Published: (2023)
by: Luo, Haoran, et al.
Published: (2023)
Robust Remote Sensing Image-Text Retrieval with Noisy Correspondence
by: Song, Qiya, et al.
Published: (2026)
by: Song, Qiya, et al.
Published: (2026)
Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation
by: Yuan, Chao, et al.
Published: (2026)
by: Yuan, Chao, et al.
Published: (2026)
EIFNet: Leveraging Event-Image Fusion for Robust Semantic Segmentation
by: Li, Zhijiang, et al.
Published: (2025)
by: Li, Zhijiang, et al.
Published: (2025)
Integrated Library System (ILS) Challenges and Opportunities: A Survey of U.S. Academic Libraries with Migration Projects
by: Wang, Zhonghong
Published: (2009)
by: Wang, Zhonghong
Published: (2009)
Enhance Multimodal Consistency and Coherence for Text-Image Plan Generation
by: Lu, Xiaoxin, et al.
Published: (2025)
by: Lu, Xiaoxin, et al.
Published: (2025)
CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval
by: Kang, Bin, et al.
Published: (2025)
by: Kang, Bin, et al.
Published: (2025)
Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval
by: Liu, Delong, et al.
Published: (2023)
by: Liu, Delong, et al.
Published: (2023)
Text2NKG: Fine-Grained N-ary Relation Extraction for N-ary relational Knowledge Graph Construction
by: Luo, Haoran, et al.
Published: (2023)
by: Luo, Haoran, et al.
Published: (2023)
Zero Shot Domain Adaptive Semantic Segmentation by Synthetic Data Generation and Progressive Adaptation
by: Luo, Jun, et al.
Published: (2025)
by: Luo, Jun, et al.
Published: (2025)
Rebalancing Contrastive Alignment with Bottlenecked Semantic Increments in Text-Video Retrieval
by: Xiao, Jian, et al.
Published: (2025)
by: Xiao, Jian, et al.
Published: (2025)
Hierarchical Semantic Compression for Consistent Image Semantic Restoration
by: Li, Shengxi, et al.
Published: (2025)
by: Li, Shengxi, et al.
Published: (2025)
EHRAG: Bridging Semantic Gaps in Lightweight GraphRAG via Hybrid Hypergraph Construction and Retrieval
by: Song, Yifan, et al.
Published: (2026)
by: Song, Yifan, et al.
Published: (2026)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
by: Luo, Jingzhou, et al.
Published: (2026)
by: Luo, Jingzhou, et al.
Published: (2026)
Asynchronous Denoising Diffusion Models for Aligning Text-to-Image Generation
by: Hu, Zijing, et al.
Published: (2025)
by: Hu, Zijing, et al.
Published: (2025)
SETR: A Two-Stage Semantic-Enhanced Framework for Zero-Shot Composed Image Retrieval
by: Xiao, Yuqi, et al.
Published: (2025)
by: Xiao, Yuqi, et al.
Published: (2025)
Entropy-Based Decoding for Retrieval-Augmented Large Language Models
by: Qiu, Zexuan, et al.
Published: (2024)
by: Qiu, Zexuan, et al.
Published: (2024)
Artificial Intelligence for Computer Science Education in Higher Education: A Systematic Review of Empirical Research Published in 2003-2023
by: Meina Zhu, et al.
Published: (2025)
by: Meina Zhu, et al.
Published: (2025)
PM-MOE: Mixture of Experts on Private Model Parameters for Personalized Federated Learning
by: Feng, Yu, et al.
Published: (2025)
by: Feng, Yu, et al.
Published: (2025)
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search
by: Luo, Haoran, et al.
Published: (2025)
by: Luo, Haoran, et al.
Published: (2025)
Benchmark Granularity and Model Robustness for Image-Text Retrieval
by: Hendriksen, Mariya, et al.
Published: (2024)
by: Hendriksen, Mariya, et al.
Published: (2024)
DPC: Training-Free Text-to-SQL Candidate Selection via Dual-Paradigm Consistency
by: Li, Boyan, et al.
Published: (2026)
by: Li, Boyan, et al.
Published: (2026)
SLICE: Semantic Latent Injection via Compartmentalized Embedding for Image Watermarking
by: Gao, Zheng, et al.
Published: (2026)
by: Gao, Zheng, et al.
Published: (2026)
Sequential Visual and Semantic Consistency for Semi-supervised Text Recognition
by: Yang, Mingkun, et al.
Published: (2024)
by: Yang, Mingkun, et al.
Published: (2024)
Similar Items
-
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL
by: Lyu, Shuai, et al.
Published: (2025) -
Structures Meet Semantics: Multimodal Fusion via Graph Contrastive Learning
by: Sun, Jiangfeng, et al.
Published: (2025) -
OmniRAG-Agent: Agentic Omnimodal Reasoning for Low-Resource Long Audio-Video Question Answering
by: Zhu, Yifan, et al.
Published: (2026) -
Text-Video Retrieval with Global-Local Semantic Consistent Learning
by: Zhang, Haonan, et al.
Published: (2024) -
CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning
by: Feng, Yu, et al.
Published: (2024)