TSVC:Tripartite Learning with Semantic Variation Consistency for Robust Image-Text Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Lyu, Shuai, Tian, Zijing, Ou, Zhonghong, Zhu, Yifan, Zhang, Xiao, Ha, Qiankun, Luo, Haoran, Song, Meina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL
von: Lyu, Shuai, et al.
Veröffentlicht: (2025)
von: Lyu, Shuai, et al.
Veröffentlicht: (2025)
Structures Meet Semantics: Multimodal Fusion via Graph Contrastive Learning
von: Sun, Jiangfeng, et al.
Veröffentlicht: (2025)
von: Sun, Jiangfeng, et al.
Veröffentlicht: (2025)
OmniRAG-Agent: Agentic Omnimodal Reasoning for Low-Resource Long Audio-Video Question Answering
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
Text-Video Retrieval with Global-Local Semantic Consistent Learning
von: Zhang, Haonan, et al.
Veröffentlicht: (2024)
von: Zhang, Haonan, et al.
Veröffentlicht: (2024)
CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning
von: Feng, Yu, et al.
Veröffentlicht: (2024)
von: Feng, Yu, et al.
Veröffentlicht: (2024)
ERGeoBench:A Comprehensive Benchmark for Embodied Reasoning and Geo-localization in Multimodal Large Language Models
von: Xue, Kaiwen, et al.
Veröffentlicht: (2026)
von: Xue, Kaiwen, et al.
Veröffentlicht: (2026)
Evaluating Semantic Variation in Text-to-Image Synthesis: A Causal Perspective
von: Zhu, Xiangru, et al.
Veröffentlicht: (2024)
von: Zhu, Xiangru, et al.
Veröffentlicht: (2024)
Leveraging Text-Driven Semantic Variation for Robust OOD Segmentation
von: Song, Seungheon, et al.
Veröffentlicht: (2025)
von: Song, Seungheon, et al.
Veröffentlicht: (2025)
FGU3R: Fine-Grained Fusion via Unified 3D Representation for Multimodal 3D Object Detection
von: Zhang, Guoxin, et al.
Veröffentlicht: (2025)
von: Zhang, Guoxin, et al.
Veröffentlicht: (2025)
RevGNN: Negative Sampling Enhanced Contrastive Graph Learning for Academic Reviewer Recommendation
von: Liao, Weibin, et al.
Veröffentlicht: (2024)
von: Liao, Weibin, et al.
Veröffentlicht: (2024)
Consistency and Discrepancy-Based Contrastive Tripartite Graph Learning for Recommendations
von: Guo, Linxin, et al.
Veröffentlicht: (2024)
von: Guo, Linxin, et al.
Veröffentlicht: (2024)
LungCURE: Benchmarking Multimodal Real-World Clinical Reasoning for Precision Lung Cancer Diagnosis and Treatment
von: Hao, Fangyu, et al.
Veröffentlicht: (2026)
von: Hao, Fangyu, et al.
Veröffentlicht: (2026)
Boosting Text-to-Chart Retrieval through Training with Synthesized Semantic Insights
von: Wu, Yifan, et al.
Veröffentlicht: (2025)
von: Wu, Yifan, et al.
Veröffentlicht: (2025)
HyperGraphRAG: Retrieval-Augmented Generation via Hypergraph-Structured Knowledge Representation
von: Luo, Haoran, et al.
Veröffentlicht: (2025)
von: Luo, Haoran, et al.
Veröffentlicht: (2025)
VarDiU: A Variational Diffusive Upper Bound for One-Step Diffusion Distillation
von: Wang, Leyang, et al.
Veröffentlicht: (2025)
von: Wang, Leyang, et al.
Veröffentlicht: (2025)
HiFlow: Hierarchical Feedback-Driven Optimization for Constrained Long-Form Text Generation
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
von: Zhu, Yifan, et al.
Veröffentlicht: (2026)
ChatKBQA: A Generate-then-Retrieve Framework for Knowledge Base Question Answering with Fine-tuned Large Language Models
von: Luo, Haoran, et al.
Veröffentlicht: (2023)
von: Luo, Haoran, et al.
Veröffentlicht: (2023)
Robust Remote Sensing Image-Text Retrieval with Noisy Correspondence
von: Song, Qiya, et al.
Veröffentlicht: (2026)
von: Song, Qiya, et al.
Veröffentlicht: (2026)
Towards Robust Text-to-Image Person Retrieval: Multi-View Reformulation for Semantic Compensation
von: Yuan, Chao, et al.
Veröffentlicht: (2026)
von: Yuan, Chao, et al.
Veröffentlicht: (2026)
EIFNet: Leveraging Event-Image Fusion for Robust Semantic Segmentation
von: Li, Zhijiang, et al.
Veröffentlicht: (2025)
von: Li, Zhijiang, et al.
Veröffentlicht: (2025)
Integrated Library System (ILS) Challenges and Opportunities: A Survey of U.S. Academic Libraries with Migration Projects
von: Wang, Zhonghong
Veröffentlicht: (2009)
von: Wang, Zhonghong
Veröffentlicht: (2009)
Enhance Multimodal Consistency and Coherence for Text-Image Plan Generation
von: Lu, Xiaoxin, et al.
Veröffentlicht: (2025)
von: Lu, Xiaoxin, et al.
Veröffentlicht: (2025)
CalibCLIP: Contextual Calibration of Dominant Semantics for Text-Driven Image Retrieval
von: Kang, Bin, et al.
Veröffentlicht: (2025)
von: Kang, Bin, et al.
Veröffentlicht: (2025)
Text-guided Image Restoration and Semantic Enhancement for Text-to-Image Person Retrieval
von: Liu, Delong, et al.
Veröffentlicht: (2023)
von: Liu, Delong, et al.
Veröffentlicht: (2023)
Text2NKG: Fine-Grained N-ary Relation Extraction for N-ary relational Knowledge Graph Construction
von: Luo, Haoran, et al.
Veröffentlicht: (2023)
von: Luo, Haoran, et al.
Veröffentlicht: (2023)
Zero Shot Domain Adaptive Semantic Segmentation by Synthetic Data Generation and Progressive Adaptation
von: Luo, Jun, et al.
Veröffentlicht: (2025)
von: Luo, Jun, et al.
Veröffentlicht: (2025)
Rebalancing Contrastive Alignment with Bottlenecked Semantic Increments in Text-Video Retrieval
von: Xiao, Jian, et al.
Veröffentlicht: (2025)
von: Xiao, Jian, et al.
Veröffentlicht: (2025)
Hierarchical Semantic Compression for Consistent Image Semantic Restoration
von: Li, Shengxi, et al.
Veröffentlicht: (2025)
von: Li, Shengxi, et al.
Veröffentlicht: (2025)
EHRAG: Bridging Semantic Gaps in Lightweight GraphRAG via Hybrid Hypergraph Construction and Retrieval
von: Song, Yifan, et al.
Veröffentlicht: (2026)
von: Song, Yifan, et al.
Veröffentlicht: (2026)
RoVLA: Multi-Consistency Constraints for Robust Vision-Language-Action Models
von: Luo, Jingzhou, et al.
Veröffentlicht: (2026)
von: Luo, Jingzhou, et al.
Veröffentlicht: (2026)
Asynchronous Denoising Diffusion Models for Aligning Text-to-Image Generation
von: Hu, Zijing, et al.
Veröffentlicht: (2025)
von: Hu, Zijing, et al.
Veröffentlicht: (2025)
SETR: A Two-Stage Semantic-Enhanced Framework for Zero-Shot Composed Image Retrieval
von: Xiao, Yuqi, et al.
Veröffentlicht: (2025)
von: Xiao, Yuqi, et al.
Veröffentlicht: (2025)
Entropy-Based Decoding for Retrieval-Augmented Large Language Models
von: Qiu, Zexuan, et al.
Veröffentlicht: (2024)
von: Qiu, Zexuan, et al.
Veröffentlicht: (2024)
Artificial Intelligence for Computer Science Education in Higher Education: A Systematic Review of Empirical Research Published in 2003-2023
von: Meina Zhu, et al.
Veröffentlicht: (2025)
von: Meina Zhu, et al.
Veröffentlicht: (2025)
PM-MOE: Mixture of Experts on Private Model Parameters for Personalized Federated Learning
von: Feng, Yu, et al.
Veröffentlicht: (2025)
von: Feng, Yu, et al.
Veröffentlicht: (2025)
KBQA-o1: Agentic Knowledge Base Question Answering with Monte Carlo Tree Search
von: Luo, Haoran, et al.
Veröffentlicht: (2025)
von: Luo, Haoran, et al.
Veröffentlicht: (2025)
Benchmark Granularity and Model Robustness for Image-Text Retrieval
von: Hendriksen, Mariya, et al.
Veröffentlicht: (2024)
von: Hendriksen, Mariya, et al.
Veröffentlicht: (2024)
DPC: Training-Free Text-to-SQL Candidate Selection via Dual-Paradigm Consistency
von: Li, Boyan, et al.
Veröffentlicht: (2026)
von: Li, Boyan, et al.
Veröffentlicht: (2026)
SLICE: Semantic Latent Injection via Compartmentalized Embedding for Image Watermarking
von: Gao, Zheng, et al.
Veröffentlicht: (2026)
von: Gao, Zheng, et al.
Veröffentlicht: (2026)
Sequential Visual and Semantic Consistency for Semi-supervised Text Recognition
von: Yang, Mingkun, et al.
Veröffentlicht: (2024)
von: Yang, Mingkun, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
SQL-o1: A Self-Reward Heuristic Dynamic Search Method for Text-to-SQL
von: Lyu, Shuai, et al.
Veröffentlicht: (2025) -
Structures Meet Semantics: Multimodal Fusion via Graph Contrastive Learning
von: Sun, Jiangfeng, et al.
Veröffentlicht: (2025) -
OmniRAG-Agent: Agentic Omnimodal Reasoning for Low-Resource Long Audio-Video Question Answering
von: Zhu, Yifan, et al.
Veröffentlicht: (2026) -
Text-Video Retrieval with Global-Local Semantic Consistent Learning
von: Zhang, Haonan, et al.
Veröffentlicht: (2024) -
CP-Prompt: Composition-Based Cross-modal Prompting for Domain-Incremental Continual Learning
von: Feng, Yu, et al.
Veröffentlicht: (2024)