Semi-Supervised Image Captioning Considering Wasserstein Graph Matching
Fuente:
arXiv
Salvato in:
| Autore principale: | Yang, Yang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Integration of Self-Supervised BYOL in Semi-Supervised Medical Image Recognition
di: Feng, Hao, et al.
Pubblicazione: (2024)
di: Feng, Hao, et al.
Pubblicazione: (2024)
CaptionFormer: Unified Segmentation, Tracking, and Captioning for Spatio-Temporal Objects
di: Fiastre, Gabriel, et al.
Pubblicazione: (2025)
di: Fiastre, Gabriel, et al.
Pubblicazione: (2025)
Graph-Based Captioning: Enhancing Visual Descriptions by Interconnecting Region Captions
di: Hsieh, Yu-Guan, et al.
Pubblicazione: (2024)
di: Hsieh, Yu-Guan, et al.
Pubblicazione: (2024)
SemiOccam: A Robust Semi-Supervised Image Recognition Network Using Sparse Labels
di: Yann, Rui, et al.
Pubblicazione: (2025)
di: Yann, Rui, et al.
Pubblicazione: (2025)
Generalizable Geometric Image Caption Synthesis
di: Xin, Yue, et al.
Pubblicazione: (2025)
di: Xin, Yue, et al.
Pubblicazione: (2025)
Online Reward-Weighted Fine-Tuning of Flow Matching with Wasserstein Regularization
di: Fan, Jiajun, et al.
Pubblicazione: (2025)
di: Fan, Jiajun, et al.
Pubblicazione: (2025)
Image Captions are Natural Prompts for Text-to-Image Models
di: Lei, Shiye, et al.
Pubblicazione: (2023)
di: Lei, Shiye, et al.
Pubblicazione: (2023)
Harmonizing Generalization and Specialization: Uncertainty-Informed Collaborative Learning for Semi-supervised Medical Image Segmentation
di: Lu, Wenjing, et al.
Pubblicazione: (2025)
di: Lu, Wenjing, et al.
Pubblicazione: (2025)
Revisit Large-Scale Image-Caption Data in Pre-training Multimodal Foundation Models
di: Lai, Zhengfeng, et al.
Pubblicazione: (2024)
di: Lai, Zhengfeng, et al.
Pubblicazione: (2024)
Semi-Supervised 3D Medical Segmentation from 2D Natural Images Pretrained Model
di: Yeung, Pak-Hei, et al.
Pubblicazione: (2025)
di: Yeung, Pak-Hei, et al.
Pubblicazione: (2025)
ConformalSAM: Unlocking the Potential of Foundational Segmentation Models in Semi-Supervised Semantic Segmentation with Conformal Prediction
di: Chen, Danhui, et al.
Pubblicazione: (2025)
di: Chen, Danhui, et al.
Pubblicazione: (2025)
Wasserstein Distance Rivals Kullback-Leibler Divergence for Knowledge Distillation
di: Lv, Jiaming, et al.
Pubblicazione: (2024)
di: Lv, Jiaming, et al.
Pubblicazione: (2024)
Semi-Supervised Learning for Deep Causal Generative Models
di: Ibrahim, Yasin, et al.
Pubblicazione: (2024)
di: Ibrahim, Yasin, et al.
Pubblicazione: (2024)
An Embarrassingly Simple Baseline for Imbalanced Semi-Supervised Learning
di: Chen, Hao, et al.
Pubblicazione: (2022)
di: Chen, Hao, et al.
Pubblicazione: (2022)
IG Captioner: Information Gain Captioners are Strong Zero-shot Classifiers
di: Yang, Chenglin, et al.
Pubblicazione: (2023)
di: Yang, Chenglin, et al.
Pubblicazione: (2023)
Hyperdimensional Cross-Modal Alignment of Frozen Language and Image Models for Efficient Image Captioning
di: Dalvi, Abhishek, et al.
Pubblicazione: (2026)
di: Dalvi, Abhishek, et al.
Pubblicazione: (2026)
Style-Extracting Diffusion Models for Semi-Supervised Histopathology Segmentation
di: Öttl, Mathias, et al.
Pubblicazione: (2024)
di: Öttl, Mathias, et al.
Pubblicazione: (2024)
SynC: Synthetic Image Caption Dataset Refinement with One-to-many Mapping for Zero-shot Image Captioning
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
di: Kim, Si-Woo, et al.
Pubblicazione: (2025)
RubiCap: Rubric-Guided Reinforcement Learning for Dense Image Captioning
di: Huang, Tzu-Heng, et al.
Pubblicazione: (2026)
di: Huang, Tzu-Heng, et al.
Pubblicazione: (2026)
Aligning Forest and Trees in Images & Long Captions for Visually Grounded Understanding
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
di: Woo, Byeongju, et al.
Pubblicazione: (2026)
Prompt-Driven Feature Diffusion for Open-World Semi-Supervised Learning
di: Heidari, Marzi, et al.
Pubblicazione: (2024)
di: Heidari, Marzi, et al.
Pubblicazione: (2024)
Overcoming Data Inequality across Domains with Semi-Supervised Domain Generalization
di: Park, Jinha, et al.
Pubblicazione: (2024)
di: Park, Jinha, et al.
Pubblicazione: (2024)
From Semantic To Instance: A Semi-Self-Supervised Learning Approach
di: Najafian, Keyhan, et al.
Pubblicazione: (2025)
di: Najafian, Keyhan, et al.
Pubblicazione: (2025)
IPixMatch: Boost Semi-supervised Semantic Segmentation with Inter-Pixel Relation
di: Wu, Kebin, et al.
Pubblicazione: (2024)
di: Wu, Kebin, et al.
Pubblicazione: (2024)
Hardness-Aware Scene Synthesis for Semi-Supervised 3D Object Detection
di: Zeng, Shuai, et al.
Pubblicazione: (2024)
di: Zeng, Shuai, et al.
Pubblicazione: (2024)
Safe Semi-Supervised Contrastive Learning Using In-Distribution Data as Positive Examples
di: Kwak, Min Gu, et al.
Pubblicazione: (2024)
di: Kwak, Min Gu, et al.
Pubblicazione: (2024)
Semi-Supervised Masked Autoencoders: Unlocking Vision Transformer Potential with Limited Data
di: Faysal, Atik, et al.
Pubblicazione: (2026)
di: Faysal, Atik, et al.
Pubblicazione: (2026)
Curvature-Aware Captioning:Leveraging Geodesic Attention for 3D Scene Understanding
di: He, Ziyao, et al.
Pubblicazione: (2026)
di: He, Ziyao, et al.
Pubblicazione: (2026)
Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M)
di: Merchant, Nicholas, et al.
Pubblicazione: (2025)
di: Merchant, Nicholas, et al.
Pubblicazione: (2025)
Enhancing AI Diagnostics: Autonomous Lesion Masking via Semi-Supervised Deep Learning
di: Wei, Ting-Ruen, et al.
Pubblicazione: (2024)
di: Wei, Ting-Ruen, et al.
Pubblicazione: (2024)
CW-BASS: Confidence-Weighted Boundary-Aware Learning for Semi-Supervised Semantic Segmentation
di: Tarubinga, Ebenezer, et al.
Pubblicazione: (2025)
di: Tarubinga, Ebenezer, et al.
Pubblicazione: (2025)
Training-Only Heterogeneous Image-Patch-Text Graph Supervision for Advancing Few-Shot Learning Adapters
di: Mohammad, Mohammed Rahman Sherif Khan, et al.
Pubblicazione: (2026)
di: Mohammad, Mohammed Rahman Sherif Khan, et al.
Pubblicazione: (2026)
Video Latent Flow Matching: Optimal Polynomial Projections for Video Interpolation and Extrapolation
di: Cao, Yang, et al.
Pubblicazione: (2025)
di: Cao, Yang, et al.
Pubblicazione: (2025)
Matrix Information Theory for Self-Supervised Learning
di: Zhang, Yifan, et al.
Pubblicazione: (2023)
di: Zhang, Yifan, et al.
Pubblicazione: (2023)
How to Train your Text-to-Image Model: Evaluating Design Choices for Synthetic Training Captions
di: Brack, Manuel, et al.
Pubblicazione: (2025)
di: Brack, Manuel, et al.
Pubblicazione: (2025)
CompCap: Improving Multimodal Large Language Models with Composite Captions
di: Chen, Xiaohui, et al.
Pubblicazione: (2024)
di: Chen, Xiaohui, et al.
Pubblicazione: (2024)
VeCLIP: Improving CLIP Training via Visual-enriched Captions
di: Lai, Zhengfeng, et al.
Pubblicazione: (2023)
di: Lai, Zhengfeng, et al.
Pubblicazione: (2023)
Learning Unlabeled Clients Divergence for Federated Semi-Supervised Learning via Anchor Model Aggregation
di: Elbatel, Marawan, et al.
Pubblicazione: (2024)
di: Elbatel, Marawan, et al.
Pubblicazione: (2024)
An Examination of the Robustness of Reference-Free Image Captioning Evaluation Metrics
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
di: Ahmadi, Saba, et al.
Pubblicazione: (2023)
TA-Prompting: Enhancing Video Large Language Models for Dense Video Captioning via Temporal Anchors
di: Cheng, Wei-Yuan, et al.
Pubblicazione: (2026)
di: Cheng, Wei-Yuan, et al.
Pubblicazione: (2026)
Documenti analoghi
-
Integration of Self-Supervised BYOL in Semi-Supervised Medical Image Recognition
di: Feng, Hao, et al.
Pubblicazione: (2024) -
CaptionFormer: Unified Segmentation, Tracking, and Captioning for Spatio-Temporal Objects
di: Fiastre, Gabriel, et al.
Pubblicazione: (2025) -
Graph-Based Captioning: Enhancing Visual Descriptions by Interconnecting Region Captions
di: Hsieh, Yu-Guan, et al.
Pubblicazione: (2024) -
SemiOccam: A Robust Semi-Supervised Image Recognition Network Using Sparse Labels
di: Yann, Rui, et al.
Pubblicazione: (2025) -
Generalizable Geometric Image Caption Synthesis
di: Xin, Yue, et al.
Pubblicazione: (2025)