Continual Learning for Image Captioning through Improved Image-Text Alignment
Fuente:
arXiv
Salvato in:
| Autori principali: | Taetz, Bertram, Bordelius, Gal |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Removing Distributional Discrepancies in Captions Improves Image-Text Alignment
di: Li, Yuheng, et al.
Pubblicazione: (2024)
di: Li, Yuheng, et al.
Pubblicazione: (2024)
Improving Text Generation on Images with Synthetic Captions
di: Koh, Jun Young, et al.
Pubblicazione: (2024)
di: Koh, Jun Young, et al.
Pubblicazione: (2024)
Generating an Image From 1,000 Words: Enhancing Text-to-Image With Structured Captions
di: Gutflaish, Eyal, et al.
Pubblicazione: (2025)
di: Gutflaish, Eyal, et al.
Pubblicazione: (2025)
SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning
di: Zhang, Lin, et al.
Pubblicazione: (2025)
di: Zhang, Lin, et al.
Pubblicazione: (2025)
Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning
di: Ye, Qinghao, et al.
Pubblicazione: (2025)
di: Ye, Qinghao, et al.
Pubblicazione: (2025)
Benchmarking and Improving Detail Image Caption
di: Dong, Hongyuan, et al.
Pubblicazione: (2024)
di: Dong, Hongyuan, et al.
Pubblicazione: (2024)
Panoptic Captioning: An Equivalence Bridge for Image and Text
di: Lin, Kun-Yu, et al.
Pubblicazione: (2025)
di: Lin, Kun-Yu, et al.
Pubblicazione: (2025)
ITIScore: An Image-to-Text-to-Image Rating Framework for the Image Captioning Ability of MLLMs
di: Xu, Zitong, et al.
Pubblicazione: (2026)
di: Xu, Zitong, et al.
Pubblicazione: (2026)
Evaluating Image Caption via Cycle-consistent Text-to-Image Generation
di: Cui, Tianyu, et al.
Pubblicazione: (2025)
di: Cui, Tianyu, et al.
Pubblicazione: (2025)
Text Data-Centric Image Captioning with Interactive Prompts
di: Wang, Yiyu, et al.
Pubblicazione: (2024)
di: Wang, Yiyu, et al.
Pubblicazione: (2024)
Mining Fine-Grained Image-Text Alignment for Zero-Shot Captioning via Text-Only Training
di: Qiu, Longtian, et al.
Pubblicazione: (2024)
di: Qiu, Longtian, et al.
Pubblicazione: (2024)
Harnessing Caption Detailness for Data-Efficient Text-to-Image Generation
di: Wang, Xinran, et al.
Pubblicazione: (2025)
di: Wang, Xinran, et al.
Pubblicazione: (2025)
AGIC: Attention-Guided Image Captioning to Improve Caption Relevance
di: Teja, L. D. M. S. Sai, et al.
Pubblicazione: (2025)
di: Teja, L. D. M. S. Sai, et al.
Pubblicazione: (2025)
Learning to Rank Caption Chains for Video-Text Alignment
di: Blume, Ansel, et al.
Pubblicazione: (2026)
di: Blume, Ansel, et al.
Pubblicazione: (2026)
Text-only Synthesis for Image Captioning
di: Zhou, Qing, et al.
Pubblicazione: (2024)
di: Zhou, Qing, et al.
Pubblicazione: (2024)
Precision or Recall? An Analysis of Image Captions for Training Text-to-Image Generation Model
di: Cheng, Sheng, et al.
Pubblicazione: (2024)
di: Cheng, Sheng, et al.
Pubblicazione: (2024)
Text-to-Image Alignment in Denoising-Based Models through Step Selection
di: Grimal, Paul, et al.
Pubblicazione: (2025)
di: Grimal, Paul, et al.
Pubblicazione: (2025)
Policy Optimized Text-to-Image Pipeline Design
di: Gadot, Uri, et al.
Pubblicazione: (2025)
di: Gadot, Uri, et al.
Pubblicazione: (2025)
LCM-Lookahead for Encoder-based Text-to-Image Personalization
di: Gal, Rinon, et al.
Pubblicazione: (2024)
di: Gal, Rinon, et al.
Pubblicazione: (2024)
VIXEN: Visual Text Comparison Network for Image Difference Captioning
di: Black, Alexander, et al.
Pubblicazione: (2024)
di: Black, Alexander, et al.
Pubblicazione: (2024)
Amortized Inverse Kinematics via Graph Attention for Real-Time Human Avatar Animation
di: Khan, Muhammad Saif Ullah, et al.
Pubblicazione: (2026)
di: Khan, Muhammad Saif Ullah, et al.
Pubblicazione: (2026)
CaptionQA: Is Your Caption as Useful as the Image Itself?
di: Yang, Shijia, et al.
Pubblicazione: (2025)
di: Yang, Shijia, et al.
Pubblicazione: (2025)
Improving Long-Text Alignment for Text-to-Image Diffusion Models
di: Liu, Luping, et al.
Pubblicazione: (2024)
di: Liu, Luping, et al.
Pubblicazione: (2024)
Linear Alignment of Vision-language Models for Image Captioning
di: Paischer, Fabian, et al.
Pubblicazione: (2023)
di: Paischer, Fabian, et al.
Pubblicazione: (2023)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
di: Kim, Seoyeon, et al.
Pubblicazione: (2023)
di: Kim, Seoyeon, et al.
Pubblicazione: (2023)
Transformers in Medicine: Improving Vision-Language Alignment for Medical Image Captioning
di: Suresh, Yogesh Thakku, et al.
Pubblicazione: (2025)
di: Suresh, Yogesh Thakku, et al.
Pubblicazione: (2025)
Pretrained Image-Text Models are Secretly Video Captioners
di: Zhang, Chunhui, et al.
Pubblicazione: (2025)
di: Zhang, Chunhui, et al.
Pubblicazione: (2025)
Is Your Text-to-Image Model Robust to Caption Noise?
di: Yu, Weichen, et al.
Pubblicazione: (2024)
di: Yu, Weichen, et al.
Pubblicazione: (2024)
Image Captions are Natural Prompts for Text-to-Image Models
di: Lei, Shiye, et al.
Pubblicazione: (2023)
di: Lei, Shiye, et al.
Pubblicazione: (2023)
Data-Driven Loss Functions for Inference-Time Optimization in Text-to-Image
di: Yiflach, Sapir Esther, et al.
Pubblicazione: (2025)
di: Yiflach, Sapir Esther, et al.
Pubblicazione: (2025)
EAMA : Entity-Aware Multimodal Alignment Based Approach for News Image Captioning
di: Zhang, Junzhe, et al.
Pubblicazione: (2024)
di: Zhang, Junzhe, et al.
Pubblicazione: (2024)
CaptionSmiths: Flexibly Controlling Language Pattern in Image Captioning
di: Saito, Kuniaki, et al.
Pubblicazione: (2025)
di: Saito, Kuniaki, et al.
Pubblicazione: (2025)
Structured Captions Improve Prompt Adherence in Text-to-Image Models (Re-LAION-Caption 19M)
di: Merchant, Nicholas, et al.
Pubblicazione: (2025)
di: Merchant, Nicholas, et al.
Pubblicazione: (2025)
Language-Image Alignment with Fixed Text Encoders
di: Yang, Jingfeng, et al.
Pubblicazione: (2025)
di: Yang, Jingfeng, et al.
Pubblicazione: (2025)
Generalizing Alignment Paradigm of Text-to-Image Generation with Preferences through $f$-divergence Minimization
di: Sun, Haoyuan, et al.
Pubblicazione: (2024)
di: Sun, Haoyuan, et al.
Pubblicazione: (2024)
Image Generation from Image Captioning -- Invertible Approach
di: Menon, Nandakishore S, et al.
Pubblicazione: (2024)
di: Menon, Nandakishore S, et al.
Pubblicazione: (2024)
Caption-Driven Explorations: Aligning Image and Text Embeddings through Human-Inspired Foveated Vision
di: Zanca, Dario, et al.
Pubblicazione: (2024)
di: Zanca, Dario, et al.
Pubblicazione: (2024)
Key-Locked Rank One Editing for Text-to-Image Personalization
di: Tewel, Yoad, et al.
Pubblicazione: (2023)
di: Tewel, Yoad, et al.
Pubblicazione: (2023)
TIT-Score: Evaluating Long-Prompt Based Text-to-Image Alignment via Text-to-Image-to-Text Consistency
di: Wang, Juntong, et al.
Pubblicazione: (2025)
di: Wang, Juntong, et al.
Pubblicazione: (2025)
Continual Alignment for SAM: Rethinking Foundation Models for Medical Image Segmentation in Continual Learning
di: Wang, Jiayi, et al.
Pubblicazione: (2025)
di: Wang, Jiayi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Removing Distributional Discrepancies in Captions Improves Image-Text Alignment
di: Li, Yuheng, et al.
Pubblicazione: (2024) -
Improving Text Generation on Images with Synthetic Captions
di: Koh, Jun Young, et al.
Pubblicazione: (2024) -
Generating an Image From 1,000 Words: Enhancing Text-to-Image With Structured Captions
di: Gutflaish, Eyal, et al.
Pubblicazione: (2025) -
SC-Captioner: Improving Image Captioning with Self-Correction by Reinforcement Learning
di: Zhang, Lin, et al.
Pubblicazione: (2025) -
Painting with Words: Elevating Detailed Image Captioning with Benchmark and Alignment Learning
di: Ye, Qinghao, et al.
Pubblicazione: (2025)