Learning Hyperspectral Images with Curated Text Prompts for Efficient Multimodal Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Chatterjee, Abhiroop, Ghosh, Susmita |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-supervised Learning for Hyperspectral Images of Trees
von: Rahman, Moqsadur, et al.
Veröffentlicht: (2025)
von: Rahman, Moqsadur, et al.
Veröffentlicht: (2025)
Image Captions are Natural Prompts for Text-to-Image Models
von: Lei, Shiye, et al.
Veröffentlicht: (2023)
von: Lei, Shiye, et al.
Veröffentlicht: (2023)
Revisiting Aerial Scene Classification on the AID Benchmark
von: Das, Subhajeet, et al.
Veröffentlicht: (2026)
von: Das, Subhajeet, et al.
Veröffentlicht: (2026)
Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2026)
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2026)
Hybrid Deep Learning for Hyperspectral Single Image Super-Resolution
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)
Gramian Multimodal Representation Learning and Alignment
von: Cicchetti, Giordano, et al.
Veröffentlicht: (2024)
von: Cicchetti, Giordano, et al.
Veröffentlicht: (2024)
Improving GFlowNets for Text-to-Image Diffusion Alignment
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
von: Zhang, Dinghuai, et al.
Veröffentlicht: (2024)
Towards Evaluating Robustness of Prompt Adherence in Text to Image Models
von: Vemishetty, Sujith, et al.
Veröffentlicht: (2025)
von: Vemishetty, Sujith, et al.
Veröffentlicht: (2025)
An Efficient and Explanatory Image and Text Clustering System with Multimodal Autoencoder Architecture
von: Shi, Tiancheng, et al.
Veröffentlicht: (2024)
von: Shi, Tiancheng, et al.
Veröffentlicht: (2024)
AlignGuard: Scalable Safety Alignment for Text-to-Image Generation
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
von: Liu, Runtao, et al.
Veröffentlicht: (2024)
Minority-Focused Text-to-Image Generation via Prompt Optimization
von: Um, Soobin, et al.
Veröffentlicht: (2024)
von: Um, Soobin, et al.
Veröffentlicht: (2024)
Test-Time Alignment of Text-to-Image Diffusion Models via Null-Text Embedding Optimisation
von: Kim, Taehoon, et al.
Veröffentlicht: (2025)
von: Kim, Taehoon, et al.
Veröffentlicht: (2025)
Label Semantics for Robust Hyperspectral Image Classification
von: Hassan, Rafin, et al.
Veröffentlicht: (2025)
von: Hassan, Rafin, et al.
Veröffentlicht: (2025)
Optimizing Negative Prompts for Enhanced Aesthetics and Fidelity in Text-To-Image Generation
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
von: Ogezi, Michael, et al.
Veröffentlicht: (2024)
Prompt Optimizer of Text-to-Image Diffusion Models for Abstract Concept Understanding
von: Fan, Zezhong, et al.
Veröffentlicht: (2024)
von: Fan, Zezhong, et al.
Veröffentlicht: (2024)
Naïve PAINE: Lightweight Text-to-Image Generation Improvement with Prompt Evaluation
von: Kim, Joong Ho, et al.
Veröffentlicht: (2026)
von: Kim, Joong Ho, et al.
Veröffentlicht: (2026)
Distilled Prompt Learning for Incomplete Multimodal Survival Prediction
von: Xu, Yingxue, et al.
Veröffentlicht: (2025)
von: Xu, Yingxue, et al.
Veröffentlicht: (2025)
One-Prompt-One-Story: Free-Lunch Consistent Text-to-Image Generation Using a Single Prompt
von: Liu, Tao, et al.
Veröffentlicht: (2025)
von: Liu, Tao, et al.
Veröffentlicht: (2025)
Zero-Residual Concept Erasure via Progressive Alignment in Text-to-Image Model
von: Chen, Hongxu, et al.
Veröffentlicht: (2025)
von: Chen, Hongxu, et al.
Veröffentlicht: (2025)
Text-to-Image Diffusion Models Cannot Count, and Prompt Refinement Cannot Help
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
von: Guo, Xuyang, et al.
Veröffentlicht: (2025)
Prompt-Based Safety Guidance Is Ineffective for Unlearned Text-to-Image Diffusion Models
von: Shin, Jiwoo, et al.
Veröffentlicht: (2025)
von: Shin, Jiwoo, et al.
Veröffentlicht: (2025)
CLEFT: Language-Image Contrastive Learning with Efficient Large Language Model and Prompt Fine-Tuning
von: Du, Yuexi, et al.
Veröffentlicht: (2024)
von: Du, Yuexi, et al.
Veröffentlicht: (2024)
Cross-Domain Few-Shot Learning for Hyperspectral Image Classification Based on Mixup Foundation Model
von: Paeedeh, Naeem, et al.
Veröffentlicht: (2026)
von: Paeedeh, Naeem, et al.
Veröffentlicht: (2026)
Mastering Text-to-Image Diffusion: Recaptioning, Planning, and Generating with Multimodal LLMs
von: Yang, Ling, et al.
Veröffentlicht: (2024)
von: Yang, Ling, et al.
Veröffentlicht: (2024)
DiffBlender: Composable and Versatile Multimodal Text-to-Image Diffusion Models
von: Kim, Sungnyun, et al.
Veröffentlicht: (2023)
von: Kim, Sungnyun, et al.
Veröffentlicht: (2023)
Free$^2$Guide: Training-Free Text-to-Video Alignment using Image LVLM
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
von: Kim, Jaemin, et al.
Veröffentlicht: (2024)
Hyperdimensional Cross-Modal Alignment of Frozen Language and Image Models for Efficient Image Captioning
von: Dalvi, Abhishek, et al.
Veröffentlicht: (2026)
von: Dalvi, Abhishek, et al.
Veröffentlicht: (2026)
Lightweight Cloud Masking Models for On-Board Inference in Hyperspectral Imaging
von: Ali, Mazen, et al.
Veröffentlicht: (2025)
von: Ali, Mazen, et al.
Veröffentlicht: (2025)
Learning Relative Representations for Fine-Grained Multimodal Alignment with Limited Data
von: Kim, Shiwon, et al.
Veröffentlicht: (2026)
von: Kim, Shiwon, et al.
Veröffentlicht: (2026)
Secure and Storage-Efficient Deep Learning Models for Edge AI Using Automatic Weight Generation
von: Rahaman, Habibur, et al.
Veröffentlicht: (2025)
von: Rahaman, Habibur, et al.
Veröffentlicht: (2025)
Whose View of Safety? A Deep DIVE Dataset for Pluralistic Alignment of Text-to-Image Models
von: Rastogi, Charvi, et al.
Veröffentlicht: (2025)
von: Rastogi, Charvi, et al.
Veröffentlicht: (2025)
AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2026)
von: Kao, Kuei-Chun, et al.
Veröffentlicht: (2026)
TIAM -- A Metric for Evaluating Alignment in Text-to-Image Generation
von: Grimal, Paul, et al.
Veröffentlicht: (2023)
von: Grimal, Paul, et al.
Veröffentlicht: (2023)
PromptLoop: Plug-and-Play Prompt Refinement via Latent Feedback for Diffusion Model Alignment
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
von: Lee, Suhyeon, et al.
Veröffentlicht: (2025)
Skrr: Skip and Re-use Text Encoder Layers for Memory Efficient Text-to-Image Generation
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
von: Seo, Hoigi, et al.
Veröffentlicht: (2025)
Efficient Multi-turn RL for GUI Agents via Decoupled Training and Adaptive Data Curation
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
von: Li, Pengxiang, et al.
Veröffentlicht: (2025)
REP: Resource-Efficient Prompting for Rehearsal-Free Continual Learning
von: Jeon, Sungho, et al.
Veröffentlicht: (2024)
von: Jeon, Sungho, et al.
Veröffentlicht: (2024)
SCITUNE: Aligning Large Language Models with Human-Curated Scientific Multimodal Instructions
von: Horawalavithana, Sameera, et al.
Veröffentlicht: (2023)
von: Horawalavithana, Sameera, et al.
Veröffentlicht: (2023)
Automatic Data Curation for Self-Supervised Learning: A Clustering-Based Approach
von: Vo, Huy V., et al.
Veröffentlicht: (2024)
von: Vo, Huy V., et al.
Veröffentlicht: (2024)
Reconstruction Alignment Improves Unified Multimodal Models
von: Xie, Ji, et al.
Veröffentlicht: (2025)
von: Xie, Ji, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Self-supervised Learning for Hyperspectral Images of Trees
von: Rahman, Moqsadur, et al.
Veröffentlicht: (2025) -
Image Captions are Natural Prompts for Text-to-Image Models
von: Lei, Shiye, et al.
Veröffentlicht: (2023) -
Revisiting Aerial Scene Classification on the AID Benchmark
von: Das, Subhajeet, et al.
Veröffentlicht: (2026) -
Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models
von: Lee, Jaa-Yeon, et al.
Veröffentlicht: (2026) -
Hybrid Deep Learning for Hyperspectral Single Image Super-Resolution
von: Muhammad, Usman, et al.
Veröffentlicht: (2025)