Clustering-based Image-Text Graph Matching for Domain Generalization
Fuente:
arXiv
Salvato in:
| Autori principali: | Park, Nokyung, Chae, Daewon, Shim, Jeongyong, Kim, Sangpil, Kim, Eun-Sol, Kim, Jinkyu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2023
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
InstructBooth: Instruction-following Personalized Text-to-Image Generation
di: Chae, Daewon, et al.
Pubblicazione: (2023)
di: Chae, Daewon, et al.
Pubblicazione: (2023)
DiffExp: Efficient Exploration in Reward Fine-tuning for Text-to-Image Diffusion Models
di: Chae, Daewon, et al.
Pubblicazione: (2025)
di: Chae, Daewon, et al.
Pubblicazione: (2025)
SemanticControl: A Training-Free Approach for Handling Loosely Aligned Visual Conditions in ControlNet
di: Joung, Woosung, et al.
Pubblicazione: (2025)
di: Joung, Woosung, et al.
Pubblicazione: (2025)
Scene Graph-Guided Proactive Replanning for Failure-Resilient Embodied Agent
di: Yu, Che Rin, et al.
Pubblicazione: (2025)
di: Yu, Che Rin, et al.
Pubblicazione: (2025)
FPANet: Frequency-based Video Demoireing using Frame-level Post Alignment
di: Oh, Gyeongrok, et al.
Pubblicazione: (2023)
di: Oh, Gyeongrok, et al.
Pubblicazione: (2023)
Finetuning Pre-trained Model with Limited Data for LiDAR-based 3D Object Detection by Bridging Domain Gaps
di: Jang, Jiyun, et al.
Pubblicazione: (2024)
di: Jang, Jiyun, et al.
Pubblicazione: (2024)
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
di: Park, Minjeong, et al.
Pubblicazione: (2025)
di: Park, Minjeong, et al.
Pubblicazione: (2025)
Text-Aware Image Restoration with Diffusion Models
di: Min, Jaewon, et al.
Pubblicazione: (2025)
di: Min, Jaewon, et al.
Pubblicazione: (2025)
MambaEye: A Size-Agnostic Visual Encoder with Causal Sequential Processing
di: Choi, Changho, et al.
Pubblicazione: (2025)
di: Choi, Changho, et al.
Pubblicazione: (2025)
Image Clustering Conditioned on Text Criteria
di: Kwon, Sehyun, et al.
Pubblicazione: (2023)
di: Kwon, Sehyun, et al.
Pubblicazione: (2023)
MEVG: Multi-event Video Generation with Text-to-Video Models
di: Oh, Gyeongrok, et al.
Pubblicazione: (2023)
di: Oh, Gyeongrok, et al.
Pubblicazione: (2023)
3D Occupancy Prediction with Low-Resolution Queries via Prototype-aware View Transformation
di: Oh, Gyeongrok, et al.
Pubblicazione: (2025)
di: Oh, Gyeongrok, et al.
Pubblicazione: (2025)
PASTA: Part-Aware Sketch-to-3D Shape Generation with Text-Aligned Prior
di: Lee, Seunggwan, et al.
Pubblicazione: (2025)
di: Lee, Seunggwan, et al.
Pubblicazione: (2025)
Bridging the Domain Gap: A Simple Domain Matching Method for Reference-based Image Super-Resolution in Remote Sensing
di: Min, Jeongho, et al.
Pubblicazione: (2024)
di: Min, Jeongho, et al.
Pubblicazione: (2024)
Image-Guided Semantic Pseudo-LiDAR Point Generation for 3D Object Detection
di: Lee, Minseung, et al.
Pubblicazione: (2024)
di: Lee, Minseung, et al.
Pubblicazione: (2024)
Rethinking Data Augmentation for Robust LiDAR Semantic Segmentation in Adverse Weather
di: Park, Junsung, et al.
Pubblicazione: (2024)
di: Park, Junsung, et al.
Pubblicazione: (2024)
LEAP:D -- A Novel Prompt-based Approach for Domain-Generalized Aerial Object Detection
di: Park, Chanyeong, et al.
Pubblicazione: (2024)
di: Park, Chanyeong, et al.
Pubblicazione: (2024)
Addressing Image Hallucination in Text-to-Image Generation through Factual Image Retrieval
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
Real-Time Person Image Synthesis Using a Flow Matching Model
di: Jeong, Jiwoo, et al.
Pubblicazione: (2025)
di: Jeong, Jiwoo, et al.
Pubblicazione: (2025)
Improving Text-to-Image Generation with Intrinsic Self-Confidence Rewards
di: Kim, Seungwook, et al.
Pubblicazione: (2026)
di: Kim, Seungwook, et al.
Pubblicazione: (2026)
Dynamic-eDiTor: Training-Free Text-Driven 4D Scene Editing with Multimodal Diffusion Transformer
di: Lee, Dong In, et al.
Pubblicazione: (2025)
di: Lee, Dong In, et al.
Pubblicazione: (2025)
Guiding What Not to Generate: Automated Negative Prompting for Text-Image Alignment
di: Park, Sangha, et al.
Pubblicazione: (2025)
di: Park, Sangha, et al.
Pubblicazione: (2025)
Unified Domain Generalization and Adaptation for Multi-View 3D Object Detection
di: Chang, Gyusam, et al.
Pubblicazione: (2024)
di: Chang, Gyusam, et al.
Pubblicazione: (2024)
Evaluating Image Hallucination in Text-to-Image Generation with Question-Answering
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
di: Lim, Youngsun, et al.
Pubblicazione: (2024)
EditSplat: Multi-View Fusion and Attention-Guided Optimization for View-Consistent 3D Scene Editing with 3D Gaussian Splatting
di: Lee, Dong In, et al.
Pubblicazione: (2024)
di: Lee, Dong In, et al.
Pubblicazione: (2024)
Classifier-guided CLIP Distillation for Unsupervised Multi-label Classification
di: Kim, Dongseob, et al.
Pubblicazione: (2025)
di: Kim, Dongseob, et al.
Pubblicazione: (2025)
Semantic Diversity-aware Prototype-based Learning for Unbiased Scene Graph Generation
di: Jeon, Jaehyeong, et al.
Pubblicazione: (2024)
di: Jeon, Jaehyeong, et al.
Pubblicazione: (2024)
RaDL: Relation-aware Disentangled Learning for Multi-Instance Text-to-Image Generation
di: Park, Geon, et al.
Pubblicazione: (2025)
di: Park, Geon, et al.
Pubblicazione: (2025)
Anchoring and Rescaling Attention for Semantically Coherent Inbetweening
di: Choi, Tae Eun, et al.
Pubblicazione: (2026)
di: Choi, Tae Eun, et al.
Pubblicazione: (2026)
CompMarkGS: Robust Watermarking for Compressed 3D Gaussian Splatting
di: In, Sumin, et al.
Pubblicazione: (2025)
di: In, Sumin, et al.
Pubblicazione: (2025)
Knowledge-based learning in Text-RAG and Image-RAG
di: Shim, Alexander, et al.
Pubblicazione: (2026)
di: Shim, Alexander, et al.
Pubblicazione: (2026)
Adaptive Self-training Framework for Fine-grained Scene Graph Generation
di: Kim, Kibum, et al.
Pubblicazione: (2024)
di: Kim, Kibum, et al.
Pubblicazione: (2024)
Overcoming Data Inequality across Domains with Semi-Supervised Domain Generalization
di: Park, Jinha, et al.
Pubblicazione: (2024)
di: Park, Jinha, et al.
Pubblicazione: (2024)
Watermarking for Factuality: Guiding Vision-Language Models Toward Truth via Tri-layer Contrastive Decoding
di: Back, Kyungryul, et al.
Pubblicazione: (2025)
di: Back, Kyungryul, et al.
Pubblicazione: (2025)
Learning Primitive Relations for Compositional Zero-Shot Learning
di: Lee, Insu, et al.
Pubblicazione: (2025)
di: Lee, Insu, et al.
Pubblicazione: (2025)
TextBoost: Boosting Text Encoder for Personalized Text-to-Image Generation
di: Park, NaHyeon, et al.
Pubblicazione: (2024)
di: Park, NaHyeon, et al.
Pubblicazione: (2024)
Unexplored Faces of Robustness and Out-of-Distribution: Covariate Shifts in Environment and Sensor Domains
di: Baek, Eunsu, et al.
Pubblicazione: (2024)
di: Baek, Eunsu, et al.
Pubblicazione: (2024)
LaMoGen: Laban Movement-Guided Diffusion for Text-to-Motion Generation
di: Kim, Heechang, et al.
Pubblicazione: (2025)
di: Kim, Heechang, et al.
Pubblicazione: (2025)
Local Representative Token Guided Merging for Text-to-Image Generation
di: Lee, Min-Jeong, et al.
Pubblicazione: (2025)
di: Lee, Min-Jeong, et al.
Pubblicazione: (2025)
Focus, Don't Prune: Identifying Instruction-Relevant Regions for Information-Rich Image Understanding
di: Kwon, Mincheol, et al.
Pubblicazione: (2026)
di: Kwon, Mincheol, et al.
Pubblicazione: (2026)
Documenti analoghi
-
InstructBooth: Instruction-following Personalized Text-to-Image Generation
di: Chae, Daewon, et al.
Pubblicazione: (2023) -
DiffExp: Efficient Exploration in Reward Fine-tuning for Text-to-Image Diffusion Models
di: Chae, Daewon, et al.
Pubblicazione: (2025) -
SemanticControl: A Training-Free Approach for Handling Loosely Aligned Visual Conditions in ControlNet
di: Joung, Woosung, et al.
Pubblicazione: (2025) -
Scene Graph-Guided Proactive Replanning for Failure-Resilient Embodied Agent
di: Yu, Che Rin, et al.
Pubblicazione: (2025) -
FPANet: Frequency-based Video Demoireing using Frame-level Post Alignment
di: Oh, Gyeongrok, et al.
Pubblicazione: (2023)