Salvato in:
| Autori principali: | Chung, Jiwan, Lim, Seungwon, Lee, Sangkyu, Yu, Youngjae |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2501.11469 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Can visual language models resolve textual ambiguity with visual cues? Let visual puns tell you!
di: Chung, Jiwan, et al.
Pubblicazione: (2024)
di: Chung, Jiwan, et al.
Pubblicazione: (2024)
Teaching Metric Distance to Discrete Autoregressive Language Models
di: Chung, Jiwan, et al.
Pubblicazione: (2025)
di: Chung, Jiwan, et al.
Pubblicazione: (2025)
Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
di: Kim, Youngmin, et al.
Pubblicazione: (2025)
di: Kim, Youngmin, et al.
Pubblicazione: (2025)
Towards Visual Text Design Transfer Across Languages
di: Choi, Yejin, et al.
Pubblicazione: (2024)
di: Choi, Yejin, et al.
Pubblicazione: (2024)
CovMatch: Cross-Covariance Guided Multimodal Dataset Distillation with Trainable Text Encoder
di: Lee, Yongmin, et al.
Pubblicazione: (2025)
di: Lee, Yongmin, et al.
Pubblicazione: (2025)
CANVAS: Commonsense-Aware Navigation System for Intuitive Human-Robot Interaction
di: Choi, Suhwan, et al.
Pubblicazione: (2024)
di: Choi, Suhwan, et al.
Pubblicazione: (2024)
SelMatch: Effectively Scaling Up Dataset Distillation via Selection-Based Initialization and Partial Updates by Trajectory Matching
di: Lee, Yongmin, et al.
Pubblicazione: (2024)
di: Lee, Yongmin, et al.
Pubblicazione: (2024)
BiasConnect: Investigating Bias Interactions in Text-to-Image Models
di: Shukla, Pushkar, et al.
Pubblicazione: (2025)
di: Shukla, Pushkar, et al.
Pubblicazione: (2025)
VAGUE: Visual Contexts Clarify Ambiguous Expressions
di: Nam, Heejeong, et al.
Pubblicazione: (2024)
di: Nam, Heejeong, et al.
Pubblicazione: (2024)
Specify and Edit: Overcoming Ambiguity in Text-Based Image Editing
di: Iakovleva, Ekaterina, et al.
Pubblicazione: (2024)
di: Iakovleva, Ekaterina, et al.
Pubblicazione: (2024)
Active Learning for Finely-Categorized Image-Text Retrieval by Selecting Hard Negative Unpaired Samples
di: Jo, Dae Ung, et al.
Pubblicazione: (2024)
di: Jo, Dae Ung, et al.
Pubblicazione: (2024)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
di: Chung, Jiwan, et al.
Pubblicazione: (2025)
di: Chung, Jiwan, et al.
Pubblicazione: (2025)
FameBias: Embedding Manipulation Bias Attack in Text-to-Image Models
di: Roh, Jaechul, et al.
Pubblicazione: (2024)
di: Roh, Jaechul, et al.
Pubblicazione: (2024)
Alignment-Guided Score Matching for Text-to-Image Alignment in Diffusion Models
di: Lee, Jaa-Yeon, et al.
Pubblicazione: (2026)
di: Lee, Jaa-Yeon, et al.
Pubblicazione: (2026)
CVA: Context-aware Video-text Alignment for Video Temporal Grounding
di: Moon, Sungho, et al.
Pubblicazione: (2026)
di: Moon, Sungho, et al.
Pubblicazione: (2026)
Interaction Field Matching: Overcoming Limitations of Electrostatic Models
di: Manukhov, Stepan I., et al.
Pubblicazione: (2025)
di: Manukhov, Stepan I., et al.
Pubblicazione: (2025)
Model-Agnostic Gender Bias Control for Text-to-Image Generation via Sparse Autoencoder
di: Wu, Chao, et al.
Pubblicazione: (2025)
di: Wu, Chao, et al.
Pubblicazione: (2025)
Advanced Multimodal Deep Learning Architecture for Image-Text Matching
di: Wang, Jinyin, et al.
Pubblicazione: (2024)
di: Wang, Jinyin, et al.
Pubblicazione: (2024)
Rate-Adaptive Quantization: A Multi-Rate Codebook Adaptation for Vector Quantization-based Generative Models
di: Seo, Jiwan, et al.
Pubblicazione: (2024)
di: Seo, Jiwan, et al.
Pubblicazione: (2024)
FairImagen: Post-Processing for Bias Mitigation in Text-to-Image Models
di: Fu, Zihao, et al.
Pubblicazione: (2025)
di: Fu, Zihao, et al.
Pubblicazione: (2025)
Selective Vision is the Challenge for Visual Reasoning: A Benchmark for Visual Argument Understanding
di: Chung, Jiwan, et al.
Pubblicazione: (2024)
di: Chung, Jiwan, et al.
Pubblicazione: (2024)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
di: Lee, Dongyeun, et al.
Pubblicazione: (2025)
di: Lee, Dongyeun, et al.
Pubblicazione: (2025)
Overcome Modal Bias in Multi-modal Federated Learning via Balanced Modality Selection
di: Fan, Yunfeng, et al.
Pubblicazione: (2023)
di: Fan, Yunfeng, et al.
Pubblicazione: (2023)
BiasMap: Leveraging Cross-Attentions to Discover and Mitigate Hidden Social Biases in Text-to-Image Generation
di: Chakraborty, Rajatsubhra, et al.
Pubblicazione: (2025)
di: Chakraborty, Rajatsubhra, et al.
Pubblicazione: (2025)
Aligned but Stereotypical? The Hidden Influence of System Prompts on Social Bias in LVLM-Based Text-to-Image Models
di: Park, NaHyeon, et al.
Pubblicazione: (2025)
di: Park, NaHyeon, et al.
Pubblicazione: (2025)
Mitigating the Bias in the Model for Continual Test-Time Adaptation
di: Chung, Inseop, et al.
Pubblicazione: (2024)
di: Chung, Inseop, et al.
Pubblicazione: (2024)
Receler: Reliable Concept Erasing of Text-to-Image Diffusion Models via Lightweight Erasers
di: Huang, Chi-Pin, et al.
Pubblicazione: (2023)
di: Huang, Chi-Pin, et al.
Pubblicazione: (2023)
Concept Pinpoint Eraser for Text-to-image Diffusion Models via Residual Attention Gate
di: Lee, Byung Hyun, et al.
Pubblicazione: (2025)
di: Lee, Byung Hyun, et al.
Pubblicazione: (2025)
PMPGuard: Catching Pseudo-Matched Pairs in Remote Sensing Image-Text Retrieval
di: Ouyang, Pengxiang, et al.
Pubblicazione: (2025)
di: Ouyang, Pengxiang, et al.
Pubblicazione: (2025)
Hidden Bias in the Machine: Stereotypes in Text-to-Image Models
di: Porikli, Sedat, et al.
Pubblicazione: (2025)
di: Porikli, Sedat, et al.
Pubblicazione: (2025)
CoreMatching: A Co-adaptive Sparse Inference Framework with Token and Neuron Pruning for Comprehensive Acceleration of Vision-Language Models
di: Wang, Qinsi, et al.
Pubblicazione: (2025)
di: Wang, Qinsi, et al.
Pubblicazione: (2025)
PRISM: Video Dataset Condensation with Progressive Refinement and Insertion for Sparse Motion
di: Choi, Jaehyun, et al.
Pubblicazione: (2025)
di: Choi, Jaehyun, et al.
Pubblicazione: (2025)
Spanning Tree Autoregressive Visual Generation
di: Lee, Sangkyu, et al.
Pubblicazione: (2025)
di: Lee, Sangkyu, et al.
Pubblicazione: (2025)
Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies
di: Smith, Megan, et al.
Pubblicazione: (2026)
di: Smith, Megan, et al.
Pubblicazione: (2026)
Learned Image Compression with Text Quality Enhancement
di: Lai, Chih-Yu, et al.
Pubblicazione: (2024)
di: Lai, Chih-Yu, et al.
Pubblicazione: (2024)
ReflexSplit: Single Image Reflection Separation via Layer Fusion-Separation
di: Lee, Chia-Ming, et al.
Pubblicazione: (2026)
di: Lee, Chia-Ming, et al.
Pubblicazione: (2026)
MASS: MoErging through Adaptive Subspace Selection
di: Crisostomi, Donato, et al.
Pubblicazione: (2025)
di: Crisostomi, Donato, et al.
Pubblicazione: (2025)
Data Attribution for Text-to-Image Models by Unlearning Synthesized Images
di: Wang, Sheng-Yu, et al.
Pubblicazione: (2024)
di: Wang, Sheng-Yu, et al.
Pubblicazione: (2024)
Dual-Granularity Cross-Modal Identity Association for Weakly-Supervised Text-to-Person Image Matching
di: Zhang, Yafei, et al.
Pubblicazione: (2025)
di: Zhang, Yafei, et al.
Pubblicazione: (2025)
Bias Analysis in Unconditional Image Generative Models
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2025)
di: Zhang, Xiaofeng, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Can visual language models resolve textual ambiguity with visual cues? Let visual puns tell you!
di: Chung, Jiwan, et al.
Pubblicazione: (2024) -
Teaching Metric Distance to Discrete Autoregressive Language Models
di: Chung, Jiwan, et al.
Pubblicazione: (2025) -
Speaking Beyond Language: A Large-Scale Multimodal Dataset for Learning Nonverbal Cues from Video-Grounded Dialogues
di: Kim, Youngmin, et al.
Pubblicazione: (2025) -
Towards Visual Text Design Transfer Across Languages
di: Choi, Yejin, et al.
Pubblicazione: (2024) -
CovMatch: Cross-Covariance Guided Multimodal Dataset Distillation with Trainable Text Encoder
di: Lee, Yongmin, et al.
Pubblicazione: (2025)