Prompt Augmentation for Self-supervised Text-guided Image Manipulation
Fuente:
arXiv
Saved in:
| Main Authors: | Bodur, Rumeysa, Bhattarai, Binod, Kim, Tae-Kyun |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SRHand: Super-Resolving Hand Images and 3D Shapes via View/Pose-aware Neural Image Representations and Explicit 3D Meshes
by: Kim, Minje, et al.
Published: (2025)
by: Kim, Minje, et al.
Published: (2025)
Arbitrary-Scale Image Generation and Upsampling using Latent Diffusion Model and Implicit Neural Decoder
by: Kim, Jinseok, et al.
Published: (2024)
by: Kim, Jinseok, et al.
Published: (2024)
BiTT: Bi-directional Texture Reconstruction of Interacting Two Hands from a Single Image
by: Kim, Minje, et al.
Published: (2024)
by: Kim, Minje, et al.
Published: (2024)
Semi-Supervised 3D Object Detection with Channel Augmentation using Transformation Equivariance
by: Kang, Minju, et al.
Published: (2024)
by: Kang, Minju, et al.
Published: (2024)
LieHMR: Autoregressive Human Mesh Recovery with $SO(3)$ Diffusion
by: Kim, Donghwan, et al.
Published: (2025)
by: Kim, Donghwan, et al.
Published: (2025)
Estimating 2D Keypoints of Surgical Tools Using Vision-Language Models with Low-Rank Adaptation
by: Duangprom, Krit, et al.
Published: (2025)
by: Duangprom, Krit, et al.
Published: (2025)
Learning Adaptive Pseudo-Label Selection for Semi-Supervised 3D Object Detection
by: Kong, Taehun, et al.
Published: (2025)
by: Kong, Taehun, et al.
Published: (2025)
Body-Hand Modality Expertized Networks with Cross-attention for Fine-grained Skeleton Action Recognition
by: Cho, Seungyeon, et al.
Published: (2025)
by: Cho, Seungyeon, et al.
Published: (2025)
Joint Learning of Pose Regression and Denoising Diffusion with Score Scaling Sampling for Category-level 6D Pose Estimation
by: Lee, Seunghyun, et al.
Published: (2025)
by: Lee, Seunghyun, et al.
Published: (2025)
Multi-hypotheses Conditioned Point Cloud Diffusion for 3D Human Reconstruction from Occluded Images
by: Kim, Donghwan, et al.
Published: (2024)
by: Kim, Donghwan, et al.
Published: (2024)
CAR-MFL: Cross-Modal Augmentation by Retrieval for Multimodal Federated Learning with Missing Modalities
by: Poudel, Pranav, et al.
Published: (2024)
by: Poudel, Pranav, et al.
Published: (2024)
Active Label Refinement for Robust Training of Imbalanced Medical Image Classification Tasks in the Presence of High Label Noise
by: Khanal, Bidur, et al.
Published: (2024)
by: Khanal, Bidur, et al.
Published: (2024)
How does self-supervised pretraining improve robustness against noisy labels across various medical image classification datasets?
by: Khanal, Bidur, et al.
Published: (2024)
by: Khanal, Bidur, et al.
Published: (2024)
Co-learning Single-Step Diffusion Upsampler and Downsampler with Two Discriminators and Distillation
by: Kim, Sohwi, et al.
Published: (2024)
by: Kim, Sohwi, et al.
Published: (2024)
Cascaded Diffusion Framework for Probabilistic Coarse-to-Fine Hand Pose Estimation
by: Woo, Taeyun, et al.
Published: (2025)
by: Woo, Taeyun, et al.
Published: (2025)
Generalist Multi-Class Anomaly Detection via Distillation to Two Heterogeneous Student Networks
by: Park, Hangil, et al.
Published: (2025)
by: Park, Hangil, et al.
Published: (2025)
Effect of Data Augmentation on Conformal Prediction for Diabetic Retinopathy
by: Ahamed, Rizwan, et al.
Published: (2025)
by: Ahamed, Rizwan, et al.
Published: (2025)
DySurface: Consistent 4D Surface Reconstruction via Bridging Explicit Gaussians and Implicit Functions
by: Kim, Minje, et al.
Published: (2026)
by: Kim, Minje, et al.
Published: (2026)
TE-SSL: Time and Event-aware Self Supervised Learning for Alzheimer's Disease Progression Analysis
by: Thrasher, Jacob, et al.
Published: (2024)
by: Thrasher, Jacob, et al.
Published: (2024)
SSiT: Saliency-guided Self-supervised Image Transformer for Diabetic Retinopathy Grading
by: Huang, Yijin, et al.
Published: (2022)
by: Huang, Yijin, et al.
Published: (2022)
TTA-OOD: Test-time Augmentation for Improving Out-of-Distribution Detection in Gastrointestinal Vision
by: Pokhrel, Sandesh, et al.
Published: (2024)
by: Pokhrel, Sandesh, et al.
Published: (2024)
Dynamic Full-body Motion Agent with Object Interaction via Blending Pre-trained Modular Controllers
by: Nam, Sanghyeok, et al.
Published: (2026)
by: Nam, Sanghyeok, et al.
Published: (2026)
MPMAvatar: Learning 3D Gaussian Avatars with Accurate and Robust Physics-Based Dynamics
by: Lee, Changmin, et al.
Published: (2025)
by: Lee, Changmin, et al.
Published: (2025)
NERO: Explainable Out-of-Distribution Detection with Neuron-level Relevance
by: Chhetri, Anju, et al.
Published: (2025)
by: Chhetri, Anju, et al.
Published: (2025)
Linguistics-aware Masked Image Modeling for Self-supervised Scene Text Recognition
by: Zhang, Yifei, et al.
Published: (2025)
by: Zhang, Yifei, et al.
Published: (2025)
3DPhysVideo: Consistency-Guided Flow SDE for Video Generation via 3D Scene Reconstruction and Physical Simulation
by: Kim, Hwidong, et al.
Published: (2026)
by: Kim, Hwidong, et al.
Published: (2026)
OmniText: A Training-Free Generalist for Controllable Text-Image Manipulation
by: Gunawan, Agus, et al.
Published: (2025)
by: Gunawan, Agus, et al.
Published: (2025)
Energy-based Domain-Adaptive Segmentation with Depth Guidance
by: Zhu, Jinjing, et al.
Published: (2024)
by: Zhu, Jinjing, et al.
Published: (2024)
Text-guided Visual Prompt DINO for Generic Segmentation
by: Guan, Yuchen, et al.
Published: (2025)
by: Guan, Yuchen, et al.
Published: (2025)
Hand-object reconstruction via interaction-aware graph attention mechanism
by: Woo, Taeyun, et al.
Published: (2024)
by: Woo, Taeyun, et al.
Published: (2024)
MGHanD: Multi-modal Guidance for authentic Hand Diffusion
by: Eum, Taehyeon, et al.
Published: (2025)
by: Eum, Taehyeon, et al.
Published: (2025)
Relational Self-supervised Distillation with Compact Descriptors for Image Copy Detection
by: Kim, Juntae, et al.
Published: (2024)
by: Kim, Juntae, et al.
Published: (2024)
Visual Delta Generator with Large Multi-modal Models for Semi-supervised Composed Image Retrieval
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Semantic-guided Adversarial Diffusion Model for Self-supervised Shadow Removal
by: Zeng, Ziqi, et al.
Published: (2024)
by: Zeng, Ziqi, et al.
Published: (2024)
A Review of Predictive and Contrastive Self-supervised Learning for Medical Images
by: Wang, Wei-Chien, et al.
Published: (2023)
by: Wang, Wei-Chien, et al.
Published: (2023)
MATE: Meet At The Embedding -- Connecting Images with Long Texts
by: Jang, Young Kyun, et al.
Published: (2024)
by: Jang, Young Kyun, et al.
Published: (2024)
Cross-Task Data Augmentation by Pseudo-label Generation for Region Based Coronary Artery Instance Segmentation
by: Pokhrel, Sandesh, et al.
Published: (2023)
by: Pokhrel, Sandesh, et al.
Published: (2023)
Towards Scalable Human-aligned Benchmark for Text-guided Image Editing
by: Ryu, Suho, et al.
Published: (2025)
by: Ryu, Suho, et al.
Published: (2025)
FAGStyle: Feature Augmentation on Geodesic Surface for Zero-shot Text-guided Diffusion Image Style Transfer
by: Han, Yuexing, et al.
Published: (2024)
by: Han, Yuexing, et al.
Published: (2024)
InterHandGen: Two-Hand Interaction Generation via Cascaded Reverse Diffusion
by: Lee, Jihyun, et al.
Published: (2024)
by: Lee, Jihyun, et al.
Published: (2024)
Similar Items
-
SRHand: Super-Resolving Hand Images and 3D Shapes via View/Pose-aware Neural Image Representations and Explicit 3D Meshes
by: Kim, Minje, et al.
Published: (2025) -
Arbitrary-Scale Image Generation and Upsampling using Latent Diffusion Model and Implicit Neural Decoder
by: Kim, Jinseok, et al.
Published: (2024) -
BiTT: Bi-directional Texture Reconstruction of Interacting Two Hands from a Single Image
by: Kim, Minje, et al.
Published: (2024) -
Semi-Supervised 3D Object Detection with Channel Augmentation using Transformation Equivariance
by: Kang, Minju, et al.
Published: (2024) -
LieHMR: Autoregressive Human Mesh Recovery with $SO(3)$ Diffusion
by: Kim, Donghwan, et al.
Published: (2025)