Saved in:
| Main Authors: | Lee, Taekyung, Lee, Donggyu, Kang, Myungjoo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2506.01370 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Spatial-and-Frequency-aware Restoration method for Images based on Diffusion Models
by: Lee, Kyungsung, et al.
Published: (2024)
by: Lee, Kyungsung, et al.
Published: (2024)
Unsupervised Point Cloud Completion through Unbalanced Optimal Transport
by: Lee, Taekyung, et al.
Published: (2024)
by: Lee, Taekyung, et al.
Published: (2024)
MINR: Implicit Neural Representations with Masked Image Modelling
by: Lee, Sua, et al.
Published: (2025)
by: Lee, Sua, et al.
Published: (2025)
FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model
by: Lee, Yebin, et al.
Published: (2024)
by: Lee, Yebin, et al.
Published: (2024)
Analyzing and Improving Optimal-Transport-based Adversarial Networks
by: Choi, Jaemoo, et al.
Published: (2023)
by: Choi, Jaemoo, et al.
Published: (2023)
Unsupervised training of keypoint-agnostic descriptors for flexible retinal image registration
by: Rivas-Villar, David, et al.
Published: (2025)
by: Rivas-Villar, David, et al.
Published: (2025)
ConKeD: Multiview contrastive descriptor learning for keypoint-based retinal image registration
by: Rivas-Villar, David, et al.
Published: (2024)
by: Rivas-Villar, David, et al.
Published: (2024)
Affine steerers for structured keypoint description
by: Bökman, Georg, et al.
Published: (2024)
by: Bökman, Georg, et al.
Published: (2024)
Spatial regularisation for improved accuracy and interpretability in keypoint-based registration
by: Billot, Benjamin, et al.
Published: (2025)
by: Billot, Benjamin, et al.
Published: (2025)
DEAL: Decoupled Classifier with Adaptive Linear Modulation for Group Robust Early Diagnosis of MCI to AD Conversion
by: Lee, Donggyu, et al.
Published: (2024)
by: Lee, Donggyu, et al.
Published: (2024)
UKDM: Underwater keypoint detection and matching using underwater image enhancement techniques
by: Diaz-Garcia, Pedro, et al.
Published: (2025)
by: Diaz-Garcia, Pedro, et al.
Published: (2025)
Universal Image Immunization against Diffusion-based Image Editing via Semantic Injection
by: Lee, Chanhui, et al.
Published: (2026)
by: Lee, Chanhui, et al.
Published: (2026)
Retaining and Enhancing Pre-trained Knowledge in Vision-Language Models with Prompt Ensembling
by: Kim, Donggeun, et al.
Published: (2024)
by: Kim, Donggeun, et al.
Published: (2024)
Consistent text-to-image generation via scene de-contextualization
by: Tang, Song, et al.
Published: (2025)
by: Tang, Song, et al.
Published: (2025)
Steerers: A framework for rotation equivariant keypoint descriptors
by: Bökman, Georg, et al.
Published: (2023)
by: Bökman, Georg, et al.
Published: (2023)
Generative Modeling through the Semi-dual Formulation of Unbalanced Optimal Transport
by: Choi, Jaemoo, et al.
Published: (2023)
by: Choi, Jaemoo, et al.
Published: (2023)
Scalable Wasserstein Gradient Flow for Generative Modeling through Unbalanced Optimal Transport
by: Choi, Jaemoo, et al.
Published: (2024)
by: Choi, Jaemoo, et al.
Published: (2024)
Visual question answering based evaluation metrics for text-to-image generation
by: Miyamoto, Mizuki, et al.
Published: (2024)
by: Miyamoto, Mizuki, et al.
Published: (2024)
Sequential keypoint density estimator: an overlooked baseline of skeleton-based video anomaly detection
by: Delić, Anja, et al.
Published: (2025)
by: Delić, Anja, et al.
Published: (2025)
Exploring text-to-image generation for historical document image retrieval
by: Cote, Melissa, et al.
Published: (2025)
by: Cote, Melissa, et al.
Published: (2025)
Leveraging Prior Knowledge of Diffusion Model for Person Search
by: Kim, Giyeol, et al.
Published: (2025)
by: Kim, Giyeol, et al.
Published: (2025)
Short-Window Sliding Learning for Real-Time Violence Detection via LLM-based Auto-Labeling
by: Jung, Seoik, et al.
Published: (2025)
by: Jung, Seoik, et al.
Published: (2025)
New keypoint-based approach for recognising British Sign Language (BSL) from sequences
by: Deb, Oishi, et al.
Published: (2024)
by: Deb, Oishi, et al.
Published: (2024)
IMAGE-ALCHEMY: Advancing subject fidelity in personalised text-to-image generation
by: Tiwari, Amritanshu, et al.
Published: (2025)
by: Tiwari, Amritanshu, et al.
Published: (2025)
Pseudo-keypoint RKHS Learning for Self-supervised 6DoF Pose Estimation
by: Wu, Yangzheng, et al.
Published: (2023)
by: Wu, Yangzheng, et al.
Published: (2023)
A Billion-scale Foundation Model for Remote Sensing Images
by: Cha, Keumgang, et al.
Published: (2023)
by: Cha, Keumgang, et al.
Published: (2023)
Trinity Detector:text-assisted and attention mechanisms based spectral fusion for diffusion generation image detection
by: Song, Jiawei, et al.
Published: (2024)
by: Song, Jiawei, et al.
Published: (2024)
DUAL-VAD: Dual Benchmarks and Anomaly-Focused Sampling for Video Anomaly Detection
by: Jung, Seoik, et al.
Published: (2025)
by: Jung, Seoik, et al.
Published: (2025)
Dark Miner: Defend against undesirable generation for text-to-image diffusion models
by: Meng, Zheling, et al.
Published: (2024)
by: Meng, Zheling, et al.
Published: (2024)
Attend-and-Refine: Interactive keypoint estimation and quantitative cervical vertebrae analysis for bone age assessment
by: Kim, Jinhee, et al.
Published: (2025)
by: Kim, Jinhee, et al.
Published: (2025)
FastPoint: Accelerating 3D Point Cloud Model Inference via Sample Point Distance Prediction
by: Lee, Donghyun, et al.
Published: (2025)
by: Lee, Donghyun, et al.
Published: (2025)
Training-free Detection of AI-generated images via Cropping Robustness
by: Choi, Sungik, et al.
Published: (2025)
by: Choi, Sungik, et al.
Published: (2025)
Cross-Modal Emotion Transfer for Emotion Editing in Talking Face Video
by: Choi, Chanhyuk, et al.
Published: (2026)
by: Choi, Chanhyuk, et al.
Published: (2026)
Environmental Understanding Vision-Language Model for Embodied Agent
by: Bang, Jinsik, et al.
Published: (2026)
by: Bang, Jinsik, et al.
Published: (2026)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
by: Ki, Taekyung, et al.
Published: (2023)
by: Ki, Taekyung, et al.
Published: (2023)
Generalized Zero-Shot Learning for Point Cloud Segmentation with Evidence-Based Dynamic Calibration
by: Kim, Hyeonseok, et al.
Published: (2025)
by: Kim, Hyeonseok, et al.
Published: (2025)
Enhancement of text recognition for hanja handwritten documents of Ancient Korea
by: Ahna, Joonmo, et al.
Published: (2024)
by: Ahna, Joonmo, et al.
Published: (2024)
Unsupervised learning of Data-driven Facial Expression Coding System (DFECS) using keypoint tracking
by: Tripathi, Shivansh Chandra, et al.
Published: (2024)
by: Tripathi, Shivansh Chandra, et al.
Published: (2024)
Map the Flow: Revealing Hidden Pathways of Information in VideoLLMs
by: Kim, Minji, et al.
Published: (2025)
by: Kim, Minji, et al.
Published: (2025)
Morphing Tokens Draw Strong Masked Image Models
by: Kim, Taekyung, et al.
Published: (2023)
by: Kim, Taekyung, et al.
Published: (2023)
Similar Items
-
Spatial-and-Frequency-aware Restoration method for Images based on Diffusion Models
by: Lee, Kyungsung, et al.
Published: (2024) -
Unsupervised Point Cloud Completion through Unbalanced Optimal Transport
by: Lee, Taekyung, et al.
Published: (2024) -
MINR: Implicit Neural Representations with Masked Image Modelling
by: Lee, Sua, et al.
Published: (2025) -
FLEUR: An Explainable Reference-Free Evaluation Metric for Image Captioning Using a Large Multimodal Model
by: Lee, Yebin, et al.
Published: (2024) -
Analyzing and Improving Optimal-Transport-based Adversarial Networks
by: Choi, Jaemoo, et al.
Published: (2023)