Saved in:
| Main Authors: | Sohn, Jongwon, Moon, Juhyeon, Jung, Hyunjoon, Nam, Jaewook |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2512.01268 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
TLDR: Text Based Last-layer Retraining for Debiasing Image Classifiers
by: Park, Juhyeon, et al.
Published: (2023)
by: Park, Juhyeon, et al.
Published: (2023)
DEAL: Decoupled Classifier with Adaptive Linear Modulation for Group Robust Early Diagnosis of MCI to AD Conversion
by: Lee, Donggyu, et al.
Published: (2024)
by: Lee, Donggyu, et al.
Published: (2024)
Joint Reconstruction of 3D Human and Object via Contact-Based Refinement Transformer
by: Nam, Hyeongjin, et al.
Published: (2024)
by: Nam, Hyeongjin, et al.
Published: (2024)
Designing Extremely Memory-Efficient CNNs for On-device Vision Tasks
by: Lee, Jaewook, et al.
Published: (2024)
by: Lee, Jaewook, et al.
Published: (2024)
SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding
by: Park, Juhyeon, et al.
Published: (2025)
by: Park, Juhyeon, et al.
Published: (2025)
COREA: Coupled Relightable 3D Gaussians and SDFs for Efficient Normal Alignment
by: Lee, Jaeyoon, et al.
Published: (2025)
by: Lee, Jaeyoon, et al.
Published: (2025)
Can Natural Image Autoencoders Compactly Tokenize fMRI Volumes for Long-Range Dynamics Modeling?
by: Kim, Peter Yongho, et al.
Published: (2026)
by: Kim, Peter Yongho, et al.
Published: (2026)
GSFeatLoc: Visual Localization Using Feature Correspondence on 3D Gaussian Splatting
by: Lee, Jongwon, et al.
Published: (2025)
by: Lee, Jongwon, et al.
Published: (2025)
Group-wise Scaling and Orthogonal Decomposition for Domain-Invariant Feature Extraction in Face Anti-Spoofing
by: Jung, Seungjin, et al.
Published: (2025)
by: Jung, Seungjin, et al.
Published: (2025)
Domain-generalizable Face Anti-Spoofing with Patch-based Multi-tasking and Artifact Pattern Conversion
by: Jung, Seungjin, et al.
Published: (2026)
by: Jung, Seungjin, et al.
Published: (2026)
BEAF: Observing BEfore-AFter Changes to Evaluate Hallucination in Vision-language Models
by: Ye-Bin, Moon, et al.
Published: (2024)
by: Ye-Bin, Moon, et al.
Published: (2024)
Text-Guided Variational Image Generation for Industrial Anomaly Detection and Segmentation
by: Lee, Mingyu, et al.
Published: (2024)
by: Lee, Mingyu, et al.
Published: (2024)
CF3: Compact and Fast 3D Feature Fields
by: Lee, Hyunjoon, et al.
Published: (2025)
by: Lee, Hyunjoon, et al.
Published: (2025)
Entropy is not Enough for Test-Time Adaptation: From the Perspective of Disentangled Factors
by: Lee, Jonghyun, et al.
Published: (2024)
by: Lee, Jonghyun, et al.
Published: (2024)
LoREnc: Low-Rank Encryption for Securing Foundation Models and LoRA Adapters
by: Ahn, Beomjin, et al.
Published: (2026)
by: Ahn, Beomjin, et al.
Published: (2026)
Towards Model-Agnostic Dataset Condensation by Heterogeneous Models
by: Moon, Jun-Yeong, et al.
Published: (2024)
by: Moon, Jun-Yeong, et al.
Published: (2024)
VisionTrap: Vision-Augmented Trajectory Prediction Guided by Textual Descriptions
by: Moon, Seokha, et al.
Published: (2024)
by: Moon, Seokha, et al.
Published: (2024)
Bridging Vision and Language Spaces with Assignment Prediction
by: Park, Jungin, et al.
Published: (2024)
by: Park, Jungin, et al.
Published: (2024)
VLM's Eye Examination: Instruct and Inspect Visual Competency of Vision Language Models
by: Hyeon-Woo, Nam, et al.
Published: (2024)
by: Hyeon-Woo, Nam, et al.
Published: (2024)
How to Move Your Dragon: Text-to-Motion Synthesis for Large-Vocabulary Objects
by: Lee, Wonkwang, et al.
Published: (2025)
by: Lee, Wonkwang, et al.
Published: (2025)
PARTE: Part-Guided Texturing for 3D Human Reconstruction from a Single Image
by: Nam, Hyeongjin, et al.
Published: (2025)
by: Nam, Hyeongjin, et al.
Published: (2025)
M2SFormer: Multi-Spectral and Multi-Scale Attention with Edge-Aware Difficulty Guidance for Image Forgery Localization
by: Nam, Ju-Hyeon, et al.
Published: (2025)
by: Nam, Ju-Hyeon, et al.
Published: (2025)
Boost Your Human Image Generation Model via Direct Preference Optimization
by: Na, Sanghyeon, et al.
Published: (2024)
by: Na, Sanghyeon, et al.
Published: (2024)
InsideOut: Integrated RGB-Radiative Gaussian Splatting for Comprehensive 3D Object Representation
by: Lee, Jungmin, et al.
Published: (2025)
by: Lee, Jungmin, et al.
Published: (2025)
AJAHR: Amputated Joint Aware 3D Human Mesh Recovery
by: Cho, Hyunjin, et al.
Published: (2025)
by: Cho, Hyunjin, et al.
Published: (2025)
RePIC: Reinforced Post-Training for Personalizing Multi-Modal Language Models
by: Oh, Yeongtak, et al.
Published: (2025)
by: Oh, Yeongtak, et al.
Published: (2025)
Pick-or-Mix: Dynamic Channel Sampling for ConvNets
by: Kumar, Ashish, et al.
Published: (2024)
by: Kumar, Ashish, et al.
Published: (2024)
SF(DA)$^2$: Source-free Domain Adaptation Through the Lens of Data Augmentation
by: Hwang, Uiwon, et al.
Published: (2024)
by: Hwang, Uiwon, et al.
Published: (2024)
GLOS: Sign Language Generation with Temporally Aligned Gloss-Level Conditioning
by: Lee, Taeryung, et al.
Published: (2025)
by: Lee, Taeryung, et al.
Published: (2025)
Priority-Aware Clinical Pathology Hierarchy Training for Multiple Instance Learning
by: Hong, Sungrae, et al.
Published: (2025)
by: Hong, Sungrae, et al.
Published: (2025)
TPE-Net: Track Point Extraction and Association Network for Rail Path Proposal Generation
by: Kang, Jungwon, et al.
Published: (2023)
by: Kang, Jungwon, et al.
Published: (2023)
Embodied4C: Measuring What Matters for Embodied Vision-Language Navigation
by: Sohn, Tin Stribor, et al.
Published: (2025)
by: Sohn, Tin Stribor, et al.
Published: (2025)
MMIS-Net for Retinal Fluid Segmentation and Detection
by: Ndipenocha, Nchongmaje, et al.
Published: (2025)
by: Ndipenocha, Nchongmaje, et al.
Published: (2025)
How Can Objects Help Video-Language Understanding?
by: Tang, Zitian, et al.
Published: (2025)
by: Tang, Zitian, et al.
Published: (2025)
Pre-trained Vision and Language Transformers Are Few-Shot Incremental Learners
by: Park, Keon-Hee, et al.
Published: (2024)
by: Park, Keon-Hee, et al.
Published: (2024)
Transformer Based Building Boundary Reconstruction using Attraction Field Maps
by: Kamran, Muhammad, et al.
Published: (2025)
by: Kamran, Muhammad, et al.
Published: (2025)
DualFluidNet: an Attention-based Dual-pipeline Network for FLuid Simulation
by: Chen, Yu, et al.
Published: (2023)
by: Chen, Yu, et al.
Published: (2023)
R4: Retrieval-Augmented Reasoning for Vision-Language Models in 4D Spatio-Temporal Space
by: Sohn, Tin Stribor, et al.
Published: (2025)
by: Sohn, Tin Stribor, et al.
Published: (2025)
QTSeg: A Query Token-Based Dual-Mix Attention Framework with Multi-Level Feature Distribution for Medical Image Segmentation
by: Tran, Phuong-Nam, et al.
Published: (2024)
by: Tran, Phuong-Nam, et al.
Published: (2024)
Flaws of ImageNet, Computer Vision's Favourite Dataset
by: Kisel, Nikita, et al.
Published: (2024)
by: Kisel, Nikita, et al.
Published: (2024)
Similar Items
-
TLDR: Text Based Last-layer Retraining for Debiasing Image Classifiers
by: Park, Juhyeon, et al.
Published: (2023) -
DEAL: Decoupled Classifier with Adaptive Linear Modulation for Group Robust Early Diagnosis of MCI to AD Conversion
by: Lee, Donggyu, et al.
Published: (2024) -
Joint Reconstruction of 3D Human and Object via Contact-Based Refinement Transformer
by: Nam, Hyeongjin, et al.
Published: (2024) -
Designing Extremely Memory-Efficient CNNs for On-device Vision Tasks
by: Lee, Jaewook, et al.
Published: (2024) -
SEED: Towards More Accurate Semantic Evaluation for Visual Brain Decoding
by: Park, Juhyeon, et al.
Published: (2025)