Exploring and Leveraging Class Vectors for Classifier Editing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Jaeik, Do, Jaeyoung |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
SECOND: Mitigating Perceptual Hallucination in Vision-Language Models via Selective and Contrastive Decoding
von: Park, Woohyeon, et al.
Veröffentlicht: (2025)
von: Park, Woohyeon, et al.
Veröffentlicht: (2025)
MMPB: It's Time for Multi-Modal Personalization
von: Kim, Jaeik, et al.
Veröffentlicht: (2025)
von: Kim, Jaeik, et al.
Veröffentlicht: (2025)
VIRST: Video-Instructed Reasoning Assistant for SpatioTemporal Segmentation
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
3rd Place of MeViS-Audio Track of the 5th PVUW: VIRST-Audio
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
von: Hong, Jihwan, et al.
Veröffentlicht: (2026)
MEDIC-AD: Towards Medical Vision-Language Model's Clinical Intelligence
von: Park, Woohyeon, et al.
Veröffentlicht: (2026)
von: Park, Woohyeon, et al.
Veröffentlicht: (2026)
MI-CXR: A Benchmark for Longitudinal Reasoning over Multi-Interval Chest X-rays
von: Cho, Sunghwan Steve, et al.
Veröffentlicht: (2026)
von: Cho, Sunghwan Steve, et al.
Veröffentlicht: (2026)
DALDA: Data Augmentation Leveraging Diffusion Model and LLM with Adaptive Guidance Scaling
von: Jung, Kyuheon, et al.
Veröffentlicht: (2024)
von: Jung, Kyuheon, et al.
Veröffentlicht: (2024)
Uni-Classifier: Leveraging Video Diffusion Priors for Universal Guidance Classifier
von: Zhou, Yujie, et al.
Veröffentlicht: (2026)
von: Zhou, Yujie, et al.
Veröffentlicht: (2026)
Exploring Multimodal Diffusion Transformers for Enhanced Prompt-based Image Editing
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
von: Shin, Joonghyuk, et al.
Veröffentlicht: (2025)
Predicting the Reliability of an Image Classifier under Image Distortion
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
von: Nguyen, Dang, et al.
Veröffentlicht: (2024)
EggHand: A Multimodal Foundation Model for Egocentric Hand Pose Forecasting
von: Choi, Jaeyoung, et al.
Veröffentlicht: (2026)
von: Choi, Jaeyoung, et al.
Veröffentlicht: (2026)
Class Prototypes based Contrastive Learning for Classifying Multi-Label and Fine-Grained Educational Videos
von: Gupta, Rohit, et al.
Veröffentlicht: (2025)
von: Gupta, Rohit, et al.
Veröffentlicht: (2025)
Leveraging Verifier-Based Reinforcement Learning in Image Editing
von: Guo, Hanzhong, et al.
Veröffentlicht: (2026)
von: Guo, Hanzhong, et al.
Veröffentlicht: (2026)
Image-to-Image Translation with Disentangled Latent Vectors for Face Editing
von: Dalva, Yusuf, et al.
Veröffentlicht: (2023)
von: Dalva, Yusuf, et al.
Veröffentlicht: (2023)
Adaptive Margin Global Classifier for Exemplar-Free Class-Incremental Learning
von: Yao, Zhongren, et al.
Veröffentlicht: (2024)
von: Yao, Zhongren, et al.
Veröffentlicht: (2024)
Salvaging the Overlooked: Leveraging Class-Aware Contrastive Learning for Multi-Class Anomaly Detection
von: Fan, Lei, et al.
Veröffentlicht: (2024)
von: Fan, Lei, et al.
Veröffentlicht: (2024)
Multi-Class Textual-Inversion Secretly Yields a Semantic-Agnostic Classifier
von: Wang, Kai, et al.
Veröffentlicht: (2024)
von: Wang, Kai, et al.
Veröffentlicht: (2024)
Leveraging Enhanced Queries of Point Sets for Vectorized Map Construction
von: Liu, Zihao, et al.
Veröffentlicht: (2024)
von: Liu, Zihao, et al.
Veröffentlicht: (2024)
ACPV-Net: All-Class Polygonal Vectorization for Seamless Vector Map Generation from Aerial Imagery
von: Jiao, Weiqin, et al.
Veröffentlicht: (2026)
von: Jiao, Weiqin, et al.
Veröffentlicht: (2026)
LIVE: Leveraging Image Manipulation Priors for Instruction-based Video Editing
von: Wang, Weicheng, et al.
Veröffentlicht: (2026)
von: Wang, Weicheng, et al.
Veröffentlicht: (2026)
h-Edit: Effective and Flexible Diffusion-Based Editing via Doob's h-Transform
von: Nguyen, Toan, et al.
Veröffentlicht: (2025)
von: Nguyen, Toan, et al.
Veröffentlicht: (2025)
FastVideoEdit: Leveraging Consistency Models for Efficient Text-to-Video Editing
von: Zhang, Youyuan, et al.
Veröffentlicht: (2024)
von: Zhang, Youyuan, et al.
Veröffentlicht: (2024)
Personalized Scientific Figure Caption Generation: An Empirical Study on Author-Specific Writing Style Transfer
von: Kim, Jaeyoung, et al.
Veröffentlicht: (2025)
von: Kim, Jaeyoung, et al.
Veröffentlicht: (2025)
Embedding Space Allocation with Angle-Norm Joint Classifiers for Few-Shot Class-Incremental Learning
von: Tu, Dunwei, et al.
Veröffentlicht: (2024)
von: Tu, Dunwei, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models For Scalable Vector Graphics Processing: A Review
von: Malashenko, Boris, et al.
Veröffentlicht: (2025)
von: Malashenko, Boris, et al.
Veröffentlicht: (2025)
Vectorized Video Representation with Easy Editing via Hierarchical Spatio-Temporally Consistent Proxy Embedding
von: Chen, Ye, et al.
Veröffentlicht: (2025)
von: Chen, Ye, et al.
Veröffentlicht: (2025)
Exploring Iterative Manifold Constraint for Zero-shot Image Editing
von: Li, Maomao, et al.
Veröffentlicht: (2025)
von: Li, Maomao, et al.
Veröffentlicht: (2025)
Bridging the Skeleton-Text Modality Gap: Diffusion-Powered Modality Alignment for Zero-shot Skeleton-based Action Recognition
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
SkateFormer: Skeletal-Temporal Transformer for Human Action Recognition
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
VLMDiff: Leveraging Vision-Language Models for Multi-Class Anomaly Detection with Diffusion
von: Hicsonmez, Samet, et al.
Veröffentlicht: (2025)
von: Hicsonmez, Samet, et al.
Veröffentlicht: (2025)
C-DiffSET: Leveraging Latent Diffusion for SAR-to-EO Image Translation with Confidence-Guided Reliable Object Generation
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
von: Do, Jeonghyeok, et al.
Veröffentlicht: (2024)
Sculpting Margin Penalty: Intra-Task Adapter Merging and Classifier Calibration for Few-Shot Class-Incremental Learning
von: Bai, Liang, et al.
Veröffentlicht: (2025)
von: Bai, Liang, et al.
Veröffentlicht: (2025)
Exploring Text-Guided Single Image Editing for Remote Sensing Images
von: Han, Fangzhou, et al.
Veröffentlicht: (2024)
von: Han, Fangzhou, et al.
Veröffentlicht: (2024)
CDMAD: Class-Distribution-Mismatch-Aware Debiasing for Class-Imbalanced Semi-Supervised Learning
von: Lee, Hyuck, et al.
Veröffentlicht: (2024)
von: Lee, Hyuck, et al.
Veröffentlicht: (2024)
Leveraging Textual Anatomical Knowledge for Class-Imbalanced Semi-Supervised Multi-Organ Segmentation
von: Gu, Yuliang, et al.
Veröffentlicht: (2025)
von: Gu, Yuliang, et al.
Veröffentlicht: (2025)
CPAM: Context-Preserving Adaptive Manipulation for Zero-Shot Real Image Editing
von: Vo, Dinh-Khoi, et al.
Veröffentlicht: (2025)
von: Vo, Dinh-Khoi, et al.
Veröffentlicht: (2025)
v1: Learning to Point Visual Tokens for Multimodal Grounded Reasoning
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
von: Chung, Jiwan, et al.
Veröffentlicht: (2025)
Accurate Explanation Model for Image Classifiers using Class Association Embedding
von: Xie, Ruitao, et al.
Veröffentlicht: (2024)
von: Xie, Ruitao, et al.
Veröffentlicht: (2024)
TCFG: Tangential Damping Classifier-free Guidance
von: Kwon, Mingi, et al.
Veröffentlicht: (2025)
von: Kwon, Mingi, et al.
Veröffentlicht: (2025)
CADRef: Robust Out-of-Distribution Detection via Class-Aware Decoupled Relative Feature Leveraging
von: Ling, Zhiwei, et al.
Veröffentlicht: (2025)
von: Ling, Zhiwei, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
SECOND: Mitigating Perceptual Hallucination in Vision-Language Models via Selective and Contrastive Decoding
von: Park, Woohyeon, et al.
Veröffentlicht: (2025) -
MMPB: It's Time for Multi-Modal Personalization
von: Kim, Jaeik, et al.
Veröffentlicht: (2025) -
VIRST: Video-Instructed Reasoning Assistant for SpatioTemporal Segmentation
von: Hong, Jihwan, et al.
Veröffentlicht: (2026) -
3rd Place of MeViS-Audio Track of the 5th PVUW: VIRST-Audio
von: Hong, Jihwan, et al.
Veröffentlicht: (2026) -
MEDIC-AD: Towards Medical Vision-Language Model's Clinical Intelligence
von: Park, Woohyeon, et al.
Veröffentlicht: (2026)