ARNet: Self-Supervised FG-SBIR with Unified Sample Feature Alignment and Multi-Scale Token Recycling
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Jiang, Jianan, Tang, Hao, Jiang, Zhilin, Yu, Weiren, Wu, Di |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
HAIFIT: Human-to-AI Fashion Image Translation
von: Jiang, Jianan, et al.
Veröffentlicht: (2024)
von: Jiang, Jianan, et al.
Veröffentlicht: (2024)
Token Transforming: A Unified and Training-Free Token Compression Framework for Vision Transformer Acceleration
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025)
Learning Accurate Segmentation Purely from Self-Supervision
von: You, Zuyao, et al.
Veröffentlicht: (2026)
von: You, Zuyao, et al.
Veröffentlicht: (2026)
FG-CLIP: Fine-Grained Visual and Textual Alignment
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
von: Xie, Chunyu, et al.
Veröffentlicht: (2025)
MIRROR: Multi-Modal Pathological Self-Supervised Representation Learning via Modality Alignment and Retention
von: Wang, Tianyi, et al.
Veröffentlicht: (2025)
von: Wang, Tianyi, et al.
Veröffentlicht: (2025)
Multi-Prompt Alignment for Multi-Source Unsupervised Domain Adaptation
von: Chen, Haoran, et al.
Veröffentlicht: (2022)
von: Chen, Haoran, et al.
Veröffentlicht: (2022)
MSNeRV: Neural Video Representation with Multi-Scale Feature Fusion
von: Zhu, Jun, et al.
Veröffentlicht: (2025)
von: Zhu, Jun, et al.
Veröffentlicht: (2025)
SynerMedGen: Synergizing Medical Multimodal Understanding with Generation via Task Alignment
von: Zhao, Weiren, et al.
Veröffentlicht: (2026)
von: Zhao, Weiren, et al.
Veröffentlicht: (2026)
Enhanced Textual Feature Extraction for Visual Question Answering: A Simple Convolutional Approach
von: Zhang, Zhilin, et al.
Veröffentlicht: (2024)
von: Zhang, Zhilin, et al.
Veröffentlicht: (2024)
MSSSeg: Learning Multi-Scale Structural Complexity for Self-Supervised Segmentation
von: Li, Haotang, et al.
Veröffentlicht: (2025)
von: Li, Haotang, et al.
Veröffentlicht: (2025)
Self-Supervised Discriminative Feature Learning for Deep Multi-View Clustering
von: Xu, Jie, et al.
Veröffentlicht: (2021)
von: Xu, Jie, et al.
Veröffentlicht: (2021)
Spatiotemporal Blind-Spot Network with Calibrated Flow Alignment for Self-Supervised Video Denoising
von: Chen, Zikang, et al.
Veröffentlicht: (2024)
von: Chen, Zikang, et al.
Veröffentlicht: (2024)
Learning Contrastive Self-Distillation for Ultra-Fine-Grained Visual Categorization Targeting Limited Samples
von: Fang, Ziye, et al.
Veröffentlicht: (2023)
von: Fang, Ziye, et al.
Veröffentlicht: (2023)
OmniTokenizer: A Joint Image-Video Tokenizer for Visual Generation
von: Wang, Junke, et al.
Veröffentlicht: (2024)
von: Wang, Junke, et al.
Veröffentlicht: (2024)
Self-Supervised Implicit Attention Priors for Point Cloud Reconstruction
von: Fogarty, Kyle, et al.
Veröffentlicht: (2025)
von: Fogarty, Kyle, et al.
Veröffentlicht: (2025)
Advancing Comprehensive Aesthetic Insight with Multi-Scale Text-Guided Self-Supervised Learning
von: Liu, Yuti, et al.
Veröffentlicht: (2024)
von: Liu, Yuti, et al.
Veröffentlicht: (2024)
Multi-Prompt Progressive Alignment for Multi-Source Unsupervised Domain Adaptation
von: Chen, Haoran, et al.
Veröffentlicht: (2025)
von: Chen, Haoran, et al.
Veröffentlicht: (2025)
FG$^2$: Fine-Grained Cross-View Localization by Fine-Grained Feature Matching
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
von: Xia, Zimin, et al.
Veröffentlicht: (2025)
Infra-YOLO: Efficient Neural Network Structure with Model Compression for Real-Time Infrared Small Object Detection
von: Chen, Zhonglin, et al.
Veröffentlicht: (2024)
von: Chen, Zhonglin, et al.
Veröffentlicht: (2024)
Labeled-to-Unlabeled Distribution Alignment for Partially-Supervised Multi-Organ Medical Image Segmentation
von: Jiang, Xixi, et al.
Veröffentlicht: (2024)
von: Jiang, Xixi, et al.
Veröffentlicht: (2024)
Attentive Convolution: Unifying the Expressivity of Self-Attention with Convolutional Efficiency
von: Yu, Hao, et al.
Veröffentlicht: (2025)
von: Yu, Hao, et al.
Veröffentlicht: (2025)
Self-Supervised Multi-Scale Network for Blind Image Deblurring via Alternating Optimization
von: Guo, Lening, et al.
Veröffentlicht: (2024)
von: Guo, Lening, et al.
Veröffentlicht: (2024)
On the Theory of Conditional Feature Alignment for Unsupervised Domain-Adaptive Counting
von: Liang, Zhuonan, et al.
Veröffentlicht: (2025)
von: Liang, Zhuonan, et al.
Veröffentlicht: (2025)
MetaSeg: Content-Aware Meta-Net for Omni-Supervised Semantic Segmentation
von: Jiang, Shenwang, et al.
Veröffentlicht: (2024)
von: Jiang, Shenwang, et al.
Veröffentlicht: (2024)
Supervised Classification Heads as Semantic Prototypes: Unlocking Vision-Language Alignment via Weight Recycling
von: Méndez, David, et al.
Veröffentlicht: (2026)
von: Méndez, David, et al.
Veröffentlicht: (2026)
Exploiting Self-Supervised Constraints in Image Super-Resolution
von: Wu, Gang, et al.
Veröffentlicht: (2024)
von: Wu, Gang, et al.
Veröffentlicht: (2024)
Improving Video Diffusion Transformer Training by Multi-Feature Fusion and Alignment from Self-Supervised Vision Encoders
von: Lee, Dohun, et al.
Veröffentlicht: (2025)
von: Lee, Dohun, et al.
Veröffentlicht: (2025)
Towards Sequence Modeling Alignment between Tokenizer and Autoregressive Model
von: Wu, Pingyu, et al.
Veröffentlicht: (2025)
von: Wu, Pingyu, et al.
Veröffentlicht: (2025)
FocusTrack: A Self-Adaptive Local Sampling Algorithm for Efficient Anti-UAV Tracking
von: Wang, Ying, et al.
Veröffentlicht: (2025)
von: Wang, Ying, et al.
Veröffentlicht: (2025)
SEA: Supervised Embedding Alignment for Token-Level Visual-Textual Integration in MLLMs
von: Yin, Yuanyang, et al.
Veröffentlicht: (2024)
von: Yin, Yuanyang, et al.
Veröffentlicht: (2024)
Back to Fundamentals: Low-Level Visual Features Guided Progressive Token Pruning
von: Ouyang, Yuanbing, et al.
Veröffentlicht: (2025)
von: Ouyang, Yuanbing, et al.
Veröffentlicht: (2025)
UniTok: A Unified Tokenizer for Visual Generation and Understanding
von: Ma, Chuofan, et al.
Veröffentlicht: (2025)
von: Ma, Chuofan, et al.
Veröffentlicht: (2025)
Bridging Supervision Gaps: A Unified Framework for Remote Sensing Change Detection
von: Jiang, Kaixuan, et al.
Veröffentlicht: (2026)
von: Jiang, Kaixuan, et al.
Veröffentlicht: (2026)
TerraGen: A Unified Multi-Task Layout Generation Framework for Remote Sensing Data Augmentation
von: Tang, Datao, et al.
Veröffentlicht: (2025)
von: Tang, Datao, et al.
Veröffentlicht: (2025)
TokenFlow: Unified Image Tokenizer for Multimodal Understanding and Generation
von: Qu, Liao, et al.
Veröffentlicht: (2024)
von: Qu, Liao, et al.
Veröffentlicht: (2024)
Attention Deep Model with Multi-Scale Deep Supervision for Person Re-Identification
von: Wu, Di, et al.
Veröffentlicht: (2019)
von: Wu, Di, et al.
Veröffentlicht: (2019)
Joint Self-Supervised Video Alignment and Action Segmentation
von: Ali, Ali Shah, et al.
Veröffentlicht: (2025)
von: Ali, Ali Shah, et al.
Veröffentlicht: (2025)
Backdooring Self-Supervised Contrastive Learning by Noisy Alignment
von: Chen, Tuo, et al.
Veröffentlicht: (2025)
von: Chen, Tuo, et al.
Veröffentlicht: (2025)
Self-Supervised Alignment Learning for Medical Image Segmentation
von: Li, Haofeng, et al.
Veröffentlicht: (2024)
von: Li, Haofeng, et al.
Veröffentlicht: (2024)
Le MuMo JEPA: Multi-Modal Self-Supervised Representation Learning with Learnable Fusion Tokens
von: Cornelissen, Ciem, et al.
Veröffentlicht: (2026)
von: Cornelissen, Ciem, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
HAIFIT: Human-to-AI Fashion Image Translation
von: Jiang, Jianan, et al.
Veröffentlicht: (2024) -
Token Transforming: A Unified and Training-Free Token Compression Framework for Vision Transformer Acceleration
von: Zeng, Fanhu, et al.
Veröffentlicht: (2025) -
Learning Accurate Segmentation Purely from Self-Supervision
von: You, Zuyao, et al.
Veröffentlicht: (2026) -
FG-CLIP: Fine-Grained Visual and Textual Alignment
von: Xie, Chunyu, et al.
Veröffentlicht: (2025) -
MIRROR: Multi-Modal Pathological Self-Supervised Representation Learning via Modality Alignment and Retention
von: Wang, Tianyi, et al.
Veröffentlicht: (2025)