Super-class guided Transformer for Zero-Shot Attribute Classification
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Sehyung, Yang, Chanhyeong, Park, Jihwan, Song, Taehoon, Kim, Hyunwoo J. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Visual Diversity and Region-aware Prompt Learning for Zero-shot HOI Detection
von: Yang, Chanhyeong, et al.
Veröffentlicht: (2025)
von: Yang, Chanhyeong, et al.
Veröffentlicht: (2025)
RegFormer: Transferable Relational Grounding for Efficient Weakly-Supervised Human-Object Interaction Detection
von: Park, Jihwan, et al.
Veröffentlicht: (2026)
von: Park, Jihwan, et al.
Veröffentlicht: (2026)
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
von: Kim, Jongha, et al.
Veröffentlicht: (2024)
von: Kim, Jongha, et al.
Veröffentlicht: (2024)
Retrieval-Augmented Open-Vocabulary Object Detection
von: Kim, Jooyeon, et al.
Veröffentlicht: (2024)
von: Kim, Jooyeon, et al.
Veröffentlicht: (2024)
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
von: Park, Jihwan, et al.
Veröffentlicht: (2025)
von: Park, Jihwan, et al.
Veröffentlicht: (2025)
Blockwise Flow Matching: Improving Flow Matching Models For Efficient High-Quality Generation
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
von: Park, Dogyun, et al.
Veröffentlicht: (2025)
DocPrune:Efficient Document Question Answering via Background, Question, and Comprehension-aware Token Pruning
von: Choi, Joonmyung, et al.
Veröffentlicht: (2026)
von: Choi, Joonmyung, et al.
Veröffentlicht: (2026)
Retrieve What's Missing: Coverage-Maximizing Retrieval for Consistent Long Video Generation
von: Joo, Minseok, et al.
Veröffentlicht: (2026)
von: Joo, Minseok, et al.
Veröffentlicht: (2026)
Robust Multimodal 3D Object Detection via Modality-Agnostic Decoding and Proximity-based Modality Ensemble
von: Cha, Juhan, et al.
Veröffentlicht: (2024)
von: Cha, Juhan, et al.
Veröffentlicht: (2024)
Constant Acceleration Flow
von: Park, Dogyun, et al.
Veröffentlicht: (2024)
von: Park, Dogyun, et al.
Veröffentlicht: (2024)
MADS: Multi-Attribute Document Supervision for Zero-Shot Image Classification
von: Qu, Xiangyan, et al.
Veröffentlicht: (2025)
von: Qu, Xiangyan, et al.
Veröffentlicht: (2025)
Model Synthesis for Zero-Shot Model Attribution
von: Yang, Tianyun, et al.
Veröffentlicht: (2023)
von: Yang, Tianyun, et al.
Veröffentlicht: (2023)
Future-Proof Yourself: An AI Era Survival Guide
von: Kim, Taehoon
Veröffentlicht: (2025)
von: Kim, Taehoon
Veröffentlicht: (2025)
Cross-Class Feature Augmentation for Class Incremental Learning
von: Kim, Taehoon, et al.
Veröffentlicht: (2023)
von: Kim, Taehoon, et al.
Veröffentlicht: (2023)
Zero-Shot Scene Change Detection
von: Cho, Kyusik, et al.
Veröffentlicht: (2024)
von: Cho, Kyusik, et al.
Veröffentlicht: (2024)
Long-term Pre-training for Temporal Action Detection with Transformers
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
von: Kim, Jihwan, et al.
Veröffentlicht: (2024)
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024)
von: Song, Yeji, et al.
Veröffentlicht: (2024)
Prompt Learning via Meta-Regularization
von: Park, Jinyoung, et al.
Veröffentlicht: (2024)
von: Park, Jinyoung, et al.
Veröffentlicht: (2024)
DropGaussian: Structural Regularization for Sparse-view Gaussian Splatting
von: Park, Hyunwoo, et al.
Veröffentlicht: (2025)
von: Park, Hyunwoo, et al.
Veröffentlicht: (2025)
Multi-criteria Token Fusion with One-step-ahead Attention for Efficient Vision Transformers
von: Lee, Sanghyeok, et al.
Veröffentlicht: (2024)
von: Lee, Sanghyeok, et al.
Veröffentlicht: (2024)
Follow the Saliency: Supervised Saliency for Retrieval-augmented Dense Video Captioning
von: Choi, Seung hee, et al.
Veröffentlicht: (2026)
von: Choi, Seung hee, et al.
Veröffentlicht: (2026)
MAVIS: A Benchmark for Multimodal Source Attribution in Long-form Visual Question Answering
von: Song, Seokwon, et al.
Veröffentlicht: (2025)
von: Song, Seokwon, et al.
Veröffentlicht: (2025)
Text-guided Zero-Shot Object Localization
von: Wang, Jingjing, et al.
Veröffentlicht: (2024)
von: Wang, Jingjing, et al.
Veröffentlicht: (2024)
Will It Zero-Shot?: Predicting Zero-Shot Classification Performance For Arbitrary Queries
von: Robbins, Kevin, et al.
Veröffentlicht: (2026)
von: Robbins, Kevin, et al.
Veröffentlicht: (2026)
MAC: A Benchmark for Multiple Attributes Compositional Zero-Shot Learning
von: Xu, Shuo, et al.
Veröffentlicht: (2024)
von: Xu, Shuo, et al.
Veröffentlicht: (2024)
CXR-LT 2026 Challenge: Projection-Aware Multi-Label and Zero-Shot Chest X-Ray Classification
von: Cho, Juno, et al.
Veröffentlicht: (2026)
von: Cho, Juno, et al.
Veröffentlicht: (2026)
When Model Knowledge meets Diffusion Model: Diffusion-assisted Data-free Image Synthesis with Alignment of Domain and Class
von: Kim, Yujin, et al.
Veröffentlicht: (2025)
von: Kim, Yujin, et al.
Veröffentlicht: (2025)
Mutually-Aware Feature Learning for Few-Shot Object Counting
von: Jeon, Yerim, et al.
Veröffentlicht: (2024)
von: Jeon, Yerim, et al.
Veröffentlicht: (2024)
Efficient multi-view training for 3D Gaussian Splatting
von: Choi, Minhyuk, et al.
Veröffentlicht: (2025)
von: Choi, Minhyuk, et al.
Veröffentlicht: (2025)
ViTA-PAR: Visual and Textual Attribute Alignment with Attribute Prompting for Pedestrian Attribute Recognition
von: Park, Minjeong, et al.
Veröffentlicht: (2025)
von: Park, Minjeong, et al.
Veröffentlicht: (2025)
Enhancing Spatio-Temporal Zero-shot Action Recognition with Language-driven Description Attributes
von: Kim, Yehna, et al.
Veröffentlicht: (2025)
von: Kim, Yehna, et al.
Veröffentlicht: (2025)
Fine-Grained Zero-Shot Learning with Attribute-Centric Representations
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
Attributed Synthetic Data Generation for Zero-shot Domain-specific Image Classification
von: Wang, Shijian, et al.
Veröffentlicht: (2025)
von: Wang, Shijian, et al.
Veröffentlicht: (2025)
Stochastic Conditional Diffusion Models for Robust Semantic Image Synthesis
von: Ko, Juyeon, et al.
Veröffentlicht: (2024)
von: Ko, Juyeon, et al.
Veröffentlicht: (2024)
Diffusion Prior-Based Amortized Variational Inference for Noisy Inverse Problems
von: Lee, Sojin, et al.
Veröffentlicht: (2024)
von: Lee, Sojin, et al.
Veröffentlicht: (2024)
Accelerating Image Super-Resolution Networks with Pixel-Level Classification
von: Jeong, Jinho, et al.
Veröffentlicht: (2024)
von: Jeong, Jinho, et al.
Veröffentlicht: (2024)
SAT: Selective Aggregation Transformer for Image Super-Resolution
von: Tran, Dinh Phu, et al.
Veröffentlicht: (2026)
von: Tran, Dinh Phu, et al.
Veröffentlicht: (2026)
DeepVideo-R1: Video Reinforcement Fine-Tuning via Difficulty-aware Regressive GRPO
von: Park, Jinyoung, et al.
Veröffentlicht: (2025)
von: Park, Jinyoung, et al.
Veröffentlicht: (2025)
Early Timestep Zero-Shot Candidate Selection for Instruction-Guided Image Editing
von: Kim, Joowon, et al.
Veröffentlicht: (2025)
von: Kim, Joowon, et al.
Veröffentlicht: (2025)
VidChain: Chain-of-Tasks with Metric-based Direct Preference Optimization for Dense Video Captioning
von: Lee, Ji Soo, et al.
Veröffentlicht: (2025)
von: Lee, Ji Soo, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Visual Diversity and Region-aware Prompt Learning for Zero-shot HOI Detection
von: Yang, Chanhyeong, et al.
Veröffentlicht: (2025) -
RegFormer: Transferable Relational Grounding for Efficient Weakly-Supervised Human-Object Interaction Detection
von: Park, Jihwan, et al.
Veröffentlicht: (2026) -
Groupwise Query Specialization and Quality-Aware Multi-Assignment for Transformer-based Visual Relationship Detection
von: Kim, Jongha, et al.
Veröffentlicht: (2024) -
Retrieval-Augmented Open-Vocabulary Object Detection
von: Kim, Jooyeon, et al.
Veröffentlicht: (2024) -
Transferable Model-agnostic Vision-Language Model Adaptation for Efficient Weak-to-Strong Generalization
von: Park, Jihwan, et al.
Veröffentlicht: (2025)