Bootstrapping Top-down Information for Self-modulating Slot Attention
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kim, Dongwon, Kim, Seoyeon, Kwak, Suha |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PLOT: Text-based Person Search with Part Slot Attention for Corresponding Part Discovery
von: Park, Jicheol, et al.
Veröffentlicht: (2024)
von: Park, Jicheol, et al.
Veröffentlicht: (2024)
Improving Text-based Person Search via Part-level Cross-modal Correspondence
von: Park, Jicheol, et al.
Veröffentlicht: (2024)
von: Park, Jicheol, et al.
Veröffentlicht: (2024)
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023)
Efficient and Versatile Robust Fine-Tuning of Zero-shot Models
von: Kim, Sungyeon, et al.
Veröffentlicht: (2024)
von: Kim, Sungyeon, et al.
Veröffentlicht: (2024)
Learning Unified Distance Metric Across Diverse Data Distributions with Parameter-Efficient Transfer Learning
von: Kim, Sungyeon, et al.
Veröffentlicht: (2023)
von: Kim, Sungyeon, et al.
Veröffentlicht: (2023)
Improving Sound Source Localization with Joint Slot Attention on Image and Audio
von: Kim, Inho, et al.
Veröffentlicht: (2025)
von: Kim, Inho, et al.
Veröffentlicht: (2025)
Improving Robustness to Multiple Spurious Correlations by Multi-Objective Optimization
von: Kim, Nayeong, et al.
Veröffentlicht: (2024)
von: Kim, Nayeong, et al.
Veröffentlicht: (2024)
Enhancing Cost Efficiency in Active Learning with Candidate Set Query
von: Gwon, Yeho, et al.
Veröffentlicht: (2025)
von: Gwon, Yeho, et al.
Veröffentlicht: (2025)
ActFusion: a Unified Diffusion Model for Action Segmentation and Anticipation
von: Gong, Dayoung, et al.
Veröffentlicht: (2024)
von: Gong, Dayoung, et al.
Veröffentlicht: (2024)
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
von: Fan, Ke, et al.
Veröffentlicht: (2024)
von: Fan, Ke, et al.
Veröffentlicht: (2024)
Concept Unlearning via Cross-Attention Activation Projection for Diffusion Models
von: Moon, Saemi, et al.
Veröffentlicht: (2026)
von: Moon, Saemi, et al.
Veröffentlicht: (2026)
Tempered Self-Similarity Alignment for Physically Plausible Video Generation
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
von: Kim, Manjin, et al.
Veröffentlicht: (2026)
SYNAuG: Exploiting Synthetic Data for Data Imbalance Problems
von: Ye-Bin, Moon, et al.
Veröffentlicht: (2023)
von: Ye-Bin, Moon, et al.
Veröffentlicht: (2023)
Structured State-Space Regularization for Generation-Friendly Image Tokenization
von: Lee, Jinsung, et al.
Veröffentlicht: (2026)
von: Lee, Jinsung, et al.
Veröffentlicht: (2026)
Learning Audio-guided Video Representation with Gated Attention for Video-Text Retrieval
von: Jeong, Boseung, et al.
Veröffentlicht: (2025)
von: Jeong, Boseung, et al.
Veröffentlicht: (2025)
FORLA: Federated Object-centric Representation Learning with Slot Attention
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
Planning in 8 Tokens: A Compact Discrete Tokenizer for Latent World Model
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
von: Kim, Dongwon, et al.
Veröffentlicht: (2026)
Federated Learning for Face Recognition via Intra-subject Self-supervised Learning
von: Kim, Hansol, et al.
Veröffentlicht: (2024)
von: Kim, Hansol, et al.
Veröffentlicht: (2024)
Part-Aware Bottom-Up Group Reasoning for Fine-Grained Social Interaction Detection
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
von: Kim, Dongkeun, et al.
Veröffentlicht: (2025)
Text Embedding is Not All You Need: Attention Control for Text-to-Image Semantic Alignment with Text Self-Attention Maps
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
von: Kim, Jeeyung, et al.
Veröffentlicht: (2024)
GENIUS: A Generative Framework for Universal Multimodal Search
von: Kim, Sungyeon, et al.
Veröffentlicht: (2025)
von: Kim, Sungyeon, et al.
Veröffentlicht: (2025)
SpargeAttention2: Trainable Sparse Attention via Hybrid Top-k+Top-p Masking and Distillation Fine-Tuning
von: Zhang, Jintao, et al.
Veröffentlicht: (2026)
von: Zhang, Jintao, et al.
Veröffentlicht: (2026)
MetaSlot: Break Through the Fixed Number of Slots in Object-Centric Learning
von: Liu, Hongjia, et al.
Veröffentlicht: (2025)
von: Liu, Hongjia, et al.
Veröffentlicht: (2025)
Self-Bootstrapping for Versatile Test-Time Adaptation
von: Niu, Shuaicheng, et al.
Veröffentlicht: (2025)
von: Niu, Shuaicheng, et al.
Veröffentlicht: (2025)
FREST: Feature RESToration for Semantic Segmentation under Multiple Adverse Conditions
von: Lee, Sohyun, et al.
Veröffentlicht: (2024)
von: Lee, Sohyun, et al.
Veröffentlicht: (2024)
Mixture-of-Top-k Attention: Efficient Attention via Scalable Fast Weights
von: Wen, Qishuai, et al.
Veröffentlicht: (2026)
von: Wen, Qishuai, et al.
Veröffentlicht: (2026)
Slot Structured World Models
von: Collu, Jonathan, et al.
Veröffentlicht: (2024)
von: Collu, Jonathan, et al.
Veröffentlicht: (2024)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
Democratizing Text-to-Image Masked Generative Models with Compact Text-Aware One-Dimensional Tokens
von: Kim, Dongwon, et al.
Veröffentlicht: (2025)
von: Kim, Dongwon, et al.
Veröffentlicht: (2025)
Divided Attention: Unsupervised Multi-Object Discovery with Contextually Separated Slots
von: Lao, Dong, et al.
Veröffentlicht: (2023)
von: Lao, Dong, et al.
Veröffentlicht: (2023)
Regularizing Attention Scores with Bootstrapping
von: Chung, Neo Christopher, et al.
Veröffentlicht: (2026)
von: Chung, Neo Christopher, et al.
Veröffentlicht: (2026)
Self-Rectifying Diffusion Sampling with Perturbed-Attention Guidance
von: Ahn, Donghoon, et al.
Veröffentlicht: (2024)
von: Ahn, Donghoon, et al.
Veröffentlicht: (2024)
Stabilizing Consistency Training: A Flow Map Analysis and Self-Distillation
von: Kim, Youngjoong, et al.
Veröffentlicht: (2026)
von: Kim, Youngjoong, et al.
Veröffentlicht: (2026)
Online Temporal Action Localization with Memory-Augmented Transformer
von: Song, Youngkil, et al.
Veröffentlicht: (2024)
von: Song, Youngkil, et al.
Veröffentlicht: (2024)
Active Label Correction for Semantic Segmentation with Foundation Models
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
von: Kim, Hoyoung, et al.
Veröffentlicht: (2024)
Towards More Practical Group Activity Detection: A New Benchmark and Model
von: Kim, Dongkeun, et al.
Veröffentlicht: (2023)
von: Kim, Dongkeun, et al.
Veröffentlicht: (2023)
Slot Abstractors: Toward Scalable Abstract Visual Reasoning
von: Mondal, Shanka Subhra, et al.
Veröffentlicht: (2024)
von: Mondal, Shanka Subhra, et al.
Veröffentlicht: (2024)
Open-world Instance Segmentation: Top-down Learning with Bottom-up Supervision
von: Kalluri, Tarun, et al.
Veröffentlicht: (2023)
von: Kalluri, Tarun, et al.
Veröffentlicht: (2023)
Coreset Selection for Object Detection
von: Lee, Hojun, et al.
Veröffentlicht: (2024)
von: Lee, Hojun, et al.
Veröffentlicht: (2024)
CrIBo: Self-Supervised Learning via Cross-Image Object-Level Bootstrapping
von: Lebailly, Tim, et al.
Veröffentlicht: (2023)
von: Lebailly, Tim, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
PLOT: Text-based Person Search with Part Slot Attention for Corresponding Part Discovery
von: Park, Jicheol, et al.
Veröffentlicht: (2024) -
Improving Text-based Person Search via Part-level Cross-modal Correspondence
von: Park, Jicheol, et al.
Veröffentlicht: (2024) -
Extending CLIP's Image-Text Alignment to Referring Image Segmentation
von: Kim, Seoyeon, et al.
Veröffentlicht: (2023) -
Efficient and Versatile Robust Fine-Tuning of Zero-shot Models
von: Kim, Sungyeon, et al.
Veröffentlicht: (2024) -
Learning Unified Distance Metric Across Diverse Data Distributions with Parameter-Efficient Transfer Learning
von: Kim, Sungyeon, et al.
Veröffentlicht: (2023)