Neural Slot Interpreters: Grounding Object Semantics in Emergent Slot Representations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dedhia, Bhishma, Jha, Niraj K. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Zero-TPrune: Zero-Shot Token Pruning through Leveraging of the Attention Graph in Pre-Trained Transformers
von: Wang, Hongjie, et al.
Veröffentlicht: (2023)
von: Wang, Hongjie, et al.
Veröffentlicht: (2023)
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
von: Grigore, Diana-Nicoleta, et al.
Veröffentlicht: (2025)
von: Grigore, Diana-Nicoleta, et al.
Veröffentlicht: (2025)
Object-Centric Learning with Slot Mixture Module
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023)
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
von: Chi, Donghwan, et al.
Veröffentlicht: (2025)
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
von: Liu, Yu, et al.
Veröffentlicht: (2024)
von: Liu, Yu, et al.
Veröffentlicht: (2024)
Bottom-up Domain-specific Superintelligence: A Reliable Knowledge Graph is What We Need
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2025)
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2025)
CGSA: Class-Guided Slot-Aware Adaptation for Source-Free Object Detection
von: Dai, Boyang, et al.
Veröffentlicht: (2026)
von: Dai, Boyang, et al.
Veröffentlicht: (2026)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
von: Lai, Yixuan, et al.
Veröffentlicht: (2026)
Masked Multi-Query Slot Attention for Unsupervised Object Discovery
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
von: Pramanik, Rishav, et al.
Veröffentlicht: (2024)
SlotPi: Physics-informed Object-centric Reasoning Models
von: Li, Jian, et al.
Veröffentlicht: (2025)
von: Li, Jian, et al.
Veröffentlicht: (2025)
TC-SSA: Token Compression via Semantic Slot Aggregation for Gigapixel Pathology Reasoning
von: Chen, Zhuo, et al.
Veröffentlicht: (2026)
von: Chen, Zhuo, et al.
Veröffentlicht: (2026)
Temporally Consistent Object-Centric Learning by Contrasting Slots
von: Manasyan, Anna, et al.
Veröffentlicht: (2024)
von: Manasyan, Anna, et al.
Veröffentlicht: (2024)
Divided Attention: Unsupervised Multi-Object Discovery with Contextually Separated Slots
von: Lao, Dong, et al.
Veröffentlicht: (2023)
von: Lao, Dong, et al.
Veröffentlicht: (2023)
Slot-VAE: Object-Centric Scene Generation with Slot Attention
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
von: Wang, Yanbo, et al.
Veröffentlicht: (2023)
When Slots Compete: Slot Merging in Object-Centric Learning
von: Chatzisavvas, Christos, et al.
Veröffentlicht: (2026)
von: Chatzisavvas, Christos, et al.
Veröffentlicht: (2026)
Anatomy-Slot: Unsupervised Anatomical Factorization for Homologous Bilateral Reasoning in Retinal Diagnosis
von: Ma, Yingzhe, et al.
Veröffentlicht: (2026)
von: Ma, Yingzhe, et al.
Veröffentlicht: (2026)
Transformers and Slot Encoding for Sample Efficient Physical World Modelling
von: Petri, Francesco, et al.
Veröffentlicht: (2024)
von: Petri, Francesco, et al.
Veröffentlicht: (2024)
Adaptive Slot Attention: Object Discovery with Dynamic Slot Number
von: Fan, Ke, et al.
Veröffentlicht: (2024)
von: Fan, Ke, et al.
Veröffentlicht: (2024)
SlotVTG: Object-Centric Adapter for Generalizable Video Temporal Grounding
von: Han, Jiwook, et al.
Veröffentlicht: (2026)
von: Han, Jiwook, et al.
Veröffentlicht: (2026)
Generating, Fast and Slow: Scalable Parallel Video Generation with Video Interface Networks
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2025)
von: Dedhia, Bhishma, et al.
Veröffentlicht: (2025)
Multi-View Slot Attention Using Paraphrased Texts for Face Anti-Spoofing
von: Yu, Jeongmin, et al.
Veröffentlicht: (2025)
von: Yu, Jeongmin, et al.
Veröffentlicht: (2025)
Learning Global Object-Centric Representations via Disentangled Slot Attention
von: Chen, Tonglin, et al.
Veröffentlicht: (2024)
von: Chen, Tonglin, et al.
Veröffentlicht: (2024)
Learning Object-Centric Representations Based on Slots in Real World Scenarios
von: Akan, Adil Kaan
Veröffentlicht: (2025)
von: Akan, Adil Kaan
Veröffentlicht: (2025)
MetaSlot: Break Through the Fixed Number of Slots in Object-Centric Learning
von: Liu, Hongjia, et al.
Veröffentlicht: (2025)
von: Liu, Hongjia, et al.
Veröffentlicht: (2025)
Slot-guided Volumetric Object Radiance Fields
von: Qi, Di, et al.
Veröffentlicht: (2024)
von: Qi, Di, et al.
Veröffentlicht: (2024)
Slot-VLM: SlowFast Slots for Video-Language Modeling
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
von: Xu, Jiaqi, et al.
Veröffentlicht: (2024)
FORLA: Federated Object-centric Representation Learning with Slot Attention
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
SlotVLA: Towards Modeling of Object-Relation Representations in Robotic Manipulation
von: Hanyu, Taisei, et al.
Veröffentlicht: (2025)
von: Hanyu, Taisei, et al.
Veröffentlicht: (2025)
Guided Slot Attention for Unsupervised Video Object Segmentation
von: Lee, Minhyeok, et al.
Veröffentlicht: (2023)
von: Lee, Minhyeok, et al.
Veröffentlicht: (2023)
Predicting Video Slot Attention Queries from Random Slot-Feature Pairs
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
Future Slot Prediction for Unsupervised Object Discovery in Surgical Video
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
von: Liao, Guiqiu, et al.
Veröffentlicht: (2025)
LinMU: Multimodal Understanding Made Linear
von: Wang, Hongjie, et al.
Veröffentlicht: (2026)
von: Wang, Hongjie, et al.
Veröffentlicht: (2026)
OpenSlot: Mixed Open-Set Recognition with Object-Centric Learning
von: Yin, Xu, et al.
Veröffentlicht: (2024)
von: Yin, Xu, et al.
Veröffentlicht: (2024)
Smoothing Slot Attention Iterations and Recurrences
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
von: Zhao, Rongzhen, et al.
Veröffentlicht: (2025)
SlotMemory: Object-Centric KV Memory for Streaming Long-Video Generation
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
von: Dou, Weijia, et al.
Veröffentlicht: (2026)
GLASS: Guided Latent Slot Diffusion for Object-Centric Learning
von: Singh, Krishnakant, et al.
Veröffentlicht: (2024)
von: Singh, Krishnakant, et al.
Veröffentlicht: (2024)
Augmented Commonsense Knowledge for Remote Object Grounding
von: Mohammadi, Bahram, et al.
Veröffentlicht: (2024)
von: Mohammadi, Bahram, et al.
Veröffentlicht: (2024)
Grounding Continuous Representations in Geometry: Equivariant Neural Fields
von: Wessels, David R, et al.
Veröffentlicht: (2024)
von: Wessels, David R, et al.
Veröffentlicht: (2024)
Emergent Visual Grounding in Large Multimodal Models Without Grounding Supervision
von: Cao, Shengcao, et al.
Veröffentlicht: (2024)
von: Cao, Shengcao, et al.
Veröffentlicht: (2024)
GroundCount: Grounding Vision-Language Models with Object Detection for Mitigating Counting Hallucinations
von: Chen, Boyuan, et al.
Veröffentlicht: (2026)
von: Chen, Boyuan, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Zero-TPrune: Zero-Shot Token Pruning through Leveraging of the Attention Graph in Pre-Trained Transformers
von: Wang, Hongjie, et al.
Veröffentlicht: (2023) -
SlotMatch: Distilling Object-Centric Representations for Unsupervised Video Segmentation
von: Grigore, Diana-Nicoleta, et al.
Veröffentlicht: (2025) -
Object-Centric Learning with Slot Mixture Module
von: Kirilenko, Daniil, et al.
Veröffentlicht: (2023) -
Slot-MLLM: Object-Centric Visual Tokenization for Multimodal LLM
von: Chi, Donghwan, et al.
Veröffentlicht: (2025) -
SlotLifter: Slot-guided Feature Lifting for Learning Object-centric Radiance Fields
von: Liu, Yu, et al.
Veröffentlicht: (2024)