Set2Seq Transformer: Temporal and Position-Aware Set Representations for Sequential Multiple-Instance Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Efthymiou, Athanasios, Rudinac, Stevan, Kackovic, Monika, Wijnberg, Nachoem, Worring, Marcel |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Graph Neural Networks for Knowledge Enhanced Visual Representation of Paintings
by: Efthymiou, Athanasios, et al.
Published: (2021)
by: Efthymiou, Athanasios, et al.
Published: (2021)
VL-KGE: Vision-Language Models Meet Knowledge Graph Embeddings
by: Efthymiou, Athanasios, et al.
Published: (2026)
by: Efthymiou, Athanasios, et al.
Published: (2026)
ArtRAG: Retrieval-Augmented Generation with Structured Context for Visual Art Understanding
by: Wang, Shuai, et al.
Published: (2025)
by: Wang, Shuai, et al.
Published: (2025)
Ada-HGNN: Adaptive Sampling for Scalable Hypergraph Neural Networks
by: Wang, Shuai, et al.
Published: (2024)
by: Wang, Shuai, et al.
Published: (2024)
A-MAR: Agent-based Multimodal Art Retrieval for Fine-Grained Artwork Understanding
by: Wang, Shuai, et al.
Published: (2026)
by: Wang, Shuai, et al.
Published: (2026)
SetFlow: Generating Structured Sets of Representations for Multiple Instance Learning
by: Jovišić, Nikola, et al.
Published: (2026)
by: Jovišić, Nikola, et al.
Published: (2026)
Looking Beyond the Obvious: A Survey on Abstract Concept Recognition for Video Understanding
by: Mago, Gowreesh, et al.
Published: (2025)
by: Mago, Gowreesh, et al.
Published: (2025)
SeqPE: Transformer with Sequential Position Encoding
by: Li, Huayang, et al.
Published: (2025)
by: Li, Huayang, et al.
Published: (2025)
Enhancing Interactive Image Retrieval With Query Rewriting Using Large Language Models and Vision Language Models
by: Zhu, Hongyi, et al.
Published: (2024)
by: Zhu, Hongyi, et al.
Published: (2024)
Seq2Time: Sequential Knowledge Transfer for Video LLM Temporal Grounding
by: Deng, Andong, et al.
Published: (2024)
by: Deng, Andong, et al.
Published: (2024)
Image2Text2Image: A Novel Framework for Label-Free Evaluation of Image-to-Text Generation with Text-to-Image Diffusion Models
by: Huang, Jia-Hong, et al.
Published: (2024)
by: Huang, Jia-Hong, et al.
Published: (2024)
CAPRMIL: Context-Aware Patch Representations for Multiple Instance Learning
by: Lolos, Andreas, et al.
Published: (2025)
by: Lolos, Andreas, et al.
Published: (2025)
Event Voxel Set Transformer for Spatiotemporal Representation Learning on Event Streams
by: Xie, Bochen, et al.
Published: (2023)
by: Xie, Bochen, et al.
Published: (2023)
A Novel Evaluation Framework for Image2Text Generation
by: Huang, Jia-Hong, et al.
Published: (2024)
by: Huang, Jia-Hong, et al.
Published: (2024)
Priority-Aware Clinical Pathology Hierarchy Training for Multiple Instance Learning
by: Hong, Sungrae, et al.
Published: (2025)
by: Hong, Sungrae, et al.
Published: (2025)
Conformal Prediction Sets for Instance Segmentation
by: Lu, Kerri, et al.
Published: (2026)
by: Lu, Kerri, et al.
Published: (2026)
Fourier Transform Multiple Instance Learning for Whole Slide Image Classification
by: Bilic, Anthony, et al.
Published: (2025)
by: Bilic, Anthony, et al.
Published: (2025)
Temporal2Seq: A Unified Framework for Temporal Video Understanding Tasks
by: Yang, Min, et al.
Published: (2024)
by: Yang, Min, et al.
Published: (2024)
Forest2Seq: Revitalizing Order Prior for Sequential Indoor Scene Synthesis
by: Sun, Qi, et al.
Published: (2024)
by: Sun, Qi, et al.
Published: (2024)
Open-Set Vein Biometric Recognition with Deep Metric Learning
by: Pilarek, Paweł, et al.
Published: (2026)
by: Pilarek, Paweł, et al.
Published: (2026)
InstanceRSR: Real-World Super-Resolution via Instance-Aware Representation Alignment
by: Guo, Zixin, et al.
Published: (2026)
by: Guo, Zixin, et al.
Published: (2026)
Layout-Aware Representation Learning for Open-Set ID Fraud Discovery
by: Li, Jinxing, et al.
Published: (2026)
by: Li, Jinxing, et al.
Published: (2026)
SeqCSIST: Sequential Closely-Spaced Infrared Small Target Unmixing
by: Zhai, Ximeng, et al.
Published: (2025)
by: Zhai, Ximeng, et al.
Published: (2025)
Instance-free Text to Point Cloud Localization with Relative Position Awareness
by: Wang, Lichao, et al.
Published: (2024)
by: Wang, Lichao, et al.
Published: (2024)
A Multimodal Seq2Seq Transformer for Predicting Brain Responses to Naturalistic Stimuli
by: He, Qianyi, et al.
Published: (2025)
by: He, Qianyi, et al.
Published: (2025)
Instance-Level Generation for Representation Learning
by: Wu, Yankun, et al.
Published: (2025)
by: Wu, Yankun, et al.
Published: (2025)
Exploring Diverse Representations for Open Set Recognition
by: Wang, Yu, et al.
Published: (2024)
by: Wang, Yu, et al.
Published: (2024)
MicroMIL: Graph-Based Multiple Instance Learning for Context-Aware Diagnosis with Microscopic Images
by: Kim, Jongwoo, et al.
Published: (2024)
by: Kim, Jongwoo, et al.
Published: (2024)
MiCo: Multiple Instance Learning with Context-Aware Clustering for Whole Slide Image Analysis
by: Li, Junjian, et al.
Published: (2025)
by: Li, Junjian, et al.
Published: (2025)
CAMIL: Context-Aware Multiple Instance Learning for Cancer Detection and Subtyping in Whole Slide Images
by: Fourkioti, Olga, et al.
Published: (2023)
by: Fourkioti, Olga, et al.
Published: (2023)
Semantic Positive Pairs for Enhancing Visual Representation Learning of Instance Discrimination Methods
by: Alkhalefi, Mohammad, et al.
Published: (2023)
by: Alkhalefi, Mohammad, et al.
Published: (2023)
Do Multiple Instance Learning Models Transfer?
by: Shao, Daniel, et al.
Published: (2025)
by: Shao, Daniel, et al.
Published: (2025)
Instance Segmentation for Point Sets
by: Talwar, Abhimanyu, et al.
Published: (2025)
by: Talwar, Abhimanyu, et al.
Published: (2025)
Video Set Distillation: Information Diversification and Temporal Densification
by: Zhao, Yinjie, et al.
Published: (2024)
by: Zhao, Yinjie, et al.
Published: (2024)
Positive Semi-definite Latent Factor Grouping-Boosted Cluster-reasoning Instance Disentangled Learning for WSI Representation
by: Li, Chentao, et al.
Published: (2025)
by: Li, Chentao, et al.
Published: (2025)
A Spatially-Aware Multiple Instance Learning Framework for Digital Pathology
by: Keshvarikhojasteh, Hassan, et al.
Published: (2025)
by: Keshvarikhojasteh, Hassan, et al.
Published: (2025)
SGPMIL: Sparse Gaussian Process Multiple Instance Learning
by: Lolos, Andreas, et al.
Published: (2025)
by: Lolos, Andreas, et al.
Published: (2025)
Of Great Importance
by: Wijnberg, Nachoem M.
Published: (2019)
by: Wijnberg, Nachoem M.
Published: (2019)
The Jews
by: Wijnberg, Nachoem M.
Published: (2019)
by: Wijnberg, Nachoem M.
Published: (2019)
Instance-Aware Group Quantization for Vision Transformers
by: Moon, Jaehyeon, et al.
Published: (2024)
by: Moon, Jaehyeon, et al.
Published: (2024)
Similar Items
-
Graph Neural Networks for Knowledge Enhanced Visual Representation of Paintings
by: Efthymiou, Athanasios, et al.
Published: (2021) -
VL-KGE: Vision-Language Models Meet Knowledge Graph Embeddings
by: Efthymiou, Athanasios, et al.
Published: (2026) -
ArtRAG: Retrieval-Augmented Generation with Structured Context for Visual Art Understanding
by: Wang, Shuai, et al.
Published: (2025) -
Ada-HGNN: Adaptive Sampling for Scalable Hypergraph Neural Networks
by: Wang, Shuai, et al.
Published: (2024) -
A-MAR: Agent-based Multimodal Art Retrieval for Fine-Grained Artwork Understanding
by: Wang, Shuai, et al.
Published: (2026)