Investigating Permutation-Invariant Discrete Representation Learning for Spatially Aligned Images
Fuente:
arXiv
Saved in:
| Main Authors: | Stirling, Jamie S. J., Al-Moubayed, Noura, Shum, Hubert P. H. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Controllable Image Generation with Composed Parallel Token Prediction
by: Stirling, Jamie, et al.
Published: (2024)
by: Stirling, Jamie, et al.
Published: (2024)
Controllable Image Generation with Composed Parallel Token Prediction
by: Stirling, Jamie, et al.
Published: (2026)
by: Stirling, Jamie, et al.
Published: (2026)
Disentangling Racial Phenotypes: Fine-Grained Control of Race-related Facial Phenotype Characteristics
by: Yucer, Seyma, et al.
Published: (2024)
by: Yucer, Seyma, et al.
Published: (2024)
Video Prediction of Dynamic Physical Simulations With Pixel-Space Spatiotemporal Transformers
by: Slack, Dean L, et al.
Published: (2025)
by: Slack, Dean L, et al.
Published: (2025)
Everything is a Video: Unifying Modalities through Next-Frame Prediction
by: Hudson, G. Thomas, et al.
Published: (2024)
by: Hudson, G. Thomas, et al.
Published: (2024)
TraIL-Det: Transformation-Invariant Local Feature Networks for 3D LiDAR Object Detection with Unsupervised Pre-Training
by: Li, Li, et al.
Published: (2024)
by: Li, Li, et al.
Published: (2024)
Textual Localization: Decomposing Multi-concept Images for Subject-Driven Text-to-Image Generation
by: Shentu, Junjie, et al.
Published: (2024)
by: Shentu, Junjie, et al.
Published: (2024)
AttenCraft: Attention-guided Disentanglement of Multiple Concepts for Text-to-Image Customization
by: Shentu, Junjie, et al.
Published: (2024)
by: Shentu, Junjie, et al.
Published: (2024)
RAPiD-Seg: Range-Aware Pointwise Distance Distribution Networks for 3D LiDAR Segmentation
by: Li, Li, et al.
Published: (2024)
by: Li, Li, et al.
Published: (2024)
MIEB: Massive Image Embedding Benchmark
by: Xiao, Chenghao, et al.
Published: (2025)
by: Xiao, Chenghao, et al.
Published: (2025)
One-Index Vector Quantization Based Adversarial Attack on Image Classification
by: Fan, Haiju, et al.
Published: (2024)
by: Fan, Haiju, et al.
Published: (2024)
Clustered Patch Embeddings for Permutation-Invariant Classification of Whole Slide Images
by: Gupta, Ravi Kant, et al.
Published: (2024)
by: Gupta, Ravi Kant, et al.
Published: (2024)
Grouped Discrete Representation for Object-Centric Learning
by: Zhao, Rongzhen, et al.
Published: (2024)
by: Zhao, Rongzhen, et al.
Published: (2024)
On the Role of Discrete Tokenization in Visual Representation Learning
by: Du, Tianqi, et al.
Published: (2024)
by: Du, Tianqi, et al.
Published: (2024)
Cross-View Graph Consistency Learning for Invariant Graph Representations
by: Chen, Jie, et al.
Published: (2023)
by: Chen, Jie, et al.
Published: (2023)
COD: Learning Conditional Invariant Representation for Domain Adaptation Regression
by: Yang, Hao-Ran, et al.
Published: (2024)
by: Yang, Hao-Ran, et al.
Published: (2024)
Maximizing Information in Domain-Invariant Representation Improves Transfer Learning
by: Li, Adrian Shuai, et al.
Published: (2023)
by: Li, Adrian Shuai, et al.
Published: (2023)
On the Design Fundamentals of Diffusion Models: A Survey
by: Chang, Ziyi, et al.
Published: (2023)
by: Chang, Ziyi, et al.
Published: (2023)
The Power of Next-Frame Prediction for Learning Physical Laws
by: Winterbottom, Thomas, et al.
Published: (2024)
by: Winterbottom, Thomas, et al.
Published: (2024)
Mixture of Group Experts for Learning Invariant Representations
by: Kang, Lei, et al.
Published: (2025)
by: Kang, Lei, et al.
Published: (2025)
Subject Invariant Contrastive Learning for Human Activity Recognition
by: Yarici, Yavuz, et al.
Published: (2025)
by: Yarici, Yavuz, et al.
Published: (2025)
Revisiting Transformation Invariant Geometric Deep Learning: An Initial Representation Perspective
by: Zhang, Ziwei, et al.
Published: (2021)
by: Zhang, Ziwei, et al.
Published: (2021)
When Invariant Representation Learning Meets Label Shift: Insufficiency and Theoretical Insights
by: Luo, You-Wei, et al.
Published: (2024)
by: Luo, You-Wei, et al.
Published: (2024)
Investigating the Benefits of Projection Head for Representation Learning
by: Xue, Yihao, et al.
Published: (2024)
by: Xue, Yihao, et al.
Published: (2024)
Stain-Invariant Representation for Tissue Classification in Histology Images
by: Raza, Manahil, et al.
Published: (2024)
by: Raza, Manahil, et al.
Published: (2024)
Self-Organising Neural Discrete Representation Learning à la Kohonen
by: Irie, Kazuki, et al.
Published: (2023)
by: Irie, Kazuki, et al.
Published: (2023)
Semi-Supervised Crowd Counting from Unlabeled Data
by: Duan, Haoran, et al.
Published: (2021)
by: Duan, Haoran, et al.
Published: (2021)
Sensor-Invariant Tactile Representation
by: Gupta, Harsh, et al.
Published: (2025)
by: Gupta, Harsh, et al.
Published: (2025)
Representation learning from OCT images
by: Tabia, Hedi, et al.
Published: (2026)
by: Tabia, Hedi, et al.
Published: (2026)
TRAJGANR: Trajectory-Centric Urban Multimodal Learning via Geospatially Aligned Neural Representations
by: Siampou, Maria Despoina, et al.
Published: (2026)
by: Siampou, Maria Despoina, et al.
Published: (2026)
UrbanFusion: Stochastic Multimodal Fusion for Contrastive Learning of Robust Spatial Representations
by: Mühlematter, Dominik J., et al.
Published: (2025)
by: Mühlematter, Dominik J., et al.
Published: (2025)
SwordBench: Evaluating Orthogonality of Steering Image Representations
by: Zaigrajew, Vladimir, et al.
Published: (2026)
by: Zaigrajew, Vladimir, et al.
Published: (2026)
Enhancing JEPAs with Spatial Conditioning: Robust and Efficient Representation Learning
by: Littwin, Etai, et al.
Published: (2024)
by: Littwin, Etai, et al.
Published: (2024)
Global Context-aware Representation Learning for Spatially Resolved Transcriptomics
by: Oh, Yunhak, et al.
Published: (2025)
by: Oh, Yunhak, et al.
Published: (2025)
Semantic Residual for Multimodal Unified Discrete Representation
by: Huang, Hai, et al.
Published: (2024)
by: Huang, Hai, et al.
Published: (2024)
Data-Driven Stochastic Motion Evaluation and Optimization with Image by Spatially-Aligned Temporal Encoding
by: Oba, Takeru, et al.
Published: (2023)
by: Oba, Takeru, et al.
Published: (2023)
LGQ: Learning Discretization Geometry for Scalable and Stable Image Tokenization
by: Altun, Idil Bilge, et al.
Published: (2026)
by: Altun, Idil Bilge, et al.
Published: (2026)
Text-Guided Image Invariant Feature Learning for Robust Image Watermarking
by: Ahtesham, Muhammad, et al.
Published: (2025)
by: Ahtesham, Muhammad, et al.
Published: (2025)
Spectral and Spatial Graph Learning for Multispectral Solar Image Compression
by: Siwakoti, Prasiddha, et al.
Published: (2025)
by: Siwakoti, Prasiddha, et al.
Published: (2025)
SONIC: Spectral Oriented Neural Invariant Convolutions
by: Moens, Gijs Joppe, et al.
Published: (2026)
by: Moens, Gijs Joppe, et al.
Published: (2026)
Similar Items
-
Controllable Image Generation with Composed Parallel Token Prediction
by: Stirling, Jamie, et al.
Published: (2024) -
Controllable Image Generation with Composed Parallel Token Prediction
by: Stirling, Jamie, et al.
Published: (2026) -
Disentangling Racial Phenotypes: Fine-Grained Control of Race-related Facial Phenotype Characteristics
by: Yucer, Seyma, et al.
Published: (2024) -
Video Prediction of Dynamic Physical Simulations With Pixel-Space Spatiotemporal Transformers
by: Slack, Dean L, et al.
Published: (2025) -
Everything is a Video: Unifying Modalities through Next-Frame Prediction
by: Hudson, G. Thomas, et al.
Published: (2024)