DReX: Pure Vision Fusion of Self-Supervised and Convolutional Representations for Image Complexity Prediction
Fuente:
arXiv
Saved in:
| Main Authors: | Skaza, Jonathan, Madinei, Parsa, Wen, Ziqi, Eckstein, Miguel |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
INTERLACE: Interleaved Layer Pruning and Efficient Adaptation in Large Vision-Language Models
by: Madinei, Parsa, et al.
Published: (2025)
by: Madinei, Parsa, et al.
Published: (2025)
Revealing the Gap in Human and VLM Scene Perception through Counterfactual Semantic Saliency
by: Wen, Ziqi, et al.
Published: (2026)
by: Wen, Ziqi, et al.
Published: (2026)
Predicting Reaction Time to Comprehend Scenes with Foveated Scene Understanding Maps
by: Wen, Ziqi, et al.
Published: (2025)
by: Wen, Ziqi, et al.
Published: (2025)
IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models
by: Madinei, Parsa, et al.
Published: (2026)
by: Madinei, Parsa, et al.
Published: (2026)
SSCR: Iterative Language-Based Image Editing via Self-Supervised Counterfactual Reasoning
by: Fu, Tsu-Jui, et al.
Published: (2020)
by: Fu, Tsu-Jui, et al.
Published: (2020)
Self-Supervised Learning of Plant Image Representations
by: Moummad, Ilyass, et al.
Published: (2026)
by: Moummad, Ilyass, et al.
Published: (2026)
Learning Accurate Segmentation Purely from Self-Supervision
by: You, Zuyao, et al.
Published: (2026)
by: You, Zuyao, et al.
Published: (2026)
Fusion from Decomposition: A Self-Supervised Approach for Image Fusion and Beyond
by: Liang, Pengwei, et al.
Published: (2024)
by: Liang, Pengwei, et al.
Published: (2024)
Self-Supervised Ultrasound-Video Segmentation with Feature Prediction and 3D Localised Loss
by: Ellis, Edward, et al.
Published: (2025)
by: Ellis, Edward, et al.
Published: (2025)
Why We Look Where We Look: Emergent Human-like Fixations of a Foveated Visual Language Model Maximizing Scene Understanding
by: Murlidaran, Shravan, et al.
Published: (2026)
by: Murlidaran, Shravan, et al.
Published: (2026)
Geometric Analysis of Self-Supervised Vision Representations for Semantic Image Retrieval
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
Self-Supervised Scene Flow Estimation with Point-Voxel Fusion and Surface Representation
by: Xiang, Xuezhi, et al.
Published: (2024)
by: Xiang, Xuezhi, et al.
Published: (2024)
SwinIA: Self-Supervised Blind-Spot Image Denoising without Convolutions
by: Papkov, Mikhail, et al.
Published: (2023)
by: Papkov, Mikhail, et al.
Published: (2023)
AFiRe: Anatomy-Driven Self-Supervised Learning for Fine-Grained Representation in Radiographic Images
by: Liu, Yihang, et al.
Published: (2025)
by: Liu, Yihang, et al.
Published: (2025)
SSVIF: Self-Supervised Segmentation-Oriented Visible and Infrared Image Fusion
by: Zhao, Zixian, et al.
Published: (2025)
by: Zhao, Zixian, et al.
Published: (2025)
From Local Cues to Global Percepts: Emergent Gestalt Organization in Self-Supervised Vision Models
by: Li, Tianqin, et al.
Published: (2025)
by: Li, Tianqin, et al.
Published: (2025)
LEGO: Self-Supervised Representation Learning for Scene Text Images
by: Ren, Yujin, et al.
Published: (2024)
by: Ren, Yujin, et al.
Published: (2024)
Enhancing Representations through Heterogeneous Self-Supervised Learning
by: Li, Zhong-Yu, et al.
Published: (2023)
by: Li, Zhong-Yu, et al.
Published: (2023)
Hierarchical Text-to-Vision Self Supervised Alignment for Improved Histopathology Representation Learning
by: Watawana, Hasindri, et al.
Published: (2024)
by: Watawana, Hasindri, et al.
Published: (2024)
Supervised and Contrastive Self-Supervised In-Domain Representation Learning for Dense Prediction Problems in Remote Sensing
by: Ghanbarzade, Ali, et al.
Published: (2023)
by: Ghanbarzade, Ali, et al.
Published: (2023)
On Convolutional Vision Transformers for Yield Prediction
by: Inderka, Alvin, et al.
Published: (2024)
by: Inderka, Alvin, et al.
Published: (2024)
SITUATE: Indoor Human Trajectory Prediction through Geometric Features and Self-Supervised Vision Representation
by: Capogrosso, Luigi, et al.
Published: (2024)
by: Capogrosso, Luigi, et al.
Published: (2024)
CNN-JEPA: Self-Supervised Pretraining Convolutional Neural Networks Using Joint Embedding Predictive Architecture
by: Kalapos, András, et al.
Published: (2024)
by: Kalapos, András, et al.
Published: (2024)
Robust Human Trajectory Prediction via Self-Supervised Skeleton Representation Learning
by: Arashima, Taishu, et al.
Published: (2026)
by: Arashima, Taishu, et al.
Published: (2026)
Le MuMo JEPA: Multi-Modal Self-Supervised Representation Learning with Learnable Fusion Tokens
by: Cornelissen, Ciem, et al.
Published: (2026)
by: Cornelissen, Ciem, et al.
Published: (2026)
Weakly-Supervised Image Forgery Localization via Vision-Language Collaborative Reasoning Framework
by: Sheng, Ziqi, et al.
Published: (2025)
by: Sheng, Ziqi, et al.
Published: (2025)
On the Discriminability of Self-Supervised Representation Learning
by: Song, Zeen, et al.
Published: (2024)
by: Song, Zeen, et al.
Published: (2024)
MixDiff: Mixing Natural and Synthetic Images for Robust Self-Supervised Representations
by: Bafghi, Reza Akbarian, et al.
Published: (2024)
by: Bafghi, Reza Akbarian, et al.
Published: (2024)
Information-Maximized Soft Variable Discretization for Self-Supervised Image Representation Learning
by: Niu, Chuang, et al.
Published: (2025)
by: Niu, Chuang, et al.
Published: (2025)
Robust Fusion of Object-Level V2X for Learned 3D Object Detection
by: Ostendorf, Lukas, et al.
Published: (2026)
by: Ostendorf, Lukas, et al.
Published: (2026)
Depth-Wise Representation Development Under Blockwise Self-Supervised Learning for Video Vision Transformers
by: Römer, Jonas, et al.
Published: (2026)
by: Römer, Jonas, et al.
Published: (2026)
Self-Supervised Ultrasound Representation Learning for Renal Anomaly Prediction in Prenatal Imaging
by: Megahed, Youssef, et al.
Published: (2025)
by: Megahed, Youssef, et al.
Published: (2025)
Adapting Self-Supervised Representations as a Latent Space for Efficient Generation
by: Gui, Ming, et al.
Published: (2025)
by: Gui, Ming, et al.
Published: (2025)
Learning Self-Prior for Mesh Inpainting Using Self-Supervised Graph Convolutional Networks
by: Hattori, Shota, et al.
Published: (2023)
by: Hattori, Shota, et al.
Published: (2023)
Adaptive 3D Convolution for Remote Sensing Image Fusion
by: Peng, Siran, et al.
Published: (2026)
by: Peng, Siran, et al.
Published: (2026)
Equivariant Representation Learning for Augmentation-based Self-Supervised Learning via Image Reconstruction
by: Wang, Qin, et al.
Published: (2024)
by: Wang, Qin, et al.
Published: (2024)
Self-Supervised Learning Based on Transformed Image Reconstruction for Equivariance-Coherent Feature Representation
by: Wang, Qin, et al.
Published: (2025)
by: Wang, Qin, et al.
Published: (2025)
CrossVideoMAE: Self-Supervised Image-Video Representation Learning with Masked Autoencoders
by: Ahamed, Shihab Aaqil, et al.
Published: (2025)
by: Ahamed, Shihab Aaqil, et al.
Published: (2025)
Self-Supervised Anatomical Consistency Learning for Vision-Grounded Medical Report Generation
by: Yang, Longzhen, et al.
Published: (2025)
by: Yang, Longzhen, et al.
Published: (2025)
CoMiX: Cross-Modal Fusion with Deformable Convolutions for HSI-X Semantic Segmentation
by: Zhang, Xuming, et al.
Published: (2024)
by: Zhang, Xuming, et al.
Published: (2024)
Similar Items
-
INTERLACE: Interleaved Layer Pruning and Efficient Adaptation in Large Vision-Language Models
by: Madinei, Parsa, et al.
Published: (2025) -
Revealing the Gap in Human and VLM Scene Perception through Counterfactual Semantic Saliency
by: Wen, Ziqi, et al.
Published: (2026) -
Predicting Reaction Time to Comprehend Scenes with Foveated Scene Understanding Maps
by: Wen, Ziqi, et al.
Published: (2025) -
IRIS: Intent Resolution via Inference-time Saccades for Open-Ended VQA in Large Vision-Language Models
by: Madinei, Parsa, et al.
Published: (2026) -
SSCR: Iterative Language-Based Image Editing via Self-Supervised Counterfactual Reasoning
by: Fu, Tsu-Jui, et al.
Published: (2020)