Dynamic Mode Decomposition along Depth in Vision Transformers
Fuente:
arXiv
Saved in:
| Main Authors: | Aswani, Nishant Suresh, Jabari, Saif Eddin |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Koopman Autoencoders Learn Neural Representation Dynamics
by: Aswani, Nishant Suresh, et al.
Published: (2025)
by: Aswani, Nishant Suresh, et al.
Published: (2025)
Exploring the Interplay of Interpretability and Robustness in Deep Neural Networks: A Saliency-guided Approach
by: Guesmi, Amira, et al.
Published: (2024)
by: Guesmi, Amira, et al.
Published: (2024)
Representing Neural Network Layers as Linear Operations via Koopman Operator Theory
by: Aswani, Nishant Suresh, et al.
Published: (2024)
by: Aswani, Nishant Suresh, et al.
Published: (2024)
ModeT: Learning Deformable Image Registration via Motion Decomposition Transformer
by: Wang, Haiqiao, et al.
Published: (2023)
by: Wang, Haiqiao, et al.
Published: (2023)
Real-Time Motion Detection Using Dynamic Mode Decomposition
by: Mignacca, Marco, et al.
Published: (2024)
by: Mignacca, Marco, et al.
Published: (2024)
Rethinking Vision Transformer Depth via Structural Reparameterization
by: Zhou, Chengwei, et al.
Published: (2025)
by: Zhou, Chengwei, et al.
Published: (2025)
Interpretable Vision Transformers in Monocular Depth Estimation via SVDA
by: Arampatzakis, Vasileios, et al.
Published: (2026)
by: Arampatzakis, Vasileios, et al.
Published: (2026)
Depth-Wise Convolutions in Vision Transformers for Efficient Training on Small Datasets
by: Zhang, Tianxiao, et al.
Published: (2024)
by: Zhang, Tianxiao, et al.
Published: (2024)
RD-ViT: Recurrent-Depth Vision Transformer for Semantic Segmentation with Reduced Data Dependence Extending the Recurrent-Depth Transformer Architecture to Dense Prediction
by: He, Renjie
Published: (2026)
by: He, Renjie
Published: (2026)
Bi-Orthogonal Factor Decomposition for Vision Transformers
by: Doshi, Fenil R., et al.
Published: (2026)
by: Doshi, Fenil R., et al.
Published: (2026)
ModeTv2: GPU-accelerated Motion Decomposition Transformer for Pairwise Optimization in Medical Image Registration
by: Wang, Haiqiao, et al.
Published: (2024)
by: Wang, Haiqiao, et al.
Published: (2024)
Dynamic Weight Adjustment for Knowledge Distillation: Leveraging Vision Transformer for High-Accuracy Lung Cancer Detection and Real-Time Deployment
by: Khan, Saif Ur Rehman, et al.
Published: (2025)
by: Khan, Saif Ur Rehman, et al.
Published: (2025)
Identifying Spatio-Temporal Drivers of Extreme Events
by: Eddin, Mohamad Hakam Shams, et al.
Published: (2024)
by: Eddin, Mohamad Hakam Shams, et al.
Published: (2024)
Recursive Vision Transformer with Dynamic Depth and Width Adjustment for Resource-Efficient Image Semantic Communication
by: Zhang, Zhilong, et al.
Published: (2026)
by: Zhang, Zhilong, et al.
Published: (2026)
Light-Field Dataset for Disparity Based Depth Estimation
by: Nehra, Suresh, et al.
Published: (2025)
by: Nehra, Suresh, et al.
Published: (2025)
Tri-Perspective View Decomposition for Geometry-Aware Depth Completion
by: Yan, Zhiqiang, et al.
Published: (2024)
by: Yan, Zhiqiang, et al.
Published: (2024)
Illustrator's Depth: Monocular Layer Index Prediction for Image Decomposition
by: Maruani, Nissim, et al.
Published: (2025)
by: Maruani, Nissim, et al.
Published: (2025)
DepthLM: Metric Depth From Vision Language Models
by: Cai, Zhipeng, et al.
Published: (2025)
by: Cai, Zhipeng, et al.
Published: (2025)
Depth-Wise Representation Development Under Blockwise Self-Supervised Learning for Video Vision Transformers
by: Römer, Jonas, et al.
Published: (2026)
by: Römer, Jonas, et al.
Published: (2026)
UAV-VLN: End-to-End Vision Language guided Navigation for UAVs
by: Saxena, Pranav, et al.
Published: (2025)
by: Saxena, Pranav, et al.
Published: (2025)
From Edges to Depth: Probing the Spatial Hierarchy in Vision Transformers
by: Sanghavi, Jainum
Published: (2026)
by: Sanghavi, Jainum
Published: (2026)
A Dynamic Mode Decomposition Approach to Morphological Component Analysis
by: Huber, Owen T., et al.
Published: (2025)
by: Huber, Owen T., et al.
Published: (2025)
Image Safeguarding: Reasoning with Conditional Vision Language Model and Obfuscating Unsafe Content Counterfactually
by: Bethany, Mazal, et al.
Published: (2024)
by: Bethany, Mazal, et al.
Published: (2024)
DepthCues: Evaluating Monocular Depth Perception in Large Vision Models
by: Danier, Duolikun, et al.
Published: (2024)
by: Danier, Duolikun, et al.
Published: (2024)
Automatic Cardiac Pathology Recognition in Echocardiography Images Using Higher Order Dynamic Mode Decomposition and a Vision Transformer for Small Datasets
by: Bell-Navas, Andrés, et al.
Published: (2024)
by: Bell-Navas, Andrés, et al.
Published: (2024)
Towards Depth Foundation Model: Recent Trends in Vision-Based Depth Estimation
by: Xu, Zhen, et al.
Published: (2025)
by: Xu, Zhen, et al.
Published: (2025)
UPDP: A Unified Progressive Depth Pruner for CNN and Vision Transformer
by: Liu, Ji, et al.
Published: (2024)
by: Liu, Ji, et al.
Published: (2024)
αDepth: Learning Single-Pass Soft Boundary Decomposition for Stereo Conversion
by: Zhang, Xiang, et al.
Published: (2026)
by: Zhang, Xiang, et al.
Published: (2026)
Depth-wise Decomposition for Accelerating Separable Convolutions in Efficient Convolutional Neural Networks
by: He, Yihui, et al.
Published: (2019)
by: He, Yihui, et al.
Published: (2019)
Shape2.5D: A Dataset of Texture-less Surfaces for Depth and Normals Estimation
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
by: Khan, Muhammad Saif Ullah, et al.
Published: (2024)
Fisheye Stereo Vision: Depth and Range Error
by: Jiang, Leaf, et al.
Published: (2026)
by: Jiang, Leaf, et al.
Published: (2026)
4DVGGT-D: 4D Visual Geometry Transformer with Improved Dynamic Depth Estimation
by: Zang, Ying, et al.
Published: (2026)
by: Zang, Ying, et al.
Published: (2026)
Vision-Language Embodiment for Monocular Depth Estimation
by: Zhang, Jinchang, et al.
Published: (2025)
by: Zhang, Jinchang, et al.
Published: (2025)
EndoDepthL: Lightweight Endoscopic Monocular Depth Estimation with CNN-Transformer
by: Li, Yangke
Published: (2023)
by: Li, Yangke
Published: (2023)
DepthVLA: Enhancing Vision-Language-Action Models with Depth-Aware Spatial Reasoning
by: Yuan, Tianyuan, et al.
Published: (2025)
by: Yuan, Tianyuan, et al.
Published: (2025)
SHADeS: Self-supervised Monocular Depth Estimation Through Non-Lambertian Image Decomposition
by: Daher, Rema, et al.
Published: (2025)
by: Daher, Rema, et al.
Published: (2025)
FiffDepth: Feed-forward Transformation of Diffusion-Based Generators for Detailed Depth Estimation
by: Bai, Yunpeng, et al.
Published: (2024)
by: Bai, Yunpeng, et al.
Published: (2024)
RVLM: Recursive Vision-Language Models with Adaptive Depth
by: Mayumu, Nicanor, et al.
Published: (2026)
by: Mayumu, Nicanor, et al.
Published: (2026)
DORA: Dynamic Online Reinforcement Agent for Token Merging in Vision Transformers
by: He, Kaixuan, et al.
Published: (2026)
by: He, Kaixuan, et al.
Published: (2026)
Distillation Dynamics: Towards Understanding Feature-Based Distillation in Vision Transformers
by: Tian, Huiyuan, et al.
Published: (2025)
by: Tian, Huiyuan, et al.
Published: (2025)
Similar Items
-
Koopman Autoencoders Learn Neural Representation Dynamics
by: Aswani, Nishant Suresh, et al.
Published: (2025) -
Exploring the Interplay of Interpretability and Robustness in Deep Neural Networks: A Saliency-guided Approach
by: Guesmi, Amira, et al.
Published: (2024) -
Representing Neural Network Layers as Linear Operations via Koopman Operator Theory
by: Aswani, Nishant Suresh, et al.
Published: (2024) -
ModeT: Learning Deformable Image Registration via Motion Decomposition Transformer
by: Wang, Haiqiao, et al.
Published: (2023) -
Real-Time Motion Detection Using Dynamic Mode Decomposition
by: Mignacca, Marco, et al.
Published: (2024)