Gespeichert in:
| Hauptverfasser: | Schäfer, Frederik, Mandl, Luis, Kälber, Lars, Ricken, Tim |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2605.05908 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Self-supervised pretraining for an iterative image size agnostic vision transformer
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2026)
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2026)
an interpretable vision transformer framework for automated brain tumor classification
von: Mbonu, Chinedu Emmanuel, et al.
Veröffentlicht: (2026)
von: Mbonu, Chinedu Emmanuel, et al.
Veröffentlicht: (2026)
HSFusion: A high-level vision task-driven infrared and visible image fusion network via semantic and geometric domain transformation
von: Jiang, Chengjie, et al.
Veröffentlicht: (2024)
von: Jiang, Chengjie, et al.
Veröffentlicht: (2024)
A novel network for classification of cuneiform tablet metadata
von: Hagelskjær, Frederik
Veröffentlicht: (2026)
von: Hagelskjær, Frederik
Veröffentlicht: (2026)
TREAD: Token Routing for Efficient Architecture-agnostic Diffusion Training
von: Krause, Felix, et al.
Veröffentlicht: (2025)
von: Krause, Felix, et al.
Veröffentlicht: (2025)
Architecture and evaluation protocol for transformer-based visual object tracking in UAV applications
von: Borne, Augustin, et al.
Veröffentlicht: (2026)
von: Borne, Augustin, et al.
Veröffentlicht: (2026)
A hierarchical semantic segmentation framework for computer vision-based bridge damage detection
von: Liu, Jingxiao, et al.
Veröffentlicht: (2022)
von: Liu, Jingxiao, et al.
Veröffentlicht: (2022)
Enhancing the vision-language foundation model with key semantic knowledge-emphasized report refinement
von: Huang, Weijian, et al.
Veröffentlicht: (2024)
von: Huang, Weijian, et al.
Veröffentlicht: (2024)
On the effectiveness of multimodal privileged knowledge distillation in two vision transformer based diagnostic applications
von: Baur, Simon, et al.
Veröffentlicht: (2025)
von: Baur, Simon, et al.
Veröffentlicht: (2025)
STARS: Sensor-agnostic Transformer Architecture for Remote Sensing
von: King, Ethan, et al.
Veröffentlicht: (2024)
von: King, Ethan, et al.
Veröffentlicht: (2024)
Separable DeepONet: Breaking the Curse of Dimensionality in Physics-Informed Machine Learning
von: Mandl, Luis, et al.
Veröffentlicht: (2024)
von: Mandl, Luis, et al.
Veröffentlicht: (2024)
Physics-Informed Time-Integrated DeepONet: Temporal Tangent Space Operator Learning for High-Accuracy Inference
von: Mandl, Luis, et al.
Veröffentlicht: (2025)
von: Mandl, Luis, et al.
Veröffentlicht: (2025)
MinkOcc: Towards real-time label-efficient semantic occupancy prediction
von: Sze, Samuel, et al.
Veröffentlicht: (2025)
von: Sze, Samuel, et al.
Veröffentlicht: (2025)
NormalView: sensor-agnostic tree species classification from backpack and aerial lidar data using geometric projections
von: Korkeala, Juho, et al.
Veröffentlicht: (2025)
von: Korkeala, Juho, et al.
Veröffentlicht: (2025)
A comprehensive overview of deep learning techniques for 3D point cloud classification and semantic segmentation
von: Sarker, Sushmita, et al.
Veröffentlicht: (2024)
von: Sarker, Sushmita, et al.
Veröffentlicht: (2024)
Real-time 3D semantic occupancy prediction for autonomous vehicles using memory-efficient sparse convolution
von: Sze, Samuel, et al.
Veröffentlicht: (2024)
von: Sze, Samuel, et al.
Veröffentlicht: (2024)
Cross multiscale vision transformer for deep fake detection
von: P, Akhshan, et al.
Veröffentlicht: (2025)
von: P, Akhshan, et al.
Veröffentlicht: (2025)
Automated diagnosis of lung diseases using vision transformer: a comparative study on chest x-ray classification
von: Ahmad, Muhammad, et al.
Veröffentlicht: (2025)
von: Ahmad, Muhammad, et al.
Veröffentlicht: (2025)
Beyond the final layer: Attentive multilayer fusion for vision transformers
von: Ciernik, Laure, et al.
Veröffentlicht: (2026)
von: Ciernik, Laure, et al.
Veröffentlicht: (2026)
Category-aware EEG image generation based on wavelet transform and contrast semantic loss
von: Zhang, Enshang, et al.
Veröffentlicht: (2025)
von: Zhang, Enshang, et al.
Veröffentlicht: (2025)
Depth-agnostic Single Image Dehazing
von: Xu, Honglei, et al.
Veröffentlicht: (2024)
von: Xu, Honglei, et al.
Veröffentlicht: (2024)
Towards Domain-agnostic Depth Completion
von: Xu, Guangkai, et al.
Veröffentlicht: (2022)
von: Xu, Guangkai, et al.
Veröffentlicht: (2022)
FISHing in Uncertainty: Synthetic Contrastive Learning for Genetic Aberration Detection
von: Gutwein, Simon, et al.
Veröffentlicht: (2024)
von: Gutwein, Simon, et al.
Veröffentlicht: (2024)
Interpreting vision transformers via residual replacement model
von: Kim, Jinyeong, et al.
Veröffentlicht: (2025)
von: Kim, Jinyeong, et al.
Veröffentlicht: (2025)
A survey on efficient vision transformers: algorithms, techniques, and performance benchmarking
von: Papa, Lorenzo, et al.
Veröffentlicht: (2023)
von: Papa, Lorenzo, et al.
Veröffentlicht: (2023)
METER: a mobile vision transformer architecture for monocular depth estimation
von: Papa, L., et al.
Veröffentlicht: (2024)
von: Papa, L., et al.
Veröffentlicht: (2024)
Steering CLIP's vision transformer with sparse autoencoders
von: Joseph, Sonia, et al.
Veröffentlicht: (2025)
von: Joseph, Sonia, et al.
Veröffentlicht: (2025)
SPRINT: Script-agnostic Structure Recognition in Tables
von: Kudale, Dhruv, et al.
Veröffentlicht: (2025)
von: Kudale, Dhruv, et al.
Veröffentlicht: (2025)
Clothing agnostic Pre-inpainting Virtual Try-ON
von: Kim, Sehyun, et al.
Veröffentlicht: (2025)
von: Kim, Sehyun, et al.
Veröffentlicht: (2025)
Scene-agnostic Pose Regression for Visual Localization
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
von: Zheng, Junwei, et al.
Veröffentlicht: (2025)
SAFER: Sharpness Aware layer-selective Finetuning for Enhanced Robustness in vision transformers
von: Gopal, Bhavna, et al.
Veröffentlicht: (2025)
von: Gopal, Bhavna, et al.
Veröffentlicht: (2025)
Low-latency vision transformers via large-scale multi-head attention
von: Gross, Ronit D., et al.
Veröffentlicht: (2025)
von: Gross, Ronit D., et al.
Veröffentlicht: (2025)
Initialization matters in few-shot adaptation of vision-language models for histopathological image classification
von: Meseguer, Pablo, et al.
Veröffentlicht: (2026)
von: Meseguer, Pablo, et al.
Veröffentlicht: (2026)
Are vision-language models ready to zero-shot replace supervised classification models in agriculture?
von: Ranario, Earl, et al.
Veröffentlicht: (2025)
von: Ranario, Earl, et al.
Veröffentlicht: (2025)
On the application of the Wasserstein metric to 2D curves classification
von: Kaliszewska, Agnieszka, et al.
Veröffentlicht: (2026)
von: Kaliszewska, Agnieszka, et al.
Veröffentlicht: (2026)
PIV3CAMS: a multi-camera dataset for multiple computer vision problems and its application to novel view-point synthesis
von: Kim, Sohyeong, et al.
Veröffentlicht: (2024)
von: Kim, Sohyeong, et al.
Veröffentlicht: (2024)
$L^3$:Scene-agnostic Visual Localization in the Wild
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
von: Zhang, Yu, et al.
Veröffentlicht: (2026)
Preserving Marker Specificity with Lightweight Channel-Independent Representation Learning
von: Gutwein, Simon, et al.
Veröffentlicht: (2025)
von: Gutwein, Simon, et al.
Veröffentlicht: (2025)
PTQ4ViT: Post-training quantization for vision transformers with twin uniform quantization
von: Yuan, Zhihang, et al.
Veröffentlicht: (2021)
von: Yuan, Zhihang, et al.
Veröffentlicht: (2021)
Two-Stream temporal transformer for video action classification
von: Kurpukdee, Nattapong, et al.
Veröffentlicht: (2026)
von: Kurpukdee, Nattapong, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Self-supervised pretraining for an iterative image size agnostic vision transformer
von: Prisadnikov, Nedyalko, et al.
Veröffentlicht: (2026) -
an interpretable vision transformer framework for automated brain tumor classification
von: Mbonu, Chinedu Emmanuel, et al.
Veröffentlicht: (2026) -
HSFusion: A high-level vision task-driven infrared and visible image fusion network via semantic and geometric domain transformation
von: Jiang, Chengjie, et al.
Veröffentlicht: (2024) -
A novel network for classification of cuneiform tablet metadata
von: Hagelskjær, Frederik
Veröffentlicht: (2026) -
TREAD: Token Routing for Efficient Architecture-agnostic Diffusion Training
von: Krause, Felix, et al.
Veröffentlicht: (2025)