HEAL-SWIN: A Vision Transformer On The Sphere
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Carlsson, Oscar, Gerken, Jan E., Linander, Hampus, Spieß, Heiner, Ohlsson, Fredrik, Petersson, Christoffer, Persson, Daniel |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PEAR: Equal Area Weather Forecasting on the Sphere
von: Linander, Hampus, et al.
Veröffentlicht: (2025)
von: Linander, Hampus, et al.
Veröffentlicht: (2025)
Bayesian posterior approximation with stochastic ensembles
von: Balabanov, Oleksandr, et al.
Veröffentlicht: (2022)
von: Balabanov, Oleksandr, et al.
Veröffentlicht: (2022)
Raw or Cooked? Object Detection on RAW Images
von: Ljungbergh, William, et al.
Veröffentlicht: (2023)
von: Ljungbergh, William, et al.
Veröffentlicht: (2023)
HEAL-ViT: Vision Transformers on a spherical mesh for medium-range weather forecasting
von: Ramavajjala, Vivek
Veröffentlicht: (2024)
von: Ramavajjala, Vivek
Veröffentlicht: (2024)
Domain Generalized Recaptured Screen Image Identification Using SWIN Transformer
von: Mehta, Preeti, et al.
Veröffentlicht: (2024)
von: Mehta, Preeti, et al.
Veröffentlicht: (2024)
Faster Inference of Integer SWIN Transformer by Removing the GELU Activation
von: Tayaranian, Mohammadreza, et al.
Veröffentlicht: (2024)
von: Tayaranian, Mohammadreza, et al.
Veröffentlicht: (2024)
Learning Chern Numbers of Topological Insulators with Gauge Equivariant Neural Networks
von: Huang, Longde, et al.
Veröffentlicht: (2025)
von: Huang, Longde, et al.
Veröffentlicht: (2025)
SKGE-SWIN: End-To-End Autonomous Vehicle Waypoint Prediction and Navigation Using Skip Stage Swin Transformer
von: Kartiman, Fachri Najm Noer, et al.
Veröffentlicht: (2025)
von: Kartiman, Fachri Najm Noer, et al.
Veröffentlicht: (2025)
R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation
von: Ljungbergh, William, et al.
Veröffentlicht: (2025)
von: Ljungbergh, William, et al.
Veröffentlicht: (2025)
Learning from Noisy Labels with Contrastive Co-Transformer
von: Han, Yan, et al.
Veröffentlicht: (2025)
von: Han, Yan, et al.
Veröffentlicht: (2025)
Octic Vision Transformers: Quicker ViTs Through Equivariance
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
SplatAD: Real-Time Lidar and Camera Rendering with 3D Gaussian Splatting for Autonomous Driving
von: Hess, Georg, et al.
Veröffentlicht: (2024)
von: Hess, Georg, et al.
Veröffentlicht: (2024)
NeuRAD: Neural Rendering for Autonomous Driving
von: Tonderski, Adam, et al.
Veröffentlicht: (2023)
von: Tonderski, Adam, et al.
Veröffentlicht: (2023)
R3D-SWIN:Use Shifted Window Attention for Single-View 3D Reconstruction
von: Li, Chenhuan, et al.
Veröffentlicht: (2023)
von: Li, Chenhuan, et al.
Veröffentlicht: (2023)
Uncertainty quantification in fine-tuned LLMs using LoRA ensembles
von: Balabanov, Oleksandr, et al.
Veröffentlicht: (2024)
von: Balabanov, Oleksandr, et al.
Veröffentlicht: (2024)
Parabolic Position Encoding: Vision-Centric, Principled, Extrapolatable, General
von: Øhrstrøm, Christoffer Koo, et al.
Veröffentlicht: (2026)
von: Øhrstrøm, Christoffer Koo, et al.
Veröffentlicht: (2026)
LIB-KD: Teaching Inductive Bias for Efficient Vision Transformer Distillation and Compression
von: Habib, Gousia, et al.
Veröffentlicht: (2023)
von: Habib, Gousia, et al.
Veröffentlicht: (2023)
Muon in Vision Transformers: Optimizer-Recipe Interactions and Gradient Spectra
von: Southworth, Ben S., et al.
Veröffentlicht: (2026)
von: Southworth, Ben S., et al.
Veröffentlicht: (2026)
Native Segmentation Vision Transformers
von: Brasó, Guillem, et al.
Veröffentlicht: (2025)
von: Brasó, Guillem, et al.
Veröffentlicht: (2025)
NeuroNCAP: Photorealistic Closed-loop Safety Testing for Autonomous Driving
von: Ljungbergh, William, et al.
Veröffentlicht: (2024)
von: Ljungbergh, William, et al.
Veröffentlicht: (2024)
DiffiT: Diffusion Vision Transformers for Image Generation
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
von: Hatamizadeh, Ali, et al.
Veröffentlicht: (2023)
ViTNT-FIQA: Training-Free Face Image Quality Assessment with Vision Transformers
von: Ozgur, Guray, et al.
Veröffentlicht: (2026)
von: Ozgur, Guray, et al.
Veröffentlicht: (2026)
Iwin Transformer: Hierarchical Vision Transformer using Interleaved Windows
von: Huo, Simin, et al.
Veröffentlicht: (2025)
von: Huo, Simin, et al.
Veröffentlicht: (2025)
Learning Temporal Saliency for Time Series Forecasting with Cross-Scale Attention
von: Delibasoglu, Ibrahim, et al.
Veröffentlicht: (2025)
von: Delibasoglu, Ibrahim, et al.
Veröffentlicht: (2025)
Slicing Vision Transformer for Flexible Inference
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
von: Zhang, Yitian, et al.
Veröffentlicht: (2024)
RAViT: Resolution-Adaptive Vision Transformer
von: Guidez, Martial, et al.
Veröffentlicht: (2026)
von: Guidez, Martial, et al.
Veröffentlicht: (2026)
Rotary Position Embedding for Vision Transformer
von: Heo, Byeongho, et al.
Veröffentlicht: (2024)
von: Heo, Byeongho, et al.
Veröffentlicht: (2024)
Decorrelation Speeds Up Vision Transformers
von: Carrigg, Kieran, et al.
Veröffentlicht: (2025)
von: Carrigg, Kieran, et al.
Veröffentlicht: (2025)
Hybrid Convolution and Vision Transformer NAS Search Space for TinyML Image Classification
von: Djajapermana, Mikhael, et al.
Veröffentlicht: (2025)
von: Djajapermana, Mikhael, et al.
Veröffentlicht: (2025)
OmniEarth-Bench: Towards Holistic Evaluation of Earth's Six Spheres and Cross-Spheres Interactions with Multimodal Observational Earth Data
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
von: Wang, Fengxiang, et al.
Veröffentlicht: (2025)
MuM: Multi-View Masked Image Modeling for 3D Vision
von: Nordström, David, et al.
Veröffentlicht: (2025)
von: Nordström, David, et al.
Veröffentlicht: (2025)
Weather-Robust Scene Semantics with Vision-Aligned 4D Radar
von: Hamilton, Kali, et al.
Veröffentlicht: (2026)
von: Hamilton, Kali, et al.
Veröffentlicht: (2026)
Video-Based Inpatient Fall Risk Assessment: A Case Study
von: Wang, Ziqing, et al.
Veröffentlicht: (2021)
von: Wang, Ziqing, et al.
Veröffentlicht: (2021)
In-Bed Human Pose Estimation from Unseen and Privacy-Preserving Image Domains
von: Cao, Ting, et al.
Veröffentlicht: (2021)
von: Cao, Ting, et al.
Veröffentlicht: (2021)
Timeline-based Process Discovery
von: Kaur, Harleen, et al.
Veröffentlicht: (2023)
von: Kaur, Harleen, et al.
Veröffentlicht: (2023)
Faster-HEAL: An Efficient and Privacy-Preserving Collaborative Perception Framework for Heterogeneous Autonomous Vehicles
von: Maleki, Armin, et al.
Veröffentlicht: (2026)
von: Maleki, Armin, et al.
Veröffentlicht: (2026)
SPoT: Subpixel Placement of Tokens in Vision Transformers
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2025)
von: Hjelkrem-Tan, Martine, et al.
Veröffentlicht: (2025)
Instance-Aware Group Quantization for Vision Transformers
von: Moon, Jaehyeon, et al.
Veröffentlicht: (2024)
von: Moon, Jaehyeon, et al.
Veröffentlicht: (2024)
Vision Transformer-based Adversarial Domain Adaptation
von: Li, Yahan, et al.
Veröffentlicht: (2024)
von: Li, Yahan, et al.
Veröffentlicht: (2024)
Elastic Attention Cores for Scalable Vision Transformers
von: Song, Alan Z., et al.
Veröffentlicht: (2026)
von: Song, Alan Z., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
PEAR: Equal Area Weather Forecasting on the Sphere
von: Linander, Hampus, et al.
Veröffentlicht: (2025) -
Bayesian posterior approximation with stochastic ensembles
von: Balabanov, Oleksandr, et al.
Veröffentlicht: (2022) -
Raw or Cooked? Object Detection on RAW Images
von: Ljungbergh, William, et al.
Veröffentlicht: (2023) -
HEAL-ViT: Vision Transformers on a spherical mesh for medium-range weather forecasting
von: Ramavajjala, Vivek
Veröffentlicht: (2024) -
Domain Generalized Recaptured Screen Image Identification Using SWIN Transformer
von: Mehta, Preeti, et al.
Veröffentlicht: (2024)