HEAL-SWIN: A Vision Transformer On The Sphere
Fuente:
arXiv
Guardado en:
| Autores principales: | Carlsson, Oscar, Gerken, Jan E., Linander, Hampus, Spieß, Heiner, Ohlsson, Fredrik, Petersson, Christoffer, Persson, Daniel |
|---|---|
| Formato: | Preprint |
| Publicado: |
2023
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
PEAR: Equal Area Weather Forecasting on the Sphere
por: Linander, Hampus, et al.
Publicado: (2025)
por: Linander, Hampus, et al.
Publicado: (2025)
Bayesian posterior approximation with stochastic ensembles
por: Balabanov, Oleksandr, et al.
Publicado: (2022)
por: Balabanov, Oleksandr, et al.
Publicado: (2022)
Raw or Cooked? Object Detection on RAW Images
por: Ljungbergh, William, et al.
Publicado: (2023)
por: Ljungbergh, William, et al.
Publicado: (2023)
HEAL-ViT: Vision Transformers on a spherical mesh for medium-range weather forecasting
por: Ramavajjala, Vivek
Publicado: (2024)
por: Ramavajjala, Vivek
Publicado: (2024)
Domain Generalized Recaptured Screen Image Identification Using SWIN Transformer
por: Mehta, Preeti, et al.
Publicado: (2024)
por: Mehta, Preeti, et al.
Publicado: (2024)
Faster Inference of Integer SWIN Transformer by Removing the GELU Activation
por: Tayaranian, Mohammadreza, et al.
Publicado: (2024)
por: Tayaranian, Mohammadreza, et al.
Publicado: (2024)
Learning Chern Numbers of Topological Insulators with Gauge Equivariant Neural Networks
por: Huang, Longde, et al.
Publicado: (2025)
por: Huang, Longde, et al.
Publicado: (2025)
SKGE-SWIN: End-To-End Autonomous Vehicle Waypoint Prediction and Navigation Using Skip Stage Swin Transformer
por: Kartiman, Fachri Najm Noer, et al.
Publicado: (2025)
por: Kartiman, Fachri Najm Noer, et al.
Publicado: (2025)
R3D2: Realistic 3D Asset Insertion via Diffusion for Autonomous Driving Simulation
por: Ljungbergh, William, et al.
Publicado: (2025)
por: Ljungbergh, William, et al.
Publicado: (2025)
Learning from Noisy Labels with Contrastive Co-Transformer
por: Han, Yan, et al.
Publicado: (2025)
por: Han, Yan, et al.
Publicado: (2025)
Octic Vision Transformers: Quicker ViTs Through Equivariance
por: Nordström, David, et al.
Publicado: (2025)
por: Nordström, David, et al.
Publicado: (2025)
SplatAD: Real-Time Lidar and Camera Rendering with 3D Gaussian Splatting for Autonomous Driving
por: Hess, Georg, et al.
Publicado: (2024)
por: Hess, Georg, et al.
Publicado: (2024)
NeuRAD: Neural Rendering for Autonomous Driving
por: Tonderski, Adam, et al.
Publicado: (2023)
por: Tonderski, Adam, et al.
Publicado: (2023)
R3D-SWIN:Use Shifted Window Attention for Single-View 3D Reconstruction
por: Li, Chenhuan, et al.
Publicado: (2023)
por: Li, Chenhuan, et al.
Publicado: (2023)
Uncertainty quantification in fine-tuned LLMs using LoRA ensembles
por: Balabanov, Oleksandr, et al.
Publicado: (2024)
por: Balabanov, Oleksandr, et al.
Publicado: (2024)
Parabolic Position Encoding: Vision-Centric, Principled, Extrapolatable, General
por: Øhrstrøm, Christoffer Koo, et al.
Publicado: (2026)
por: Øhrstrøm, Christoffer Koo, et al.
Publicado: (2026)
LIB-KD: Teaching Inductive Bias for Efficient Vision Transformer Distillation and Compression
por: Habib, Gousia, et al.
Publicado: (2023)
por: Habib, Gousia, et al.
Publicado: (2023)
Muon in Vision Transformers: Optimizer-Recipe Interactions and Gradient Spectra
por: Southworth, Ben S., et al.
Publicado: (2026)
por: Southworth, Ben S., et al.
Publicado: (2026)
Native Segmentation Vision Transformers
por: Brasó, Guillem, et al.
Publicado: (2025)
por: Brasó, Guillem, et al.
Publicado: (2025)
NeuroNCAP: Photorealistic Closed-loop Safety Testing for Autonomous Driving
por: Ljungbergh, William, et al.
Publicado: (2024)
por: Ljungbergh, William, et al.
Publicado: (2024)
DiffiT: Diffusion Vision Transformers for Image Generation
por: Hatamizadeh, Ali, et al.
Publicado: (2023)
por: Hatamizadeh, Ali, et al.
Publicado: (2023)
ViTNT-FIQA: Training-Free Face Image Quality Assessment with Vision Transformers
por: Ozgur, Guray, et al.
Publicado: (2026)
por: Ozgur, Guray, et al.
Publicado: (2026)
Iwin Transformer: Hierarchical Vision Transformer using Interleaved Windows
por: Huo, Simin, et al.
Publicado: (2025)
por: Huo, Simin, et al.
Publicado: (2025)
Learning Temporal Saliency for Time Series Forecasting with Cross-Scale Attention
por: Delibasoglu, Ibrahim, et al.
Publicado: (2025)
por: Delibasoglu, Ibrahim, et al.
Publicado: (2025)
Slicing Vision Transformer for Flexible Inference
por: Zhang, Yitian, et al.
Publicado: (2024)
por: Zhang, Yitian, et al.
Publicado: (2024)
RAViT: Resolution-Adaptive Vision Transformer
por: Guidez, Martial, et al.
Publicado: (2026)
por: Guidez, Martial, et al.
Publicado: (2026)
Rotary Position Embedding for Vision Transformer
por: Heo, Byeongho, et al.
Publicado: (2024)
por: Heo, Byeongho, et al.
Publicado: (2024)
Decorrelation Speeds Up Vision Transformers
por: Carrigg, Kieran, et al.
Publicado: (2025)
por: Carrigg, Kieran, et al.
Publicado: (2025)
Hybrid Convolution and Vision Transformer NAS Search Space for TinyML Image Classification
por: Djajapermana, Mikhael, et al.
Publicado: (2025)
por: Djajapermana, Mikhael, et al.
Publicado: (2025)
OmniEarth-Bench: Towards Holistic Evaluation of Earth's Six Spheres and Cross-Spheres Interactions with Multimodal Observational Earth Data
por: Wang, Fengxiang, et al.
Publicado: (2025)
por: Wang, Fengxiang, et al.
Publicado: (2025)
MuM: Multi-View Masked Image Modeling for 3D Vision
por: Nordström, David, et al.
Publicado: (2025)
por: Nordström, David, et al.
Publicado: (2025)
Weather-Robust Scene Semantics with Vision-Aligned 4D Radar
por: Hamilton, Kali, et al.
Publicado: (2026)
por: Hamilton, Kali, et al.
Publicado: (2026)
Video-Based Inpatient Fall Risk Assessment: A Case Study
por: Wang, Ziqing, et al.
Publicado: (2021)
por: Wang, Ziqing, et al.
Publicado: (2021)
In-Bed Human Pose Estimation from Unseen and Privacy-Preserving Image Domains
por: Cao, Ting, et al.
Publicado: (2021)
por: Cao, Ting, et al.
Publicado: (2021)
Timeline-based Process Discovery
por: Kaur, Harleen, et al.
Publicado: (2023)
por: Kaur, Harleen, et al.
Publicado: (2023)
Faster-HEAL: An Efficient and Privacy-Preserving Collaborative Perception Framework for Heterogeneous Autonomous Vehicles
por: Maleki, Armin, et al.
Publicado: (2026)
por: Maleki, Armin, et al.
Publicado: (2026)
SPoT: Subpixel Placement of Tokens in Vision Transformers
por: Hjelkrem-Tan, Martine, et al.
Publicado: (2025)
por: Hjelkrem-Tan, Martine, et al.
Publicado: (2025)
Instance-Aware Group Quantization for Vision Transformers
por: Moon, Jaehyeon, et al.
Publicado: (2024)
por: Moon, Jaehyeon, et al.
Publicado: (2024)
Vision Transformer-based Adversarial Domain Adaptation
por: Li, Yahan, et al.
Publicado: (2024)
por: Li, Yahan, et al.
Publicado: (2024)
Elastic Attention Cores for Scalable Vision Transformers
por: Song, Alan Z., et al.
Publicado: (2026)
por: Song, Alan Z., et al.
Publicado: (2026)
Ejemplares similares
-
PEAR: Equal Area Weather Forecasting on the Sphere
por: Linander, Hampus, et al.
Publicado: (2025) -
Bayesian posterior approximation with stochastic ensembles
por: Balabanov, Oleksandr, et al.
Publicado: (2022) -
Raw or Cooked? Object Detection on RAW Images
por: Ljungbergh, William, et al.
Publicado: (2023) -
HEAL-ViT: Vision Transformers on a spherical mesh for medium-range weather forecasting
por: Ramavajjala, Vivek
Publicado: (2024) -
Domain Generalized Recaptured Screen Image Identification Using SWIN Transformer
por: Mehta, Preeti, et al.
Publicado: (2024)