Salvato in:
| Autori principali: | Jiang, Liangyan, Zhu, Chuang, Chen, Yanxu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2407.15708 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
MV-Swin-T: Mammogram Classification with Multi-view Swin Transformer
di: Sarker, Sushmita, et al.
Pubblicazione: (2024)
di: Sarker, Sushmita, et al.
Pubblicazione: (2024)
Swin-TUNA : A Novel PEFT Approach for Accurate Food Image Segmentation
di: Chen, Haotian, et al.
Pubblicazione: (2025)
di: Chen, Haotian, et al.
Pubblicazione: (2025)
SceneGen: Single-Image 3D Scene Generation in One Feedforward Pass
di: Meng, Yanxu, et al.
Pubblicazione: (2025)
di: Meng, Yanxu, et al.
Pubblicazione: (2025)
USP-Gaussian: Unifying Spike-based Image Reconstruction, Pose Correction and Gaussian Splatting
di: Chen, Kang, et al.
Pubblicazione: (2024)
di: Chen, Kang, et al.
Pubblicazione: (2024)
Yuan-TecSwin: A text conditioned Diffusion model with Swin-transformer blocks
di: Wu, Shaohua, et al.
Pubblicazione: (2025)
di: Wu, Shaohua, et al.
Pubblicazione: (2025)
SwinIFS: Landmark Guided Swin Transformer For Identity Preserving Face Super Resolution
di: Kausar, Habiba, et al.
Pubblicazione: (2026)
di: Kausar, Habiba, et al.
Pubblicazione: (2026)
Barlow-Swin: Toward a novel siamese-based segmentation architecture using Swin-Transformers
di: Haftlang, Morteza Kiani, et al.
Pubblicazione: (2025)
di: Haftlang, Morteza Kiani, et al.
Pubblicazione: (2025)
STSA: Spatial-Temporal Semantic Alignment for Visual Dubbing
di: Ding, Zijun, et al.
Pubblicazione: (2025)
di: Ding, Zijun, et al.
Pubblicazione: (2025)
Text Embedded Swin-UMamba for DeepLesion Segmentation
di: Cheng, Ruida, et al.
Pubblicazione: (2025)
di: Cheng, Ruida, et al.
Pubblicazione: (2025)
TCJA-SNN: Temporal-Channel Joint Attention for Spiking Neural Networks
di: Zhu, Rui-Jie, et al.
Pubblicazione: (2022)
di: Zhu, Rui-Jie, et al.
Pubblicazione: (2022)
SparseSwin: Swin Transformer with Sparse Transformer Block
di: Pinasthika, Krisna, et al.
Pubblicazione: (2023)
di: Pinasthika, Krisna, et al.
Pubblicazione: (2023)
Spatial-Temporal Deep Embedding for Vehicle Trajectory Reconstruction from High-Angle Video
di: D., Tianya T. Zhang Ph., et al.
Pubblicazione: (2022)
di: D., Tianya T. Zhang Ph., et al.
Pubblicazione: (2022)
SpikeNVS: Enhancing Novel View Synthesis from Blurry Images via Spike Camera
di: Dai, Gaole, et al.
Pubblicazione: (2024)
di: Dai, Gaole, et al.
Pubblicazione: (2024)
3D Nephrographic Image Synthesis in CT Urography with the Diffusion Model and Swin Transformer
di: Yu, Hongkun, et al.
Pubblicazione: (2025)
di: Yu, Hongkun, et al.
Pubblicazione: (2025)
GCA-SUNet: A Gated Context-Aware Swin-UNet for Exemplar-Free Counting
di: Wu, Yuzhe, et al.
Pubblicazione: (2024)
di: Wu, Yuzhe, et al.
Pubblicazione: (2024)
WinT3R: Window-Based Streaming Reconstruction with Camera Token Pool
di: Li, Zizun, et al.
Pubblicazione: (2025)
di: Li, Zizun, et al.
Pubblicazione: (2025)
Temporal Reversal Regularization for Spiking Neural Networks: Hybrid Spatio-Temporal Invariance for Generalization
di: Zuo, Lin, et al.
Pubblicazione: (2024)
di: Zuo, Lin, et al.
Pubblicazione: (2024)
RS-FME-SwinT: A Novel Feature Map Enhancement Framework Integrating Customized SwinT with Residual and Spatial CNN for Monkeypox Diagnosis
di: Khan, Saddam Hussain, et al.
Pubblicazione: (2024)
di: Khan, Saddam Hussain, et al.
Pubblicazione: (2024)
SpikeVAEDiff: Neural Spike-based Natural Visual Scene Reconstruction via VD-VAE and Versatile Diffusion
di: Li, Jialu, et al.
Pubblicazione: (2026)
di: Li, Jialu, et al.
Pubblicazione: (2026)
RSwinV2-MD: An Enhanced Residual SwinV2 Transformer for Monkeypox Detection from Skin Images
di: Iqbal, Rashid, et al.
Pubblicazione: (2026)
di: Iqbal, Rashid, et al.
Pubblicazione: (2026)
Temporal-adaptive Weight Quantization for Spiking Neural Networks
di: Zhang, Han, et al.
Pubblicazione: (2025)
di: Zhang, Han, et al.
Pubblicazione: (2025)
Point3R: Streaming 3D Reconstruction with Explicit Spatial Pointer Memory
di: Wu, Yuqi, et al.
Pubblicazione: (2025)
di: Wu, Yuqi, et al.
Pubblicazione: (2025)
SweetTok: Semantic-Aware Spatial-Temporal Tokenizer for Compact Video Discretization
di: Tan, Zhentao, et al.
Pubblicazione: (2024)
di: Tan, Zhentao, et al.
Pubblicazione: (2024)
SF-Mamba: Rethinking State Space Model for Vision
di: Yoshimura, Masakazu, et al.
Pubblicazione: (2026)
di: Yoshimura, Masakazu, et al.
Pubblicazione: (2026)
uSF: Learning Neural Semantic Field with Uncertainty
di: Skorokhodov, Vsevolod, et al.
Pubblicazione: (2023)
di: Skorokhodov, Vsevolod, et al.
Pubblicazione: (2023)
Leveraging Swin Transformer for Local-to-Global Weakly Supervised Semantic Segmentation
di: Ahmadi, Rozhan, et al.
Pubblicazione: (2024)
di: Ahmadi, Rozhan, et al.
Pubblicazione: (2024)
DST-Net: A Dual-Stream Transformer with Illumination-Independent Feature Guidance and Multi-Scale Spatial Convolution for Low-Light Image Enhancement
di: Shi, Yicui, et al.
Pubblicazione: (2026)
di: Shi, Yicui, et al.
Pubblicazione: (2026)
Noisy Label Processing for Classification: A Survey
di: Li, Mengting, et al.
Pubblicazione: (2024)
di: Li, Mengting, et al.
Pubblicazione: (2024)
CheX-DS: Improving Chest X-ray Image Classification with Ensemble Learning Based on DenseNet and Swin Transformer
di: Li, Xinran, et al.
Pubblicazione: (2025)
di: Li, Xinran, et al.
Pubblicazione: (2025)
SwinTF3D: A Lightweight Multimodal Fusion Approach for Text-Guided 3D Medical Image Segmentation
di: Khan, Hasan Faraz, et al.
Pubblicazione: (2025)
di: Khan, Hasan Faraz, et al.
Pubblicazione: (2025)
Audio-Sync Video Generation with Multi-Stream Temporal Control
di: Weng, Shuchen, et al.
Pubblicazione: (2025)
di: Weng, Shuchen, et al.
Pubblicazione: (2025)
StreamGaze: Gaze-Guided Temporal Reasoning and Proactive Understanding in Streaming Videos
di: Lee, Daeun, et al.
Pubblicazione: (2025)
di: Lee, Daeun, et al.
Pubblicazione: (2025)
Select2Col: Leveraging Spatial-Temporal Importance of Semantic Information for Efficient Collaborative Perception
di: Liu, Yuntao, et al.
Pubblicazione: (2023)
di: Liu, Yuntao, et al.
Pubblicazione: (2023)
Unified Spatial-Temporal Edge-Enhanced Graph Networks for Pedestrian Trajectory Prediction
di: Li, Ruochen, et al.
Pubblicazione: (2025)
di: Li, Ruochen, et al.
Pubblicazione: (2025)
SatSwinMAE: Efficient Autoencoding for Multiscale Time-series Satellite Imagery
di: Nakayama, Yohei, et al.
Pubblicazione: (2024)
di: Nakayama, Yohei, et al.
Pubblicazione: (2024)
DualSwinUnet++: An Enhanced Swin-Unet Architecture With Dual Decoders For PTMC Segmentation
di: Dialameh, Maryam, et al.
Pubblicazione: (2024)
di: Dialameh, Maryam, et al.
Pubblicazione: (2024)
7DGS: Unified Spatial-Temporal-Angular Gaussian Splatting
di: Gao, Zhongpai, et al.
Pubblicazione: (2025)
di: Gao, Zhongpai, et al.
Pubblicazione: (2025)
REL-SF4PASS: Panoramic Semantic Segmentation with REL Depth Representation and Spherical Fusion
di: Li, Xuewei, et al.
Pubblicazione: (2026)
di: Li, Xuewei, et al.
Pubblicazione: (2026)
Spatial Transcriptomics as Images for Large-Scale Pretraining
di: Zhu, Yishun, et al.
Pubblicazione: (2026)
di: Zhu, Yishun, et al.
Pubblicazione: (2026)
HieraTok: Multi-Scale Visual Tokenizer Improves Image Reconstruction and Generation
di: Chen, Cong, et al.
Pubblicazione: (2025)
di: Chen, Cong, et al.
Pubblicazione: (2025)
Documenti analoghi
-
MV-Swin-T: Mammogram Classification with Multi-view Swin Transformer
di: Sarker, Sushmita, et al.
Pubblicazione: (2024) -
Swin-TUNA : A Novel PEFT Approach for Accurate Food Image Segmentation
di: Chen, Haotian, et al.
Pubblicazione: (2025) -
SceneGen: Single-Image 3D Scene Generation in One Feedforward Pass
di: Meng, Yanxu, et al.
Pubblicazione: (2025) -
USP-Gaussian: Unifying Spike-based Image Reconstruction, Pose Correction and Gaussian Splatting
di: Chen, Kang, et al.
Pubblicazione: (2024) -
Yuan-TecSwin: A text conditioned Diffusion model with Swin-transformer blocks
di: Wu, Shaohua, et al.
Pubblicazione: (2025)