Effects of Different Attention Mechanisms Applied on 3D Models in Video Classification
Fuente:
arXiv
Guardado en:
| Autores principales: | Rasras, Mohammad, Marin, Iuliana, Radu, Serban, Mocanu, Irina |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Evaluating Data Augmentation Techniques for Coffee Leaf Disease Classification
por: Gheorghiu, Adrian, et al.
Publicado: (2024)
por: Gheorghiu, Adrian, et al.
Publicado: (2024)
Online Continual Domain Adaptation for Semantic Image Segmentation Using Internal Representations
por: Stan, Serban, et al.
Publicado: (2024)
por: Stan, Serban, et al.
Publicado: (2024)
Generative 3D Cardiac Shape Modelling for In-Silico Trials
por: Gasparovici, Andrei, et al.
Publicado: (2024)
por: Gasparovici, Andrei, et al.
Publicado: (2024)
Understanding Attention Mechanism in Video Diffusion Models
por: Liu, Bingyan, et al.
Publicado: (2025)
por: Liu, Bingyan, et al.
Publicado: (2025)
EgoCVR: An Egocentric Benchmark for Fine-Grained Composed Video Retrieval
por: Hummel, Thomas, et al.
Publicado: (2024)
por: Hummel, Thomas, et al.
Publicado: (2024)
E2ENet: Dynamic Sparse Feature Fusion for Accurate and Efficient 3D Medical Image Segmentation
por: Wu, Boqian, et al.
Publicado: (2023)
por: Wu, Boqian, et al.
Publicado: (2023)
ImVideoEdit: Image-learning Video Editing via 2D Spatial Difference Attention Blocks
por: Xu, Jiayang, et al.
Publicado: (2026)
por: Xu, Jiayang, et al.
Publicado: (2026)
X-Aligner: Composed Visual Retrieval without the Bells and Whistles
por: Zheng, Yuqian, et al.
Publicado: (2026)
por: Zheng, Yuqian, et al.
Publicado: (2026)
Video Compression Meets Video Generation: Latent Inter-Frame Pruning with Attention Recovery
por: Menn, Dennis, et al.
Publicado: (2026)
por: Menn, Dennis, et al.
Publicado: (2026)
Road Obstacle Video Segmentation
por: Rai, Shyam Nandan, et al.
Publicado: (2025)
por: Rai, Shyam Nandan, et al.
Publicado: (2025)
Efficient Dynamic Attention 3D Convolution for Hyperspectral Image Classification
por: Li, Guandong, et al.
Publicado: (2025)
por: Li, Guandong, et al.
Publicado: (2025)
ConceptVAE: Self-Supervised Fine-Grained Concept Disentanglement from 2D Echocardiographies
por: Ciusdel, Costin F., et al.
Publicado: (2025)
por: Ciusdel, Costin F., et al.
Publicado: (2025)
Enhancing Few-Shot Image Classification through Learnable Multi-Scale Embedding and Attention Mechanisms
por: Askari, Fatemeh, et al.
Publicado: (2024)
por: Askari, Fatemeh, et al.
Publicado: (2024)
Audiovisual Masked Autoencoders
por: Georgescu, Mariana-Iuliana, et al.
Publicado: (2022)
por: Georgescu, Mariana-Iuliana, et al.
Publicado: (2022)
Classification Matters: Improving Video Action Detection with Class-Specific Attention
por: Lee, Jinsung, et al.
Publicado: (2024)
por: Lee, Jinsung, et al.
Publicado: (2024)
Axial-Centric Cross-Plane Attention for 3D Medical Image Classification
por: Park, Doyoung, et al.
Publicado: (2026)
por: Park, Doyoung, et al.
Publicado: (2026)
Brain Tumor Classification using Vision Transformer with Selective Cross-Attention Mechanism and Feature Calibration
por: Khaniki, Mohammad Ali Labbaf, et al.
Publicado: (2024)
por: Khaniki, Mohammad Ali Labbaf, et al.
Publicado: (2024)
CA3D: Convolutional-Attentional 3D Nets for Efficient Video Activity Recognition on the Edge
por: Lagani, Gabriele, et al.
Publicado: (2025)
por: Lagani, Gabriele, et al.
Publicado: (2025)
Deformable Attention Mechanisms Applied to Object Detection, case of Remote Sensing
por: Boutayeb, Anasse, et al.
Publicado: (2025)
por: Boutayeb, Anasse, et al.
Publicado: (2025)
Weight Copy and Low-Rank Adaptation for Few-Shot Distillation of Vision Transformers
por: Grigore, Diana-Nicoleta, et al.
Publicado: (2024)
por: Grigore, Diana-Nicoleta, et al.
Publicado: (2024)
IlluSign: Illustrating Sign Language Videos by Leveraging the Attention Mechanism
por: Bruner, Janna, et al.
Publicado: (2025)
por: Bruner, Janna, et al.
Publicado: (2025)
Dynamic Try-On: Taming Video Virtual Try-on with Dynamic Attention Mechanism
por: Zheng, Jun, et al.
Publicado: (2024)
por: Zheng, Jun, et al.
Publicado: (2024)
A Hierarchical Slice Attention Network for Appendicitis Classification in 3D CT Scans
por: Huang, Chia-Wen, et al.
Publicado: (2025)
por: Huang, Chia-Wen, et al.
Publicado: (2025)
African or European Swallow? Benchmarking Large Vision-Language Models for Fine-Grained Object Classification
por: Geigle, Gregor, et al.
Publicado: (2024)
por: Geigle, Gregor, et al.
Publicado: (2024)
Spatiotemporal Analysis of Forest Machine Operations Using 3D Video Classification
por: Wielgosz, Maciej, et al.
Publicado: (2025)
por: Wielgosz, Maciej, et al.
Publicado: (2025)
Cross-Modal Fusion and Attention Mechanism for Weakly Supervised Video Anomaly Detection
por: Ghadiya, Ayush, et al.
Publicado: (2024)
por: Ghadiya, Ayush, et al.
Publicado: (2024)
Efficient Multi-scale Masked Autoencoders with Hybrid-Attention Mechanism for Breast Lesion Classification
por: Vo, Hung Q., et al.
Publicado: (2025)
por: Vo, Hung Q., et al.
Publicado: (2025)
Benchmarking of Different YOLO Models for CAPTCHAs Detection and Classification
por: Wysocki, Mikołaj, et al.
Publicado: (2025)
por: Wysocki, Mikołaj, et al.
Publicado: (2025)
Content-based Video Retrieval in Traffic Videos using Latent Dirichlet Allocation Topic Model
por: Kianpisheh, Mohammad
Publicado: (2025)
por: Kianpisheh, Mohammad
Publicado: (2025)
Interspatial Attention for Efficient 4D Human Video Generation
por: Shao, Ruizhi, et al.
Publicado: (2025)
por: Shao, Ruizhi, et al.
Publicado: (2025)
CustomVideoX: 3D Reference Attention Driven Dynamic Adaptation for Zero-Shot Customized Video Diffusion Transformers
por: She, D., et al.
Publicado: (2025)
por: She, D., et al.
Publicado: (2025)
Imitating Radiological Scrolling: A Global-Local Attention Model for 3D Chest CT Volumes Multi-Label Anomaly Classification
por: Di Piazza, Theo, et al.
Publicado: (2025)
por: Di Piazza, Theo, et al.
Publicado: (2025)
A 3D mesh convolution-based autoencoder for geometry compression
por: Bregeon, Germain, et al.
Publicado: (2026)
por: Bregeon, Germain, et al.
Publicado: (2026)
Optimizing Violence Detection in Video Classification Accuracy through 3D Convolutional Neural Networks
por: Kavathia, Aarjav, et al.
Publicado: (2024)
por: Kavathia, Aarjav, et al.
Publicado: (2024)
MeshConv3D: Efficient convolution and pooling operators for triangular 3D meshes
por: Bregeon, Germain, et al.
Publicado: (2025)
por: Bregeon, Germain, et al.
Publicado: (2025)
RISurConv: Rotation Invariant Surface Attention-Augmented Convolutions for 3D Point Cloud Classification and Segmentation
por: Zhang, Zhiyuan, et al.
Publicado: (2024)
por: Zhang, Zhiyuan, et al.
Publicado: (2024)
CRAG: Can 3D Generative Models Help 3D Assembly?
por: Jiang, Zeyu, et al.
Publicado: (2026)
por: Jiang, Zeyu, et al.
Publicado: (2026)
V3D: Video Diffusion Models are Effective 3D Generators
por: Chen, Zilong, et al.
Publicado: (2024)
por: Chen, Zilong, et al.
Publicado: (2024)
HAtt-Flow: Hierarchical Attention-Flow Mechanism for Group Activity Scene Graph Generation in Videos
por: Chappa, Naga VS Raviteja, et al.
Publicado: (2023)
por: Chappa, Naga VS Raviteja, et al.
Publicado: (2023)
VoD: Learning Volume of Differences for Video-Based Deepfake Detection
por: Xu, Ying, et al.
Publicado: (2025)
por: Xu, Ying, et al.
Publicado: (2025)
Ejemplares similares
-
Evaluating Data Augmentation Techniques for Coffee Leaf Disease Classification
por: Gheorghiu, Adrian, et al.
Publicado: (2024) -
Online Continual Domain Adaptation for Semantic Image Segmentation Using Internal Representations
por: Stan, Serban, et al.
Publicado: (2024) -
Generative 3D Cardiac Shape Modelling for In-Silico Trials
por: Gasparovici, Andrei, et al.
Publicado: (2024) -
Understanding Attention Mechanism in Video Diffusion Models
por: Liu, Bingyan, et al.
Publicado: (2025) -
EgoCVR: An Egocentric Benchmark for Fine-Grained Composed Video Retrieval
por: Hummel, Thomas, et al.
Publicado: (2024)