LoLA-SpecViT: Local Attention SwiGLU Vision Transformer with LoRA for Hyperspectral Imaging
Fuente:
arXiv
Guardado en:
| Autores principales: | Zidi, Fadi Abdeladhim, Boukhari, Djamel Eddine, Sellam, Abdellah Zakaria, Ouafi, Abdelkrim, Distante, Cosimo, Bekhouche, Salah Eddine, Taleb-Ahmed, Abdelmalik |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Beyond Linear Bottlenecks: Spline-Based Knowledge Distillation for Culturally Diverse Art Style Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
Advancing Wheat Crop Analysis: A Survey of Deep Learning Approaches Using Hyperspectral Imaging
por: Zidi, Fadi Abdeladhim, et al.
Publicado: (2025)
por: Zidi, Fadi Abdeladhim, et al.
Publicado: (2025)
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
C-DiffDet+: Fusing Global Scene Context with Generative Denoising for High-Fidelity Car Damage Detection
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
RF-HiT: Rectified Flow Hierarchical Transformer for General Medical Image Segmentation
por: Djouama, Ahmed Marouane, et al.
Publicado: (2026)
por: Djouama, Ahmed Marouane, et al.
Publicado: (2026)
Boosting House Price Estimations with Multi-Head Gated Attention
por: Sellam, Zakaria Abdellah, et al.
Publicado: (2024)
por: Sellam, Zakaria Abdellah, et al.
Publicado: (2024)
Mamba Adaptive Anomaly Transformer with association discrepancy for time series
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025)
SynergyNet: Fusing Generative Priors and State-Space Models for Facial Beauty Prediction
por: Boukhari, Djamel Eddine
Publicado: (2025)
por: Boukhari, Djamel Eddine
Publicado: (2025)
FairViT-GAN: A Hybrid Vision Transformer with Adversarial Debiasing for Fair and Explainable Facial Beauty Prediction
por: Boukhari, Djamel Eddine
Publicado: (2025)
por: Boukhari, Djamel Eddine
Publicado: (2025)
Mamba-CNN: A Hybrid Architecture for Efficient and Accurate Facial Beauty Prediction
por: Boukhari, Djamel Eddine
Publicado: (2025)
por: Boukhari, Djamel Eddine
Publicado: (2025)
Scale-interaction transformer: a hybrid cnn-transformer model for facial beauty prediction
por: Boukhari, Djamel Eddine
Publicado: (2025)
por: Boukhari, Djamel Eddine
Publicado: (2025)
VM-BeautyNet: A Synergistic Ensemble of Vision Transformer and Mamba for Facial Beauty Prediction
por: Boukhari, Djamel Eddine
Publicado: (2025)
por: Boukhari, Djamel Eddine
Publicado: (2025)
Conflict-Aware Multimodal Fusion for Ambivalence and Hesitancy Recognition
por: Bekhouche, Salah Eddine, et al.
Publicado: (2026)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2026)
Integrating ConvNeXt and Vision Transformers for Enhancing Facial Age Estimation
por: Maroun, Gaby, et al.
Publicado: (2025)
por: Maroun, Gaby, et al.
Publicado: (2025)
Generative Pre-training for Subjective Tasks: A Diffusion Transformer-Based Framework for Facial Beauty Prediction
por: Boukhari, Djamel Eddine, et al.
Publicado: (2025)
por: Boukhari, Djamel Eddine, et al.
Publicado: (2025)
D-TrAttUnet: Toward Hybrid CNN-Transformer Architecture for Generic and Subtle Segmentation in Medical Images
por: Bougourzi, Fares, et al.
Publicado: (2024)
por: Bougourzi, Fares, et al.
Publicado: (2024)
AudioMAE++: learning better masked audio representations with SwiGLU FFNs
por: Yadav, Sarthak, et al.
Publicado: (2025)
por: Yadav, Sarthak, et al.
Publicado: (2025)
CVPD at QIAS 2026: RAG-Guided LLM Reasoning for Al-Mawarith Share Computation and Heir Allocation
por: Swaileh, Wassim, et al.
Publicado: (2026)
por: Swaileh, Wassim, et al.
Publicado: (2026)
SegDT: A Diffusion Transformer-Based Segmentation Model for Medical Imaging
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
Dual Path Attribution: Efficient Attribution for SwiGLU-Transformers through Layer-Wise Target Propagation
por: Jantsch, Lasse Marten, et al.
Publicado: (2026)
por: Jantsch, Lasse Marten, et al.
Publicado: (2026)
Depth Registers Unlock W4A4 on SwiGLU: A Reader/Generator Decomposition
por: Liu, Ziyang
Publicado: (2026)
por: Liu, Ziyang
Publicado: (2026)
Symmetry-Compatible Principle for Optimizer Design: Embeddings, LM Heads, SwiGLU MLPs, and MoE Routers
por: Lau, Tim Tsz-Kit, et al.
Publicado: (2026)
por: Lau, Tim Tsz-Kit, et al.
Publicado: (2026)
SPARK-IL: Spectral Retrieval-Augmented RAG for Knowledge-driven Deepfake Detection via Incremental Learning
por: Eutamene, Hessen Bougueffa, et al.
Publicado: (2026)
por: Eutamene, Hessen Bougueffa, et al.
Publicado: (2026)
Enhanced Arabic Text Retrieval with Attentive Relevance Scoring
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)
Brecha digital y per¿les de uso de las tic en México: Un estudio exploratorio con microdatos
por: Djamel Eddine Toudert
Publicado: (2015)
por: Djamel Eddine Toudert
Publicado: (2015)
Aprovechamiento de las TIC en México: una aproximación empírica mediante el uso de microdatos y la aplicación de la modelación PLS
por: Djamel Eddine Toudert
Publicado: (2014)
por: Djamel Eddine Toudert
Publicado: (2014)
Boosting Hyperspectral Image Classification with Gate-Shift-Fuse Mechanisms in a Novel CNN-Transformer Approach
por: Guerri, Mohamed Fadhlallah, et al.
Publicado: (2024)
por: Guerri, Mohamed Fadhlallah, et al.
Publicado: (2024)
Cross-Modal Mapping and Dual-Branch Reconstruction for 2D-3D Multimodal Industrial Anomaly Detection
por: Daci, Radia, et al.
Publicado: (2026)
por: Daci, Radia, et al.
Publicado: (2026)
CineInfini: Adaptive Multi-Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
CineInfini: Adaptive Multi‑Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
CineInfini: Adaptive Multi‑Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
CineInfini: Adaptive Multi-Stage Video Quality Audit Pipeline
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
por: BENBRAHIM, Salah-Eddine
Publicado: (2026)
LoLA: Low-Rank Linear Attention With Sparse Caching
por: McDermott, Luke, et al.
Publicado: (2025)
por: McDermott, Luke, et al.
Publicado: (2025)
Thermodynamic properties of an ideal Quark-Gluon plasma under quantum gravitational effects
por: Zenkhri, Djamel Eddine, et al.
Publicado: (2025)
por: Zenkhri, Djamel Eddine, et al.
Publicado: (2025)
Ideal Fermi gas in the Dunkl formalism
por: Zenkhri, Djamel Eddine, et al.
Publicado: (2025)
por: Zenkhri, Djamel Eddine, et al.
Publicado: (2025)
The surjection property and computable type
por: Amir, Djamel Eddine, et al.
Publicado: (2023)
por: Amir, Djamel Eddine, et al.
Publicado: (2023)
LoLA: Long Horizon Latent Action Learning for General Robot Manipulation
por: Wang, Xiaofan, et al.
Publicado: (2025)
por: Wang, Xiaofan, et al.
Publicado: (2025)
DATASHI: A Parallel English-Tashlhiyt Corpus for Orthography Normalization and Low-Resource Language Processing
por: Monir, Nasser-Eddine, et al.
Publicado: (2026)
por: Monir, Nasser-Eddine, et al.
Publicado: (2026)
Ejemplares similares
-
Beyond Linear Bottlenecks: Spline-Based Knowledge Distillation for Culturally Diverse Art Style Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025) -
Advancing Wheat Crop Analysis: A Survey of Deep Learning Approaches Using Hyperspectral Imaging
por: Zidi, Fadi Abdeladhim, et al.
Publicado: (2025) -
VP-Hype: A Hybrid Mamba-Transformer Framework with Visual-Textual Prompting for Hyperspectral Image Classification
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2026) -
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
por: Sellam, Abdellah Zakaria, et al.
Publicado: (2025) -
CVPD at QIAS 2025 Shared Task: An Efficient Encoder-Based Approach for Islamic Inheritance Reasoning
por: Bekhouche, Salah Eddine, et al.
Publicado: (2025)