Lightweight Structure-Aware Attention for Visual Understanding
Fuente:
arXiv
Saved in:
| Main Authors: | Kwon, Heeseung, Castro, Francisco M., Marin-Jimenez, Manuel J., Guil, Nicolas, Alahari, Karteek |
|---|---|
| Format: | Preprint |
| Published: |
2022
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Exploring High-Order Self-Similarity for Video Understanding
by: Kim, Manjin, et al.
Published: (2026)
by: Kim, Manjin, et al.
Published: (2026)
Source-free Video Domain Adaptation by Learning from Noisy Labels
by: Dasgupta, Avijit, et al.
Published: (2023)
by: Dasgupta, Avijit, et al.
Published: (2023)
Online In-Context Distillation for Low-Resource Vision Language Models
by: Kang, Zhiqi, et al.
Published: (2025)
by: Kang, Zhiqi, et al.
Published: (2025)
Advancing Prompt-Based Methods for Replay-Independent General Continual Learning
by: Kang, Zhiqi, et al.
Published: (2025)
by: Kang, Zhiqi, et al.
Published: (2025)
Unlocking Pre-trained Image Backbones for Semantic Image Synthesis
by: Berrada, Tariq, et al.
Published: (2023)
by: Berrada, Tariq, et al.
Published: (2023)
Evaluating the Label Efficiency of Contrastive Self-Supervised Learning for Multi-Resolution Satellite Imagery
by: Bourcier, Jules, et al.
Published: (2022)
by: Bourcier, Jules, et al.
Published: (2022)
On the Shift Invariance of Max Pooling Feature Maps in Convolutional Neural Networks
by: Leterme, Hubert, et al.
Published: (2022)
by: Leterme, Hubert, et al.
Published: (2022)
From CNNs to Shift-Invariant Twin Models Based on Complex Wavelets
by: Leterme, Hubert, et al.
Published: (2022)
by: Leterme, Hubert, et al.
Published: (2022)
Self-Supervised Pretraining on Satellite Imagery: a Case Study on Label-Efficient Vehicle Detection
by: BOURCIER, Jules, et al.
Published: (2022)
by: BOURCIER, Jules, et al.
Published: (2022)
Entropy Rectifying Guidance for Diffusion and Flow Models
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
Flowception: Temporally Expansive Flow Matching for Video Generation
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
by: Ifriqi, Tariq Berrada, et al.
Published: (2025)
Boosting Latent Diffusion with Perceptual Objectives
by: Berrada, Tariq, et al.
Published: (2024)
by: Berrada, Tariq, et al.
Published: (2024)
FADE: Forecasting for Anomaly Detection on ECG
by: Ruiz-Barroso, Paula, et al.
Published: (2025)
by: Ruiz-Barroso, Paula, et al.
Published: (2025)
Gaze Beyond the Frame: Forecasting Egocentric 3D Visual Span
by: Yun, Heeseung, et al.
Published: (2025)
by: Yun, Heeseung, et al.
Published: (2025)
SCSegamba: Lightweight Structure-Aware Vision Mamba for Crack Segmentation in Structures
by: Liu, Hui, et al.
Published: (2025)
by: Liu, Hui, et al.
Published: (2025)
Attention-Enhanced Lightweight Hourglass Network for Human Pose Estimation
by: Kappan, Marsha Mariya, et al.
Published: (2024)
by: Kappan, Marsha Mariya, et al.
Published: (2024)
Dual Prototype Attention for Unsupervised Video Object Segmentation
by: Cho, Suhwan, et al.
Published: (2022)
by: Cho, Suhwan, et al.
Published: (2022)
LFRA-Net: A Lightweight Focal and Region-Aware Attention Network for Retinal Vessel Segmentatio
by: Mehmood, Mehwish, et al.
Published: (2025)
by: Mehmood, Mehwish, et al.
Published: (2025)
Lightweight Channel Attention for Efficient CNNs
by: Kanaparthi, Prem Babu, et al.
Published: (2026)
by: Kanaparthi, Prem Babu, et al.
Published: (2026)
LSSF-Net: Lightweight Segmentation with Self-Awareness, Spatial Attention, and Focal Modulation
by: Farooq, Hamza, et al.
Published: (2024)
by: Farooq, Hamza, et al.
Published: (2024)
InfoGaussian: Structure-Aware Dynamic Gaussians through Lightweight Information Shaping
by: Zhang, Yunchao, et al.
Published: (2024)
by: Zhang, Yunchao, et al.
Published: (2024)
Lightweight Backbone Networks Only Require Adaptive Lightweight Self-Attention Mechanisms
by: Li, Fengyun, et al.
Published: (2025)
by: Li, Fengyun, et al.
Published: (2025)
On Improved Conditioning Mechanisms and Pre-training Strategies for Diffusion Models
by: Ifriqi, Tariq Berrada, et al.
Published: (2024)
by: Ifriqi, Tariq Berrada, et al.
Published: (2024)
Scale-Aware Pre-Training for Human-Centric Visual Perception: Enabling Lightweight and Generalizable Models
by: Wang, Xuanhan, et al.
Published: (2025)
by: Wang, Xuanhan, et al.
Published: (2025)
Lightweight Spatiotemporal Highway Lane Detection via 3D-ResNet and PINet with ROI-Aware Attention
by: Raja, Sorna Shanmuga, et al.
Published: (2026)
by: Raja, Sorna Shanmuga, et al.
Published: (2026)
AD-DINO: Attention-Dynamic DINO for Distance-Aware Embodied Reference Understanding
by: Guo, Hao, et al.
Published: (2024)
by: Guo, Hao, et al.
Published: (2024)
Large-Scale Text-to-Image Model with Inpainting is a Zero-Shot Subject-Driven Image Generator
by: Shin, Chaehun, et al.
Published: (2024)
by: Shin, Chaehun, et al.
Published: (2024)
Inconsistency-Aware Cross-Attention for Audio-Visual Fusion in Dimensional Emotion Recognition
by: Rajasekhar, G, et al.
Published: (2024)
by: Rajasekhar, G, et al.
Published: (2024)
Lightweight Wasserstein Audio-Visual Model for Unified Speech Enhancement and Separation
by: Park, Jisoo, et al.
Published: (2025)
by: Park, Jisoo, et al.
Published: (2025)
Spherical World-Locking for Audio-Visual Localization in Egocentric Videos
by: Yun, Heeseung, et al.
Published: (2024)
by: Yun, Heeseung, et al.
Published: (2024)
LightAVSeg: Lightweight Audio-Visual Segmentation
by: Zhong, Qing, et al.
Published: (2026)
by: Zhong, Qing, et al.
Published: (2026)
PokeFusion Attention: A Lightweight Cross-Attention Mechanism for Style-Conditioned Image Generation
by: Tang, Jingbang
Published: (2026)
by: Tang, Jingbang
Published: (2026)
Decoding with Structured Awareness: Integrating Directional, Frequency-Spatial, and Structural Attention for Medical Image Segmentation
by: Zhang, Fan, et al.
Published: (2025)
by: Zhang, Fan, et al.
Published: (2025)
MATANet: A Multi-context Attention and Taxonomy-Aware Network for Fine-Grained Underwater Recognition of Marine Species
by: Lee, Donghwan, et al.
Published: (2026)
by: Lee, Donghwan, et al.
Published: (2026)
Head-Aware Visual Cropping: Enhancing Fine-Grained VQA with Attention-Guided Subimage
by: Xie, Junfei, et al.
Published: (2026)
by: Xie, Junfei, et al.
Published: (2026)
Motion-Adaptive Temporal Attention for Lightweight Video Generation with Stable Diffusion
by: Hong, Rui, et al.
Published: (2026)
by: Hong, Rui, et al.
Published: (2026)
Lightweight Vision Transformer with Window and Spatial Attention for Food Image Classification
by: Gao, Xinle, et al.
Published: (2025)
by: Gao, Xinle, et al.
Published: (2025)
Omni-Persona: Systematic Benchmarking and Improving Omnimodal Personalization
by: Oh, Yeongtak, et al.
Published: (2026)
by: Oh, Yeongtak, et al.
Published: (2026)
Multi-Scale Visual Prompting for Lightweight Small-Image Classification
by: Khazem, Salim
Published: (2025)
by: Khazem, Salim
Published: (2025)
ReCap: Lightweight Referential Grounding for Coherent Story Visualization
by: Arora, Aditya, et al.
Published: (2026)
by: Arora, Aditya, et al.
Published: (2026)
Similar Items
-
Exploring High-Order Self-Similarity for Video Understanding
by: Kim, Manjin, et al.
Published: (2026) -
Source-free Video Domain Adaptation by Learning from Noisy Labels
by: Dasgupta, Avijit, et al.
Published: (2023) -
Online In-Context Distillation for Low-Resource Vision Language Models
by: Kang, Zhiqi, et al.
Published: (2025) -
Advancing Prompt-Based Methods for Replay-Independent General Continual Learning
by: Kang, Zhiqi, et al.
Published: (2025) -
Unlocking Pre-trained Image Backbones for Semantic Image Synthesis
by: Berrada, Tariq, et al.
Published: (2023)