Feature boosting with efficient attention for scene parsing
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Singh, Vivek, Sharma, Shailza, Cuzzolin, Fabio |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Structured prototype regularization for synthetic-to-real driving scene parsing
von: Fan, Jiahe, et al.
Veröffentlicht: (2026)
von: Fan, Jiahe, et al.
Veröffentlicht: (2026)
3D scene generation from scene graphs and self-attention
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2024)
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2024)
Epistemic Generative Adversarial Networks
von: Mubashar, Muhammad, et al.
Veröffentlicht: (2026)
von: Mubashar, Muhammad, et al.
Veröffentlicht: (2026)
Functionality understanding and segmentation in 3D scenes
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
A neurosymbolic Approach with Epistemic Deep Learning for Hierarchical Image Classification
von: Kilicdere, Ezel, et al.
Veröffentlicht: (2026)
von: Kilicdere, Ezel, et al.
Veröffentlicht: (2026)
BiomedParse: a biomedical foundation model for image parsing of everything everywhere all at once
von: Zhao, Theodore, et al.
Veröffentlicht: (2024)
von: Zhao, Theodore, et al.
Veröffentlicht: (2024)
ROAD-Waymo: Action Awareness at Scale for Autonomous Driving
von: Khan, Salman, et al.
Veröffentlicht: (2024)
von: Khan, Salman, et al.
Veröffentlicht: (2024)
Implicit 3D scene reconstruction using deep learning towards efficient collision understanding in autonomous driving
von: Ramanayake, Akarshani, et al.
Veröffentlicht: (2025)
von: Ramanayake, Akarshani, et al.
Veröffentlicht: (2025)
A Hybrid Transformer-Sequencer approach for Age and Gender classification from in-wild facial images
von: Singh, Aakash, et al.
Veröffentlicht: (2024)
von: Singh, Aakash, et al.
Veröffentlicht: (2024)
ASTRA: A Scene-aware TRAnsformer-based model for trajectory prediction
von: Teeti, Izzeddin, et al.
Veröffentlicht: (2025)
von: Teeti, Izzeddin, et al.
Veröffentlicht: (2025)
A Novel Vision Transformer with Residual in Self-attention for Biomedical Image Classification
von: Sharma, Arun K., et al.
Veröffentlicht: (2023)
von: Sharma, Arun K., et al.
Veröffentlicht: (2023)
FALFormer: Feature-aware Landmarks self-attention for Whole-slide Image Classification
von: Bui, Doanh C., et al.
Veröffentlicht: (2024)
von: Bui, Doanh C., et al.
Veröffentlicht: (2024)
A transition towards virtual representations of visual scenes
von: Pereira, Américo, et al.
Veröffentlicht: (2024)
von: Pereira, Américo, et al.
Veröffentlicht: (2024)
Swift4D:Adaptive divide-and-conquer Gaussian Splatting for compact and efficient reconstruction of dynamic scene
von: Wu, Jiahao, et al.
Veröffentlicht: (2025)
von: Wu, Jiahao, et al.
Veröffentlicht: (2025)
CrowdFormer: Weakly-supervised Crowd counting with Improved Generalizability
von: Savner, Siddharth Singh, et al.
Veröffentlicht: (2022)
von: Savner, Siddharth Singh, et al.
Veröffentlicht: (2022)
Indoor scene recognition from images under visual corruptions
von: Costa, Willams de Lima, et al.
Veröffentlicht: (2024)
von: Costa, Willams de Lima, et al.
Veröffentlicht: (2024)
A Versatile Framework for Multi-scene Person Re-identification
von: Zheng, Wei-Shi, et al.
Veröffentlicht: (2024)
von: Zheng, Wei-Shi, et al.
Veröffentlicht: (2024)
Efficient scene text image super-resolution with semantic guidance
von: TomyEnrique, LeoWu, et al.
Veröffentlicht: (2024)
von: TomyEnrique, LeoWu, et al.
Veröffentlicht: (2024)
Semantic UV mapping to improve texture inpainting for indoor scenes
von: Vermandere, Jelle, et al.
Veröffentlicht: (2024)
von: Vermandere, Jelle, et al.
Veröffentlicht: (2024)
Consistent text-to-image generation via scene de-contextualization
von: Tang, Song, et al.
Veröffentlicht: (2025)
von: Tang, Song, et al.
Veröffentlicht: (2025)
Open-vocabulary 3D scene perception in industrial environments
von: Moenck, Keno, et al.
Veröffentlicht: (2026)
von: Moenck, Keno, et al.
Veröffentlicht: (2026)
Teaching in adverse scenes: a statistically feedback-driven threshold and mask adjustment teacher-student framework for object detection in UAV images under adverse scenes
von: Chen, Hongyu, et al.
Veröffentlicht: (2025)
von: Chen, Hongyu, et al.
Veröffentlicht: (2025)
PMFSNet: Polarized Multi-scale Feature Self-attention Network For Lightweight Medical Image Segmentation
von: Zhong, Jiahui, et al.
Veröffentlicht: (2024)
von: Zhong, Jiahui, et al.
Veröffentlicht: (2024)
Local positional graphs and attentive local features for a data and runtime-efficient hierarchical place recognition pipeline
von: Yuan, Fangming, et al.
Veröffentlicht: (2024)
von: Yuan, Fangming, et al.
Veröffentlicht: (2024)
MessyKitchens: Contact-rich object-level 3D scene reconstruction
von: Ansari, Junaid Ahmed, et al.
Veröffentlicht: (2026)
von: Ansari, Junaid Ahmed, et al.
Veröffentlicht: (2026)
VWise: A novel benchmark for evaluating scene classification for vehicular applications
von: Azevedo, Pedro, et al.
Veröffentlicht: (2024)
von: Azevedo, Pedro, et al.
Veröffentlicht: (2024)
Motion-guided small MAV detection in complex and non-planar scenes
von: Guo, Hanqing, et al.
Veröffentlicht: (2024)
von: Guo, Hanqing, et al.
Veröffentlicht: (2024)
Visual enhancement and 3D representation for underwater scenes: a review
von: Huang, Guoxi, et al.
Veröffentlicht: (2025)
von: Huang, Guoxi, et al.
Veröffentlicht: (2025)
GazeLT: Visual attention-guided long-tailed disease classification in chest radiographs
von: Bhattacharya, Moinak, et al.
Veröffentlicht: (2025)
von: Bhattacharya, Moinak, et al.
Veröffentlicht: (2025)
ReCoGS: Real-time ReColoring for Gaussian Splatting scenes
von: Rutayisire, Lorenzo, et al.
Veröffentlicht: (2025)
von: Rutayisire, Lorenzo, et al.
Veröffentlicht: (2025)
Generating metamers of human scene understanding
von: Raina, Ritik, et al.
Veröffentlicht: (2026)
von: Raina, Ritik, et al.
Veröffentlicht: (2026)
VistaDream: Sampling multiview consistent images for single-view scene reconstruction
von: Wang, Haiping, et al.
Veröffentlicht: (2024)
von: Wang, Haiping, et al.
Veröffentlicht: (2024)
RobustSurg: Tackling domain generalisation for out-of-distribution surgical scene segmentation
von: Ali, Mansoor, et al.
Veröffentlicht: (2025)
von: Ali, Mansoor, et al.
Veröffentlicht: (2025)
VIZOR: Viewpoint-Invariant Zero-Shot Scene Graph Generation for 3D Scene Reasoning
von: Madhavaram, Vivek, et al.
Veröffentlicht: (2026)
von: Madhavaram, Vivek, et al.
Veröffentlicht: (2026)
Deep learning-based approach for tomato classification in complex scenes
von: Mousse, Mikael A., et al.
Veröffentlicht: (2024)
von: Mousse, Mikael A., et al.
Veröffentlicht: (2024)
Uncertainty-boosted Robust Video Activity Anticipation
von: Qi, Zhaobo, et al.
Veröffentlicht: (2024)
von: Qi, Zhaobo, et al.
Veröffentlicht: (2024)
DEF-oriCORN: efficient 3D scene understanding for robust language-directed manipulation without demonstrations
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
von: Son, Dongwon, et al.
Veröffentlicht: (2024)
SituationalLLM: Proactive language models with scene awareness for dynamic, contextual task guidance
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
von: Khan, Muhammad Saif Ullah, et al.
Veröffentlicht: (2024)
Enhancing the quality of gauge images captured in smoke and haze scenes through deep learning
von: Ramírez-Agudelo, Oscar H., et al.
Veröffentlicht: (2026)
von: Ramírez-Agudelo, Oscar H., et al.
Veröffentlicht: (2026)
3DOF+Quantization: 3DGS quantization for large scenes with limited Degrees of Freedom
von: Gendrin, Matthieu, et al.
Veröffentlicht: (2025)
von: Gendrin, Matthieu, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Structured prototype regularization for synthetic-to-real driving scene parsing
von: Fan, Jiahe, et al.
Veröffentlicht: (2026) -
3D scene generation from scene graphs and self-attention
von: Bonazzi, Pietro, et al.
Veröffentlicht: (2024) -
Epistemic Generative Adversarial Networks
von: Mubashar, Muhammad, et al.
Veröffentlicht: (2026) -
Functionality understanding and segmentation in 3D scenes
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024) -
A neurosymbolic Approach with Epistemic Deep Learning for Hierarchical Image Classification
von: Kilicdere, Ezel, et al.
Veröffentlicht: (2026)