EMBANet: A Flexible Efffcient Multi-branch Attention Network
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zu, Keke, Zhang, Hu, Lu, Jian, Zhang, Lei, Xu, Chen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models
von: Feng, Kailai, et al.
Veröffentlicht: (2024)
von: Feng, Kailai, et al.
Veröffentlicht: (2024)
PriorNet: A Novel Lightweight Network with Multidimensional Interactive Attention for Efficient Image Dehazing
von: Chen, Yutong, et al.
Veröffentlicht: (2024)
von: Chen, Yutong, et al.
Veröffentlicht: (2024)
Multi-Agent Image Restoration
von: Jiang, Xu, et al.
Veröffentlicht: (2025)
von: Jiang, Xu, et al.
Veröffentlicht: (2025)
DynamicID: Zero-Shot Multi-ID Image Personalization with Flexible Facial Editability
von: Hu, Xirui, et al.
Veröffentlicht: (2025)
von: Hu, Xirui, et al.
Veröffentlicht: (2025)
FlexGen: Flexible Multi-View Generation from Text and Image Inputs
von: Xu, Xinli, et al.
Veröffentlicht: (2024)
von: Xu, Xinli, et al.
Veröffentlicht: (2024)
Effective Attention-Guided Multi-Scale Medical Network for Skin Lesion Segmentation
von: Wang, Siyu, et al.
Veröffentlicht: (2025)
von: Wang, Siyu, et al.
Veröffentlicht: (2025)
Cross-Stage Attention Multi-Expert Network for Radiologist-Inspired Breast Ultrasound Diagnosis
von: Zhai, Xinyang, et al.
Veröffentlicht: (2026)
von: Zhai, Xinyang, et al.
Veröffentlicht: (2026)
A Global-Local Cross-Attention Network for Ultra-high Resolution Remote Sensing Image Semantic Segmentation
von: Yi, Chen, et al.
Veröffentlicht: (2025)
von: Yi, Chen, et al.
Veröffentlicht: (2025)
Text-Guided Attention is All You Need for Zero-Shot Robustness in Vision-Language Models
von: Yu, Lu, et al.
Veröffentlicht: (2024)
von: Yu, Lu, et al.
Veröffentlicht: (2024)
CookAnything: A Framework for Flexible and Consistent Multi-Step Recipe Image Generation
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Ruoxuan, et al.
Veröffentlicht: (2025)
TCJA-SNN: Temporal-Channel Joint Attention for Spiking Neural Networks
von: Zhu, Rui-Jie, et al.
Veröffentlicht: (2022)
von: Zhu, Rui-Jie, et al.
Veröffentlicht: (2022)
CMHANet: A Cross-Modal Hybrid Attention Network for Point Cloud Registration
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
A4-Unet: Deformable Multi-Scale Attention Network for Brain Tumor Segmentation
von: Wang, Ruoxin, et al.
Veröffentlicht: (2024)
von: Wang, Ruoxin, et al.
Veröffentlicht: (2024)
Training-Free Representation Guidance for Diffusion Models with a Representation Alignment Projector
von: Zu, Wenqiang, et al.
Veröffentlicht: (2026)
von: Zu, Wenqiang, et al.
Veröffentlicht: (2026)
H-CNN-ViT: A Hierarchical Gated Attention Multi-Branch Model for Bladder Cancer Recurrence Prediction
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
von: Li, Xueyang, et al.
Veröffentlicht: (2025)
Visual Document Understanding and Reasoning: A Multi-Agent Collaboration Framework with Agent-Wise Adaptive Test-Time Scaling
von: Yu, Xinlei, et al.
Veröffentlicht: (2025)
von: Yu, Xinlei, et al.
Veröffentlicht: (2025)
MGHFT: Multi-Granularity Hierarchical Fusion Transformer for Cross-Modal Sticker Emotion Recognition
von: Chen, Jian, et al.
Veröffentlicht: (2025)
von: Chen, Jian, et al.
Veröffentlicht: (2025)
PainterNet: Adaptive Image Inpainting with Actual-Token Attention and Diverse Mask Control
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
von: Wang, Ruichen, et al.
Veröffentlicht: (2024)
HGNet: High-Order Spatial Awareness Hypergraph and Multi-Scale Context Attention Network for Colorectal Polyp Detection
von: Liu, Xiaofang, et al.
Veröffentlicht: (2025)
von: Liu, Xiaofang, et al.
Veröffentlicht: (2025)
Image Forgery Localization via Guided Noise and Multi-Scale Feature Aggregation
von: Niu, Yakun, et al.
Veröffentlicht: (2024)
von: Niu, Yakun, et al.
Veröffentlicht: (2024)
Object-level Cross-view Geo-localization with Location Enhancement and Multi-Head Cross Attention
von: Huang, Zheyang, et al.
Veröffentlicht: (2025)
von: Huang, Zheyang, et al.
Veröffentlicht: (2025)
UniShield: An Adaptive Multi-Agent Framework for Unified Forgery Image Detection and Localization
von: Huang, Qing, et al.
Veröffentlicht: (2025)
von: Huang, Qing, et al.
Veröffentlicht: (2025)
Attention Retention for Continual Learning with Vision Transformers
von: Lu, Yue, et al.
Veröffentlicht: (2026)
von: Lu, Yue, et al.
Veröffentlicht: (2026)
M2DA: Multi-Modal Fusion Transformer Incorporating Driver Attention for Autonomous Driving
von: Xu, Dongyang, et al.
Veröffentlicht: (2024)
von: Xu, Dongyang, et al.
Veröffentlicht: (2024)
Revisiting the Integration of Convolution and Attention for Vision Backbone
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
von: Zhu, Lei, et al.
Veröffentlicht: (2024)
Ultra3D: Efficient and High-Fidelity 3D Generation with Part Attention
von: Chen, Yiwen, et al.
Veröffentlicht: (2025)
von: Chen, Yiwen, et al.
Veröffentlicht: (2025)
A-VL: Adaptive Attention for Large Vision-Language Models
von: Zhang, Junyang, et al.
Veröffentlicht: (2024)
von: Zhang, Junyang, et al.
Veröffentlicht: (2024)
A triple-branch network for latent fingerprint enhancement guided by orientation fields and minutiae
von: Wang, Yurun, et al.
Veröffentlicht: (2025)
von: Wang, Yurun, et al.
Veröffentlicht: (2025)
An Efficient Dual-Line Decoder Network with Multi-Scale Convolutional Attention for Multi-organ Segmentation
von: Hassan, Riad, et al.
Veröffentlicht: (2025)
von: Hassan, Riad, et al.
Veröffentlicht: (2025)
MHLA: Restoring Expressivity of Linear Attention via Token-Level Multi-Head
von: Zhang, Kewei, et al.
Veröffentlicht: (2026)
von: Zhang, Kewei, et al.
Veröffentlicht: (2026)
Dynamic Attention Mechanism in Spatiotemporal Memory Networks for Object Tracking
von: Zhou, Meng, et al.
Veröffentlicht: (2025)
von: Zhou, Meng, et al.
Veröffentlicht: (2025)
HDBN: A Novel Hybrid Dual-branch Network for Robust Skeleton-based Action Recognition
von: Liu, Jinfu, et al.
Veröffentlicht: (2024)
von: Liu, Jinfu, et al.
Veröffentlicht: (2024)
Multi-scale Information Sharing and Selection Network with Boundary Attention for Polyp Segmentation
von: Kang, Xiaolu, et al.
Veröffentlicht: (2024)
von: Kang, Xiaolu, et al.
Veröffentlicht: (2024)
Beyond Attention or Similarity: Maximizing Conditional Diversity for Token Pruning in MLLMs
von: Zhang, Qizhe, et al.
Veröffentlicht: (2025)
von: Zhang, Qizhe, et al.
Veröffentlicht: (2025)
FakeShield: Explainable Image Forgery Detection and Localization via Multi-modal Large Language Models
von: Xu, Zhipei, et al.
Veröffentlicht: (2024)
von: Xu, Zhipei, et al.
Veröffentlicht: (2024)
Endo-TTAP: Robust Endoscopic Tissue Tracking via Multi-Facet Guided Attention and Hybrid Flow-point Supervision
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
von: Zhou, Rulin, et al.
Veröffentlicht: (2025)
IGASA: Integrated Geometry-Aware and Skip-Attention Modules for Enhanced Point Cloud Registration
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
von: Zhang, Dongxu, et al.
Veröffentlicht: (2026)
MSLoRA: Multi-Scale Low-Rank Adaptation via Attention Reweighting
von: Yang, Xu, et al.
Veröffentlicht: (2025)
von: Yang, Xu, et al.
Veröffentlicht: (2025)
PulseMind: A Multi-Modal Medical Model for Real-World Clinical Diagnosis
von: Xu, Jiao, et al.
Veröffentlicht: (2026)
von: Xu, Jiao, et al.
Veröffentlicht: (2026)
A Low-Rank Method for Vision Language Model Hallucination Mitigation in Autonomous Driving
von: Long, Keke, et al.
Veröffentlicht: (2025)
von: Long, Keke, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
VitaGlyph: Vitalizing Artistic Typography with Flexible Dual-branch Diffusion Models
von: Feng, Kailai, et al.
Veröffentlicht: (2024) -
PriorNet: A Novel Lightweight Network with Multidimensional Interactive Attention for Efficient Image Dehazing
von: Chen, Yutong, et al.
Veröffentlicht: (2024) -
Multi-Agent Image Restoration
von: Jiang, Xu, et al.
Veröffentlicht: (2025) -
DynamicID: Zero-Shot Multi-ID Image Personalization with Flexible Facial Editability
von: Hu, Xirui, et al.
Veröffentlicht: (2025) -
FlexGen: Flexible Multi-View Generation from Text and Image Inputs
von: Xu, Xinli, et al.
Veröffentlicht: (2024)