ASR: Attention-alike Structural Re-parameterization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhong, Shanshan, Huang, Zhongzhan, Wen, Wushao, Qin, Jinghui, Lin, Liang |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Let's Think Outside the Box: Exploring Leap-of-Thought in Large Language Models with Creative Humor Generation
von: Zhong, Shanshan, et al.
Veröffentlicht: (2023)
von: Zhong, Shanshan, et al.
Veröffentlicht: (2023)
A Generic Shared Attention Mechanism for Various Backbone Neural Networks
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2022)
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2022)
MoExtend: Tuning New Experts for Modality and Task Extension
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024)
CIDER: A Causal Cure for Brand-Obsessed Text-to-Image Models
von: Shen, Fangjian, et al.
Veröffentlicht: (2025)
von: Shen, Fangjian, et al.
Veröffentlicht: (2025)
RVAFM: Re-parameterizing Vertical Attention Fusion Module for Handwritten Paragraph Text Recognition
von: Zheng, Jinhui, et al.
Veröffentlicht: (2025)
von: Zheng, Jinhui, et al.
Veröffentlicht: (2025)
Multimodal Representation-disentangled Information Bottleneck for Multimodal Recommendation
von: Wang, Hui, et al.
Veröffentlicht: (2025)
von: Wang, Hui, et al.
Veröffentlicht: (2025)
Multi-Branch Auxiliary Fusion YOLO with Re-parameterization Heterogeneous Convolutional for accurate object detection
von: Yang, Zhiqiang, et al.
Veröffentlicht: (2024)
von: Yang, Zhiqiang, et al.
Veröffentlicht: (2024)
VMonarch: Efficient Video Diffusion Transformers with Structured Attention
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
von: Liang, Cheng, et al.
Veröffentlicht: (2026)
MIA-Mind: A Multidimensional Interactive Attention Mechanism Based on MindSpore
von: Qin, Zhenkai, et al.
Veröffentlicht: (2025)
von: Qin, Zhenkai, et al.
Veröffentlicht: (2025)
Re-Attentional Controllable Video Diffusion Editing
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2024)
von: Wang, Yuanzhi, et al.
Veröffentlicht: (2024)
CyberHost: Taming Audio-driven Avatar Diffusion Model with Region Codebook Attention
von: Lin, Gaojie, et al.
Veröffentlicht: (2024)
von: Lin, Gaojie, et al.
Veröffentlicht: (2024)
ReCorD: Reasoning and Correcting Diffusion for HOI Generation
von: Jiang-Lin, Jian-Yu, et al.
Veröffentlicht: (2024)
von: Jiang-Lin, Jian-Yu, et al.
Veröffentlicht: (2024)
Chasing Better Deep Image Priors between Over- and Under-parameterization
von: Wu, Qiming, et al.
Veröffentlicht: (2024)
von: Wu, Qiming, et al.
Veröffentlicht: (2024)
Boundary-Driven Table-Filling with Cross-Granularity Contrastive Learning for Aspect Sentiment Triplet Extraction
von: Li, Qingling, et al.
Veröffentlicht: (2025)
von: Li, Qingling, et al.
Veröffentlicht: (2025)
ReLE: A Scalable System and Structured Benchmark for Diagnosing Capability Anisotropy in Chinese LLMs
von: Fang, Rui, et al.
Veröffentlicht: (2026)
von: Fang, Rui, et al.
Veröffentlicht: (2026)
SDIGLM: Leveraging Large Language Models and Multi-Modal Chain of Thought for Structural Damage Identification
von: Zhang, Yunkai, et al.
Veröffentlicht: (2025)
von: Zhang, Yunkai, et al.
Veröffentlicht: (2025)
Visual Attention Never Fades: Selective Progressive Attention ReCalibration for Detailed Image Captioning in Multimodal Large Language Models
von: Jung, Mingi, et al.
Veröffentlicht: (2025)
von: Jung, Mingi, et al.
Veröffentlicht: (2025)
Vision-Enhanced Time Series Forecasting via Latent Diffusion Models
von: Ruan, Weilin, et al.
Veröffentlicht: (2025)
von: Ruan, Weilin, et al.
Veröffentlicht: (2025)
Breaking the SFT Plateau: Multimodal Structured Reinforcement Learning for Chart-to-Code Generation
von: Chen, Lei, et al.
Veröffentlicht: (2025)
von: Chen, Lei, et al.
Veröffentlicht: (2025)
ATA: Bridging Implicit Reasoning with Attention-Guided and Action-Guided Inference for Vision-Language Action Models
von: Yang, Cheng, et al.
Veröffentlicht: (2026)
von: Yang, Cheng, et al.
Veröffentlicht: (2026)
AdvI2I: Adversarial Image Attack on Image-to-Image Diffusion models
von: Zeng, Yaopei, et al.
Veröffentlicht: (2024)
von: Zeng, Yaopei, et al.
Veröffentlicht: (2024)
ReMamber: Referring Image Segmentation with Mamba Twister
von: Yang, Yuhuan, et al.
Veröffentlicht: (2024)
von: Yang, Yuhuan, et al.
Veröffentlicht: (2024)
Probing Routing-Conditional Calibration in Attention-Residual Transformers
von: Liang, Wenhao, et al.
Veröffentlicht: (2026)
von: Liang, Wenhao, et al.
Veröffentlicht: (2026)
Ultra3D: Efficient and High-Fidelity 3D Generation with Part Attention
von: Chen, Yiwen, et al.
Veröffentlicht: (2025)
von: Chen, Yiwen, et al.
Veröffentlicht: (2025)
SOAP: Enhancing Spatio-Temporal Relation and Motion Information Capturing for Few-Shot Action Recognition
von: Huang, Wenbo, et al.
Veröffentlicht: (2024)
von: Huang, Wenbo, et al.
Veröffentlicht: (2024)
Large Vision-Language Models Get Lost in Attention
von: Xi, Gongli, et al.
Veröffentlicht: (2026)
von: Xi, Gongli, et al.
Veröffentlicht: (2026)
Mitigating Object Hallucinations in LVLMs via Attention Imbalance Rectification
von: Sun, Han, et al.
Veröffentlicht: (2026)
von: Sun, Han, et al.
Veröffentlicht: (2026)
Generalizable Object Re-Identification via Visual In-Context Prompting
von: Huang, Zhizhong, et al.
Veröffentlicht: (2025)
von: Huang, Zhizhong, et al.
Veröffentlicht: (2025)
SNP: Structured Neuron-level Pruning to Preserve Attention Scores
von: Shim, Kyunghwan, et al.
Veröffentlicht: (2024)
von: Shim, Kyunghwan, et al.
Veröffentlicht: (2024)
Frequency-Dynamic Attention Modulation for Dense Prediction
von: Chen, Linwei, et al.
Veröffentlicht: (2025)
von: Chen, Linwei, et al.
Veröffentlicht: (2025)
STAR: STacked AutoRegressive Scheme for Unified Multimodal Learning
von: Qin, Jie, et al.
Veröffentlicht: (2025)
von: Qin, Jie, et al.
Veröffentlicht: (2025)
ENA: Efficient N-dimensional Attention
von: Zhong, Yibo
Veröffentlicht: (2025)
von: Zhong, Yibo
Veröffentlicht: (2025)
When Looking Is Not Enough: Visual Attention Structure Reveals Hallucination in MLLMs
von: Cao, Fanpu, et al.
Veröffentlicht: (2026)
von: Cao, Fanpu, et al.
Veröffentlicht: (2026)
RelayFormer: A Unified Local-Global Attention Framework for Scalable Image and Video Manipulation Localization
von: Huang, Wen, et al.
Veröffentlicht: (2025)
von: Huang, Wen, et al.
Veröffentlicht: (2025)
BridgeDrive: Diffusion Bridge Policy for Closed-Loop Trajectory Planning in Autonomous Driving
von: Liu, Shu, et al.
Veröffentlicht: (2025)
von: Liu, Shu, et al.
Veröffentlicht: (2025)
ReFlow: Self-correction Motion Learning for Dynamic Scene Reconstruction
von: Liang, Yanzhe, et al.
Veröffentlicht: (2026)
von: Liang, Yanzhe, et al.
Veröffentlicht: (2026)
Attention Retention for Continual Learning with Vision Transformers
von: Lu, Yue, et al.
Veröffentlicht: (2026)
von: Lu, Yue, et al.
Veröffentlicht: (2026)
ReGenNet: Towards Human Action-Reaction Synthesis
von: Xu, Liang, et al.
Veröffentlicht: (2024)
von: Xu, Liang, et al.
Veröffentlicht: (2024)
Camera-Invariant Meta-Learning Network for Single-Camera-Training Person Re-identification
von: Pei, Jiangbo, et al.
Veröffentlicht: (2024)
von: Pei, Jiangbo, et al.
Veröffentlicht: (2024)
TdAttenMix: Top-Down Attention Guided Mixup
von: Wang, Zhiming, et al.
Veröffentlicht: (2025)
von: Wang, Zhiming, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Let's Think Outside the Box: Exploring Leap-of-Thought in Large Language Models with Creative Humor Generation
von: Zhong, Shanshan, et al.
Veröffentlicht: (2023) -
A Generic Shared Attention Mechanism for Various Backbone Neural Networks
von: Huang, Zhongzhan, et al.
Veröffentlicht: (2022) -
MoExtend: Tuning New Experts for Modality and Task Extension
von: Zhong, Shanshan, et al.
Veröffentlicht: (2024) -
CIDER: A Causal Cure for Brand-Obsessed Text-to-Image Models
von: Shen, Fangjian, et al.
Veröffentlicht: (2025) -
RVAFM: Re-parameterizing Vertical Attention Fusion Module for Handwritten Paragraph Text Recognition
von: Zheng, Jinhui, et al.
Veröffentlicht: (2025)