Rectifying Magnitude Neglect in Linear Attention
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Fan, Qihang, Huang, Huaibo, Ai, Yuang, He, Ran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Breaking Complexity Barriers: High-Resolution Image Restoration with Rank Enhanced Linear Attention
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
Breaking the Low-Rank Dilemma of Linear Attention
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
Random Wins All: Rethinking Grouping Strategies for Vision Tokens
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
von: Ai, Yuang, et al.
Veröffentlicht: (2025)
Expand and Prune: Maximizing Trajectory Diversity for Effective GRPO in Generative Models
von: Ge, Shiran, et al.
Veröffentlicht: (2025)
von: Ge, Shiran, et al.
Veröffentlicht: (2025)
Semantic Equitable Clustering: A Simple and Effective Strategy for Clustering Vision Tokens
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
Lightweight Vision Transformer with Bidirectional Interaction
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
Multimodal Prompt Perceiver: Empower Adaptiveness, Generalizability and Fidelity for All-in-One Image Restoration
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
Uncertainty-Aware Source-Free Adaptive Image Super-Resolution with Wavelet Augmentation Transformer
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
von: Ai, Yuang, et al.
Veröffentlicht: (2023)
RMT: Retentive Networks Meet Vision Transformers
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
von: Fan, Qihang, et al.
Veröffentlicht: (2023)
Advancing Vision Transformer with Enhanced Spatial Priors
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
von: Fan, Qihang, et al.
Veröffentlicht: (2026)
Vision Transformer with Sparse Scan Prior
von: Zhang, Yuguang, et al.
Veröffentlicht: (2024)
von: Zhang, Yuguang, et al.
Veröffentlicht: (2024)
InfiMM-WebMath-40B: Advancing Multimodal Pre-Training for Enhanced Mathematical Reasoning
von: Han, Xiaotian, et al.
Veröffentlicht: (2024)
von: Han, Xiaotian, et al.
Veröffentlicht: (2024)
ViTAR: Vision Transformer with Any Resolution
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
von: Fan, Qihang, et al.
Veröffentlicht: (2024)
DreamClear: High-Capacity Real-World Image Restoration with Privacy-Safe Dataset Curation
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
von: Ai, Yuang, et al.
Veröffentlicht: (2024)
NOFT: Test-Time Noise Finetune via Information Bottleneck for Highly Correlated Asset Creation
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
ZePo: Zero-Shot Portrait Stylization with Faster Sampling
von: Liu, Jin, et al.
Veröffentlicht: (2024)
von: Liu, Jin, et al.
Veröffentlicht: (2024)
DeVAn: Dense Video Annotation for Video-Language Models
von: Liu, Tingkai, et al.
Veröffentlicht: (2023)
von: Liu, Tingkai, et al.
Veröffentlicht: (2023)
Marmot: Object-Level Self-Correction via Multi-Agent Reasoning
von: Sun, Jiayang, et al.
Veröffentlicht: (2025)
von: Sun, Jiayang, et al.
Veröffentlicht: (2025)
InfoBFR: Real-World Blind Face Restoration via Information Bottleneck
von: Gao, Nan, et al.
Veröffentlicht: (2025)
von: Gao, Nan, et al.
Veröffentlicht: (2025)
Think 360°: Evaluating the Width-centric Reasoning Capability of MLLMs Beyond Depth
von: Chen, Mingrui, et al.
Veröffentlicht: (2026)
von: Chen, Mingrui, et al.
Veröffentlicht: (2026)
Parallel Augmentation and Dual Enhancement for Occluded Person Re-identification
von: Wang, Zi, et al.
Veröffentlicht: (2022)
von: Wang, Zi, et al.
Veröffentlicht: (2022)
Vision Transformer with Super Token Sampling
von: Huang, Huaibo, et al.
Veröffentlicht: (2022)
von: Huang, Huaibo, et al.
Veröffentlicht: (2022)
Unlocking the Potential of Difficulty Prior in RL-based Multimodal Reasoning
von: Chen, Mingrui, et al.
Veröffentlicht: (2025)
von: Chen, Mingrui, et al.
Veröffentlicht: (2025)
MVPBench: A Multi-Video Perception Evaluation Benchmark for Multi-Modal Video Understanding
von: Bai, Purui, et al.
Veröffentlicht: (2026)
von: Bai, Purui, et al.
Veröffentlicht: (2026)
Band-Attention Modulated RetNet for Face Forgery Detection
von: Zhang, Zhida, et al.
Veröffentlicht: (2024)
von: Zhang, Zhida, et al.
Veröffentlicht: (2024)
Quantize-then-Rectify: Efficient VQ-VAE Training
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
von: Zhang, Borui, et al.
Veröffentlicht: (2025)
ProxyTransformation: Preshaping Point Cloud Manifold With Proxy Attention For 3D Visual Grounding
von: Peng, Qihang, et al.
Veröffentlicht: (2025)
von: Peng, Qihang, et al.
Veröffentlicht: (2025)
DiffMAC: Diffusion Manifold Hallucination Correction for High Generalization Blind Face Restoration
von: Gao, Nan, et al.
Veröffentlicht: (2024)
von: Gao, Nan, et al.
Veröffentlicht: (2024)
Visual Anchors Are Strong Information Aggregators For Multimodal Large Language Model
von: Liu, Haogeng, et al.
Veröffentlicht: (2024)
von: Liu, Haogeng, et al.
Veröffentlicht: (2024)
GenVideoLens: Where LVLMs Fall Short in AI-Generated Video Detection?
von: Zou, Yueying, et al.
Veröffentlicht: (2026)
von: Zou, Yueying, et al.
Veröffentlicht: (2026)
Inversion-Free Style Transfer with Dual Rectified Flows
von: Deng, Yingying, et al.
Veröffentlicht: (2025)
von: Deng, Yingying, et al.
Veröffentlicht: (2025)
Rectified Diffusion: Straightness Is Not Your Need in Rectified Flow
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
von: Wang, Fu-Yun, et al.
Veröffentlicht: (2024)
Rectified Diffusion Guidance for Conditional Generation
von: Xia, Mengfei, et al.
Veröffentlicht: (2024)
von: Xia, Mengfei, et al.
Veröffentlicht: (2024)
Survey on AI-Generated Media Detection: From Non-MLLM to MLLM
von: Zou, Yueying, et al.
Veröffentlicht: (2025)
von: Zou, Yueying, et al.
Veröffentlicht: (2025)
Linear-Time Global Visual Modeling without Explicit Attention
von: He, Ruize, et al.
Veröffentlicht: (2026)
von: He, Ruize, et al.
Veröffentlicht: (2026)
FireFlow: Fast Inversion of Rectified Flow for Image Semantic Editing
von: Deng, Yingying, et al.
Veröffentlicht: (2024)
von: Deng, Yingying, et al.
Veröffentlicht: (2024)
Runge-Kutta Approximation and Decoupled Attention for Rectified Flow Inversion and Semantic Editing
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
von: Chen, Weiming, et al.
Veröffentlicht: (2025)
BitDance: Scaling Autoregressive Generative Models with Binary Tokens
von: Ai, Yuang, et al.
Veröffentlicht: (2026)
von: Ai, Yuang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Breaking Complexity Barriers: High-Resolution Image Restoration with Rank Enhanced Linear Attention
von: Ai, Yuang, et al.
Veröffentlicht: (2025) -
Breaking the Low-Rank Dilemma of Linear Attention
von: Fan, Qihang, et al.
Veröffentlicht: (2024) -
Random Wins All: Rethinking Grouping Strategies for Vision Tokens
von: Fan, Qihang, et al.
Veröffentlicht: (2026) -
LoRA-IR: Taming Low-Rank Experts for Efficient All-in-One Image Restoration
von: Ai, Yuang, et al.
Veröffentlicht: (2024) -
DiCo: Revitalizing ConvNets for Scalable and Efficient Diffusion Modeling
von: Ai, Yuang, et al.
Veröffentlicht: (2025)