Orthogonal Quadratic Complements for Vision Transformer Feed-Forward Networks
Fuente:
arXiv
Salvato in:
| Autore principale: | Zixian, Wang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
CascadedViT: Cascaded Chunk-FeedForward and Cascaded Group Attention Vision Transformer
di: Sivakumar, Srivathsan, et al.
Pubblicazione: (2025)
di: Sivakumar, Srivathsan, et al.
Pubblicazione: (2025)
Precise, Fast, and Low-cost Concept Erasure in Value Space: Orthogonal Complement Matters
di: Wang, Yuan, et al.
Pubblicazione: (2024)
di: Wang, Yuan, et al.
Pubblicazione: (2024)
Bi-Orthogonal Factor Decomposition for Vision Transformers
di: Doshi, Fenil R., et al.
Pubblicazione: (2026)
di: Doshi, Fenil R., et al.
Pubblicazione: (2026)
You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass
di: Yang, Yinuo, et al.
Pubblicazione: (2026)
di: Yang, Yinuo, et al.
Pubblicazione: (2026)
Efficient Adaptation of Pre-trained Vision Transformer underpinned by Approximately Orthogonal Fine-Tuning Strategy
di: Yang, Yiting, et al.
Pubblicazione: (2025)
di: Yang, Yiting, et al.
Pubblicazione: (2025)
Particulate: Feed-Forward 3D Object Articulation
di: Li, Ruining, et al.
Pubblicazione: (2025)
di: Li, Ruining, et al.
Pubblicazione: (2025)
AREA3D: Active Reconstruction Agent with Unified Feed-Forward 3D Perception and Vision-Language Guidance
di: Xu, Tianling, et al.
Pubblicazione: (2025)
di: Xu, Tianling, et al.
Pubblicazione: (2025)
DenoiseSplat: Feed-Forward Gaussian Splatting for Noisy 3D Scene Reconstruction
di: Jiang, Fuzhen, et al.
Pubblicazione: (2026)
di: Jiang, Fuzhen, et al.
Pubblicazione: (2026)
Feed-Forward 3D Scene Modeling: A Problem-Driven Perspective
di: Wang, Weijie, et al.
Pubblicazione: (2026)
di: Wang, Weijie, et al.
Pubblicazione: (2026)
Training Convolutional Neural Networks with the Forward-Forward algorithm
di: Scodellaro, Riccardo, et al.
Pubblicazione: (2023)
di: Scodellaro, Riccardo, et al.
Pubblicazione: (2023)
Group Orthogonalization Regularization For Vision Models Adaptation and Robustness
di: Kurtz, Yoav, et al.
Pubblicazione: (2023)
di: Kurtz, Yoav, et al.
Pubblicazione: (2023)
Feed-Forward Bullet-Time Reconstruction of Dynamic Scenes from Monocular Videos
di: Liang, Hanxue, et al.
Pubblicazione: (2024)
di: Liang, Hanxue, et al.
Pubblicazione: (2024)
Resource-efficient Medical Image Analysis with Self-adapting Forward-Forward Networks
di: Müller, Johanna P., et al.
Pubblicazione: (2024)
di: Müller, Johanna P., et al.
Pubblicazione: (2024)
Learning to Complement and to Defer to Multiple Users
di: Zhang, Zheng, et al.
Pubblicazione: (2024)
di: Zhang, Zheng, et al.
Pubblicazione: (2024)
Enhancing Breast Cancer Detection with Vision Transformers and Graph Neural Networks
di: Cai, Yeming, et al.
Pubblicazione: (2025)
di: Cai, Yeming, et al.
Pubblicazione: (2025)
Concept Complement Bottleneck Model for Interpretable Medical Image Diagnosis
di: Wang, Hongmei, et al.
Pubblicazione: (2024)
di: Wang, Hongmei, et al.
Pubblicazione: (2024)
ArtSplat: Feed-Forward Articulated 3D Gaussian Splatting from Sparse Multi-State Uncalibrated Views
di: Lee, Inseo, et al.
Pubblicazione: (2026)
di: Lee, Inseo, et al.
Pubblicazione: (2026)
ViT-Linearizer: Distilling Quadratic Knowledge into Linear-Time Vision Models
di: Wei, Guoyizhe, et al.
Pubblicazione: (2025)
di: Wei, Guoyizhe, et al.
Pubblicazione: (2025)
Vision Bridge Transformer at Scale
di: Tan, Zhenxiong, et al.
Pubblicazione: (2025)
di: Tan, Zhenxiong, et al.
Pubblicazione: (2025)
CLQ: Cross-Layer Guided Orthogonal-based Quantization for Diffusion Transformers
di: Liu, Kai, et al.
Pubblicazione: (2025)
di: Liu, Kai, et al.
Pubblicazione: (2025)
GeoDecoder: Empowering Multimodal Map Understanding
di: Qi, Feng, et al.
Pubblicazione: (2024)
di: Qi, Feng, et al.
Pubblicazione: (2024)
FPRF: Feed-Forward Photorealistic Style Transfer of Large-Scale 3D Neural Radiance Fields
di: Kim, GeonU, et al.
Pubblicazione: (2024)
di: Kim, GeonU, et al.
Pubblicazione: (2024)
AT-SNN: Adaptive Tokens for Vision Transformer on Spiking Neural Network
di: Kang, Donghwa, et al.
Pubblicazione: (2024)
di: Kang, Donghwa, et al.
Pubblicazione: (2024)
VFIG: Vectorizing Complex Figures in SVG with Vision-Language Models
di: He, Qijia, et al.
Pubblicazione: (2026)
di: He, Qijia, et al.
Pubblicazione: (2026)
Hierarchical Vision Transformer Enhanced by Graph Convolutional Network for Image Classification
di: Jiao, Haibin
Pubblicazione: (2026)
di: Jiao, Haibin
Pubblicazione: (2026)
ProVision: Programmatically Scaling Vision-centric Instruction Data for Multimodal Language Models
di: Zhang, Jieyu, et al.
Pubblicazione: (2024)
di: Zhang, Jieyu, et al.
Pubblicazione: (2024)
Weierstrass Positional Encoding for Vision Transformers
di: Xin, Zhihang, et al.
Pubblicazione: (2026)
di: Xin, Zhihang, et al.
Pubblicazione: (2026)
Efficient Feature-Free Initialization for Monocular Visual-Inertial Systems Using a Feed-Forward 3D Model
di: Zhang, Yuantai, et al.
Pubblicazione: (2026)
di: Zhang, Yuantai, et al.
Pubblicazione: (2026)
MapAnything: Universal Feed-Forward Metric 3D Reconstruction
di: Keetha, Nikhil, et al.
Pubblicazione: (2025)
di: Keetha, Nikhil, et al.
Pubblicazione: (2025)
Revisiting Residual Connections: Orthogonal Updates for Stable and Efficient Deep Networks
di: Oh, Giyeong, et al.
Pubblicazione: (2025)
di: Oh, Giyeong, et al.
Pubblicazione: (2025)
RePaViT: Scalable Vision Transformer Acceleration via Structural Reparameterization on Feedforward Network Layers
di: Xu, Xuwei, et al.
Pubblicazione: (2025)
di: Xu, Xuwei, et al.
Pubblicazione: (2025)
IoT Botnet Detection: Application of Vision Transformer to Classification of Network Flow Traffic
di: Wasswa, Hassan, et al.
Pubblicazione: (2025)
di: Wasswa, Hassan, et al.
Pubblicazione: (2025)
Spiking Vision Transformer with Saccadic Attention
di: Wang, Shuai, et al.
Pubblicazione: (2025)
di: Wang, Shuai, et al.
Pubblicazione: (2025)
Efficient Adaptation of Pre-trained Vision Transformer via Householder Transformation
di: Dong, Wei, et al.
Pubblicazione: (2024)
di: Dong, Wei, et al.
Pubblicazione: (2024)
Benchmarking Unlearning for Vision Transformers
di: Zhao, Kairan, et al.
Pubblicazione: (2026)
di: Zhao, Kairan, et al.
Pubblicazione: (2026)
Any4D: Unified Feed-Forward Metric 4D Reconstruction
di: Karhade, Jay, et al.
Pubblicazione: (2025)
di: Karhade, Jay, et al.
Pubblicazione: (2025)
Multi-Level Heterogeneous Knowledge Transfer Network on Forward Scattering Center Model for Limited Samples SAR ATR
di: Zhao, Chenxi, et al.
Pubblicazione: (2025)
di: Zhao, Chenxi, et al.
Pubblicazione: (2025)
Can Cross-Layer Transcoders Replace Vision Transformer Activations? An Interpretable Perspective on Vision
di: Chatzoudis, Gerasimos, et al.
Pubblicazione: (2026)
di: Chatzoudis, Gerasimos, et al.
Pubblicazione: (2026)
TinyDrop: Tiny Model Guided Token Dropping for Vision Transformers
di: Wang, Guoxin, et al.
Pubblicazione: (2025)
di: Wang, Guoxin, et al.
Pubblicazione: (2025)
MRI Embeddings Complement Clinical Predictors for Cognitive Decline Modeling in Alzheimer's Disease Cohorts
di: Putera, Nathaniel, et al.
Pubblicazione: (2025)
di: Putera, Nathaniel, et al.
Pubblicazione: (2025)
Documenti analoghi
-
CascadedViT: Cascaded Chunk-FeedForward and Cascaded Group Attention Vision Transformer
di: Sivakumar, Srivathsan, et al.
Pubblicazione: (2025) -
Precise, Fast, and Low-cost Concept Erasure in Value Space: Orthogonal Complement Matters
di: Wang, Yuan, et al.
Pubblicazione: (2024) -
Bi-Orthogonal Factor Decomposition for Vision Transformers
di: Doshi, Fenil R., et al.
Pubblicazione: (2026) -
You Only Judge Once: Multi-response Reward Modeling in a Single Forward Pass
di: Yang, Yinuo, et al.
Pubblicazione: (2026) -
Efficient Adaptation of Pre-trained Vision Transformer underpinned by Approximately Orthogonal Fine-Tuning Strategy
di: Yang, Yiting, et al.
Pubblicazione: (2025)