Memory Efficient Transformer Adapter for Dense Predictions
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhang, Dong, Yan, Rui, Dong, Pingcheng, Cheng, Kwang-Ting |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Towards Customized Knowledge Distillation for Chip-Level Dense Image Predictions
von: Zhang, Dong, et al.
Veröffentlicht: (2024)
von: Zhang, Dong, et al.
Veröffentlicht: (2024)
Quantization Variation: A New Perspective on Training Transformers with Low-Bit Precision
von: Huang, Xijie, et al.
Veröffentlicht: (2023)
von: Huang, Xijie, et al.
Veröffentlicht: (2023)
Token Merging via Spatiotemporal Information Mining for Surgical Video Understanding
von: Jiang, Xixi, et al.
Veröffentlicht: (2025)
von: Jiang, Xixi, et al.
Veröffentlicht: (2025)
LLM-FP4: 4-Bit Floating-Point Quantized Transformers
von: Liu, Shih-yang, et al.
Veröffentlicht: (2023)
von: Liu, Shih-yang, et al.
Veröffentlicht: (2023)
Generalized Task-Driven Medical Image Quality Enhancement with Gradient Promotion
von: Zhang, Dong, et al.
Veröffentlicht: (2025)
von: Zhang, Dong, et al.
Veröffentlicht: (2025)
Efficient Adaptation of Large Vision Transformer via Adapter Re-Composing
von: Dong, Wei, et al.
Veröffentlicht: (2023)
von: Dong, Wei, et al.
Veröffentlicht: (2023)
Online Dense Point Tracking with Streaming Memory
von: Dong, Qiaole, et al.
Veröffentlicht: (2025)
von: Dong, Qiaole, et al.
Veröffentlicht: (2025)
BiDense: Binarization for Dense Prediction
von: Yin, Rui, et al.
Veröffentlicht: (2024)
von: Yin, Rui, et al.
Veröffentlicht: (2024)
ConSept: Continual Semantic Segmentation via Adapter-based Vision Transformer
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
von: Dong, Bowen, et al.
Veröffentlicht: (2024)
Dynamic Adapter with Semantics Disentangling for Cross-lingual Cross-modal Retrieval
von: Cai, Rui, et al.
Veröffentlicht: (2024)
von: Cai, Rui, et al.
Veröffentlicht: (2024)
Hierarchical Awareness Adapters with Hybrid Pyramid Feature Fusion for Dense Depth Prediction
von: Su, Wuqi, et al.
Veröffentlicht: (2026)
von: Su, Wuqi, et al.
Veröffentlicht: (2026)
CAD: Memory Efficient Convolutional Adapter for Segment Anything
von: Kim, Joohyeok, et al.
Veröffentlicht: (2024)
von: Kim, Joohyeok, et al.
Veröffentlicht: (2024)
BoNuS: Boundary Mining for Nuclei Segmentation with Partial Point Labels
von: Lin, Yi, et al.
Veröffentlicht: (2024)
von: Lin, Yi, et al.
Veröffentlicht: (2024)
Aligning Medical Images with General Knowledge from Large Language Models
von: Fang, Xiao, et al.
Veröffentlicht: (2024)
von: Fang, Xiao, et al.
Veröffentlicht: (2024)
Boosting Convolution with Efficient MLP-Permutation for Volumetric Medical Image Segmentation
von: Lin, Yi, et al.
Veröffentlicht: (2023)
von: Lin, Yi, et al.
Veröffentlicht: (2023)
Non-parametric regularization for class imbalance federated medical image classification
von: Wicaksana, Jeffry, et al.
Veröffentlicht: (2024)
von: Wicaksana, Jeffry, et al.
Veröffentlicht: (2024)
iDAT: inverse Distillation Adapter-Tuning
von: Ruan, Jiacheng, et al.
Veröffentlicht: (2024)
von: Ruan, Jiacheng, et al.
Veröffentlicht: (2024)
A Study of Finetuning Video Transformers for Multi-view Geometry Tasks
von: Wu, Huimin, et al.
Veröffentlicht: (2025)
von: Wu, Huimin, et al.
Veröffentlicht: (2025)
Prompt-Guided Adaptive Model Transformation for Whole Slide Image Classification
von: Lin, Yi, et al.
Veröffentlicht: (2024)
von: Lin, Yi, et al.
Veröffentlicht: (2024)
Beyond Prompt Learning: Continual Adapter for Efficient Rehearsal-Free Continual Learning
von: Gao, Xinyuan, et al.
Veröffentlicht: (2024)
von: Gao, Xinyuan, et al.
Veröffentlicht: (2024)
RelationAdapter: Learning and Transferring Visual Relation with Diffusion Transformers
von: Gong, Yan, et al.
Veröffentlicht: (2025)
von: Gong, Yan, et al.
Veröffentlicht: (2025)
Labeled-to-Unlabeled Distribution Alignment for Partially-Supervised Multi-Organ Medical Image Segmentation
von: Jiang, Xixi, et al.
Veröffentlicht: (2024)
von: Jiang, Xixi, et al.
Veröffentlicht: (2024)
Sparse-Dense Mixture of Experts Adapter for Multi-Modal Tracking
von: Zhu, Yabin, et al.
Veröffentlicht: (2026)
von: Zhu, Yabin, et al.
Veröffentlicht: (2026)
MemFlow: Optical Flow Estimation and Prediction with Memory
von: Dong, Qiaole, et al.
Veröffentlicht: (2024)
von: Dong, Qiaole, et al.
Veröffentlicht: (2024)
Vision Transformers: From Semantic Segmentation to Dense Prediction
von: Zhang, Li, et al.
Veröffentlicht: (2022)
von: Zhang, Li, et al.
Veröffentlicht: (2022)
Cyclic Contrastive Knowledge Transfer for Open-Vocabulary Object Detection
von: Zhang, Chuhan, et al.
Veröffentlicht: (2025)
von: Zhang, Chuhan, et al.
Veröffentlicht: (2025)
ExFusion: Efficient Transformer Training via Multi-Experts Fusion
von: Ruan, Jiacheng, et al.
Veröffentlicht: (2026)
von: Ruan, Jiacheng, et al.
Veröffentlicht: (2026)
Efficient Prediction of Dense Visual Embeddings via Distillation and RGB-D Transformers
von: Fischedick, Söhnke Benedikt, et al.
Veröffentlicht: (2026)
von: Fischedick, Söhnke Benedikt, et al.
Veröffentlicht: (2026)
Inf-DiT: Upsampling Any-Resolution Image with Memory-Efficient Diffusion Transformer
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2024)
von: Yang, Zhuoyi, et al.
Veröffentlicht: (2024)
Task Indicating Transformer for Task-conditional Dense Predictions
von: Lu, Yuxiang, et al.
Veröffentlicht: (2024)
von: Lu, Yuxiang, et al.
Veröffentlicht: (2024)
EVCtrl: Efficient Control Adapter for Visual Generation
von: Yang, Zixiang, et al.
Veröffentlicht: (2025)
von: Yang, Zixiang, et al.
Veröffentlicht: (2025)
DeepFake-Adapter: Dual-Level Adapter for DeepFake Detection
von: Shao, Rui, et al.
Veröffentlicht: (2023)
von: Shao, Rui, et al.
Veröffentlicht: (2023)
CLIPSelf: Vision Transformer Distills Itself for Open-Vocabulary Dense Prediction
von: Wu, Size, et al.
Veröffentlicht: (2023)
von: Wu, Size, et al.
Veröffentlicht: (2023)
ThreshNet: An Efficient DenseNet Using Threshold Mechanism to Reduce Connections
von: Ju, Rui-Yang, et al.
Veröffentlicht: (2022)
von: Ju, Rui-Yang, et al.
Veröffentlicht: (2022)
From Frames to Sequences: Temporally Consistent Human-Centric Dense Prediction
von: Miao, Xingyu, et al.
Veröffentlicht: (2026)
von: Miao, Xingyu, et al.
Veröffentlicht: (2026)
Dense-Face: Personalized Face Generation Model via Dense Annotation Prediction
von: Guo, Xiao, et al.
Veröffentlicht: (2024)
von: Guo, Xiao, et al.
Veröffentlicht: (2024)
Box2Poly: Memory-Efficient Polygon Prediction of Arbitrarily Shaped and Rotated Text
von: Chen, Xuyang, et al.
Veröffentlicht: (2023)
von: Chen, Xuyang, et al.
Veröffentlicht: (2023)
Towards Efficient General Feature Prediction in Masked Skeleton Modeling
von: Sun, Shengkai, et al.
Veröffentlicht: (2025)
von: Sun, Shengkai, et al.
Veröffentlicht: (2025)
OptiSAR-Net++: A Large-Scale Benchmark and Transformer-Free Framework for Cross-Domain Remote Sensing Visual Grounding
von: Tang, Xiaoyu, et al.
Veröffentlicht: (2026)
von: Tang, Xiaoyu, et al.
Veröffentlicht: (2026)
A Memory-Efficient Framework for Deformable Transformer with Neural Architecture Search
von: Mao, Wendong, et al.
Veröffentlicht: (2025)
von: Mao, Wendong, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Towards Customized Knowledge Distillation for Chip-Level Dense Image Predictions
von: Zhang, Dong, et al.
Veröffentlicht: (2024) -
Quantization Variation: A New Perspective on Training Transformers with Low-Bit Precision
von: Huang, Xijie, et al.
Veröffentlicht: (2023) -
Token Merging via Spatiotemporal Information Mining for Surgical Video Understanding
von: Jiang, Xixi, et al.
Veröffentlicht: (2025) -
LLM-FP4: 4-Bit Floating-Point Quantized Transformers
von: Liu, Shih-yang, et al.
Veröffentlicht: (2023) -
Generalized Task-Driven Medical Image Quality Enhancement with Gradient Promotion
von: Zhang, Dong, et al.
Veröffentlicht: (2025)