Gespeichert in:
| Hauptverfasser: | Snyder, Thomas, Yang, H. Lexie, Schnake, Stefan, Schotthöfer, Steffen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2601.08882 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Dynamical Low-Rank Compression of Neural Networks with Robustness under Adversarial Attacks
von: Schotthöfer, Steffen, et al.
Veröffentlicht: (2025)
von: Schotthöfer, Steffen, et al.
Veröffentlicht: (2025)
Global Context Compression with Interleaved Vision-Text Transformation
von: Jiao, Dian, et al.
Veröffentlicht: (2026)
von: Jiao, Dian, et al.
Veröffentlicht: (2026)
Proximal Vision Transformer: Enhancing Feature Representation through Two-Stage Manifold Geometry
von: Yun, Haoyu, et al.
Veröffentlicht: (2025)
von: Yun, Haoyu, et al.
Veröffentlicht: (2025)
ButterflyViT: 354$\times$ Expert Compression for Edge Vision Transformers
von: Karmore, Aryan
Veröffentlicht: (2026)
von: Karmore, Aryan
Veröffentlicht: (2026)
Towards Difficulty-Agnostic Efficient Transfer Learning for Vision-Language Models
von: Yang, Yongjin, et al.
Veröffentlicht: (2023)
von: Yang, Yongjin, et al.
Veröffentlicht: (2023)
Lossy Neural Compression for Geospatial Analytics: A Review
von: Gomes, Carlos, et al.
Veröffentlicht: (2025)
von: Gomes, Carlos, et al.
Veröffentlicht: (2025)
Deep Extrinsic Manifold Representation for Vision Tasks
von: Zhang, Tongtong, et al.
Veröffentlicht: (2024)
von: Zhang, Tongtong, et al.
Veröffentlicht: (2024)
Patch Rebirth: Toward Fast and Transferable Model Inversion of Vision Transformers
von: Heo, Seongsoo, et al.
Veröffentlicht: (2025)
von: Heo, Seongsoo, et al.
Veröffentlicht: (2025)
EFTViT: Efficient Federated Training of Vision Transformers with Masked Images on Resource-Constrained Clients
von: Wu, Meihan, et al.
Veröffentlicht: (2024)
von: Wu, Meihan, et al.
Veröffentlicht: (2024)
MMEarth: Exploring Multi-Modal Pretext Tasks For Geospatial Representation Learning
von: Nedungadi, Vishal, et al.
Veröffentlicht: (2024)
von: Nedungadi, Vishal, et al.
Veröffentlicht: (2024)
Learning to Transform Dynamically for Better Adversarial Transferability
von: Zhu, Rongyi, et al.
Veröffentlicht: (2024)
von: Zhu, Rongyi, et al.
Veröffentlicht: (2024)
Downstream Transfer Attack: Adversarial Attacks on Downstream Models with Pre-trained Vision Transformers
von: Zheng, Weijie, et al.
Veröffentlicht: (2024)
von: Zheng, Weijie, et al.
Veröffentlicht: (2024)
End-to-End Optimized Image Compression with the Frequency-Oriented Transform
von: Zhang, Yuefeng, et al.
Veröffentlicht: (2024)
von: Zhang, Yuefeng, et al.
Veröffentlicht: (2024)
Vision Bridge Transformer at Scale
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2025)
von: Tan, Zhenxiong, et al.
Veröffentlicht: (2025)
FlattenGPT: Depth Compression for Transformer with Layer Flattening
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
von: Xu, Ruihan, et al.
Veröffentlicht: (2026)
Attention Retention for Continual Learning with Vision Transformers
von: Lu, Yue, et al.
Veröffentlicht: (2026)
von: Lu, Yue, et al.
Veröffentlicht: (2026)
Smartflow: Enabling Scalable Spatiotemporal Geospatial Research
von: McVicar, David, et al.
Veröffentlicht: (2025)
von: McVicar, David, et al.
Veröffentlicht: (2025)
Efficient Adaptation of Pre-trained Vision Transformer via Householder Transformation
von: Dong, Wei, et al.
Veröffentlicht: (2024)
von: Dong, Wei, et al.
Veröffentlicht: (2024)
OReole-FM: successes and challenges toward billion-parameter foundation models for high-resolution satellite imagery
von: Dias, Philipe, et al.
Veröffentlicht: (2024)
von: Dias, Philipe, et al.
Veröffentlicht: (2024)
Bi-Orthogonal Factor Decomposition for Vision Transformers
von: Doshi, Fenil R., et al.
Veröffentlicht: (2026)
von: Doshi, Fenil R., et al.
Veröffentlicht: (2026)
Unlocking Feature Visualization for Deeper Networks with MAgnitude Constrained Optimization
von: Fel, Thomas, et al.
Veröffentlicht: (2023)
von: Fel, Thomas, et al.
Veröffentlicht: (2023)
Spiking Vision Transformer with Saccadic Attention
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
von: Wang, Shuai, et al.
Veröffentlicht: (2025)
Vision Transformers for Zero-Shot Clustering of Animal Images: A Comparative Benchmarking Study
von: Markoff, Hugo, et al.
Veröffentlicht: (2026)
von: Markoff, Hugo, et al.
Veröffentlicht: (2026)
ACPO: Anchor-Constrained Perceptual Optimization for Diffusion Models with No-Reference Quality Guidance
von: Yang, Yang, et al.
Veröffentlicht: (2026)
von: Yang, Yang, et al.
Veröffentlicht: (2026)
Manifold-Aware Exploration for Reinforcement Learning in Video Generation
von: Zheng, Mingzhe, et al.
Veröffentlicht: (2026)
von: Zheng, Mingzhe, et al.
Veröffentlicht: (2026)
Multi-Context Fusion Transformer for Pedestrian Crossing Intention Prediction in Urban Environments
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025)
von: Li, Yuanzhe, et al.
Veröffentlicht: (2025)
Towards Lossless Ultimate Vision Token Compression for VLMs
von: Zheng, Dehua, et al.
Veröffentlicht: (2025)
von: Zheng, Dehua, et al.
Veröffentlicht: (2025)
Memory-Efficient Vision Transformers: An Activation-Aware Mixed-Rank Compression Strategy
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2024)
von: Azizi, Seyedarmin, et al.
Veröffentlicht: (2024)
TVE: Learning Meta-attribution for Transferable Vision Explainer
von: Wang, Guanchu, et al.
Veröffentlicht: (2023)
von: Wang, Guanchu, et al.
Veröffentlicht: (2023)
Vision without Images: End-to-End Computer Vision from Single Compressive Measurements
von: Pan, Fengpu, et al.
Veröffentlicht: (2025)
von: Pan, Fengpu, et al.
Veröffentlicht: (2025)
Boosting Adversarial Transferability across Model Genus by Deformation-Constrained Warping
von: Lin, Qinliang, et al.
Veröffentlicht: (2024)
von: Lin, Qinliang, et al.
Veröffentlicht: (2024)
Geometrically Constrained and Token-Based Probabilistic Spatial Transformers
von: Schmidt, Johann, et al.
Veröffentlicht: (2025)
von: Schmidt, Johann, et al.
Veröffentlicht: (2025)
Transfer Learning Applied to Computer Vision Problems: Survey on Current Progress, Limitations, and Opportunities
von: Panda, Aaryan, et al.
Veröffentlicht: (2024)
von: Panda, Aaryan, et al.
Veröffentlicht: (2024)
Understanding the Transfer Limits of Vision Foundation Models
von: Huang, Shiqi, et al.
Veröffentlicht: (2026)
von: Huang, Shiqi, et al.
Veröffentlicht: (2026)
Charting New Territories: Exploring the Geographic and Geospatial Capabilities of Multimodal LLMs
von: Roberts, Jonathan, et al.
Veröffentlicht: (2023)
von: Roberts, Jonathan, et al.
Veröffentlicht: (2023)
Seg the HAB: Language-Guided Geospatial Algae Bloom Reasoning and Segmentation
von: Hsieh, Patterson, et al.
Veröffentlicht: (2025)
von: Hsieh, Patterson, et al.
Veröffentlicht: (2025)
Towards Scalable Foundation Model for Multi-modal and Hyperspectral Geospatial Data
von: Si, Haozhe, et al.
Veröffentlicht: (2025)
von: Si, Haozhe, et al.
Veröffentlicht: (2025)
Learning to Merge Tokens via Decoupled Embedding for Efficient Vision Transformers
von: Lee, Dong Hoon, et al.
Veröffentlicht: (2024)
von: Lee, Dong Hoon, et al.
Veröffentlicht: (2024)
Forensic License Plate Recognition with Compression-Informed Transformers
von: Moussa, Denise, et al.
Veröffentlicht: (2022)
von: Moussa, Denise, et al.
Veröffentlicht: (2022)
PQV-Mobile: A Combined Pruning and Quantization Toolkit to Optimize Vision Transformers for Mobile Applications
von: Bhardwaj, Kshitij
Veröffentlicht: (2024)
von: Bhardwaj, Kshitij
Veröffentlicht: (2024)
Ähnliche Einträge
-
Dynamical Low-Rank Compression of Neural Networks with Robustness under Adversarial Attacks
von: Schotthöfer, Steffen, et al.
Veröffentlicht: (2025) -
Global Context Compression with Interleaved Vision-Text Transformation
von: Jiao, Dian, et al.
Veröffentlicht: (2026) -
Proximal Vision Transformer: Enhancing Feature Representation through Two-Stage Manifold Geometry
von: Yun, Haoyu, et al.
Veröffentlicht: (2025) -
ButterflyViT: 354$\times$ Expert Compression for Edge Vision Transformers
von: Karmore, Aryan
Veröffentlicht: (2026) -
Towards Difficulty-Agnostic Efficient Transfer Learning for Vision-Language Models
von: Yang, Yongjin, et al.
Veröffentlicht: (2023)