Gespeichert in:
| Hauptverfasser: | Inkawhich, Matthew, Inkawhich, Nathan, Yang, Hao, Zhang, Jingyang, Linderman, Randolph, Chen, Yiran |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | https://arxiv.org/abs/2404.10865 |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Tunable Hybrid Proposal Networks for the Open World
von: Inkawhich, Matthew, et al.
Veröffentlicht: (2022)
von: Inkawhich, Matthew, et al.
Veröffentlicht: (2022)
On the Status of Foundation Models for SAR Imagery
von: Inkawhich, Nathan
Veröffentlicht: (2025)
von: Inkawhich, Nathan
Veröffentlicht: (2025)
Comprehensive OOD Detection Improvements
von: Lakkapragada, Anish, et al.
Veröffentlicht: (2024)
von: Lakkapragada, Anish, et al.
Veröffentlicht: (2024)
Out-of-Distribution Detection via Deep Multi-Comprehension Ensemble
von: Xu, Chenhui, et al.
Veröffentlicht: (2024)
von: Xu, Chenhui, et al.
Veröffentlicht: (2024)
Multi-layer Radial Basis Function Networks for Out-of-distribution Detection
von: Khanna, Amol, et al.
Veröffentlicht: (2025)
von: Khanna, Amol, et al.
Veröffentlicht: (2025)
Modulating CNN Features with Pre-Trained ViT Representations for Open-Vocabulary Object Detection
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
von: Gao, Xiangyu, et al.
Veröffentlicht: (2025)
LiFT: A Surprisingly Simple Lightweight Feature Transform for Dense ViT Descriptors
von: Suri, Saksham, et al.
Veröffentlicht: (2024)
von: Suri, Saksham, et al.
Veröffentlicht: (2024)
Knowledge Distillation in YOLOX-ViT for Side-Scan Sonar Object Detection
von: Aubard, Martin, et al.
Veröffentlicht: (2024)
von: Aubard, Martin, et al.
Veröffentlicht: (2024)
Exploring Plain ViT Reconstruction for Multi-class Unsupervised Anomaly Detection
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
von: Zhang, Jiangning, et al.
Veröffentlicht: (2023)
I&S-ViT: An Inclusive & Stable Method for Pushing the Limit of Post-Training ViTs Quantization
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
von: Zhong, Yunshan, et al.
Veröffentlicht: (2023)
RepViT: Revisiting Mobile CNN From ViT Perspective
von: Wang, Ao, et al.
Veröffentlicht: (2023)
von: Wang, Ao, et al.
Veröffentlicht: (2023)
Revisiting Token Compression for Accelerating ViT-based Sparse Multi-View 3D Object Detectors
von: Ji, Mingqian, et al.
Veröffentlicht: (2026)
von: Ji, Mingqian, et al.
Veröffentlicht: (2026)
Deeper Inside Deep ViT
von: Hong, Sungrae
Veröffentlicht: (2025)
von: Hong, Sungrae
Veröffentlicht: (2025)
A Hybrid Framework Bridging CNN and ViT based on Theory of Evidence for Diabetic Retinopathy Grading
von: Qiu, Junlai, et al.
Veröffentlicht: (2025)
von: Qiu, Junlai, et al.
Veröffentlicht: (2025)
SPAR: Single-Pass Any-Resolution ViT for Open-vocabulary Segmentation
von: Kombol, Naomi, et al.
Veröffentlicht: (2026)
von: Kombol, Naomi, et al.
Veröffentlicht: (2026)
How to train your ViT for OOD Detection
von: Mueller, Maximilian, et al.
Veröffentlicht: (2024)
von: Mueller, Maximilian, et al.
Veröffentlicht: (2024)
EA-ViT: Efficient Adaptation for Elastic Vision Transformer
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
von: Zhu, Chen, et al.
Veröffentlicht: (2025)
UADet: A Remarkably Simple Yet Effective Uncertainty-Aware Open-Set Object Detection Framework
von: Cheng, Silin, et al.
Veröffentlicht: (2024)
von: Cheng, Silin, et al.
Veröffentlicht: (2024)
Unsupervised Object Localization in the Era of Self-Supervised ViTs: A Survey
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
von: Siméoni, Oriane, et al.
Veröffentlicht: (2023)
ViT-5: Vision Transformers for The Mid-2020s
von: Wang, Feng, et al.
Veröffentlicht: (2026)
von: Wang, Feng, et al.
Veröffentlicht: (2026)
LAMM-ViT: AI Face Detection via Layer-Aware Modulation of Region-Guided Attention
von: Zhang, Jiangling, et al.
Veröffentlicht: (2025)
von: Zhang, Jiangling, et al.
Veröffentlicht: (2025)
STRAP-ViT: Segregated Tokens with Randomized -- Transformations for Defense against Adversarial Patches in ViTs
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
von: Chattopadhyay, Nandish, et al.
Veröffentlicht: (2026)
Mobile U-ViT: Revisiting large kernel and U-shaped ViT for efficient medical image segmentation
von: Tang, Fenghe, et al.
Veröffentlicht: (2025)
von: Tang, Fenghe, et al.
Veröffentlicht: (2025)
Scene-Graph ViT: End-to-End Open-Vocabulary Visual Relationship Detection
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
von: Salzmann, Tim, et al.
Veröffentlicht: (2024)
ViTCAE: ViT-based Class-conditioned Autoencoder
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
von: Jebraeeli, Vahid, et al.
Veröffentlicht: (2025)
Exploiting Lightweight Hierarchical ViT and Dynamic Framework for Efficient Visual Tracking
von: Kang, Ben, et al.
Veröffentlicht: (2025)
von: Kang, Ben, et al.
Veröffentlicht: (2025)
Rethinking Random Masking in Self-Distillation on ViT
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
von: Seong, Jihyeon, et al.
Veröffentlicht: (2025)
YOLO-Former: YOLO Shakes Hand With ViT
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
von: Khoramdel, Javad, et al.
Veröffentlicht: (2024)
Your ViT is Secretly an Image Segmentation Model
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
von: Kerssies, Tommie, et al.
Veröffentlicht: (2025)
LogitDynamics: Reliable ViT Error Detection from Layerwise Logit Trajectories
von: Beigelman, Ido, et al.
Veröffentlicht: (2026)
von: Beigelman, Ido, et al.
Veröffentlicht: (2026)
U-REPA: Aligning Diffusion U-Nets to ViTs
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
von: Tian, Yuchuan, et al.
Veröffentlicht: (2025)
Dynamic Tuning Towards Parameter and Inference Efficiency for ViT Adaptation
von: Zhao, Wangbo, et al.
Veröffentlicht: (2024)
von: Zhao, Wangbo, et al.
Veröffentlicht: (2024)
CAS-ViT: Convolutional Additive Self-attention Vision Transformers for Efficient Mobile Applications
von: Zhang, Tianfang, et al.
Veröffentlicht: (2024)
von: Zhang, Tianfang, et al.
Veröffentlicht: (2024)
Intriguing Frequency Interpretation of Adversarial Robustness for CNNs and ViTs
von: Chen, Lu, et al.
Veröffentlicht: (2025)
von: Chen, Lu, et al.
Veröffentlicht: (2025)
Purrturbed but Stable: Human-Cat Invariant Representations Across CNNs, ViTs and Self-Supervised ViTs
von: Shah, Arya, et al.
Veröffentlicht: (2025)
von: Shah, Arya, et al.
Veröffentlicht: (2025)
A Hybrid CNN-ViT-GNN Framework with GAN-Based Augmentation for Intelligent Weed Detection in Precision Agriculture
von: V, Pandiyaraju, et al.
Veröffentlicht: (2025)
von: V, Pandiyaraju, et al.
Veröffentlicht: (2025)
ViT-Lens: Towards Omni-modal Representations
von: Lei, Weixian, et al.
Veröffentlicht: (2023)
von: Lei, Weixian, et al.
Veröffentlicht: (2023)
Harnessing the Computation Redundancy in ViTs to Boost Adversarial Transferability
von: Liu, Jiani, et al.
Veröffentlicht: (2025)
von: Liu, Jiani, et al.
Veröffentlicht: (2025)
Multimodal Informative ViT: Information Aggregation and Distribution for Hyperspectral and LiDAR Classification
von: Zhang, Jiaqing, et al.
Veröffentlicht: (2024)
von: Zhang, Jiaqing, et al.
Veröffentlicht: (2024)
Vanilla ViT for Automotive Point Cloud Semantic Segmentation
von: Puy, Gilles, et al.
Veröffentlicht: (2026)
von: Puy, Gilles, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Tunable Hybrid Proposal Networks for the Open World
von: Inkawhich, Matthew, et al.
Veröffentlicht: (2022) -
On the Status of Foundation Models for SAR Imagery
von: Inkawhich, Nathan
Veröffentlicht: (2025) -
Comprehensive OOD Detection Improvements
von: Lakkapragada, Anish, et al.
Veröffentlicht: (2024) -
Out-of-Distribution Detection via Deep Multi-Comprehension Ensemble
von: Xu, Chenhui, et al.
Veröffentlicht: (2024) -
Multi-layer Radial Basis Function Networks for Out-of-distribution Detection
von: Khanna, Amol, et al.
Veröffentlicht: (2025)