VORTEX: Challenging CNNs at Texture Recognition by using Vision Transformers with Orderless and Randomized Token Encodings
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Scabini, Leonardo, Zielinski, Kallil M., Konuk, Emir, Fares, Ricardo T., Ribas, Lucas C., Smith, Kevin, Bruno, Odemir M. |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Comparative Survey of Vision Transformers for Feature Extraction in Texture Analysis
von: Scabini, Leonardo, et al.
Veröffentlicht: (2024)
von: Scabini, Leonardo, et al.
Veröffentlicht: (2024)
Advanced wood species identification based on multiple anatomical sections and using deep feature transfer and fusion
von: Zielinski, Kallil M., et al.
Veröffentlicht: (2024)
von: Zielinski, Kallil M., et al.
Veröffentlicht: (2024)
MIXER: Mixed Hyperspherical Random Embedding Neural Network for Texture Recognition
von: Fares, Ricardo T., et al.
Veröffentlicht: (2025)
von: Fares, Ricardo T., et al.
Veröffentlicht: (2025)
APLA: A Simple Adaptation Method for Vision Transformers
von: Sorkhei, Moein, et al.
Veröffentlicht: (2025)
von: Sorkhei, Moein, et al.
Veröffentlicht: (2025)
The Cost of Reasoning: Chain-of-Thought Induces Overconfidence in Vision-Language Models
von: Welch, Robert, et al.
Veröffentlicht: (2026)
von: Welch, Robert, et al.
Veröffentlicht: (2026)
A Comprehensive Taxonomy of Cellular Automata
von: Rollier, Michiel, et al.
Veröffentlicht: (2024)
von: Rollier, Michiel, et al.
Veröffentlicht: (2024)
Learning What Helps: Task-Aligned Context Selection for Vision Tasks
von: Guo, Jingyu, et al.
Veröffentlicht: (2025)
von: Guo, Jingyu, et al.
Veröffentlicht: (2025)
Learning from Offline Foundation Features with Tensor Augmentations
von: Konuk, Emir, et al.
Veröffentlicht: (2024)
von: Konuk, Emir, et al.
Veröffentlicht: (2024)
k-NN as a Simple and Effective Estimator of Transferability
von: Sorkhei, Moein, et al.
Veröffentlicht: (2025)
von: Sorkhei, Moein, et al.
Veröffentlicht: (2025)
Locally Orderless Images for Optimization in Differentiable Rendering
von: Mehta, Ishit, et al.
Veröffentlicht: (2025)
von: Mehta, Ishit, et al.
Veröffentlicht: (2025)
Efficient Self-Supervised Adaptation for Medical Image Analysis
von: Sorkhei, Moein, et al.
Veröffentlicht: (2025)
von: Sorkhei, Moein, et al.
Veröffentlicht: (2025)
Evaluation of Activated Sludge Settling Characteristics from Microscopy Images with Deep Convolutional Neural Networks and Transfer Learning
von: Borzooei, Sina, et al.
Veröffentlicht: (2024)
von: Borzooei, Sina, et al.
Veröffentlicht: (2024)
Network classification through random walks
von: Travieso, Gonzalo, et al.
Veröffentlicht: (2025)
von: Travieso, Gonzalo, et al.
Veröffentlicht: (2025)
Essential metrics for Life on graphs
von: Rollier, Michiel, et al.
Veröffentlicht: (2025)
von: Rollier, Michiel, et al.
Veröffentlicht: (2025)
A Benchmarking Framework for Network Classification Methods
von: Merenda, Joao V., et al.
Veröffentlicht: (2025)
von: Merenda, Joao V., et al.
Veröffentlicht: (2025)
Pose Matters: Evaluating Vision Transformers and CNNs for Human Action Recognition on Small COCO Subsets
von: Tang, MingZe, et al.
Veröffentlicht: (2025)
von: Tang, MingZe, et al.
Veröffentlicht: (2025)
La Psicología y La Psicoterapia en Otros Países "La Psicología Tiene un Largo Pasado Pero una Historia Corta": Turquía, como un ejemplo
von: Emre Konuk
Veröffentlicht: (2011)
von: Emre Konuk
Veröffentlicht: (2011)
Efficient Simulation of Non-uniform Cellular Automata with a Convolutional Neural Network
von: Rollier, Michiel, et al.
Veröffentlicht: (2024)
von: Rollier, Michiel, et al.
Veröffentlicht: (2024)
On The Potential of The Fractal Geometry and The CNNs Ability to Encode it
von: Zini, Julia El, et al.
Veröffentlicht: (2024)
von: Zini, Julia El, et al.
Veröffentlicht: (2024)
B-cos Alignment for Inherently Interpretable CNNs and Vision Transformers
von: Böhle, Moritz, et al.
Veröffentlicht: (2023)
von: Böhle, Moritz, et al.
Veröffentlicht: (2023)
From CNNs to Transformers in Multimodal Human Action Recognition: A Survey
von: Shaikh, Muhammad Bilal, et al.
Veröffentlicht: (2024)
von: Shaikh, Muhammad Bilal, et al.
Veröffentlicht: (2024)
TORE: Token Recycling in Vision Transformers for Efficient Active Visual Exploration
von: Olszewski, Jan, et al.
Veröffentlicht: (2023)
von: Olszewski, Jan, et al.
Veröffentlicht: (2023)
VISUALIZATION STUDIES OF THE VORTEX SYSTEM AROUND 3-D RECTANGULAR BUILDINGS
von: Stephanie C. Zucoloto
Veröffentlicht: (2013)
von: Stephanie C. Zucoloto
Veröffentlicht: (2013)
Joint Encoding of KV-Cache Blocks for Scalable LLM Serving
von: Kampeas, Joseph, et al.
Veröffentlicht: (2026)
von: Kampeas, Joseph, et al.
Veröffentlicht: (2026)
Random Token Fusion for Multi-View Medical Diagnosis
von: Guo, Jingyu, et al.
Veröffentlicht: (2024)
von: Guo, Jingyu, et al.
Veröffentlicht: (2024)
CNNs, Transformers, Hybrid, and Vision Language Models for Skin Cancer Detection
von: Dey, Durjoy, et al.
Veröffentlicht: (2026)
von: Dey, Durjoy, et al.
Veröffentlicht: (2026)
Shape and Texture Recognition in Large Vision-Language Models
von: Eppel, Sagi, et al.
Veröffentlicht: (2025)
von: Eppel, Sagi, et al.
Veröffentlicht: (2025)
Deep Network Pruning: A Comparative Study on CNNs in Face Recognition
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2024)
von: Alonso-Fernandez, Fernando, et al.
Veröffentlicht: (2024)
X-VORTEX: Spatio-Temporal Contrastive Learning for Wake Vortex Trajectory Forecasting
von: Qu, Zhan, et al.
Veröffentlicht: (2026)
von: Qu, Zhan, et al.
Veröffentlicht: (2026)
Zeros Of Random Analytic Functions And Spectral Properties Of Perturbed Unitary Matrices
von: Fares, Aniss
Veröffentlicht: (2025)
von: Fares, Aniss
Veröffentlicht: (2025)
RNNs, CNNs and Transformers in Human Action Recognition: A Survey and a Hybrid Model
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
von: Alomar, Khaled, et al.
Veröffentlicht: (2024)
Vision Transformers for Kidney Stone Image Classification: A Comparative Study with CNNs
von: Reyes-Amezcua, Ivan, et al.
Veröffentlicht: (2025)
von: Reyes-Amezcua, Ivan, et al.
Veröffentlicht: (2025)
Towards Optimal Trade-offs in Knowledge Distillation for CNNs and Vision Transformers at the Edge
von: Violos, John, et al.
Veröffentlicht: (2024)
von: Violos, John, et al.
Veröffentlicht: (2024)
Surface‐aware Mesh Texture Synthesis with Pre‐trained 2D CNNs
von: Áron Samuel Kovács, et al.
Veröffentlicht: (2024)
von: Áron Samuel Kovács, et al.
Veröffentlicht: (2024)
Surface-aware Mesh Texture Synthesis with Pre-trained 2D CNNs
von: Kovács, Áron Samuel, et al.
Veröffentlicht: (2024)
von: Kovács, Áron Samuel, et al.
Veröffentlicht: (2024)
Commitment and Randomization in Communication
von: Kamenica, Emir, et al.
Veröffentlicht: (2024)
von: Kamenica, Emir, et al.
Veröffentlicht: (2024)
Weierstrass Positional Encoding for Vision Transformers
von: Xin, Zhihang, et al.
Veröffentlicht: (2026)
von: Xin, Zhihang, et al.
Veröffentlicht: (2026)
Elastic least‐squares reverse time migration from topography through anisotropic tensorial elastodynamics
von: Tugrul Konuk, et al.
Veröffentlicht: (2024)
von: Tugrul Konuk, et al.
Veröffentlicht: (2024)
VORTEX: Aligning Task Utility and Human Preferences through LLM-Guided Reward Shaping
von: Xiong, Guojun, et al.
Veröffentlicht: (2025)
von: Xiong, Guojun, et al.
Veröffentlicht: (2025)
Trapped acoustic energy and resonances in spherical scatterers
von: Rodrigues, Naruna E., et al.
Veröffentlicht: (2024)
von: Rodrigues, Naruna E., et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
A Comparative Survey of Vision Transformers for Feature Extraction in Texture Analysis
von: Scabini, Leonardo, et al.
Veröffentlicht: (2024) -
Advanced wood species identification based on multiple anatomical sections and using deep feature transfer and fusion
von: Zielinski, Kallil M., et al.
Veröffentlicht: (2024) -
MIXER: Mixed Hyperspherical Random Embedding Neural Network for Texture Recognition
von: Fares, Ricardo T., et al.
Veröffentlicht: (2025) -
APLA: A Simple Adaptation Method for Vision Transformers
von: Sorkhei, Moein, et al.
Veröffentlicht: (2025) -
The Cost of Reasoning: Chain-of-Thought Induces Overconfidence in Vision-Language Models
von: Welch, Robert, et al.
Veröffentlicht: (2026)