Differentiable Hierarchical Visual Tokenization
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Aasan, Marius, Hjelkrem-Tan, Martine, Catalano, Nico, Choi, Changkyu, Rivera, Adín Ramírez |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
von: Aasan, Marius, et al.
Veröffentlicht: (2024)
ROI-NeRFs: Hi-Fi Visualization of Objects of Interest within a Scene by NeRFs Composition
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
Graph-PiT: Enhancing Structural Coherence in Part-Based Image Synthesis via Graph Priors
von: Zhang, Junbin, et al.
Veröffentlicht: (2026)
von: Zhang, Junbin, et al.
Veröffentlicht: (2026)
Towards Onboard Continuous Change Detection for Floods
von: Kyselica, Daniel, et al.
Veröffentlicht: (2026)
von: Kyselica, Daniel, et al.
Veröffentlicht: (2026)
Non-Robust Features are Not Always Useful in One-Class Classification
von: Lau, Matthew, et al.
Veröffentlicht: (2024)
von: Lau, Matthew, et al.
Veröffentlicht: (2024)
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
von: Li, Jinhao, et al.
Veröffentlicht: (2024)
ROI-GS: Interest-based Local Quality 3D Gaussian Splatting
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
von: Zhang, Jian, et al.
Veröffentlicht: (2024)
Gaussian Splatting: 3D Reconstruction and Novel View Synthesis, a Review
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
von: Dalal, Anurag, et al.
Veröffentlicht: (2024)
Semantic Prioritization in Visual Counterfactual Explanations with Weighted Segmentation and Auto-Adaptive Region Selection
von: Zhang, Lintong, et al.
Veröffentlicht: (2025)
von: Zhang, Lintong, et al.
Veröffentlicht: (2025)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
von: Komurcu, Kursat, et al.
Veröffentlicht: (2026)
Parking Space Detection in the City of Granada
von: Luis, Crespo-Orti, et al.
Veröffentlicht: (2025)
von: Luis, Crespo-Orti, et al.
Veröffentlicht: (2025)
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
von: Benfeghoul, Martin, et al.
Veröffentlicht: (2024)
Feature-Augmented Deep Networks for Multiscale Building Segmentation in High-Resolution UAV and Satellite Imagery
von: Maniyar, Chintan B., et al.
Veröffentlicht: (2025)
von: Maniyar, Chintan B., et al.
Veröffentlicht: (2025)
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
von: Adiban, Mohammad, et al.
Veröffentlicht: (2023)
von: Adiban, Mohammad, et al.
Veröffentlicht: (2023)
Training a Student Expert via Semi-Supervised Foundation Model Distillation
von: Taghavi, Pardis, et al.
Veröffentlicht: (2026)
von: Taghavi, Pardis, et al.
Veröffentlicht: (2026)
PCRI: Measuring Context Robustness in Multimodal Models for Enterprise Applications
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
von: Patel, Hitesh Laxmichand, et al.
Veröffentlicht: (2025)
Advancing Brain Tumor Segmentation via Attention-based 3D U-Net Architecture and Digital Image Processing
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
von: Gad, Eyad, et al.
Veröffentlicht: (2025)
Consistent Zero-shot 3D Texture Synthesis Using Geometry-aware Diffusion and Temporal Video Models
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
von: Kang, Donggoo, et al.
Veröffentlicht: (2025)
FLD+: Data-efficient Evaluation Metric for Generative Models
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
WaveMixSR-V2: Enhancing Super-resolution with Higher Efficiency
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Normalizing Flow-Based Metric for Image Generation
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
TexTile: A Differentiable Metric for Texture Tileability
von: Rodriguez-Pardo, Carlos, et al.
Veröffentlicht: (2024)
von: Rodriguez-Pardo, Carlos, et al.
Veröffentlicht: (2024)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
von: Raoufi, Behnam, et al.
Veröffentlicht: (2025)
A Landmark-Aware Visual Navigation Dataset
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
von: Johnson, Faith, et al.
Veröffentlicht: (2024)
Learning 3D object-centric representation through prediction
von: Day, John, et al.
Veröffentlicht: (2024)
von: Day, John, et al.
Veröffentlicht: (2024)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
von: Gupta, Sunny, et al.
Veröffentlicht: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
von: Kashyap, Pankhi, et al.
Veröffentlicht: (2024)
μ-Net: A Deep Learning-Based Architecture for μ-CT Segmentation
von: Bruno, Pierangela, et al.
Veröffentlicht: (2024)
von: Bruno, Pierangela, et al.
Veröffentlicht: (2024)
Learning Association via Track-Detection Matching for Multi-Object Tracking
von: Adžemović, Momir
Veröffentlicht: (2025)
von: Adžemović, Momir
Veröffentlicht: (2025)
Smelly, dense, and spreaded: The Object Detection for Olfactory References (ODOR) dataset
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
von: Zinnen, Mathias, et al.
Veröffentlicht: (2025)
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
von: Yasuno, Takato
Veröffentlicht: (2026)
von: Yasuno, Takato
Veröffentlicht: (2026)
Addressing Issues with Working Memory in Video Object Segmentation
von: Bromley, Clayton, et al.
Veröffentlicht: (2024)
von: Bromley, Clayton, et al.
Veröffentlicht: (2024)
WaveMix: A Resource-efficient Neural Network for Image Analysis
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2022)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
von: Jeevan, Pranav, et al.
Veröffentlicht: (2024)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
von: Semenov, Andrei, et al.
Veröffentlicht: (2024)
Hierarchical Spatial Algorithms for High-Resolution Image Quantization and Feature Extraction
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
von: Mohammad, Noor Islam S.
Veröffentlicht: (2025)
JVLGS: Joint Vision-Language Gas Leak Segmentation
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
von: Zhao, Xinlong, et al.
Veröffentlicht: (2025)
Fusing Structure from Motion and Simulation-Augmented Pose Regression from Optical Flow for Challenging Indoor Environments
von: Ott, Felix, et al.
Veröffentlicht: (2023)
von: Ott, Felix, et al.
Veröffentlicht: (2023)
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
von: Zhang, Xinyi, et al.
Veröffentlicht: (2026)
von: Zhang, Xinyi, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
von: Aasan, Marius, et al.
Veröffentlicht: (2024) -
ROI-NeRFs: Hi-Fi Visualization of Objects of Interest within a Scene by NeRFs Composition
von: Bui, Quoc-Anh, et al.
Veröffentlicht: (2025) -
Graph-PiT: Enhancing Structural Coherence in Part-Based Image Synthesis via Graph Priors
von: Zhang, Junbin, et al.
Veröffentlicht: (2026) -
Towards Onboard Continuous Change Detection for Floods
von: Kyselica, Daniel, et al.
Veröffentlicht: (2026) -
Non-Robust Features are Not Always Useful in One-Class Classification
von: Lau, Matthew, et al.
Veröffentlicht: (2024)