Saved in:
| Main Authors: | Kapon, Danielle, Fire, Michael, Gordin, Shai |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2406.04039 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
by: Yasuno, Takato
Published: (2026)
by: Yasuno, Takato
Published: (2026)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
by: Semenov, Andrei, et al.
Published: (2024)
by: Semenov, Andrei, et al.
Published: (2024)
Symmetry Awareness Encoded Deep Learning Framework for Brain Imaging Analysis
by: Ma, Yang, et al.
Published: (2024)
by: Ma, Yang, et al.
Published: (2024)
Supervised Contrastive Learning for Few-Shot AI-Generated Image Detection and Attribution
by: Urueña, Jaime Álvarez, et al.
Published: (2025)
by: Urueña, Jaime Álvarez, et al.
Published: (2025)
Cora: Correspondence-aware image editing using few step diffusion
by: Alimohammadi, Amirhossein, et al.
Published: (2025)
by: Alimohammadi, Amirhossein, et al.
Published: (2025)
Pointing-Based Object Recognition
by: Hajdúch, Lukáš, et al.
Published: (2026)
by: Hajdúch, Lukáš, et al.
Published: (2026)
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
by: Zhang, Xinyi, et al.
Published: (2026)
by: Zhang, Xinyi, et al.
Published: (2026)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
by: Raoufi, Behnam, et al.
Published: (2025)
by: Raoufi, Behnam, et al.
Published: (2025)
When in Doubt, Think Slow: Iterative Reasoning with Latent Imagination
by: Benfeghoul, Martin, et al.
Published: (2024)
by: Benfeghoul, Martin, et al.
Published: (2024)
Differentiable Hierarchical Visual Tokenization
by: Aasan, Marius, et al.
Published: (2025)
by: Aasan, Marius, et al.
Published: (2025)
Feature-Augmented Deep Networks for Multiscale Building Segmentation in High-Resolution UAV and Satellite Imagery
by: Maniyar, Chintan B., et al.
Published: (2025)
by: Maniyar, Chintan B., et al.
Published: (2025)
STimage-1K4M: A histopathology image-gene expression dataset for spatial transcriptomics
by: Chen, Jiawen, et al.
Published: (2024)
by: Chen, Jiawen, et al.
Published: (2024)
Scalable Face Security Vision Foundation Model for Deepfake, Diffusion, and Spoofing Detection
by: Wang, Gaojian, et al.
Published: (2025)
by: Wang, Gaojian, et al.
Published: (2025)
Towards Onboard Continuous Change Detection for Floods
by: Kyselica, Daniel, et al.
Published: (2026)
by: Kyselica, Daniel, et al.
Published: (2026)
Non-Robust Features are Not Always Useful in One-Class Classification
by: Lau, Matthew, et al.
Published: (2024)
by: Lau, Matthew, et al.
Published: (2024)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
by: Gupta, Sunny, et al.
Published: (2024)
by: Gupta, Sunny, et al.
Published: (2024)
EDSNet: Efficient-DSNet for Video Summarization
by: Prasad, Ashish, et al.
Published: (2024)
by: Prasad, Ashish, et al.
Published: (2024)
Learning 3D object-centric representation through prediction
by: Day, John, et al.
Published: (2024)
by: Day, John, et al.
Published: (2024)
DUDF: Differentiable Unsigned Distance Fields with Hyperbolic Scaling
by: Fainstein, Miguel, et al.
Published: (2024)
by: Fainstein, Miguel, et al.
Published: (2024)
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
by: Adiban, Mohammad, et al.
Published: (2023)
by: Adiban, Mohammad, et al.
Published: (2023)
GeoPos: A Minimal Positional Encoding for Enhanced Fine-Grained Details in Image Synthesis Using Convolutional Neural Networks
by: Hosseini, Mehran, et al.
Published: (2024)
by: Hosseini, Mehran, et al.
Published: (2024)
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
by: Masrourisaadat, Nila, et al.
Published: (2024)
by: Masrourisaadat, Nila, et al.
Published: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
by: Kashyap, Pankhi, et al.
Published: (2024)
by: Kashyap, Pankhi, et al.
Published: (2024)
Visual-Text Cross Alignment: Refining the Similarity Score in Vision-Language Models
by: Li, Jinhao, et al.
Published: (2024)
by: Li, Jinhao, et al.
Published: (2024)
WaveMix: A Resource-efficient Neural Network for Image Analysis
by: Jeevan, Pranav, et al.
Published: (2022)
by: Jeevan, Pranav, et al.
Published: (2022)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
by: Shahin, Nada, et al.
Published: (2025)
by: Shahin, Nada, et al.
Published: (2025)
NV3D: Leveraging Spatial Shape Through Normal Vector-based 3D Object Detection
by: Chaowakarn, Krittin, et al.
Published: (2025)
by: Chaowakarn, Krittin, et al.
Published: (2025)
Hierarchical Multi-Positive Contrastive Learning for Patent Image Retrieval
by: Kavimandan, Kshitij, et al.
Published: (2025)
by: Kavimandan, Kshitij, et al.
Published: (2025)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
by: Komurcu, Kursat, et al.
Published: (2026)
by: Komurcu, Kursat, et al.
Published: (2026)
Gaussian Splatting: 3D Reconstruction and Novel View Synthesis, a Review
by: Dalal, Anurag, et al.
Published: (2024)
by: Dalal, Anurag, et al.
Published: (2024)
Fusing Structure from Motion and Simulation-Augmented Pose Regression from Optical Flow for Challenging Indoor Environments
by: Ott, Felix, et al.
Published: (2023)
by: Ott, Felix, et al.
Published: (2023)
HATL: Hierarchical Adaptive-Transfer Learning Framework for Sign Language Machine Translation
by: Shahin, Nada, et al.
Published: (2026)
by: Shahin, Nada, et al.
Published: (2026)
HOSC: A Periodic Activation Function for Preserving Sharp Features in Implicit Neural Representations
by: Serrano, Danzel, et al.
Published: (2024)
by: Serrano, Danzel, et al.
Published: (2024)
BG-YOLO: A Bidirectional-Guided Method for Underwater Object Detection
by: Zhang, Jian, et al.
Published: (2024)
by: Zhang, Jian, et al.
Published: (2024)
InterChart: Benchmarking Visual Reasoning Across Decomposed and Distributed Chart Information
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025)
by: Iyengar, Anirudh Iyengar Kaniyar Narayana, et al.
Published: (2025)
A Spitting Image: Modular Superpixel Tokenization in Vision Transformers
by: Aasan, Marius, et al.
Published: (2024)
by: Aasan, Marius, et al.
Published: (2024)
Which Backbone to Use: A Resource-efficient Domain Specific Comparison for Computer Vision
by: Jeevan, Pranav, et al.
Published: (2024)
by: Jeevan, Pranav, et al.
Published: (2024)
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation
by: Shahin, Nada, et al.
Published: (2025)
by: Shahin, Nada, et al.
Published: (2025)
GLEaN: A Text-to-image Bias Detection Approach for Public Comprehension
by: Ding, Bochu, et al.
Published: (2026)
by: Ding, Bochu, et al.
Published: (2026)
FOCUS on Contamination: Hydrology-Informed Noise-Aware Learning for Geospatial PFAS Mapping
by: Khan, Jowaria, et al.
Published: (2025)
by: Khan, Jowaria, et al.
Published: (2025)
Similar Items
-
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
by: Yasuno, Takato
Published: (2026) -
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
by: Semenov, Andrei, et al.
Published: (2024) -
Symmetry Awareness Encoded Deep Learning Framework for Brain Imaging Analysis
by: Ma, Yang, et al.
Published: (2024) -
Supervised Contrastive Learning for Few-Shot AI-Generated Image Detection and Attribution
by: Urueña, Jaime Álvarez, et al.
Published: (2025) -
Cora: Correspondence-aware image editing using few step diffusion
by: Alimohammadi, Amirhossein, et al.
Published: (2025)