Pointing-Based Object Recognition
Fuente:
arXiv
Salvato in:
| Autori principali: | Hajdúch, Lukáš, Kocur, Viktor |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
di: Masrourisaadat, Nila, et al.
Pubblicazione: (2024)
di: Masrourisaadat, Nila, et al.
Pubblicazione: (2024)
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
di: Semenov, Andrei, et al.
Pubblicazione: (2024)
Learning 3D object-centric representation through prediction
di: Day, John, et al.
Pubblicazione: (2024)
di: Day, John, et al.
Pubblicazione: (2024)
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
di: Shahin, Nada, et al.
Pubblicazione: (2025)
di: Shahin, Nada, et al.
Pubblicazione: (2025)
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
di: Adiban, Mohammad, et al.
Pubblicazione: (2023)
di: Adiban, Mohammad, et al.
Pubblicazione: (2023)
Neuromorphic Monocular Depth Estimation with Uncertainty Modeling
di: Bergkvist, Viktor, et al.
Pubblicazione: (2026)
di: Bergkvist, Viktor, et al.
Pubblicazione: (2026)
ADAT: Time-Series-Aware Adaptive Transformer Architecture for Sign Language Translation
di: Shahin, Nada, et al.
Pubblicazione: (2025)
di: Shahin, Nada, et al.
Pubblicazione: (2025)
Adaptive Self-Training for Object Detection
di: Vandeghen, Renaud, et al.
Pubblicazione: (2022)
di: Vandeghen, Renaud, et al.
Pubblicazione: (2022)
Quantized Vision-Language Models for Damage Assessment: A Comparative Study of LLaVA-1.5-7B Quantization Levels
di: Yasuno, Takato
Pubblicazione: (2026)
di: Yasuno, Takato
Pubblicazione: (2026)
NV3D: Leveraging Spatial Shape Through Normal Vector-based 3D Object Detection
di: Chaowakarn, Krittin, et al.
Pubblicazione: (2025)
di: Chaowakarn, Krittin, et al.
Pubblicazione: (2025)
Selection, Not Fusion: Radar-Modulated State Space Models for Radar-Camera Depth Estimation
di: Hou, Zhangcheng, et al.
Pubblicazione: (2026)
di: Hou, Zhangcheng, et al.
Pubblicazione: (2026)
Once-For-All: A Train-Once and Select-Anytime Framework for Multimodal Instruction Tuning
di: Dong, Mingkang, et al.
Pubblicazione: (2026)
di: Dong, Mingkang, et al.
Pubblicazione: (2026)
Exposing Blindspots: Cultural Bias Evaluation in Generative Image Models
di: Seo, Huichan, et al.
Pubblicazione: (2025)
di: Seo, Huichan, et al.
Pubblicazione: (2025)
Smooth regularization for efficient video recognition
di: Goldman, Gil, et al.
Pubblicazione: (2025)
di: Goldman, Gil, et al.
Pubblicazione: (2025)
FAME: Feature Activation Map Explanation on Image Classification and Face Recognition
di: Zhang, Xinyi, et al.
Pubblicazione: (2026)
di: Zhang, Xinyi, et al.
Pubblicazione: (2026)
Differentiable Hierarchical Visual Tokenization
di: Aasan, Marius, et al.
Pubblicazione: (2025)
di: Aasan, Marius, et al.
Pubblicazione: (2025)
Feature-Augmented Deep Networks for Multiscale Building Segmentation in High-Resolution UAV and Satellite Imagery
di: Maniyar, Chintan B., et al.
Pubblicazione: (2025)
di: Maniyar, Chintan B., et al.
Pubblicazione: (2025)
FeedbackSTS-Det: Sparse Frames-Based Spatio-Temporal Semantic Feedback Network for Moving Infrared Small Target Detection
di: Huang, Yian, et al.
Pubblicazione: (2026)
di: Huang, Yian, et al.
Pubblicazione: (2026)
Sat-JEPA-Diff: Bridging Self-Supervised Learning and Generative Diffusion for Remote Sensing
di: Komurcu, Kursat, et al.
Pubblicazione: (2026)
di: Komurcu, Kursat, et al.
Pubblicazione: (2026)
Uncertainty quantification for White Matter Hyperintensity segmentation detects silent failures and improves automated Fazekas quantification
di: Philps, Ben, et al.
Pubblicazione: (2024)
di: Philps, Ben, et al.
Pubblicazione: (2024)
Cora: Correspondence-aware image editing using few step diffusion
di: Alimohammadi, Amirhossein, et al.
Pubblicazione: (2025)
di: Alimohammadi, Amirhossein, et al.
Pubblicazione: (2025)
Lost in Latent Space: Disentangled Models and the Challenge of Combinatorial Generalisation
di: Montero, Milton L., et al.
Pubblicazione: (2022)
di: Montero, Milton L., et al.
Pubblicazione: (2022)
Towards Onboard Continuous Change Detection for Floods
di: Kyselica, Daniel, et al.
Pubblicazione: (2026)
di: Kyselica, Daniel, et al.
Pubblicazione: (2026)
CCVA-FL: Cross-Client Variations Adaptive Federated Learning for Medical Imaging
di: Gupta, Sunny, et al.
Pubblicazione: (2024)
di: Gupta, Sunny, et al.
Pubblicazione: (2024)
Taming the Tail: Leveraging Asymmetric Loss and Pade Approximation to Overcome Medical Image Long-Tailed Class Imbalance
di: Kashyap, Pankhi, et al.
Pubblicazione: (2024)
di: Kashyap, Pankhi, et al.
Pubblicazione: (2024)
MaSC: A Masked Similarity Metric for Evaluating Concept-Driven Generation
di: Bartkowiak, Patryk, et al.
Pubblicazione: (2026)
di: Bartkowiak, Patryk, et al.
Pubblicazione: (2026)
SERA-H: Beyond Native Sentinel Spatial Limits for High-Resolution Canopy Height Mapping
di: Boudras, Thomas, et al.
Pubblicazione: (2025)
di: Boudras, Thomas, et al.
Pubblicazione: (2025)
Data Organization Matters in Multimodal Instruction Tuning: A Controlled Study of Capability Trade-offs
di: Tang, Guowei
Pubblicazione: (2026)
di: Tang, Guowei
Pubblicazione: (2026)
Training a Student Expert via Semi-Supervised Foundation Model Distillation
di: Taghavi, Pardis, et al.
Pubblicazione: (2026)
di: Taghavi, Pardis, et al.
Pubblicazione: (2026)
Mitigating Catastrophic Forgetting in the Incremental Learning of Medical Images
di: Yavari, Sara, et al.
Pubblicazione: (2025)
di: Yavari, Sara, et al.
Pubblicazione: (2025)
ReLKD: Inter-Class Relation Learning with Knowledge Distillation for Generalized Category Discovery
di: Zhou, Fang, et al.
Pubblicazione: (2025)
di: Zhou, Fang, et al.
Pubblicazione: (2025)
Efficient Image Pre-Training with Siamese Cropped Masked Autoencoders
di: Eymaël, Alexandre, et al.
Pubblicazione: (2024)
di: Eymaël, Alexandre, et al.
Pubblicazione: (2024)
RDPO: Real Data Preference Optimization for Physics Consistency Video Generation
di: Qian, Wenxu, et al.
Pubblicazione: (2025)
di: Qian, Wenxu, et al.
Pubblicazione: (2025)
CLIP-Joint-Detect: End-to-End Joint Training of Object Detectors with Contrastive Vision-Language Supervision
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
di: Raoufi, Behnam, et al.
Pubblicazione: (2025)
Supervised Contrastive Learning for Few-Shot AI-Generated Image Detection and Attribution
di: Urueña, Jaime Álvarez, et al.
Pubblicazione: (2025)
di: Urueña, Jaime Álvarez, et al.
Pubblicazione: (2025)
HATL: Hierarchical Adaptive-Transfer Learning Framework for Sign Language Machine Translation
di: Shahin, Nada, et al.
Pubblicazione: (2026)
di: Shahin, Nada, et al.
Pubblicazione: (2026)
Butter: Frequency Consistency and Hierarchical Fusion for Autonomous Driving Object Detection
di: Lin, Xiaojian, et al.
Pubblicazione: (2025)
di: Lin, Xiaojian, et al.
Pubblicazione: (2025)
Pointing-Guided Target Estimation via Transformer-Based Attention
di: Müller, Luca, et al.
Pubblicazione: (2025)
di: Müller, Luca, et al.
Pubblicazione: (2025)
Decoupling Vision and Language: Codebook Anchored Visual Adaptation
di: Wu, Jason, et al.
Pubblicazione: (2026)
di: Wu, Jason, et al.
Pubblicazione: (2026)
STimage-1K4M: A histopathology image-gene expression dataset for spatial transcriptomics
di: Chen, Jiawen, et al.
Pubblicazione: (2024)
di: Chen, Jiawen, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Analyzing Quality, Bias, and Performance in Text-to-Image Generative Models
di: Masrourisaadat, Nila, et al.
Pubblicazione: (2024) -
Sparse Concept Bottleneck Models: Gumbel Tricks in Contrastive Learning
di: Semenov, Andrei, et al.
Pubblicazione: (2024) -
Learning 3D object-centric representation through prediction
di: Day, John, et al.
Pubblicazione: (2024) -
GLoT: A Novel Gated-Logarithmic Transformer for Efficient Sign Language Translation
di: Shahin, Nada, et al.
Pubblicazione: (2025) -
S-HR-VQVAE: Sequential Hierarchical Residual Learning Vector Quantized Variational Autoencoder for Video Prediction
di: Adiban, Mohammad, et al.
Pubblicazione: (2023)