Salvato in:
| Autori principali: | Perez, Gustavo, Yu, Stella X. |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | https://arxiv.org/abs/2506.04401 |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Vision Harnessing Agent for Open Ad-hoc Segmentation
di: Wang, Zilin, et al.
Pubblicazione: (2026)
di: Wang, Zilin, et al.
Pubblicazione: (2026)
Co-domain Symmetry for Complex-Valued Deep Learning
di: Singhal, Utkarsh, et al.
Pubblicazione: (2021)
di: Singhal, Utkarsh, et al.
Pubblicazione: (2021)
Learning Normals of Noisy Points by Local Gradient-Aware Surface Filtering
di: Li, Qing, et al.
Pubblicazione: (2025)
di: Li, Qing, et al.
Pubblicazione: (2025)
Free-Grained Hierarchical Visual Recognition
di: Park, Seulki, et al.
Pubblicazione: (2025)
di: Park, Seulki, et al.
Pubblicazione: (2025)
Next-Embedding Prediction Makes Strong Vision Learners
di: Xu, Sihan, et al.
Pubblicazione: (2025)
di: Xu, Sihan, et al.
Pubblicazione: (2025)
Context Matters: Vision-Based Depression Detection Comparing Classical and Deep Approaches
di: Bilalpur, Maneesh, et al.
Pubblicazione: (2026)
di: Bilalpur, Maneesh, et al.
Pubblicazione: (2026)
Design description of Wisdom Computing Persperctive
di: Yu, TianYi
Pubblicazione: (2025)
di: Yu, TianYi
Pubblicazione: (2025)
SHED Light on Segmentation for Dense Prediction
di: Lee, Seung Hyun, et al.
Pubblicazione: (2026)
di: Lee, Seung Hyun, et al.
Pubblicazione: (2026)
Novel View Synthesis from A Few Glimpses via Test-Time Natural Video Completion
di: Xu, Yan, et al.
Pubblicazione: (2025)
di: Xu, Yan, et al.
Pubblicazione: (2025)
GeoSANE: Learning Geospatial Representations from Models, Not Data
di: Hanna, Joelle, et al.
Pubblicazione: (2026)
di: Hanna, Joelle, et al.
Pubblicazione: (2026)
Silicon Minds versus Human Hearts: The Wisdom of Crowds Beats the Wisdom of AI in Emotion Recognition
di: Akben, Mustafa, et al.
Pubblicazione: (2025)
di: Akben, Mustafa, et al.
Pubblicazione: (2025)
The Wisdom of a Crowd of Brains: A Universal Brain Encoder
di: Beliy, Roman, et al.
Pubblicazione: (2024)
di: Beliy, Roman, et al.
Pubblicazione: (2024)
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
di: Wang, Jiayun, et al.
Pubblicazione: (2024)
di: Wang, Jiayun, et al.
Pubblicazione: (2024)
Let Humanoids Hike! Integrative Skill Development on Complex Trails
di: Lin, Kwan-Yee, et al.
Pubblicazione: (2025)
di: Lin, Kwan-Yee, et al.
Pubblicazione: (2025)
Benchmarking Deep Learning and Vision Foundation Models for Atypical vs. Normal Mitosis Classification with Cross-Dataset Evaluation
di: Banerjee, Sweta, et al.
Pubblicazione: (2025)
di: Banerjee, Sweta, et al.
Pubblicazione: (2025)
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
di: Cuttano, Claudia, et al.
Pubblicazione: (2024)
di: Cuttano, Claudia, et al.
Pubblicazione: (2024)
Visually Consistent Hierarchical Image Classification
di: Park, Seulki, et al.
Pubblicazione: (2024)
di: Park, Seulki, et al.
Pubblicazione: (2024)
Beyond Kalman Filters: Deep Learning-Based Filters for Improved Object Tracking
di: Adžemović, Momir, et al.
Pubblicazione: (2024)
di: Adžemović, Momir, et al.
Pubblicazione: (2024)
FViT: A Focal Vision Transformer with Gabor Filter
di: Shi, Yulong, et al.
Pubblicazione: (2024)
di: Shi, Yulong, et al.
Pubblicazione: (2024)
Quantum-enhanced Computer Vision: Going Beyond Classical Algorithms
di: Meli, Natacha Kuete, et al.
Pubblicazione: (2025)
di: Meli, Natacha Kuete, et al.
Pubblicazione: (2025)
Vision Transformer-Based Deep Learning for Histologic Classification of Endometrial Cancer
di: Goyal, Manu, et al.
Pubblicazione: (2023)
di: Goyal, Manu, et al.
Pubblicazione: (2023)
Test-Time Canonicalization by Foundation Models for Robust Perception
di: Singhal, Utkarsh, et al.
Pubblicazione: (2025)
di: Singhal, Utkarsh, et al.
Pubblicazione: (2025)
Learning to Transform for Generalizable Instance-wise Invariance
di: Singhal, Utkarsh, et al.
Pubblicazione: (2023)
di: Singhal, Utkarsh, et al.
Pubblicazione: (2023)
There is More to Attention: Statistical Filtering Enhances Explanations in Vision Transformers
di: Ayyar, Meghna P, et al.
Pubblicazione: (2025)
di: Ayyar, Meghna P, et al.
Pubblicazione: (2025)
Low-Pass Filtering Improves Behavioral Alignment of Vision Models
di: Wolff, Max, et al.
Pubblicazione: (2026)
di: Wolff, Max, et al.
Pubblicazione: (2026)
The Master Key Filters Hypothesis: Deep Filters Are General
di: Babaiee, Zahra, et al.
Pubblicazione: (2024)
di: Babaiee, Zahra, et al.
Pubblicazione: (2024)
MVFormer: Diversifying Feature Normalization and Token Mixing for Efficient Vision Transformers
di: Bae, Jongseong, et al.
Pubblicazione: (2024)
di: Bae, Jongseong, et al.
Pubblicazione: (2024)
Are you In or Out (of gallery)? Wisdom from the Same-Identity Crowd
di: Bhatta, Aman, et al.
Pubblicazione: (2025)
di: Bhatta, Aman, et al.
Pubblicazione: (2025)
Sampling Strategies based on Wisdom of Crowds for Amazon Deforestation Detection
di: Resende, Hugo, et al.
Pubblicazione: (2024)
di: Resende, Hugo, et al.
Pubblicazione: (2024)
Speed-up of Vision Transformer Models by Attention-aware Token Filtering
di: Naruko, Takahiro, et al.
Pubblicazione: (2025)
di: Naruko, Takahiro, et al.
Pubblicazione: (2025)
IDTrust: Deep Identity Document Quality Detection with Bandpass Filtering
di: Al-Ghadi, Musab, et al.
Pubblicazione: (2024)
di: Al-Ghadi, Musab, et al.
Pubblicazione: (2024)
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking
di: Feng, X., et al.
Pubblicazione: (2025)
di: Feng, X., et al.
Pubblicazione: (2025)
ProReason: Multi-Modal Proactive Reasoning with Decoupled Eyesight and Wisdom
di: Zhou, Jingqi, et al.
Pubblicazione: (2024)
di: Zhou, Jingqi, et al.
Pubblicazione: (2024)
Hallucination Filtering in Radiology Vision-Language Models Using Discrete Semantic Entropy
di: Wienholt, Patrick, et al.
Pubblicazione: (2025)
di: Wienholt, Patrick, et al.
Pubblicazione: (2025)
Taxonomy-Aware Evaluation of Vision-Language Models
di: Snæbjarnarson, Vésteinn, et al.
Pubblicazione: (2025)
di: Snæbjarnarson, Vésteinn, et al.
Pubblicazione: (2025)
Bridging Classical and Modern Computer Vision: PerceptiveNet for Tree Crown Semantic Segmentation
di: Voulgaris, Georgios
Pubblicazione: (2025)
di: Voulgaris, Georgios
Pubblicazione: (2025)
FGFP: A Fractional Gaussian Filter and Pruning for Deep Neural Networks Compression
di: Tu, Kuan-Ting, et al.
Pubblicazione: (2025)
di: Tu, Kuan-Ting, et al.
Pubblicazione: (2025)
Open Ad-hoc Categorization with Contextualized Feature Learning
di: Wang, Zilin, et al.
Pubblicazione: (2025)
di: Wang, Zilin, et al.
Pubblicazione: (2025)
Exploring the Efficacy of Group-Normalization in Deep Learning Models for Alzheimer's Disease Classification
di: Habib, Gousia, et al.
Pubblicazione: (2024)
di: Habib, Gousia, et al.
Pubblicazione: (2024)
Normal and Abnormal Pathology Knowledge-Augmented Vision-Language Model for Anomaly Detection in Pathology Images
di: Song, Jinsol, et al.
Pubblicazione: (2025)
di: Song, Jinsol, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Vision Harnessing Agent for Open Ad-hoc Segmentation
di: Wang, Zilin, et al.
Pubblicazione: (2026) -
Co-domain Symmetry for Complex-Valued Deep Learning
di: Singhal, Utkarsh, et al.
Pubblicazione: (2021) -
Learning Normals of Noisy Points by Local Gradient-Aware Surface Filtering
di: Li, Qing, et al.
Pubblicazione: (2025) -
Free-Grained Hierarchical Visual Recognition
di: Park, Seulki, et al.
Pubblicazione: (2025) -
Next-Embedding Prediction Makes Strong Vision Learners
di: Xu, Sihan, et al.
Pubblicazione: (2025)