Guardado en:
| Autores principales: | Perez, Gustavo, Yu, Stella X. |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2506.04401 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Vision Harnessing Agent for Open Ad-hoc Segmentation
por: Wang, Zilin, et al.
Publicado: (2026)
por: Wang, Zilin, et al.
Publicado: (2026)
Co-domain Symmetry for Complex-Valued Deep Learning
por: Singhal, Utkarsh, et al.
Publicado: (2021)
por: Singhal, Utkarsh, et al.
Publicado: (2021)
Learning Normals of Noisy Points by Local Gradient-Aware Surface Filtering
por: Li, Qing, et al.
Publicado: (2025)
por: Li, Qing, et al.
Publicado: (2025)
Free-Grained Hierarchical Visual Recognition
por: Park, Seulki, et al.
Publicado: (2025)
por: Park, Seulki, et al.
Publicado: (2025)
Next-Embedding Prediction Makes Strong Vision Learners
por: Xu, Sihan, et al.
Publicado: (2025)
por: Xu, Sihan, et al.
Publicado: (2025)
Context Matters: Vision-Based Depression Detection Comparing Classical and Deep Approaches
por: Bilalpur, Maneesh, et al.
Publicado: (2026)
por: Bilalpur, Maneesh, et al.
Publicado: (2026)
Design description of Wisdom Computing Persperctive
por: Yu, TianYi
Publicado: (2025)
por: Yu, TianYi
Publicado: (2025)
SHED Light on Segmentation for Dense Prediction
por: Lee, Seung Hyun, et al.
Publicado: (2026)
por: Lee, Seung Hyun, et al.
Publicado: (2026)
Novel View Synthesis from A Few Glimpses via Test-Time Natural Video Completion
por: Xu, Yan, et al.
Publicado: (2025)
por: Xu, Yan, et al.
Publicado: (2025)
GeoSANE: Learning Geospatial Representations from Models, Not Data
por: Hanna, Joelle, et al.
Publicado: (2026)
por: Hanna, Joelle, et al.
Publicado: (2026)
Silicon Minds versus Human Hearts: The Wisdom of Crowds Beats the Wisdom of AI in Emotion Recognition
por: Akben, Mustafa, et al.
Publicado: (2025)
por: Akben, Mustafa, et al.
Publicado: (2025)
The Wisdom of a Crowd of Brains: A Universal Brain Encoder
por: Beliy, Roman, et al.
Publicado: (2024)
por: Beliy, Roman, et al.
Publicado: (2024)
Pose-Aware Self-Supervised Learning with Viewpoint Trajectory Regularization
por: Wang, Jiayun, et al.
Publicado: (2024)
por: Wang, Jiayun, et al.
Publicado: (2024)
Let Humanoids Hike! Integrative Skill Development on Complex Trails
por: Lin, Kwan-Yee, et al.
Publicado: (2025)
por: Lin, Kwan-Yee, et al.
Publicado: (2025)
Benchmarking Deep Learning and Vision Foundation Models for Atypical vs. Normal Mitosis Classification with Cross-Dataset Evaluation
por: Banerjee, Sweta, et al.
Publicado: (2025)
por: Banerjee, Sweta, et al.
Publicado: (2025)
SAMWISE: Infusing Wisdom in SAM2 for Text-Driven Video Segmentation
por: Cuttano, Claudia, et al.
Publicado: (2024)
por: Cuttano, Claudia, et al.
Publicado: (2024)
Visually Consistent Hierarchical Image Classification
por: Park, Seulki, et al.
Publicado: (2024)
por: Park, Seulki, et al.
Publicado: (2024)
Beyond Kalman Filters: Deep Learning-Based Filters for Improved Object Tracking
por: Adžemović, Momir, et al.
Publicado: (2024)
por: Adžemović, Momir, et al.
Publicado: (2024)
FViT: A Focal Vision Transformer with Gabor Filter
por: Shi, Yulong, et al.
Publicado: (2024)
por: Shi, Yulong, et al.
Publicado: (2024)
Quantum-enhanced Computer Vision: Going Beyond Classical Algorithms
por: Meli, Natacha Kuete, et al.
Publicado: (2025)
por: Meli, Natacha Kuete, et al.
Publicado: (2025)
Vision Transformer-Based Deep Learning for Histologic Classification of Endometrial Cancer
por: Goyal, Manu, et al.
Publicado: (2023)
por: Goyal, Manu, et al.
Publicado: (2023)
Test-Time Canonicalization by Foundation Models for Robust Perception
por: Singhal, Utkarsh, et al.
Publicado: (2025)
por: Singhal, Utkarsh, et al.
Publicado: (2025)
Learning to Transform for Generalizable Instance-wise Invariance
por: Singhal, Utkarsh, et al.
Publicado: (2023)
por: Singhal, Utkarsh, et al.
Publicado: (2023)
There is More to Attention: Statistical Filtering Enhances Explanations in Vision Transformers
por: Ayyar, Meghna P, et al.
Publicado: (2025)
por: Ayyar, Meghna P, et al.
Publicado: (2025)
Low-Pass Filtering Improves Behavioral Alignment of Vision Models
por: Wolff, Max, et al.
Publicado: (2026)
por: Wolff, Max, et al.
Publicado: (2026)
The Master Key Filters Hypothesis: Deep Filters Are General
por: Babaiee, Zahra, et al.
Publicado: (2024)
por: Babaiee, Zahra, et al.
Publicado: (2024)
MVFormer: Diversifying Feature Normalization and Token Mixing for Efficient Vision Transformers
por: Bae, Jongseong, et al.
Publicado: (2024)
por: Bae, Jongseong, et al.
Publicado: (2024)
Are you In or Out (of gallery)? Wisdom from the Same-Identity Crowd
por: Bhatta, Aman, et al.
Publicado: (2025)
por: Bhatta, Aman, et al.
Publicado: (2025)
Sampling Strategies based on Wisdom of Crowds for Amazon Deforestation Detection
por: Resende, Hugo, et al.
Publicado: (2024)
por: Resende, Hugo, et al.
Publicado: (2024)
Speed-up of Vision Transformer Models by Attention-aware Token Filtering
por: Naruko, Takahiro, et al.
Publicado: (2025)
por: Naruko, Takahiro, et al.
Publicado: (2025)
IDTrust: Deep Identity Document Quality Detection with Bandpass Filtering
por: Al-Ghadi, Musab, et al.
Publicado: (2024)
por: Al-Ghadi, Musab, et al.
Publicado: (2024)
ATCTrack: Aligning Target-Context Cues with Dynamic Target States for Robust Vision-Language Tracking
por: Feng, X., et al.
Publicado: (2025)
por: Feng, X., et al.
Publicado: (2025)
ProReason: Multi-Modal Proactive Reasoning with Decoupled Eyesight and Wisdom
por: Zhou, Jingqi, et al.
Publicado: (2024)
por: Zhou, Jingqi, et al.
Publicado: (2024)
Hallucination Filtering in Radiology Vision-Language Models Using Discrete Semantic Entropy
por: Wienholt, Patrick, et al.
Publicado: (2025)
por: Wienholt, Patrick, et al.
Publicado: (2025)
Taxonomy-Aware Evaluation of Vision-Language Models
por: Snæbjarnarson, Vésteinn, et al.
Publicado: (2025)
por: Snæbjarnarson, Vésteinn, et al.
Publicado: (2025)
Bridging Classical and Modern Computer Vision: PerceptiveNet for Tree Crown Semantic Segmentation
por: Voulgaris, Georgios
Publicado: (2025)
por: Voulgaris, Georgios
Publicado: (2025)
FGFP: A Fractional Gaussian Filter and Pruning for Deep Neural Networks Compression
por: Tu, Kuan-Ting, et al.
Publicado: (2025)
por: Tu, Kuan-Ting, et al.
Publicado: (2025)
Open Ad-hoc Categorization with Contextualized Feature Learning
por: Wang, Zilin, et al.
Publicado: (2025)
por: Wang, Zilin, et al.
Publicado: (2025)
Exploring the Efficacy of Group-Normalization in Deep Learning Models for Alzheimer's Disease Classification
por: Habib, Gousia, et al.
Publicado: (2024)
por: Habib, Gousia, et al.
Publicado: (2024)
Normal and Abnormal Pathology Knowledge-Augmented Vision-Language Model for Anomaly Detection in Pathology Images
por: Song, Jinsol, et al.
Publicado: (2025)
por: Song, Jinsol, et al.
Publicado: (2025)
Ejemplares similares
-
Vision Harnessing Agent for Open Ad-hoc Segmentation
por: Wang, Zilin, et al.
Publicado: (2026) -
Co-domain Symmetry for Complex-Valued Deep Learning
por: Singhal, Utkarsh, et al.
Publicado: (2021) -
Learning Normals of Noisy Points by Local Gradient-Aware Surface Filtering
por: Li, Qing, et al.
Publicado: (2025) -
Free-Grained Hierarchical Visual Recognition
por: Park, Seulki, et al.
Publicado: (2025) -
Next-Embedding Prediction Makes Strong Vision Learners
por: Xu, Sihan, et al.
Publicado: (2025)