This Looks Distinctly Like That: Grounding Interpretable Recognition in Stiefel Geometry against Neural Collapse
Fuente:
arXiv
Saved in:
| Main Authors: | Jia, Junhao, Wang, Jiaqi, Liu, Yunyou, Jing, Haodong, Wu, Yueyi, Wu, Xian, Zheng, Yefeng |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unsupervised Causal Prototypical Networks for De-biased Interpretable Dermoscopy Diagnosis
by: Jia, Junhao, et al.
Published: (2026)
by: Jia, Junhao, et al.
Published: (2026)
Geodesic Prototype Matching via Diffusion Maps for Interpretable Fine-Grained Recognition
by: Jia, Junhao, et al.
Published: (2025)
by: Jia, Junhao, et al.
Published: (2025)
Reusing Fusion-Time Spectral Reliability for Adaptive Fusion and Expert Routing in RGB-Infrared Object Detection
by: Wu, Yefeng
Published: (2026)
by: Wu, Yefeng
Published: (2026)
Brain-HGCN: A Hyperbolic Graph Convolutional Network for Brain Functional Network Analysis
by: Jia, Junhao, et al.
Published: (2025)
by: Jia, Junhao, et al.
Published: (2025)
BrainSegDMlF: A Dynamic Fusion-enhanced SAM for Brain Lesion Segmentation
by: Wang, Hongming, et al.
Published: (2025)
by: Wang, Hongming, et al.
Published: (2025)
MVGGT: Multimodal Visual Geometry Grounded Transformer for Multiview 3D Referring Expression Segmentation
by: Wu, Changli, et al.
Published: (2026)
by: Wu, Changli, et al.
Published: (2026)
IQE-CLIP: Instance-aware Query Embedding for Zero-/Few-shot Anomaly Detection in Medical Domain
by: Huang, Hong, et al.
Published: (2025)
by: Huang, Hong, et al.
Published: (2025)
Federated Modality-specific Encoders and Partially Personalized Fusion Decoder for Multimodal Brain Tumor Segmentation
by: Liu, Hong, et al.
Published: (2026)
by: Liu, Hong, et al.
Published: (2026)
RTGMFF: Enhanced fMRI-based Brain Disorder Diagnosis via ROI-driven Text Generation and Multimodal Feature Fusion
by: Jia, Junhao, et al.
Published: (2025)
by: Jia, Junhao, et al.
Published: (2025)
A Simple and Better Baseline for Visual Grounding
by: Wang, Jingchao, et al.
Published: (2025)
by: Wang, Jingchao, et al.
Published: (2025)
FineCLIPER: Multi-modal Fine-grained CLIP for Dynamic Facial Expression Recognition with AdaptERs
by: Chen, Haodong, et al.
Published: (2024)
by: Chen, Haodong, et al.
Published: (2024)
EIANet: A Novel Domain Adaptation Approach to Maximize Class Distinction with Neural Collapse Principles
by: Pan, Zicheng, et al.
Published: (2024)
by: Pan, Zicheng, et al.
Published: (2024)
Structure Observation Driven Image-Text Contrastive Learning for Computed Tomography Report Generation
by: Liu, Hong, et al.
Published: (2026)
by: Liu, Hong, et al.
Published: (2026)
RegionReasoner: Region-Grounded Multi-Round Visual Reasoning
by: Sun, Wenfang, et al.
Published: (2026)
by: Sun, Wenfang, et al.
Published: (2026)
DistinctAD: Distinctive Audio Description Generation in Contexts
by: Fang, Bo, et al.
Published: (2024)
by: Fang, Bo, et al.
Published: (2024)
DAONet-YOLOv8: An Occlusion-Aware Dual-Attention Network for Tea Leaf Pest and Disease Detection
by: Wu, Yefeng, et al.
Published: (2025)
by: Wu, Yefeng, et al.
Published: (2025)
GeoMVD: Geometry-Enhanced Multi-View Generation Model Based on Geometric Information Extraction
by: Wu, Jiaqi, et al.
Published: (2025)
by: Wu, Jiaqi, et al.
Published: (2025)
Quantized Visual Geometry Grounded Transformer
by: Feng, Weilun, et al.
Published: (2025)
by: Feng, Weilun, et al.
Published: (2025)
Reloc-VGGT: Visual Re-localization with Geometry Grounded Transformer
by: Deng, Tianchen, et al.
Published: (2025)
by: Deng, Tianchen, et al.
Published: (2025)
Beyond Single Models: Mitigating Multimodal Hallucinations via Adaptive Token Ensemble Decoding
by: Li, Jinlin, et al.
Published: (2025)
by: Li, Jinlin, et al.
Published: (2025)
Inteval Analysis for two spherical functions arising from robust Perspective-n-Lines problem
by: Zheng, Xiang, et al.
Published: (2025)
by: Zheng, Xiang, et al.
Published: (2025)
X-ray Insights Unleashed: Pioneering the Enhancement of Multi-Label Long-Tail Data
by: Yang, Xinquan, et al.
Published: (2025)
by: Yang, Xinquan, et al.
Published: (2025)
Language and Geometry Grounded Sparse Voxel Representations for Holistic Scene Understanding
by: Wu, Guile, et al.
Published: (2026)
by: Wu, Guile, et al.
Published: (2026)
Knowledge-Informed Neural Network for Complex-Valued SAR Image Recognition
by: Yang, Haodong, et al.
Published: (2025)
by: Yang, Haodong, et al.
Published: (2025)
4DLangVGGT: 4D Language-Visual Geometry Grounded Transformer
by: Wu, Xianfeng, et al.
Published: (2025)
by: Wu, Xianfeng, et al.
Published: (2025)
Efficient Event-Based Semantic Segmentation via Exploiting Frame-Event Fusion: A Hybrid Neural Network Approach
by: Li, Hebei, et al.
Published: (2025)
by: Li, Hebei, et al.
Published: (2025)
D-FINE: Redefine Regression Task in DETRs as Fine-grained Distribution Refinement
by: Peng, Yansong, et al.
Published: (2024)
by: Peng, Yansong, et al.
Published: (2024)
SP-SLAM: Neural Real-Time Dense SLAM With Scene Priors
by: Hong, Zhen, et al.
Published: (2025)
by: Hong, Zhen, et al.
Published: (2025)
CGF-DETR: Cross-Gated Fusion DETR for Enhanced Pneumonia Detection in Chest X-rays
by: Wu, Yefeng, et al.
Published: (2025)
by: Wu, Yefeng, et al.
Published: (2025)
Scene Adaptive Sparse Transformer for Event-based Object Detection
by: Peng, Yansong, et al.
Published: (2024)
by: Peng, Yansong, et al.
Published: (2024)
Single Domain Generalization for Multimodal Cross-Cancer Prognosis via Dirac Rebalancer and Distribution Entanglement
by: Jiang, Jia-Xuan, et al.
Published: (2025)
by: Jiang, Jia-Xuan, et al.
Published: (2025)
Non-Negative Stiefel Approximating Flow: Orthogonalish Matrix Optimization for Interpretable Embeddings
by: Avants, Brian B., et al.
Published: (2025)
by: Avants, Brian B., et al.
Published: (2025)
Reliable Active Learning from Unreliable Labels via Neural Collapse Geometry
by: Goel, Atharv, et al.
Published: (2025)
by: Goel, Atharv, et al.
Published: (2025)
Rethinking Facial Expression Recognition in the Era of Multimodal Large Language Models: Benchmark, Datasets, and Beyond
by: Zhang, Fan, et al.
Published: (2025)
by: Zhang, Fan, et al.
Published: (2025)
Progressive Language-guided Visual Learning for Multi-Task Visual Grounding
by: Wang, Jingchao, et al.
Published: (2025)
by: Wang, Jingchao, et al.
Published: (2025)
Video-Holmes: Can MLLM Think Like Holmes for Complex Video Reasoning?
by: Cheng, Junhao, et al.
Published: (2025)
by: Cheng, Junhao, et al.
Published: (2025)
VPTracker: Global Vision-Language Tracking via Visual Prompt
by: Wang, Jingchao, et al.
Published: (2025)
by: Wang, Jingchao, et al.
Published: (2025)
UniPR-3D: Towards Universal Visual Place Recognition with Visual Geometry Grounded Transformer
by: Deng, Tianchen, et al.
Published: (2025)
by: Deng, Tianchen, et al.
Published: (2025)
Geometry-Grounded Gaussian Splatting
by: Zhang, Baowen, et al.
Published: (2026)
by: Zhang, Baowen, et al.
Published: (2026)
EVA: Mixture-of-Experts Semantic Variant Alignment for Compositional Zero-Shot Learning
by: Zhang, Xiao, et al.
Published: (2025)
by: Zhang, Xiao, et al.
Published: (2025)
Similar Items
-
Unsupervised Causal Prototypical Networks for De-biased Interpretable Dermoscopy Diagnosis
by: Jia, Junhao, et al.
Published: (2026) -
Geodesic Prototype Matching via Diffusion Maps for Interpretable Fine-Grained Recognition
by: Jia, Junhao, et al.
Published: (2025) -
Reusing Fusion-Time Spectral Reliability for Adaptive Fusion and Expert Routing in RGB-Infrared Object Detection
by: Wu, Yefeng
Published: (2026) -
Brain-HGCN: A Hyperbolic Graph Convolutional Network for Brain Functional Network Analysis
by: Jia, Junhao, et al.
Published: (2025) -
BrainSegDMlF: A Dynamic Fusion-enhanced SAM for Brain Lesion Segmentation
by: Wang, Hongming, et al.
Published: (2025)