Wasserstein-Aligned Hyperbolic Multi-View Clustering
Fuente:
arXiv
Salvato in:
| Autori principali: | Wang, Rui, Jiang, Yuting, Luo, Xiaoqing, Wu, Xiao-Jun, Sebe, Nicu, Chen, Ziheng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Hyperbolic Busemann Neural Networks
di: Chen, Ziheng, et al.
Pubblicazione: (2026)
di: Chen, Ziheng, et al.
Pubblicazione: (2026)
Understanding Matrix Function Normalizations in Covariance Pooling through the Lens of Riemannian Geometry
di: Chen, Ziheng, et al.
Pubblicazione: (2024)
di: Chen, Ziheng, et al.
Pubblicazione: (2024)
AlignCAT: Visual-Linguistic Alignment of Category and Attribute for Weakly Supervised Visual Grounding
di: Wang, Yidan, et al.
Pubblicazione: (2025)
di: Wang, Yidan, et al.
Pubblicazione: (2025)
PoInit-of-View: Poisoning Initialization of Views Transfers Across Multiple 3D Reconstruction Systems
di: Wang, Weijie, et al.
Pubblicazione: (2026)
di: Wang, Weijie, et al.
Pubblicazione: (2026)
Multi-focal Conditioned Latent Diffusion for Person Image Synthesis
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)
Token Reduction via Local and Global Contexts Optimization for Efficient Video Large Language Models
di: Li, Jinlong, et al.
Pubblicazione: (2026)
di: Li, Jinlong, et al.
Pubblicazione: (2026)
CLIP is Strong Enough to Fight Back: Test-time Counterattacks towards Zero-shot Adversarial Robustness of CLIP
di: Xing, Songlong, et al.
Pubblicazione: (2025)
di: Xing, Songlong, et al.
Pubblicazione: (2025)
Vision+X: A Survey on Multimodal Learning in the Light of Data
di: Zhu, Ye, et al.
Pubblicazione: (2022)
di: Zhu, Ye, et al.
Pubblicazione: (2022)
Hierarchical Cross-Attention Network for Virtual Try-On
di: Tang, Hao, et al.
Pubblicazione: (2024)
di: Tang, Hao, et al.
Pubblicazione: (2024)
Enhanced Multi-Scale Cross-Attention for Person Image Generation
di: Tang, Hao, et al.
Pubblicazione: (2025)
di: Tang, Hao, et al.
Pubblicazione: (2025)
RankFeat&RankWeight: Rank-1 Feature/Weight Removal for Out-of-distribution Detection
di: Song, Yue, et al.
Pubblicazione: (2023)
di: Song, Yue, et al.
Pubblicazione: (2023)
Rethinking the Learning Paradigm for Facial Expression Recognition
di: Wang, Weijie, et al.
Pubblicazione: (2022)
di: Wang, Weijie, et al.
Pubblicazione: (2022)
Reverse Personalization
di: Kung, Han-Wei, et al.
Pubblicazione: (2025)
di: Kung, Han-Wei, et al.
Pubblicazione: (2025)
Probabilistically Aligned View-unaligned Clustering with Adaptive Template Selection
di: Dong, Wenhua, et al.
Pubblicazione: (2024)
di: Dong, Wenhua, et al.
Pubblicazione: (2024)
Generalized Fine-Grained Category Discovery with Multi-Granularity Conceptual Experts
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
di: Zheng, Haiyang, et al.
Pubblicazione: (2025)
Asymmetric GANs for Image-to-Image Translation
di: Tang, Hao, et al.
Pubblicazione: (2019)
di: Tang, Hao, et al.
Pubblicazione: (2019)
Masked Clustering Prediction for Unsupervised Point Cloud Pre-training
di: Ren, Bin, et al.
Pubblicazione: (2025)
di: Ren, Bin, et al.
Pubblicazione: (2025)
Cross-Modal and Uncertainty-Aware Agglomeration for Open-Vocabulary 3D Scene Understanding
di: Li, Jinlong, et al.
Pubblicazione: (2025)
di: Li, Jinlong, et al.
Pubblicazione: (2025)
High-Fidelity 3D Facial Avatar Synthesis with Controllable Fine-Grained Expressions
di: He, Yikang, et al.
Pubblicazione: (2026)
di: He, Yikang, et al.
Pubblicazione: (2026)
Anti-Forgetting Adaptation for Unsupervised Person Re-identification
di: Chen, Hao, et al.
Pubblicazione: (2024)
di: Chen, Hao, et al.
Pubblicazione: (2024)
Inverse Virtual Try-On: Generating Multi-Category Product-Style Images from Clothed Individuals
di: Lobba, Davide, et al.
Pubblicazione: (2025)
di: Lobba, Davide, et al.
Pubblicazione: (2025)
NullFace: Training-Free Localized Face Anonymization
di: Kung, Han-Wei, et al.
Pubblicazione: (2025)
di: Kung, Han-Wei, et al.
Pubblicazione: (2025)
Graph Transformer GANs with Graph Masked Modeling for Architectural Layout Generation
di: Tang, Hao, et al.
Pubblicazione: (2024)
di: Tang, Hao, et al.
Pubblicazione: (2024)
MTGA: Multi-View Temporal Granularity Aligned Aggregation for Event-Based Lip-Reading
di: Zhang, Wenhao, et al.
Pubblicazione: (2024)
di: Zhang, Wenhao, et al.
Pubblicazione: (2024)
Generate, Refine, and Encode: Leveraging Synthesized Novel Samples for On-the-Fly Fine-Grained Category Discovery
di: Liu, Xiao, et al.
Pubblicazione: (2025)
di: Liu, Xiao, et al.
Pubblicazione: (2025)
Cues3D: Unleashing the Power of Sole NeRF for Consistent and Unique Instances in Open-Vocabulary 3D Panoptic Segmentation
di: Xue, Feng, et al.
Pubblicazione: (2025)
di: Xue, Feng, et al.
Pubblicazione: (2025)
RAGME: Retrieval Augmented Video Generation for Enhanced Motion Realism
di: Peruzzo, Elia, et al.
Pubblicazione: (2025)
di: Peruzzo, Elia, et al.
Pubblicazione: (2025)
Prototypical Hash Encoding for On-the-Fly Fine-Grained Category Discovery
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
Optimizing Resource Consumption in Diffusion Models through Hallucination Early Detection
di: Betti, Federico, et al.
Pubblicazione: (2024)
di: Betti, Federico, et al.
Pubblicazione: (2024)
Hallucination Early Detection in Diffusion Models
di: Betti, Federico, et al.
Pubblicazione: (2026)
di: Betti, Federico, et al.
Pubblicazione: (2026)
Textual Knowledge Matters: Cross-Modality Co-Teaching for Generalized Visual Class Discovery
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
di: Zheng, Haiyang, et al.
Pubblicazione: (2024)
Transferable-guided Attention Is All You Need for Video Domain Adaptation
di: Sacilotti, André, et al.
Pubblicazione: (2024)
di: Sacilotti, André, et al.
Pubblicazione: (2024)
MVReward: Better Aligning and Evaluating Multi-View Diffusion Models with Human Preferences
di: Wang, Weitao, et al.
Pubblicazione: (2024)
di: Wang, Weitao, et al.
Pubblicazione: (2024)
Towards End-to-End Explainable Facial Action Unit Recognition via Vision-Language Joint Learning
di: Ge, Xuri, et al.
Pubblicazione: (2024)
di: Ge, Xuri, et al.
Pubblicazione: (2024)
FisherTune: Fisher-Guided Robust Tuning of Vision Foundation Models for Domain Generalized Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2025)
di: Zhao, Dong, et al.
Pubblicazione: (2025)
In defense of the two-stage framework for open-set domain adaptive semantic segmentation
di: Ren, Wenqi, et al.
Pubblicazione: (2026)
di: Ren, Wenqi, et al.
Pubblicazione: (2026)
FreeInsert: Disentangled Text-Guided Object Insertion in 3D Gaussian Scene without Spatial Priors
di: Li, Chenxi, et al.
Pubblicazione: (2025)
di: Li, Chenxi, et al.
Pubblicazione: (2025)
Fully-Geometric Cross-Attention for Point Cloud Registration
di: Wang, Weijie, et al.
Pubblicazione: (2025)
di: Wang, Weijie, et al.
Pubblicazione: (2025)
Causal Disentanglement for Robust Long-tail Medical Image Generation
di: Nie, Weizhi, et al.
Pubblicazione: (2025)
di: Nie, Weizhi, et al.
Pubblicazione: (2025)
Stable Neighbor Denoising for Source-free Domain Adaptive Segmentation
di: Zhao, Dong, et al.
Pubblicazione: (2024)
di: Zhao, Dong, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Hyperbolic Busemann Neural Networks
di: Chen, Ziheng, et al.
Pubblicazione: (2026) -
Understanding Matrix Function Normalizations in Covariance Pooling through the Lens of Riemannian Geometry
di: Chen, Ziheng, et al.
Pubblicazione: (2024) -
AlignCAT: Visual-Linguistic Alignment of Category and Attribute for Weakly Supervised Visual Grounding
di: Wang, Yidan, et al.
Pubblicazione: (2025) -
PoInit-of-View: Poisoning Initialization of Views Transfers Across Multiple 3D Reconstruction Systems
di: Wang, Weijie, et al.
Pubblicazione: (2026) -
Multi-focal Conditioned Latent Diffusion for Person Image Synthesis
di: Liu, Jiaqi, et al.
Pubblicazione: (2025)