Robust Fuzzy Multi-view Learning under View Conflict
Fuente:
arXiv
Salvato in:
| Autori principali: | Duan, Siyuan, Sun, Yuan, Peng, Dezhong, Chen, Yingke, Peng, Xi, Hu, Peng |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Noisy-Correspondence Learning for Text-to-Image Person Re-identification
di: Qin, Yang, et al.
Pubblicazione: (2023)
di: Qin, Yang, et al.
Pubblicazione: (2023)
Robust Duality Learning for Unsupervised Visible-Infrared Person Re-Identification
di: Li, Yongxiang, et al.
Pubblicazione: (2025)
di: Li, Yongxiang, et al.
Pubblicazione: (2025)
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
di: Pu, Ruitao, et al.
Pubblicazione: (2025)
di: Pu, Ruitao, et al.
Pubblicazione: (2025)
Enhancing Visible-Infrared Person Re-identification with Modality- and Instance-aware Visual Prompt Learning
di: Wu, Ruiqi, et al.
Pubblicazione: (2024)
di: Wu, Ruiqi, et al.
Pubblicazione: (2024)
Advancing 3D Scene Understanding with MV-ScanQA Multi-View Reasoning Evaluation and TripAlign Pre-training Dataset
di: Mo, Wentao, et al.
Pubblicazione: (2025)
di: Mo, Wentao, et al.
Pubblicazione: (2025)
ASR-enhanced Multimodal Representation Learning for Cross-Domain Product Retrieval
di: Zhao, Ruixiang, et al.
Pubblicazione: (2024)
di: Zhao, Ruixiang, et al.
Pubblicazione: (2024)
Deep Reversible Consistency Learning for Cross-modal Retrieval
di: Pu, Ruitao, et al.
Pubblicazione: (2025)
di: Pu, Ruitao, et al.
Pubblicazione: (2025)
Q-Ponder: A Unified Training Pipeline for Reasoning-based Visual Quality Assessment
di: Cai, Zhuoxuan, et al.
Pubblicazione: (2025)
di: Cai, Zhuoxuan, et al.
Pubblicazione: (2025)
Knowledge-enhanced Multi-perspective Video Representation Learning for Scene Recognition
di: Yu, Xuzheng, et al.
Pubblicazione: (2024)
di: Yu, Xuzheng, et al.
Pubblicazione: (2024)
MM-Point: Multi-View Information-Enhanced Multi-Modal Self-Supervised 3D Point Cloud Understanding
di: Yu, Hai-Tao, et al.
Pubblicazione: (2024)
di: Yu, Hai-Tao, et al.
Pubblicazione: (2024)
OVFoodSeg: Elevating Open-Vocabulary Food Image Segmentation via Image-Informed Textual Representation
di: Wu, Xiongwei, et al.
Pubblicazione: (2024)
di: Wu, Xiongwei, et al.
Pubblicazione: (2024)
Robust Latent Representation Tuning for Image-text Classification
di: Sun, Hao, et al.
Pubblicazione: (2024)
di: Sun, Hao, et al.
Pubblicazione: (2024)
ContextDet: Temporal Action Detection with Adaptive Context Aggregation
di: Wang, Ning, et al.
Pubblicazione: (2024)
di: Wang, Ning, et al.
Pubblicazione: (2024)
Navigating the Mirage: A Dual-Path Agentic Framework for Robust Misleading Chart Question Answering
di: Zhang, Yanjie, et al.
Pubblicazione: (2026)
di: Zhang, Yanjie, et al.
Pubblicazione: (2026)
Learning to Rematch Mismatched Pairs for Robust Cross-Modal Retrieval
di: Han, Haochen, et al.
Pubblicazione: (2024)
di: Han, Haochen, et al.
Pubblicazione: (2024)
Multi-modal and Multi-scale Spatial Environment Understanding for Immersive Visual Text-to-Speech
di: Liu, Rui, et al.
Pubblicazione: (2024)
di: Liu, Rui, et al.
Pubblicazione: (2024)
MuKV: Multi-Grained KV Cache Compression for Long Streaming Video Question-Answering
di: Xiao, Junbin, et al.
Pubblicazione: (2026)
di: Xiao, Junbin, et al.
Pubblicazione: (2026)
AutoAWG: Adverse Weather Generation with Adaptive Multi-Controls for Automotive Videos
di: Hu, Jiagao, et al.
Pubblicazione: (2026)
di: Hu, Jiagao, et al.
Pubblicazione: (2026)
Crab$^{+}$: A Scalable and Unified Audio-Visual Scene Understanding Model with Explicit Cooperation
di: Cai, Dongnuan, et al.
Pubblicazione: (2026)
di: Cai, Dongnuan, et al.
Pubblicazione: (2026)
Semantic-Consistent Bidirectional Contrastive Hashing for Noisy Multi-Label Cross-Modal Retrieval
di: Peng, Likang, et al.
Pubblicazione: (2025)
di: Peng, Likang, et al.
Pubblicazione: (2025)
Reliable Disentanglement Multi-view Learning Against View Adversarial Attacks
di: Wang, Xuyang, et al.
Pubblicazione: (2025)
di: Wang, Xuyang, et al.
Pubblicazione: (2025)
Human-Centric Foundation Models: Perception, Generation and Agentic Modeling
di: Tang, Shixiang, et al.
Pubblicazione: (2025)
di: Tang, Shixiang, et al.
Pubblicazione: (2025)
Diagnosing and Re-learning for Balanced Multimodal Learning
di: Wei, Yake, et al.
Pubblicazione: (2024)
di: Wei, Yake, et al.
Pubblicazione: (2024)
TIDE : Temporal-Aware Sparse Autoencoders for Interpretable Diffusion Transformers in Image Generation
di: Huang, Victor Shea-Jay, et al.
Pubblicazione: (2025)
di: Huang, Victor Shea-Jay, et al.
Pubblicazione: (2025)
Deep Contrastive Multi-view Clustering under Semantic Feature Guidance
di: Liu, Siwen, et al.
Pubblicazione: (2024)
di: Liu, Siwen, et al.
Pubblicazione: (2024)
Audio Visual Segmentation Through Text Embeddings
di: Lee, Kyungbok, et al.
Pubblicazione: (2025)
di: Lee, Kyungbok, et al.
Pubblicazione: (2025)
KAN Text to Vision? The Exploration of Kolmogorov-Arnold Networks for Multi-Scale Sequence-Based Pose Animation from Sign Language Notation
di: Du, Guanyi, et al.
Pubblicazione: (2026)
di: Du, Guanyi, et al.
Pubblicazione: (2026)
DIRECT: Video Mashup Creation via Hierarchical Multi-Agent Planning and Intent-Guided Editing
di: Li, Ke, et al.
Pubblicazione: (2026)
di: Li, Ke, et al.
Pubblicazione: (2026)
MIRROR: Multi-Modal Pathological Self-Supervised Representation Learning via Modality Alignment and Retention
di: Wang, Tianyi, et al.
Pubblicazione: (2025)
di: Wang, Tianyi, et al.
Pubblicazione: (2025)
From Satellite to Street: A Hybrid Framework Integrating Stable Diffusion and PanoGAN for Consistent Cross-View Synthesis
di: Bajbaa, Khawlah, et al.
Pubblicazione: (2025)
di: Bajbaa, Khawlah, et al.
Pubblicazione: (2025)
Learning Segment Similarity and Alignment in Large-Scale Content Based Video Retrieval
di: Jiang, Chen, et al.
Pubblicazione: (2023)
di: Jiang, Chen, et al.
Pubblicazione: (2023)
Bootstrap3D: Improving Multi-view Diffusion Model with Synthetic Data
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
di: Sun, Zeyi, et al.
Pubblicazione: (2024)
DreamStory: Open-Domain Story Visualization by LLM-Guided Multi-Subject Consistent Diffusion
di: He, Huiguo, et al.
Pubblicazione: (2024)
di: He, Huiguo, et al.
Pubblicazione: (2024)
AKiRa: Augmentation Kit on Rays for optical video generation
di: Wang, Xi, et al.
Pubblicazione: (2024)
di: Wang, Xi, et al.
Pubblicazione: (2024)
URMF: Uncertainty-aware Robust Multimodal Fusion for Multimodal Sarcasm Detection
di: Wang, Zhenyu, et al.
Pubblicazione: (2026)
di: Wang, Zhenyu, et al.
Pubblicazione: (2026)
Rethinking Multi-Condition DiTs: Eliminating Redundant Attention via Position-Alignment and Keyword-Scoping
di: Zhou, Chao, et al.
Pubblicazione: (2026)
di: Zhou, Chao, et al.
Pubblicazione: (2026)
DiffMesh: A Motion-aware Diffusion Framework for Human Mesh Recovery from Videos
di: Zheng, Ce, et al.
Pubblicazione: (2023)
di: Zheng, Ce, et al.
Pubblicazione: (2023)
Large-scale Multi-Modal Pre-trained Models: A Comprehensive Survey
di: Wang, Xiao, et al.
Pubblicazione: (2023)
di: Wang, Xiao, et al.
Pubblicazione: (2023)
PointCoT: A Multi-modal Benchmark for Explicit 3D Geometric Reasoning
di: Zhang, Dongxu, et al.
Pubblicazione: (2026)
di: Zhang, Dongxu, et al.
Pubblicazione: (2026)
SoundVista: Novel-View Ambient Sound Synthesis via Visual-Acoustic Binding
di: Chen, Mingfei, et al.
Pubblicazione: (2025)
di: Chen, Mingfei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Noisy-Correspondence Learning for Text-to-Image Person Re-identification
di: Qin, Yang, et al.
Pubblicazione: (2023) -
Robust Duality Learning for Unsupervised Visible-Infrared Person Re-Identification
di: Li, Yongxiang, et al.
Pubblicazione: (2025) -
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
di: Pu, Ruitao, et al.
Pubblicazione: (2025) -
Enhancing Visible-Infrared Person Re-identification with Modality- and Instance-aware Visual Prompt Learning
di: Wu, Ruiqi, et al.
Pubblicazione: (2024) -
Advancing 3D Scene Understanding with MV-ScanQA Multi-View Reasoning Evaluation and TripAlign Pre-training Dataset
di: Mo, Wentao, et al.
Pubblicazione: (2025)