Structure-Aware Residual-Center Representation for Self-Supervised Open-Set 3D Cross-Modal Retrieval
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Xu, Yang, Feng, Yifan, Jiang, Yu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
von: Pu, Ruitao, et al.
Veröffentlicht: (2025)
von: Pu, Ruitao, et al.
Veröffentlicht: (2025)
Towards Unified Representation of Multi-Modal Pre-training for 3D Understanding via Differentiable Rendering
von: Fei, Ben, et al.
Veröffentlicht: (2024)
von: Fei, Ben, et al.
Veröffentlicht: (2024)
Towards Unbiased Cross-Modal Representation Learning for Food Image-to-Recipe Retrieval
von: Wang, Qing, et al.
Veröffentlicht: (2025)
von: Wang, Qing, et al.
Veröffentlicht: (2025)
Improving the Consistency in Cross-Lingual Cross-Modal Retrieval with 1-to-K Contrastive Learning
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
von: Nie, Zhijie, et al.
Veröffentlicht: (2024)
AsCL: An Asymmetry-sensitive Contrastive Learning Method for Image-Text Retrieval with Cross-Modal Fusion
von: Gong, Ziyu, et al.
Veröffentlicht: (2024)
von: Gong, Ziyu, et al.
Veröffentlicht: (2024)
MIRROR: Multi-Modal Pathological Self-Supervised Representation Learning via Modality Alignment and Retention
von: Wang, Tianyi, et al.
Veröffentlicht: (2025)
von: Wang, Tianyi, et al.
Veröffentlicht: (2025)
Start from Video-Music Retrieval: An Inter-Intra Modal Loss for Cross Modal Retrieval
von: Chen, Zeyu, et al.
Veröffentlicht: (2024)
von: Chen, Zeyu, et al.
Veröffentlicht: (2024)
Towards Multimodal Sentiment Analysis via Contrastive Cross-modal Retrieval Augmentation and Hierachical Prompts
von: Zhao, Xianbing, et al.
Veröffentlicht: (2025)
von: Zhao, Xianbing, et al.
Veröffentlicht: (2025)
Music2Palette: Emotion-aligned Color Palette Generation via Cross-Modal Representation Learning
von: Hu, Jiayun, et al.
Veröffentlicht: (2025)
von: Hu, Jiayun, et al.
Veröffentlicht: (2025)
Cross-Modal Retrieval: A Systematic Review of Methods and Future Directions
von: Wang, Tianshi, et al.
Veröffentlicht: (2023)
von: Wang, Tianshi, et al.
Veröffentlicht: (2023)
Token-Level Contrastive Learning with Modality-Aware Prompting for Multimodal Intent Recognition
von: Zhou, Qianrui, et al.
Veröffentlicht: (2023)
von: Zhou, Qianrui, et al.
Veröffentlicht: (2023)
Towards Temporal-Aware Multi-Modal Retrieval Augmented Generation in Finance
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
von: Zhu, Fengbin, et al.
Veröffentlicht: (2025)
Towards Identity-Aware Cross-Modal Retrieval: a Dataset and a Baseline
von: Messina, Nicola, et al.
Veröffentlicht: (2024)
von: Messina, Nicola, et al.
Veröffentlicht: (2024)
GestureHYDRA: Semantic Co-speech Gesture Synthesis via Hybrid Modality Diffusion Transformer and Cascaded-Synchronized Retrieval-Augmented Generation
von: Yang, Quanwei, et al.
Veröffentlicht: (2025)
von: Yang, Quanwei, et al.
Veröffentlicht: (2025)
MM-Point: Multi-View Information-Enhanced Multi-Modal Self-Supervised 3D Point Cloud Understanding
von: Yu, Hai-Tao, et al.
Veröffentlicht: (2024)
von: Yu, Hai-Tao, et al.
Veröffentlicht: (2024)
Cross-Modal and Uni-Modal Soft-Label Alignment for Image-Text Retrieval
von: Huang, Hailang, et al.
Veröffentlicht: (2024)
von: Huang, Hailang, et al.
Veröffentlicht: (2024)
Cross-Modality and Within-Modality Regularization for Audio-Visual DeepFake Detection
von: Zou, Heqing, et al.
Veröffentlicht: (2024)
von: Zou, Heqing, et al.
Veröffentlicht: (2024)
SonicGauss: Position-Aware Physical Sound Synthesis for 3D Gaussian Representations
von: Wang, Chunshi, et al.
Veröffentlicht: (2025)
von: Wang, Chunshi, et al.
Veröffentlicht: (2025)
SemCORE: A Semantic-Enhanced Generative Cross-Modal Retrieval Framework with MLLMs
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
von: Li, Haoxuan, et al.
Veröffentlicht: (2025)
StreamOptix: A Cross-layer Adaptive Video Delivery Scheme
von: Liu, Mufan, et al.
Veröffentlicht: (2024)
von: Liu, Mufan, et al.
Veröffentlicht: (2024)
Cross-Modal Retrieval with Cauchy-Schwarz Divergence
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
von: Zhang, Jiahao, et al.
Veröffentlicht: (2025)
Teacher-Guided Pseudo Supervision and Cross-Modal Alignment for Audio-Visual Video Parsing
von: Chen, Yaru, et al.
Veröffentlicht: (2025)
von: Chen, Yaru, et al.
Veröffentlicht: (2025)
Mitigating Cross-modal Representation Bias for Multicultural Image-to-Recipe Retrieval
von: Wang, Qing, et al.
Veröffentlicht: (2025)
von: Wang, Qing, et al.
Veröffentlicht: (2025)
Cross Modal Fine-Grained Alignment via Granularity-Aware and Region-Uncertain Modeling
von: Liu, Jiale, et al.
Veröffentlicht: (2025)
von: Liu, Jiale, et al.
Veröffentlicht: (2025)
SRA: Semantic Relation-Aware Flowchart Question Answering
von: Li, Xinyu, et al.
Veröffentlicht: (2026)
von: Li, Xinyu, et al.
Veröffentlicht: (2026)
MagicLens: Self-Supervised Image Retrieval with Open-Ended Instructions
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
von: Zhang, Kai, et al.
Veröffentlicht: (2024)
Modality-Aware Contrastive and Uncertainty-Regularized Emotion Recognition
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
von: Zhuang, Yan, et al.
Veröffentlicht: (2026)
Cross-Modal Coordination Across a Diverse Set of Input Modalities
von: Sánchez, Jorge, et al.
Veröffentlicht: (2024)
von: Sánchez, Jorge, et al.
Veröffentlicht: (2024)
Multimodal Information Retrieval for Open World with Edit Distance Weak Supervision
von: Solaiman, KMA, et al.
Veröffentlicht: (2025)
von: Solaiman, KMA, et al.
Veröffentlicht: (2025)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
von: Xu, Junhao, et al.
Veröffentlicht: (2025)
von: Xu, Junhao, et al.
Veröffentlicht: (2025)
Evolutionary Multimodal Reasoning via Hierarchical Semantic Representation for Intent Recognition
von: Zhou, Qianrui, et al.
Veröffentlicht: (2026)
von: Zhou, Qianrui, et al.
Veröffentlicht: (2026)
CL2CM: Improving Cross-Lingual Cross-Modal Retrieval via Cross-Lingual Knowledge Transfer
von: Wang, Yabing, et al.
Veröffentlicht: (2023)
von: Wang, Yabing, et al.
Veröffentlicht: (2023)
M3TR: Temporal Retrieval Enhanced Multi-Modal Micro-video Popularity Prediction
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024)
von: Lu, Jiacheng, et al.
Veröffentlicht: (2024)
PrismSSL: One Interface, Many Modalities; A Single-Interface Library for Multimodal Self-Supervised Learning
von: Shirian, Melika, et al.
Veröffentlicht: (2025)
von: Shirian, Melika, et al.
Veröffentlicht: (2025)
FLARE: Full-Modality Long-Video Audiovisual Retrieval Benchmark with User-Simulated Queries
von: You, Qijie, et al.
Veröffentlicht: (2026)
von: You, Qijie, et al.
Veröffentlicht: (2026)
High-Fidelity 3D Gaussian Human Reconstruction via Region-Aware Initialization and Geometric Priors
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
EAR: Enhancing Uni-Modal Representations for Weakly Supervised Audio-Visual Video Parsing
von: Li, Huilai, et al.
Veröffentlicht: (2026)
von: Li, Huilai, et al.
Veröffentlicht: (2026)
A Survey on Music Generation from Single-Modal, Cross-Modal, and Multi-Modal Perspectives
von: Li, Shuyu, et al.
Veröffentlicht: (2025)
von: Li, Shuyu, et al.
Veröffentlicht: (2025)
GaussianCross: Cross-modal Self-supervised 3D Representation Learning via Gaussian Splatting
von: Yao, Lei, et al.
Veröffentlicht: (2025)
von: Yao, Lei, et al.
Veröffentlicht: (2025)
AFL-Net: Integrating Audio, Facial, and Lip Modalities with a Two-step Cross-attention for Robust Speaker Diarization in the Wild
von: Yin, Yongkang, et al.
Veröffentlicht: (2023)
von: Yin, Yongkang, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Robust Self-Paced Hashing for Cross-Modal Retrieval with Noisy Labels
von: Pu, Ruitao, et al.
Veröffentlicht: (2025) -
Towards Unified Representation of Multi-Modal Pre-training for 3D Understanding via Differentiable Rendering
von: Fei, Ben, et al.
Veröffentlicht: (2024) -
Towards Unbiased Cross-Modal Representation Learning for Food Image-to-Recipe Retrieval
von: Wang, Qing, et al.
Veröffentlicht: (2025) -
Improving the Consistency in Cross-Lingual Cross-Modal Retrieval with 1-to-K Contrastive Learning
von: Nie, Zhijie, et al.
Veröffentlicht: (2024) -
AsCL: An Asymmetry-sensitive Contrastive Learning Method for Image-Text Retrieval with Cross-Modal Fusion
von: Gong, Ziyu, et al.
Veröffentlicht: (2024)