Comparing Contrastive and Triplet Loss: Variance Analysis and Optimization Behavior
Fuente:
arXiv
Salvato in:
| Autore principale: | Zeng, Donghuo |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Learning Audio-Visual Embeddings with Inferred Latent Interaction Graphs
di: Zeng, Donghuo, et al.
Pubblicazione: (2026)
di: Zeng, Donghuo, et al.
Pubblicazione: (2026)
EquiAV: Leveraging Equivariance for Audio-Visual Contrastive Learning
di: Kim, Jongsuk, et al.
Pubblicazione: (2024)
di: Kim, Jongsuk, et al.
Pubblicazione: (2024)
Hierarchical Semantic Correlation-Aware Masked Autoencoder for Unsupervised Audio-Visual Representation Learning
di: Zeng, Donghuo, et al.
Pubblicazione: (2026)
di: Zeng, Donghuo, et al.
Pubblicazione: (2026)
Low-Resolution Object Recognition with Cross-Resolution Relational Contrastive Distillation
di: Zhang, Kangkai, et al.
Pubblicazione: (2024)
di: Zhang, Kangkai, et al.
Pubblicazione: (2024)
DRAGON: Distributional Rewards Optimize Diffusion Generative Models
di: Bai, Yatong, et al.
Pubblicazione: (2025)
di: Bai, Yatong, et al.
Pubblicazione: (2025)
Contrastive Regularization over LoRA for Multimodal Biomedical Image Incremental Learning
di: Zhang, Haojie, et al.
Pubblicazione: (2025)
di: Zhang, Haojie, et al.
Pubblicazione: (2025)
Harmony: A Unified Framework for Modality Incremental Learning
di: Song, Yaguang, et al.
Pubblicazione: (2025)
di: Song, Yaguang, et al.
Pubblicazione: (2025)
GRAFT: GRaPH and Table Reasoning for Textual Alignment -- A Benchmark for Structured Instruction Following and Visual Reasoning
di: Verma, Abhigya, et al.
Pubblicazione: (2025)
di: Verma, Abhigya, et al.
Pubblicazione: (2025)
Identifying Multi-modal Knowledge Neurons in Pretrained Transformers via Two-stage Filtering
di: Sato, Yugen, et al.
Pubblicazione: (2025)
di: Sato, Yugen, et al.
Pubblicazione: (2025)
Time-RA: Towards Time Series Reasoning for Anomaly Diagnosis with LLM Feedback
di: Yang, Yiyuan, et al.
Pubblicazione: (2025)
di: Yang, Yiyuan, et al.
Pubblicazione: (2025)
OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition
di: Yan, Xueming, et al.
Pubblicazione: (2025)
di: Yan, Xueming, et al.
Pubblicazione: (2025)
HeLo: Heterogeneous Multi-Modal Fusion with Label Correlation for Emotion Distribution Learning
di: Zheng, Chuhang, et al.
Pubblicazione: (2025)
di: Zheng, Chuhang, et al.
Pubblicazione: (2025)
AI-Integrated Decision Support System for Real-Time Market Growth Forecasting and Multi-Source Content Diffusion Analytics
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
di: Yin, Ziqing, et al.
Pubblicazione: (2025)
FedNano: Toward Lightweight Federated Tuning for Pretrained Multimodal Large Language Models
di: Zhang, Yao, et al.
Pubblicazione: (2025)
di: Zhang, Yao, et al.
Pubblicazione: (2025)
Detecting Multimedia Generated by Large AI Models: A Survey
di: Lin, Li, et al.
Pubblicazione: (2024)
di: Lin, Li, et al.
Pubblicazione: (2024)
Unveiling Covert Toxicity in Multimodal Data via Toxicity Association Graphs: A Graph-Based Metric and Interpretable Detection Framework
di: Wu, Guanzong, et al.
Pubblicazione: (2026)
di: Wu, Guanzong, et al.
Pubblicazione: (2026)
MixEval-X: Any-to-Any Evaluations from Real-World Data Mixtures
di: Ni, Jinjie, et al.
Pubblicazione: (2024)
di: Ni, Jinjie, et al.
Pubblicazione: (2024)
Predicting Outcomes in Video Games with Long Short Term Memory Networks
di: Chulajata, Kittimate, et al.
Pubblicazione: (2024)
di: Chulajata, Kittimate, et al.
Pubblicazione: (2024)
A review on Machine Learning based User-Centric Multimedia Streaming Techniques
di: Ghosh, Monalisa, et al.
Pubblicazione: (2024)
di: Ghosh, Monalisa, et al.
Pubblicazione: (2024)
Enhancing Modality Representation and Alignment for Multimodal Cold-start Active Learning
di: Shen, Meng, et al.
Pubblicazione: (2024)
di: Shen, Meng, et al.
Pubblicazione: (2024)
BI-MDRG: Bridging Image History in Multimodal Dialogue Response Generation
di: Yoon, Hee Suk, et al.
Pubblicazione: (2024)
di: Yoon, Hee Suk, et al.
Pubblicazione: (2024)
Harmful Visual Content Manipulation Matters in Misinformation Detection Under Multimedia Scenarios
di: Wang, Bing, et al.
Pubblicazione: (2026)
di: Wang, Bing, et al.
Pubblicazione: (2026)
Metric Learning with Progressive Self-Distillation for Audio-Visual Embedding Learning
di: Zeng, Donghuo, et al.
Pubblicazione: (2025)
di: Zeng, Donghuo, et al.
Pubblicazione: (2025)
MuPHI: Learning Implicit Multimodal Harm Reasoning via Semantically Grounded Reward Optimization
di: Saha, Anisha, et al.
Pubblicazione: (2026)
di: Saha, Anisha, et al.
Pubblicazione: (2026)
DLF: Disentangled-Language-Focused Multimodal Sentiment Analysis
di: Wang, Pan, et al.
Pubblicazione: (2024)
di: Wang, Pan, et al.
Pubblicazione: (2024)
Multimodal Multi-loss Fusion Network for Sentiment Analysis
di: Wu, Zehui, et al.
Pubblicazione: (2023)
di: Wu, Zehui, et al.
Pubblicazione: (2023)
Anchor-aware Deep Metric Learning for Audio-visual Retrieval
di: Zeng, Donghuo, et al.
Pubblicazione: (2024)
di: Zeng, Donghuo, et al.
Pubblicazione: (2024)
SABR: A Stable Adaptive Bitrate Framework Using Behavior Cloning Pretraining and Reinforcement Learning Fine-Tuning
di: Luo, Pengcheng, et al.
Pubblicazione: (2025)
di: Luo, Pengcheng, et al.
Pubblicazione: (2025)
Counterfactual Reasoning Using Predicted Latent Personality Dimensions for Optimizing Persuasion Outcome
di: Zeng, Donghuo, et al.
Pubblicazione: (2024)
di: Zeng, Donghuo, et al.
Pubblicazione: (2024)
FISHER: A Foundation Model for Multi-Modal Industrial Signal Comprehensive Representation
di: Fan, Pingyi, et al.
Pubblicazione: (2025)
di: Fan, Pingyi, et al.
Pubblicazione: (2025)
LASPA: Language Agnostic Speaker Disentanglement with Prefix-Tuned Cross-Attention
di: Menon, Aditya Srinivas, et al.
Pubblicazione: (2025)
di: Menon, Aditya Srinivas, et al.
Pubblicazione: (2025)
Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators
di: Novack, Zachary, et al.
Pubblicazione: (2026)
di: Novack, Zachary, et al.
Pubblicazione: (2026)
See, Think, Act: Online Shopper Behavior Simulation with VLM Agents
di: Zhang, Yimeng, et al.
Pubblicazione: (2025)
di: Zhang, Yimeng, et al.
Pubblicazione: (2025)
Contrastive Visual Data Augmentation
di: Zhou, Yu, et al.
Pubblicazione: (2025)
di: Zhou, Yu, et al.
Pubblicazione: (2025)
EEG2TEXT-CN: An Exploratory Study of Open-Vocabulary Chinese Text-EEG Alignment via Large Language Model and Contrastive Learning on ChineseEEG
di: Lu, Jacky Tai-Yu, et al.
Pubblicazione: (2025)
di: Lu, Jacky Tai-Yu, et al.
Pubblicazione: (2025)
End-to-end Semantic-centric Video-based Multimodal Affective Computing
di: Lin, Ronghao, et al.
Pubblicazione: (2024)
di: Lin, Ronghao, et al.
Pubblicazione: (2024)
MultiScript30k: Leveraging Multilingual Embeddings to Extend Cross Script Parallel Data
di: Driggers-Ellis, Christopher, et al.
Pubblicazione: (2025)
di: Driggers-Ellis, Christopher, et al.
Pubblicazione: (2025)
Adaptive Prototype Knowledge Transfer for Federated Learning with Mixed Modalities and Heterogeneous Tasks
di: Gai, Keke, et al.
Pubblicazione: (2025)
di: Gai, Keke, et al.
Pubblicazione: (2025)
Multi-Modal Multi-Task Federated Foundation Models for Next-Generation Extended Reality Systems: Towards Privacy-Preserving Distributed Intelligence in AR/VR/MR
di: Nadimi, Fardis, et al.
Pubblicazione: (2025)
di: Nadimi, Fardis, et al.
Pubblicazione: (2025)
MultiMedEdit: A Scenario-Aware Benchmark for Evaluating Knowledge Editing in Medical VQA
di: Wen, Shengtao, et al.
Pubblicazione: (2025)
di: Wen, Shengtao, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Learning Audio-Visual Embeddings with Inferred Latent Interaction Graphs
di: Zeng, Donghuo, et al.
Pubblicazione: (2026) -
EquiAV: Leveraging Equivariance for Audio-Visual Contrastive Learning
di: Kim, Jongsuk, et al.
Pubblicazione: (2024) -
Hierarchical Semantic Correlation-Aware Masked Autoencoder for Unsupervised Audio-Visual Representation Learning
di: Zeng, Donghuo, et al.
Pubblicazione: (2026) -
Low-Resolution Object Recognition with Cross-Resolution Relational Contrastive Distillation
di: Zhang, Kangkai, et al.
Pubblicazione: (2024) -
DRAGON: Distributional Rewards Optimize Diffusion Generative Models
di: Bai, Yatong, et al.
Pubblicazione: (2025)