Enhancing Few-Shot Classification without Forgetting through Multi-Level Contrastive Constraints
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Bingzhi, Zhou, Haoming, Liu, Yishu, Zeng, Biqing, Pan, Jiahui, Lu, Guangming |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Cross Modal Fine-Grained Alignment via Granularity-Aware and Region-Uncertain Modeling
di: Liu, Jiale, et al.
Pubblicazione: (2025)
di: Liu, Jiale, et al.
Pubblicazione: (2025)
Multi-view Hypergraph-based Contrastive Learning Model for Cold-Start Micro-video Recommendation
di: Lyu, Sisuo, et al.
Pubblicazione: (2024)
di: Lyu, Sisuo, et al.
Pubblicazione: (2024)
MSVBench: Towards Human-Level Evaluation of Multi-Shot Video Generation
di: Shi, Haoyuan, et al.
Pubblicazione: (2026)
di: Shi, Haoyuan, et al.
Pubblicazione: (2026)
Semantic Compensation via Adversarial Removal for Robust Zero-Shot ECG Diagnosis
di: Liu, Hongjun, et al.
Pubblicazione: (2026)
di: Liu, Hongjun, et al.
Pubblicazione: (2026)
Token-Level Contrastive Learning with Modality-Aware Prompting for Multimodal Intent Recognition
di: Zhou, Qianrui, et al.
Pubblicazione: (2023)
di: Zhou, Qianrui, et al.
Pubblicazione: (2023)
Multimodal Pre-training Framework for Sequential Recommendation via Contrastive Learning
di: Zhang, Lingzi, et al.
Pubblicazione: (2023)
di: Zhang, Lingzi, et al.
Pubblicazione: (2023)
TALDS-Net: Task-Aware Adaptive Local Descriptors Selection for Few-shot Image Classification
di: Qiao, Qian, et al.
Pubblicazione: (2023)
di: Qiao, Qian, et al.
Pubblicazione: (2023)
Multi-Reference Generative Face Video Compression with Contrastive Learning
di: Konuko, Goluck, et al.
Pubblicazione: (2024)
di: Konuko, Goluck, et al.
Pubblicazione: (2024)
FineBadminton: A Multi-Level Dataset for Fine-Grained Badminton Video Understanding
di: He, Xusheng, et al.
Pubblicazione: (2025)
di: He, Xusheng, et al.
Pubblicazione: (2025)
CARAT: Contrastive Feature Reconstruction and Aggregation for Multi-Modal Multi-Label Emotion Recognition
di: Peng, Cheng, et al.
Pubblicazione: (2023)
di: Peng, Cheng, et al.
Pubblicazione: (2023)
Private Speech Classification without Collapse: Stabilized DP Training and Offline Distillation
di: Wen, Yadi, et al.
Pubblicazione: (2026)
di: Wen, Yadi, et al.
Pubblicazione: (2026)
Multi-source Knowledge Enhanced Graph Attention Networks for Multimodal Fact Verification
di: Cao, Han, et al.
Pubblicazione: (2024)
di: Cao, Han, et al.
Pubblicazione: (2024)
PRM-BAS: Enhancing Multimodal Reasoning through PRM-guided Beam Annealing Search
di: Hu, Pengfei, et al.
Pubblicazione: (2025)
di: Hu, Pengfei, et al.
Pubblicazione: (2025)
Contrast then Memorize: Semantic Neighbor Retrieval-Enhanced Inductive Multimodal Knowledge Graph Completion
di: Zhao, Yu, et al.
Pubblicazione: (2024)
di: Zhao, Yu, et al.
Pubblicazione: (2024)
Modality-Aware Contrastive and Uncertainty-Regularized Emotion Recognition
di: Zhuang, Yan, et al.
Pubblicazione: (2026)
di: Zhuang, Yan, et al.
Pubblicazione: (2026)
Multimodal Graph-Based Variational Mixture of Experts Network for Zero-Shot Multimodal Information Extraction
di: Zhou, Baohang, et al.
Pubblicazione: (2025)
di: Zhou, Baohang, et al.
Pubblicazione: (2025)
Exploring the Robustness of Decision-Level Through Adversarial Attacks on LLM-Based Embodied Models
di: Liu, Shuyuan, et al.
Pubblicazione: (2024)
di: Liu, Shuyuan, et al.
Pubblicazione: (2024)
Connecting Giants: Synergistic Knowledge Transfer of Large Multimodal Models for Few-Shot Learning
di: Tang, Hao, et al.
Pubblicazione: (2025)
di: Tang, Hao, et al.
Pubblicazione: (2025)
Beyond Patches: Global-aware Autoregressive Model for Multimodal Few-Shot Font Generation
di: Cai, Haonan, et al.
Pubblicazione: (2026)
di: Cai, Haonan, et al.
Pubblicazione: (2026)
Hyperbolic Multimodal Generative Representation Learning for Generalized Zero-Shot Multimodal Information Extraction
di: Zhou, Baohang, et al.
Pubblicazione: (2026)
di: Zhou, Baohang, et al.
Pubblicazione: (2026)
CineAGI: Character-Consistent Movie Creation through LLM-Orchestrated Multi-Modal Generation and Cross-Scene Integration
di: Xie, Tianyidan, et al.
Pubblicazione: (2026)
di: Xie, Tianyidan, et al.
Pubblicazione: (2026)
Comparing Contrastive and Triplet Loss: Variance Analysis and Optimization Behavior
di: Zeng, Donghuo
Pubblicazione: (2025)
di: Zeng, Donghuo
Pubblicazione: (2025)
EEmo-Bench: A Benchmark for Multi-modal Large Language Models on Image Evoked Emotion Assessment
di: Gao, Lancheng, et al.
Pubblicazione: (2025)
di: Gao, Lancheng, et al.
Pubblicazione: (2025)
Multimodal Classification and Out-of-distribution Detection for Multimodal Intent Understanding
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
TeMTG: Text-Enhanced Multi-Hop Temporal Graph Modeling for Audio-Visual Video Parsing
di: Chen, Yaru, et al.
Pubblicazione: (2025)
di: Chen, Yaru, et al.
Pubblicazione: (2025)
ZO-ASR: Zeroth-Order Fine-Tuning of Speech Foundation Models without Back-Propagation
di: Peng, Yuezhang, et al.
Pubblicazione: (2025)
di: Peng, Yuezhang, et al.
Pubblicazione: (2025)
Contrastive Knowledge Distillation for Robust Multimodal Sentiment Analysis
di: Sang, Zhongyi, et al.
Pubblicazione: (2024)
di: Sang, Zhongyi, et al.
Pubblicazione: (2024)
Beyond Isolated Utterances: Cue-Guided Interaction for Context-Dependent Conversational Multimodal Understanding
di: Pan, Zhaoyan, et al.
Pubblicazione: (2026)
di: Pan, Zhaoyan, et al.
Pubblicazione: (2026)
ALMol: Aligned Language-Molecule Translation LLMs through Offline Preference Contrastive Optimisation
di: Gkoumas, Dimitris
Pubblicazione: (2024)
di: Gkoumas, Dimitris
Pubblicazione: (2024)
Rethinking Few-Shot Adaptation of Vision-Language Models in Two Stages
di: Farina, Matteo, et al.
Pubblicazione: (2025)
di: Farina, Matteo, et al.
Pubblicazione: (2025)
Retrieval Augmented Verification for Zero-Shot Detection of Multimodal Disinformation
di: Dey, Arka Ujjal, et al.
Pubblicazione: (2024)
di: Dey, Arka Ujjal, et al.
Pubblicazione: (2024)
Characterizing Multimedia Information Environment through Multi-modal Clustering of YouTube Videos
di: Yousefi, Niloofar, et al.
Pubblicazione: (2024)
di: Yousefi, Niloofar, et al.
Pubblicazione: (2024)
Across-Game Engagement Modelling via Few-Shot Learning
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
di: Pinitas, Kosmas, et al.
Pubblicazione: (2024)
Deep Contrastive Multi-view Clustering under Semantic Feature Guidance
di: Liu, Siwen, et al.
Pubblicazione: (2024)
di: Liu, Siwen, et al.
Pubblicazione: (2024)
PSA-MF: Personality-Sentiment Aligned Multi-Level Fusion for Multimodal Sentiment Analysis
di: Xie, Heng, et al.
Pubblicazione: (2025)
di: Xie, Heng, et al.
Pubblicazione: (2025)
Simple but Effective Raw-Data Level Multimodal Fusion for Composed Image Retrieval
di: Wen, Haokun, et al.
Pubblicazione: (2024)
di: Wen, Haokun, et al.
Pubblicazione: (2024)
HDA-SELD: Hierarchical Cross-Modal Distillation with Multi-Level Data Augmentation for Low-Resource Audio-Visual Sound Event Localization and Detection
di: Wang, Qing, et al.
Pubblicazione: (2025)
di: Wang, Qing, et al.
Pubblicazione: (2025)
TimeNeRF: Building Generalizable Neural Radiance Fields across Time from Few-Shot Input Views
di: Hung, Hsiang-Hui, et al.
Pubblicazione: (2025)
di: Hung, Hsiang-Hui, et al.
Pubblicazione: (2025)
Multimodal Fusion via Hypergraph Autoencoder and Contrastive Learning for Emotion Recognition in Conversation
di: Yi, Zijian, et al.
Pubblicazione: (2024)
di: Yi, Zijian, et al.
Pubblicazione: (2024)
Deep Learning Classification of Photoplethysmogram Signal for Hypertension Levels
di: Nasir, Nida, et al.
Pubblicazione: (2024)
di: Nasir, Nida, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Cross Modal Fine-Grained Alignment via Granularity-Aware and Region-Uncertain Modeling
di: Liu, Jiale, et al.
Pubblicazione: (2025) -
Multi-view Hypergraph-based Contrastive Learning Model for Cold-Start Micro-video Recommendation
di: Lyu, Sisuo, et al.
Pubblicazione: (2024) -
MSVBench: Towards Human-Level Evaluation of Multi-Shot Video Generation
di: Shi, Haoyuan, et al.
Pubblicazione: (2026) -
Semantic Compensation via Adversarial Removal for Robust Zero-Shot ECG Diagnosis
di: Liu, Hongjun, et al.
Pubblicazione: (2026) -
Token-Level Contrastive Learning with Modality-Aware Prompting for Multimodal Intent Recognition
di: Zhou, Qianrui, et al.
Pubblicazione: (2023)