Concept Drift Guided LayerNorm Tuning for Efficient Multimodal Metaphor Identification
Fuente:
arXiv
Salvato in:
| Autori principali: | Qian, Wenhao, Hu, Zhenzhen, Song, Zijie, Li, Jia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition
di: Yu, Yangchen, et al.
Pubblicazione: (2026)
di: Yu, Yangchen, et al.
Pubblicazione: (2026)
Balanced Multimodal Learning: An Unidirectional Dynamic Interaction Perspective
di: Wang, Shijie, et al.
Pubblicazione: (2025)
di: Wang, Shijie, et al.
Pubblicazione: (2025)
Geometry and Dynamics of LayerNorm
di: Riechers, Paul M.
Pubblicazione: (2024)
di: Riechers, Paul M.
Pubblicazione: (2024)
Multimodal Representation Learning and Fusion
di: Jin, Qihang, et al.
Pubblicazione: (2025)
di: Jin, Qihang, et al.
Pubblicazione: (2025)
Listening to the Unspoken: Exploring "365" Aspects of Multimodal Interview Performance Assessment
di: Li, Jia, et al.
Pubblicazione: (2025)
di: Li, Jia, et al.
Pubblicazione: (2025)
L3GS: Layered 3D Gaussian Splats for Efficient 3D Scene Delivery
di: Tsai, Yi-Zhen, et al.
Pubblicazione: (2025)
di: Tsai, Yi-Zhen, et al.
Pubblicazione: (2025)
FedNano: Toward Lightweight Federated Tuning for Pretrained Multimodal Large Language Models
di: Zhang, Yao, et al.
Pubblicazione: (2025)
di: Zhang, Yao, et al.
Pubblicazione: (2025)
On the Role of Attention Masks and LayerNorm in Transformers
di: Wu, Xinyi, et al.
Pubblicazione: (2024)
di: Wu, Xinyi, et al.
Pubblicazione: (2024)
HyperFusion: Hierarchical Multimodal Ensemble Learning for Social Media Popularity Prediction
di: Ye, Liliang, et al.
Pubblicazione: (2025)
di: Ye, Liliang, et al.
Pubblicazione: (2025)
Exploring Modality Disruption in Multimodal Fake News Detection
di: Liu, Moyang, et al.
Pubblicazione: (2025)
di: Liu, Moyang, et al.
Pubblicazione: (2025)
Hybrid Feedback-Guided Optimal Learning for Wireless Interactive Panoramic Scene Delivery
di: Wu, Xiaoyi, et al.
Pubblicazione: (2026)
di: Wu, Xiaoyi, et al.
Pubblicazione: (2026)
Transformers Don't Need LayerNorm at Inference Time: Scaling LayerNorm Removal to GPT-2 XL and the Implications for Mechanistic Interpretability
di: Baroni, Luca, et al.
Pubblicazione: (2025)
di: Baroni, Luca, et al.
Pubblicazione: (2025)
Deconfounded Reasoning for Multimodal Fake News Detection via Causal Intervention
di: Liu, Moyang, et al.
Pubblicazione: (2025)
di: Liu, Moyang, et al.
Pubblicazione: (2025)
Token-Level Contrastive Learning with Modality-Aware Prompting for Multimodal Intent Recognition
di: Zhou, Qianrui, et al.
Pubblicazione: (2023)
di: Zhou, Qianrui, et al.
Pubblicazione: (2023)
Hyper-modal Imputation Diffusion Embedding with Dual-Distillation for Federated Multimodal Knowledge Graph Completion
di: Zhang, Ying, et al.
Pubblicazione: (2025)
di: Zhang, Ying, et al.
Pubblicazione: (2025)
Zero-Shot Relational Learning for Multimodal Knowledge Graphs
di: Cai, Rui, et al.
Pubblicazione: (2024)
di: Cai, Rui, et al.
Pubblicazione: (2024)
MCAD: Multimodal Context-Aware Audio Description Generation For Soccer
di: Chaudhary, Lipisha, et al.
Pubblicazione: (2025)
di: Chaudhary, Lipisha, et al.
Pubblicazione: (2025)
MCIGLE: Multimodal Exemplar-Free Class-Incremental Graph Learning
di: You, Haochen, et al.
Pubblicazione: (2025)
di: You, Haochen, et al.
Pubblicazione: (2025)
To Align or Not to Align: Strategic Multimodal Representation Alignment for Optimal Performance
di: Fang, Wanlong, et al.
Pubblicazione: (2025)
di: Fang, Wanlong, et al.
Pubblicazione: (2025)
Cross-Space Synergy: A Unified Framework for Multimodal Emotion Recognition in Conversation
di: Lyu, Xiaosen, et al.
Pubblicazione: (2025)
di: Lyu, Xiaosen, et al.
Pubblicazione: (2025)
Multimodal Methods for Analyzing Learning and Training Environments: A Systematic Literature Review
di: Cohn, Clayton, et al.
Pubblicazione: (2024)
di: Cohn, Clayton, et al.
Pubblicazione: (2024)
GAME-ON: Graph Attention Network based Multimodal Fusion for Fake News Detection
di: Dhawan, Mudit, et al.
Pubblicazione: (2022)
di: Dhawan, Mudit, et al.
Pubblicazione: (2022)
Stable Multimodal Graph Unlearning via Feature-Dimension Aware Quantile Selection
di: Zhou, Jingjing, et al.
Pubblicazione: (2026)
di: Zhou, Jingjing, et al.
Pubblicazione: (2026)
SLaNC: Static LayerNorm Calibration
di: Salmani, Mahsa, et al.
Pubblicazione: (2024)
di: Salmani, Mahsa, et al.
Pubblicazione: (2024)
PrismSSL: One Interface, Many Modalities; A Single-Interface Library for Multimodal Self-Supervised Learning
di: Shirian, Melika, et al.
Pubblicazione: (2025)
di: Shirian, Melika, et al.
Pubblicazione: (2025)
E2LVLM:Evidence-Enhanced Large Vision-Language Model for Multimodal Out-of-Context Misinformation Detection
di: Wu, Junjie, et al.
Pubblicazione: (2025)
di: Wu, Junjie, et al.
Pubblicazione: (2025)
LayerNorm Induces Recency Bias in Transformer Decoders
di: Kim, Junu, et al.
Pubblicazione: (2025)
di: Kim, Junu, et al.
Pubblicazione: (2025)
Atom: Efficient On-Device Video-Language Pipelines Through Modular Reuse
di: Panchal, Kunjal, et al.
Pubblicazione: (2025)
di: Panchal, Kunjal, et al.
Pubblicazione: (2025)
Efficient Distributed Training through Gradient Compression with Sparsification and Quantization Techniques
di: Singh, Shruti, et al.
Pubblicazione: (2024)
di: Singh, Shruti, et al.
Pubblicazione: (2024)
Enhancing Modality Representation and Alignment for Multimodal Cold-start Active Learning
di: Shen, Meng, et al.
Pubblicazione: (2024)
di: Shen, Meng, et al.
Pubblicazione: (2024)
OmniMER: Auxiliary-Enhanced LLM Adaptation for Indonesian Multimodal Emotion Recognition
di: Yan, Xueming, et al.
Pubblicazione: (2025)
di: Yan, Xueming, et al.
Pubblicazione: (2025)
Live Music Diffusion Models: Efficient Fine-Tuning and Post-Training of Interactive Diffusion Music Generators
di: Novack, Zachary, et al.
Pubblicazione: (2026)
di: Novack, Zachary, et al.
Pubblicazione: (2026)
Training Data Efficiency in Multimodal Process Reward Models
di: Li, Jinyuan, et al.
Pubblicazione: (2026)
di: Li, Jinyuan, et al.
Pubblicazione: (2026)
CLAIP-Emo: Parameter-Efficient Adaptation of Language-supervised models for In-the-Wild Audiovisual Emotion Recognition
di: Chen, Yin, et al.
Pubblicazione: (2025)
di: Chen, Yin, et al.
Pubblicazione: (2025)
ChemDFM-X: Towards Large Multimodal Model for Chemistry
di: Zhao, Zihan, et al.
Pubblicazione: (2024)
di: Zhao, Zihan, et al.
Pubblicazione: (2024)
Enhancing Cross-Prompt Transferability in Vision-Language Models through Contextual Injection of Target Tokens
di: Yang, Xikang, et al.
Pubblicazione: (2024)
di: Yang, Xikang, et al.
Pubblicazione: (2024)
Traits Run Deep: Enhancing Personality Assessment via Psychology-Guided LLM Representations and Multimodal Apparent Behaviors
di: Li, Jia, et al.
Pubblicazione: (2025)
di: Li, Jia, et al.
Pubblicazione: (2025)
Post-LayerNorm Is Back: Stable, ExpressivE, and Deep
di: Chen, Chen, et al.
Pubblicazione: (2026)
di: Chen, Chen, et al.
Pubblicazione: (2026)
Human-centered Interactive Learning via MLLMs for Text-to-Image Person Re-identification
di: Qin, Yang, et al.
Pubblicazione: (2025)
di: Qin, Yang, et al.
Pubblicazione: (2025)
InfoMAE: Pair-Efficient Cross-Modal Alignment for Multimodal Time-Series Sensing Signals
di: Kimura, Tomoyoshi, et al.
Pubblicazione: (2025)
di: Kimura, Tomoyoshi, et al.
Pubblicazione: (2025)
Documenti analoghi
-
To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition
di: Yu, Yangchen, et al.
Pubblicazione: (2026) -
Balanced Multimodal Learning: An Unidirectional Dynamic Interaction Perspective
di: Wang, Shijie, et al.
Pubblicazione: (2025) -
Geometry and Dynamics of LayerNorm
di: Riechers, Paul M.
Pubblicazione: (2024) -
Multimodal Representation Learning and Fusion
di: Jin, Qihang, et al.
Pubblicazione: (2025) -
Listening to the Unspoken: Exploring "365" Aspects of Multimodal Interview Performance Assessment
di: Li, Jia, et al.
Pubblicazione: (2025)