DALR: Dual-level Alignment Learning for Multimodal Sentence Representation Learning
Fuente:
arXiv
Saved in:
| Main Authors: | He, Kang, Ding, Yuzhe, Wang, Haining, Li, Fei, Teng, Chong, Ji, Donghong |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Enhance-then-Balance Modality Collaboration for Robust Multimodal Sentiment Analysis
by: He, Kang, et al.
Published: (2026)
by: He, Kang, et al.
Published: (2026)
Dynamic Emotion and Personality Profiling for Multimodal Deception Detection
by: Zheng, Li, et al.
Published: (2026)
by: Zheng, Li, et al.
Published: (2026)
Zero-Shot Conversational Stance Detection: Dataset and Approaches
by: Ding, Yuzhe, et al.
Published: (2025)
by: Ding, Yuzhe, et al.
Published: (2025)
PaSE: Prototype-aligned Calibration and Shapley-based Equilibrium for Multimodal Sentiment Analysis
by: He, Kang, et al.
Published: (2025)
by: He, Kang, et al.
Published: (2025)
CMNER: A Chinese Multimodal NER Dataset based on Social Media
by: Ji, Yuanze, et al.
Published: (2024)
by: Ji, Yuanze, et al.
Published: (2024)
Multi-Granular Multimodal Clue Fusion for Meme Understanding
by: Zheng, Li, et al.
Published: (2025)
by: Zheng, Li, et al.
Published: (2025)
Code-MIE: A Code-style Model for Multimodal Information Extraction with Scene Graph and Entity Attribute Knowledge Enhancement
by: Liu, Jiang, et al.
Published: (2026)
by: Liu, Jiang, et al.
Published: (2026)
Revisiting Structured Sentiment Analysis as Latent Dependency Graph Parsing
by: Zhou, Chengjie, et al.
Published: (2024)
by: Zhou, Chengjie, et al.
Published: (2024)
DiscoSG: Towards Discourse-Level Text Scene Graph Parsing through Iterative Graph Refinement
by: Lin, Shaoqing, et al.
Published: (2025)
by: Lin, Shaoqing, et al.
Published: (2025)
Are Emotion and Rhetoric Neurons in LLM? Neuron Recognition and Adaptive Masking for Emotion-Rhetoric Prediction Steering
by: Zheng, Li, et al.
Published: (2026)
by: Zheng, Li, et al.
Published: (2026)
Modeling Unified Semantic Discourse Structure for High-quality Headline Generation
by: Xu, Minghui, et al.
Published: (2024)
by: Xu, Minghui, et al.
Published: (2024)
Improving Multimodal Contrastive Learning of Sentence Embeddings with Object-Phrase Alignment
by: Zhao, Kaiyan, et al.
Published: (2025)
by: Zhao, Kaiyan, et al.
Published: (2025)
Enhancing Hyperbole and Metaphor Detection with Their Bidirectional Dynamic Interaction and Emotion Knowledge
by: Zheng, Li, et al.
Published: (2025)
by: Zheng, Li, et al.
Published: (2025)
M$^{3}$D: A Multimodal, Multilingual and Multitask Dataset for Grounded Document-level Information Extraction
by: Liu, Jiang, et al.
Published: (2024)
by: Liu, Jiang, et al.
Published: (2024)
Harvesting Events from Multiple Sources: Towards a Cross-Document Event Extraction Paradigm
by: Gao, Qiang, et al.
Published: (2024)
by: Gao, Qiang, et al.
Published: (2024)
SDA: Simple Discrete Augmentation for Contrastive Sentence Representation Learning
by: Zhu, Dongsheng, et al.
Published: (2022)
by: Zhu, Dongsheng, et al.
Published: (2022)
TRIDENT: Enhancing Large Language Model Safety with Tri-Dimensional Diversified Red-Teaming Data Synthesis
by: Wu, Xiaorui, et al.
Published: (2025)
by: Wu, Xiaorui, et al.
Published: (2025)
Pixel Sentence Representation Learning
by: Xiao, Chenghao, et al.
Published: (2024)
by: Xiao, Chenghao, et al.
Published: (2024)
Cross-Document Cross-Lingual NLI via RST-Enhanced Graph Fusion and Interpretability Prediction
by: Yuan, Mengying, et al.
Published: (2025)
by: Yuan, Mengying, et al.
Published: (2025)
DenoSent: A Denoising Objective for Self-Supervised Sentence Representation Learning
by: Wang, Xinghao, et al.
Published: (2024)
by: Wang, Xinghao, et al.
Published: (2024)
Enhancing Cross-Document Event Coreference Resolution by Discourse Structure and Semantic Information
by: Gao, Qiang, et al.
Published: (2024)
by: Gao, Qiang, et al.
Published: (2024)
Data-CUBE: Data Curriculum for Instruction-based Sentence Representation Learning
by: Min, Yingqian, et al.
Published: (2024)
by: Min, Yingqian, et al.
Published: (2024)
Towards Linguistic Neural Representation Learning and Sentence Retrieval from Electroencephalogram Recordings
by: Zhou, Jinzhao, et al.
Published: (2024)
by: Zhou, Jinzhao, et al.
Published: (2024)
Dual Alignment Between Language Model Layers and Human Sentence Processing
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
by: Kuribayashi, Tatsuki, et al.
Published: (2026)
Large Language Models can Contrastively Refine their Generation for Better Sentence Representation Learning
by: Wang, Huiming, et al.
Published: (2023)
by: Wang, Huiming, et al.
Published: (2023)
S2Sent: Nested Selectivity Aware Sentence Representation Learning
by: Zang, Jianxiang, et al.
Published: (2025)
by: Zang, Jianxiang, et al.
Published: (2025)
Towards Better Understanding of Contrastive Sentence Representation Learning: A Unified Paradigm for Gradient
by: Li, Mingxin, et al.
Published: (2024)
by: Li, Mingxin, et al.
Published: (2024)
PropNet: a White-Box and Human-Like Network for Sentence Representation
by: Yang, Fei
Published: (2025)
by: Yang, Fei
Published: (2025)
CSE-SFP: Enabling Unsupervised Sentence Representation Learning via a Single Forward Pass
by: Zhang, Bowen, et al.
Published: (2025)
by: Zhang, Bowen, et al.
Published: (2025)
SETUP: Sentence-level English-To-Uniform Meaning Representation Parser
by: Markle, Emma, et al.
Published: (2025)
by: Markle, Emma, et al.
Published: (2025)
One Sentence, Two Embeddings: Contrastive Learning of Explicit and Implicit Semantic Representations
by: Oda, Kohei, et al.
Published: (2025)
by: Oda, Kohei, et al.
Published: (2025)
Label Confidence Weighted Learning for Target-level Sentence Simplification
by: Qiu, Xinying, et al.
Published: (2024)
by: Qiu, Xinying, et al.
Published: (2024)
CodeBind: Decoupled Representation Learning for Multimodal Alignment with Unified Compositional Codebook
by: Chen, Zeyu, et al.
Published: (2026)
by: Chen, Zeyu, et al.
Published: (2026)
Improving In-context Learning of Multilingual Generative Language Models with Cross-lingual Alignment
by: Li, Chong, et al.
Published: (2023)
by: Li, Chong, et al.
Published: (2023)
Stream Aligner: Efficient Sentence-Level Alignment via Distribution Induction
by: Lou, Hantao, et al.
Published: (2025)
by: Lou, Hantao, et al.
Published: (2025)
KDMCSE: Knowledge Distillation Multimodal Sentence Embeddings with Adaptive Angular margin Contrastive Learning
by: Nguyen, Cong-Duy, et al.
Published: (2024)
by: Nguyen, Cong-Duy, et al.
Published: (2024)
Decoding the Multimodal Mind: Generalizable Brain-to-Text Translation via Multimodal Alignment and Adaptive Routing
by: Ye, Chunyu, et al.
Published: (2025)
by: Ye, Chunyu, et al.
Published: (2025)
SetCSE: Set Operations using Contrastive Learning of Sentence Embeddings
by: Liu, Kang
Published: (2024)
by: Liu, Kang
Published: (2024)
Expand BERT Representation with Visual Information via Grounded Language Learning with Multimodal Partial Alignment
by: Nguyen, Cong-Duy, et al.
Published: (2023)
by: Nguyen, Cong-Duy, et al.
Published: (2023)
Sentence Representations via Gaussian Embedding
by: Yoda, Shohei, et al.
Published: (2023)
by: Yoda, Shohei, et al.
Published: (2023)
Similar Items
-
Enhance-then-Balance Modality Collaboration for Robust Multimodal Sentiment Analysis
by: He, Kang, et al.
Published: (2026) -
Dynamic Emotion and Personality Profiling for Multimodal Deception Detection
by: Zheng, Li, et al.
Published: (2026) -
Zero-Shot Conversational Stance Detection: Dataset and Approaches
by: Ding, Yuzhe, et al.
Published: (2025) -
PaSE: Prototype-aligned Calibration and Shapley-based Equilibrium for Multimodal Sentiment Analysis
by: He, Kang, et al.
Published: (2025) -
CMNER: A Chinese Multimodal NER Dataset based on Social Media
by: Ji, Yuanze, et al.
Published: (2024)