CPCLDETECTOR: Knowledge Enhancement and Alignment Selection for Chinese Patronizing and Condescending Language Detection
Fuente:
arXiv
Salvato in:
| Autori principali: | Yang, Jiaxun, Han, Yifei, Zhang, Long, Liu, Yujie, Li, Bin, Gao, Bo, He, Yangfan, Zhan, Kejia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
SonicSense: Object Perception from In-Hand Acoustic Vibration
di: Liu, Jiaxun, et al.
Pubblicazione: (2024)
di: Liu, Jiaxun, et al.
Pubblicazione: (2024)
DiffCL: A Diffusion-Based Contrastive Learning Framework with Semantic Alignment for Multimodal Recommendations
di: Song, Qiya, et al.
Pubblicazione: (2025)
di: Song, Qiya, et al.
Pubblicazione: (2025)
Adaptive Offloading and Enhancement for Low-Light Video Analytics on Mobile Devices
di: He, Yuanyi, et al.
Pubblicazione: (2024)
di: He, Yuanyi, et al.
Pubblicazione: (2024)
COPA: Efficient Vision-Language Pre-training Through Collaborative Object- and Patch-Text Alignment
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
di: Jiang, Chaoya, et al.
Pubblicazione: (2023)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
SEER: Semantic Enhancement and Emotional Reasoning Network for Multimodal Fake News Detection
di: Zhu, Peican, et al.
Pubblicazione: (2025)
di: Zhu, Peican, et al.
Pubblicazione: (2025)
ISMAF: Intrinsic-Social Modality Alignment and Fusion for Multimodal Rumor Detection
di: Yu, Zihao, et al.
Pubblicazione: (2025)
di: Yu, Zihao, et al.
Pubblicazione: (2025)
Distilling Implicit Multimodal Knowledge into Large Language Models for Zero-Resource Dialogue Generation
di: Zhang, Bo, et al.
Pubblicazione: (2024)
di: Zhang, Bo, et al.
Pubblicazione: (2024)
EmotionTalk: An Interactive Chinese Multimodal Emotion Dataset With Rich Annotations
di: Sun, Haoqin, et al.
Pubblicazione: (2025)
di: Sun, Haoqin, et al.
Pubblicazione: (2025)
MM-InstructEval: Zero-Shot Evaluation of (Multimodal) Large Language Models on Multimodal Reasoning Tasks
di: Yang, Xiaocui, et al.
Pubblicazione: (2024)
di: Yang, Xiaocui, et al.
Pubblicazione: (2024)
MAGE: Multimodal Alignment and Generation Enhancement via Bridging Visual and Semantic Spaces
di: E, Shaojun, et al.
Pubblicazione: (2025)
di: E, Shaojun, et al.
Pubblicazione: (2025)
Synchronized Video Storytelling: Generating Video Narrations with Structured Storyline
di: Yang, Dingyi, et al.
Pubblicazione: (2024)
di: Yang, Dingyi, et al.
Pubblicazione: (2024)
Muse: A Multimodal Conversational Recommendation Dataset with Scenario-Grounded User Profiles
di: Wang, Zihan, et al.
Pubblicazione: (2024)
di: Wang, Zihan, et al.
Pubblicazione: (2024)
Towards Patronizing and Condescending Language in Chinese Videos: A Multimodal Dataset and Detector
di: Wang, Hongbo, et al.
Pubblicazione: (2024)
di: Wang, Hongbo, et al.
Pubblicazione: (2024)
Towards Alleviating Text-to-Image Retrieval Hallucination for CLIP in Zero-shot Learning
di: Wang, Hanyao, et al.
Pubblicazione: (2024)
di: Wang, Hanyao, et al.
Pubblicazione: (2024)
MMPKUBase: A Comprehensive and High-quality Chinese Multi-modal Knowledge Graph
di: Yi, Xuan, et al.
Pubblicazione: (2024)
di: Yi, Xuan, et al.
Pubblicazione: (2024)
Video Echoed in Music: Semantic, Temporal, and Rhythmic Alignment for Video-to-Music Generation
di: Tong, Xinyi, et al.
Pubblicazione: (2025)
di: Tong, Xinyi, et al.
Pubblicazione: (2025)
Identity-Aware Vision-Language Model for Explainable Face Forgery Detection
di: Xu, Junhao, et al.
Pubblicazione: (2025)
di: Xu, Junhao, et al.
Pubblicazione: (2025)
Knowledge-aware Diffusion-Enhanced Multimedia Recommendation
di: Mo, Xian, et al.
Pubblicazione: (2025)
di: Mo, Xian, et al.
Pubblicazione: (2025)
Multi-source Knowledge Enhanced Graph Attention Networks for Multimodal Fact Verification
di: Cao, Han, et al.
Pubblicazione: (2024)
di: Cao, Han, et al.
Pubblicazione: (2024)
Fine-grained Knowledge Graph-driven Video-Language Learning for Action Recognition
di: Zhang, Rui, et al.
Pubblicazione: (2024)
di: Zhang, Rui, et al.
Pubblicazione: (2024)
PclGPT: A Large Language Model for Patronizing and Condescending Language Detection
di: Wang, Hongbo, et al.
Pubblicazione: (2024)
di: Wang, Hongbo, et al.
Pubblicazione: (2024)
Audio-Visual Separation with Hierarchical Fusion and Representation Alignment
di: Hu, Han, et al.
Pubblicazione: (2025)
di: Hu, Han, et al.
Pubblicazione: (2025)
ConCLVD: Controllable Chinese Landscape Video Generation via Diffusion Model
di: Liu, Dingming, et al.
Pubblicazione: (2024)
di: Liu, Dingming, et al.
Pubblicazione: (2024)
An Emotion Recognition Framework via Cross-modal Alignment of EEG and Eye Movement Data
di: Wang, Jianlu, et al.
Pubblicazione: (2025)
di: Wang, Jianlu, et al.
Pubblicazione: (2025)
EmpathyEar: An Open-source Avatar Multimodal Empathetic Chatbot
di: Fei, Hao, et al.
Pubblicazione: (2024)
di: Fei, Hao, et al.
Pubblicazione: (2024)
Is One-Shot In-Context Learning Helpful for Data Selection in Task-Specific Fine-Tuning of Multimodal LLMs?
di: An, Xiao, et al.
Pubblicazione: (2026)
di: An, Xiao, et al.
Pubblicazione: (2026)
Identity-Preserving Text-to-Video Generation via Training-Free Prompt, Image, and Guidance Enhancement
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
di: Gao, Jiayi, et al.
Pubblicazione: (2025)
Towards Structure-aware Model for Multi-modal Knowledge Graph Completion
di: Li, Linyu, et al.
Pubblicazione: (2025)
di: Li, Linyu, et al.
Pubblicazione: (2025)
Decompose and Transfer: CoT-Prompting Enhanced Alignment for Open-Vocabulary Temporal Action Detection
di: Zhu, Sa, et al.
Pubblicazione: (2026)
di: Zhu, Sa, et al.
Pubblicazione: (2026)
Signal: Selective Interaction and Global-local Alignment for Multi-Modal Object Re-Identification
di: Liu, Yangyang, et al.
Pubblicazione: (2025)
di: Liu, Yangyang, et al.
Pubblicazione: (2025)
Hybrid CNN-Mamba Enhancement Network for Robust Multimodal Sentiment Analysis
di: Li, Xiang, et al.
Pubblicazione: (2025)
di: Li, Xiang, et al.
Pubblicazione: (2025)
From Natural Alignment to Conditional Controllability in Multimodal Dialogue
di: Jin, Zeyu, et al.
Pubblicazione: (2026)
di: Jin, Zeyu, et al.
Pubblicazione: (2026)
Integrated Semantic and Temporal Alignment for Interactive Video Retrieval
di: Luu, Thanh-Danh, et al.
Pubblicazione: (2025)
di: Luu, Thanh-Danh, et al.
Pubblicazione: (2025)
Enhancing Partially Relevant Video Retrieval with Robust Alignment Learning
di: Zhang, Long, et al.
Pubblicazione: (2025)
di: Zhang, Long, et al.
Pubblicazione: (2025)
Exploring the Role of Audio in Multimodal Misinformation Detection
di: Liu, Moyang, et al.
Pubblicazione: (2024)
di: Liu, Moyang, et al.
Pubblicazione: (2024)
Advancing Unsupervised Low-light Image Enhancement: Noise Estimation, Illumination Interpolation, and Self-Regulation
di: Liu, Xiaofeng, et al.
Pubblicazione: (2023)
di: Liu, Xiaofeng, et al.
Pubblicazione: (2023)
Multimodal Classification and Out-of-distribution Detection for Multimodal Intent Understanding
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
CLIP-PCQA: Exploring Subjective-Aligned Vision-Language Modeling for Point Cloud Quality Assessment
di: Liu, Yating, et al.
Pubblicazione: (2025)
di: Liu, Yating, et al.
Pubblicazione: (2025)
Identity-Driven Multimedia Forgery Detection via Reference Assistance
di: Xu, Junhao, et al.
Pubblicazione: (2024)
di: Xu, Junhao, et al.
Pubblicazione: (2024)
Documenti analoghi
-
SonicSense: Object Perception from In-Hand Acoustic Vibration
di: Liu, Jiaxun, et al.
Pubblicazione: (2024) -
DiffCL: A Diffusion-Based Contrastive Learning Framework with Semantic Alignment for Multimodal Recommendations
di: Song, Qiya, et al.
Pubblicazione: (2025) -
Adaptive Offloading and Enhancement for Low-Light Video Analytics on Mobile Devices
di: He, Yuanyi, et al.
Pubblicazione: (2024) -
COPA: Efficient Vision-Language Pre-training Through Collaborative Object- and Patch-Text Alignment
di: Jiang, Chaoya, et al.
Pubblicazione: (2023) -
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)