MORALISE: A Structured Benchmark for Moral Alignment in Visual Language Models
Fuente:
arXiv
Salvato in:
| Autori principali: | Lin, Xiao, Liu, Zhining, Yang, Ze, Li, Gaotang, Qiu, Ruizhong, Wang, Shuke, Liu, Hui, Li, Haotian, Keswani, Sumit, Pardeshi, Vishwa, Zhao, Huijun, Fan, Wei, Tong, Hanghang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Do VLMs Have a Moral Backbone? A Study on the Fragile Morality of Vision-Language Models
di: Liu, Zhining, et al.
Pubblicazione: (2026)
di: Liu, Zhining, et al.
Pubblicazione: (2026)
A Collaborative Extended Reality Prototype for 3D Surgical Planning and Visualization
di: Qiu, Shi, et al.
Pubblicazione: (2026)
di: Qiu, Shi, et al.
Pubblicazione: (2026)
MMAPS: End-to-End Multi-Grained Multi-Modal Attribute-Aware Product Summarization
di: Chen, Tao, et al.
Pubblicazione: (2023)
di: Chen, Tao, et al.
Pubblicazione: (2023)
MAC-SLU: Multi-Intent Automotive Cabin Spoken Language Understanding Benchmark
di: Peng, Yuezhang, et al.
Pubblicazione: (2025)
di: Peng, Yuezhang, et al.
Pubblicazione: (2025)
V-FAT: Benchmarking Visual Fidelity Against Text-bias
di: Wang, Ziteng, et al.
Pubblicazione: (2026)
di: Wang, Ziteng, et al.
Pubblicazione: (2026)
CvhSlicer 2.0: Immersive and Interactive Visualization of Chinese Visible Human Data in XR Environments
di: Qiu, Yue, et al.
Pubblicazione: (2025)
di: Qiu, Yue, et al.
Pubblicazione: (2025)
LoginMEA: Local-to-Global Interaction Network for Multi-modal Entity Alignment
di: Su, Taoyu, et al.
Pubblicazione: (2024)
di: Su, Taoyu, et al.
Pubblicazione: (2024)
IBMEA: Exploring Variational Information Bottleneck for Multi-modal Entity Alignment
di: Su, Taoyu, et al.
Pubblicazione: (2024)
di: Su, Taoyu, et al.
Pubblicazione: (2024)
AlignVSR: Audio-Visual Cross-Modal Alignment for Visual Speech Recognition
di: Liu, Zehua, et al.
Pubblicazione: (2024)
di: Liu, Zehua, et al.
Pubblicazione: (2024)
Teacher-Guided Pseudo Supervision and Cross-Modal Alignment for Audio-Visual Video Parsing
di: Chen, Yaru, et al.
Pubblicazione: (2025)
di: Chen, Yaru, et al.
Pubblicazione: (2025)
MS2Mesh-XR: Multi-modal Sketch-to-Mesh Generation in XR Environments
di: Tong, Yuqi, et al.
Pubblicazione: (2024)
di: Tong, Yuqi, et al.
Pubblicazione: (2024)
Dependency Structure Augmented Contextual Scoping Framework for Multimodal Aspect-Based Sentiment Analysis
di: Liu, Hao, et al.
Pubblicazione: (2025)
di: Liu, Hao, et al.
Pubblicazione: (2025)
Shorter Is Different: Characterizing the Dynamics of Short-Form Video Platforms
di: Chen, Zhilong, et al.
Pubblicazione: (2024)
di: Chen, Zhilong, et al.
Pubblicazione: (2024)
XY-Cut++: Advanced Layout Ordering via Hierarchical Mask Mechanism on a Novel Benchmark
di: Liu, Shuai, et al.
Pubblicazione: (2025)
di: Liu, Shuai, et al.
Pubblicazione: (2025)
Interpretable Multimodal Misinformation Detection with Logic Reasoning
di: Liu, Hui, et al.
Pubblicazione: (2023)
di: Liu, Hui, et al.
Pubblicazione: (2023)
ASAP: Advancing Semantic Alignment Promotes Multi-Modal Manipulation Detecting and Grounding
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
di: Zhang, Zhenxing, et al.
Pubblicazione: (2024)
P2P: Automated Paper-to-Poster Generation and Fine-Grained Benchmark
di: Sun, Tao, et al.
Pubblicazione: (2025)
di: Sun, Tao, et al.
Pubblicazione: (2025)
MMESGBench: Pioneering Multimodal Understanding and Complex Reasoning Benchmark for ESG Tasks
di: Zhang, Lei, et al.
Pubblicazione: (2025)
di: Zhang, Lei, et al.
Pubblicazione: (2025)
VC-Bench: Pioneering the Video Connecting Benchmark with a Dataset and Evaluation Metrics
di: Yin, Zhiyu, et al.
Pubblicazione: (2026)
di: Yin, Zhiyu, et al.
Pubblicazione: (2026)
Save It for the "Hot" Day: An LLM-Empowered Visual Analytics System for Heat Risk Management
di: Li, Haobo, et al.
Pubblicazione: (2024)
di: Li, Haobo, et al.
Pubblicazione: (2024)
LaF-GRPO: In-Situ Navigation Instruction Generation for the Visually Impaired via GRPO with LLM-as-Follower Reward
di: Zhao, Yi, et al.
Pubblicazione: (2025)
di: Zhao, Yi, et al.
Pubblicazione: (2025)
Learned Image Compression with Hierarchical Progressive Context Modeling
di: Li, Yuqi, et al.
Pubblicazione: (2025)
di: Li, Yuqi, et al.
Pubblicazione: (2025)
PixCLIP: Achieving Fine-grained Visual Language Understanding via Any-granularity Pixel-Text Alignment Learning
di: Xiao, Yicheng, et al.
Pubblicazione: (2025)
di: Xiao, Yicheng, et al.
Pubblicazione: (2025)
Dual Attribute-Spatial Relation Alignment for 3D Visual Grounding
di: Xu, Yue, et al.
Pubblicazione: (2024)
di: Xu, Yue, et al.
Pubblicazione: (2024)
Taming Knowledge Conflicts in Language Models
di: Li, Gaotang, et al.
Pubblicazione: (2025)
di: Li, Gaotang, et al.
Pubblicazione: (2025)
Enhancing Generalization in Medical Visual Question Answering Tasks via Gradient-Guided Model Perturbation
di: Liu, Gang, et al.
Pubblicazione: (2024)
di: Liu, Gang, et al.
Pubblicazione: (2024)
RealX3D: A Physically-Degraded 3D Benchmark for Multi-view Visual Restoration and Reconstruction
di: Liu, Shuhong, et al.
Pubblicazione: (2025)
di: Liu, Shuhong, et al.
Pubblicazione: (2025)
MAGE: Multimodal Alignment and Generation Enhancement via Bridging Visual and Semantic Spaces
di: E, Shaojun, et al.
Pubblicazione: (2025)
di: E, Shaojun, et al.
Pubblicazione: (2025)
VCap: Hypergeometric Rewards for Weak-to-Strong Visual Captioning
di: Lu, Xingyu, et al.
Pubblicazione: (2026)
di: Lu, Xingyu, et al.
Pubblicazione: (2026)
Visual Set Program Synthesizer
di: Cheng, Zehua, et al.
Pubblicazione: (2026)
di: Cheng, Zehua, et al.
Pubblicazione: (2026)
Mitigating Modality Bias in Multi-modal Entity Alignment from a Causal Perspective
di: Su, Taoyu, et al.
Pubblicazione: (2025)
di: Su, Taoyu, et al.
Pubblicazione: (2025)
Cross Modal Fine-Grained Alignment via Granularity-Aware and Region-Uncertain Modeling
di: Liu, Jiale, et al.
Pubblicazione: (2025)
di: Liu, Jiale, et al.
Pubblicazione: (2025)
PROVE: A Perceptual RemOVal cohErence Benchmark for Visual Media
di: Li, Fuhao, et al.
Pubblicazione: (2026)
di: Li, Fuhao, et al.
Pubblicazione: (2026)
Hierarchical Aligned Multimodal Learning for NER on Tweet Posts
di: Liu, Peipei, et al.
Pubblicazione: (2023)
di: Liu, Peipei, et al.
Pubblicazione: (2023)
UniCVR: From Alignment to Reranking for Unified Zero-Shot Composed Visual Retrieval
di: Wen, Haokun, et al.
Pubblicazione: (2026)
di: Wen, Haokun, et al.
Pubblicazione: (2026)
Retrieving Any Relevant Moments: Benchmark and Models for Generalized Moment Retrieval
di: Ding, Yiming, et al.
Pubblicazione: (2026)
di: Ding, Yiming, et al.
Pubblicazione: (2026)
MIntRec2.0: A Large-scale Benchmark Dataset for Multimodal Intent Recognition and Out-of-scope Detection in Conversations
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
di: Zhang, Hanlei, et al.
Pubblicazione: (2024)
Zero-Shot Visual Grounding in 3D Gaussians via View Retrieval
di: Liao, Liwei, et al.
Pubblicazione: (2025)
di: Liao, Liwei, et al.
Pubblicazione: (2025)
PoEmotion: Can AI Utilize Chinese Calligraphy to Express Emotion from Poems?
di: Liu, Tiancheng, et al.
Pubblicazione: (2025)
di: Liu, Tiancheng, et al.
Pubblicazione: (2025)
JointAVBench: A Benchmark for Joint Audio-Visual Reasoning Evaluation
di: Chao, Jianghan, et al.
Pubblicazione: (2025)
di: Chao, Jianghan, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Do VLMs Have a Moral Backbone? A Study on the Fragile Morality of Vision-Language Models
di: Liu, Zhining, et al.
Pubblicazione: (2026) -
A Collaborative Extended Reality Prototype for 3D Surgical Planning and Visualization
di: Qiu, Shi, et al.
Pubblicazione: (2026) -
MMAPS: End-to-End Multi-Grained Multi-Modal Attribute-Aware Product Summarization
di: Chen, Tao, et al.
Pubblicazione: (2023) -
MAC-SLU: Multi-Intent Automotive Cabin Spoken Language Understanding Benchmark
di: Peng, Yuezhang, et al.
Pubblicazione: (2025) -
V-FAT: Benchmarking Visual Fidelity Against Text-bias
di: Wang, Ziteng, et al.
Pubblicazione: (2026)