MiSCHiEF: A Benchmark in Minimal-Pairs of Safety and Culture for Holistic Evaluation of Fine-Grained Image-Caption Alignment
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Banerjee, Sagarika, Madi, Tangatar, Swaminathan, Advait, Anh, Nguyen Dao Minh, Garg, Shivank, Zhu, Kevin, Sharma, Vasu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Captioning: Task-Specific Prompting for Improved VLM Performance in Mathematical Reasoning
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Unmasking the Veil: An Investigation into Concept Ablation for Privacy and Copyright Protection in Images
von: Garg, Shivank, et al.
Veröffentlicht: (2024)
von: Garg, Shivank, et al.
Veröffentlicht: (2024)
Attention Shift: Steering AI Away from Unsafe Content
von: Garg, Shivank, et al.
Veröffentlicht: (2024)
von: Garg, Shivank, et al.
Veröffentlicht: (2024)
EF1 for Mixed Manna with Unequal Entitlements
von: Garg, Jugal, et al.
Veröffentlicht: (2024)
von: Garg, Jugal, et al.
Veröffentlicht: (2024)
Text2Arch: A Dataset for Generating Scientific Architecture Diagrams from Natural Language Descriptions
von: Garg, Shivank, et al.
Veröffentlicht: (2026)
von: Garg, Shivank, et al.
Veröffentlicht: (2026)
Mirage: Unveiling Hidden Artifacts in Synthetic Images with Large Vision-Language Models
von: Sharma, Pranav, et al.
Veröffentlicht: (2025)
von: Sharma, Pranav, et al.
Veröffentlicht: (2025)
LoRA-Mini : Adaptation Matrices Decomposition and Selective Training
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Are VLMs Really Blind
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
MedBLIP: Fine-tuning BLIP for Medical Image Captioning
von: Limbu, Manshi, et al.
Veröffentlicht: (2025)
von: Limbu, Manshi, et al.
Veröffentlicht: (2025)
Mining Fine-Grained Image-Text Alignment for Zero-Shot Captioning via Text-Only Training
von: Qiu, Longtian, et al.
Veröffentlicht: (2024)
von: Qiu, Longtian, et al.
Veröffentlicht: (2024)
MoBind: Motion Binding for Fine-Grained IMU-Video Pose Alignment
von: Nguyen, Duc Duy, et al.
Veröffentlicht: (2026)
von: Nguyen, Duc Duy, et al.
Veröffentlicht: (2026)
SIDiffAgent: Self-Improving Diffusion Agent
von: Garg, Shivank, et al.
Veröffentlicht: (2026)
von: Garg, Shivank, et al.
Veröffentlicht: (2026)
Bridging the Multilingual Safety Divide: Efficient, Culturally-Aware Alignment for Global South Languages
von: Banerjee, Somnath, et al.
Veröffentlicht: (2026)
von: Banerjee, Somnath, et al.
Veröffentlicht: (2026)
Towards Fine-Grained Human Motion Video Captioning
von: Song, Guorui, et al.
Veröffentlicht: (2025)
von: Song, Guorui, et al.
Veröffentlicht: (2025)
Minimal mechanism for flocking in phoretically interacting active particles
von: Subramaniam, Arvin Gopal, et al.
Veröffentlicht: (2025)
von: Subramaniam, Arvin Gopal, et al.
Veröffentlicht: (2025)
CodeWiki: Evaluating AI's Ability to Generate Holistic Documentation for Large-Scale Codebases
von: Hoang, Anh Nguyen, et al.
Veröffentlicht: (2025)
von: Hoang, Anh Nguyen, et al.
Veröffentlicht: (2025)
Solving Non-Monotone Inclusions Using Monotonicity of Pairs of Operators
von: Le, Ba Khiet, et al.
Veröffentlicht: (2025)
von: Le, Ba Khiet, et al.
Veröffentlicht: (2025)
Training-Free Cultural Alignment of Large Language Models via Persona Disagreement
von: Kiet, Huynh Trung, et al.
Veröffentlicht: (2026)
von: Kiet, Huynh Trung, et al.
Veröffentlicht: (2026)
TrafficVLM: A Controllable Visual Language Model for Traffic Video Captioning
von: Dinh, Quang Minh, et al.
Veröffentlicht: (2024)
von: Dinh, Quang Minh, et al.
Veröffentlicht: (2024)
Weighted EF1 and PO Allocations with Few Types of Agents or Chores
von: Garg, Jugal, et al.
Veröffentlicht: (2024)
von: Garg, Jugal, et al.
Veröffentlicht: (2024)
FAID: Fine-Grained AI-Generated Text Detection Using Multi-Task Auxiliary and Multi-Level Contrastive Learning
von: Ta, Minh Ngoc, et al.
Veröffentlicht: (2025)
von: Ta, Minh Ngoc, et al.
Veröffentlicht: (2025)
Fine-Grained Alignment in Vision-and-Language Navigation through Bayesian Optimization
von: Song, Yuhang, et al.
Veröffentlicht: (2024)
von: Song, Yuhang, et al.
Veröffentlicht: (2024)
Snowy Scenes,Clear Detections: A Robust Model for Traffic Light Detection in Adverse Weather Conditions
von: Garg, Shivank, et al.
Veröffentlicht: (2024)
von: Garg, Shivank, et al.
Veröffentlicht: (2024)
ViT Registers and Fractal ViT
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2026)
von: Chou, Jason Chuan-Chih, et al.
Veröffentlicht: (2026)
IPO: Your Language Model is Secretly a Preference Classifier
von: Garg, Shivank, et al.
Veröffentlicht: (2025)
von: Garg, Shivank, et al.
Veröffentlicht: (2025)
When Prompt Optimization Becomes Jailbreaking: Adaptive Red-Teaming of Large Language Models
von: Shamsi, Zafir, et al.
Veröffentlicht: (2026)
von: Shamsi, Zafir, et al.
Veröffentlicht: (2026)
Adaptive Urban Planning: A Hybrid Framework for Balanced City Development
von: Singla, Pratham, et al.
Veröffentlicht: (2024)
von: Singla, Pratham, et al.
Veröffentlicht: (2024)
Figure 1 from: Pham A, Dao QT, Tran A, Le M, Nguyen T (2026) New records and an updated list of snakes from Ha Tinh Province, Vietnam. Biodiversity Data Journal 14: e177493. https://doi.org/10.3897/BDJ.14.e177493
von: Pham, Anh, et al.
Veröffentlicht: (2026)
von: Pham, Anh, et al.
Veröffentlicht: (2026)
Scene Graph-guided SegCaptioning Transformer with Fine-grained Alignment for Controllable Video Segmentation and Captioning
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
von: Zhang, Xu, et al.
Veröffentlicht: (2026)
Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models
von: Singla, Pratham, et al.
Veröffentlicht: (2025)
von: Singla, Pratham, et al.
Veröffentlicht: (2025)
NeuronTune: Fine-Grained Neuron Modulation for Balanced Safety-Utility Alignment in LLMs
von: Pan, Birong, et al.
Veröffentlicht: (2025)
von: Pan, Birong, et al.
Veröffentlicht: (2025)
Language Models are Homer Simpson! Safety Re-Alignment of Fine-tuned Language Models through Task Arithmetic
von: Bhardwaj, Rishabh, et al.
Veröffentlicht: (2024)
von: Bhardwaj, Rishabh, et al.
Veröffentlicht: (2024)
Fine-Grained Captioning of Long Videos through Scene Graph Consolidation
von: Chu, Sanghyeok, et al.
Veröffentlicht: (2025)
von: Chu, Sanghyeok, et al.
Veröffentlicht: (2025)
Incorporating Talker Identity Aids With Improving Speech Recognition in Adversarial Environments
von: Alavilli, Sagarika, et al.
Veröffentlicht: (2024)
von: Alavilli, Sagarika, et al.
Veröffentlicht: (2024)
Cosmological Dynamics of Accelerating Model in $f(T)$ Gravity with Special Forms of Deceleration Parameter
von: Pal, Shivank, et al.
Veröffentlicht: (2025)
von: Pal, Shivank, et al.
Veröffentlicht: (2025)
Do Biased Models Have Biased Thoughts?
von: Rajwal, Swati, et al.
Veröffentlicht: (2025)
von: Rajwal, Swati, et al.
Veröffentlicht: (2025)
CLIP with Quality Captions: A Strong Pretraining for Vision Tasks
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
von: Vasu, Pavan Kumar Anasosalu, et al.
Veröffentlicht: (2024)
Beyond Vision: Contextually Enriched Image Captioning with Multi-Modal Retrieval
von: Quy, Nguyen Lam Phu, et al.
Veröffentlicht: (2025)
von: Quy, Nguyen Lam Phu, et al.
Veröffentlicht: (2025)
HFrEF and HFpEF: are mitochondria at the heart of the matter?
von: Amanda Groenewald, et al.
Veröffentlicht: (2024)
von: Amanda Groenewald, et al.
Veröffentlicht: (2024)
BrainSCUBA: Fine-Grained Natural Language Captions of Visual Cortex Selectivity
von: Luo, Andrew F., et al.
Veröffentlicht: (2023)
von: Luo, Andrew F., et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
Beyond Captioning: Task-Specific Prompting for Improved VLM Performance in Mathematical Reasoning
von: Singh, Ayush, et al.
Veröffentlicht: (2024) -
Unmasking the Veil: An Investigation into Concept Ablation for Privacy and Copyright Protection in Images
von: Garg, Shivank, et al.
Veröffentlicht: (2024) -
Attention Shift: Steering AI Away from Unsafe Content
von: Garg, Shivank, et al.
Veröffentlicht: (2024) -
EF1 for Mixed Manna with Unequal Entitlements
von: Garg, Jugal, et al.
Veröffentlicht: (2024) -
Text2Arch: A Dataset for Generating Scientific Architecture Diagrams from Natural Language Descriptions
von: Garg, Shivank, et al.
Veröffentlicht: (2026)