Which private attributes do VLMs agree on and predict well?
Fuente:
arXiv
Saved in:
| Main Authors: | Hrynenko, Olena, Baranouskaya, Darya, Baia, Alina Elena, Cavallaro, Andrea |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
PrivLEX: Detecting legal concepts in images through Vision-Language Models
by: Baranouskaya, Darya, et al.
Published: (2026)
by: Baranouskaya, Darya, et al.
Published: (2026)
The impact of abstract and object tags on image privacy classification
by: Baranouskaya, Darya, et al.
Published: (2025)
by: Baranouskaya, Darya, et al.
Published: (2025)
Image-guided topic modeling for interpretable privacy classification
by: Baia, Alina Elena, et al.
Published: (2024)
by: Baia, Alina Elena, et al.
Published: (2024)
Cross-modal Counterfactual Explanations: Uncovering Decision Factors and Dataset Biases in Subjective Classification
by: Baia, Alina Elena, et al.
Published: (2025)
by: Baia, Alina Elena, et al.
Published: (2025)
Black-box Attacks on Image Activity Prediction and its Natural Language Explanations
by: Baia, Alina Elena, et al.
Published: (2023)
by: Baia, Alina Elena, et al.
Published: (2023)
Zero-shot image privacy classification with Vision-Language Models
by: Baia, Alina Elena, et al.
Published: (2025)
by: Baia, Alina Elena, et al.
Published: (2025)
Identifying Privacy Personas
by: Hrynenko, Olena, et al.
Published: (2024)
by: Hrynenko, Olena, et al.
Published: (2024)
FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection
by: Wei, Yao, et al.
Published: (2026)
by: Wei, Yao, et al.
Published: (2026)
Learning Privacy from Visual Entities
by: Xompero, Alessio, et al.
Published: (2025)
by: Xompero, Alessio, et al.
Published: (2025)
Sparse multi-view hand-object reconstruction for unseen environments
by: Pang, Yik Lung, et al.
Published: (2024)
by: Pang, Yik Lung, et al.
Published: (2024)
Visual Affordance Prediction: Survey and Reproducibility
by: Apicella, Tommaso, et al.
Published: (2025)
by: Apicella, Tommaso, et al.
Published: (2025)
Segmenting Object Affordances: Reproducibility and Sensitivity to Scale
by: Apicella, Tommaso, et al.
Published: (2024)
by: Apicella, Tommaso, et al.
Published: (2024)
Differentially private fine-tuned NF-Net to predict GI cancer type
by: Chilukoti, Sai Venkatesh, et al.
Published: (2025)
by: Chilukoti, Sai Venkatesh, et al.
Published: (2025)
Improving Generalization of Language-Conditioned Robot Manipulation
by: Cui, Chenglin, et al.
Published: (2025)
by: Cui, Chenglin, et al.
Published: (2025)
Open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2023)
by: Corsetti, Jaime, et al.
Published: (2023)
Deep Pre-Alignment for VLMs
by: Yu, Tianyu, et al.
Published: (2026)
by: Yu, Tianyu, et al.
Published: (2026)
Explaining models relating objects and privacy
by: Xompero, Alessio, et al.
Published: (2024)
by: Xompero, Alessio, et al.
Published: (2024)
Learning human-to-robot handovers through 3D scene reconstruction
by: Wu, Yuekun, et al.
Published: (2025)
by: Wu, Yuekun, et al.
Published: (2025)
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
by: Pang, Yik Lung, et al.
Published: (2024)
by: Pang, Yik Lung, et al.
Published: (2024)
Are VLMs Really Blind
by: Singh, Ayush, et al.
Published: (2024)
by: Singh, Ayush, et al.
Published: (2024)
Preserving Localized Patch Semantics in VLMs
by: Esmaeilkhani, Parsa, et al.
Published: (2026)
by: Esmaeilkhani, Parsa, et al.
Published: (2026)
Regression in EO: Are VLMs Up to the Challenge?
by: Xue, Xizhe, et al.
Published: (2025)
by: Xue, Xizhe, et al.
Published: (2025)
3D Face Reconstruction Error Decomposed: A Modular Benchmark for Fair and Fast Method Evaluation
by: Sariyanidi, Evangelos, et al.
Published: (2025)
by: Sariyanidi, Evangelos, et al.
Published: (2025)
High-resolution open-vocabulary object 6D pose estimation
by: Corsetti, Jaime, et al.
Published: (2024)
by: Corsetti, Jaime, et al.
Published: (2024)
Affordance segmentation of hand-occluded containers from exocentric images
by: Apicella, Tommaso, et al.
Published: (2023)
by: Apicella, Tommaso, et al.
Published: (2023)
Counting to Four is still a Chore for VLMs
by: Anh, Duy Le Dinh, et al.
Published: (2026)
by: Anh, Duy Le Dinh, et al.
Published: (2026)
Unlocking Dense Metric Depth Estimation in VLMs
by: Yu, Hanxun, et al.
Published: (2026)
by: Yu, Hanxun, et al.
Published: (2026)
CLGRPO: Reasoning Ability Enhancement for Small VLMs
by: Wang, Fanyi, et al.
Published: (2025)
by: Wang, Fanyi, et al.
Published: (2025)
Generate, Transduct, Adapt: Iterative Transduction with VLMs
by: Saha, Oindrila, et al.
Published: (2025)
by: Saha, Oindrila, et al.
Published: (2025)
Clapper: Compact Learning and Video Representation in VLMs
by: Kong, Lingyu, et al.
Published: (2025)
by: Kong, Lingyu, et al.
Published: (2025)
Image Recognition with Vision and Language Embeddings of VLMs
by: Volkov, Illia, et al.
Published: (2025)
by: Volkov, Illia, et al.
Published: (2025)
Should VLMs be Pre-trained with Image Data?
by: Keh, Sedrick, et al.
Published: (2025)
by: Keh, Sedrick, et al.
Published: (2025)
Gaze-Regularized VLMs for Ego-Centric Behavior Understanding
by: Pani, Anupam, et al.
Published: (2026)
by: Pani, Anupam, et al.
Published: (2026)
Linear Scaling Video VLMs for Long Video Understanding
by: Eyzaguirre, Cristobal, et al.
Published: (2026)
by: Eyzaguirre, Cristobal, et al.
Published: (2026)
MMTok: Multimodal Coverage Maximization for Efficient Inference of VLMs
by: Dong, Sixun, et al.
Published: (2025)
by: Dong, Sixun, et al.
Published: (2025)
How Auxiliary Reasoning Unleashes GUI Grounding in VLMs
by: Li, Weiming, et al.
Published: (2025)
by: Li, Weiming, et al.
Published: (2025)
VLMs Guided Interpretable Decision Making for Autonomous Driving
by: Hu, Xin, et al.
Published: (2025)
by: Hu, Xin, et al.
Published: (2025)
MyVLM: Personalizing VLMs for User-Specific Queries
by: Alaluf, Yuval, et al.
Published: (2024)
by: Alaluf, Yuval, et al.
Published: (2024)
Are VLMs Ready for Lane Topology Awareness in Autonomous Driving?
by: Chen, Xin, et al.
Published: (2025)
by: Chen, Xin, et al.
Published: (2025)
Beyond the Linear Separability Ceiling: Aligning Representations in VLMs
by: Vompa, Enrico, et al.
Published: (2025)
by: Vompa, Enrico, et al.
Published: (2025)
Similar Items
-
PrivLEX: Detecting legal concepts in images through Vision-Language Models
by: Baranouskaya, Darya, et al.
Published: (2026) -
The impact of abstract and object tags on image privacy classification
by: Baranouskaya, Darya, et al.
Published: (2025) -
Image-guided topic modeling for interpretable privacy classification
by: Baia, Alina Elena, et al.
Published: (2024) -
Cross-modal Counterfactual Explanations: Uncovering Decision Factors and Dataset Biases in Subjective Classification
by: Baia, Alina Elena, et al.
Published: (2025) -
Black-box Attacks on Image Activity Prediction and its Natural Language Explanations
by: Baia, Alina Elena, et al.
Published: (2023)