Which private attributes do VLMs agree on and predict well?
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Hrynenko, Olena, Baranouskaya, Darya, Baia, Alina Elena, Cavallaro, Andrea |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
PrivLEX: Detecting legal concepts in images through Vision-Language Models
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2026)
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2026)
The impact of abstract and object tags on image privacy classification
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2025)
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2025)
Image-guided topic modeling for interpretable privacy classification
von: Baia, Alina Elena, et al.
Veröffentlicht: (2024)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2024)
Cross-modal Counterfactual Explanations: Uncovering Decision Factors and Dataset Biases in Subjective Classification
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
Black-box Attacks on Image Activity Prediction and its Natural Language Explanations
von: Baia, Alina Elena, et al.
Veröffentlicht: (2023)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2023)
Zero-shot image privacy classification with Vision-Language Models
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025)
Identifying Privacy Personas
von: Hrynenko, Olena, et al.
Veröffentlicht: (2024)
von: Hrynenko, Olena, et al.
Veröffentlicht: (2024)
FlowOVD: Learning Generative Latent Flows for Zero-shot Open-vocabulary Detection
von: Wei, Yao, et al.
Veröffentlicht: (2026)
von: Wei, Yao, et al.
Veröffentlicht: (2026)
Learning Privacy from Visual Entities
von: Xompero, Alessio, et al.
Veröffentlicht: (2025)
von: Xompero, Alessio, et al.
Veröffentlicht: (2025)
Sparse multi-view hand-object reconstruction for unseen environments
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
Visual Affordance Prediction: Survey and Reproducibility
von: Apicella, Tommaso, et al.
Veröffentlicht: (2025)
von: Apicella, Tommaso, et al.
Veröffentlicht: (2025)
Segmenting Object Affordances: Reproducibility and Sensitivity to Scale
von: Apicella, Tommaso, et al.
Veröffentlicht: (2024)
von: Apicella, Tommaso, et al.
Veröffentlicht: (2024)
Differentially private fine-tuned NF-Net to predict GI cancer type
von: Chilukoti, Sai Venkatesh, et al.
Veröffentlicht: (2025)
von: Chilukoti, Sai Venkatesh, et al.
Veröffentlicht: (2025)
Improving Generalization of Language-Conditioned Robot Manipulation
von: Cui, Chenglin, et al.
Veröffentlicht: (2025)
von: Cui, Chenglin, et al.
Veröffentlicht: (2025)
Open-vocabulary object 6D pose estimation
von: Corsetti, Jaime, et al.
Veröffentlicht: (2023)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2023)
Deep Pre-Alignment for VLMs
von: Yu, Tianyu, et al.
Veröffentlicht: (2026)
von: Yu, Tianyu, et al.
Veröffentlicht: (2026)
Explaining models relating objects and privacy
von: Xompero, Alessio, et al.
Veröffentlicht: (2024)
von: Xompero, Alessio, et al.
Veröffentlicht: (2024)
Learning human-to-robot handovers through 3D scene reconstruction
von: Wu, Yuekun, et al.
Veröffentlicht: (2025)
von: Wu, Yuekun, et al.
Veröffentlicht: (2025)
Stereo Hand-Object Reconstruction for Human-to-Robot Handover
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
von: Pang, Yik Lung, et al.
Veröffentlicht: (2024)
Are VLMs Really Blind
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
von: Singh, Ayush, et al.
Veröffentlicht: (2024)
Preserving Localized Patch Semantics in VLMs
von: Esmaeilkhani, Parsa, et al.
Veröffentlicht: (2026)
von: Esmaeilkhani, Parsa, et al.
Veröffentlicht: (2026)
Regression in EO: Are VLMs Up to the Challenge?
von: Xue, Xizhe, et al.
Veröffentlicht: (2025)
von: Xue, Xizhe, et al.
Veröffentlicht: (2025)
3D Face Reconstruction Error Decomposed: A Modular Benchmark for Fair and Fast Method Evaluation
von: Sariyanidi, Evangelos, et al.
Veröffentlicht: (2025)
von: Sariyanidi, Evangelos, et al.
Veröffentlicht: (2025)
High-resolution open-vocabulary object 6D pose estimation
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
von: Corsetti, Jaime, et al.
Veröffentlicht: (2024)
Affordance segmentation of hand-occluded containers from exocentric images
von: Apicella, Tommaso, et al.
Veröffentlicht: (2023)
von: Apicella, Tommaso, et al.
Veröffentlicht: (2023)
Counting to Four is still a Chore for VLMs
von: Anh, Duy Le Dinh, et al.
Veröffentlicht: (2026)
von: Anh, Duy Le Dinh, et al.
Veröffentlicht: (2026)
Unlocking Dense Metric Depth Estimation in VLMs
von: Yu, Hanxun, et al.
Veröffentlicht: (2026)
von: Yu, Hanxun, et al.
Veröffentlicht: (2026)
CLGRPO: Reasoning Ability Enhancement for Small VLMs
von: Wang, Fanyi, et al.
Veröffentlicht: (2025)
von: Wang, Fanyi, et al.
Veröffentlicht: (2025)
Generate, Transduct, Adapt: Iterative Transduction with VLMs
von: Saha, Oindrila, et al.
Veröffentlicht: (2025)
von: Saha, Oindrila, et al.
Veröffentlicht: (2025)
Clapper: Compact Learning and Video Representation in VLMs
von: Kong, Lingyu, et al.
Veröffentlicht: (2025)
von: Kong, Lingyu, et al.
Veröffentlicht: (2025)
Image Recognition with Vision and Language Embeddings of VLMs
von: Volkov, Illia, et al.
Veröffentlicht: (2025)
von: Volkov, Illia, et al.
Veröffentlicht: (2025)
Should VLMs be Pre-trained with Image Data?
von: Keh, Sedrick, et al.
Veröffentlicht: (2025)
von: Keh, Sedrick, et al.
Veröffentlicht: (2025)
Gaze-Regularized VLMs for Ego-Centric Behavior Understanding
von: Pani, Anupam, et al.
Veröffentlicht: (2026)
von: Pani, Anupam, et al.
Veröffentlicht: (2026)
Linear Scaling Video VLMs for Long Video Understanding
von: Eyzaguirre, Cristobal, et al.
Veröffentlicht: (2026)
von: Eyzaguirre, Cristobal, et al.
Veröffentlicht: (2026)
MMTok: Multimodal Coverage Maximization for Efficient Inference of VLMs
von: Dong, Sixun, et al.
Veröffentlicht: (2025)
von: Dong, Sixun, et al.
Veröffentlicht: (2025)
How Auxiliary Reasoning Unleashes GUI Grounding in VLMs
von: Li, Weiming, et al.
Veröffentlicht: (2025)
von: Li, Weiming, et al.
Veröffentlicht: (2025)
VLMs Guided Interpretable Decision Making for Autonomous Driving
von: Hu, Xin, et al.
Veröffentlicht: (2025)
von: Hu, Xin, et al.
Veröffentlicht: (2025)
MyVLM: Personalizing VLMs for User-Specific Queries
von: Alaluf, Yuval, et al.
Veröffentlicht: (2024)
von: Alaluf, Yuval, et al.
Veröffentlicht: (2024)
Are VLMs Ready for Lane Topology Awareness in Autonomous Driving?
von: Chen, Xin, et al.
Veröffentlicht: (2025)
von: Chen, Xin, et al.
Veröffentlicht: (2025)
Beyond the Linear Separability Ceiling: Aligning Representations in VLMs
von: Vompa, Enrico, et al.
Veröffentlicht: (2025)
von: Vompa, Enrico, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
PrivLEX: Detecting legal concepts in images through Vision-Language Models
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2026) -
The impact of abstract and object tags on image privacy classification
von: Baranouskaya, Darya, et al.
Veröffentlicht: (2025) -
Image-guided topic modeling for interpretable privacy classification
von: Baia, Alina Elena, et al.
Veröffentlicht: (2024) -
Cross-modal Counterfactual Explanations: Uncovering Decision Factors and Dataset Biases in Subjective Classification
von: Baia, Alina Elena, et al.
Veröffentlicht: (2025) -
Black-box Attacks on Image Activity Prediction and its Natural Language Explanations
von: Baia, Alina Elena, et al.
Veröffentlicht: (2023)