Compositional Attribute Imbalance in Vision Datasets
Fuente:
arXiv
Salvato in:
| Autori principali: | Chen, Jiayi, Ma, Yanbiao, Zhang, Andi, Tang, Weidong, Dai, Wei, Liu, Bowei |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Geometric Origins of Bias in Deep Neural Networks: A Human Visual System Perspective
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
Pursuing Better Decision Boundaries for Long-Tailed Object Detection via Category Information Amount
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
Geometric Knowledge-Guided Localized Global Distribution Alignment for Federated Learning
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
di: Ma, Yanbiao, et al.
Pubblicazione: (2025)
MathDoc: Benchmarking Structured Extraction and Active Refusal on Noisy Mathematics Exam Papers
di: Zhou, Chenyue, et al.
Pubblicazione: (2026)
di: Zhou, Chenyue, et al.
Pubblicazione: (2026)
Pedestrian Attribute Recognition via CLIP based Prompt Vision-Language Fusion
di: Wang, Xiao, et al.
Pubblicazione: (2023)
di: Wang, Xiao, et al.
Pubblicazione: (2023)
Learning Quantised Structure-Preserving Motion Representations for Dance Fingerprinting
di: Kharlamova, Arina, et al.
Pubblicazione: (2026)
di: Kharlamova, Arina, et al.
Pubblicazione: (2026)
XAI-guided Insulator Anomaly Detection for Imbalanced Datasets
di: Hoefler, Maximilian Andreas, et al.
Pubblicazione: (2024)
di: Hoefler, Maximilian Andreas, et al.
Pubblicazione: (2024)
Improving Diagnostic Performance on Small and Imbalanced Datasets Using Class-Based Input Image Composition
di: Azzeddine, Hlali, et al.
Pubblicazione: (2025)
di: Azzeddine, Hlali, et al.
Pubblicazione: (2025)
Exploring Beyond Logits: Hierarchical Dynamic Labeling Based on Embeddings for Semi-Supervised Classification
di: Ma, Yanbiao, et al.
Pubblicazione: (2024)
di: Ma, Yanbiao, et al.
Pubblicazione: (2024)
Unveiling and Mitigating Generalized Biases of DNNs through the Intrinsic Dimensions of Perceptual Manifolds
di: Ma, Yanbiao, et al.
Pubblicazione: (2024)
di: Ma, Yanbiao, et al.
Pubblicazione: (2024)
Towards Automated Differential Diagnosis of Skin Diseases Using Deep Learning and Imbalance-Aware Strategies
di: Anaissi, Ali, et al.
Pubblicazione: (2026)
di: Anaissi, Ali, et al.
Pubblicazione: (2026)
Multimodal Causal Reasoning Benchmark: Challenging Vision Large Language Models to Discern Causal Links Across Modalities
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
di: Li, Zhiyuan, et al.
Pubblicazione: (2024)
Predicting and Enhancing the Fairness of DNNs with the Curvature of Perceptual Manifolds
di: Ma, Yanbiao, et al.
Pubblicazione: (2023)
di: Ma, Yanbiao, et al.
Pubblicazione: (2023)
Evading Visual Aphasia: Contrastive Adaptive Semantic Token Pruning for Vision-Language Models
di: Ma, Jie, et al.
Pubblicazione: (2026)
di: Ma, Jie, et al.
Pubblicazione: (2026)
Generating Attribution Reports for Manipulated Facial Images: A Dataset and Baseline
di: Lian, Jingchun, et al.
Pubblicazione: (2024)
di: Lian, Jingchun, et al.
Pubblicazione: (2024)
Benchmarking Large Vision-Language Models on CFMME: A Comprehensive Chinese Financial Multimodal Evaluation Dataset
di: Chen, Qian, et al.
Pubblicazione: (2026)
di: Chen, Qian, et al.
Pubblicazione: (2026)
MEET: A Million-Scale Dataset for Fine-Grained Geospatial Scene Classification with Zoom-Free Remote Sensing Imagery
di: Li, Yansheng, et al.
Pubblicazione: (2025)
di: Li, Yansheng, et al.
Pubblicazione: (2025)
Leveraging MLLM Embeddings and Attribute Smoothing for Compositional Zero-Shot Learning
di: Yan, Xudong, et al.
Pubblicazione: (2024)
di: Yan, Xudong, et al.
Pubblicazione: (2024)
15M Multimodal Facial Image-Text Dataset
di: Dai, Dawei, et al.
Pubblicazione: (2024)
di: Dai, Dawei, et al.
Pubblicazione: (2024)
Unsupervised Synthetic Image Attribution: Alignment and Disentanglement
di: Liu, Zongfang, et al.
Pubblicazione: (2026)
di: Liu, Zongfang, et al.
Pubblicazione: (2026)
Hybrid Discriminative Attribute-Object Embedding Network for Compositional Zero-Shot Learning
di: Liu, Yang, et al.
Pubblicazione: (2024)
di: Liu, Yang, et al.
Pubblicazione: (2024)
3DCoMPaT$^{++}$: An improved Large-scale 3D Vision Dataset for Compositional Recognition
di: Slim, Habib, et al.
Pubblicazione: (2023)
di: Slim, Habib, et al.
Pubblicazione: (2023)
DO-Bench: An Attributable Benchmark for Diagnosing Object Hallucination in Vision-Language Models
di: Wang, JiYang, et al.
Pubblicazione: (2026)
di: Wang, JiYang, et al.
Pubblicazione: (2026)
DanQing: An Up-to-Date Large-Scale Chinese Vision-Language Pre-training Dataset
di: Shen, Hengyu, et al.
Pubblicazione: (2026)
di: Shen, Hengyu, et al.
Pubblicazione: (2026)
World-to-Words: Grounded Open Vocabulary Acquisition through Fast Mapping in Vision-Language Models
di: Ma, Ziqiao, et al.
Pubblicazione: (2023)
di: Ma, Ziqiao, et al.
Pubblicazione: (2023)
UniVCD: A New Method for Unsupervised Change Detection in the Open-Vocabulary Era
di: Zhu, Ziqiang, et al.
Pubblicazione: (2025)
di: Zhu, Ziqiang, et al.
Pubblicazione: (2025)
Luminark: Training-free, Probabilistically-Certified Watermarking for General Vision Generative Models
di: Xu, Jiayi, et al.
Pubblicazione: (2026)
di: Xu, Jiayi, et al.
Pubblicazione: (2026)
Smile on the Face, Sadness in the Eyes: Bridging the Emotion Gap with a Multimodal Dataset of Eye and Facial Behaviors
di: Liu, Kejun, et al.
Pubblicazione: (2025)
di: Liu, Kejun, et al.
Pubblicazione: (2025)
AttrSeg: Open-Vocabulary Semantic Segmentation via Attribute Decomposition-Aggregation
di: Ma, Chaofan, et al.
Pubblicazione: (2023)
di: Ma, Chaofan, et al.
Pubblicazione: (2023)
Landsat30-AU: A Vision-Language Dataset for Australian Landsat Imagery
di: Ma, Sai, et al.
Pubblicazione: (2025)
di: Ma, Sai, et al.
Pubblicazione: (2025)
U-Face: An Efficient and Generalizable Framework for Unsupervised Facial Attribute Editing via Subspace Learning
di: Liu, Bo, et al.
Pubblicazione: (2026)
di: Liu, Bo, et al.
Pubblicazione: (2026)
An Empirical Study of Mamba-based Pedestrian Attribute Recognition
di: Wang, Xiao, et al.
Pubblicazione: (2024)
di: Wang, Xiao, et al.
Pubblicazione: (2024)
DGTRSD & DGTRS-CLIP: A Dual-Granularity Remote Sensing Image-Text Dataset and Vision Language Foundation Model for Alignment
di: Chen, Weizhi, et al.
Pubblicazione: (2025)
di: Chen, Weizhi, et al.
Pubblicazione: (2025)
Imbalance in Balance: Online Concept Balancing in Generation Models
di: Shi, Yukai, et al.
Pubblicazione: (2025)
di: Shi, Yukai, et al.
Pubblicazione: (2025)
VLM-PAR: A Vision Language Model for Pedestrian Attribute Recognition
di: Sellam, Abdellah Zakaria, et al.
Pubblicazione: (2025)
di: Sellam, Abdellah Zakaria, et al.
Pubblicazione: (2025)
RGB-Event based Pedestrian Attribute Recognition: A Benchmark Dataset and An Asymmetric RWKV Fusion Framework
di: Wang, Xiao, et al.
Pubblicazione: (2025)
di: Wang, Xiao, et al.
Pubblicazione: (2025)
Hunting Attributes: Context Prototype-Aware Learning for Weakly Supervised Semantic Segmentation
di: Tang, Feilong, et al.
Pubblicazione: (2024)
di: Tang, Feilong, et al.
Pubblicazione: (2024)
An Examination of the Compositionality of Large Generative Vision-Language Models
di: Ma, Teli, et al.
Pubblicazione: (2023)
di: Ma, Teli, et al.
Pubblicazione: (2023)
Addressing Domain Shift via Imbalance-Aware Domain Adaptation in Embryo Development Assessment
di: Li, Lei, et al.
Pubblicazione: (2025)
di: Li, Lei, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Geometric Origins of Bias in Deep Neural Networks: A Human Visual System Perspective
di: Ma, Yanbiao, et al.
Pubblicazione: (2025) -
Pursuing Better Decision Boundaries for Long-Tailed Object Detection via Category Information Amount
di: Ma, Yanbiao, et al.
Pubblicazione: (2025) -
Geometric Knowledge-Guided Localized Global Distribution Alignment for Federated Learning
di: Ma, Yanbiao, et al.
Pubblicazione: (2025) -
Calibrating Biased Distribution in VFM-derived Latent Space via Cross-Domain Geometric Consistency
di: Ma, Yanbiao, et al.
Pubblicazione: (2025) -
MathDoc: Benchmarking Structured Extraction and Active Refusal on Noisy Mathematics Exam Papers
di: Zhou, Chenyue, et al.
Pubblicazione: (2026)