Do Vision Models Encode Object-Level Semantic Relatedness? A Cognitive Psychology-Inspired Benchmark
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Hansang, Lee, Haeil, Kim, Junmo |
|---|---|
| Format: | Preprint |
| Published: |
2017
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
The Effects of Mixed Sample Data Augmentation are Class Dependent
by: Lee, Haeil, et al.
Published: (2023)
by: Lee, Haeil, et al.
Published: (2023)
Noisy Label Classification using Label Noise Selection with Test-Time Augmentation Cross-Entropy and NoiseMix Learning
by: Lee, Hansang, et al.
Published: (2022)
by: Lee, Hansang, et al.
Published: (2022)
Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis
by: Lee, Haeil, et al.
Published: (2024)
by: Lee, Haeil, et al.
Published: (2024)
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
by: Lee, Hansang, et al.
Published: (2022)
by: Lee, Hansang, et al.
Published: (2022)
GenMix: Combining Generative and Mixture Data Augmentation for Medical Image Classification
by: Lee, Hansang, et al.
Published: (2024)
by: Lee, Hansang, et al.
Published: (2024)
Inspecting Explainability of Transformer Models with Additional Statistical Information
by: Nguyen, Hoang C., et al.
Published: (2023)
by: Nguyen, Hoang C., et al.
Published: (2023)
ContextMix: A context-aware data augmentation method for industrial visual inspection systems
by: Kim, Hyungmin, et al.
Published: (2024)
by: Kim, Hyungmin, et al.
Published: (2024)
Inlier-Centric Post-Training Quantization for Object Detection Models
by: Kim, Minsu, et al.
Published: (2026)
by: Kim, Minsu, et al.
Published: (2026)
IWP: Token Pruning as Implicit Weight Pruning in Large Vision Language Models
by: Lee, Dong-Jae, et al.
Published: (2026)
by: Lee, Dong-Jae, et al.
Published: (2026)
DeFloMat: Detection with Flow Matching for Stable and Efficient Generative Object Localization
by: Lee, Hansang, et al.
Published: (2025)
by: Lee, Hansang, et al.
Published: (2025)
B4DL: A Benchmark for 4D LiDAR LLM in Spatio-Temporal Understanding
by: Choi, Changho, et al.
Published: (2025)
by: Choi, Changho, et al.
Published: (2025)
AI-KD: Adversarial learning and Implicit regularization for self-Knowledge Distillation
by: Kim, Hyungmin, et al.
Published: (2022)
by: Kim, Hyungmin, et al.
Published: (2022)
Pygmalion Effect in Vision: Image-to-Clay Translation for Reflective Geometry Reconstruction
by: Lee, Gayoung, et al.
Published: (2025)
by: Lee, Gayoung, et al.
Published: (2025)
SOCO: Benchmarking Semantic Object Correspondence in Vision Foundation Models
by: Dünkel, Olaf, et al.
Published: (2026)
by: Dünkel, Olaf, et al.
Published: (2026)
Beyond Semantics: Disentangling Information Scope in Sparse Autoencoders for CLIP
by: Ro, Yusung, et al.
Published: (2026)
by: Ro, Yusung, et al.
Published: (2026)
Do Pre-trained Vision-Language Models Encode Object States?
by: Newman, Kaleb, et al.
Published: (2024)
by: Newman, Kaleb, et al.
Published: (2024)
ConceptPrism: Concept Disentanglement in Personalized Diffusion Models via Residual Token Optimization
by: Kim, Minseo, et al.
Published: (2026)
by: Kim, Minseo, et al.
Published: (2026)
Modeling Stereo-Confidence Out of the End-to-End Stereo-Matching Network via Disparity Plane Sweep
by: Lee, Jae Young, et al.
Published: (2024)
by: Lee, Jae Young, et al.
Published: (2024)
MSG-Loc: Multi-Label Likelihood-based Semantic Graph Matching for Object-Level Global Localization
by: Lee, Gihyeon, et al.
Published: (2025)
by: Lee, Gihyeon, et al.
Published: (2025)
FRED: Towards a Full Rotation-Equivariance in Aerial Image Object Detection
by: Lee, Chanho, et al.
Published: (2023)
by: Lee, Chanho, et al.
Published: (2023)
MDS-DETR: DETR with Masked Duplicate Suppressor
by: Lee, Chanho, et al.
Published: (2026)
by: Lee, Chanho, et al.
Published: (2026)
Frequency-Aware Token Reduction for Efficient Vision Transformer
by: Lee, Dong-Jae, et al.
Published: (2025)
by: Lee, Dong-Jae, et al.
Published: (2025)
Video Inference for Human Mesh Recovery with Vision Transformer
by: Cho, Hanbyel, et al.
Published: (2025)
by: Cho, Hanbyel, et al.
Published: (2025)
Cognition-Inspired Dual-Stream Semantic Enhancement for Vision-Based Dynamic Emotion Modeling
by: Wang, Huanzhen, et al.
Published: (2026)
by: Wang, Huanzhen, et al.
Published: (2026)
Comparison Reveals Commonality: Customized Image Generation through Contrastive Inversion
by: Kim, Minseo, et al.
Published: (2025)
by: Kim, Minseo, et al.
Published: (2025)
Stereo-Matching Knowledge Distilled Monocular Depth Estimation Filtered by Multiple Disparity Consistency
by: Ka, Woonghyun, et al.
Published: (2024)
by: Ka, Woonghyun, et al.
Published: (2024)
A Decoding Scheme with Successive Aggregation of Multi-Level Features for Light-Weight Semantic Segmentation
by: Yoo, Jiwon, et al.
Published: (2024)
by: Yoo, Jiwon, et al.
Published: (2024)
Do Vision Models Develop Human-Like Progressive Difficulty Understanding?
by: Huang, Zeyi, et al.
Published: (2025)
by: Huang, Zeyi, et al.
Published: (2025)
AH-OCDA: Amplitude-based Curriculum Learning and Hopfield Segmentation Model for Open Compound Domain Adaptation
by: Choi, Jaehyun, et al.
Published: (2024)
by: Choi, Jaehyun, et al.
Published: (2024)
ImageNet-D: Benchmarking Neural Network Robustness on Diffusion Synthetic Object
by: Zhang, Chenshuang, et al.
Published: (2024)
by: Zhang, Chenshuang, et al.
Published: (2024)
Unlocking the Capabilities of Masked Generative Models for Image Synthesis via Self-Guidance
by: Hur, Jiwan, et al.
Published: (2024)
by: Hur, Jiwan, et al.
Published: (2024)
Superpixel Tokenization for Vision Transformers: Preserving Semantic Integrity in Visual Tokens
by: Lew, Jaihyun, et al.
Published: (2024)
by: Lew, Jaihyun, et al.
Published: (2024)
HoliSafe: Holistic Safety Benchmarking and Modeling for Vision-Language Model
by: Lee, Youngwan, et al.
Published: (2025)
by: Lee, Youngwan, et al.
Published: (2025)
SPACE: SPAtial-aware Consistency rEgularization for anomaly detection in Industrial applications
by: Kim, Daehwan, et al.
Published: (2024)
by: Kim, Daehwan, et al.
Published: (2024)
Object-Level Verbalized Confidence Calibration in Vision-Language Models via Semantic Perturbation
by: Zhao, Yunpu, et al.
Published: (2025)
by: Zhao, Yunpu, et al.
Published: (2025)
What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?
by: Ryu, Koki, et al.
Published: (2026)
by: Ryu, Koki, et al.
Published: (2026)
DMQ: Dissecting Outliers of Diffusion Models for Post-Training Quantization
by: Lee, Dongyeun, et al.
Published: (2025)
by: Lee, Dongyeun, et al.
Published: (2025)
Benchmarking Direct Preference Optimization for Medical Large Vision-Language Models
by: Kim, Dain, et al.
Published: (2026)
by: Kim, Dain, et al.
Published: (2026)
THE-Pose: Topological Prior with Hybrid Graph Fusion for Estimating Category-Level 6D Object Pose
by: Lee, Eunho, et al.
Published: (2025)
by: Lee, Eunho, et al.
Published: (2025)
Instruct-4DGS: Efficient Dynamic Scene Editing via 4D Gaussian-based Static-Dynamic Separation
by: Kwon, Joohyun, et al.
Published: (2025)
by: Kwon, Joohyun, et al.
Published: (2025)
Similar Items
-
The Effects of Mixed Sample Data Augmentation are Class Dependent
by: Lee, Haeil, et al.
Published: (2023) -
Noisy Label Classification using Label Noise Selection with Test-Time Augmentation Cross-Entropy and NoiseMix Learning
by: Lee, Hansang, et al.
Published: (2022) -
Beta Sampling is All You Need: Efficient Image Generation Strategy for Diffusion Models using Stepwise Spectral Analysis
by: Lee, Haeil, et al.
Published: (2024) -
Test-Time Mixup Augmentation for Data and Class-Specific Uncertainty Estimation in Deep Learning Image Classification
by: Lee, Hansang, et al.
Published: (2022) -
GenMix: Combining Generative and Mixture Data Augmentation for Medical Image Classification
by: Lee, Hansang, et al.
Published: (2024)