Computer Vision Models Show Human-Like Sensitivity to Geometric and Topological Concepts
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Zekun, Varma, Sashank |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Computer Vision Modeling of the Development of Geometric and Numerical Concepts in Humans
by: Wang, Zekun, et al.
Published: (2025)
by: Wang, Zekun, et al.
Published: (2025)
When Visuals Aren't the Problem: Evaluating Vision-Language Models on Misleading Data Visualizations
by: Lalai, Harsh Nishant, et al.
Published: (2026)
by: Lalai, Harsh Nishant, et al.
Published: (2026)
Understanding Graphical Perception in Data Visualization through Zero-shot Prompting of Vision-Language Models
by: Guo, Grace, et al.
Published: (2024)
by: Guo, Grace, et al.
Published: (2024)
NNDM: NN_UNet Diffusion Model for Brain Tumor Segmentation
by: Makanaboyina, Sashank
Published: (2025)
by: Makanaboyina, Sashank
Published: (2025)
TraceVision: Trajectory-Aware Vision-Language Model for Human-Like Spatial Understanding
by: Yang, Fan, et al.
Published: (2026)
by: Yang, Fan, et al.
Published: (2026)
Can Large Vision Language Models Read Maps Like a Human?
by: Xing, Shuo, et al.
Published: (2025)
by: Xing, Shuo, et al.
Published: (2025)
Contour Integration Underlies Human-Like Vision
by: Lonnqvist, Ben, et al.
Published: (2025)
by: Lonnqvist, Ben, et al.
Published: (2025)
Do Vision Models Develop Human-Like Progressive Difficulty Understanding?
by: Huang, Zeyi, et al.
Published: (2025)
by: Huang, Zeyi, et al.
Published: (2025)
ShowFlow: From Robust Single Concept to Condition-Free Multi-Concept Generation
by: Hoang, Trong-Vu, et al.
Published: (2025)
by: Hoang, Trong-Vu, et al.
Published: (2025)
Human-Like Coarse Object Representations in Vision Models
by: Gizdov, Andrey, et al.
Published: (2026)
by: Gizdov, Andrey, et al.
Published: (2026)
Learning Like Humans: Analogical Concept Learning for Generalized Category Discovery
by: Han, Jizhou, et al.
Published: (2026)
by: Han, Jizhou, et al.
Published: (2026)
Concept Replacer: Replacing Sensitive Concepts in Diffusion Models via Precision Localization
by: Zhang, Lingyun, et al.
Published: (2024)
by: Zhang, Lingyun, et al.
Published: (2024)
ShowUI-Aloha: Human-Taught GUI Agent
by: Zhang, Yichun, et al.
Published: (2026)
by: Zhang, Yichun, et al.
Published: (2026)
Show and Guide: Instructional-Plan Grounded Vision and Language Model
by: Glória-Silva, Diogo, et al.
Published: (2024)
by: Glória-Silva, Diogo, et al.
Published: (2024)
Improving Concept Alignment in Vision-Language Concept Bottleneck Models
by: Selvaraj, Nithish Muthuchamy, et al.
Published: (2024)
by: Selvaraj, Nithish Muthuchamy, et al.
Published: (2024)
Show and Tell: Visually Explainable Deep Neural Nets via Spatially-Aware Concept Bottleneck Models
by: Benou, Itay, et al.
Published: (2025)
by: Benou, Itay, et al.
Published: (2025)
Do Vision Transformers See Like Humans? Evaluating their Perceptual Alignment
by: Hernández-Cámara, Pablo, et al.
Published: (2025)
by: Hernández-Cámara, Pablo, et al.
Published: (2025)
HumanGenesis: Agent-Based Geometric and Generative Modeling for Synthetic Human Dynamics
by: Li, Weiqi, et al.
Published: (2025)
by: Li, Weiqi, et al.
Published: (2025)
LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation
by: Peng, Jiankun, et al.
Published: (2026)
by: Peng, Jiankun, et al.
Published: (2026)
Lift3D: Zero-Shot Lifting of Any 2D Vision Model to 3D
by: T, Mukund Varma, et al.
Published: (2024)
by: T, Mukund Varma, et al.
Published: (2024)
DrVD-Bench: Do Vision-Language Models Reason Like Human Doctors in Medical Image Diagnosis?
by: Zhou, Tianhong, et al.
Published: (2025)
by: Zhou, Tianhong, et al.
Published: (2025)
Benchmarking PhD-Level Coding in 3D Geometric Computer Vision
by: Li, Wenyi, et al.
Published: (2026)
by: Li, Wenyi, et al.
Published: (2026)
Preserving Topological and Geometric Embeddings for Point Cloud Recovery
by: Zhou, Kaiyue, et al.
Published: (2025)
by: Zhou, Kaiyue, et al.
Published: (2025)
CLIPSwarm: Generating Drone Shows from Text Prompts with Vision-Language Models
by: Pueyo, Pablo, et al.
Published: (2024)
by: Pueyo, Pablo, et al.
Published: (2024)
Concept-Based Explanations in Computer Vision: Where Are We and Where Could We Go?
by: Lee, Jae Hee, et al.
Published: (2024)
by: Lee, Jae Hee, et al.
Published: (2024)
Bridging Human Concepts and Computer Vision for Explainable Face Verification
by: Doh, Miriam, et al.
Published: (2024)
by: Doh, Miriam, et al.
Published: (2024)
Reading Images Like Texts: Sequential Image Understanding in Vision-Language Models
by: Li, Yueyan, et al.
Published: (2025)
by: Li, Yueyan, et al.
Published: (2025)
CVGL: Causal Learning and Geometric Topology
by: Ouyang, Songsong, et al.
Published: (2026)
by: Ouyang, Songsong, et al.
Published: (2026)
Computer Vision-Driven Gesture Recognition: Toward Natural and Intuitive Human-Computer
by: Shao, Fenghua, et al.
Published: (2024)
by: Shao, Fenghua, et al.
Published: (2024)
Test-Time Hinting for Black-Box Vision-Language Models
by: Hou, Kaihua, et al.
Published: (2026)
by: Hou, Kaihua, et al.
Published: (2026)
Show or Tell? Effectively prompting Vision-Language Models for semantic segmentation
by: Avogaro, Niccolo, et al.
Published: (2025)
by: Avogaro, Niccolo, et al.
Published: (2025)
Concept-Guided Prompt Learning for Generalization in Vision-Language Models
by: Zhang, Yi, et al.
Published: (2024)
by: Zhang, Yi, et al.
Published: (2024)
Which Concepts to Forget and How to Refuse? Decomposing Concepts for Continual Unlearning in Large Vision-Language Models
by: Jin, Hyundong, et al.
Published: (2026)
by: Jin, Hyundong, et al.
Published: (2026)
Like Humans to Few-Shot Learning through Knowledge Permeation of Vision and Text
by: Jia, Yuyu, et al.
Published: (2024)
by: Jia, Yuyu, et al.
Published: (2024)
TRoVe: Discovering Error-Inducing Static Feature Biases in Temporal Vision-Language Models
by: Varma, Maya, et al.
Published: (2025)
by: Varma, Maya, et al.
Published: (2025)
Topology-Aware Layer Pruning for Large Vision-Language Models
by: Zheng, Pengcheng, et al.
Published: (2026)
by: Zheng, Pengcheng, et al.
Published: (2026)
OmniShow: Unifying Multimodal Conditions for Human-Object Interaction Video Generation
by: Zhou, Donghao, et al.
Published: (2026)
by: Zhou, Donghao, et al.
Published: (2026)
SER-Diff: Synthetic Error Replay Diffusion for Incremental Brain Tumor Segmentation
by: Makanaboyina, Sashank
Published: (2025)
by: Makanaboyina, Sashank
Published: (2025)
Analyzing the Sensitivity of Vision Language Models in Visual Question Answering
by: Shah, Monika, et al.
Published: (2025)
by: Shah, Monika, et al.
Published: (2025)
Enhancing Zero-Shot Image Recognition in Vision-Language Models through Human-like Concept Guidance
by: Liu, Hui, et al.
Published: (2025)
by: Liu, Hui, et al.
Published: (2025)
Similar Items
-
Computer Vision Modeling of the Development of Geometric and Numerical Concepts in Humans
by: Wang, Zekun, et al.
Published: (2025) -
When Visuals Aren't the Problem: Evaluating Vision-Language Models on Misleading Data Visualizations
by: Lalai, Harsh Nishant, et al.
Published: (2026) -
Understanding Graphical Perception in Data Visualization through Zero-shot Prompting of Vision-Language Models
by: Guo, Grace, et al.
Published: (2024) -
NNDM: NN_UNet Diffusion Model for Brain Tumor Segmentation
by: Makanaboyina, Sashank
Published: (2025) -
TraceVision: Trajectory-Aware Vision-Language Model for Human-Like Spatial Understanding
by: Yang, Fan, et al.
Published: (2026)