"I Know It When I See It": Mood Spaces for Connecting and Expressing Visual Concepts
Fuente:
arXiv
Saved in:
| Main Authors: | Yang, Huzheng, Xu, Katherine, Grossberg, Michael D., Bai, Yutong, Shi, Jianbo |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Vibe Spaces for Creatively Connecting and Expressing Visual Concepts
by: Yang, Huzheng, et al.
Published: (2025)
by: Yang, Huzheng, et al.
Published: (2025)
AlignedCut: Visual Concepts Discovery on Brain-Guided Universal Feature Space
by: Yang, Huzheng, et al.
Published: (2024)
by: Yang, Huzheng, et al.
Published: (2024)
Brain Decodes Deep Nets
by: Yang, Huzheng, et al.
Published: (2023)
by: Yang, Huzheng, et al.
Published: (2023)
Artifacts and Attention Sinks: Structured Approximations for Efficient Vision Transformers
by: Lu, Andrew, et al.
Published: (2025)
by: Lu, Andrew, et al.
Published: (2025)
Good Seed Makes a Good Crop: Discovering Secret Seeds in Text-to-Image Diffusion Models
by: Xu, Katherine, et al.
Published: (2024)
by: Xu, Katherine, et al.
Published: (2024)
Detecting Origin Attribution for Text-to-Image Diffusion Models
by: Xu, Katherine, et al.
Published: (2024)
by: Xu, Katherine, et al.
Published: (2024)
Seeing Isn't Knowing: Do VLMs Know When Not to Answer Spatial Questions (and Why)?
by: Zhang, Yue, et al.
Published: (2026)
by: Zhang, Yue, et al.
Published: (2026)
Explorations of the Softmax Space: Knowing When the Neural Network Doesn't Know
by: Sikar, Daniel, et al.
Published: (2025)
by: Sikar, Daniel, et al.
Published: (2025)
When Seeing Overrides Knowing: Disentangling Knowledge Conflicts in Vision-Language Models
by: Ortu, Francesco, et al.
Published: (2025)
by: Ortu, Francesco, et al.
Published: (2025)
Does Seeing More Mean Knowing More? Mono-Anchored Advantage Normalization for Multi-Source Visual Reasoning
by: Zeng, Fanhu, et al.
Published: (2026)
by: Zeng, Fanhu, et al.
Published: (2026)
Visually Dehallucinative Instruction Generation: Know What You Don't Know
by: Cha, Sungguk, et al.
Published: (2024)
by: Cha, Sungguk, et al.
Published: (2024)
Do You See What I Say? Generalizable Deepfake Detection based on Visual Speech Recognition
by: Bora, Maheswar, et al.
Published: (2025)
by: Bora, Maheswar, et al.
Published: (2025)
Visual Concept Connectome (VCC): Open World Concept Discovery and their Interlayer Connections in Deep Models
by: Kowal, Matthew, et al.
Published: (2024)
by: Kowal, Matthew, et al.
Published: (2024)
MambaOcc: Visual State Space Model for BEV-based Occupancy Prediction with Local Adaptive Reordering
by: Tian, Yonglin, et al.
Published: (2024)
by: Tian, Yonglin, et al.
Published: (2024)
Where Am I and What Will I See: An Auto-Regressive Model for Spatial Localization and View Prediction
by: Chen, Junyi, et al.
Published: (2024)
by: Chen, Junyi, et al.
Published: (2024)
ConceptExpress: Harnessing Diffusion Models for Single-image Unsupervised Concept Extraction
by: Hao, Shaozhe, et al.
Published: (2024)
by: Hao, Shaozhe, et al.
Published: (2024)
Finding Visual Task Vectors
by: Hojel, Alberto, et al.
Published: (2024)
by: Hojel, Alberto, et al.
Published: (2024)
Exploiting Text-Image Latent Spaces for the Description of Visual Concepts
by: Schmalwasser, Laines, et al.
Published: (2024)
by: Schmalwasser, Laines, et al.
Published: (2024)
CusConcept: Customized Visual Concept Decomposition with Diffusion Models
by: Xu, Zhi, et al.
Published: (2024)
by: Xu, Zhi, et al.
Published: (2024)
When Robots Should Say "I Don't Know": Benchmarking Abstention in Embodied Question Answering
by: Wu, Tao, et al.
Published: (2025)
by: Wu, Tao, et al.
Published: (2025)
Thinking in Space: How Multimodal Large Language Models See, Remember, and Recall Spaces
by: Yang, Jihan, et al.
Published: (2024)
by: Yang, Jihan, et al.
Published: (2024)
See Further When Clear: Curriculum Consistency Model
by: Liu, Yunpeng, et al.
Published: (2024)
by: Liu, Yunpeng, et al.
Published: (2024)
An Exploratory Study on Abstract Images and Visual Representations Learned from Them
by: Li, Haotian, et al.
Published: (2025)
by: Li, Haotian, et al.
Published: (2025)
Seeing Through Words: Controlling Visual Retrieval Quality with Language Models
by: Lu, Jianglin, et al.
Published: (2026)
by: Lu, Jianglin, et al.
Published: (2026)
Are VLMs Seeing or Just Saying? Uncovering the Illusion of Visual Re-examination
by: Shi, Chufan, et al.
Published: (2026)
by: Shi, Chufan, et al.
Published: (2026)
Visual Embodied Brain: Let Multimodal Large Language Models See, Think, and Control in Spaces
by: Luo, Gen, et al.
Published: (2025)
by: Luo, Gen, et al.
Published: (2025)
Multi-Scale VMamba: Hierarchy in Hierarchy Visual State Space Model
by: Shi, Yuheng, et al.
Published: (2024)
by: Shi, Yuheng, et al.
Published: (2024)
Seeing and Knowing in the Wild: Open-domain Visual Entity Recognition with Large-scale Knowledge Graphs via Contrastive Learning
by: Zhou, Hongkuan, et al.
Published: (2025)
by: Zhou, Hongkuan, et al.
Published: (2025)
See in Depth: Training-Free Surgical Scene Segmentation with Monocular Depth Priors
by: Yang, Kunyi, et al.
Published: (2025)
by: Yang, Kunyi, et al.
Published: (2025)
Seeing is Believing? Enhancing Vision-Language Navigation using Visual Perturbations
by: Zhang, Xuesong, et al.
Published: (2024)
by: Zhang, Xuesong, et al.
Published: (2024)
SeeGround: See and Ground for Zero-Shot Open-Vocabulary 3D Visual Grounding
by: Li, Rong, et al.
Published: (2024)
by: Li, Rong, et al.
Published: (2024)
Exploring Image Representation with Decoupled Classical Visual Descriptors
by: Qu, Chenyuan, et al.
Published: (2025)
by: Qu, Chenyuan, et al.
Published: (2025)
Enhancing Visual In-Context Learning by Multi-Faceted Fusion
by: Liao, Wenwen, et al.
Published: (2026)
by: Liao, Wenwen, et al.
Published: (2026)
Knowing When Not to Answer: Evaluating Abstention in Multimodal Reasoning Systems
by: Madhusudhan, Nishanth, et al.
Published: (2026)
by: Madhusudhan, Nishanth, et al.
Published: (2026)
Seeing What Matters: Visual Preference Policy Optimization for Visual Generation
by: Ni, Ziqi, et al.
Published: (2025)
by: Ni, Ziqi, et al.
Published: (2025)
Generation Models Know Space: Unleashing Implicit 3D Priors for Scene Understanding
by: Wu, Xianjin, et al.
Published: (2026)
by: Wu, Xianjin, et al.
Published: (2026)
Seeing Space and Motion: Enhancing Latent Actions with Geometric and Dynamic Awareness for Vision-Language-Action Models
by: Cai, Zhejia, et al.
Published: (2025)
by: Cai, Zhejia, et al.
Published: (2025)
Beyond Global Scanning: Adaptive Visual State Space Modeling for Salient Object Detection in Optical Remote Sensing Images
by: Ren, Mengyu, et al.
Published: (2025)
by: Ren, Mengyu, et al.
Published: (2025)
Seeing Candidates at Scale: Multimodal LLMs for Visual Political Communication on Instagram
by: Achmann-Denkler, Michael, et al.
Published: (2026)
by: Achmann-Denkler, Michael, et al.
Published: (2026)
Do You See What I Am Pointing At? Gesture-Based Egocentric Video Question Answering
by: Choi, Yura, et al.
Published: (2026)
by: Choi, Yura, et al.
Published: (2026)
Similar Items
-
Vibe Spaces for Creatively Connecting and Expressing Visual Concepts
by: Yang, Huzheng, et al.
Published: (2025) -
AlignedCut: Visual Concepts Discovery on Brain-Guided Universal Feature Space
by: Yang, Huzheng, et al.
Published: (2024) -
Brain Decodes Deep Nets
by: Yang, Huzheng, et al.
Published: (2023) -
Artifacts and Attention Sinks: Structured Approximations for Efficient Vision Transformers
by: Lu, Andrew, et al.
Published: (2025) -
Good Seed Makes a Good Crop: Discovering Secret Seeds in Text-to-Image Diffusion Models
by: Xu, Katherine, et al.
Published: (2024)