StableSemantics: A Synthetic Language-Vision Dataset of Semantic Representations in Naturalistic Images
Fuente:
arXiv
Saved in:
| Main Authors: | Zawar, Rushikesh, Dewan, Shaurya, Luo, Andrew F., Henderson, Margaret M., Tarr, Michael J., Wehbe, Leila |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Brain Mapping with Dense Features: Grounding Cortical Semantic Selectivity in Natural Images With Vision Transformers
by: Luo, Andrew F., et al.
Published: (2024)
by: Luo, Andrew F., et al.
Published: (2024)
DiffusionPID: Interpreting Diffusion via Partial Information Decomposition
by: Zawar, Rushikesh, et al.
Published: (2024)
by: Zawar, Rushikesh, et al.
Published: (2024)
Reanimating Images using Neural Representations of Dynamic Stimuli
by: Yeung, Jacob, et al.
Published: (2024)
by: Yeung, Jacob, et al.
Published: (2024)
BrainSCUBA: Fine-Grained Natural Language Captions of Visual Cortex Selectivity
by: Luo, Andrew F., et al.
Published: (2023)
by: Luo, Andrew F., et al.
Published: (2023)
UrbanTwin: Synthetic Roadside LiDAR Datasets
by: Shahbaz, Muhammad, et al.
Published: (2025)
by: Shahbaz, Muhammad, et al.
Published: (2025)
Fast-FoundationStereo: Real-Time Zero-Shot Stereo Matching
by: Wen, Bowen, et al.
Published: (2025)
by: Wen, Bowen, et al.
Published: (2025)
SynthmanticLiDAR: A Synthetic Dataset for Semantic Segmentation on LiDAR Imaging
by: Montalvo, Javier, et al.
Published: (2025)
by: Montalvo, Javier, et al.
Published: (2025)
NeRF-To-Real Tester: Neural Radiance Fields as Test Image Generators for Vision of Autonomous Systems
by: Weihl, Laura, et al.
Published: (2024)
by: Weihl, Laura, et al.
Published: (2024)
Discovering Meaningful Units with Visually Grounded Semantics from Image Captions
by: Behjati, Melika, et al.
Published: (2025)
by: Behjati, Melika, et al.
Published: (2025)
Language-Oriented Semantic Latent Representation for Image Transmission
by: Cicchetti, Giordano, et al.
Published: (2024)
by: Cicchetti, Giordano, et al.
Published: (2024)
MetaSegNet: Metadata-collaborative Vision-Language Representation Learning for Semantic Segmentation of Remote Sensing Images
by: Wang, Libo, et al.
Published: (2023)
by: Wang, Libo, et al.
Published: (2023)
Noisy Label Refinement with Semantically Reliable Synthetic Images
by: Li, Yingxuan, et al.
Published: (2025)
by: Li, Yingxuan, et al.
Published: (2025)
Geometric Analysis of Self-Supervised Vision Representations for Semantic Image Retrieval
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
by: Rodríguez-Betancourt, Esteban, et al.
Published: (2026)
Tuned Compositional Feature Replays for Efficient Stream Learning
by: Talbot, Morgan B., et al.
Published: (2021)
by: Talbot, Morgan B., et al.
Published: (2021)
Semantic Alignment of Unimodal Medical Text and Vision Representations
by: Di Folco, Maxime, et al.
Published: (2025)
by: Di Folco, Maxime, et al.
Published: (2025)
Vision Transformers with Natural Language Semantics
by: Kim, Young Kyung, et al.
Published: (2024)
by: Kim, Young Kyung, et al.
Published: (2024)
GraSP-VL: Length as a Semantic Granularity Interface for Vision-Language Representations
by: Li, Zesheng, et al.
Published: (2026)
by: Li, Zesheng, et al.
Published: (2026)
Superpixel Semantics Representation and Pre-training for Vision-Language Task
by: Zhang, Siyu, et al.
Published: (2023)
by: Zhang, Siyu, et al.
Published: (2023)
Representation Separation for Semantic Segmentation with Vision Transformers
by: Hong, Yuanduo, et al.
Published: (2022)
by: Hong, Yuanduo, et al.
Published: (2022)
Synthetic Instance Segmentation from Semantic Image Segmentation Masks
by: Shen, Yuchen, et al.
Published: (2023)
by: Shen, Yuchen, et al.
Published: (2023)
Are Vision Transformer Representations Semantically Meaningful? A Case Study in Medical Imaging
by: Shams, Montasir, et al.
Published: (2025)
by: Shams, Montasir, et al.
Published: (2025)
SemanticMIM: Marring Masked Image Modeling with Semantics Compression for General Visual Representation
by: Yuan, Yike, et al.
Published: (2024)
by: Yuan, Yike, et al.
Published: (2024)
Resolving Inconsistent Semantics in Multi-Dataset Image Segmentation
by: Zhangli, Qilong, et al.
Published: (2024)
by: Zhangli, Qilong, et al.
Published: (2024)
Diffusion Attack: Leveraging Stable Diffusion for Naturalistic Image Attacking
by: Guo, Qianyu, et al.
Published: (2024)
by: Guo, Qianyu, et al.
Published: (2024)
Toward Semantic-Agnostic and Shape-Aware Vision-Language Segmentation Models
by: Seutin, Corentin, et al.
Published: (2026)
by: Seutin, Corentin, et al.
Published: (2026)
Vision-Language Semantic Aggregation Leveraging Foundation Model for Generalizable Medical Image Segmentation
by: Yu, Wenjun, et al.
Published: (2025)
by: Yu, Wenjun, et al.
Published: (2025)
Semantics-enhanced Cross-modal Masked Image Modeling for Vision-Language Pre-training
by: Liu, Haowei, et al.
Published: (2024)
by: Liu, Haowei, et al.
Published: (2024)
Agent Journey Beyond RGB: Hierarchical Semantic-Spatial Representation Enrichment for Vision-and-Language Navigation
by: Zhang, Xuesong, et al.
Published: (2024)
by: Zhang, Xuesong, et al.
Published: (2024)
Vision-Language Models as Differentiable Semantic and Spatial Rewards for Text-to-3D Generation
by: Bai, Weimin, et al.
Published: (2025)
by: Bai, Weimin, et al.
Published: (2025)
Semantic-Aware Ship Detection with Vision-Language Integration
by: Li, Jiahao, et al.
Published: (2025)
by: Li, Jiahao, et al.
Published: (2025)
Unleashing Vision-Language Semantics for Deepfake Video Detection
by: Zhu, Jiawen, et al.
Published: (2026)
by: Zhu, Jiawen, et al.
Published: (2026)
Semantically Grounded QFormer for Efficient Vision Language Understanding
by: Choraria, Moulik, et al.
Published: (2023)
by: Choraria, Moulik, et al.
Published: (2023)
SGHA-Attack: Semantic-Guided Hierarchical Alignment for Transferable Targeted Attacks on Vision-Language Models
by: Wang, Haobo, et al.
Published: (2026)
by: Wang, Haobo, et al.
Published: (2026)
MaterialFusion: Enhancing Inverse Rendering with Material Diffusion Priors
by: Litman, Yehonathan, et al.
Published: (2024)
by: Litman, Yehonathan, et al.
Published: (2024)
Suppressing Non-Semantic Noise in Masked Image Modeling Representations
by: Hjelkrem-Tan, Martine, et al.
Published: (2026)
by: Hjelkrem-Tan, Martine, et al.
Published: (2026)
Improving Human Image Animation via Semantic Representation Alignment
by: Liu, Chang, et al.
Published: (2026)
by: Liu, Chang, et al.
Published: (2026)
Semantic Segmentation for Real-World and Synthetic Vehicle's Forward-Facing Camera Images
by: Nguyen, Tuan T., et al.
Published: (2024)
by: Nguyen, Tuan T., et al.
Published: (2024)
CO-SPY: Combining Semantic and Pixel Features to Detect Synthetic Images by AI
by: Cheng, Siyuan, et al.
Published: (2025)
by: Cheng, Siyuan, et al.
Published: (2025)
Leveraging Stable Diffusion for Monocular Depth Estimation via Image Semantic Encoding
by: Xia, Jingming, et al.
Published: (2025)
by: Xia, Jingming, et al.
Published: (2025)
Hierarchical Neural Semantic Representation for 3D Semantic Correspondence
by: Du, Keyu, et al.
Published: (2025)
by: Du, Keyu, et al.
Published: (2025)
Similar Items
-
Brain Mapping with Dense Features: Grounding Cortical Semantic Selectivity in Natural Images With Vision Transformers
by: Luo, Andrew F., et al.
Published: (2024) -
DiffusionPID: Interpreting Diffusion via Partial Information Decomposition
by: Zawar, Rushikesh, et al.
Published: (2024) -
Reanimating Images using Neural Representations of Dynamic Stimuli
by: Yeung, Jacob, et al.
Published: (2024) -
BrainSCUBA: Fine-Grained Natural Language Captions of Visual Cortex Selectivity
by: Luo, Andrew F., et al.
Published: (2023) -
UrbanTwin: Synthetic Roadside LiDAR Datasets
by: Shahbaz, Muhammad, et al.
Published: (2025)