DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data
Fuente:
arXiv
Saved in:
| Main Authors: | Fu, Stephanie, Tamir, Netanel, Sundaram, Shobhita, Chai, Lucy, Zhang, Richard, Dekel, Tali, Isola, Phillip |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
When Does Perceptual Alignment Benefit Vision Representations?
by: Sundaram, Shobhita, et al.
Published: (2024)
by: Sundaram, Shobhita, et al.
Published: (2024)
What Makes for a Good Stereoscopic Image?
by: Tamir, Netanel Y., et al.
Published: (2024)
by: Tamir, Netanel Y., et al.
Published: (2024)
Better Together: Leveraging Unpaired Multimodal Data for Stronger Unimodal Models
by: Gupta, Sharut, et al.
Published: (2025)
by: Gupta, Sharut, et al.
Published: (2025)
Personalized Representation from Personalized Generation
by: Sundaram, Shobhita, et al.
Published: (2024)
by: Sundaram, Shobhita, et al.
Published: (2024)
Scalable Real2Sim: Physics-Aware Asset Generation Via Robotic Pick-and-Place Setups
by: Pfaff, Nicholas, et al.
Published: (2025)
by: Pfaff, Nicholas, et al.
Published: (2025)
Match-and-Fuse: Consistent Generation from Unstructured Image Sets
by: Feingold, Kate, et al.
Published: (2025)
by: Feingold, Kate, et al.
Published: (2025)
What's in the Image? A Deep-Dive into the Vision of Vision Language Models
by: Kaduri, Omri, et al.
Published: (2024)
by: Kaduri, Omri, et al.
Published: (2024)
Cycle Consistency as Reward: Learning Image-Text Alignment without Human Preferences
by: Bahng, Hyojin, et al.
Published: (2025)
by: Bahng, Hyojin, et al.
Published: (2025)
Universality of Packing Dimension Estimates for Spectral Measures of Quasiperiodic Operators: Monotone Potentials
by: Levi, Netanel
Published: (2025)
by: Levi, Netanel
Published: (2025)
Neural Thickets: Diverse Task Experts Are Dense Around Pretrained Weights
by: Gan, Yulu, et al.
Published: (2026)
by: Gan, Yulu, et al.
Published: (2026)
Improved Representation of Asymmetrical Distances with Interval Quasimetric Embeddings
by: Wang, Tongzhou, et al.
Published: (2022)
by: Wang, Tongzhou, et al.
Published: (2022)
Learning Visual Parkour from Generated Images
by: Yu, Alan, et al.
Published: (2024)
by: Yu, Alan, et al.
Published: (2024)
Teaching Models to Teach Themselves: Reasoning at the Edge of Learnability
by: Sundaram, Shobhita, et al.
Published: (2026)
by: Sundaram, Shobhita, et al.
Published: (2026)
DINO-Tracker: Taming DINO for Self-Supervised Point Tracking in a Single Video
by: Tumanyan, Narek, et al.
Published: (2024)
by: Tumanyan, Narek, et al.
Published: (2024)
Learning Dynamical Systems by Leveraging Data from Similar Systems
by: Xin, Lei, et al.
Published: (2023)
by: Xin, Lei, et al.
Published: (2023)
Learning Vision from Models Rivals Learning Vision from Data
by: Tian, Yonglong, et al.
Published: (2023)
by: Tian, Yonglong, et al.
Published: (2023)
SimClone: Detecting Tabular Data Clones using Value Similarity
by: Yang, Xu, et al.
Published: (2024)
by: Yang, Xu, et al.
Published: (2024)
The Cognitive Divergence: AI Context Windows, Human Attention Decline, and the Delegation Feedback Loop
by: Eliav, Netanel
Published: (2026)
by: Eliav, Netanel
Published: (2026)
Separating Knowledge and Perception with Procedural Data
by: Rodríguez-Muñoz, Adrián, et al.
Published: (2025)
by: Rodríguez-Muñoz, Adrián, et al.
Published: (2025)
DynVFX: Augmenting Real Videos with Dynamic Content
by: Yatim, Danah, et al.
Published: (2025)
by: Yatim, Danah, et al.
Published: (2025)
DiffSim: Taming Diffusion Models for Evaluating Visual Similarity
by: Song, Yiren, et al.
Published: (2024)
by: Song, Yiren, et al.
Published: (2024)
ID-Sim: An Identity-Focused Similarity Metric
by: Chae, Julia, et al.
Published: (2026)
by: Chae, Julia, et al.
Published: (2026)
The Truth Lies Somewhere in the Middle (of the Generated Tokens)
by: Wang, Sophie L., et al.
Published: (2026)
by: Wang, Sophie L., et al.
Published: (2026)
Words That Make Language Models Perceive
by: Wang, Sophie L., et al.
Published: (2025)
by: Wang, Sophie L., et al.
Published: (2025)
SimPHonY source code and analysis
by: Papastefanou, Phillip
Published: (2026)
by: Papastefanou, Phillip
Published: (2026)
Canonicalizing Multimodal Contrastive Representation Learning
by: Gupta, Sharut, et al.
Published: (2026)
by: Gupta, Sharut, et al.
Published: (2026)
Eye2Eye: A Simple Approach for Monocular-to-Stereo Video Synthesis
by: Geyer, Michal, et al.
Published: (2025)
by: Geyer, Michal, et al.
Published: (2025)
Dreaming of disability‐as‐possibility as a humanistic STEM education futurity
by: Phillip A. Boda
Published: (2024)
by: Phillip A. Boda
Published: (2024)
Self-Similar Topological Fractals
by: Cipriani, Fabio E. G., et al.
Published: (2024)
by: Cipriani, Fabio E. G., et al.
Published: (2024)
SimCert: Probabilistic Certification for Behavioral Similarity in Deep Neural Network Compression
by: Li, Jingyang, et al.
Published: (2026)
by: Li, Jingyang, et al.
Published: (2026)
SimGrade: Using Code Similarity Measures for More Accurate Human Grading
by: Johnson-Yu, Sonja, et al.
Published: (2024)
by: Johnson-Yu, Sonja, et al.
Published: (2024)
Ascribe New Dimensions to Scientific Data Visualization with VR
by: Ushizima, Daniela, et al.
Published: (2025)
by: Ushizima, Daniela, et al.
Published: (2025)
The Platonic Representation Hypothesis
by: Huh, Minyoung, et al.
Published: (2024)
by: Huh, Minyoung, et al.
Published: (2024)
Adaptive Length Image Tokenization via Recurrent Allocation
by: Duggal, Shivam, et al.
Published: (2024)
by: Duggal, Shivam, et al.
Published: (2024)
MRFP: Learning Generalizable Semantic Segmentation from Sim-2-Real with Multi-Resolution Feature Perturbation
by: Udupa, Sumanth, et al.
Published: (2023)
by: Udupa, Sumanth, et al.
Published: (2023)
Contrastive Similarity Learning for Market Forecasting: The ContraSim Framework
by: Vinden, Nicholas, et al.
Published: (2025)
by: Vinden, Nicholas, et al.
Published: (2025)
VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images
by: Zhou, Guanyu, et al.
Published: (2026)
by: Zhou, Guanyu, et al.
Published: (2026)
Automatic Extraction of Disease Risk Factors from Medical Publications
by: Rubchinsky, Maxim, et al.
Published: (2024)
by: Rubchinsky, Maxim, et al.
Published: (2024)
DreamMask: Boosting Open-vocabulary Panoptic Segmentation with Synthetic Data
by: Tu, Yuanpeng, et al.
Published: (2025)
by: Tu, Yuanpeng, et al.
Published: (2025)
Versatile Editing of Video Content, Actions, and Dynamics without Training
by: Kulikov, Vladimir, et al.
Published: (2026)
by: Kulikov, Vladimir, et al.
Published: (2026)
Similar Items
-
When Does Perceptual Alignment Benefit Vision Representations?
by: Sundaram, Shobhita, et al.
Published: (2024) -
What Makes for a Good Stereoscopic Image?
by: Tamir, Netanel Y., et al.
Published: (2024) -
Better Together: Leveraging Unpaired Multimodal Data for Stronger Unimodal Models
by: Gupta, Sharut, et al.
Published: (2025) -
Personalized Representation from Personalized Generation
by: Sundaram, Shobhita, et al.
Published: (2024) -
Scalable Real2Sim: Physics-Aware Asset Generation Via Robotic Pick-and-Place Setups
by: Pfaff, Nicholas, et al.
Published: (2025)