LipSim: A Provably Robust Perceptual Similarity Metric
Fuente:
arXiv
Saved in:
| Main Authors: | Ghazanfari, Sara, Araujo, Alexandre, Krishnamurthy, Prashanth, Khorrami, Farshad, Garg, Siddharth |
|---|---|
| Format: | Preprint |
| Published: |
2023
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Towards Unified Benchmark and Models for Multi-Modal Perceptual Metrics
by: Ghazanfari, Sara, et al.
Published: (2024)
by: Ghazanfari, Sara, et al.
Published: (2024)
EMMA: Efficient Visual Alignment in Multi-Modal LLMs
by: Ghazanfari, Sara, et al.
Published: (2024)
by: Ghazanfari, Sara, et al.
Published: (2024)
SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding
by: Ghazanfari, Sara, et al.
Published: (2026)
by: Ghazanfari, Sara, et al.
Published: (2026)
Chain-of-Frames: Advancing Video Understanding in Multimodal LLMs via Frame-Aware Reasoning
by: Ghazanfari, Sara, et al.
Published: (2025)
by: Ghazanfari, Sara, et al.
Published: (2025)
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation
by: Patel, Naman, et al.
Published: (2025)
by: Patel, Naman, et al.
Published: (2025)
CLIPScope: Enhancing Zero-Shot OOD Detection with Bayesian Scoring
by: Fu, Hao, et al.
Published: (2024)
by: Fu, Hao, et al.
Published: (2024)
FlashMix: Fast Map-Free LiDAR Localization via Feature Mixing and Contrastive-Constrained Accelerated Training
by: Goswami, Raktim Gautam, et al.
Published: (2024)
by: Goswami, Raktim Gautam, et al.
Published: (2024)
Efficient and Distributed Large-Scale 3D Map Registration using Tomographic Features
by: Unlu, Halil Utku, et al.
Published: (2024)
by: Unlu, Halil Utku, et al.
Published: (2024)
SALSA: Swift Adaptive Lightweight Self-Attention for Enhanced LiDAR Place Recognition
by: Goswami, Raktim Gautam, et al.
Published: (2024)
by: Goswami, Raktim Gautam, et al.
Published: (2024)
Out-of-Distribution Detection with Overlap Index
by: Fu, Hao, et al.
Published: (2024)
by: Fu, Hao, et al.
Published: (2024)
An Upper Bound for the Distribution Overlap Index and Its Applications
by: Fu, Hao, et al.
Published: (2022)
by: Fu, Hao, et al.
Published: (2022)
RoboPEPP: Vision-Based Robot Pose and Joint Angle Estimation through Embedding Predictive Pre-Training
by: Goswami, Raktim Gautam, et al.
Published: (2024)
by: Goswami, Raktim Gautam, et al.
Published: (2024)
Adversarially Robust CLIP Models Can Induce Better (Robust) Perceptual Metrics
by: Croce, Francesco, et al.
Published: (2025)
by: Croce, Francesco, et al.
Published: (2025)
3D CAVLA: Leveraging Depth and 3D Context to Generalize Vision Language Action Models for Unseen Tasks
by: Bhat, Vineet, et al.
Published: (2025)
by: Bhat, Vineet, et al.
Published: (2025)
SpotEdit: Evaluating Visually-Guided Image Editing Methods
by: Ghazanfari, Sara, et al.
Published: (2025)
by: Ghazanfari, Sara, et al.
Published: (2025)
LASER: Lip Landmark Assisted Speaker Detection for Robustness
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
by: Nguyen, Le Thien Phuc, et al.
Published: (2025)
DreamSim: Learning New Dimensions of Human Visual Similarity using Synthetic Data
by: Fu, Stephanie, et al.
Published: (2023)
by: Fu, Stephanie, et al.
Published: (2023)
Image Statistics Predict the Sensitivity of Perceptual Quality Metrics
by: Hepburn, Alexander, et al.
Published: (2023)
by: Hepburn, Alexander, et al.
Published: (2023)
Evaluation of Audio-Visual Alignments in Visually Grounded Speech Models
by: Khorrami, Khazar, et al.
Published: (2021)
by: Khorrami, Khazar, et al.
Published: (2021)
Rethinking Perceptual Metrics for Medical Image Translation
by: Konz, Nicholas, et al.
Published: (2024)
by: Konz, Nicholas, et al.
Published: (2024)
Novel Quadratic Constraints for Extending LipSDP beyond Slope-Restricted Activations
by: Pauli, Patricia, et al.
Published: (2024)
by: Pauli, Patricia, et al.
Published: (2024)
Which Models have Perceptually-Aligned Gradients? An Explanation via Off-Manifold Robustness
by: Srinivas, Suraj, et al.
Published: (2023)
by: Srinivas, Suraj, et al.
Published: (2023)
ID-Sim: An Identity-Focused Similarity Metric
by: Chae, Julia, et al.
Published: (2026)
by: Chae, Julia, et al.
Published: (2026)
SinSim: Sinkhorn-Regularized SimCLR
by: Sepanj, M. Hadi, et al.
Published: (2025)
by: Sepanj, M. Hadi, et al.
Published: (2025)
StyleLipSync: Style-based Personalized Lip-sync Video Generation
by: Ki, Taekyung, et al.
Published: (2023)
by: Ki, Taekyung, et al.
Published: (2023)
LipShiFT: A Certifiably Robust Shift-based Vision Transformer
by: Menon, Rohan, et al.
Published: (2025)
by: Menon, Rohan, et al.
Published: (2025)
Provably Robust Conformal Prediction with Improved Efficiency
by: Yan, Ge, et al.
Published: (2024)
by: Yan, Ge, et al.
Published: (2024)
BOP-ASK: Object-Interaction Reasoning for Vision-Language Models
by: Bhat, Vineet, et al.
Published: (2025)
by: Bhat, Vineet, et al.
Published: (2025)
Foundation Models Boost Low-Level Perceptual Similarity Metrics
by: Ghildyal, Abhijay, et al.
Published: (2024)
by: Ghildyal, Abhijay, et al.
Published: (2024)
How Robust Are Energy-Based Models Trained With Equilibrium Propagation?
by: Mansingh, Siddharth, et al.
Published: (2024)
by: Mansingh, Siddharth, et al.
Published: (2024)
Generalizable Blood Cell Detection via Unified Dataset and Faster R-CNN
by: Sahay, Siddharth
Published: (2025)
by: Sahay, Siddharth
Published: (2025)
Out-Of-Distribution Detection with Diversification (Provably)
by: Yao, Haiyun, et al.
Published: (2024)
by: Yao, Haiyun, et al.
Published: (2024)
SplatSim: Zero-Shot Sim2Real Transfer of RGB Manipulation Policies Using Gaussian Splatting
by: Qureshi, Mohammad Nomaan, et al.
Published: (2024)
by: Qureshi, Mohammad Nomaan, et al.
Published: (2024)
Counterfactual Explanations on Robust Perceptual Geodesics
by: Zaher, Eslam, et al.
Published: (2026)
by: Zaher, Eslam, et al.
Published: (2026)
Defending Against Gradient Inversion Attacks for Biomedical Images via Learnable Data Perturbation
by: Jiang, Shiyi, et al.
Published: (2025)
by: Jiang, Shiyi, et al.
Published: (2025)
When Does Perceptual Alignment Benefit Vision Representations?
by: Sundaram, Shobhita, et al.
Published: (2024)
by: Sundaram, Shobhita, et al.
Published: (2024)
Robustness of Practical Perceptual Hashing Algorithms to Hash-Evasion and Hash-Inversion Attacks
by: Madden, Jordan, et al.
Published: (2024)
by: Madden, Jordan, et al.
Published: (2024)
GD doesn't make the cut: Three ways that non-differentiability affects neural network training
by: Kumar, Siddharth Krishna
Published: (2024)
by: Kumar, Siddharth Krishna
Published: (2024)
Sim-to-Real Causal Transfer: A Metric Learning Approach to Causally-Aware Interaction Representations
by: Rahimi, Ahmad, et al.
Published: (2023)
by: Rahimi, Ahmad, et al.
Published: (2023)
Provable Uncertainty Decomposition via Higher-Order Calibration
by: Ahdritz, Gustaf, et al.
Published: (2024)
by: Ahdritz, Gustaf, et al.
Published: (2024)
Similar Items
-
Towards Unified Benchmark and Models for Multi-Modal Perceptual Metrics
by: Ghazanfari, Sara, et al.
Published: (2024) -
EMMA: Efficient Visual Alignment in Multi-Modal LLMs
by: Ghazanfari, Sara, et al.
Published: (2024) -
SYNCR: A Cross-Video Reasoning Benchmark with Synthetic Grounding
by: Ghazanfari, Sara, et al.
Published: (2026) -
Chain-of-Frames: Advancing Video Understanding in Multimodal LLMs via Frame-Aware Reasoning
by: Ghazanfari, Sara, et al.
Published: (2025) -
RAZER: Robust Accelerated Zero-Shot 3D Open-Vocabulary Panoptic Reconstruction with Spatio-Temporal Aggregation
by: Patel, Naman, et al.
Published: (2025)