ID-Sim: An Identity-Focused Similarity Metric
Fuente:
arXiv
Guardado en:
| Autores principales: | Chae, Julia, Kolkin, Nicholas, Wang, Jui-Hsien, Zhang, Richard, Beery, Sara, Ham, Cusuh |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Finer-Personalization Rank: Fine-Grained Retrieval Examines Identity Preservation for Personalized Generation
por: Kilrain, Connor, et al.
Publicado: (2025)
por: Kilrain, Connor, et al.
Publicado: (2025)
DreamLoop: Controllable Cinemagraph Generation from a Single Photograph
por: Mahapatra, Aniruddha, et al.
Publicado: (2026)
por: Mahapatra, Aniruddha, et al.
Publicado: (2026)
Personalized Residuals for Concept-Driven Text-to-Image Generation
por: Ham, Cusuh, et al.
Publicado: (2024)
por: Ham, Cusuh, et al.
Publicado: (2024)
FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models
por: Lim, Youngsun, et al.
Publicado: (2026)
por: Lim, Youngsun, et al.
Publicado: (2026)
Relational Visual Similarity
por: Nguyen, Thao, et al.
Publicado: (2025)
por: Nguyen, Thao, et al.
Publicado: (2025)
Generative Models: What Do They Know? Do They Know Things? Let's Find Out!
por: Du, Xiaodan, et al.
Publicado: (2023)
por: Du, Xiaodan, et al.
Publicado: (2023)
Adapting the re-ID challenge for static sensors
por: Sundaresan, Avirath, et al.
Publicado: (2024)
por: Sundaresan, Avirath, et al.
Publicado: (2024)
Pairwise Matching of Intermediate Representations for Fine-grained Explainability
por: Shrack, Lauren, et al.
Publicado: (2025)
por: Shrack, Lauren, et al.
Publicado: (2025)
ST-SimDiff: Balancing Spatiotemporal Similarity and Difference for Efficient Video Understanding with MLLMs
por: Luo, Bingjun, et al.
Publicado: (2026)
por: Luo, Bingjun, et al.
Publicado: (2026)
Are They the Same Picture? Adapting Concept Bottleneck Models for Human-AI Collaboration in Image Retrieval
por: Balloli, Vaibhav, et al.
Publicado: (2024)
por: Balloli, Vaibhav, et al.
Publicado: (2024)
SAMScore: A Content Structural Similarity Metric for Image Translation Evaluation
por: Li, Yunxiang, et al.
Publicado: (2023)
por: Li, Yunxiang, et al.
Publicado: (2023)
TokenDial: Continuous Attribute Control in Text-to-Video via Spatiotemporal Token Offsets
por: Liu, Zhixuan, et al.
Publicado: (2026)
por: Liu, Zhixuan, et al.
Publicado: (2026)
InstantID: Zero-shot Identity-Preserving Generation in Seconds
por: Wang, Qixun, et al.
Publicado: (2024)
por: Wang, Qixun, et al.
Publicado: (2024)
Slot-ID: Identity-Preserving Video Generation from Reference Videos via Slot-Based Temporal Identity Encoding
por: Lai, Yixuan, et al.
Publicado: (2026)
por: Lai, Yixuan, et al.
Publicado: (2026)
Concat-ID: Towards Universal Identity-Preserving Video Synthesis
por: Zhong, Yong, et al.
Publicado: (2025)
por: Zhong, Yong, et al.
Publicado: (2025)
Case-Enhanced Vision Transformer: Improving Explanations of Image Similarity with a ViT-based Similarity Metric
por: Zhao, Ziwei, et al.
Publicado: (2024)
por: Zhao, Ziwei, et al.
Publicado: (2024)
ID-Aligner: Enhancing Identity-Preserving Text-to-Image Generation with Reward Feedback Learning
por: Chen, Weifeng, et al.
Publicado: (2024)
por: Chen, Weifeng, et al.
Publicado: (2024)
ConsistentID: Portrait Generation with Multimodal Fine-Grained Identity Preserving
por: Huang, Jiehui, et al.
Publicado: (2024)
por: Huang, Jiehui, et al.
Publicado: (2024)
DiffusionBrowser: Interactive Diffusion Previews via Multi-Branch Decoders
por: Hong, Susung, et al.
Publicado: (2025)
por: Hong, Susung, et al.
Publicado: (2025)
MITO: A Millimeter-Wave Dataset and Simulator for Non-Line-of-Sight Perception
por: Dodds, Laura, et al.
Publicado: (2025)
por: Dodds, Laura, et al.
Publicado: (2025)
Sim4Seg: Boosting Multimodal Multi-disease Medical Diagnosis Segmentation with Region-Aware Vision-Language Similarity Masks
por: Song, Lingran, et al.
Publicado: (2025)
por: Song, Lingran, et al.
Publicado: (2025)
Consensus-Driven Active Model Selection
por: Kay, Justin, et al.
Publicado: (2025)
por: Kay, Justin, et al.
Publicado: (2025)
AdaIR: Exploiting Underlying Similarities of Image Restoration Tasks with Adapters
por: Chen, Hao-Wei, et al.
Publicado: (2024)
por: Chen, Hao-Wei, et al.
Publicado: (2024)
Evaluating Facial Expression Recognition Datasets for Deep Learning: A Benchmark Study with Novel Similarity Metrics
por: Gaya-Morey, F. Xavier, et al.
Publicado: (2025)
por: Gaya-Morey, F. Xavier, et al.
Publicado: (2025)
DynamicID: Zero-Shot Multi-ID Image Personalization with Flexible Facial Editability
por: Hu, Xirui, et al.
Publicado: (2025)
por: Hu, Xirui, et al.
Publicado: (2025)
Semantic Similarity Score for Measuring Visual Similarity at Semantic Level
por: Fan, Senran, et al.
Publicado: (2024)
por: Fan, Senran, et al.
Publicado: (2024)
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
por: Chen, Xinyu, et al.
Publicado: (2026)
por: Chen, Xinyu, et al.
Publicado: (2026)
Focus on Focus: Focus-oriented Representation Learning and Multi-view Cross-modal Alignment for Glioma Grading
por: Pan, Li, et al.
Publicado: (2024)
por: Pan, Li, et al.
Publicado: (2024)
Tuning-Free Multi-Event Long Video Generation via Synchronized Coupled Sampling
por: Kim, Subin, et al.
Publicado: (2025)
por: Kim, Subin, et al.
Publicado: (2025)
FaceID-6M: A Large-Scale, Open-Source FaceID Customization Dataset
por: Wang, Shuhe, et al.
Publicado: (2025)
por: Wang, Shuhe, et al.
Publicado: (2025)
Sim2Radar: Toward Bridging the Radar Sim-to-Real Gap with VLM-Guided Scene Reconstruction
por: Bejerano, Emily, et al.
Publicado: (2026)
por: Bejerano, Emily, et al.
Publicado: (2026)
OccSim: Multi-kilometer Simulation with Long-horizon Occupancy World Models
por: Liu, Tianran, et al.
Publicado: (2026)
por: Liu, Tianran, et al.
Publicado: (2026)
Multi-Focused Video Group Activities Hashing
por: Qi, Zhongmiao, et al.
Publicado: (2025)
por: Qi, Zhongmiao, et al.
Publicado: (2025)
Personalized Representation from Personalized Generation
por: Sundaram, Shobhita, et al.
Publicado: (2024)
por: Sundaram, Shobhita, et al.
Publicado: (2024)
EgoSim: Egocentric World Simulator for Embodied Interaction Generation
por: Hao, Jinkun, et al.
Publicado: (2026)
por: Hao, Jinkun, et al.
Publicado: (2026)
Simple Unsupervised Knowledge Distillation With Space Similarity
por: Singh, Aditya, et al.
Publicado: (2024)
por: Singh, Aditya, et al.
Publicado: (2024)
LipSim: A Provably Robust Perceptual Similarity Metric
por: Ghazanfari, Sara, et al.
Publicado: (2023)
por: Ghazanfari, Sara, et al.
Publicado: (2023)
FENet: Focusing Enhanced Network for Lane Detection
por: Wang, Liman, et al.
Publicado: (2023)
por: Wang, Liman, et al.
Publicado: (2023)
Learning Joint ID-Textual Representation for ID-Preserving Image Synthesis
por: Liu, Zichuan, et al.
Publicado: (2025)
por: Liu, Zichuan, et al.
Publicado: (2025)
MetricAnything: Scaling Metric Depth Pretraining with Noisy Heterogeneous Sources
por: Ma, Baorui, et al.
Publicado: (2026)
por: Ma, Baorui, et al.
Publicado: (2026)
Ejemplares similares
-
Finer-Personalization Rank: Fine-Grained Retrieval Examines Identity Preservation for Personalized Generation
por: Kilrain, Connor, et al.
Publicado: (2025) -
DreamLoop: Controllable Cinemagraph Generation from a Single Photograph
por: Mahapatra, Aniruddha, et al.
Publicado: (2026) -
Personalized Residuals for Concept-Driven Text-to-Image Generation
por: Ham, Cusuh, et al.
Publicado: (2024) -
FAGER: Factually Grounded Evaluation and Refinement of Text-to-Image Models
por: Lim, Youngsun, et al.
Publicado: (2026) -
Relational Visual Similarity
por: Nguyen, Thao, et al.
Publicado: (2025)