NearID: Identity Representation Learning via Near-identity Distractors
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Cvejic, Aleksandar, Abdal, Rameen, Eldesokey, Abdelrahman, Ghanem, Bernard, Wonka, Peter |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Mind-the-Glitch: Visual Correspondence for Detecting Inconsistencies in Subject-Driven Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2025)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2025)
EditCLIP: Representation Learning for Image Editing
von: Wang, Qian, et al.
Veröffentlicht: (2025)
von: Wang, Qian, et al.
Veröffentlicht: (2025)
PartEdit: Fine-Grained Image Editing using Pre-Trained Diffusion Models
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2025)
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2025)
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2024)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2024)
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2023)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2023)
Skill-Aligned Annotation for Reliable Evaluation in Text-to-Image Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2026)
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2026)
Zero-Shot Video Semantic Segmentation based on Pre-Trained Diffusion Models
von: Wang, Qian, et al.
Veröffentlicht: (2024)
von: Wang, Qian, et al.
Veröffentlicht: (2024)
ZeroKey: Point-Level Reasoning and Zero-Shot 3D Keypoint Detection from Large Language Models
von: Gong, Bingchen, et al.
Veröffentlicht: (2024)
von: Gong, Bingchen, et al.
Veröffentlicht: (2024)
AvatarMMC: 3D Head Avatar Generation and Editing with Multi-Modal Conditioning
von: Para, Wamiq Reyaz, et al.
Veröffentlicht: (2024)
von: Para, Wamiq Reyaz, et al.
Veröffentlicht: (2024)
SoccerLens: Grounded Soccer Video Understanding Beyond Accuracy
von: Elsharkawi, Ismael, et al.
Veröffentlicht: (2026)
von: Elsharkawi, Ismael, et al.
Veröffentlicht: (2026)
CounterCount: A Diagnostic Framework for Counting Bias in Vision Language Models
von: Alzahrani, Reem, et al.
Veröffentlicht: (2026)
von: Alzahrani, Reem, et al.
Veröffentlicht: (2026)
Helix4D: Complex 4D Mesh Generation
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2026)
von: Yenphraphai, Jiraphon, et al.
Veröffentlicht: (2026)
Visual Personalization Turing Test
von: Abdal, Rameen, et al.
Veröffentlicht: (2026)
von: Abdal, Rameen, et al.
Veröffentlicht: (2026)
Out-of-Distribution Segmentation via Wasserstein-Based Evidential Uncertainty
von: Brosch, Arnold, et al.
Veröffentlicht: (2025)
von: Brosch, Arnold, et al.
Veröffentlicht: (2025)
LaGeM: A Large Geometry Model for 3D Representation Learning and Diffusion
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
von: Zhang, Biao, et al.
Veröffentlicht: (2024)
TrackMAE: Video Representation Learning via Track Mask and Predict
von: Vandeghen, Renaud, et al.
Veröffentlicht: (2026)
von: Vandeghen, Renaud, et al.
Veröffentlicht: (2026)
efunc: An Efficient Function Representation without Neural Networks
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
von: Zhang, Biao, et al.
Veröffentlicht: (2025)
Vivid-ZOO: Multi-View Video Generation with Diffusion Model
von: Li, Bing, et al.
Veröffentlicht: (2024)
von: Li, Bing, et al.
Veröffentlicht: (2024)
PlaceIt3D: Language-Guided Object Placement in Real 3D Scenes
von: Abdelreheem, Ahmed, et al.
Veröffentlicht: (2025)
von: Abdelreheem, Ahmed, et al.
Veröffentlicht: (2025)
Tuning-free Visual Effect Transfer across Videos
von: Jones, Maxwell, et al.
Veröffentlicht: (2026)
von: Jones, Maxwell, et al.
Veröffentlicht: (2026)
Distractors-Immune Representation Learning with Cross-modal Contrastive Regularization for Change Captioning
von: Tu, Yunbin, et al.
Veröffentlicht: (2024)
von: Tu, Yunbin, et al.
Veröffentlicht: (2024)
LASPA: Latent Spatial Alignment for Fast Training-free Single Image Editing
von: Alharbi, Yazeed, et al.
Veröffentlicht: (2024)
von: Alharbi, Yazeed, et al.
Veröffentlicht: (2024)
Omni-ID: Holistic Identity Representation Designed for Generative Tasks
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
von: Qian, Guocheng, et al.
Veröffentlicht: (2024)
ImmersePro: End-to-End Stereo Video Synthesis Via Implicit Disparity Learning
von: Shi, Jian, et al.
Veröffentlicht: (2024)
von: Shi, Jian, et al.
Veröffentlicht: (2024)
Infinite-ID: Identity-preserved Personalization via ID-semantics Decoupling Paradigm
von: Wu, Yi, et al.
Veröffentlicht: (2024)
von: Wu, Yi, et al.
Veröffentlicht: (2024)
EasyV2V: A High-quality Instruction-based Video Editing Framework
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
von: Mai, Jinjie, et al.
Veröffentlicht: (2025)
Dynamic Concepts Personalization from Single Videos
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
Zero-Shot Dynamic Concept Personalization with Grid-Based LoRA
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
von: Abdal, Rameen, et al.
Veröffentlicht: (2025)
SEVERE++: Evaluating Benchmark Sensitivity in Generalization of Video Representation Learning
von: Thoker, Fida Mohammad, et al.
Veröffentlicht: (2025)
von: Thoker, Fida Mohammad, et al.
Veröffentlicht: (2025)
$β$-CLIP: Text-Conditioned Contrastive Learning for Multi-Granular Vision-Language Alignment
von: Zohra, Fatimah, et al.
Veröffentlicht: (2025)
von: Zohra, Fatimah, et al.
Veröffentlicht: (2025)
PoseGAM: Robust Unseen Object Pose Estimation via Geometry-Aware Multi-View Reasoning
von: Chen, Jianqi, et al.
Veröffentlicht: (2025)
von: Chen, Jianqi, et al.
Veröffentlicht: (2025)
Video Self-Stitching Graph Network for Temporal Action Localization
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
von: Zhao, Chen, et al.
Veröffentlicht: (2020)
Interpreting the Weight Space of Customized Diffusion Models
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
von: Dravid, Amil, et al.
Veröffentlicht: (2024)
Generative Human Geometry Distribution
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
von: Tang, Xiangjun, et al.
Veröffentlicht: (2025)
SMILE: Infusing Spatial and Motion Semantics in Masked Video Learning
von: Thoker, Fida Mohammad, et al.
Veröffentlicht: (2025)
von: Thoker, Fida Mohammad, et al.
Veröffentlicht: (2025)
ArtifactLens: Hundreds of Labels Are Enough for Artifact Detection with VLMs
von: Burgess, James, et al.
Veröffentlicht: (2026)
von: Burgess, James, et al.
Veröffentlicht: (2026)
ID-Booth: Identity-consistent Face Generation with Diffusion Models
von: Tomašević, Darian, et al.
Veröffentlicht: (2025)
von: Tomašević, Darian, et al.
Veröffentlicht: (2025)
No Mesh, No Problem: Estimating Coral Volume and Surface from Sparse Multi-View Images
von: Farchione, Diego Eustachio, et al.
Veröffentlicht: (2025)
von: Farchione, Diego Eustachio, et al.
Veröffentlicht: (2025)
PatchRefiner: Leveraging Synthetic Data for Real-Domain High-Resolution Monocular Metric Depth Estimation
von: Li, Zhenyu, et al.
Veröffentlicht: (2024)
von: Li, Zhenyu, et al.
Veröffentlicht: (2024)
Improving the Diffusability of Autoencoders
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
von: Skorokhodov, Ivan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Mind-the-Glitch: Visual Correspondence for Detecting Inconsistencies in Subject-Driven Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2025) -
EditCLIP: Representation Learning for Image Editing
von: Wang, Qian, et al.
Veröffentlicht: (2025) -
PartEdit: Fine-Grained Image Editing using Pre-Trained Diffusion Models
von: Cvejic, Aleksandar, et al.
Veröffentlicht: (2025) -
Build-A-Scene: Interactive 3D Layout Control for Diffusion-Based Image Generation
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2024) -
LatentMan: Generating Consistent Animated Characters using Image Diffusion Models
von: Eldesokey, Abdelrahman, et al.
Veröffentlicht: (2023)