Zero-Shot Personalization of Objects via Textual Inversion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Roy, Aniket, Suin, Maitreya, Chellappa, Rama |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MultLFG: Training-free Multi-LoRA composition using Frequency-domain Guidance
von: Roy, Aniket, et al.
Veröffentlicht: (2025)
von: Roy, Aniket, et al.
Veröffentlicht: (2025)
CLR-Face: Conditional Latent Refinement for Blind Face Restoration Using Score-Based Diffusion Models
von: Suin, Maitreya, et al.
Veröffentlicht: (2024)
von: Suin, Maitreya, et al.
Veröffentlicht: (2024)
Spatially-Attentive Patch-Hierarchical Network with Adaptive Sampling for Motion Deblurring
von: Suin, Maitreya, et al.
Veröffentlicht: (2024)
von: Suin, Maitreya, et al.
Veröffentlicht: (2024)
Test-Time-Scaling for Zero-Shot Diagnosis with Visual-Language Reasoning
von: Byun, Ji Young, et al.
Veröffentlicht: (2025)
von: Byun, Ji Young, et al.
Veröffentlicht: (2025)
Fine-grained Textual Inversion Network for Zero-Shot Composed Image Retrieval
von: Lin, Haoqiang, et al.
Veröffentlicht: (2025)
von: Lin, Haoqiang, et al.
Veröffentlicht: (2025)
DiffInf: Influence-Guided Diffusion for Supervision Alignment in Facial Attribute Learning
von: Pal, Basudha, et al.
Veröffentlicht: (2026)
von: Pal, Basudha, et al.
Veröffentlicht: (2026)
Template-based Multi-Domain Face Recognition
von: Nanduri, Anirudh, et al.
Veröffentlicht: (2024)
von: Nanduri, Anirudh, et al.
Veröffentlicht: (2024)
MimicGait: A Model Agnostic approach for Occluded Gait Recognition using Correlational Knowledge Distillation
von: Gupta, Ayush, et al.
Veröffentlicht: (2025)
von: Gupta, Ayush, et al.
Veröffentlicht: (2025)
Mind the Gap: Bridging Occlusion in Gait Recognition via Residual Gap Correction
von: Gupta, Ayush, et al.
Veröffentlicht: (2025)
von: Gupta, Ayush, et al.
Veröffentlicht: (2025)
iSEARLE: Improving Textual Inversion for Zero-Shot Composed Image Retrieval
von: Agnolucci, Lorenzo, et al.
Veröffentlicht: (2024)
von: Agnolucci, Lorenzo, et al.
Veröffentlicht: (2024)
VILLS -- Video-Image Learning to Learn Semantics for Person Re-Identification
von: Huang, Siyuan, et al.
Veröffentlicht: (2023)
von: Huang, Siyuan, et al.
Veröffentlicht: (2023)
ViT-Linearizer: Distilling Quadratic Knowledge into Linear-Time Vision Models
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
von: Wei, Guoyizhe, et al.
Veröffentlicht: (2025)
Zero-Shot Object-Centric Representation Learning
von: Didolkar, Aniket, et al.
Veröffentlicht: (2024)
von: Didolkar, Aniket, et al.
Veröffentlicht: (2024)
Zero-Shot Textual Explanations via Translating Decision-Critical Features
von: Yamauchi, Toshinori, et al.
Veröffentlicht: (2025)
von: Yamauchi, Toshinori, et al.
Veröffentlicht: (2025)
A Quantitative Evaluation of the Expressivity of BMI, Pose and Gender in Body Embeddings for Recognition and Identification
von: Pal, Basudha, et al.
Veröffentlicht: (2025)
von: Pal, Basudha, et al.
Veröffentlicht: (2025)
Cross-Spectral Body Recognition with Side Information Embedding: Benchmarks on LLCM and Analyzing Range-Induced Occlusions on IJB-MDF
von: Nanduri, Anirudh, et al.
Veröffentlicht: (2025)
von: Nanduri, Anirudh, et al.
Veröffentlicht: (2025)
Multi-Domain Biometric Recognition using Body Embeddings
von: Nanduri, Anirudh, et al.
Veröffentlicht: (2025)
von: Nanduri, Anirudh, et al.
Veröffentlicht: (2025)
Zero-Shot Faithful Textual Explanations via Directional-Derivative Influence on Predictions
von: Yamauchi, Toshinori, et al.
Veröffentlicht: (2026)
von: Yamauchi, Toshinori, et al.
Veröffentlicht: (2026)
Directional Textual Inversion for Personalized Text-to-Image Generation
von: Kim, Kunhee, et al.
Veröffentlicht: (2025)
von: Kim, Kunhee, et al.
Veröffentlicht: (2025)
Zero-Shot Temporal Action Localization Through Textual Guidance
von: Liberatori, Benedetta, et al.
Veröffentlicht: (2026)
von: Liberatori, Benedetta, et al.
Veröffentlicht: (2026)
Training Free Zero-Shot Visual Anomaly Localization via Diffusion Inversion
von: Hicsonmez, Samet, et al.
Veröffentlicht: (2026)
von: Hicsonmez, Samet, et al.
Veröffentlicht: (2026)
Textual Inversion for Efficient Adaptation of Open-Vocabulary Object Detectors Without Forgetting
von: Ruis, Frank, et al.
Veröffentlicht: (2025)
von: Ruis, Frank, et al.
Veröffentlicht: (2025)
SyncFix: Fixing 3D Reconstructions via Multi-View Synchronization
von: Li, Deming, et al.
Veröffentlicht: (2026)
von: Li, Deming, et al.
Veröffentlicht: (2026)
Harmonizing Visual and Textual Embeddings for Zero-Shot Text-to-Image Customization
von: Song, Yeji, et al.
Veröffentlicht: (2024)
von: Song, Yeji, et al.
Veröffentlicht: (2024)
SubZero: Composing Subject, Style, and Action via Zero-Shot Personalization
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
von: Borse, Shubhankar, et al.
Veröffentlicht: (2025)
CAM3R: Camera-Agnostic Model for 3D Reconstruction
von: Guruprasad, Namitha, et al.
Veröffentlicht: (2026)
von: Guruprasad, Namitha, et al.
Veröffentlicht: (2026)
SPIDER: Spatial Image CorresponDence Estimator for Robust Calibration
von: Shao, Zhimin, et al.
Veröffentlicht: (2025)
von: Shao, Zhimin, et al.
Veröffentlicht: (2025)
Prompt-driven Transferable Adversarial Attack on Person Re-Identification with Attribute-aware Textual Inversion
von: Bian, Yuan, et al.
Veröffentlicht: (2025)
von: Bian, Yuan, et al.
Veröffentlicht: (2025)
SPIQA: A Dataset for Multimodal Question Answering on Scientific Papers
von: Pramanick, Shraman, et al.
Veröffentlicht: (2024)
von: Pramanick, Shraman, et al.
Veröffentlicht: (2024)
ReplicateAnyScene: Zero-Shot Video-to-3D Composition via Textual-Visual-Spatial Alignment
von: Dong, Mingyu, et al.
Veröffentlicht: (2026)
von: Dong, Mingyu, et al.
Veröffentlicht: (2026)
Uncertainty-quantified Pulse Signal Recovery from Facial Video using Regularized Stochastic Interpolants
von: Shenoy, Vineet R., et al.
Veröffentlicht: (2026)
von: Shenoy, Vineet R., et al.
Veröffentlicht: (2026)
FaceXFormer: A Unified Transformer for Facial Analysis
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
von: Narayan, Kartik, et al.
Veröffentlicht: (2024)
FusionRF: High-Fidelity Satellite Neural Radiance Fields from Multispectral and Panchromatic Acquisitions
von: Sprintson, Michael, et al.
Veröffentlicht: (2024)
von: Sprintson, Michael, et al.
Veröffentlicht: (2024)
DegustaBot: Zero-Shot Visual Preference Estimation for Personalized Multi-Object Rearrangement
von: Newman, Benjamin A., et al.
Veröffentlicht: (2024)
von: Newman, Benjamin A., et al.
Veröffentlicht: (2024)
ID-EA: Identity-driven Text Enhancement and Adaptation with Textual Inversion for Personalized Text-to-Image Generation
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
von: Jin, Hyun-Jun, et al.
Veröffentlicht: (2025)
DiffRegCD: Integrated Registration and Change Detection with Diffusion Features
von: Madani, Seyedehanita, et al.
Veröffentlicht: (2025)
von: Madani, Seyedehanita, et al.
Veröffentlicht: (2025)
Fine-Grained Zero-Shot Object Detection
von: Ma, Hongxu, et al.
Veröffentlicht: (2025)
von: Ma, Hongxu, et al.
Veröffentlicht: (2025)
Text-guided Zero-Shot Object Localization
von: Wang, Jingjing, et al.
Veröffentlicht: (2024)
von: Wang, Jingjing, et al.
Veröffentlicht: (2024)
Zero-Shot Multi-Object Scene Completion
von: Iwase, Shun, et al.
Veröffentlicht: (2024)
von: Iwase, Shun, et al.
Veröffentlicht: (2024)
AttriBE: Quantifying Attribute Expressivity in Body Embeddings for Recognition and Identification
von: Pal, Basudha, et al.
Veröffentlicht: (2026)
von: Pal, Basudha, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MultLFG: Training-free Multi-LoRA composition using Frequency-domain Guidance
von: Roy, Aniket, et al.
Veröffentlicht: (2025) -
CLR-Face: Conditional Latent Refinement for Blind Face Restoration Using Score-Based Diffusion Models
von: Suin, Maitreya, et al.
Veröffentlicht: (2024) -
Spatially-Attentive Patch-Hierarchical Network with Adaptive Sampling for Motion Deblurring
von: Suin, Maitreya, et al.
Veröffentlicht: (2024) -
Test-Time-Scaling for Zero-Shot Diagnosis with Visual-Language Reasoning
von: Byun, Ji Young, et al.
Veröffentlicht: (2025) -
Fine-grained Textual Inversion Network for Zero-Shot Composed Image Retrieval
von: Lin, Haoqiang, et al.
Veröffentlicht: (2025)