SubZero: Composing Subject, Style, and Action via Zero-Shot Personalization
Fuente:
arXiv
Saved in:
| Main Authors: | Borse, Shubhankar, Bhardwaj, Kartikeya, Dastjerdi, Mohammad Reza Karimi, Park, Hyojin, Kadambi, Shreya, Shivakumar, Shobitha, Mandke, Prathamesh, Nayak, Ankita, Teague, Harris, Hayat, Munawar, Porikli, Fatih |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DuoLoRA : Cycle-consistent and Rank-disentangled Content-Style Personalization
by: Roy, Aniket, et al.
Published: (2025)
by: Roy, Aniket, et al.
Published: (2025)
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
by: Kadambi, Shreya, et al.
Published: (2025)
by: Kadambi, Shreya, et al.
Published: (2025)
Resolving the Identity Crisis in Text-to-Image Generation
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
FouRA: Fourier Low Rank Adaptation
by: Borse, Shubhankar, et al.
Published: (2024)
by: Borse, Shubhankar, et al.
Published: (2024)
PosSAM: Panoptic Open-vocabulary Segment Anything
by: VS, Vibashan, et al.
Published: (2024)
by: VS, Vibashan, et al.
Published: (2024)
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
Zero-Shot Adaptation of Parameter-Efficient Fine-Tuning in Diffusion Models
by: Farhadzadeh, Farzad, et al.
Published: (2025)
by: Farhadzadeh, Farzad, et al.
Published: (2025)
Personalized OVSS: Understanding Personal Concept in Open-Vocabulary Semantic Segmentation
by: Park, Sunghyun, et al.
Published: (2025)
by: Park, Sunghyun, et al.
Published: (2025)
Do-Undo Bench: Reversibility for Action Understanding in Image Generation
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
LoRA-X: Bridging Foundation Models with Training-Free Cross-Model Adaptation
by: Farhadzadeh, Farzad, et al.
Published: (2025)
by: Farhadzadeh, Farzad, et al.
Published: (2025)
Attention Guided Alignment in Efficient Vision-Language Models
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
by: Garrepalli, Risheek, et al.
Published: (2024)
by: Garrepalli, Risheek, et al.
Published: (2024)
Ar2Can: An Architect and an Artist Leveraging a Canvas for Multi-Human Generation
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
Generalized Contrastive Learning for Universal Multimodal Retrieval
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
Sparse High Rank Adapters
by: Bhardwaj, Kartikeya, et al.
Published: (2024)
by: Bhardwaj, Kartikeya, et al.
Published: (2024)
Rapid Switching and Multi-Adapter Fusion via Sparse High Rank Adapters
by: Bhardwaj, Kartikeya, et al.
Published: (2024)
by: Bhardwaj, Kartikeya, et al.
Published: (2024)
Video Reasoning without Training
by: Sridhar, Deepak, et al.
Published: (2025)
by: Sridhar, Deepak, et al.
Published: (2025)
HyperNet Fields: Efficiently Training Hypernetworks without Ground Truth by Learning Weight Trajectories
by: Hedlin, Eric, et al.
Published: (2024)
by: Hedlin, Eric, et al.
Published: (2024)
CustomKD: Customizing Large Vision Foundation for Edge Model Improvement via Knowledge Distillation
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
FLoC: Facility Location-Based Efficient Visual Token Compression for Long Video Understanding
by: Cho, Janghoon, et al.
Published: (2025)
by: Cho, Janghoon, et al.
Published: (2025)
ConsNoTrainLoRA: Data-driven Weight Initialization of Low-rank Adapters using Constraints
by: Das, Debasmit, et al.
Published: (2025)
by: Das, Debasmit, et al.
Published: (2025)
OCAI: Improving Optical Flow Estimation by Occlusion and Consistency Aware Interpolation
by: Jeong, Jisoo, et al.
Published: (2024)
by: Jeong, Jisoo, et al.
Published: (2024)
ForeSea: AI Forensic Search with Multi-modal Queries for Video Surveillance
by: Park, Hyojin, et al.
Published: (2026)
by: Park, Hyojin, et al.
Published: (2026)
Zero-Shot Neural Architecture Search: Challenges, Solutions, and Opportunities
by: Li, Guihong, et al.
Published: (2023)
by: Li, Guihong, et al.
Published: (2023)
Comprehensive Design and Transient Analysis of Novel Off‐Grid Zero Energy and Nearly Zero Emission Building with Hydrogen‐Integrated Storage System
by: Sajad Maleki Dastjerdi, et al.
Published: (2024)
by: Sajad Maleki Dastjerdi, et al.
Published: (2024)
Memory-Efficient Fine-Tuning Diffusion Transformers via Dynamic Patch Sampling and Block Skipping
by: Park, Sunghyun, et al.
Published: (2026)
by: Park, Sunghyun, et al.
Published: (2026)
Graph Attention for Heterogeneous Graphs with Positional Encoding
by: Nayak, Nikhil Shivakumar
Published: (2025)
by: Nayak, Nikhil Shivakumar
Published: (2025)
Mathematical Modeling of Option Pricing with an Extended Black-Scholes Framework
by: Nayak, Nikhil Shivakumar
Published: (2025)
by: Nayak, Nikhil Shivakumar
Published: (2025)
Oh! We Freeze: Improving Quantized Knowledge Distillation via Signal Propagation Analysis for Large Language Models
by: Bhardwaj, Kartikeya, et al.
Published: (2024)
by: Bhardwaj, Kartikeya, et al.
Published: (2024)
FlowComposer: Composable Flows for Compositional Zero-Shot Learning
by: He, Zhenqi, et al.
Published: (2026)
by: He, Zhenqi, et al.
Published: (2026)
Zero Shot Composed Image Retrieval
by: Kakarla, Santhosh, et al.
Published: (2025)
by: Kakarla, Santhosh, et al.
Published: (2025)
Zero-Shot Head Swapping in Real-World Scenarios
by: Kang, Taewoong, et al.
Published: (2025)
by: Kang, Taewoong, et al.
Published: (2025)
Tripartite Weight-Space Ensemble for Few-Shot Class-Incremental Learning
by: Lee, Juntae, et al.
Published: (2025)
by: Lee, Juntae, et al.
Published: (2025)
Zero-shot Composed Text-Image Retrieval
by: Liu, Yikun, et al.
Published: (2023)
by: Liu, Yikun, et al.
Published: (2023)
CA-LoRA: Concept-Aware LoRA for Domain-Aligned Segmentation Dataset Generation
by: Park, Minho, et al.
Published: (2025)
by: Park, Minho, et al.
Published: (2025)
Neural 5G Indoor Localization with IMU Supervision
by: Ermolov, Aleksandr, et al.
Published: (2024)
by: Ermolov, Aleksandr, et al.
Published: (2024)
Hermes Seal: Zero-Knowledge Assurance for Autonomous Vehicle Communications
by: Hasan, Munawar, et al.
Published: (2026)
by: Hasan, Munawar, et al.
Published: (2026)
Temporally and Spatially Adaptive Models for Estimating Transient Urban Populations: Integrating Interpolation Techniques With Spatial Criteria
by: Mina Sadeghi, et al.
Published: (2026)
by: Mina Sadeghi, et al.
Published: (2026)
Emolysis: A Multimodal Open-Source Group Emotion Analysis and Visualization Toolkit
by: Ghosh, Shreya, et al.
Published: (2023)
by: Ghosh, Shreya, et al.
Published: (2023)
1M-Deepfakes Detection Challenge
by: Cai, Zhixi, et al.
Published: (2024)
by: Cai, Zhixi, et al.
Published: (2024)
Similar Items
-
DuoLoRA : Cycle-consistent and Rank-disentangled Content-Style Personalization
by: Roy, Aniket, et al.
Published: (2025) -
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
by: Kadambi, Shreya, et al.
Published: (2025) -
Resolving the Identity Crisis in Text-to-Image Generation
by: Borse, Shubhankar, et al.
Published: (2025) -
FouRA: Fourier Low Rank Adaptation
by: Borse, Shubhankar, et al.
Published: (2024) -
PosSAM: Panoptic Open-vocabulary Segment Anything
by: VS, Vibashan, et al.
Published: (2024)