Ar2Can: An Architect and an Artist Leveraging a Canvas for Multi-Human Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Borse, Shubhankar, Pham, Phuc, Farhadzadeh, Farzad, Choi, Seokeon, Nguyen, Phong Ha, Tran, Anh Tuan, Yun, Sungrack, Hayat, Munawar, Porikli, Fatih |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Resolving the Identity Crisis in Text-to-Image Generation
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
Zero-Shot Adaptation of Parameter-Efficient Fine-Tuning in Diffusion Models
by: Farhadzadeh, Farzad, et al.
Published: (2025)
by: Farhadzadeh, Farzad, et al.
Published: (2025)
LoRA-X: Bridging Foundation Models with Training-Free Cross-Model Adaptation
by: Farhadzadeh, Farzad, et al.
Published: (2025)
by: Farhadzadeh, Farzad, et al.
Published: (2025)
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
ConsNoTrainLoRA: Data-driven Weight Initialization of Low-rank Adapters using Constraints
by: Das, Debasmit, et al.
Published: (2025)
by: Das, Debasmit, et al.
Published: (2025)
Attention Guided Alignment in Efficient Vision-Language Models
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
Memory-Efficient Fine-Tuning Diffusion Transformers via Dynamic Patch Sampling and Block Skipping
by: Park, Sunghyun, et al.
Published: (2026)
by: Park, Sunghyun, et al.
Published: (2026)
Personalized OVSS: Understanding Personal Concept in Open-Vocabulary Semantic Segmentation
by: Park, Sunghyun, et al.
Published: (2025)
by: Park, Sunghyun, et al.
Published: (2025)
MADI: Masking-Augmented Diffusion with Inference-Time Scaling for Visual Editing
by: Kadambi, Shreya, et al.
Published: (2025)
by: Kadambi, Shreya, et al.
Published: (2025)
PosSAM: Panoptic Open-vocabulary Segment Anything
by: VS, Vibashan, et al.
Published: (2024)
by: VS, Vibashan, et al.
Published: (2024)
Hollowed Net for On-Device Personalization of Text-to-Image Diffusion Models
by: Cho, Wonguk, et al.
Published: (2024)
by: Cho, Wonguk, et al.
Published: (2024)
Tripartite Weight-Space Ensemble for Few-Shot Class-Incremental Learning
by: Lee, Juntae, et al.
Published: (2025)
by: Lee, Juntae, et al.
Published: (2025)
From Wardrobe to Canvas: Wardrobe Polyptych LoRA for Part-level Controllable Human Image Generation
by: Kim, Jeongho, et al.
Published: (2025)
by: Kim, Jeongho, et al.
Published: (2025)
Steering Guidance for Personalized Text-to-Image Diffusion Models
by: Park, Sunghyun, et al.
Published: (2025)
by: Park, Sunghyun, et al.
Published: (2025)
DDIL: Diversity Enhancing Diffusion Distillation With Imitation Learning
by: Garrepalli, Risheek, et al.
Published: (2024)
by: Garrepalli, Risheek, et al.
Published: (2024)
FouRA: Fourier Low Rank Adaptation
by: Borse, Shubhankar, et al.
Published: (2024)
by: Borse, Shubhankar, et al.
Published: (2024)
DuoLoRA : Cycle-consistent and Rank-disentangled Content-Style Personalization
by: Roy, Aniket, et al.
Published: (2025)
by: Roy, Aniket, et al.
Published: (2025)
Memory-Efficient Personalization of Text-to-Image Diffusion Models via Selective Optimization Strategies
by: Choi, Seokeon, et al.
Published: (2025)
by: Choi, Seokeon, et al.
Published: (2025)
CustomKD: Customizing Large Vision Foundation for Edge Model Improvement via Knowledge Distillation
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
FLoC: Facility Location-Based Efficient Visual Token Compression for Long Video Understanding
by: Cho, Janghoon, et al.
Published: (2025)
by: Cho, Janghoon, et al.
Published: (2025)
Feature Diversification and Adaptation for Federated Domain Generalization
by: Yang, Seunghan, et al.
Published: (2024)
by: Yang, Seunghan, et al.
Published: (2024)
SubZero: Composing Subject, Style, and Action via Zero-Shot Personalization
by: Borse, Shubhankar, et al.
Published: (2025)
by: Borse, Shubhankar, et al.
Published: (2025)
Neural Graphics Texture Compression Supporting Random Access
by: Farhadzadeh, Farzad, et al.
Published: (2024)
by: Farhadzadeh, Farzad, et al.
Published: (2024)
HyperNet Fields: Efficiently Training Hypernetworks without Ground Truth by Learning Weight Trajectories
by: Hedlin, Eric, et al.
Published: (2024)
by: Hedlin, Eric, et al.
Published: (2024)
Generalized Contrastive Learning for Universal Multimodal Retrieval
by: Lee, Jungsoo, et al.
Published: (2025)
by: Lee, Jungsoo, et al.
Published: (2025)
Sort-free Gaussian Splatting via Weighted Sum Rendering
by: Hou, Qiqi, et al.
Published: (2024)
by: Hou, Qiqi, et al.
Published: (2024)
OCAI: Improving Optical Flow Estimation by Occlusion and Consistency Aware Interpolation
by: Jeong, Jisoo, et al.
Published: (2024)
by: Jeong, Jisoo, et al.
Published: (2024)
PixelRush: Ultra-Fast, Training-Free High-Resolution Image Generation via One-step Diffusion
by: Lai, Hong-Phuc, et al.
Published: (2026)
by: Lai, Hong-Phuc, et al.
Published: (2026)
SepFormer: Coarse-to-fine Separator Regression Network for Table Structure Recognition
by: Nguyen, Nam Quan, et al.
Published: (2025)
by: Nguyen, Nam Quan, et al.
Published: (2025)
Do-Undo Bench: Reversibility for Action Understanding in Image Generation
by: Mahajan, Shweta, et al.
Published: (2025)
by: Mahajan, Shweta, et al.
Published: (2025)
Georegistration Improvements Using ArUco Markers as Ground Control Points for Fast Deployment of 3D Map Reconstruction
by: Anh Quang Nguyen, et al.
Published: (2025)
by: Anh Quang Nguyen, et al.
Published: (2025)
Improving impact strength and fracture toughness of epoxy resin through oligoester—A byproduct derived from the unsaturated polyester resin manufacturing process
by: Anh‐Tuan Pham, et al.
Published: (2024)
by: Anh‐Tuan Pham, et al.
Published: (2024)
SwiftTailor: Efficient 3D Garment Generation with Geometry Image Representation
by: Pham, Phuc, et al.
Published: (2026)
by: Pham, Phuc, et al.
Published: (2026)
FedHide: Federated Learning by Hiding in the Neighbors
by: Park, Hyunsin, et al.
Published: (2024)
by: Park, Hyunsin, et al.
Published: (2024)
ForeSea: AI Forensic Search with Multi-modal Queries for Video Surveillance
by: Park, Hyojin, et al.
Published: (2026)
by: Park, Hyojin, et al.
Published: (2026)
A4O: All Trigger for One sample
by: Vu, Duc Anh, et al.
Published: (2025)
by: Vu, Duc Anh, et al.
Published: (2025)
CodeLSI: Leveraging Foundation Models for Automated Code Generation with Low-Rank Optimization and Domain-Specific Instruction Tuning
by: Le, Huy, et al.
Published: (2025)
by: Le, Huy, et al.
Published: (2025)
Blur2Blur: Blur Conversion for Unsupervised Image Deblurring on Unknown Domains
by: Pham, Bang-Dang, et al.
Published: (2024)
by: Pham, Bang-Dang, et al.
Published: (2024)
The orbit method for the Virasoro algebra
by: Pham, Tuan Anh
Published: (2025)
by: Pham, Tuan Anh
Published: (2025)
Ultrasound‐guided posterior lateral branches steroid injections of the sacral foramina for chronic new‐onset sacroiliac joint pain management after spinal fusion surgery: A prospective study of single‐center in Vietnam
by: Tuan Anh Pham, et al.
Published: (2024)
by: Tuan Anh Pham, et al.
Published: (2024)
Similar Items
-
Resolving the Identity Crisis in Text-to-Image Generation
by: Borse, Shubhankar, et al.
Published: (2025) -
Zero-Shot Adaptation of Parameter-Efficient Fine-Tuning in Diffusion Models
by: Farhadzadeh, Farzad, et al.
Published: (2025) -
LoRA-X: Bridging Foundation Models with Training-Free Cross-Model Adaptation
by: Farhadzadeh, Farzad, et al.
Published: (2025) -
MultiHuman-Testbench: Benchmarking Image Generation for Multiple Humans
by: Borse, Shubhankar, et al.
Published: (2025) -
ConsNoTrainLoRA: Data-driven Weight Initialization of Low-rank Adapters using Constraints
by: Das, Debasmit, et al.
Published: (2025)