Generative Phomosaic with Structure-Aligned and Personalized Diffusion
Fuente:
arXiv
Salvato in:
| Autori principali: | Chung, Jaeyoung, Son, Hyunjin, Lee, Kyoung Mu |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2026
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
InpaintSLat: Inpainting Structured 3D Latents via Initial Noise Optimization
di: Chung, Jaeyoung, et al.
Pubblicazione: (2026)
di: Chung, Jaeyoung, et al.
Pubblicazione: (2026)
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation
di: Choi, Jeongsoo, et al.
Pubblicazione: (2025)
di: Choi, Jeongsoo, et al.
Pubblicazione: (2025)
MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration
di: Lee, Kanggeon, et al.
Pubblicazione: (2026)
di: Lee, Kanggeon, et al.
Pubblicazione: (2026)
Auto-regressive transformation for image alignment
di: Lee, Kanggeon, et al.
Pubblicazione: (2025)
di: Lee, Kanggeon, et al.
Pubblicazione: (2025)
Depth-Regularized Optimization for 3D Gaussian Splatting in Few-Shot Images
di: Chung, Jaeyoung, et al.
Pubblicazione: (2023)
di: Chung, Jaeyoung, et al.
Pubblicazione: (2023)
CLOSER: Towards Better Representation Learning for Few-Shot Class-Incremental Learning
di: Oh, Junghun, et al.
Pubblicazione: (2024)
di: Oh, Junghun, et al.
Pubblicazione: (2024)
ODGS: 3D Scene Reconstruction from Omnidirectional Images with 3D Gaussian Splattings
di: Lee, Suyoung, et al.
Pubblicazione: (2024)
di: Lee, Suyoung, et al.
Pubblicazione: (2024)
MEIL-NeRF: Memory-Efficient Incremental Learning of Neural Radiance Fields
di: Chung, Jaeyoung, et al.
Pubblicazione: (2022)
di: Chung, Jaeyoung, et al.
Pubblicazione: (2022)
DeblurGS: Gaussian Splatting for Camera Motion Blur
di: Oh, Jeongtaek, et al.
Pubblicazione: (2024)
di: Oh, Jeongtaek, et al.
Pubblicazione: (2024)
Map2World: Segment Map Conditioned Text to 3D World Generation
di: Chung, Jaeyoung, et al.
Pubblicazione: (2026)
di: Chung, Jaeyoung, et al.
Pubblicazione: (2026)
TeHOR: Text-Guided 3D Human and Object Reconstruction with Textures
di: Nam, Hyeongjin, et al.
Pubblicazione: (2026)
di: Nam, Hyeongjin, et al.
Pubblicazione: (2026)
MMPB: It's Time for Multi-Modal Personalization
di: Kim, Jaeik, et al.
Pubblicazione: (2025)
di: Kim, Jaeik, et al.
Pubblicazione: (2025)
RadiomicsRetrieval: A Customizable Framework for Medical Image Retrieval Using Radiomics Features
di: Na, Inye, et al.
Pubblicazione: (2025)
di: Na, Inye, et al.
Pubblicazione: (2025)
DeClotH: Decomposable 3D Cloth and Human Body Reconstruction from a Single Image
di: Nam, Hyeongjin, et al.
Pubblicazione: (2025)
di: Nam, Hyeongjin, et al.
Pubblicazione: (2025)
Personalized Reward Modeling for Text-to-Image Generation
di: Lee, Jeongeun, et al.
Pubblicazione: (2025)
di: Lee, Jeongeun, et al.
Pubblicazione: (2025)
Fusion Embedding for Pose-Guided Person Image Synthesis with Diffusion Model
di: Lee, Donghwna, et al.
Pubblicazione: (2024)
di: Lee, Donghwna, et al.
Pubblicazione: (2024)
Video Diffusion Models Excel at Tracking Similar-Looking Objects Without Supervision
di: Zhang, Chenshuang, et al.
Pubblicazione: (2025)
di: Zhang, Chenshuang, et al.
Pubblicazione: (2025)
Efficient Personalization of Quantized Diffusion Model without Backpropagation
di: Seo, Hoigi, et al.
Pubblicazione: (2025)
di: Seo, Hoigi, et al.
Pubblicazione: (2025)
Aligning Diffusion Models with Noise-Conditioned Perception
di: Gambashidze, Alexander, et al.
Pubblicazione: (2024)
di: Gambashidze, Alexander, et al.
Pubblicazione: (2024)
Keep What Audio Cannot Say: Context-Preserving Token Pruning for Omni-LLMs
di: Jung, Chaeyoung, et al.
Pubblicazione: (2026)
di: Jung, Chaeyoung, et al.
Pubblicazione: (2026)
Rotation-Aligned Key Channel Pruning for Efficient Vision-Language Model Inference
di: Kang, Beomseok, et al.
Pubblicazione: (2026)
di: Kang, Beomseok, et al.
Pubblicazione: (2026)
Latent Expression Generation for Referring Image Segmentation and Grounding
di: Yu, Seonghoon, et al.
Pubblicazione: (2025)
di: Yu, Seonghoon, et al.
Pubblicazione: (2025)
Multilingual Text-to-Image Person Retrieval via Bidirectional Relation Reasoning and Aligning
di: Cao, Min, et al.
Pubblicazione: (2025)
di: Cao, Min, et al.
Pubblicazione: (2025)
NOVA3D: Normal Aligned Video Diffusion Model for Single Image to 3D Generation
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
di: Yang, Yuxiao, et al.
Pubblicazione: (2025)
PersonaCraft: Personalized and Controllable Full-Body Multi-Human Scene Generation Using Occlusion-Aware 3D-Conditioned Diffusion
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
di: Kim, Gwanghyun, et al.
Pubblicazione: (2024)
Pixel-Aligned Multi-View Generation with Depth Guided Decoder
di: Tang, Zhenggang, et al.
Pubblicazione: (2024)
di: Tang, Zhenggang, et al.
Pubblicazione: (2024)
DiffLoRA: Generating Personalized Low-Rank Adaptation Weights with Diffusion
di: Wu, Yujia, et al.
Pubblicazione: (2024)
di: Wu, Yujia, et al.
Pubblicazione: (2024)
Disrupting Diffusion-based Inpainters with Semantic Digression
di: Son, Geonho, et al.
Pubblicazione: (2024)
di: Son, Geonho, et al.
Pubblicazione: (2024)
Person Re-ID in 2025: Supervised, Self-Supervised, and Language-Aligned. What Works?
di: Balasubramanian, Lakshman
Pubblicazione: (2026)
di: Balasubramanian, Lakshman
Pubblicazione: (2026)
MANGO: A Global Single-Date Paired Dataset for Mangrove Segmentation
di: Heo, Junhyuk, et al.
Pubblicazione: (2026)
di: Heo, Junhyuk, et al.
Pubblicazione: (2026)
Short-term Object Interaction Anticipation with Disentangled Object Detection @ Ego4D Short Term Object Interaction Anticipation Challenge
di: Cho, Hyunjin, et al.
Pubblicazione: (2024)
di: Cho, Hyunjin, et al.
Pubblicazione: (2024)
OmniSplat: Taming Feed-Forward 3D Gaussian Splatting for Omnidirectional Images with Editable Capabilities
di: Lee, Suyoung, et al.
Pubblicazione: (2024)
di: Lee, Suyoung, et al.
Pubblicazione: (2024)
FocusDPO: Dynamic Preference Optimization for Multi-Subject Personalized Image Generation via Adaptive Focus
di: Jin, Qiaoqiao, et al.
Pubblicazione: (2025)
di: Jin, Qiaoqiao, et al.
Pubblicazione: (2025)
YoChameleon: Personalized Vision and Language Generation
di: Nguyen, Thao, et al.
Pubblicazione: (2025)
di: Nguyen, Thao, et al.
Pubblicazione: (2025)
InstructBooth: Instruction-following Personalized Text-to-Image Generation
di: Chae, Daewon, et al.
Pubblicazione: (2023)
di: Chae, Daewon, et al.
Pubblicazione: (2023)
Utilizing Graph Generation for Enhanced Domain Adaptive Object Detection
di: Wang, Mu
Pubblicazione: (2024)
di: Wang, Mu
Pubblicazione: (2024)
PASTA: Part-Aware Sketch-to-3D Shape Generation with Text-Aligned Prior
di: Lee, Seunggwan, et al.
Pubblicazione: (2025)
di: Lee, Seunggwan, et al.
Pubblicazione: (2025)
GLOS: Sign Language Generation with Temporally Aligned Gloss-Level Conditioning
di: Lee, Taeryung, et al.
Pubblicazione: (2025)
di: Lee, Taeryung, et al.
Pubblicazione: (2025)
AlignDiff: Learning Physically-Grounded Camera Alignment via Diffusion
di: Xie, Liuyue, et al.
Pubblicazione: (2025)
di: Xie, Liuyue, et al.
Pubblicazione: (2025)
AJAHR: Amputated Joint Aware 3D Human Mesh Recovery
di: Cho, Hyunjin, et al.
Pubblicazione: (2025)
di: Cho, Hyunjin, et al.
Pubblicazione: (2025)
Documenti analoghi
-
InpaintSLat: Inpainting Structured 3D Latents via Initial Noise Optimization
di: Chung, Jaeyoung, et al.
Pubblicazione: (2026) -
AlignDiT: Multimodal Aligned Diffusion Transformer for Synchronized Speech Generation
di: Choi, Jeongsoo, et al.
Pubblicazione: (2025) -
MatRes: Zero-Shot Test-Time Model Adaptation for Simultaneous Matching and Restoration
di: Lee, Kanggeon, et al.
Pubblicazione: (2026) -
Auto-regressive transformation for image alignment
di: Lee, Kanggeon, et al.
Pubblicazione: (2025) -
Depth-Regularized Optimization for 3D Gaussian Splatting in Few-Shot Images
di: Chung, Jaeyoung, et al.
Pubblicazione: (2023)