OPRO: Orthogonal Panel-Relative Operators for Panel-Aware In-Context Image Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Lee, Sanghyeon, Lee, Minwoo, Shin, Euijin, Kim, Kangyeol, Choi, Seunghwan, Choo, Jaegul |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning
por: Cho, Wonwoo, et al.
Publicado: (2024)
por: Cho, Wonwoo, et al.
Publicado: (2024)
Enabling Region-Specific Control via Lassos in Point-Based Colorization
por: Lee, Sanghyeon, et al.
Publicado: (2024)
por: Lee, Sanghyeon, et al.
Publicado: (2024)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
por: Kim, Kangyeol, et al.
Publicado: (2024)
por: Kim, Kangyeol, et al.
Publicado: (2024)
Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects
por: Jo, Kyungmin, et al.
Publicado: (2024)
por: Jo, Kyungmin, et al.
Publicado: (2024)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
por: Kim, Kinam, et al.
Publicado: (2025)
por: Kim, Kinam, et al.
Publicado: (2025)
Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
por: Yun, Jooyeol, et al.
Publicado: (2024)
por: Yun, Jooyeol, et al.
Publicado: (2024)
GaussianMotion: End-to-End Learning of Animatable Gaussian Avatars with Pose Guidance from Text
por: Shim, Gyumin, et al.
Publicado: (2025)
por: Shim, Gyumin, et al.
Publicado: (2025)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
por: Kim, Jeongho, et al.
Publicado: (2024)
por: Kim, Jeongho, et al.
Publicado: (2024)
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
por: Jo, Kyungmin, et al.
Publicado: (2025)
por: Jo, Kyungmin, et al.
Publicado: (2025)
CAVIS: Context-Aware Video Instance Segmentation
por: Lee, Seunghun, et al.
Publicado: (2024)
por: Lee, Seunghun, et al.
Publicado: (2024)
From Wardrobe to Canvas: Wardrobe Polyptych LoRA for Part-level Controllable Human Image Generation
por: Kim, Jeongho, et al.
Publicado: (2025)
por: Kim, Jeongho, et al.
Publicado: (2025)
Good Noise Makes Good Edits: A Training-Free Diffusion-Based Video Editing with Image and Text Prompts
por: Choi, Saemee, et al.
Publicado: (2025)
por: Choi, Saemee, et al.
Publicado: (2025)
CA-LoRA: Concept-Aware LoRA for Domain-Aligned Segmentation Dataset Generation
por: Park, Minho, et al.
Publicado: (2025)
por: Park, Minho, et al.
Publicado: (2025)
Fair Generation without Unfair Distortions: Debiasing Text-to-Image Generation with Entanglement-Free Attention
por: Park, Jeonghoon, et al.
Publicado: (2025)
por: Park, Jeonghoon, et al.
Publicado: (2025)
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
por: Yun, Jooyeol, et al.
Publicado: (2025)
por: Yun, Jooyeol, et al.
Publicado: (2025)
Investigating Pre-Training Objectives for Generalization in Vision-Based Reinforcement Learning
por: Kim, Donghu, et al.
Publicado: (2024)
por: Kim, Donghu, et al.
Publicado: (2024)
Enhancing Intrinsic Features for Debiasing via Investigating Class-Discerning Common Attributes in Bias-Contrastive Pair
por: Park, Jeonghoon, et al.
Publicado: (2024)
por: Park, Jeonghoon, et al.
Publicado: (2024)
MagiCapture: High-Resolution Multi-Concept Portrait Customization
por: Hyung, Junha, et al.
Publicado: (2023)
por: Hyung, Junha, et al.
Publicado: (2023)
Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error Revision
por: Kim, Jinhee, et al.
Publicado: (2024)
por: Kim, Jinhee, et al.
Publicado: (2024)
Adapting Pretrained ViTs with Convolution Injector for Visuo-Motor Control
por: Hwang, Dongyoon, et al.
Publicado: (2024)
por: Hwang, Dongyoon, et al.
Publicado: (2024)
Boost Your Human Image Generation Model via Direct Preference Optimization
por: Na, Sanghyeon, et al.
Publicado: (2024)
por: Na, Sanghyeon, et al.
Publicado: (2024)
PoseBridge: Bridging the Skeletonization Gap for Zero-Shot Skeleton-Based Action Recognition
por: Lee, Sanghyeon, et al.
Publicado: (2026)
por: Lee, Sanghyeon, et al.
Publicado: (2026)
Orthogonal Negative Guidance in Attention Feature Space for Text-to-Image Generation
por: Ko, Jungmin, et al.
Publicado: (2026)
por: Ko, Jungmin, et al.
Publicado: (2026)
Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation
por: Kim, Min-Jung, et al.
Publicado: (2025)
por: Kim, Min-Jung, et al.
Publicado: (2025)
Self-supervised One-Stage Learning for RF-based Multi-Person Pose Estimation
por: Shin, Seunghwan, et al.
Publicado: (2025)
por: Shin, Seunghwan, et al.
Publicado: (2025)
SelfSwapper: Self-Supervised Face Swapping via Shape Agnostic Masked AutoEncoder
por: Lee, Jaeseong, et al.
Publicado: (2024)
por: Lee, Jaeseong, et al.
Publicado: (2024)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
por: Chung, Chaeyeon, et al.
Publicado: (2024)
por: Chung, Chaeyeon, et al.
Publicado: (2024)
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
por: Kim, Min-Jung, et al.
Publicado: (2025)
por: Kim, Min-Jung, et al.
Publicado: (2025)
Regularized Training with Generated Datasets for Name-Only Transfer of Vision-Language Models
por: Park, Minho, et al.
Publicado: (2024)
por: Park, Minho, et al.
Publicado: (2024)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
por: Kim, Jeongho, et al.
Publicado: (2024)
por: Kim, Jeongho, et al.
Publicado: (2024)
3D-GSW: 3D Gaussian Splatting for Robust Watermarking
por: Jang, Youngdong, et al.
Publicado: (2024)
por: Jang, Youngdong, et al.
Publicado: (2024)
Neural Clustering for Prefractured Mesh Generation in Real-time Object Destruction
por: Kim, Seunghwan, et al.
Publicado: (2025)
por: Kim, Seunghwan, et al.
Publicado: (2025)
EgoX: Egocentric Video Generation from a Single Exocentric Video
por: Kang, Taewoong, et al.
Publicado: (2025)
por: Kang, Taewoong, et al.
Publicado: (2025)
Effective Rank Analysis and Regularization for Enhanced 3D Gaussian Splatting
por: Hyung, Junha, et al.
Publicado: (2024)
por: Hyung, Junha, et al.
Publicado: (2024)
TexAvatars : Hybrid Texel-3D Representations for Stable Rigging of Photorealistic Gaussian Head Avatars
por: Lee, Jaeseong, et al.
Publicado: (2025)
por: Lee, Jaeseong, et al.
Publicado: (2025)
Sparse autoencoders reveal selective remapping of visual concepts during adaptation
por: Lim, Hyesu, et al.
Publicado: (2024)
por: Lim, Hyesu, et al.
Publicado: (2024)
ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
por: Choi, Jinho, et al.
Publicado: (2025)
por: Choi, Jinho, et al.
Publicado: (2025)
Beta-Sigma VAE: Separating beta and decoder variance in Gaussian variational autoencoder
por: Kim, Seunghwan, et al.
Publicado: (2024)
por: Kim, Seunghwan, et al.
Publicado: (2024)
Instance-Aware Test-Time Segmentation for Continual Domain Shifts
por: Lee, Seunghwan, et al.
Publicado: (2025)
por: Lee, Seunghwan, et al.
Publicado: (2025)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
por: Hyung, Junha, et al.
Publicado: (2024)
por: Hyung, Junha, et al.
Publicado: (2024)
Ejemplares similares
-
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning
por: Cho, Wonwoo, et al.
Publicado: (2024) -
Enabling Region-Specific Control via Lassos in Point-Based Colorization
por: Lee, Sanghyeon, et al.
Publicado: (2024) -
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
por: Kim, Kangyeol, et al.
Publicado: (2024) -
Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects
por: Jo, Kyungmin, et al.
Publicado: (2024) -
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
por: Kim, Kinam, et al.
Publicado: (2025)