Scaling Up Personalized Image Aesthetic Assessment via Task Vector Customization
Fuente:
arXiv
Salvato in:
| Autori principali: | Yun, Jooyeol, Choo, Jaegul |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
di: Yun, Jooyeol, et al.
Pubblicazione: (2025)
di: Yun, Jooyeol, et al.
Pubblicazione: (2025)
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
di: Jo, Kyungmin, et al.
Pubblicazione: (2025)
di: Jo, Kyungmin, et al.
Pubblicazione: (2025)
Enabling Region-Specific Control via Lassos in Point-Based Colorization
di: Lee, Sanghyeon, et al.
Pubblicazione: (2024)
di: Lee, Sanghyeon, et al.
Pubblicazione: (2024)
Regularized Training with Generated Datasets for Name-Only Transfer of Vision-Language Models
di: Park, Minho, et al.
Pubblicazione: (2024)
di: Park, Minho, et al.
Pubblicazione: (2024)
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
di: Park, Minho, et al.
Pubblicazione: (2025)
di: Park, Minho, et al.
Pubblicazione: (2025)
DesignLab: Designing Slides Through Iterative Detection and Correction
di: Yun, Jooyeol, et al.
Pubblicazione: (2025)
di: Yun, Jooyeol, et al.
Pubblicazione: (2025)
Imagining the Unseen: Generative Location Modeling for Object Placement
di: Yun, Jooyeol, et al.
Pubblicazione: (2024)
di: Yun, Jooyeol, et al.
Pubblicazione: (2024)
Skip-and-Play: Depth-Driven Pose-Preserved Image Generation for Any Objects
di: Jo, Kyungmin, et al.
Pubblicazione: (2024)
di: Jo, Kyungmin, et al.
Pubblicazione: (2024)
MagiCapture: High-Resolution Multi-Concept Portrait Customization
di: Hyung, Junha, et al.
Pubblicazione: (2023)
di: Hyung, Junha, et al.
Pubblicazione: (2023)
TV-LiVE: Training-Free, Text-Guided Video Editing via Layer Informed Vitality Exploitation
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
Temporal In-Context Fine-Tuning with Temporal Reasoning for Versatile Control of Video Diffusion Models
di: Kim, Kinam, et al.
Pubblicazione: (2025)
di: Kim, Kinam, et al.
Pubblicazione: (2025)
GaussianMotion: End-to-End Learning of Animatable Gaussian Avatars with Pose Guidance from Text
di: Shim, Gyumin, et al.
Pubblicazione: (2025)
di: Shim, Gyumin, et al.
Pubblicazione: (2025)
From Wardrobe to Canvas: Wardrobe Polyptych LoRA for Part-level Controllable Human Image Generation
di: Kim, Jeongho, et al.
Pubblicazione: (2025)
di: Kim, Jeongho, et al.
Pubblicazione: (2025)
Enhancing Intrinsic Features for Debiasing via Investigating Class-Discerning Common Attributes in Bias-Contrastive Pair
di: Park, Jeonghoon, et al.
Pubblicazione: (2024)
di: Park, Jeonghoon, et al.
Pubblicazione: (2024)
Image Aesthetics Assessment via Learnable Queries
di: Xiong, Zhiwei, et al.
Pubblicazione: (2023)
di: Xiong, Zhiwei, et al.
Pubblicazione: (2023)
Good Noise Makes Good Edits: A Training-Free Diffusion-Based Video Editing with Image and Text Prompts
di: Choi, Saemee, et al.
Pubblicazione: (2025)
di: Choi, Saemee, et al.
Pubblicazione: (2025)
Bones Can't Be Triangles: Accurate and Efficient Vertebrae Keypoint Estimation through Collaborative Error Revision
di: Kim, Jinhee, et al.
Pubblicazione: (2024)
di: Kim, Jinhee, et al.
Pubblicazione: (2024)
Training Spatial-Frequency Visual Prompts and Probabilistic Clusters for Accurate Black-Box Transfer Learning
di: Cho, Wonwoo, et al.
Pubblicazione: (2024)
di: Cho, Wonwoo, et al.
Pubblicazione: (2024)
What to Preserve and What to Transfer: Faithful, Identity-Preserving Diffusion-based Hairstyle Transfer
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
di: Chung, Chaeyeon, et al.
Pubblicazione: (2024)
Zero-Shot Head Swapping in Real-World Scenarios
di: Kang, Taewoong, et al.
Pubblicazione: (2025)
di: Kang, Taewoong, et al.
Pubblicazione: (2025)
ConceptScope: Characterizing Dataset Bias via Disentangled Visual Concepts
di: Choi, Jinho, et al.
Pubblicazione: (2025)
di: Choi, Jinho, et al.
Pubblicazione: (2025)
What Do Vision-Language Models Encode for Personalized Image Aesthetics Assessment?
di: Ryu, Koki, et al.
Pubblicazione: (2026)
di: Ryu, Koki, et al.
Pubblicazione: (2026)
SelfSwapper: Self-Supervised Face Swapping via Shape Agnostic Masked AutoEncoder
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
di: Lee, Jaeseong, et al.
Pubblicazione: (2024)
Learning to See What You Need: Gaze Attention for Multimodal Large Language Models
di: Song, Junha, et al.
Pubblicazione: (2026)
di: Song, Junha, et al.
Pubblicazione: (2026)
PromptDresser: Improving the Quality and Controllability of Virtual Try-On via Generative Textual Prompt and Prompt-aware Mask
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
TCAN: Animating Human Images with Temporally Consistent Pose Guidance using Diffusion Models
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
di: Kim, Jeongho, et al.
Pubblicazione: (2024)
OPRO: Orthogonal Panel-Relative Operators for Panel-Aware In-Context Image Generation
di: Lee, Sanghyeon, et al.
Pubblicazione: (2026)
di: Lee, Sanghyeon, et al.
Pubblicazione: (2026)
RL makes MLLMs see better than SFT
di: Song, Junha, et al.
Pubblicazione: (2025)
di: Song, Junha, et al.
Pubblicazione: (2025)
Layout-and-Retouch: A Dual-stage Framework for Improving Diversity in Personalized Image Generation
di: Kim, Kangyeol, et al.
Pubblicazione: (2024)
di: Kim, Kangyeol, et al.
Pubblicazione: (2024)
One Model, Two Minds: Task-Conditioned Reasoning for Unified Image Quality and Aesthetic Assessment
di: Yin, Wen, et al.
Pubblicazione: (2026)
di: Yin, Wen, et al.
Pubblicazione: (2026)
Anti-Aesthetics: Protecting Facial Privacy against Customized Text-to-Image Synthesis
di: Wang, Songping, et al.
Pubblicazione: (2025)
di: Wang, Songping, et al.
Pubblicazione: (2025)
Multi-modal Learnable Queries for Image Aesthetics Assessment
di: Xiong, Zhiwei, et al.
Pubblicazione: (2024)
di: Xiong, Zhiwei, et al.
Pubblicazione: (2024)
From Concepts to Judgments: Interpretable Image Aesthetic Assessment
di: Liu, Xiao-Chang, et al.
Pubblicazione: (2026)
di: Liu, Xiao-Chang, et al.
Pubblicazione: (2026)
Cross-Frame Representation Alignment for Fine-Tuning Video Diffusion Models
di: Hwang, Sungwon, et al.
Pubblicazione: (2025)
di: Hwang, Sungwon, et al.
Pubblicazione: (2025)
Enhancing Zero-shot Personalized Image Aesthetics Assessment with Profile-aware Multimodal LLM
di: Wang, Chun, et al.
Pubblicazione: (2026)
di: Wang, Chun, et al.
Pubblicazione: (2026)
MM-SeR: Multimodal Self-Refinement for Lightweight Image Captioning
di: Song, Junha, et al.
Pubblicazione: (2025)
di: Song, Junha, et al.
Pubblicazione: (2025)
Sparse autoencoders reveal selective remapping of visual concepts during adaptation
di: Lim, Hyesu, et al.
Pubblicazione: (2024)
di: Lim, Hyesu, et al.
Pubblicazione: (2024)
VEGS: View Extrapolation of Urban Scenes in 3D Gaussian Splatting using Learned Priors
di: Hwang, Sungwon, et al.
Pubblicazione: (2024)
di: Hwang, Sungwon, et al.
Pubblicazione: (2024)
Spatiotemporal Skip Guidance for Enhanced Video Diffusion Sampling
di: Hyung, Junha, et al.
Pubblicazione: (2024)
di: Hyung, Junha, et al.
Pubblicazione: (2024)
Infinite-Homography as Robust Conditioning for Camera-Controlled Video Generation
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
di: Kim, Min-Jung, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
di: Yun, Jooyeol, et al.
Pubblicazione: (2025) -
Devil is in the Detail: Towards Injecting Fine Details of Image Prompt in Image Generation via Conflict-free Guidance and Stratified Attention
di: Jo, Kyungmin, et al.
Pubblicazione: (2025) -
Enabling Region-Specific Control via Lassos in Point-Based Colorization
di: Lee, Sanghyeon, et al.
Pubblicazione: (2024) -
Regularized Training with Generated Datasets for Name-Only Transfer of Vision-Language Models
di: Park, Minho, et al.
Pubblicazione: (2024) -
SphereDiff: Tuning-free 360° Static and Dynamic Panorama Generation via Spherical Latent Representation
di: Park, Minho, et al.
Pubblicazione: (2025)