Human Aesthetic Preference-Based Large Text-to-Image Model Personalization: Kandinsky Generation as an Example
Fuente:
arXiv
Salvato in:
| Autori principali: | Zhou, Aven-Le, Wang, Yu-Ao, Wu, Wei, Zhang, Kang |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Archiving Body Movements: Collective Generation of Chinese Calligraphy
di: Zhou, Aven Le, et al.
Pubblicazione: (2023)
di: Zhou, Aven Le, et al.
Pubblicazione: (2023)
Node-Based Editing for Multimodal Generation of Text, Audio, Image, and Video
di: Kyaw, Alexander Htet, et al.
Pubblicazione: (2025)
di: Kyaw, Alexander Htet, et al.
Pubblicazione: (2025)
Steering Large Text-to-Image Model for Abstract Art Synthesis: Preference-based Prompt Optimization and Visualization
di: Zhou, Aven-Le, et al.
Pubblicazione: (2024)
di: Zhou, Aven-Le, et al.
Pubblicazione: (2024)
DreamLLM-3D: Affective Dream Reliving using Large Language Model and 3D Generative AI
di: Liu, Pinyao, et al.
Pubblicazione: (2025)
di: Liu, Pinyao, et al.
Pubblicazione: (2025)
Simulacra Naturae: Generative Ecosystem driven by Agent-Based Simulations and Brain Organoid Collective Intelligence
di: Manoudaki, Nefeli, et al.
Pubblicazione: (2025)
di: Manoudaki, Nefeli, et al.
Pubblicazione: (2025)
A Multimedia Analytics Model for the Foundation Model Era
di: Worring, Marcel, et al.
Pubblicazione: (2025)
di: Worring, Marcel, et al.
Pubblicazione: (2025)
A Rhetorical Relations-Based Framework for Tailored Multimedia Document Summarization
di: Maredj, Azze-Eddine, et al.
Pubblicazione: (2024)
di: Maredj, Azze-Eddine, et al.
Pubblicazione: (2024)
Vidmento: Creating Video Stories Through Context-Aware Expansion With Generative Video
di: Yeh, Catherine, et al.
Pubblicazione: (2026)
di: Yeh, Catherine, et al.
Pubblicazione: (2026)
Freetalker: Controllable Speech and Text-Driven Gesture Generation Based on Diffusion Models for Enhanced Speaker Naturalness
di: Yang, Sicheng, et al.
Pubblicazione: (2024)
di: Yang, Sicheng, et al.
Pubblicazione: (2024)
Signals of Provenance: Practices & Challenges of Navigating Indicators in AI-Generated Media for Sighted and Blind Individuals
di: Ide, Ayae, et al.
Pubblicazione: (2025)
di: Ide, Ayae, et al.
Pubblicazione: (2025)
IntentVLM: Open-Vocabulary Intention Recognition through Forward-Inverse Modeling with Video-Language Models
di: Rahimi, Hamed, et al.
Pubblicazione: (2026)
di: Rahimi, Hamed, et al.
Pubblicazione: (2026)
Dynamic and Super-Personalized Media Ecosystem Driven by Generative AI: Unpredictable Plays Never Repeating The Same
di: Ahn, Sungjun, et al.
Pubblicazione: (2024)
di: Ahn, Sungjun, et al.
Pubblicazione: (2024)
MetaDesigner: Advancing Artistic Typography Through AI-Driven, User-Centric, and Multilingual WordArt Synthesis
di: He, Jun-Yan, et al.
Pubblicazione: (2024)
di: He, Jun-Yan, et al.
Pubblicazione: (2024)
SimInterview: Transforming Business Education through Large Language Model-Based Simulated Multilingual Interview Training System
di: Nguyen, Truong Thanh Hung, et al.
Pubblicazione: (2025)
di: Nguyen, Truong Thanh Hung, et al.
Pubblicazione: (2025)
Counterfactual Reasoning Using Predicted Latent Personality Dimensions for Optimizing Persuasion Outcome
di: Zeng, Donghuo, et al.
Pubblicazione: (2024)
di: Zeng, Donghuo, et al.
Pubblicazione: (2024)
Focus360: Guiding User Attention in Immersive Videos for VR
di: Silva, Paulo Vitor S., et al.
Pubblicazione: (2026)
di: Silva, Paulo Vitor S., et al.
Pubblicazione: (2026)
Self-supervised Spatio-Temporal Graph Mask-Passing Attention Network for Perceptual Importance Prediction of Multi-point Tactility
di: He, Dazhong, et al.
Pubblicazione: (2024)
di: He, Dazhong, et al.
Pubblicazione: (2024)
Livia: An Emotion-Aware AR Companion Powered by Modular AI Agents and Progressive Memory Compression
di: Xi, Rui, et al.
Pubblicazione: (2025)
di: Xi, Rui, et al.
Pubblicazione: (2025)
Through the Looking-Glass: AI-Mediated Video Communication Reduces Interpersonal Trust and Confidence in Judgments
di: Fernández, Nelson Navajas, et al.
Pubblicazione: (2026)
di: Fernández, Nelson Navajas, et al.
Pubblicazione: (2026)
An Empirical Evaluation of AI-Powered Non-Player Characters' Perceived Realism and Performance in Virtual Reality Environments
di: Korkiakoski, Mikko, et al.
Pubblicazione: (2025)
di: Korkiakoski, Mikko, et al.
Pubblicazione: (2025)
AMEX: Android Multi-annotation Expo Dataset for Mobile GUI Agents
di: Chai, Yuxiang, et al.
Pubblicazione: (2024)
di: Chai, Yuxiang, et al.
Pubblicazione: (2024)
Computational Analysis of Stress, Depression and Engagement in Mental Health: A Survey
di: Kumar, Puneet, et al.
Pubblicazione: (2024)
di: Kumar, Puneet, et al.
Pubblicazione: (2024)
Virtual Reality and Artificial Intelligence as Psychological Countermeasures in Space and Other Isolated and Confined Environments: A Scoping Review
di: Sharp, Jennifer, et al.
Pubblicazione: (2025)
di: Sharp, Jennifer, et al.
Pubblicazione: (2025)
MuDoC: An Interactive Multimodal Document-grounded Conversational AI System
di: Taneja, Karan, et al.
Pubblicazione: (2025)
di: Taneja, Karan, et al.
Pubblicazione: (2025)
Chain-of-Modality: Learning Manipulation Programs from Multimodal Human Videos with Vision-Language-Models
di: Wang, Chen, et al.
Pubblicazione: (2025)
di: Wang, Chen, et al.
Pubblicazione: (2025)
Secure & Personalized Music-to-Video Generation via CHARCHA
di: Agarwal, Mehul, et al.
Pubblicazione: (2025)
di: Agarwal, Mehul, et al.
Pubblicazione: (2025)
DailyLLM: Context-Aware Activity Log Generation Using Multi-Modal Sensors and LLMs
di: Tian, Ye, et al.
Pubblicazione: (2025)
di: Tian, Ye, et al.
Pubblicazione: (2025)
Human-Machine Ritual: Synergic Performance through Real-Time Motion Recognition
di: Cai, Zhuodi, et al.
Pubblicazione: (2025)
di: Cai, Zhuodi, et al.
Pubblicazione: (2025)
Chat with AI: The Surprising Turn of Real-time Video Communication from Human to AI
di: Wu, Jiangkai, et al.
Pubblicazione: (2025)
di: Wu, Jiangkai, et al.
Pubblicazione: (2025)
Multimodal Infusion Tuning for Large Models
di: Sun, Hao, et al.
Pubblicazione: (2024)
di: Sun, Hao, et al.
Pubblicazione: (2024)
FeedQUAC: Quick Unobtrusive AI-Generated Commentary
di: Long, Tao, et al.
Pubblicazione: (2025)
di: Long, Tao, et al.
Pubblicazione: (2025)
MetaDecorator: Generating Immersive Virtual Tours through Multimodality
di: Xie, Shuang, et al.
Pubblicazione: (2025)
di: Xie, Shuang, et al.
Pubblicazione: (2025)
Emotion-Driven Personalized Recommendation for AI-Generated Content Using Multi-Modal Sentiment and Intent Analysis
di: Hu, Zheqi, et al.
Pubblicazione: (2025)
di: Hu, Zheqi, et al.
Pubblicazione: (2025)
Proceedings of The third international workshop on eXplainable AI for the Arts (XAIxArts)
di: Ford, Corey, et al.
Pubblicazione: (2025)
di: Ford, Corey, et al.
Pubblicazione: (2025)
MetaBGM: Dynamic Soundtrack Transformation For Continuous Multi-Scene Experiences With Ambient Awareness And Personalization
di: Liu, Haoxuan, et al.
Pubblicazione: (2024)
di: Liu, Haoxuan, et al.
Pubblicazione: (2024)
Human-Data Interaction, Exploration, and Visualization in the AI Era: Challenges and Opportunities
di: Fekete, Jean-Daniel, et al.
Pubblicazione: (2026)
di: Fekete, Jean-Daniel, et al.
Pubblicazione: (2026)
SkinGEN: an Explainable Dermatology Diagnosis-to-Generation Framework with Interactive Vision-Language Models
di: Lin, Bo, et al.
Pubblicazione: (2024)
di: Lin, Bo, et al.
Pubblicazione: (2024)
Coral Model Generation from Single Images for Virtual Reality Applications
di: Fu, Jie, et al.
Pubblicazione: (2024)
di: Fu, Jie, et al.
Pubblicazione: (2024)
DiffMesh: A Motion-aware Diffusion Framework for Human Mesh Recovery from Videos
di: Zheng, Ce, et al.
Pubblicazione: (2023)
di: Zheng, Ce, et al.
Pubblicazione: (2023)
Inkspire: Supporting Design Exploration with Generative AI through Analogical Sketching
di: Lin, David Chuan-En, et al.
Pubblicazione: (2025)
di: Lin, David Chuan-En, et al.
Pubblicazione: (2025)
Documenti analoghi
-
Archiving Body Movements: Collective Generation of Chinese Calligraphy
di: Zhou, Aven Le, et al.
Pubblicazione: (2023) -
Node-Based Editing for Multimodal Generation of Text, Audio, Image, and Video
di: Kyaw, Alexander Htet, et al.
Pubblicazione: (2025) -
Steering Large Text-to-Image Model for Abstract Art Synthesis: Preference-based Prompt Optimization and Visualization
di: Zhou, Aven-Le, et al.
Pubblicazione: (2024) -
DreamLLM-3D: Affective Dream Reliving using Large Language Model and 3D Generative AI
di: Liu, Pinyao, et al.
Pubblicazione: (2025) -
Simulacra Naturae: Generative Ecosystem driven by Agent-Based Simulations and Brain Organoid Collective Intelligence
di: Manoudaki, Nefeli, et al.
Pubblicazione: (2025)