Bridging Text and Image for Artist Style Transfer via Contrastive Learning
Fuente:
arXiv
Saved in:
| Main Authors: | Liu, Zhi-Song, Wang, Li-Wen, Xiao, Jun, Kalogeiton, Vicky |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Name Your Style: An Arbitrary Artist-aware Image Style Transfer
by: Liu, Zhi-Song, et al.
Published: (2022)
by: Liu, Zhi-Song, et al.
Published: (2022)
Iris Style Transfer: Enhancing Iris Recognition with Style Features and Privacy Preservation through Neural Style Transfer
by: Wang, Mengdi, et al.
Published: (2025)
by: Wang, Mengdi, et al.
Published: (2025)
Semantic Draw Engineering for Text-to-Image Creation
by: Li, Yang, et al.
Published: (2023)
by: Li, Yang, et al.
Published: (2023)
Supervised Contrastive Learning for Ordinal Engagement Measurement
by: Safa, Sadaf, et al.
Published: (2025)
by: Safa, Sadaf, et al.
Published: (2025)
Text-to-Image Representativity Fairness Evaluation Framework
by: Yamani, Asma, et al.
Published: (2024)
by: Yamani, Asma, et al.
Published: (2024)
EIT-1M: One Million EEG-Image-Text Pairs for Human Visual-textual Recognition and More
by: Zheng, Xu, et al.
Published: (2024)
by: Zheng, Xu, et al.
Published: (2024)
Intuitions of Machine Learning Researchers about Transfer Learning for Medical Image Classification
by: Lu, Yucheng, et al.
Published: (2025)
by: Lu, Yucheng, et al.
Published: (2025)
See Through Their Minds: Learning Transferable Neural Representation from Cross-Subject fMRI
by: Liu, Yulong, et al.
Published: (2024)
by: Liu, Yulong, et al.
Published: (2024)
POET: Supporting Prompting Creativity and Personalization with Automated Expansion of Text-to-Image Generation
by: Han, Evans Xu, et al.
Published: (2025)
by: Han, Evans Xu, et al.
Published: (2025)
SketchFlex: Facilitating Spatial-Semantic Coherence in Text-to-Image Generation with Region-Based Sketches
by: Lin, Haichuan, et al.
Published: (2025)
by: Lin, Haichuan, et al.
Published: (2025)
NarrativeBridge: Enhancing Video Captioning with Causal-Temporal Narrative
by: Nadeem, Asmar, et al.
Published: (2024)
by: Nadeem, Asmar, et al.
Published: (2024)
Towards Geographic Inclusion in the Evaluation of Text-to-Image Models
by: Hall, Melissa, et al.
Published: (2024)
by: Hall, Melissa, et al.
Published: (2024)
YOLOA: Real-Time Affordance Detection via LLM Adapter
by: Ji, Yuqi, et al.
Published: (2025)
by: Ji, Yuqi, et al.
Published: (2025)
Dream360: Diverse and Immersive Outdoor Virtual Scene Creation via Transformer-Based 360 Image Outpainting
by: Ai, Hao, et al.
Published: (2024)
by: Ai, Hao, et al.
Published: (2024)
GUI-R1 : A Generalist R1-Style Vision-Language Action Model For GUI Agents
by: Luo, Run, et al.
Published: (2025)
by: Luo, Run, et al.
Published: (2025)
SwipeGANSpace: Swipe-to-Compare Image Generation via Efficient Latent Space Exploration
by: Nakashima, Yuto, et al.
Published: (2024)
by: Nakashima, Yuto, et al.
Published: (2024)
CinePreGen: Camera Controllable Video Previsualization via Engine-powered Diffusion
by: Chen, Yiran, et al.
Published: (2024)
by: Chen, Yiran, et al.
Published: (2024)
GesGPT: Speech Gesture Synthesis With Text Parsing from ChatGPT
by: Gao, Nan, et al.
Published: (2023)
by: Gao, Nan, et al.
Published: (2023)
InterAnimate: Taming Region-aware Diffusion Model for Realistic Human Interaction Animation
by: Lin, Yukang, et al.
Published: (2025)
by: Lin, Yukang, et al.
Published: (2025)
L-WISE: Boosting Human Visual Category Learning Through Model-Based Image Selection and Enhancement
by: Talbot, Morgan B., et al.
Published: (2024)
by: Talbot, Morgan B., et al.
Published: (2024)
GLIMPSE : Real-Time Text Recognition and Contextual Understanding for VQA in Wearables
by: Ramachandran, Akhil, et al.
Published: (2026)
by: Ramachandran, Akhil, et al.
Published: (2026)
ChatStitch: Visualizing Through Structures via Surround-View Unsupervised Deep Image Stitching with Collaborative LLM-Agents
by: Liang, Hao, et al.
Published: (2025)
by: Liang, Hao, et al.
Published: (2025)
VizDefender: Unmasking Visualization Tampering through Proactive Localization and Intent Inference
by: Song, Sicheng, et al.
Published: (2025)
by: Song, Sicheng, et al.
Published: (2025)
Deep Learning in Mild Cognitive Impairment Diagnosis using Eye Movements and Image Content in Visual Memory Tasks
by: Rocha, Tomás Silva Santos, et al.
Published: (2025)
by: Rocha, Tomás Silva Santos, et al.
Published: (2025)
A Monocular SLAM-based Multi-User Positioning System with Image Occlusion in Augmented Reality
by: Lien, Wei-Hsiang, et al.
Published: (2024)
by: Lien, Wei-Hsiang, et al.
Published: (2024)
Text-to-Image Generation for Vocabulary Learning Using the Keyword Method
by: Attygalle, Nuwan T., et al.
Published: (2025)
by: Attygalle, Nuwan T., et al.
Published: (2025)
Using Text-to-Image Generation for Architectural Design Ideation
by: Paananen, Ville, et al.
Published: (2023)
by: Paananen, Ville, et al.
Published: (2023)
Alt4Blind: A User Interface to Simplify Charts Alt-Text Creation
by: Moured, Omar, et al.
Published: (2024)
by: Moured, Omar, et al.
Published: (2024)
BLK-Assist: A Methodological Framework for Artist-Led Co-Creation with Generative AI Models
by: Grimes, Daniel, et al.
Published: (2026)
by: Grimes, Daniel, et al.
Published: (2026)
Automated Image-Based Identification and Consistent Classification of Fire Patterns with Quantitative Shape Analysis and Spatial Location Identification
by: Liu, Pengkun, et al.
Published: (2024)
by: Liu, Pengkun, et al.
Published: (2024)
Breaking Coordinate Overfitting: Geometry-Aware WiFi Sensing for Cross-Layout 3D Pose Estimation
by: Jia, Songming, et al.
Published: (2026)
by: Jia, Songming, et al.
Published: (2026)
DiffGaze: A Diffusion Model for Continuous Gaze Sequence Generation on 360° Images
by: Jiao, Chuhan, et al.
Published: (2024)
by: Jiao, Chuhan, et al.
Published: (2024)
MedFoundationHub: A Lightweight and Secure Toolkit for Deploying Medical Vision Language Foundation Models
by: Li, Xiao, et al.
Published: (2025)
by: Li, Xiao, et al.
Published: (2025)
VideoMix: Aggregating How-To Videos for Task-Oriented Learning
by: Yang, Saelyne, et al.
Published: (2025)
by: Yang, Saelyne, et al.
Published: (2025)
Efficient Listener: Dyadic Facial Motion Synthesis via Action Diffusion
by: Wang, Zesheng, et al.
Published: (2025)
by: Wang, Zesheng, et al.
Published: (2025)
Visual Neural Decoding via Improved Visual-EEG Semantic Consistency
by: Chen, Hongzhou, et al.
Published: (2024)
by: Chen, Hongzhou, et al.
Published: (2024)
QuizRank: Picking Images by Quizzing VLMs
by: Ji, Tenghao, et al.
Published: (2025)
by: Ji, Tenghao, et al.
Published: (2025)
UniHands: Unifying Various Wild-Collected Keypoints for Personalized Hand Reconstruction
by: Zhang, Menghe, et al.
Published: (2024)
by: Zhang, Menghe, et al.
Published: (2024)
SCHEMA for Gemini 3 Pro Image: A Structured Methodology for Controlled AI Image Generation on Google's Native Multimodal Model
by: Cazzaniga, Luca
Published: (2026)
by: Cazzaniga, Luca
Published: (2026)
Transfer Learning-based Real-time Handgun Detection
by: Elmir, Youssef
Published: (2023)
by: Elmir, Youssef
Published: (2023)
Similar Items
-
Name Your Style: An Arbitrary Artist-aware Image Style Transfer
by: Liu, Zhi-Song, et al.
Published: (2022) -
Iris Style Transfer: Enhancing Iris Recognition with Style Features and Privacy Preservation through Neural Style Transfer
by: Wang, Mengdi, et al.
Published: (2025) -
Semantic Draw Engineering for Text-to-Image Creation
by: Li, Yang, et al.
Published: (2023) -
Supervised Contrastive Learning for Ordinal Engagement Measurement
by: Safa, Sadaf, et al.
Published: (2025) -
Text-to-Image Representativity Fairness Evaluation Framework
by: Yamani, Asma, et al.
Published: (2024)