TypeDance: Creating Semantic Typographic Logos from Image through Personalized Generation
Fuente:
arXiv
Guardado en:
| Autores principales: | Xiao, Shishi, Wang, Liangwei, Ma, Xiaojuan, Zeng, Wei |
|---|---|
| Formato: | Preprint |
| Publicado: |
2024
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Persian Typographical Error Type Detection Using Deep Neural Networks on Algorithmically-Generated Misspellings
por: Dehghani, Mohammad, et al.
Publicado: (2023)
por: Dehghani, Mohammad, et al.
Publicado: (2023)
ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion Map
por: Ye, Yilin, et al.
Publicado: (2024)
por: Ye, Yilin, et al.
Publicado: (2024)
Generative AI for Visualization: State of the Art and Future Directions
por: Ye, Yilin, et al.
Publicado: (2024)
por: Ye, Yilin, et al.
Publicado: (2024)
AgentTypo: Adaptive Typographic Prompt Injection Attacks against Black-box Multimodal Agents
por: Li, Yanjie, et al.
Publicado: (2025)
por: Li, Yanjie, et al.
Publicado: (2025)
Reasoning Robustness of LLMs to Adversarial Typographical Errors
por: Gan, Esther, et al.
Publicado: (2024)
por: Gan, Esther, et al.
Publicado: (2024)
MCCoder: Streamlining Motion Control with LLM-Assisted Code Generation and Rigorous Verification
por: Li, Yin, et al.
Publicado: (2024)
por: Li, Yin, et al.
Publicado: (2024)
Logos: An evolvable reasoning engine for rational molecular design
por: Wen, Haibin, et al.
Publicado: (2026)
por: Wen, Haibin, et al.
Publicado: (2026)
From Failure Modes to Reliability Awareness in Generative and Agentic AI System
por: Janet, et al.
Publicado: (2025)
por: Janet, et al.
Publicado: (2025)
ChArtist: Generating Pictorial Charts with Unified Spatial and Subject Control
por: Xiao, Shishi, et al.
Publicado: (2026)
por: Xiao, Shishi, et al.
Publicado: (2026)
Creating an AI Observer: Generative Semantic Workspaces
por: Holur, Pavan, et al.
Publicado: (2024)
por: Holur, Pavan, et al.
Publicado: (2024)
Not What You Asked For: Typographic Attacks in Household Robot Manipulation
por: Iranmanesh, Ali, et al.
Publicado: (2026)
por: Iranmanesh, Ali, et al.
Publicado: (2026)
Dyslexify: A Mechanistic Defense Against Typographic Attacks in CLIP
por: Hufe, Lorenz, et al.
Publicado: (2025)
por: Hufe, Lorenz, et al.
Publicado: (2025)
FigStep: Jailbreaking Large Vision-Language Models via Typographic Visual Prompts
por: Gong, Yichen, et al.
Publicado: (2023)
por: Gong, Yichen, et al.
Publicado: (2023)
Evaluation Metrics for Automated Typographic Poster Generation
por: Rebelo, Sérgio M., et al.
Publicado: (2024)
por: Rebelo, Sérgio M., et al.
Publicado: (2024)
Not All Layers Are Created Equal: Adaptive LoRA Ranks for Personalized Image Generation
por: Shenaj, Donald, et al.
Publicado: (2026)
por: Shenaj, Donald, et al.
Publicado: (2026)
LLM-assisted Labeling Function Generation for Semantic Type Detection
por: Li, Chenjie, et al.
Publicado: (2024)
por: Li, Chenjie, et al.
Publicado: (2024)
May the Dance be with You: Dance Generation Framework for Non-Humanoids
por: Ahn, Hyemin
Publicado: (2024)
por: Ahn, Hyemin
Publicado: (2024)
PersonaBench: Evaluating AI Models on Understanding Personal Information through Accessing (Synthetic) Private User Data
por: Tan, Juntao, et al.
Publicado: (2025)
por: Tan, Juntao, et al.
Publicado: (2025)
SCAM: A Real-World Typographic Robustness Evaluation for Multimodal Foundation Models
por: Westerhoff, Justus, et al.
Publicado: (2025)
por: Westerhoff, Justus, et al.
Publicado: (2025)
Every Image Listens, Every Image Dances: Music-Driven Image Animation
por: Dong, Zhikang, et al.
Publicado: (2025)
por: Dong, Zhikang, et al.
Publicado: (2025)
DanceCrafter: Fine-Grained Text-Driven Controllable Dance Generation via Choreographic Syntax
por: Yuan, Hang, et al.
Publicado: (2026)
por: Yuan, Hang, et al.
Publicado: (2026)
TokenDance: Token-to-Token Music-to-Dance Generation with Bidirectional Mamba
por: Yang, Ziyue, et al.
Publicado: (2026)
por: Yang, Ziyue, et al.
Publicado: (2026)
Dance Any Beat: Blending Beats with Visuals in Dance Video Generation
por: Wang, Xuanchen, et al.
Publicado: (2024)
por: Wang, Xuanchen, et al.
Publicado: (2024)
Evaluating Contextually Personalized Programming Exercises Created with Generative AI
por: Logacheva, Evanfiya, et al.
Publicado: (2024)
por: Logacheva, Evanfiya, et al.
Publicado: (2024)
Using Large Language Models to Create Personalized Networks From Therapy Sessions
por: Ong, Clarissa W., et al.
Publicado: (2025)
por: Ong, Clarissa W., et al.
Publicado: (2025)
A Real-time Endoscopic Image Denoising System
por: Xing, Yu, et al.
Publicado: (2025)
por: Xing, Yu, et al.
Publicado: (2025)
Text-based Person Search in Full Images via Semantic-Driven Proposal Generation
por: Zhang, Shizhou, et al.
Publicado: (2021)
por: Zhang, Shizhou, et al.
Publicado: (2021)
Comprehending Semantic Types in JSON Data with Graph Neural Networks
por: Wei, Shuang, et al.
Publicado: (2023)
por: Wei, Shuang, et al.
Publicado: (2023)
Deciphering Personalization: Towards Fine-Grained Explainability in Natural Language for Personalized Image Generation Models
por: Wang, Haoming, et al.
Publicado: (2025)
por: Wang, Haoming, et al.
Publicado: (2025)
ReactDance: Hierarchical Representation for High-Fidelity and Coherent Long-Form Reactive Dance Generation
por: Lin, Jingzhong, et al.
Publicado: (2025)
por: Lin, Jingzhong, et al.
Publicado: (2025)
Debiased Multimodal Personality Understanding through Dual Causal Intervention
por: Zhu, Yangfu, et al.
Publicado: (2026)
por: Zhu, Yangfu, et al.
Publicado: (2026)
DisCo: Disentangled Control for Realistic Human Dance Generation
por: Wang, Tan, et al.
Publicado: (2023)
por: Wang, Tan, et al.
Publicado: (2023)
Vision Foundation Models as Effective Visual Tokenizers for Autoregressive Image Generation
por: Zheng, Anlin, et al.
Publicado: (2025)
por: Zheng, Anlin, et al.
Publicado: (2025)
Hita: Holistic Tokenizer for Autoregressive Image Generation
por: Zheng, Anlin, et al.
Publicado: (2025)
por: Zheng, Anlin, et al.
Publicado: (2025)
ReCreate: Reasoning and Creating Domain Agents Driven by Experience
por: Hao, Zhezheng, et al.
Publicado: (2026)
por: Hao, Zhezheng, et al.
Publicado: (2026)
Brand Visibility in Packaging: A Deep Learning Approach for Logo Detection, Saliency-Map Prediction, and Logo Placement Analysis
por: Hosseini, Alireza, et al.
Publicado: (2024)
por: Hosseini, Alireza, et al.
Publicado: (2024)
Personalized Image Generation with Large Multimodal Models
por: Xu, Yiyan, et al.
Publicado: (2024)
por: Xu, Yiyan, et al.
Publicado: (2024)
Beyond Dialogue Time: Temporal Semantic Memory for Personalized LLM Agents
por: Su, Miao, et al.
Publicado: (2026)
por: Su, Miao, et al.
Publicado: (2026)
NLGR: Utilizing Neighbor Lists for Generative Rerank in Personalized Recommendation Systems
por: Wang, Shuli, et al.
Publicado: (2025)
por: Wang, Shuli, et al.
Publicado: (2025)
BitDance: Scaling Autoregressive Generative Models with Binary Tokens
por: Ai, Yuang, et al.
Publicado: (2026)
por: Ai, Yuang, et al.
Publicado: (2026)
Ejemplares similares
-
Persian Typographical Error Type Detection Using Deep Neural Networks on Algorithmically-Generated Misspellings
por: Dehghani, Mohammad, et al.
Publicado: (2023) -
ModalChorus: Visual Probing and Alignment of Multi-modal Embeddings via Modal Fusion Map
por: Ye, Yilin, et al.
Publicado: (2024) -
Generative AI for Visualization: State of the Art and Future Directions
por: Ye, Yilin, et al.
Publicado: (2024) -
AgentTypo: Adaptive Typographic Prompt Injection Attacks against Black-box Multimodal Agents
por: Li, Yanjie, et al.
Publicado: (2025) -
Reasoning Robustness of LLMs to Adversarial Typographical Errors
por: Gan, Esther, et al.
Publicado: (2024)