AutomaTikZ: Text-Guided Synthesis of Scientific Vector Graphics with TikZ
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Belouadi, Jonas, Lauscher, Anne, Eger, Steffen |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2023
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DeTikZify: Synthesizing Graphics Programs for Scientific Figures and Sketches with TikZ
von: Belouadi, Jonas, et al.
Veröffentlicht: (2024)
von: Belouadi, Jonas, et al.
Veröffentlicht: (2024)
TikZilla: Scaling Text-to-TikZ with High-Quality Data and Reinforcement Learning
von: Greisinger, Christian, et al.
Veröffentlicht: (2026)
von: Greisinger, Christian, et al.
Veröffentlicht: (2026)
TikZero: Zero-Shot Text-Guided Graphics Program Synthesis
von: Belouadi, Jonas, et al.
Veröffentlicht: (2025)
von: Belouadi, Jonas, et al.
Veröffentlicht: (2025)
Text to Automata Diagrams: Comparing TikZ Code Generation with Direct Image Synthesis
von: Young, Ethan, et al.
Veröffentlicht: (2026)
von: Young, Ethan, et al.
Veröffentlicht: (2026)
ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation?
von: Zhang, Leixin, et al.
Veröffentlicht: (2024)
von: Zhang, Leixin, et al.
Veröffentlicht: (2024)
MultiMat: Multimodal Program Synthesis for Procedural Materials using Large Multimodal Models
von: Belouadi, Jonas, et al.
Veröffentlicht: (2025)
von: Belouadi, Jonas, et al.
Veröffentlicht: (2025)
ByGPT5: End-to-End Style-conditioned Poetry Generation with Token-free Language Models
von: Belouadi, Jonas, et al.
Veröffentlicht: (2022)
von: Belouadi, Jonas, et al.
Veröffentlicht: (2022)
USCORE: An Effective Approach to Fully Unsupervised Evaluation Metrics for Machine Translation
von: Belouadi, Jonas, et al.
Veröffentlicht: (2022)
von: Belouadi, Jonas, et al.
Veröffentlicht: (2022)
NLLG Quarterly arXiv Report 09/24: What are the most influential current AI Papers?
von: Leiter, Christoph, et al.
Veröffentlicht: (2024)
von: Leiter, Christoph, et al.
Veröffentlicht: (2024)
TempViz: On the Evaluation of Temporal Knowledge in Text-to-Image Models
von: Holtermann, Carolin, et al.
Veröffentlicht: (2026)
von: Holtermann, Carolin, et al.
Veröffentlicht: (2026)
StrokeNUWA: Tokenizing Strokes for Vector Graphic Synthesis
von: Tang, Zecheng, et al.
Veröffentlicht: (2024)
von: Tang, Zecheng, et al.
Veröffentlicht: (2024)
Illustrating Finite Automata with Grail+ and TikZ
von: May, Alastair, et al.
Veröffentlicht: (2024)
von: May, Alastair, et al.
Veröffentlicht: (2024)
StarVector: Generating Scalable Vector Graphics Code from Images and Text
von: Rodriguez, Juan A., et al.
Veröffentlicht: (2023)
von: Rodriguez, Juan A., et al.
Veröffentlicht: (2023)
CROC: Evaluating and Training T2I Metrics with Pseudo- and Human-Labeled Contrastive Robustness Checks
von: Leiter, Christoph, et al.
Veröffentlicht: (2025)
von: Leiter, Christoph, et al.
Veröffentlicht: (2025)
A New Hybrid Intelligent Approach for Multimodal Detection of Suspected Disinformation on TikTok
von: Guerrero-Sosa, Jared D. T., et al.
Veröffentlicht: (2025)
von: Guerrero-Sosa, Jared D. T., et al.
Veröffentlicht: (2025)
Transforming Science with Large Language Models: A Survey on AI-assisted Scientific Discovery, Experimentation, Content Generation, and Evaluation
von: Eger, Steffen, et al.
Veröffentlicht: (2025)
von: Eger, Steffen, et al.
Veröffentlicht: (2025)
Visually Descriptive Language Model for Vector Graphics Reasoning
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2024)
von: Wang, Zhenhailong, et al.
Veröffentlicht: (2024)
Centurio: On Drivers of Multilingual Ability of Large Vision-Language Model
von: Geigle, Gregor, et al.
Veröffentlicht: (2025)
von: Geigle, Gregor, et al.
Veröffentlicht: (2025)
TikGuard: A Deep Learning Transformer-Based Solution for Detecting Unsuitable TikTok Content for Kids
von: Balat, Mazen, et al.
Veröffentlicht: (2024)
von: Balat, Mazen, et al.
Veröffentlicht: (2024)
TikArt: Stabilizing Aperture-Guided Fine-Grained Visual Reasoning with Reinforcement Learning
von: Ding, Hao, et al.
Veröffentlicht: (2026)
von: Ding, Hao, et al.
Veröffentlicht: (2026)
Why do LLaVA Vision-Language Models Reply to Images in English?
von: Hinck, Musashi, et al.
Veröffentlicht: (2024)
von: Hinck, Musashi, et al.
Veröffentlicht: (2024)
PEFT A2Z: Parameter-Efficient Fine-Tuning Survey for Large Language and Vision Models
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2025)
von: Prottasha, Nusrat Jahan, et al.
Veröffentlicht: (2025)
Scaling Concept With Text-Guided Diffusion Models
von: Huang, Chao, et al.
Veröffentlicht: (2024)
von: Huang, Chao, et al.
Veröffentlicht: (2024)
VectorPainter: Advanced Stylized Vector Graphics Synthesis Using Stroke-Style Priors
von: Hu, Juncheng, et al.
Veröffentlicht: (2024)
von: Hu, Juncheng, et al.
Veröffentlicht: (2024)
SuperSVG: Superpixel-based Scalable Vector Graphics Synthesis
von: Hu, Teng, et al.
Veröffentlicht: (2024)
von: Hu, Teng, et al.
Veröffentlicht: (2024)
Re-Thinking Inverse Graphics With Large Language Models
von: Kulits, Peter, et al.
Veröffentlicht: (2024)
von: Kulits, Peter, et al.
Veröffentlicht: (2024)
Visually Guided Generative Text-Layout Pre-training for Document Intelligence
von: Mao, Zhiming, et al.
Veröffentlicht: (2024)
von: Mao, Zhiming, et al.
Veröffentlicht: (2024)
AI Based Font Pair Suggestion Modelling For Graphic Design
von: Singh, Aryan, et al.
Veröffentlicht: (2025)
von: Singh, Aryan, et al.
Veröffentlicht: (2025)
Leveraging Large Language Models for Scalable Vector Graphics-Driven Image Understanding
von: Cai, Mu, et al.
Veröffentlicht: (2023)
von: Cai, Mu, et al.
Veröffentlicht: (2023)
Text Role Classification in Scientific Charts Using Multimodal Transformers
von: Kim, Hye Jin, et al.
Veröffentlicht: (2024)
von: Kim, Hye Jin, et al.
Veröffentlicht: (2024)
Beyond Image-Text Matching: Verb Understanding in Multimodal Transformers Using Guided Masking
von: Beňová, Ivana, et al.
Veröffentlicht: (2024)
von: Beňová, Ivana, et al.
Veröffentlicht: (2024)
3DFroMLLM: 3D Prototype Generation only from Pretrained Multimodal LLMs
von: Ahmed, Noor, et al.
Veröffentlicht: (2025)
von: Ahmed, Noor, et al.
Veröffentlicht: (2025)
Taming the Tri-Space Tension: ARC-Guided Hallucination Modeling and Control for Text-to-Image Generation
von: Yang, Jianjiang, et al.
Veröffentlicht: (2025)
von: Yang, Jianjiang, et al.
Veröffentlicht: (2025)
Scaling Text-Rich Image Understanding via Code-Guided Synthetic Multimodal Data Generation
von: Yang, Yue, et al.
Veröffentlicht: (2025)
von: Yang, Yue, et al.
Veröffentlicht: (2025)
DiffChat: Learning to Chat with Text-to-Image Synthesis Models for Interactive Image Creation
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
von: Wang, Jiapeng, et al.
Veröffentlicht: (2024)
What Is That Talk About? A Video-to-Text Summarization Dataset for Scientific Presentations
von: Liu, Dongqi, et al.
Veröffentlicht: (2025)
von: Liu, Dongqi, et al.
Veröffentlicht: (2025)
Prototypicality Bias Reveals Blindspots in Multimodal Evaluation Metrics
von: Roy, Subhadeep, et al.
Veröffentlicht: (2026)
von: Roy, Subhadeep, et al.
Veröffentlicht: (2026)
Vector Prism: Animating Vector Graphics by Stratifying Semantic Structure
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
von: Yun, Jooyeol, et al.
Veröffentlicht: (2025)
A Generalist Model for Diverse Text-Guided Medical Image Synthesis
von: Cho, Joseph, et al.
Veröffentlicht: (2024)
von: Cho, Joseph, et al.
Veröffentlicht: (2024)
Graphic Design with Large Multimodal Model
von: Cheng, Yutao, et al.
Veröffentlicht: (2024)
von: Cheng, Yutao, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DeTikZify: Synthesizing Graphics Programs for Scientific Figures and Sketches with TikZ
von: Belouadi, Jonas, et al.
Veröffentlicht: (2024) -
TikZilla: Scaling Text-to-TikZ with High-Quality Data and Reinforcement Learning
von: Greisinger, Christian, et al.
Veröffentlicht: (2026) -
TikZero: Zero-Shot Text-Guided Graphics Program Synthesis
von: Belouadi, Jonas, et al.
Veröffentlicht: (2025) -
Text to Automata Diagrams: Comparing TikZ Code Generation with Direct Image Synthesis
von: Young, Ethan, et al.
Veröffentlicht: (2026) -
ScImage: How Good Are Multimodal Large Language Models at Scientific Text-to-Image Generation?
von: Zhang, Leixin, et al.
Veröffentlicht: (2024)