DuetSVG: Unified Multimodal SVG Generation with Internal Visual Guidance
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Peiying, Zhao, Nanxuan, Fisher, Matthew, Xu, Yiran, Liao, Jing, Liu, Difan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoboSVG: A Unified Framework for Interactive SVG Generation with Multi-modal Guidance
by: Wang, Jiuniu, et al.
Published: (2025)
by: Wang, Jiuniu, et al.
Published: (2025)
InternSVG: Towards Unified SVG Tasks with Multimodal Large Language Models
by: Wang, Haomin, et al.
Published: (2025)
by: Wang, Haomin, et al.
Published: (2025)
Text-to-Vector Generation with Neural Path Representation
by: Zhang, Peiying, et al.
Published: (2024)
by: Zhang, Peiying, et al.
Published: (2024)
WildSVG: Towards Reliable SVG Generation Under Real-Word Conditions
by: Terral, Marco, et al.
Published: (2026)
by: Terral, Marco, et al.
Published: (2026)
Style Customization of Text-to-Vector Generation with Image Diffusion Priors
by: Zhang, Peiying, et al.
Published: (2025)
by: Zhang, Peiying, et al.
Published: (2025)
LiveSVG: Zero-Shot SVG Animation via Video Generation
by: Levy, Matan, et al.
Published: (2026)
by: Levy, Matan, et al.
Published: (2026)
IntroSVG: Learning from Rendering Feedback for Text-to-SVG Generation via an Introspective Generator-Critic Framework
by: Wang, Feiyu, et al.
Published: (2026)
by: Wang, Feiyu, et al.
Published: (2026)
OmniSVG: A Unified Scalable Vector Graphics Generation Model
by: Yang, Yiying, et al.
Published: (2025)
by: Yang, Yiying, et al.
Published: (2025)
T-SVG: Text-Driven Stereoscopic Video Generation
by: Jin, Qiao, et al.
Published: (2024)
by: Jin, Qiao, et al.
Published: (2024)
SVGDreamer: Text Guided SVG Generation with Diffusion Model
by: Xing, Ximing, et al.
Published: (2023)
by: Xing, Ximing, et al.
Published: (2023)
VCode: a Multimodal Coding Benchmark with SVG as Symbolic Visual Representation
by: Lin, Kevin Qinghong, et al.
Published: (2025)
by: Lin, Kevin Qinghong, et al.
Published: (2025)
SVGDreamer++: Advancing Editability and Diversity in Text-Guided SVG Generation
by: Xing, Ximing, et al.
Published: (2024)
by: Xing, Ximing, et al.
Published: (2024)
Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning
by: Xing, Ximing, et al.
Published: (2025)
by: Xing, Ximing, et al.
Published: (2025)
Chat2SVG: Vector Graphics Generation with Large Language Models and Image Diffusion Models
by: Wu, Ronghuan, et al.
Published: (2024)
by: Wu, Ronghuan, et al.
Published: (2024)
SVGThinker: Instruction-Aligned and Reasoning-Driven Text-to-SVG Generation
by: Chen, Hanqi, et al.
Published: (2025)
by: Chen, Hanqi, et al.
Published: (2025)
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
by: Banerjee, Ayan, et al.
Published: (2024)
by: Banerjee, Ayan, et al.
Published: (2024)
NeuralSVG: An Implicit Representation for Text-to-Vector Generation
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
SVGenius: Benchmarking LLMs in SVG Understanding, Editing and Generation
by: Chen, Siqi, et al.
Published: (2025)
by: Chen, Siqi, et al.
Published: (2025)
UniSVG: A Unified Dataset for Vector Graphic Understanding and Generation with Multimodal Large Language Models
by: Li, Jinke, et al.
Published: (2025)
by: Li, Jinke, et al.
Published: (2025)
SVG: 3D Stereoscopic Video Generation via Denoising Frame Matrix
by: Dai, Peng, et al.
Published: (2024)
by: Dai, Peng, et al.
Published: (2024)
NIVeL: Neural Implicit Vector Layers for Text-to-Vector Generation
by: Thamizharasan, Vikas, et al.
Published: (2024)
by: Thamizharasan, Vikas, et al.
Published: (2024)
SuperSVG: Superpixel-based Scalable Vector Graphics Synthesis
by: Hu, Teng, et al.
Published: (2024)
by: Hu, Teng, et al.
Published: (2024)
SVGauge: Towards Human-Aligned Evaluation for SVG Generation
by: Zini, Leonardo, et al.
Published: (2025)
by: Zini, Leonardo, et al.
Published: (2025)
AmodalSVG: Amodal Image Vectorization via Semantic Layer Peeling
by: Hu, Juncheng, et al.
Published: (2026)
by: Hu, Juncheng, et al.
Published: (2026)
SVG-IR: Spatially-Varying Gaussian Splatting for Inverse Rendering
by: Sun, Hanxiao, et al.
Published: (2025)
by: Sun, Hanxiao, et al.
Published: (2025)
Socratic Chart: Cooperating Multiple Agents for Robust SVG Chart Understanding
by: Ji, Yuyang, et al.
Published: (2025)
by: Ji, Yuyang, et al.
Published: (2025)
VFIG: Vectorizing Complex Figures in SVG with Vision-Language Models
by: He, Qijia, et al.
Published: (2026)
by: He, Qijia, et al.
Published: (2026)
From Tokens to Numbers: Continuous Number Modeling for SVG Generation
by: Ogezi, Michael, et al.
Published: (2026)
by: Ogezi, Michael, et al.
Published: (2026)
SVG-EAR: Parameter-Free Linear Compensation for Sparse Video Generation via Error-aware Routing
by: Zhou, Xuanyi, et al.
Published: (2026)
by: Zhou, Xuanyi, et al.
Published: (2026)
LayerTracer: Cognitive-Aligned Layered SVG Synthesis via Diffusion Transformer
by: Song, Yiren, et al.
Published: (2025)
by: Song, Yiren, et al.
Published: (2025)
Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning
by: Wang, Haomin, et al.
Published: (2026)
by: Wang, Haomin, et al.
Published: (2026)
VectorGym: A Multitask Benchmark for SVG Code Generation, Sketching, and Editing
by: Rodriguez, Juan, et al.
Published: (2026)
by: Rodriguez, Juan, et al.
Published: (2026)
AnchorFlow: Editable SVG Reconstruction via Sparse Anchor Point Fields
by: Jiang, Mengnan, et al.
Published: (2026)
by: Jiang, Mengnan, et al.
Published: (2026)
SVG-Head: Hybrid Surface-Volumetric Gaussians for High-Fidelity Head Reconstruction and Real-Time Editing
by: Sun, Heyi, et al.
Published: (2025)
by: Sun, Heyi, et al.
Published: (2025)
SVGEditBench: A Benchmark Dataset for Quantitative Assessment of LLM's SVG Editing Capabilities
by: Nishina, Kunato, et al.
Published: (2024)
by: Nishina, Kunato, et al.
Published: (2024)
SVGBuilder: Component-Based Colored SVG Generation with Text-Guided Autoregressive Transformers
by: Chen, Zehao, et al.
Published: (2024)
by: Chen, Zehao, et al.
Published: (2024)
SVG-T2I: Scaling Up Text-to-Image Latent Diffusion Model Without Variational Autoencoder
by: Shi, Minglei, et al.
Published: (2025)
by: Shi, Minglei, et al.
Published: (2025)
Semantic Document Derendering: SVG Reconstruction via Vision-Language Modeling
by: Hazimeh, Adam, et al.
Published: (2025)
by: Hazimeh, Adam, et al.
Published: (2025)
Rethinking Layered Graphic Design Generation with a Top-Down Approach
by: Chen, Jingye, et al.
Published: (2025)
by: Chen, Jingye, et al.
Published: (2025)
Illustrator's Depth: Monocular Layer Index Prediction for Image Decomposition
by: Maruani, Nissim, et al.
Published: (2025)
by: Maruani, Nissim, et al.
Published: (2025)
Similar Items
-
RoboSVG: A Unified Framework for Interactive SVG Generation with Multi-modal Guidance
by: Wang, Jiuniu, et al.
Published: (2025) -
InternSVG: Towards Unified SVG Tasks with Multimodal Large Language Models
by: Wang, Haomin, et al.
Published: (2025) -
Text-to-Vector Generation with Neural Path Representation
by: Zhang, Peiying, et al.
Published: (2024) -
WildSVG: Towards Reliable SVG Generation Under Real-Word Conditions
by: Terral, Marco, et al.
Published: (2026) -
Style Customization of Text-to-Vector Generation with Image Diffusion Priors
by: Zhang, Peiying, et al.
Published: (2025)