GeoSVG-RL: Geometry-Aware Reinforcement Learning for Layout-Constrained Text-to-SVG Diagram Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Li, Sifan, Cai, Yujun, Chen, Hongkai, Wang, Yiwei |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
RoboSVG: A Unified Framework for Interactive SVG Generation with Multi-modal Guidance
by: Wang, Jiuniu, et al.
Published: (2025)
by: Wang, Jiuniu, et al.
Published: (2025)
OptiSQL: Executable SQL Generation from Optical Tokens
by: Li, Sifan, et al.
Published: (2026)
by: Li, Sifan, et al.
Published: (2026)
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
by: Banerjee, Ayan, et al.
Published: (2024)
by: Banerjee, Ayan, et al.
Published: (2024)
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
by: Li, Sifan, et al.
Published: (2025)
by: Li, Sifan, et al.
Published: (2025)
SemVink: Advancing VLMs' Semantic Understanding of Optical Illusions via Visual Global Thinking
by: Li, Sifan, et al.
Published: (2025)
by: Li, Sifan, et al.
Published: (2025)
IntroSVG: Learning from Rendering Feedback for Text-to-SVG Generation via an Introspective Generator-Critic Framework
by: Wang, Feiyu, et al.
Published: (2026)
by: Wang, Feiyu, et al.
Published: (2026)
T-SVG: Text-Driven Stereoscopic Video Generation
by: Jin, Qiao, et al.
Published: (2024)
by: Jin, Qiao, et al.
Published: (2024)
VCode: a Multimodal Coding Benchmark with SVG as Symbolic Visual Representation
by: Lin, Kevin Qinghong, et al.
Published: (2025)
by: Lin, Kevin Qinghong, et al.
Published: (2025)
LiveSVG: Zero-Shot SVG Animation via Video Generation
by: Levy, Matan, et al.
Published: (2026)
by: Levy, Matan, et al.
Published: (2026)
DuetSVG: Unified Multimodal SVG Generation with Internal Visual Guidance
by: Zhang, Peiying, et al.
Published: (2025)
by: Zhang, Peiying, et al.
Published: (2025)
InternSVG: Towards Unified SVG Tasks with Multimodal Large Language Models
by: Wang, Haomin, et al.
Published: (2025)
by: Wang, Haomin, et al.
Published: (2025)
WildSVG: Towards Reliable SVG Generation Under Real-Word Conditions
by: Terral, Marco, et al.
Published: (2026)
by: Terral, Marco, et al.
Published: (2026)
SVGDreamer: Text Guided SVG Generation with Diffusion Model
by: Xing, Ximing, et al.
Published: (2023)
by: Xing, Ximing, et al.
Published: (2023)
SVGThinker: Instruction-Aligned and Reasoning-Driven Text-to-SVG Generation
by: Chen, Hanqi, et al.
Published: (2025)
by: Chen, Hanqi, et al.
Published: (2025)
Do "New Snow Tablets" Contain Snow? Large Language Models Over-Rely on Names to Identify Ingredients of Chinese Drugs
by: Li, Sifan, et al.
Published: (2025)
by: Li, Sifan, et al.
Published: (2025)
NeuralSVG: An Implicit Representation for Text-to-Vector Generation
by: Polaczek, Sagi, et al.
Published: (2025)
by: Polaczek, Sagi, et al.
Published: (2025)
SVGDreamer++: Advancing Editability and Diversity in Text-Guided SVG Generation
by: Xing, Ximing, et al.
Published: (2024)
by: Xing, Ximing, et al.
Published: (2024)
Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning
by: Xing, Ximing, et al.
Published: (2025)
by: Xing, Ximing, et al.
Published: (2025)
SVGBuilder: Component-Based Colored SVG Generation with Text-Guided Autoregressive Transformers
by: Chen, Zehao, et al.
Published: (2024)
by: Chen, Zehao, et al.
Published: (2024)
AudioRouter: Data Efficient Audio Understanding via RL based Dual Reasoning
by: Chen, Liyang, et al.
Published: (2026)
by: Chen, Liyang, et al.
Published: (2026)
Unveiling the Potential of Diffusion Large Language Model in Controllable Generation
by: Xiong, Zhen, et al.
Published: (2025)
by: Xiong, Zhen, et al.
Published: (2025)
Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning
by: Wang, Haomin, et al.
Published: (2026)
by: Wang, Haomin, et al.
Published: (2026)
SVGenius: Benchmarking LLMs in SVG Understanding, Editing and Generation
by: Chen, Siqi, et al.
Published: (2025)
by: Chen, Siqi, et al.
Published: (2025)
SVGauge: Towards Human-Aligned Evaluation for SVG Generation
by: Zini, Leonardo, et al.
Published: (2025)
by: Zini, Leonardo, et al.
Published: (2025)
Symbolic or Numerical? Understanding Physics Problem Solving in Reasoning LLMs
by: Dan, Nifu, et al.
Published: (2025)
by: Dan, Nifu, et al.
Published: (2025)
OmniSVG: A Unified Scalable Vector Graphics Generation Model
by: Yang, Yiying, et al.
Published: (2025)
by: Yang, Yiying, et al.
Published: (2025)
Decomate: Leveraging Generative Models for Co-Creative SVG Animation
by: Park, Jihyeon, et al.
Published: (2025)
by: Park, Jihyeon, et al.
Published: (2025)
Mapping the Minds of LLMs: A Graph-Based Analysis of Reasoning LLM
by: Xiong, Zhen, et al.
Published: (2025)
by: Xiong, Zhen, et al.
Published: (2025)
BoundRL: Efficient Structured Text Segmentation through Reinforced Boundary Generation
by: Li, Haoyuan, et al.
Published: (2025)
by: Li, Haoyuan, et al.
Published: (2025)
From Tokens to Numbers: Continuous Number Modeling for SVG Generation
by: Ogezi, Michael, et al.
Published: (2026)
by: Ogezi, Michael, et al.
Published: (2026)
Vulnerability of LLMs to Vertically Aligned Text Manipulations
by: Li, Zhecheng, et al.
Published: (2024)
by: Li, Zhecheng, et al.
Published: (2024)
Structured Attention Matters to Multimodal LLMs in Document Understanding
by: Liu, Chang, et al.
Published: (2025)
by: Liu, Chang, et al.
Published: (2025)
Detecting and Mitigating Insertion Hallucination in Video-to-Audio Generation
by: Chen, Liyang, et al.
Published: (2025)
by: Chen, Liyang, et al.
Published: (2025)
VisAnatomy: An SVG Chart Corpus with Fine-Grained Semantic Labels
by: Chen, Chen, et al.
Published: (2024)
by: Chen, Chen, et al.
Published: (2024)
SVG-IR: Spatially-Varying Gaussian Splatting for Inverse Rendering
by: Sun, Hanxiao, et al.
Published: (2025)
by: Sun, Hanxiao, et al.
Published: (2025)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
by: Wang, Baode, et al.
Published: (2025)
by: Wang, Baode, et al.
Published: (2025)
AmodalSVG: Amodal Image Vectorization via Semantic Layer Peeling
by: Hu, Juncheng, et al.
Published: (2026)
by: Hu, Juncheng, et al.
Published: (2026)
VFIG: Vectorizing Complex Figures in SVG with Vision-Language Models
by: He, Qijia, et al.
Published: (2026)
by: He, Qijia, et al.
Published: (2026)
Socratic Chart: Cooperating Multiple Agents for Robust SVG Chart Understanding
by: Ji, Yuyang, et al.
Published: (2025)
by: Ji, Yuyang, et al.
Published: (2025)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
by: Wang, Baode, et al.
Published: (2025)
by: Wang, Baode, et al.
Published: (2025)
Similar Items
-
RoboSVG: A Unified Framework for Interactive SVG Generation with Multi-modal Guidance
by: Wang, Jiuniu, et al.
Published: (2025) -
OptiSQL: Executable SQL Generation from Optical Tokens
by: Li, Sifan, et al.
Published: (2026) -
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
by: Banerjee, Ayan, et al.
Published: (2024) -
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
by: Li, Sifan, et al.
Published: (2025) -
SemVink: Advancing VLMs' Semantic Understanding of Optical Illusions via Visual Global Thinking
by: Li, Sifan, et al.
Published: (2025)