GeoSVG-RL: Geometry-Aware Reinforcement Learning for Layout-Constrained Text-to-SVG Diagram Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Li, Sifan, Cai, Yujun, Chen, Hongkai, Wang, Yiwei |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RoboSVG: A Unified Framework for Interactive SVG Generation with Multi-modal Guidance
von: Wang, Jiuniu, et al.
Veröffentlicht: (2025)
von: Wang, Jiuniu, et al.
Veröffentlicht: (2025)
OptiSQL: Executable SQL Generation from Optical Tokens
von: Li, Sifan, et al.
Veröffentlicht: (2026)
von: Li, Sifan, et al.
Veröffentlicht: (2026)
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024)
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
von: Li, Sifan, et al.
Veröffentlicht: (2025)
von: Li, Sifan, et al.
Veröffentlicht: (2025)
SemVink: Advancing VLMs' Semantic Understanding of Optical Illusions via Visual Global Thinking
von: Li, Sifan, et al.
Veröffentlicht: (2025)
von: Li, Sifan, et al.
Veröffentlicht: (2025)
IntroSVG: Learning from Rendering Feedback for Text-to-SVG Generation via an Introspective Generator-Critic Framework
von: Wang, Feiyu, et al.
Veröffentlicht: (2026)
von: Wang, Feiyu, et al.
Veröffentlicht: (2026)
T-SVG: Text-Driven Stereoscopic Video Generation
von: Jin, Qiao, et al.
Veröffentlicht: (2024)
von: Jin, Qiao, et al.
Veröffentlicht: (2024)
VCode: a Multimodal Coding Benchmark with SVG as Symbolic Visual Representation
von: Lin, Kevin Qinghong, et al.
Veröffentlicht: (2025)
von: Lin, Kevin Qinghong, et al.
Veröffentlicht: (2025)
LiveSVG: Zero-Shot SVG Animation via Video Generation
von: Levy, Matan, et al.
Veröffentlicht: (2026)
von: Levy, Matan, et al.
Veröffentlicht: (2026)
DuetSVG: Unified Multimodal SVG Generation with Internal Visual Guidance
von: Zhang, Peiying, et al.
Veröffentlicht: (2025)
von: Zhang, Peiying, et al.
Veröffentlicht: (2025)
InternSVG: Towards Unified SVG Tasks with Multimodal Large Language Models
von: Wang, Haomin, et al.
Veröffentlicht: (2025)
von: Wang, Haomin, et al.
Veröffentlicht: (2025)
WildSVG: Towards Reliable SVG Generation Under Real-Word Conditions
von: Terral, Marco, et al.
Veröffentlicht: (2026)
von: Terral, Marco, et al.
Veröffentlicht: (2026)
SVGDreamer: Text Guided SVG Generation with Diffusion Model
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
von: Xing, Ximing, et al.
Veröffentlicht: (2023)
SVGThinker: Instruction-Aligned and Reasoning-Driven Text-to-SVG Generation
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
von: Chen, Hanqi, et al.
Veröffentlicht: (2025)
Do "New Snow Tablets" Contain Snow? Large Language Models Over-Rely on Names to Identify Ingredients of Chinese Drugs
von: Li, Sifan, et al.
Veröffentlicht: (2025)
von: Li, Sifan, et al.
Veröffentlicht: (2025)
NeuralSVG: An Implicit Representation for Text-to-Vector Generation
von: Polaczek, Sagi, et al.
Veröffentlicht: (2025)
von: Polaczek, Sagi, et al.
Veröffentlicht: (2025)
SVGDreamer++: Advancing Editability and Diversity in Text-Guided SVG Generation
von: Xing, Ximing, et al.
Veröffentlicht: (2024)
von: Xing, Ximing, et al.
Veröffentlicht: (2024)
Reason-SVG: Enhancing Structured Reasoning for Vector Graphics Generation with Reinforcement Learning
von: Xing, Ximing, et al.
Veröffentlicht: (2025)
von: Xing, Ximing, et al.
Veröffentlicht: (2025)
SVGBuilder: Component-Based Colored SVG Generation with Text-Guided Autoregressive Transformers
von: Chen, Zehao, et al.
Veröffentlicht: (2024)
von: Chen, Zehao, et al.
Veröffentlicht: (2024)
AudioRouter: Data Efficient Audio Understanding via RL based Dual Reasoning
von: Chen, Liyang, et al.
Veröffentlicht: (2026)
von: Chen, Liyang, et al.
Veröffentlicht: (2026)
Unveiling the Potential of Diffusion Large Language Model in Controllable Generation
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
Reliable Reasoning in SVG-LLMs via Multi-Task Multi-Reward Reinforcement Learning
von: Wang, Haomin, et al.
Veröffentlicht: (2026)
von: Wang, Haomin, et al.
Veröffentlicht: (2026)
SVGenius: Benchmarking LLMs in SVG Understanding, Editing and Generation
von: Chen, Siqi, et al.
Veröffentlicht: (2025)
von: Chen, Siqi, et al.
Veröffentlicht: (2025)
SVGauge: Towards Human-Aligned Evaluation for SVG Generation
von: Zini, Leonardo, et al.
Veröffentlicht: (2025)
von: Zini, Leonardo, et al.
Veröffentlicht: (2025)
Symbolic or Numerical? Understanding Physics Problem Solving in Reasoning LLMs
von: Dan, Nifu, et al.
Veröffentlicht: (2025)
von: Dan, Nifu, et al.
Veröffentlicht: (2025)
OmniSVG: A Unified Scalable Vector Graphics Generation Model
von: Yang, Yiying, et al.
Veröffentlicht: (2025)
von: Yang, Yiying, et al.
Veröffentlicht: (2025)
Decomate: Leveraging Generative Models for Co-Creative SVG Animation
von: Park, Jihyeon, et al.
Veröffentlicht: (2025)
von: Park, Jihyeon, et al.
Veröffentlicht: (2025)
Mapping the Minds of LLMs: A Graph-Based Analysis of Reasoning LLM
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
von: Xiong, Zhen, et al.
Veröffentlicht: (2025)
BoundRL: Efficient Structured Text Segmentation through Reinforced Boundary Generation
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
von: Li, Haoyuan, et al.
Veröffentlicht: (2025)
From Tokens to Numbers: Continuous Number Modeling for SVG Generation
von: Ogezi, Michael, et al.
Veröffentlicht: (2026)
von: Ogezi, Michael, et al.
Veröffentlicht: (2026)
Vulnerability of LLMs to Vertically Aligned Text Manipulations
von: Li, Zhecheng, et al.
Veröffentlicht: (2024)
von: Li, Zhecheng, et al.
Veröffentlicht: (2024)
Structured Attention Matters to Multimodal LLMs in Document Understanding
von: Liu, Chang, et al.
Veröffentlicht: (2025)
von: Liu, Chang, et al.
Veröffentlicht: (2025)
Detecting and Mitigating Insertion Hallucination in Video-to-Audio Generation
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
von: Chen, Liyang, et al.
Veröffentlicht: (2025)
VisAnatomy: An SVG Chart Corpus with Fine-Grained Semantic Labels
von: Chen, Chen, et al.
Veröffentlicht: (2024)
von: Chen, Chen, et al.
Veröffentlicht: (2024)
SVG-IR: Spatially-Varying Gaussian Splatting for Inverse Rendering
von: Sun, Hanxiao, et al.
Veröffentlicht: (2025)
von: Sun, Hanxiao, et al.
Veröffentlicht: (2025)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
von: Wang, Baode, et al.
Veröffentlicht: (2025)
von: Wang, Baode, et al.
Veröffentlicht: (2025)
AmodalSVG: Amodal Image Vectorization via Semantic Layer Peeling
von: Hu, Juncheng, et al.
Veröffentlicht: (2026)
von: Hu, Juncheng, et al.
Veröffentlicht: (2026)
VFIG: Vectorizing Complex Figures in SVG with Vision-Language Models
von: He, Qijia, et al.
Veröffentlicht: (2026)
von: He, Qijia, et al.
Veröffentlicht: (2026)
Socratic Chart: Cooperating Multiple Agents for Robust SVG Chart Understanding
von: Ji, Yuyang, et al.
Veröffentlicht: (2025)
von: Ji, Yuyang, et al.
Veröffentlicht: (2025)
Infinity Parser: Layout Aware Reinforcement Learning for Scanned Document Parsing
von: Wang, Baode, et al.
Veröffentlicht: (2025)
von: Wang, Baode, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
RoboSVG: A Unified Framework for Interactive SVG Generation with Multi-modal Guidance
von: Wang, Jiuniu, et al.
Veröffentlicht: (2025) -
OptiSQL: Executable SQL Generation from Optical Tokens
von: Li, Sifan, et al.
Veröffentlicht: (2026) -
CraftSVG: Multi-Object Text-to-SVG Synthesis via Layout Guided Diffusion
von: Banerjee, Ayan, et al.
Veröffentlicht: (2024) -
Vision Language Models Map Logos to Text via Semantic Entanglement in the Visual Projector
von: Li, Sifan, et al.
Veröffentlicht: (2025) -
SemVink: Advancing VLMs' Semantic Understanding of Optical Illusions via Visual Global Thinking
von: Li, Sifan, et al.
Veröffentlicht: (2025)