From Words to Structured Visuals: A Benchmark and Framework for Text-to-Diagram Generation and Editing
Fuente:
arXiv
Saved in:
| Main Authors: | Wei, Jingxuan, Tan, Cheng, Chen, Qi, Wu, Gaowei, Li, Siyuan, Gao, Zhangyang, Sun, Linzhuang, Yu, Bihui, Guo, Ruifeng |
|---|---|
| Format: | Preprint |
| Published: |
2024
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SketchAgent: Generating Structured Diagrams from Hand-Drawn Sketches
by: Tan, Cheng, et al.
Published: (2025)
by: Tan, Cheng, et al.
Published: (2025)
Rethinking Text-to-SQL: Dynamic Multi-turn SQL Interaction for Real-world Database Exploration
by: Sun, Linzhuang, et al.
Published: (2025)
by: Sun, Linzhuang, et al.
Published: (2025)
Boosting the Power of Small Multimodal Reasoning Models to Match Larger Models with Self-Consistency Training
by: Tan, Cheng, et al.
Published: (2023)
by: Tan, Cheng, et al.
Published: (2023)
Retrieval Meets Reasoning: Even High-school Textbook Knowledge Benefits Multimodal Reasoning
by: Tan, Cheng, et al.
Published: (2024)
by: Tan, Cheng, et al.
Published: (2024)
Sentence-Level or Token-Level? A Comprehensive Study on Knowledge Distillation
by: Wei, Jingxuan, et al.
Published: (2024)
by: Wei, Jingxuan, et al.
Published: (2024)
LLaVA-NeuMT: Selective Layer-Neuron Modulation for Efficient Multilingual Multimodal Translation
by: Wei, Jingxuan, et al.
Published: (2025)
by: Wei, Jingxuan, et al.
Published: (2025)
GGBench: A Geometric Generative Reasoning Benchmark for Unified Multimodal Models
by: Wei, Jingxuan, et al.
Published: (2025)
by: Wei, Jingxuan, et al.
Published: (2025)
ResearchPulse: Building Method-Experiment Chains through Multi-Document Scientific Inference
by: Chen, Qi, et al.
Published: (2025)
by: Chen, Qi, et al.
Published: (2025)
Rational Sensibility: LLM Enhanced Empathetic Response Generation Guided by Self-presentation Theory
by: Sun, Linzhuang, et al.
Published: (2023)
by: Sun, Linzhuang, et al.
Published: (2023)
SQLStructEval: Structural Evaluation of LLM Text-to-SQL Generation
by: Zhou, Yixi, et al.
Published: (2026)
by: Zhou, Yixi, et al.
Published: (2026)
DB-GPT-Hub: Towards Open Benchmarking Text-to-SQL Empowered by Large Language Models
by: Zhou, Fan, et al.
Published: (2024)
by: Zhou, Fan, et al.
Published: (2024)
Towards Reliable Agentic Progressive Text-to-Visualization with Verification Rules
by: Xu, Wenxin, et al.
Published: (2026)
by: Xu, Wenxin, et al.
Published: (2026)
Brain-inspired Computing Based on Deep Learning for Human-computer Interaction: A Review
by: Yu, Bihui, et al.
Published: (2023)
by: Yu, Bihui, et al.
Published: (2023)
Starling: An I/O-Efficient Disk-Resident Graph Index Framework for High-Dimensional Vector Similarity Search on Data Segment
by: Wang, Mengzhao, et al.
Published: (2024)
by: Wang, Mengzhao, et al.
Published: (2024)
Geoint-R1: Formalizing Multimodal Geometric Reasoning with Dynamic Auxiliary Constructions
by: Wei, Jingxuan, et al.
Published: (2025)
by: Wei, Jingxuan, et al.
Published: (2025)
A General Framework for Per-record Differential Privacy
by: Chen, Xinghe, et al.
Published: (2025)
by: Chen, Xinghe, et al.
Published: (2025)
DIVER: A Robust Text-to-SQL System with Dynamic Interactive Value Linking and Evidence Reasoning
by: Nan, Yafeng, et al.
Published: (2026)
by: Nan, Yafeng, et al.
Published: (2026)
EGREFINE: An Execution-Grounded Optimization Framework for Text-to-SQL Schema Refinement
by: Wang, Jiaqian, et al.
Published: (2026)
by: Wang, Jiaqian, et al.
Published: (2026)
Conformance Testing of Relational DBMS Against SQL Specifications
by: Liu, Shuang, et al.
Published: (2024)
by: Liu, Shuang, et al.
Published: (2024)
A Survey on Image-text Multimodal Models
by: Guo, Ruifeng, et al.
Published: (2023)
by: Guo, Ruifeng, et al.
Published: (2023)
Towards Robustness: A Critique of Current Vector Database Assessments
by: Wang, Zikai, et al.
Published: (2025)
by: Wang, Zikai, et al.
Published: (2025)
Efficient-Empathy: Towards Efficient and Effective Selection of Empathy Data
by: Sun, Linzhuang, et al.
Published: (2024)
by: Sun, Linzhuang, et al.
Published: (2024)
StructRide: A Framework to Exploit the Structure Information of Shareability Graph in Ridesharing
by: Zhan, Jiexi, et al.
Published: (2024)
by: Zhan, Jiexi, et al.
Published: (2024)
ChartMind: A Comprehensive Benchmark for Complex Real-world Multimodal Chart Question Answering
by: Wei, Jingxuan, et al.
Published: (2025)
by: Wei, Jingxuan, et al.
Published: (2025)
LST-Bench: Benchmarking Log-Structured Tables in the Cloud
by: Camacho-Rodríguez, Jesús, et al.
Published: (2023)
by: Camacho-Rodríguez, Jesús, et al.
Published: (2023)
FDABench: A Benchmark for Data Agents on Analytical Queries over Heterogeneous Data
by: Wang, Ziting, et al.
Published: (2025)
by: Wang, Ziting, et al.
Published: (2025)
Towards Defect Phase Diagrams: From Research Data Management to Automated Workflows
by: Rejiba, Khalil, et al.
Published: (2025)
by: Rejiba, Khalil, et al.
Published: (2025)
A Word-Based Compression Technique for Text Files.
by: Vernor, Russel L., III, et al.
Published: (1978)
by: Vernor, Russel L., III, et al.
Published: (1978)
PAGER: Bridging the Semantic-Execution Gap in Point-Precise Geometric GUI Control
by: Wei, Jingxuan, et al.
Published: (2026)
by: Wei, Jingxuan, et al.
Published: (2026)
SM3-Text-to-Query: Synthetic Multi-Model Medical Text-to-Query Benchmark
by: Sivasubramaniam, Sithursan, et al.
Published: (2024)
by: Sivasubramaniam, Sithursan, et al.
Published: (2024)
V-SQL: A View-based Two-stage Text-to-SQL Framework
by: You, Zeshun, et al.
Published: (2024)
by: You, Zeshun, et al.
Published: (2024)
From Text to Databases: attribute grammar as database meta-model
by: Chabin, Jacques, et al.
Published: (2024)
by: Chabin, Jacques, et al.
Published: (2024)
Revisiting Task-Oriented Dataset Search in the Era of Large Language Models: Challenges, Benchmark, and Solution
by: Wei, Zixin, et al.
Published: (2025)
by: Wei, Zixin, et al.
Published: (2025)
SING-SQL: A Synthetic Data Generation Framework for In-Domain Text-to-SQL Translation
by: Caferoğlu, Hasan Alp, et al.
Published: (2025)
by: Caferoğlu, Hasan Alp, et al.
Published: (2025)
Dupin: A Parallel Framework for Densest Subgraph Discovery in Fraud Detection on Massive Graphs (Technical Report)
by: Jiang, Jiaxin, et al.
Published: (2025)
by: Jiang, Jiaxin, et al.
Published: (2025)
BEATS: Optimizing LLM Mathematical Capabilities with BackVerify and Adaptive Disambiguate based Efficient Tree Search
by: Sun, Linzhuang, et al.
Published: (2024)
by: Sun, Linzhuang, et al.
Published: (2024)
Agentar-Scale-SQL: Advancing Text-to-SQL through Orchestrated Test-Time Scaling
by: Wang, Pengfei, et al.
Published: (2025)
by: Wang, Pengfei, et al.
Published: (2025)
Pervasive Annotation Errors Break Text-to-SQL Benchmarks and Leaderboards
by: Jin, Tengjun, et al.
Published: (2026)
by: Jin, Tengjun, et al.
Published: (2026)
An Extensive Study on Text Serialization Formats and Methods
by: Wei, Wang, et al.
Published: (2025)
by: Wei, Wang, et al.
Published: (2025)
A Survey of Large Language Model-Based Generative AI for Text-to-SQL: Benchmarks, Applications, Use Cases, and Challenges
by: Singh, Aditi, et al.
Published: (2024)
by: Singh, Aditi, et al.
Published: (2024)
Similar Items
-
SketchAgent: Generating Structured Diagrams from Hand-Drawn Sketches
by: Tan, Cheng, et al.
Published: (2025) -
Rethinking Text-to-SQL: Dynamic Multi-turn SQL Interaction for Real-world Database Exploration
by: Sun, Linzhuang, et al.
Published: (2025) -
Boosting the Power of Small Multimodal Reasoning Models to Match Larger Models with Self-Consistency Training
by: Tan, Cheng, et al.
Published: (2023) -
Retrieval Meets Reasoning: Even High-school Textbook Knowledge Benefits Multimodal Reasoning
by: Tan, Cheng, et al.
Published: (2024) -
Sentence-Level or Token-Level? A Comprehensive Study on Knowledge Distillation
by: Wei, Jingxuan, et al.
Published: (2024)