DiagramEval: Evaluating LLM-Generated Diagrams via Graphs
Fuente:
arXiv
Saved in:
| Main Authors: | Liang, Chumeng, You, Jiaxuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Debugging Tabular Log as Dynamic Graphs
by: Liang, Chumeng, et al.
Published: (2025)
by: Liang, Chumeng, et al.
Published: (2025)
AcademicEval: Live Long-Context LLM Benchmark
by: Zhang, Haozhen, et al.
Published: (2025)
by: Zhang, Haozhen, et al.
Published: (2025)
EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution
by: Wang, Tianfu, et al.
Published: (2026)
by: Wang, Tianfu, et al.
Published: (2026)
ConsistencyChecker: Tree-based Evaluation of LLM Generalization Capabilities
by: Hong, Zhaochen, et al.
Published: (2025)
by: Hong, Zhaochen, et al.
Published: (2025)
LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling
by: Chen, Yuxin, et al.
Published: (2026)
by: Chen, Yuxin, et al.
Published: (2026)
On the Diagram of Thought
by: Zhang, Yifan, et al.
Published: (2024)
by: Zhang, Yifan, et al.
Published: (2024)
DiagrammerGPT: Generating Open-Domain, Open-Platform Diagrams via LLM Planning
by: Zala, Abhay, et al.
Published: (2023)
by: Zala, Abhay, et al.
Published: (2023)
Real-World Benchmarks Make Membership Inference Attacks Fail on Diffusion Models
by: Liang, Chumeng, et al.
Published: (2024)
by: Liang, Chumeng, et al.
Published: (2024)
LegalViz: Legal Text Visualization by Text To Diagram Generation
by: Onami, Eri, et al.
Published: (2025)
by: Onami, Eri, et al.
Published: (2025)
Draw with Thought: Unleashing Multimodal Reasoning for Scientific Diagram Generation
by: Cui, Zhiqing, et al.
Published: (2025)
by: Cui, Zhiqing, et al.
Published: (2025)
DecompressionLM: Deterministic, Diagnostic, and Zero-Shot Concept Graph Extraction from Language Models
by: Hong, Zhaochen, et al.
Published: (2026)
by: Hong, Zhaochen, et al.
Published: (2026)
AntEval: Evaluation of Social Interaction Competencies in LLM-Driven Agents
by: Liang, Yuanzhi, et al.
Published: (2024)
by: Liang, Yuanzhi, et al.
Published: (2024)
Graph of Records: Boosting Retrieval Augmented Generation for Long-context Summarization with Graphs
by: Zhang, Haozhen, et al.
Published: (2024)
by: Zhang, Haozhen, et al.
Published: (2024)
SQLStructEval: Structural Evaluation of LLM Text-to-SQL Generation
by: Zhou, Yixi, et al.
Published: (2026)
by: Zhou, Yixi, et al.
Published: (2026)
A Categorical Framework for Modeling with Stock and Flow Diagrams
by: Baez, John C., et al.
Published: (2022)
by: Baez, John C., et al.
Published: (2022)
FeedEval: Pedagogically Aligned Evaluation of LLM-Generated Essay Feedback
by: Chu, Seongyeub, et al.
Published: (2026)
by: Chu, Seongyeub, et al.
Published: (2026)
CreativEval: Evaluating Creativity of LLM-Based Hardware Code Generation
by: DeLorenzo, Matthew, et al.
Published: (2024)
by: DeLorenzo, Matthew, et al.
Published: (2024)
Bluefish: Composing Diagrams with Declarative Relations
by: Pollock, Josh, et al.
Published: (2023)
by: Pollock, Josh, et al.
Published: (2023)
RocketEval: Efficient Automated LLM Evaluation via Grading Checklist
by: Wei, Tianjun, et al.
Published: (2025)
by: Wei, Tianjun, et al.
Published: (2025)
One-Eval: An Agentic System for Automated and Traceable LLM Evaluation
by: Shen, Chengyu, et al.
Published: (2026)
by: Shen, Chengyu, et al.
Published: (2026)
Can We Improve Educational Diagram Generation with In-Context Examples? Not if a Hallucination Spoils the Bunch
by: Logacheva, Evanfiya, et al.
Published: (2026)
by: Logacheva, Evanfiya, et al.
Published: (2026)
GraphEval: A Knowledge-Graph Based LLM Hallucination Evaluation Framework
by: Sansford, Hannah, et al.
Published: (2024)
by: Sansford, Hannah, et al.
Published: (2024)
ReliableEval: A Recipe for Stochastic LLM Evaluation via Method of Moments
by: Lior, Gili, et al.
Published: (2025)
by: Lior, Gili, et al.
Published: (2025)
Text2Arch: A Dataset for Generating Scientific Architecture Diagrams from Natural Language Descriptions
by: Garg, Shivank, et al.
Published: (2026)
by: Garg, Shivank, et al.
Published: (2026)
GeoSVG-RL: Geometry-Aware Reinforcement Learning for Layout-Constrained Text-to-SVG Diagram Generation
by: Li, Sifan, et al.
Published: (2026)
by: Li, Sifan, et al.
Published: (2026)
Large-Scale Constraint Generation -- Can LLMs Parse Hundreds of Constraints?
by: Boffa, Matteo, et al.
Published: (2025)
by: Boffa, Matteo, et al.
Published: (2025)
Evaluating Compliance with Visualization Guidelines in Diagrams for Scientific Publications Using Large Vision Language Models
by: Rückert, Johannes, et al.
Published: (2025)
by: Rückert, Johannes, et al.
Published: (2025)
RepEval: Effective Text Evaluation with LLM Representation
by: Sheng, Shuqian, et al.
Published: (2024)
by: Sheng, Shuqian, et al.
Published: (2024)
SurveyEval: Towards Comprehensive Evaluation of LLM-Generated Academic Surveys
by: Zhao, Jiahao, et al.
Published: (2025)
by: Zhao, Jiahao, et al.
Published: (2025)
String Diagrams for $λ$-calculi and Functional Computation
by: Ghica, Dan, et al.
Published: (2023)
by: Ghica, Dan, et al.
Published: (2023)
DiaCBT: A Long-Periodic Dialogue Corpus Guided by Cognitive Conceptualization Diagram for CBT-based Psychological Counseling
by: Zhou, Yougen, et al.
Published: (2025)
by: Zhou, Yougen, et al.
Published: (2025)
Venn Diagram Prompting : Accelerating Comprehension with Scaffolding Effect
by: Mahendru, Sakshi, et al.
Published: (2024)
by: Mahendru, Sakshi, et al.
Published: (2024)
GraphPlanner: Graph Memory-Augmented Agentic Routing for Multi-Agent LLMs
by: Feng, Tao, et al.
Published: (2026)
by: Feng, Tao, et al.
Published: (2026)
Grounded Language Design for Lightweight Diagramming for Formal Methods
by: Prasad, Siddhartha, et al.
Published: (2024)
by: Prasad, Siddhartha, et al.
Published: (2024)
CCD-CBT: Multi-Agent Therapeutic Interaction for CBT Guided by Cognitive Conceptualization Diagram
by: Liu, Chang, et al.
Published: (2026)
by: Liu, Chang, et al.
Published: (2026)
RouteProfile: Graph-Based Profiling for Cold-Start LLM Routing
by: Xu, Jingjun, et al.
Published: (2026)
by: Xu, Jingjun, et al.
Published: (2026)
Graph2Eval: Automatic Multimodal Task Generation for Agents via Knowledge Graphs
by: Chen, Yurun, et al.
Published: (2025)
by: Chen, Yurun, et al.
Published: (2025)
Beyond Facts: Evaluating Intent Hallucination in Large Language Models
by: Hao, Yijie, et al.
Published: (2025)
by: Hao, Yijie, et al.
Published: (2025)
GraphEval: A Lightweight Graph-Based LLM Framework for Idea Evaluation
by: Feng, Tao, et al.
Published: (2025)
by: Feng, Tao, et al.
Published: (2025)
PersonaEval: Are LLM Evaluators Human Enough to Judge Role-Play?
by: Zhou, Lingfeng, et al.
Published: (2025)
by: Zhou, Lingfeng, et al.
Published: (2025)
Similar Items
-
Debugging Tabular Log as Dynamic Graphs
by: Liang, Chumeng, et al.
Published: (2025) -
AcademicEval: Live Long-Context LLM Benchmark
by: Zhang, Haozhen, et al.
Published: (2025) -
EvoDiagram: Agentic Editable Diagram Creation via Design Expertise Evolution
by: Wang, Tianfu, et al.
Published: (2026) -
ConsistencyChecker: Tree-based Evaluation of LLM Generalization Capabilities
by: Hong, Zhaochen, et al.
Published: (2025) -
LangFlow: Continuous Diffusion Rivals Discrete in Language Modeling
by: Chen, Yuxin, et al.
Published: (2026)