VIS-Shepherd: Constructing Critic for LLM-based Data Visualization Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Pan, Bo, Fu, Yixiao, Wang, Ke, Lu, Junyu, Pan, Lunke, Qian, Ziyang, Chen, Yuhan, Wang, Guoliang, Zhou, Yitao, Zheng, Li, Tang, Yinghao, Wen, Zhen, Wu, Yuchen, Lu, Junhua, Zhu, Biao, Zhu, Minfeng, Zhang, Bo, Chen, Wei |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
InterDeepResearch: Enabling Human-Agent Collaborative Information Seeking through Interactive Deep Research
by: Pan, Bo, et al.
Published: (2026)
by: Pan, Bo, et al.
Published: (2026)
AgentCoord: Visually Exploring Coordination Strategy for LLM-based Multi-Agent Collaboration
by: Pan, Bo, et al.
Published: (2024)
by: Pan, Bo, et al.
Published: (2024)
Exploring Multimodal Prompt for Visualization Authoring with Large Language Models
by: Wen, Zhen, et al.
Published: (2025)
by: Wen, Zhen, et al.
Published: (2025)
IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering
by: Wang, Yiyao, et al.
Published: (2025)
by: Wang, Yiyao, et al.
Published: (2025)
SyncVIS: Synchronized Video Instance Segmentation
by: Zheng, Rongkun, et al.
Published: (2024)
by: Zheng, Rongkun, et al.
Published: (2024)
Multimodal DeepResearcher: Generating Text-Chart Interleaved Reports From Scratch with Agentic Framework
by: Yang, Zhaorui, et al.
Published: (2025)
by: Yang, Zhaorui, et al.
Published: (2025)
AgentLens: Visual Analysis for Agent Behaviors in LLM-based Autonomous Systems
by: Lu, Jiaying, et al.
Published: (2024)
by: Lu, Jiaying, et al.
Published: (2024)
OTO Planner: An Efficient Only Travelling Once Exploration Planner for Complex and Unknown Environments
by: Zhou, Bo, et al.
Published: (2024)
by: Zhou, Bo, et al.
Published: (2024)
Topology-Enhanced Alignment for Large Language Models: Trajectory Topology Loss and Topological Preference Optimization
by: Pan, Yurui, et al.
Published: (2026)
by: Pan, Yurui, et al.
Published: (2026)
TMT-VIS: Taxonomy-aware Multi-dataset Joint Training for Video Instance Segmentation
by: Zheng, Rongkun, et al.
Published: (2023)
by: Zheng, Rongkun, et al.
Published: (2023)
RAGExplorer: A Visual Analytics System for the Comparative Diagnosis of RAG Systems
by: Tian, Haoyu, et al.
Published: (2026)
by: Tian, Haoyu, et al.
Published: (2026)
FlueBricks: A Construction Kit of Flute-like Instruments for Acoustic Reasoning
by: Chen, Bo-Yu, et al.
Published: (2026)
by: Chen, Bo-Yu, et al.
Published: (2026)
R1-Onevision: Advancing Generalized Multimodal Reasoning through Cross-Modal Formalization
by: Yang, Yi, et al.
Published: (2025)
by: Yang, Yi, et al.
Published: (2025)
X2Edit: Revisiting Arbitrary-Instruction Image Editing through Self-Constructed Data and Task-Aware Representation Learning
by: Ma, Jian, et al.
Published: (2025)
by: Ma, Jian, et al.
Published: (2025)
SuperGS: Consistent and Detailed 3D Super-Resolution Scene Reconstruction via Gaussian Splatting
by: Xie, Shiyun, et al.
Published: (2025)
by: Xie, Shiyun, et al.
Published: (2025)
SuperGS: Super-Resolution 3D Gaussian Splatting Enhanced by Variational Residual Features and Uncertainty-Augmented Learning
by: Xie, Shiyun, et al.
Published: (2024)
by: Xie, Shiyun, et al.
Published: (2024)
PA-CLIP: Enhancing Zero-Shot Anomaly Detection through Pseudo-Anomaly Awareness
by: Pan, Yurui, et al.
Published: (2025)
by: Pan, Yurui, et al.
Published: (2025)
Each Test Image Deserves A Specific Prompt: Continual Test-Time Adaptation for 2D Medical Image Segmentation
by: Chen, Ziyang, et al.
Published: (2023)
by: Chen, Ziyang, et al.
Published: (2023)
Meflex: A Multi-agent Scaffolding System for Entrepreneurial Ideation Iteration via Nonlinear Business Plan Writing
by: Luo, Lan, et al.
Published: (2026)
by: Luo, Lan, et al.
Published: (2026)
LANDeRMT: Detecting and Routing Language-Aware Neurons for Selectively Finetuning LLMs to Machine Translation
by: Zhu, Shaolin, et al.
Published: (2024)
by: Zhu, Shaolin, et al.
Published: (2024)
Edge‐Guided Dual‐Stream Collaborative Network for Salient Object Detection in Low‐Light Images
by: Wenjun Lu, et al.
Published: (2026)
by: Wenjun Lu, et al.
Published: (2026)
Alignment for Efficient Tool Calling of Large Language Models
by: Xu, Hongshen, et al.
Published: (2025)
by: Xu, Hongshen, et al.
Published: (2025)
LLM-as-a-Reviewer: Benchmarking Their Ability, Divergence, and Prompt Injection Resistance as Paper Reviewers
by: Li, Lingyao, et al.
Published: (2026)
by: Li, Lingyao, et al.
Published: (2026)
Designing the Future of Entrepreneurship Education: Exploring an AI-Empowered Scaffold System for Business Plan Development
by: Zhu, Junhua, et al.
Published: (2025)
by: Zhu, Junhua, et al.
Published: (2025)
An Impulse-formed Navier-Stokes Solver based on Long-range Particle Flow Maps
by: Li, Zhiqi, et al.
Published: (2026)
by: Li, Zhiqi, et al.
Published: (2026)
XGraphRAG: Interactive Visual Analysis for Graph-based Retrieval-Augmented Generation
by: Wang, Ke, et al.
Published: (2025)
by: Wang, Ke, et al.
Published: (2025)
ConceptViz: A Visual Analytics Approach for Exploring Concepts in Large Language Models
by: Li, Haoxuan, et al.
Published: (2025)
by: Li, Haoxuan, et al.
Published: (2025)
Towards Comprehensive Detection of Chinese Harmful Memes
by: Lu, Junyu, et al.
Published: (2024)
by: Lu, Junyu, et al.
Published: (2024)
LightM-UNet: Mamba Assists in Lightweight UNet for Medical Image Segmentation
by: Liao, Weibin, et al.
Published: (2024)
by: Liao, Weibin, et al.
Published: (2024)
Delving into Spectral Clustering with Vision-Language Representations
by: Peng, Bo, et al.
Published: (2026)
by: Peng, Bo, et al.
Published: (2026)
IGenBench: Benchmarking the Reliability of Text-to-Infographic Generation
by: Tang, Yinghao, et al.
Published: (2026)
by: Tang, Yinghao, et al.
Published: (2026)
SBCA: Cross-Modal BERT-driven Actor-Critic for Multi-Asset Portfolio Optimization
by: Pan, Jinfeng, et al.
Published: (2026)
by: Pan, Jinfeng, et al.
Published: (2026)
Med-Query: Steerable Parsing of 9-DoF Medical Anatomies with Query Embedding
by: Guo, Heng, et al.
Published: (2022)
by: Guo, Heng, et al.
Published: (2022)
Class Similarity-Based Multimodal Classification under Heterogeneous Category Sets
by: Zhu, Yangrui, et al.
Published: (2025)
by: Zhu, Yangrui, et al.
Published: (2025)
Does GenAI Rewrite How We Write? An Empirical Study on Two-Million Preprints
by: Qi, Minfeng, et al.
Published: (2025)
by: Qi, Minfeng, et al.
Published: (2025)
Olapa-MCoT: Enhancing the Chinese Mathematical Reasoning Capability of LLMs
by: Zhu, Shaojie, et al.
Published: (2023)
by: Zhu, Shaojie, et al.
Published: (2023)
V-Zero: Self-Improving Multimodal Reasoning with Zero Annotation
by: Wang, Han, et al.
Published: (2026)
by: Wang, Han, et al.
Published: (2026)
DreamJourney: Perpetual View Generation with Video Diffusion Models
by: Pan, Bo, et al.
Published: (2025)
by: Pan, Bo, et al.
Published: (2025)
Explaining latent representations of generative models with large multimodal models
by: Zhu, Mengdan, et al.
Published: (2024)
by: Zhu, Mengdan, et al.
Published: (2024)
EEA: Exploration-Exploitation Agent for Long Video Understanding
by: Yang, Te, et al.
Published: (2025)
by: Yang, Te, et al.
Published: (2025)
Similar Items
-
InterDeepResearch: Enabling Human-Agent Collaborative Information Seeking through Interactive Deep Research
by: Pan, Bo, et al.
Published: (2026) -
AgentCoord: Visually Exploring Coordination Strategy for LLM-based Multi-Agent Collaboration
by: Pan, Bo, et al.
Published: (2024) -
Exploring Multimodal Prompt for Visualization Authoring with Large Language Models
by: Wen, Zhen, et al.
Published: (2025) -
IntuiTF: MLLM-Guided Transfer Function Optimization for Direct Volume Rendering
by: Wang, Yiyao, et al.
Published: (2025) -
SyncVIS: Synchronized Video Instance Segmentation
by: Zheng, Rongkun, et al.
Published: (2024)