AIBench: Evaluating Visual-Logical Consistency in Academic Illustration Generation
Fuente:
arXiv
Saved in:
| Main Authors: | Liao, Zhaohe, Jiang, Kaixun, Liu, Zhihang, Wei, Yujie, Yu, Junqiu, Li, Quanhao, Yu, Hong-Tao, Li, Pandeng, Wang, Yuzheng, Xing, Zhen, Zhang, Shiwei, Xie, Chen-Wei, Zheng, Yun, Liu, Xihui |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models
by: Li, Quanhao, et al.
Published: (2026)
by: Li, Quanhao, et al.
Published: (2026)
ShowTable: Unlocking Creative Table Visualization with Collaborative Reflection and Refinement
by: Liu, Zhihang, et al.
Published: (2025)
by: Liu, Zhihang, et al.
Published: (2025)
GenAgent: Scaling Text-to-Image Generation via Agentic Multimodal Reasoning
by: Jiang, Kaixun, et al.
Published: (2026)
by: Jiang, Kaixun, et al.
Published: (2026)
MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation
by: Wei, Yujie, et al.
Published: (2026)
by: Wei, Yujie, et al.
Published: (2026)
TTS-VAR: A Test-Time Scaling Framework for Visual Auto-Regressive Generation
by: Chen, Zhekai, et al.
Published: (2025)
by: Chen, Zhekai, et al.
Published: (2025)
Tracking vs. Deciding: The Dual-Capability Bottleneck in Searchless Chess Transformers
by: Li, Quanhao, et al.
Published: (2026)
by: Li, Quanhao, et al.
Published: (2026)
Hybrid-Level Instruction Injection for Video Token Compression in Multi-modal Large Language Models
by: Liu, Zhihang, et al.
Published: (2025)
by: Liu, Zhihang, et al.
Published: (2025)
CAPability: A Comprehensive Visual Caption Benchmark for Evaluating Both Correctness and Thoroughness
by: Liu, Zhihang, et al.
Published: (2025)
by: Liu, Zhihang, et al.
Published: (2025)
MobileAIBench: Benchmarking LLMs and LMMs for On-Device Use Cases
by: Murthy, Rithesh, et al.
Published: (2024)
by: Murthy, Rithesh, et al.
Published: (2024)
Rethinking Video Tokenization: A Conditioned Diffusion-based Approach
by: Yang, Nianzu, et al.
Published: (2025)
by: Yang, Nianzu, et al.
Published: (2025)
Perspective of high-speed Mach-Zehnder modulators based on nonlinear optics and complex band structures
by: Li, Shuyi, et al.
Published: (2025)
by: Li, Shuyi, et al.
Published: (2025)
Align and Aggregate: Compositional Reasoning with Video Alignment and Answer Aggregation for Video Question-Answering
by: Liao, Zhaohe, et al.
Published: (2024)
by: Liao, Zhaohe, et al.
Published: (2024)
PaperBanana: Automating Academic Illustration for AI Scientists
by: Zhu, Dawei, et al.
Published: (2026)
by: Zhu, Dawei, et al.
Published: (2026)
The Chern Sectional Curvature of a Hermitian Manifold
by: Cao, Pandeng, et al.
Published: (2022)
by: Cao, Pandeng, et al.
Published: (2022)
The behavior of rich-club coefficient in scale-free networks
by: Liu, Zhihang, et al.
Published: (2023)
by: Liu, Zhihang, et al.
Published: (2023)
Routing Matters in MoE: Scaling Diffusion Transformers with Explicit Routing Guidance
by: Wei, Yujie, et al.
Published: (2025)
by: Wei, Yujie, et al.
Published: (2025)
Tensor Manifold-Based Graph-Vector Fusion for AI-Native Academic Literature Retrieval
by: Wei, Xing, et al.
Published: (2026)
by: Wei, Xing, et al.
Published: (2026)
MagicMotion: Controllable Video Generation with Dense-to-Sparse Trajectory Guidance
by: Li, Quanhao, et al.
Published: (2025)
by: Li, Quanhao, et al.
Published: (2025)
Fracture Mechanics of 2D Crystal Blisters with Irregular Geometry
by: Jiacong Cao, et al.
Published: (2025)
by: Jiacong Cao, et al.
Published: (2025)
Aligned Stable Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency
by: Wang, Yikai, et al.
Published: (2026)
by: Wang, Yikai, et al.
Published: (2026)
Text-to-TrajVis: Enabling Trajectory Data Visualizations from Natural Language Questions
by: Bai, Tian, et al.
Published: (2025)
by: Bai, Tian, et al.
Published: (2025)
LogicVista: Multimodal LLM Logical Reasoning Benchmark in Visual Contexts
by: Xiao, Yijia, et al.
Published: (2024)
by: Xiao, Yijia, et al.
Published: (2024)
Capturing Complex Rate‐Dependent Behaviors of Saturated Clays: A Fractional Consistency Kinematic Hardening Viscoplastic Approach
by: Wei Cheng, et al.
Published: (2026)
by: Wei Cheng, et al.
Published: (2026)
Basic Composition and Synthesis Process and Representative Molecule of the Branched Polyethyleneimines With Low Molecular Weight
by: Pandeng Li, et al.
Published: (2025)
by: Pandeng Li, et al.
Published: (2025)
Generic Mackey Formula for Parahoric Lusztig Functors
by: Yu, Zhihang
Published: (2025)
by: Yu, Zhihang
Published: (2025)
Taming Consistency Distillation for Accelerated Human Image Animation
by: Wang, Xiang, et al.
Published: (2025)
by: Wang, Xiang, et al.
Published: (2025)
Multimodal Visual Image Based User Association and Beamforming Using Graph Neural Networks
by: Li, Yinghan, et al.
Published: (2025)
by: Li, Yinghan, et al.
Published: (2025)
Jovian Zonal Winds Revealed from Cassini/VIMS Observations
by: Ma, Shenghan, et al.
Published: (2026)
by: Ma, Shenghan, et al.
Published: (2026)
S2ED: From Story to Executable Descriptions for Consistency-Aware Story Illustration
by: Yin, Sijing, et al.
Published: (2026)
by: Yin, Sijing, et al.
Published: (2026)
PanguMotion: Continuous Driving Motion Forecasting with Pangu Transformers
by: Ren, Quanhao, et al.
Published: (2026)
by: Ren, Quanhao, et al.
Published: (2026)
FlashMotion: Few-Step Controllable Video Generation with Trajectory Guidance
by: Li, Quanhao, et al.
Published: (2026)
by: Li, Quanhao, et al.
Published: (2026)
Polaris: Open-ended Interactive Robotic Manipulation via Syn2Real Visual Grounding and Large Language Models
by: Wang, Tianyu, et al.
Published: (2024)
by: Wang, Tianyu, et al.
Published: (2024)
Transmembrane Ion Channels: From Natural to Artificial Systems
by: Tengfei Yan, et al.
Published: (2024)
by: Tengfei Yan, et al.
Published: (2024)
Context as Memory: Scene-Consistent Interactive Long Video Generation with Memory Retrieval
by: Yu, Jiwen, et al.
Published: (2025)
by: Yu, Jiwen, et al.
Published: (2025)
SynMotion: Semantic-Visual Adaptation for Motion Customized Video Generation
by: Tan, Shuai, et al.
Published: (2025)
by: Tan, Shuai, et al.
Published: (2025)
TransFit-CSM: A Fast, Physically Consistent Framework for Interaction-Powered Transients
by: Zhang, Yu-Hao, et al.
Published: (2025)
by: Zhang, Yu-Hao, et al.
Published: (2025)
Object Isolated Attention for Consistent Story Visualization
by: Luo, Xiangyang, et al.
Published: (2025)
by: Luo, Xiangyang, et al.
Published: (2025)
Towards Enhanced Image Inpainting: Mitigating Unwanted Object Insertion and Preserving Color Consistency
by: Wang, Yikai, et al.
Published: (2023)
by: Wang, Yikai, et al.
Published: (2023)
The Power of Many: Synergistic Unification of Diverse Augmentations for Efficient Adversarial Robustness
by: Yu-Hang, Wang, et al.
Published: (2025)
by: Yu-Hang, Wang, et al.
Published: (2025)
UFO: A Unified Approach to Fine-grained Visual Perception via Open-ended Language Interface
by: Tang, Hao, et al.
Published: (2025)
by: Tang, Hao, et al.
Published: (2025)
Similar Items
-
DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models
by: Li, Quanhao, et al.
Published: (2026) -
ShowTable: Unlocking Creative Table Visualization with Collaborative Reflection and Refinement
by: Liu, Zhihang, et al.
Published: (2025) -
GenAgent: Scaling Text-to-Image Generation via Agentic Multimodal Reasoning
by: Jiang, Kaixun, et al.
Published: (2026) -
MSAVBench: Towards Comprehensive and Reliable Evaluation of Multi-Shot Audio-Video Generation
by: Wei, Yujie, et al.
Published: (2026) -
TTS-VAR: A Test-Time Scaling Framework for Visual Auto-Regressive Generation
by: Chen, Zhekai, et al.
Published: (2025)