CircuitSense: A Hierarchical MLLM Benchmark Bridging Visual Comprehension and Symbolic Reasoning in Engineering Design Process
Fuente:
arXiv
Saved in:
| Main Authors: | Akbari, Arman, Gao, Jian, Zou, Yifei, Yang, Mei, Duan, Jinru, Torbunov, Dmitrii, Wang, Yanzhi, Ren, Yihui, Zhang, Xuan |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents
by: Yu, Xi, et al.
Published: (2026)
by: Yu, Xi, et al.
Published: (2026)
Parameter Inference and Uncertainty Quantification with Diffusion Models: Extending CDI to 2D Spatial Conditioning
by: Torbunov, Dmitrii, et al.
Published: (2026)
by: Torbunov, Dmitrii, et al.
Published: (2026)
EvRT-DETR: Latent Space Adaptation of Image Detectors for Event-based Vision
by: Torbunov, Dmitrii, et al.
Published: (2024)
by: Torbunov, Dmitrii, et al.
Published: (2024)
AMSbench: A Comprehensive Benchmark for Evaluating MLLM Capabilities in AMS Circuits
by: Shi, Yichen, et al.
Published: (2025)
by: Shi, Yichen, et al.
Published: (2025)
Unpaired Image Translation to Mitigate Domain Shift in Liquid Argon Time Projection Chamber Detector Responses
by: Huang, Yi, et al.
Published: (2023)
by: Huang, Yi, et al.
Published: (2023)
Diffusion Model-based Parameter Estimation in Dynamic Power Systems
by: Zhu, Feiqin, et al.
Published: (2024)
by: Zhu, Feiqin, et al.
Published: (2024)
IE2Video: Adapting Pretrained Diffusion Models for Event-Based Video Reconstruction
by: Torbunov, Dmitrii, et al.
Published: (2025)
by: Torbunov, Dmitrii, et al.
Published: (2025)
Beyond Overall Accuracy: Pose- and Occlusion-driven Fairness Analysis in Pedestrian Detection for Autonomous Driving
by: Khoshkdahan, Mohammad, et al.
Published: (2025)
by: Khoshkdahan, Mohammad, et al.
Published: (2025)
FinReasoning: A Hierarchical Benchmark for Reliable Financial Research Reporting
by: Zhu, Yiyun, et al.
Published: (2026)
by: Zhu, Yiyun, et al.
Published: (2026)
PhyGround: Benchmarking Physical Reasoning in Generative World Models
by: Lin, Juyi, et al.
Published: (2026)
by: Lin, Juyi, et al.
Published: (2026)
Spatial-MLLM: Boosting MLLM Capabilities in Visual-based Spatial Intelligence
by: Wu, Diankun, et al.
Published: (2025)
by: Wu, Diankun, et al.
Published: (2025)
Cross-Platform Scaling of Vision-Language-Action Models from Edge to Cloud GPUs
by: Taherin, Amir, et al.
Published: (2025)
by: Taherin, Amir, et al.
Published: (2025)
Effectiveness of denoising diffusion probabilistic models for fast and high-fidelity whole-event simulation in high-energy heavy-ion experiments
by: Go, Yeonju, et al.
Published: (2024)
by: Go, Yeonju, et al.
Published: (2024)
Engineering Anisotropic Rabi Model in Circuit QED
by: Tabatabaei, S. Mojtaba, et al.
Published: (2025)
by: Tabatabaei, S. Mojtaba, et al.
Published: (2025)
ZeroSim: Zero-Shot Analog Circuit Evaluation with Unified Transformer Embeddings
by: Yang, Xiaomeng, et al.
Published: (2025)
by: Yang, Xiaomeng, et al.
Published: (2025)
Symbolic Timing Analysis of Digital Circuits Using Analytic Delay Functions
by: Thaqi, Era, et al.
Published: (2025)
by: Thaqi, Era, et al.
Published: (2025)
GoT-R1: Unleashing Reasoning Capability of MLLM for Visual Generation with Reinforcement Learning
by: Duan, Chengqi, et al.
Published: (2025)
by: Duan, Chengqi, et al.
Published: (2025)
Information Density Principle for MLLM Benchmarks
by: Li, Chunyi, et al.
Published: (2025)
by: Li, Chunyi, et al.
Published: (2025)
Cross-Hierarchical Bidirectional Consistency Learning for Fine-Grained Visual Classification
by: Gao, Pengxiang, et al.
Published: (2025)
by: Gao, Pengxiang, et al.
Published: (2025)
Benchmarking MLLM-based Web Understanding: Reasoning, Robustness and Safety
by: Liu, Junliang, et al.
Published: (2025)
by: Liu, Junliang, et al.
Published: (2025)
EngTrace: A Symbolic Benchmark for Verifiable Process Supervision of Engineering Reasoning
by: Gull, Ayesha, et al.
Published: (2025)
by: Gull, Ayesha, et al.
Published: (2025)
AbductiveMLLM: Boosting Visual Abductive Reasoning Within MLLMs
by: Chang, Boyu, et al.
Published: (2026)
by: Chang, Boyu, et al.
Published: (2026)
Look-Back: Implicit Visual Re-focusing in MLLM Reasoning
by: Yang, Shuo, et al.
Published: (2025)
by: Yang, Shuo, et al.
Published: (2025)
Palette: A Modular, Controllable, and Efficient Framework for On-demand Authorized Safety Alignment Relaxation in LLMs
by: Tan, Qitao, et al.
Published: (2026)
by: Tan, Qitao, et al.
Published: (2026)
Robust Symbolic Reasoning for Visual Narratives via Hierarchical and Semantically Normalized Knowledge Graphs
by: Chen, Yi-Chun
Published: (2025)
by: Chen, Yi-Chun
Published: (2025)
Robust and Generalizable Background Subtraction on Images of Calorimeter Jets using Unsupervised Generative Learning
by: Go, Yeonju, et al.
Published: (2025)
by: Go, Yeonju, et al.
Published: (2025)
MDF-MLLM: Deep Fusion Through Cross-Modal Feature Alignment for Contextually Aware Fundoscopic Image Classification
by: Jordan, Jason, et al.
Published: (2025)
by: Jordan, Jason, et al.
Published: (2025)
MathSticks: A Benchmark for Visual Symbolic Compositional Reasoning with Matchstick Puzzles
by: Ji, Yuheng, et al.
Published: (2025)
by: Ji, Yuheng, et al.
Published: (2025)
OPENXRD: A Comprehensive Benchmark Framework for LLM/MLLM XRD Question Answering
by: Vosoughi, Ali, et al.
Published: (2025)
by: Vosoughi, Ali, et al.
Published: (2025)
MLLM-CompBench: A Comparative Reasoning Benchmark for Multimodal LLMs
by: Kil, Jihyung, et al.
Published: (2024)
by: Kil, Jihyung, et al.
Published: (2024)
Aligning MLLM Benchmark With Human Preferences via Structural Equation Modeling
by: Xiong, Shengwu., et al.
Published: (2025)
by: Xiong, Shengwu., et al.
Published: (2025)
Saliency-Bench: A Comprehensive Benchmark for Evaluating Visual Explanations
by: Zhang, Yifei, et al.
Published: (2023)
by: Zhang, Yifei, et al.
Published: (2023)
FinDocMRE: A Benchmark for Document-Level Financial Multimodal Reasoning Evaluation
by: Zhu, Jiayong, et al.
Published: (2026)
by: Zhu, Jiayong, et al.
Published: (2026)
R-Bench: Graduate-level Multi-disciplinary Benchmarks for LLM & MLLM Complex Reasoning Evaluation
by: Guo, Meng-Hao, et al.
Published: (2025)
by: Guo, Meng-Hao, et al.
Published: (2025)
Ref-Adv: Exploring MLLM Visual Reasoning in Referring Expression Tasks
by: Dong, Qihua, et al.
Published: (2026)
by: Dong, Qihua, et al.
Published: (2026)
DesignBench: A Comprehensive Benchmark for MLLM-based Front-end Code Generation
by: Xiao, Jingyu, et al.
Published: (2025)
by: Xiao, Jingyu, et al.
Published: (2025)
Beyond Isolated Capabilities: Bridging Long CoT Reasoning and Long-Context Understanding
by: Wang, Yifei
Published: (2025)
by: Wang, Yifei
Published: (2025)
MLLM-CTBench: A Benchmark for Continual Instruction Tuning with Reasoning Process Diagnosis
by: Guo, Haiyun, et al.
Published: (2025)
by: Guo, Haiyun, et al.
Published: (2025)
RAGs to Riches: RAG-like Few-shot Learning for Large Language Model Role-playing
by: Rupprecht, Timothy, et al.
Published: (2025)
by: Rupprecht, Timothy, et al.
Published: (2025)
Improving Audio-Text Retrieval via Hierarchical Cross-Modal Interaction and Auxiliary Captions
by: Xin, Yifei, et al.
Published: (2023)
by: Xin, Yifei, et al.
Published: (2023)
Similar Items
-
AutoSizer: Automatic Sizing of Analog and Mixed-Signal Circuits via Large Language Model (LLM) Agents
by: Yu, Xi, et al.
Published: (2026) -
Parameter Inference and Uncertainty Quantification with Diffusion Models: Extending CDI to 2D Spatial Conditioning
by: Torbunov, Dmitrii, et al.
Published: (2026) -
EvRT-DETR: Latent Space Adaptation of Image Detectors for Event-based Vision
by: Torbunov, Dmitrii, et al.
Published: (2024) -
AMSbench: A Comprehensive Benchmark for Evaluating MLLM Capabilities in AMS Circuits
by: Shi, Yichen, et al.
Published: (2025) -
Unpaired Image Translation to Mitigate Domain Shift in Liquid Argon Time Projection Chamber Detector Responses
by: Huang, Yi, et al.
Published: (2023)