CFDLLMBench: A Benchmark Suite for Evaluating Large Language Models in Computational Fluid Dynamics
Fuente:
arXiv
Saved in:
| Main Authors: | Somasekharan, Nithin, Yue, Ling, Cao, Yadi, Li, Weichao, Emami, Patrick, Bhargav, Pochinapeddi Sai, Acharya, Anurag, Xie, Xingyu, Pan, Shaowu |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science
by: Somasekharan, Nithin, et al.
Published: (2026)
by: Somasekharan, Nithin, et al.
Published: (2026)
Foam-Agent 2.0: An End-to-End Composable Multi-Agent Framework for Automating CFD Simulation in OpenFOAM
by: Yue, Ling, et al.
Published: (2025)
by: Yue, Ling, et al.
Published: (2025)
Beyond the Kolmogorov Barrier: A Learnable Weighted Hybrid Autoencoder for Model Order Reduction
by: Somasekharan, Nithin, et al.
Published: (2024)
by: Somasekharan, Nithin, et al.
Published: (2024)
AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents
by: Somasekharan, Nithin, et al.
Published: (2026)
by: Somasekharan, Nithin, et al.
Published: (2026)
Foam-Agent: Towards Automated Intelligent CFD Workflows
by: Yue, Ling, et al.
Published: (2025)
by: Yue, Ling, et al.
Published: (2025)
UniFoil: A Universal Dataset of Airfoils in Transitional and Turbulent Regimes for Subsonic and Transonic Flows
by: Kanchi, Rohit Sunil, et al.
Published: (2025)
by: Kanchi, Rohit Sunil, et al.
Published: (2025)
Evaluating Memory Condensation Strategies for Coding Agents in Data-Driven Scientific Discovery
by: Chintalapati, Renuka, et al.
Published: (2026)
by: Chintalapati, Renuka, et al.
Published: (2026)
GENIUS: Generative Fluid Intelligence Evaluation Suite
by: An, Ruichuan, et al.
Published: (2026)
by: An, Ruichuan, et al.
Published: (2026)
Recent Advances on Machine Learning for Computational Fluid Dynamics: A Survey
by: Wang, Haixin, et al.
Published: (2024)
by: Wang, Haixin, et al.
Published: (2024)
VICON: Vision In-Context Operator Networks for Multi-Physics Fluid Dynamics Prediction
by: Cao, Yadi, et al.
Published: (2024)
by: Cao, Yadi, et al.
Published: (2024)
Evaluating the Robustness of Dense Retrievers in Interdisciplinary Domains
by: Chaturvedi, Sarthak, et al.
Published: (2025)
by: Chaturvedi, Sarthak, et al.
Published: (2025)
Evaluation and Benchmarking Suite for Financial Large Language Models and Agents
by: Lin, Shengyuan, et al.
Published: (2026)
by: Lin, Shengyuan, et al.
Published: (2026)
NineRec: A Benchmark Dataset Suite for Evaluating Transferable Recommendation
by: Zhang, Jiaqi, et al.
Published: (2023)
by: Zhang, Jiaqi, et al.
Published: (2023)
BuildingsBench: A Large-Scale Dataset of 900K Buildings and Benchmark for Short-Term Load Forecasting
by: Emami, Patrick, et al.
Published: (2023)
by: Emami, Patrick, et al.
Published: (2023)
On the lifting and reconstruction of nonlinear systems with multiple invariant sets
by: Pan, Shaowu, et al.
Published: (2023)
by: Pan, Shaowu, et al.
Published: (2023)
WeQA: A Benchmark for Retrieval Augmented Generation in Wind Energy Domain
by: Meyur, Rounak, et al.
Published: (2024)
by: Meyur, Rounak, et al.
Published: (2024)
A SUPERB-Style Benchmark of Self-Supervised Speech Models for Audio Deepfake Detection
by: Ali, Hashim, et al.
Published: (2026)
by: Ali, Hashim, et al.
Published: (2026)
RobotPerf: An Open-Source, Vendor-Agnostic, Benchmarking Suite for Evaluating Robotics Computing System Performance
by: Mayoral-Vilches, Víctor, et al.
Published: (2023)
by: Mayoral-Vilches, Víctor, et al.
Published: (2023)
Code2MCP: Transforming Code Repositories into MCP Services
by: Ouyang, Chaoqian, et al.
Published: (2025)
by: Ouyang, Chaoqian, et al.
Published: (2025)
Learning Noise-Robust Stable Koopman Operator for Control with Hankel DMD
by: Sakib, Shahriar Akbar, et al.
Published: (2024)
by: Sakib, Shahriar Akbar, et al.
Published: (2024)
Topological Laplace Transform and Decomposition of nc-Hodge Structures
by: Yu, Tony Yue, et al.
Published: (2024)
by: Yu, Tony Yue, et al.
Published: (2024)
Benchmarking Pedestrian Dynamics Models for Common Scenarios: An Evaluation of Force-Based Models
by: Jain, Kanika, et al.
Published: (2025)
by: Jain, Kanika, et al.
Published: (2025)
Three Pathways to Neurosymbolic Reinforcement Learning with Interpretable Model and Policy Networks
by: Graf, Peter, et al.
Published: (2024)
by: Graf, Peter, et al.
Published: (2024)
CARE: a Benchmark Suite for the Classification and Retrieval of Enzymes
by: Yang, Jason, et al.
Published: (2024)
by: Yang, Jason, et al.
Published: (2024)
Generalization of Video-Based Heart Rate Estimation Methods To Low Illumination and Elevated Heart Rates
by: Acharya, Bhargav, et al.
Published: (2025)
by: Acharya, Bhargav, et al.
Published: (2025)
Computational Fluid Dynamics on Quantum Computers
by: Syamlal, Madhava, et al.
Published: (2024)
by: Syamlal, Madhava, et al.
Published: (2024)
Real-Time Dynamic Scale-Aware Fusion Detection Network: Take Road Damage Detection as an example
by: Pan, Weichao, et al.
Published: (2024)
by: Pan, Weichao, et al.
Published: (2024)
Invoice Information Extraction: Methods and Performance Evaluation
by: Yashwant, Sai, et al.
Published: (2025)
by: Yashwant, Sai, et al.
Published: (2025)
HyQBench: A Benchmark Suite for Hybrid CV-DV Quantum Computing
by: Mohapatra, Shubdeep, et al.
Published: (2026)
by: Mohapatra, Shubdeep, et al.
Published: (2026)
HEP Benchmark Suite: Enhancing Efficiency and Sustainability in Worldwide LHC Computing Infrastructures
by: Szczepanek, Natalia, et al.
Published: (2024)
by: Szczepanek, Natalia, et al.
Published: (2024)
MTR-Suite: A Framework for Evaluating and Synthesizing Conversational Retrieval Benchmarks
by: Ruan, Junhao, et al.
Published: (2026)
by: Ruan, Junhao, et al.
Published: (2026)
A Multi-AI-agent Framework Enabling End-to-end Finite Element Analysis for Solid Mechanics Problems
by: Sarker, Titu Ranjan, et al.
Published: (2026)
by: Sarker, Titu Ranjan, et al.
Published: (2026)
AeroVerse: UAV-Agent Benchmark Suite for Simulating, Pre-training, Finetuning, and Evaluating Aerospace Embodied World Models
by: Yao, Fanglong, et al.
Published: (2024)
by: Yao, Fanglong, et al.
Published: (2024)
Quantum Computation of Fluid Dynamics
by: Bharadwaj, Sachin S., et al.
Published: (2020)
by: Bharadwaj, Sachin S., et al.
Published: (2020)
CUA-Suite: Massive Human-annotated Video Demonstrations for Computer-Use Agents
by: Jian, Xiangru, et al.
Published: (2026)
by: Jian, Xiangru, et al.
Published: (2026)
BioMamba: Domain-Adaptive Biomedical Language Models
by: Yue, Ling, et al.
Published: (2024)
by: Yue, Ling, et al.
Published: (2024)
AnesSuite: A Comprehensive Benchmark and Dataset Suite for Anesthesiology Reasoning in LLMs
by: Feng, Xiang, et al.
Published: (2025)
by: Feng, Xiang, et al.
Published: (2025)
SuiteEval: Simplifying Retrieval Benchmarks
by: Parry, Andrew, et al.
Published: (2026)
by: Parry, Andrew, et al.
Published: (2026)
Preemption-Enhanced Benchmark Suite for FPGAs
by: Malik, Arsalan Ali, et al.
Published: (2025)
by: Malik, Arsalan Ali, et al.
Published: (2025)
YOLO-ROC: A High-Precision and Ultra-Lightweight Model for Real-Time Road Damage Detection
by: Lin, Zicheng, et al.
Published: (2025)
by: Lin, Zicheng, et al.
Published: (2025)
Similar Items
-
SCICONVBENCH: Benchmarking LLMs on Multi-Turn Clarification for Task Formulation in Computational Science
by: Somasekharan, Nithin, et al.
Published: (2026) -
Foam-Agent 2.0: An End-to-End Composable Multi-Agent Framework for Automating CFD Simulation in OpenFOAM
by: Yue, Ling, et al.
Published: (2025) -
Beyond the Kolmogorov Barrier: A Learnable Weighted Hybrid Autoencoder for Model Order Reduction
by: Somasekharan, Nithin, et al.
Published: (2024) -
AI CFD Scientist: Toward Open-Ended Computational Fluid Dynamics Discovery with Physics-Aware AI Agents
by: Somasekharan, Nithin, et al.
Published: (2026) -
Foam-Agent: Towards Automated Intelligent CFD Workflows
by: Yue, Ling, et al.
Published: (2025)