SciDesignBench: Benchmarking and Improving Language Models for Scientific Inverse Design
Fuente:
arXiv
Guardado en:
| Autores principales: | van Dijk, David, Vrkic, Ivan |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
por: Rollman, John C., et al.
Publicado: (2025)
por: Rollman, John C., et al.
Publicado: (2025)
MMSciBench: Benchmarking Language Models on Chinese Multimodal Scientific Problems
por: Ye, Xinwu, et al.
Publicado: (2025)
por: Ye, Xinwu, et al.
Publicado: (2025)
CNNs for Vis-NIR Chemometrics: From Contradiction to Conditional Design
por: Passos, Dário
Publicado: (2026)
por: Passos, Dário
Publicado: (2026)
FEM-Bench: A Structured Scientific Reasoning Benchmark for Evaluating Code-Generating LLMs
por: Mohammadzadeh, Saeed, et al.
Publicado: (2025)
por: Mohammadzadeh, Saeed, et al.
Publicado: (2025)
ASD-Bench: A Four-Axis Comprehensive Benchmark of AI Models for Autism Spectrum Disorder
por: Singh, Shubhankit, et al.
Publicado: (2026)
por: Singh, Shubhankit, et al.
Publicado: (2026)
Multi-state Protein Design with DynamicMPNN
por: Abrudan, Alex, et al.
Publicado: (2025)
por: Abrudan, Alex, et al.
Publicado: (2025)
Unsupervised Novelty Detection Methods Benchmarking with Wavelet Decomposition
por: Priarone, Ariel, et al.
Publicado: (2024)
por: Priarone, Ariel, et al.
Publicado: (2024)
CaLMFlow: Volterra Flow Matching using Causal Language Models
por: He, Sizhuang, et al.
Publicado: (2024)
por: He, Sizhuang, et al.
Publicado: (2024)
Deep Causal Inference for Point-referenced Spatial Data with Continuous Treatments
por: Jiang, Ziyang, et al.
Publicado: (2024)
por: Jiang, Ziyang, et al.
Publicado: (2024)
It's All Connected: Topology-Aware Structural Graph Encoding Improves Performance on Polymer Prediction
por: Erdogan, H. Ibrahim, et al.
Publicado: (2026)
por: Erdogan, H. Ibrahim, et al.
Publicado: (2026)
Disease Entity Recognition and Normalization is Improved with Large Language Model Derived Synthetic Normalized Mentions
por: Sasse, Kuleen, et al.
Publicado: (2024)
por: Sasse, Kuleen, et al.
Publicado: (2024)
Benchmarking the Discovery Engine
por: Foxabbott, Jack, et al.
Publicado: (2025)
por: Foxabbott, Jack, et al.
Publicado: (2025)
Benchmarking Generative AI Against Bayesian Optimization for Constrained Multi-Objective Inverse Design
por: Awan, Muhammad Bilal, et al.
Publicado: (2025)
por: Awan, Muhammad Bilal, et al.
Publicado: (2025)
Data-Driven Temperature Modelling of Machine Tools by Neural Networks: A Benchmark
por: Coelho, C., et al.
Publicado: (2025)
por: Coelho, C., et al.
Publicado: (2025)
Antibody Design and Optimization with Multi-scale Equivariant Graph Diffusion Models for Accurate Complex Antigen Binding
por: Chen, Jiameng, et al.
Publicado: (2025)
por: Chen, Jiameng, et al.
Publicado: (2025)
Comprehensive Methodology for Sample Augmentation in EEG Biomarker Studies for Alzheimers Risk Classification
por: Isaza, Veronica Henao, et al.
Publicado: (2024)
por: Isaza, Veronica Henao, et al.
Publicado: (2024)
One Set to Rule Them All: How to Obtain General Chemical Conditions via Bayesian Optimization over Curried Functions
por: Schmid, Stefan P., et al.
Publicado: (2025)
por: Schmid, Stefan P., et al.
Publicado: (2025)
Bridging Artificial Intelligence and Data Assimilation: The Data-driven Ensemble Forecasting System ClimaX-LETKF
por: Takeshima, Akira, et al.
Publicado: (2025)
por: Takeshima, Akira, et al.
Publicado: (2025)
AI enhanced data assimilation and uncertainty quantification applied to Geological Carbon Storage
por: Seabra, G. S., et al.
Publicado: (2024)
por: Seabra, G. S., et al.
Publicado: (2024)
Crust Macrofracturing as the Evidence of the Last Deglaciation
por: Aleshin, Igor, et al.
Publicado: (2022)
por: Aleshin, Igor, et al.
Publicado: (2022)
Design of an basis-projected layer for sparse datasets in deep learning training using gc-ms spectra as a case study
por: Chang, Yu Tang, et al.
Publicado: (2024)
por: Chang, Yu Tang, et al.
Publicado: (2024)
ResBench: Benchmarking LLM-Generated FPGA Designs with Resource Awareness
por: Guo, Ce, et al.
Publicado: (2025)
por: Guo, Ce, et al.
Publicado: (2025)
Triplet Feature Fusion for Equipment Anomaly Prediction : An Open-Source Methodology Using Small Foundation Models
por: Yasuno, Takato
Publicado: (2026)
por: Yasuno, Takato
Publicado: (2026)
CLIC: Contextual Language-Informed Cardiac Pathology Classification
por: Lucafo, Giovani D., et al.
Publicado: (2026)
por: Lucafo, Giovani D., et al.
Publicado: (2026)
Aligning Validation with Deployment in Spatial Prediction: Target-Weighted Cross-Validation
por: Brenning, Alexander, et al.
Publicado: (2026)
por: Brenning, Alexander, et al.
Publicado: (2026)
Data-driven models for production forecasting and decision supporting in petroleum reservoirs
por: Fernandes, Mateus A., et al.
Publicado: (2025)
por: Fernandes, Mateus A., et al.
Publicado: (2025)
BioReason: Incentivizing Multimodal Biological Reasoning within a DNA-LLM Model
por: Fallahpour, Adibvafa, et al.
Publicado: (2025)
por: Fallahpour, Adibvafa, et al.
Publicado: (2025)
SciBench: Evaluating College-Level Scientific Problem-Solving Abilities of Large Language Models
por: Wang, Xiaoxuan, et al.
Publicado: (2023)
por: Wang, Xiaoxuan, et al.
Publicado: (2023)
Model-free reinforcement learning with noisy actions for automated experimental control in optics
por: Richtmann, Lea, et al.
Publicado: (2024)
por: Richtmann, Lea, et al.
Publicado: (2024)
From Learning to Analytics: Improving Model Efficacy with Goal-Directed Client Selection
por: Tong, Jingwen, et al.
Publicado: (2024)
por: Tong, Jingwen, et al.
Publicado: (2024)
Integrating GNN and Neural ODEs for Estimating Non-Reciprocal Two-Body Interactions in Mixed-Species Collective Motion
por: Uwamichi, Masahito, et al.
Publicado: (2024)
por: Uwamichi, Masahito, et al.
Publicado: (2024)
Estimating Motor Symptom Presence and Severity in Parkinson's Disease from Wrist Accelerometer Time Series using ROCKET and InceptionTime
por: Donié, Cedric, et al.
Publicado: (2023)
por: Donié, Cedric, et al.
Publicado: (2023)
Novel Development of LLM Driven mCODE Data Model for Improved Clinical Trial Matching to Enable Standardization and Interoperability in Oncology Research
por: Shekhar, Aarsh, et al.
Publicado: (2024)
por: Shekhar, Aarsh, et al.
Publicado: (2024)
Large Language Models for Water Distribution Systems Modeling and Decision-Making
por: Goldshtein, Yinon, et al.
Publicado: (2025)
por: Goldshtein, Yinon, et al.
Publicado: (2025)
Does Dimensionality Reduction via Random Projections Preserve Landscape Features?
por: Rodríguez, Iván Olarte, et al.
Publicado: (2026)
por: Rodríguez, Iván Olarte, et al.
Publicado: (2026)
Towards Leveraging Large Language Models for Automated Medical Q&A Evaluation
por: Krolik, Jack, et al.
Publicado: (2024)
por: Krolik, Jack, et al.
Publicado: (2024)
GNN-Suite: a Graph Neural Network Benchmarking Framework for Biomedical Informatics
por: Kamp, Sebestyén, et al.
Publicado: (2025)
por: Kamp, Sebestyén, et al.
Publicado: (2025)
Continually Evolved Multimodal Foundation Models for Cancer Prognosis
por: Peng, Jie, et al.
Publicado: (2025)
por: Peng, Jie, et al.
Publicado: (2025)
Fast and Interpretable Machine Learning Modelling of Atmospheric Molecular Clusters
por: Seppäläinen, Lauri, et al.
Publicado: (2025)
por: Seppäläinen, Lauri, et al.
Publicado: (2025)
A Heterogeneous Long-Micro Scale Cascading Architecture for General Aviation Health Management
por: Chen, Xinhang, et al.
Publicado: (2026)
por: Chen, Xinhang, et al.
Publicado: (2026)
Ejemplares similares
-
Practical Design and Benchmarking of Generative AI Applications for Surgical Billing and Coding
por: Rollman, John C., et al.
Publicado: (2025) -
MMSciBench: Benchmarking Language Models on Chinese Multimodal Scientific Problems
por: Ye, Xinwu, et al.
Publicado: (2025) -
CNNs for Vis-NIR Chemometrics: From Contradiction to Conditional Design
por: Passos, Dário
Publicado: (2026) -
FEM-Bench: A Structured Scientific Reasoning Benchmark for Evaluating Code-Generating LLMs
por: Mohammadzadeh, Saeed, et al.
Publicado: (2025) -
ASD-Bench: A Four-Axis Comprehensive Benchmark of AI Models for Autism Spectrum Disorder
por: Singh, Shubhankit, et al.
Publicado: (2026)