SciCustom: A Framework for Custom Evaluation of Scientific Capabilities in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Gu, Yiyang, Yang, Junwei, Luo, Junyu, Yuan, Ye, Feng, Bin, Xia, Yingce, Xie, Shufang, Liu, Kaili, Wu, Bohan, Shi, Qi, Li, Haoran, Xiao, Beier, Xiao, Zhiping, Luo, Xiao, Zhang, Weizhi, Yu, Philip S., Liu, Zequn, Zhang, Ming |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Multifaceted Evaluation of Audio-Visual Capability for MLLMs: Effectiveness, Efficiency, Generalizability and Robustness
by: Zhao, Yusheng, et al.
Published: (2025)
by: Zhao, Yusheng, et al.
Published: (2025)
Deciphering Scientific Reasoning Steps from Outcome Data for Molecule Optimization
by: Liu, Zequn, et al.
Published: (2026)
by: Liu, Zequn, et al.
Published: (2026)
ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment
by: Tian, Qiuyu, et al.
Published: (2026)
by: Tian, Qiuyu, et al.
Published: (2026)
GALA: Graph Diffusion-based Alignment with Jigsaw for Source-free Domain Adaptation
by: Luo, Junyu, et al.
Published: (2024)
by: Luo, Junyu, et al.
Published: (2024)
PolyCF: Towards the Optimal Spectral Graph Filters for Collaborative Filtering
by: Qin, Yifang, et al.
Published: (2024)
by: Qin, Yifang, et al.
Published: (2024)
MARCO: Meta-Reflection with Cross-Referencing for Code Reasoning
by: Zhao, Yusheng, et al.
Published: (2025)
by: Zhao, Yusheng, et al.
Published: (2025)
Semi-supervised Fine-tuning for Large Language Models
by: Luo, Junyu, et al.
Published: (2024)
by: Luo, Junyu, et al.
Published: (2024)
RobustFT: Robust Supervised Fine-tuning for Large Language Models under Noisy Response
by: Luo, Junyu, et al.
Published: (2024)
by: Luo, Junyu, et al.
Published: (2024)
CustomTTT: Motion and Appearance Customized Video Generation via Test-Time Training
by: Bi, Xiuli, et al.
Published: (2024)
by: Bi, Xiuli, et al.
Published: (2024)
DREAM: Dual-Standard Semantic Homogeneity with Dynamic Optimization for Graph Learning with Label Noise
by: Zhao, Yusheng, et al.
Published: (2026)
by: Zhao, Yusheng, et al.
Published: (2026)
Dynamic Bundling with Large Language Models for Zero-Shot Inference on Text-Attributed Graphs
by: Zhao, Yusheng, et al.
Published: (2025)
by: Zhao, Yusheng, et al.
Published: (2025)
Towards Graph Contrastive Learning: A Survey and Beyond
by: Ju, Wei, et al.
Published: (2024)
by: Ju, Wei, et al.
Published: (2024)
SMI-Editor: Edit-based SMILES Language Model with Fragment-level Supervision
by: Zheng, Kangjie, et al.
Published: (2024)
by: Zheng, Kangjie, et al.
Published: (2024)
ExLM: Rethinking the Impact of [MASK] Tokens in Masked Language Models
by: Zheng, Kangjie, et al.
Published: (2025)
by: Zheng, Kangjie, et al.
Published: (2025)
Rank and Align: Towards Effective Source-free Graph Domain Adaptation
by: Luo, Junyu, et al.
Published: (2024)
by: Luo, Junyu, et al.
Published: (2024)
Attention Bootstrapping for Multi-Modal Test-Time Adaptation
by: Zhao, Yusheng, et al.
Published: (2025)
by: Zhao, Yusheng, et al.
Published: (2025)
EduPlanner: LLM-Based Multi-Agent Systems for Customized and Intelligent Instructional Design
by: Zhang, Xueqiao, et al.
Published: (2025)
by: Zhang, Xueqiao, et al.
Published: (2025)
Large Language Model Agent: A Survey on Methodology, Applications and Challenges
by: Luo, Junyu, et al.
Published: (2025)
by: Luo, Junyu, et al.
Published: (2025)
Customized 3D Hierarchical Patterning for Information Encryption
by: Chong Chen, et al.
Published: (2025)
by: Chong Chen, et al.
Published: (2025)
Cross-Domain Diffusion with Progressive Alignment for Efficient Adaptive Retrieval
by: Luo, Junyu, et al.
Published: (2025)
by: Luo, Junyu, et al.
Published: (2025)
CoPA: Benchmarking Personalized Question Answering with Data-Informed Cognitive Factors
by: Su, Hang, et al.
Published: (2026)
by: Su, Hang, et al.
Published: (2026)
A Comprehensive Survey on Deep Graph Representation Learning
by: Ju, Wei, et al.
Published: (2023)
by: Ju, Wei, et al.
Published: (2023)
Sparse Causal Discovery with Generative Intervention for Unsupervised Graph Domain Adaptation
by: Luo, Junyu, et al.
Published: (2025)
by: Luo, Junyu, et al.
Published: (2025)
A Survey on Efficient Large Language Model Training: From Data-centric Perspectives
by: Luo, Junyu, et al.
Published: (2025)
by: Luo, Junyu, et al.
Published: (2025)
PGODE: Towards High-quality System Dynamics Modeling
by: Luo, Xiao, et al.
Published: (2023)
by: Luo, Xiao, et al.
Published: (2023)
Hypergraph-enhanced Dual Semi-supervised Graph Classification
by: Ju, Wei, et al.
Published: (2024)
by: Ju, Wei, et al.
Published: (2024)
Customized FinGPT Search Agents Using Foundation Models
by: Tian, Felix, et al.
Published: (2024)
by: Tian, Felix, et al.
Published: (2024)
Fast Inference of Removal-Based Node Influence
by: Li, Weikai, et al.
Published: (2024)
by: Li, Weikai, et al.
Published: (2024)
Event-Customized Image Generation
by: Wang, Zhen, et al.
Published: (2024)
by: Wang, Zhen, et al.
Published: (2024)
SciRerankBench: Benchmarking Rerankers Towards Scientific Retrieval-Augmented Generated LLMs
by: Chen, Haotian, et al.
Published: (2025)
by: Chen, Haotian, et al.
Published: (2025)
Graph Neural Networks in Intelligent Transportation Systems: Advances, Applications and Trends
by: Li, Hourun, et al.
Published: (2024)
by: Li, Hourun, et al.
Published: (2024)
A Survey of Data-Efficient Graph Learning
by: Ju, Wei, et al.
Published: (2024)
by: Ju, Wei, et al.
Published: (2024)
COMCAT: Towards Efficient Compression and Customization of Attention-Based Vision Models
by: Xiao, Jinqi, et al.
Published: (2023)
by: Xiao, Jinqi, et al.
Published: (2023)
CustomSketching: Sketch Concept Extraction for Sketch-based Image Synthesis and Editing
by: Xiao, Chufeng, et al.
Published: (2024)
by: Xiao, Chufeng, et al.
Published: (2024)
CustomSketching: Sketch Concept Extraction for Sketch‐based Image Synthesis and Editing
by: Chufeng Xiao, et al.
Published: (2024)
by: Chufeng Xiao, et al.
Published: (2024)
Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning
by: Su, Xuerui, et al.
Published: (2025)
by: Su, Xuerui, et al.
Published: (2025)
LLM-Friendly Knowledge Representation for Customer Support
by: Su, Hanchen, et al.
Published: (2025)
by: Su, Hanchen, et al.
Published: (2025)
Tora2: Motion and Appearance Customized Diffusion Transformer for Multi-Entity Video Generation
by: Zhang, Zhenghao, et al.
Published: (2025)
by: Zhang, Zhenghao, et al.
Published: (2025)
Embracing Large Language Models in Traffic Flow Forecasting
by: Zhao, Yusheng, et al.
Published: (2024)
by: Zhao, Yusheng, et al.
Published: (2024)
Towards Data-efficient Customer Intent Recognition with Prompt-based Learning Paradigm
by: Luo, Hengyu, et al.
Published: (2023)
by: Luo, Hengyu, et al.
Published: (2023)
Similar Items
-
Multifaceted Evaluation of Audio-Visual Capability for MLLMs: Effectiveness, Efficiency, Generalizability and Robustness
by: Zhao, Yusheng, et al.
Published: (2025) -
Deciphering Scientific Reasoning Steps from Outcome Data for Molecule Optimization
by: Liu, Zequn, et al.
Published: (2026) -
ForeSci: Evaluating LLM Agents for Forward-Looking AI Research Judgment
by: Tian, Qiuyu, et al.
Published: (2026) -
GALA: Graph Diffusion-based Alignment with Jigsaw for Source-free Domain Adaptation
by: Luo, Junyu, et al.
Published: (2024) -
PolyCF: Towards the Optimal Spectral Graph Filters for Collaborative Filtering
by: Qin, Yifang, et al.
Published: (2024)