Stress-Testing Multimodal Foundation Models for Crystallographic Reasoning
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Polat, Can, Kurban, Hasan, Serpedin, Erchin, Kurban, Mustafa |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
QuantumCanvas: A Multimodal Benchmark for Visual Learning of Atomic Interactions
von: Polat, Can, et al.
Veröffentlicht: (2025)
von: Polat, Can, et al.
Veröffentlicht: (2025)
SCALAR: Quantifying Structural Hallucination, Consistency, and Reasoning Gaps in Materials Foundation Models
von: Polat, Can, et al.
Veröffentlicht: (2026)
von: Polat, Can, et al.
Veröffentlicht: (2026)
Beyond Atomic Geometry Representations in Materials Science: A Human-in-the-Loop Multimodal Framework
von: Polat, Can, et al.
Veröffentlicht: (2025)
von: Polat, Can, et al.
Veröffentlicht: (2025)
Understanding the Capabilities of Molecular Graph Neural Networks in Materials Science Through Multimodal Learning and Physical Context Encoding
von: Polat, Can, et al.
Veröffentlicht: (2025)
von: Polat, Can, et al.
Veröffentlicht: (2025)
How Far Can You Grow? Characterizing the Extrapolation Frontier of Graph Generative Models for Materials Science
von: Polat, Can, et al.
Veröffentlicht: (2026)
von: Polat, Can, et al.
Veröffentlicht: (2026)
C2NP: A Benchmark for Learning Scale-Dependent Geometric Invariances in 3D Materials Generation
von: Polat, Can, et al.
Veröffentlicht: (2026)
von: Polat, Can, et al.
Veröffentlicht: (2026)
IRIS: A Real-World Benchmark for Inverse Recovery and Identification of Physical Dynamic Systems from Monocular Video
von: Khanbayov, Rasul, et al.
Veröffentlicht: (2026)
von: Khanbayov, Rasul, et al.
Veröffentlicht: (2026)
Knowing When Not to Answer: Abstention-Aware Scientific Reasoning
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2026)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2026)
xChemAgents: Agentic AI for Explainable Quantum Chemistry
von: Polat, Can, et al.
Veröffentlicht: (2025)
von: Polat, Can, et al.
Veröffentlicht: (2025)
HalluVerse25: Fine-grained Multilingual Benchmark Dataset for LLM Hallucinations
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
Audit-of-Understanding: Posterior-Constrained Inference for Mathematical Reasoning in Language Models
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
Theorem-of-Thought: A Multi-Agent Framework for Abductive, Deductive, and Inductive Reasoning in Language Models
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
Evaluating Multilingual and Code-Switched Alignment in LLMs via Synthetic Natural Language Inference
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
Halluverse-M^3: A multitask multilingual benchmark for hallucination in LLMs
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2026)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2026)
4D Synchronized Fields: Motion-Language Gaussian Splatting for Temporal Scene Understanding
von: Barhdadi, Mohamed Rayan, et al.
Veröffentlicht: (2026)
von: Barhdadi, Mohamed Rayan, et al.
Veröffentlicht: (2026)
LGQ: Learning Discretization Geometry for Scalable and Stable Image Tokenization
von: Altun, Idil Bilge, et al.
Veröffentlicht: (2026)
von: Altun, Idil Bilge, et al.
Veröffentlicht: (2026)
SINdex: Semantic INconsistency Index for Hallucination Detection in LLMs
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
MORSE-500: A Programmatically Controllable Video Benchmark to Stress-Test Multimodal Reasoning
von: Cai, Zikui, et al.
Veröffentlicht: (2025)
von: Cai, Zikui, et al.
Veröffentlicht: (2025)
Implicit Multimodal Alignment: On the Generalization of Frozen LLMs to Multimodal Inputs
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
von: Shukor, Mustafa, et al.
Veröffentlicht: (2024)
Intern-S1: A Scientific Multimodal Foundation Model
von: Bai, Lei, et al.
Veröffentlicht: (2025)
von: Bai, Lei, et al.
Veröffentlicht: (2025)
Test-Time Matching: Unlocking Compositional Reasoning in Multimodal Models
von: Zhu, Yinglun, et al.
Veröffentlicht: (2025)
von: Zhu, Yinglun, et al.
Veröffentlicht: (2025)
SAFE: A Sparse Autoencoder-Based Framework for Robust Query Enrichment and Hallucination Mitigation in LLMs
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
von: Abdaljalil, Samir, et al.
Veröffentlicht: (2025)
Intern-S1-Pro: Scientific Multimodal Foundation Model at Trillion Scale
von: Zou, Yicheng, et al.
Veröffentlicht: (2026)
von: Zou, Yicheng, et al.
Veröffentlicht: (2026)
InSight-o3: Empowering Multimodal Foundation Models with Generalized Visual Search
von: Li, Kaican, et al.
Veröffentlicht: (2025)
von: Li, Kaican, et al.
Veröffentlicht: (2025)
Zero-Shot Refinement of Buildings' Segmentation Models using SAM
von: Mayladan, Ali, et al.
Veröffentlicht: (2023)
von: Mayladan, Ali, et al.
Veröffentlicht: (2023)
A Survey of Reasoning with Foundation Models
von: Sun, Jiankai, et al.
Veröffentlicht: (2023)
von: Sun, Jiankai, et al.
Veröffentlicht: (2023)
ImageChain: Advancing Sequential Image-to-Text Reasoning in Multimodal Large Language Models
von: Villegas, Danae Sánchez, et al.
Veröffentlicht: (2025)
von: Villegas, Danae Sánchez, et al.
Veröffentlicht: (2025)
C2-Evo: Co-Evolving Multimodal Data and Model for Self-Improving Reasoning
von: Chen, Xiuwei, et al.
Veröffentlicht: (2025)
von: Chen, Xiuwei, et al.
Veröffentlicht: (2025)
HEMM: Holistic Evaluation of Multimodal Foundation Models
von: Liang, Paul Pu, et al.
Veröffentlicht: (2024)
von: Liang, Paul Pu, et al.
Veröffentlicht: (2024)
Imagine while Reasoning in Space: Multimodal Visualization-of-Thought
von: Li, Chengzu, et al.
Veröffentlicht: (2025)
von: Li, Chengzu, et al.
Veröffentlicht: (2025)
Bridging the Gap Between Multimodal Foundation Models and World Models
von: He, Xuehai
Veröffentlicht: (2025)
von: He, Xuehai
Veröffentlicht: (2025)
Many-Shot In-Context Learning in Multimodal Foundation Models
von: Jiang, Yixing, et al.
Veröffentlicht: (2024)
von: Jiang, Yixing, et al.
Veröffentlicht: (2024)
MM-Verify: Enhancing Multimodal Reasoning with Chain-of-Thought Verification
von: Sun, Linzhuang, et al.
Veröffentlicht: (2025)
von: Sun, Linzhuang, et al.
Veröffentlicht: (2025)
Cephalo: Multi-Modal Vision-Language Models for Bio-Inspired Materials Analysis and Design
von: Buehler, Markus J.
Veröffentlicht: (2024)
von: Buehler, Markus J.
Veröffentlicht: (2024)
Multilingual OCR-Aware Fine-Tuning and Prompt-Guided Chain-of-Thought Reasoning for Multimodal Large Language Models
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
von: Xu, Qinwu, et al.
Veröffentlicht: (2026)
SceneAlign: Aligning Multimodal Reasoning to Scene Graphs in Complex Visual Scenes
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
von: Wang, Chuhan, et al.
Veröffentlicht: (2026)
A Concept-Based Explainability Framework for Large Multimodal Models
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024)
von: Parekh, Jayneel, et al.
Veröffentlicht: (2024)
11Plus-Bench: Demystifying Multimodal LLM Spatial Reasoning with Cognitive-Inspired Analysis
von: Li, Chengzu, et al.
Veröffentlicht: (2025)
von: Li, Chengzu, et al.
Veröffentlicht: (2025)
GSR-BENCH: A Benchmark for Grounded Spatial Reasoning Evaluation via Multimodal LLMs
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
von: Rajabi, Navid, et al.
Veröffentlicht: (2024)
MathVista: Evaluating Mathematical Reasoning of Foundation Models in Visual Contexts
von: Lu, Pan, et al.
Veröffentlicht: (2023)
von: Lu, Pan, et al.
Veröffentlicht: (2023)
Ähnliche Einträge
-
QuantumCanvas: A Multimodal Benchmark for Visual Learning of Atomic Interactions
von: Polat, Can, et al.
Veröffentlicht: (2025) -
SCALAR: Quantifying Structural Hallucination, Consistency, and Reasoning Gaps in Materials Foundation Models
von: Polat, Can, et al.
Veröffentlicht: (2026) -
Beyond Atomic Geometry Representations in Materials Science: A Human-in-the-Loop Multimodal Framework
von: Polat, Can, et al.
Veröffentlicht: (2025) -
Understanding the Capabilities of Molecular Graph Neural Networks in Materials Science Through Multimodal Learning and Physical Context Encoding
von: Polat, Can, et al.
Veröffentlicht: (2025) -
How Far Can You Grow? Characterizing the Extrapolation Frontier of Graph Generative Models for Materials Science
von: Polat, Can, et al.
Veröffentlicht: (2026)