Saved in:
| Main Authors: | Zhang, Yiming, Tamura, Ryo, Tsuda, Koji |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | https://arxiv.org/abs/2605.28487 |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
NIMS-OS: An automation software to implement a closed loop between artificial intelligence and robotic experiments in materials science
by: Tamura, Ryo, et al.
Published: (2023)
by: Tamura, Ryo, et al.
Published: (2023)
Exploring utilization of generative AI for research and education in data-driven materials science
by: Misawa, Takahiro, et al.
Published: (2025)
by: Misawa, Takahiro, et al.
Published: (2025)
NbBench: Benchmarking Language Models for Comprehensive Nanobody Tasks
by: Zhang, Yiming, et al.
Published: (2025)
by: Zhang, Yiming, et al.
Published: (2025)
Learning from Synthetic Data via Provenance-Based Input Gradient Guidance
by: Nagano, Koshiro, et al.
Published: (2026)
by: Nagano, Koshiro, et al.
Published: (2026)
aLLoyM: A large language model for alloy phase diagram prediction
by: Oikawa, Yuna, et al.
Published: (2025)
by: Oikawa, Yuna, et al.
Published: (2025)
Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
by: Sclar, Melanie, et al.
Published: (2024)
by: Sclar, Melanie, et al.
Published: (2024)
Self-rewarding correction for mathematical reasoning
by: Xiong, Wei, et al.
Published: (2025)
by: Xiong, Wei, et al.
Published: (2025)
CoMind: Towards Community-Driven Agents for Machine Learning Engineering
by: Li, Sijie, et al.
Published: (2025)
by: Li, Sijie, et al.
Published: (2025)
Model Provenance via Model DNA
by: Mu, Xin, et al.
Published: (2023)
by: Mu, Xin, et al.
Published: (2023)
Exploring System 1 and 2 communication for latent reasoning in LLMs
by: Coda-Forno, Julian, et al.
Published: (2025)
by: Coda-Forno, Julian, et al.
Published: (2025)
CRYSIM: Prediction of Symmetric Structures of Large Crystals with GPU-based Ising Machines
by: Liang, Chen, et al.
Published: (2025)
by: Liang, Chen, et al.
Published: (2025)
Benchmarking large language models for materials synthesis: the case of atomic layer deposition
by: Yanguas-Gil, Angel, et al.
Published: (2024)
by: Yanguas-Gil, Angel, et al.
Published: (2024)
Reinforcement Learning in hyperbolic space for multi-step reasoning
by: Xu, Tao, et al.
Published: (2025)
by: Xu, Tao, et al.
Published: (2025)
Refinement Provenance Inference: Detecting LLM-Refined Training Prompts from Model Behavior
by: Yin, Bo, et al.
Published: (2026)
by: Yin, Bo, et al.
Published: (2026)
Reinforcing privacy reasoning in LLMs via normative simulacra from fiction
by: Franchi, Matt, et al.
Published: (2026)
by: Franchi, Matt, et al.
Published: (2026)
Math Takes Two: A test for emergent mathematical reasoning in communication
by: Cooper, Michael, et al.
Published: (2026)
by: Cooper, Michael, et al.
Published: (2026)
Asymmetric Proximal Policy Optimization: mini-critics boost LLM reasoning
by: Liu, Jiashun, et al.
Published: (2025)
by: Liu, Jiashun, et al.
Published: (2025)
Replacing thinking with tool usage enables reasoning in small language models
by: Rainone, Corrado, et al.
Published: (2025)
by: Rainone, Corrado, et al.
Published: (2025)
DiPT: Enhancing LLM reasoning through diversified perspective-taking
by: Just, Hoang Anh, et al.
Published: (2024)
by: Just, Hoang Anh, et al.
Published: (2024)
TBDetector:Transformer-Based Detector for Advanced Persistent Threats with Provenance Graph
by: Wang, Nan, et al.
Published: (2023)
by: Wang, Nan, et al.
Published: (2023)
Relational decomposition for program synthesis
by: Hocquette, Céline, et al.
Published: (2024)
by: Hocquette, Céline, et al.
Published: (2024)
Deep Minds and Shallow Probes
by: Lee, Su Hyeong, et al.
Published: (2026)
by: Lee, Su Hyeong, et al.
Published: (2026)
Making deep neural networks right for the right scientific reasons by interacting with their explanations
by: Schramowski, Patrick, et al.
Published: (2020)
by: Schramowski, Patrick, et al.
Published: (2020)
Modeling Others' Minds as Code
by: Jha, Kunal, et al.
Published: (2025)
by: Jha, Kunal, et al.
Published: (2025)
When can transformers reason with abstract symbols?
by: Boix-Adsera, Enric, et al.
Published: (2023)
by: Boix-Adsera, Enric, et al.
Published: (2023)
Artificial Expert Intelligence through PAC-reasoning
by: Shalev-Shwartz, Shai, et al.
Published: (2024)
by: Shalev-Shwartz, Shai, et al.
Published: (2024)
Towards Precise Action Spotting: Addressing Temporal Misalignment in Labels with Dynamic Label Assignment
by: Tamura, Masato
Published: (2025)
by: Tamura, Masato
Published: (2025)
ReactorFold: Generative discovery of nuclear reactor cores via emergent physical reasoning
by: Lee, Yoonpyo
Published: (2025)
by: Lee, Yoonpyo
Published: (2025)
Explanova: Automatically Discover Data Insights in N \times M Table via XAI Combined LLM Workflow
by: Huang, Yiming
Published: (2026)
by: Huang, Yiming
Published: (2026)
Is continuous CoT better suited for multi-lingual reasoning?
by: Bashir, Ali Hamza, et al.
Published: (2026)
by: Bashir, Ali Hamza, et al.
Published: (2026)
Are complicated loss functions necessary for teaching LLMs to reason?
by: Carrino, Gabriele, et al.
Published: (2026)
by: Carrino, Gabriele, et al.
Published: (2026)
AI scientists produce results without reasoning scientifically
by: Ríos-García, Martiño, et al.
Published: (2026)
by: Ríos-García, Martiño, et al.
Published: (2026)
Sudoku-Bench: Evaluating creative reasoning with Sudoku variants
by: Seely, Jeffrey, et al.
Published: (2025)
by: Seely, Jeffrey, et al.
Published: (2025)
HardML: A Benchmark For Evaluating Data Science And Machine Learning knowledge and reasoning in AI
by: Pricope, Tidor-Vlad
Published: (2025)
by: Pricope, Tidor-Vlad
Published: (2025)
Trust Region Preference Approximation: A simple and stable reinforcement learning algorithm for LLM reasoning
by: Su, Xuerui, et al.
Published: (2025)
by: Su, Xuerui, et al.
Published: (2025)
yProv4ML: Effortless Provenance Tracking for Machine Learning Systems
by: Padovani, Gabriele, et al.
Published: (2025)
by: Padovani, Gabriele, et al.
Published: (2025)
Mind the Gaps: Auditing and Reducing Group Inequity in Large-Scale Mobility Prediction
by: Kumar, Ashwin, et al.
Published: (2025)
by: Kumar, Ashwin, et al.
Published: (2025)
Agentic retrieval-augmented reasoning reshapes collective reliability under model variability in radiology question answering
by: Farajiamiri, Mina, et al.
Published: (2026)
by: Farajiamiri, Mina, et al.
Published: (2026)
Causes in neuron diagrams, and testing causal reasoning in Large Language Models. A glimpse of the future of philosophy?
by: Vervoort, Louis, et al.
Published: (2025)
by: Vervoort, Louis, et al.
Published: (2025)
Illusions of reflection: open-ended task reveals systematic failures in Large Language Models' reflective reasoning
by: Weatherhead, Sion, et al.
Published: (2025)
by: Weatherhead, Sion, et al.
Published: (2025)
Similar Items
-
NIMS-OS: An automation software to implement a closed loop between artificial intelligence and robotic experiments in materials science
by: Tamura, Ryo, et al.
Published: (2023) -
Exploring utilization of generative AI for research and education in data-driven materials science
by: Misawa, Takahiro, et al.
Published: (2025) -
NbBench: Benchmarking Language Models for Comprehensive Nanobody Tasks
by: Zhang, Yiming, et al.
Published: (2025) -
Learning from Synthetic Data via Provenance-Based Input Gradient Guidance
by: Nagano, Koshiro, et al.
Published: (2026) -
aLLoyM: A large language model for alloy phase diagram prediction
by: Oikawa, Yuna, et al.
Published: (2025)