CrystalReasoner: Reasoning and RL for Property-Conditioned Crystal Structure Generation
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Yuyang, Falletta, Stefano, McGrath, Delia, Yang, Sherry |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Multiple Realizability and the Rise of Deep Learning
von: McGrath, Sam Whitman, et al.
Veröffentlicht: (2024)
von: McGrath, Sam Whitman, et al.
Veröffentlicht: (2024)
DialogueReason: Rule-Based RL Sparks Dialogue Reasoning in LLMs
von: Shu, Yubo, et al.
Veröffentlicht: (2025)
von: Shu, Yubo, et al.
Veröffentlicht: (2025)
Beyond Structure: Invariant Crystal Property Prediction with Pseudo-Particle Ray Diffraction
von: Cao, Bin, et al.
Veröffentlicht: (2025)
von: Cao, Bin, et al.
Veröffentlicht: (2025)
Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026)
Continuum-Interaction-Driven Intelligence: Human-Aligned Neural Architecture via Crystallized Reasoning and Fluid Generation
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)
From Chains to Graphs: Self-Structured Reasoning for General-Domain LLMs
von: Chen, Yingjian, et al.
Veröffentlicht: (2026)
von: Chen, Yingjian, et al.
Veröffentlicht: (2026)
A Measure-Theoretic Analysis of Reasoning: Structural Generalization and Approximation Limits
von: Zhang, Yuyang, et al.
Veröffentlicht: (2026)
von: Zhang, Yuyang, et al.
Veröffentlicht: (2026)
Vero: An Open RL Recipe for General Visual Reasoning
von: Sarch, Gabriel, et al.
Veröffentlicht: (2026)
von: Sarch, Gabriel, et al.
Veröffentlicht: (2026)
Large Language Models Reasoning Abilities Under Non-Ideal Conditions After RL-Fine-Tuning
von: Tian, Chang, et al.
Veröffentlicht: (2025)
von: Tian, Chang, et al.
Veröffentlicht: (2025)
Siamese Foundation Models for Crystal Structure Prediction
von: Wu, Liming, et al.
Veröffentlicht: (2025)
von: Wu, Liming, et al.
Veröffentlicht: (2025)
Shifting the Gradient: Understanding How Defensive Training Methods Protect Language Model Integrity
von: Grant, Satchel, et al.
Veröffentlicht: (2026)
von: Grant, Satchel, et al.
Veröffentlicht: (2026)
Reasoning Matters: Mitigate Hallucination in Multimodal Large Reasoning Models via Reasoning-Conditioned Preference Optimization
von: Kong, Jiawei, et al.
Veröffentlicht: (2026)
von: Kong, Jiawei, et al.
Veröffentlicht: (2026)
Token-Efficient RL for LLM Reasoning
von: Lee, Alan, et al.
Veröffentlicht: (2025)
von: Lee, Alan, et al.
Veröffentlicht: (2025)
RL for Reasoning by Adaptively Revealing Rationales
von: Amani, Mohammad Hossein, et al.
Veröffentlicht: (2025)
von: Amani, Mohammad Hossein, et al.
Veröffentlicht: (2025)
Reasoning Core: A Scalable RL Environment for LLM Symbolic Reasoning
von: Lacombe, Valentin, et al.
Veröffentlicht: (2025)
von: Lacombe, Valentin, et al.
Veröffentlicht: (2025)
TRON: Targeted Rule-Verifiable Online Environments for Visual Reasoning RL
von: Yang, Tianze, et al.
Veröffentlicht: (2026)
von: Yang, Tianze, et al.
Veröffentlicht: (2026)
The Challenge of Teaching Reasoning to LLMs Without RL or Distillation
von: Du, Wei, et al.
Veröffentlicht: (2025)
von: Du, Wei, et al.
Veröffentlicht: (2025)
rePIRL: Learn PRM with Inverse RL for LLM Reasoning
von: Wu, Xian, et al.
Veröffentlicht: (2026)
von: Wu, Xian, et al.
Veröffentlicht: (2026)
Why Distillation can Outperform Zero-RL: The Role of Flexible Reasoning
von: Hu, Xiao, et al.
Veröffentlicht: (2025)
von: Hu, Xiao, et al.
Veröffentlicht: (2025)
The Reasoning Error About Reasoning: Why Different Types of Reasoning Require Different Representational Structures
von: Wu, Yiling
Veröffentlicht: (2026)
von: Wu, Yiling
Veröffentlicht: (2026)
Topology of Reasoning: Understanding Large Reasoning Models through Reasoning Graph Properties
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
von: Minegishi, Gouki, et al.
Veröffentlicht: (2025)
CLORE: Content-Level Optimization for Reasoning Efficiency
von: Wu, Yuyang, et al.
Veröffentlicht: (2026)
von: Wu, Yuyang, et al.
Veröffentlicht: (2026)
RL Tango: Reinforcing Generator and Verifier Together for Language Reasoning
von: Zha, Kaiwen, et al.
Veröffentlicht: (2025)
von: Zha, Kaiwen, et al.
Veröffentlicht: (2025)
Know What You Know: Metacognitive Entropy Calibration for Verifiable RL Reasoning
von: Zhao, Qiannian, et al.
Veröffentlicht: (2026)
von: Zhao, Qiannian, et al.
Veröffentlicht: (2026)
Symmetry-Driven Generation of Crystal Structures from Composition
von: Yin, Shi, et al.
Veröffentlicht: (2026)
von: Yin, Shi, et al.
Veröffentlicht: (2026)
How critically can an AI think? A framework for evaluating the quality of thinking of generative artificial intelligence
von: Zaphir, Luke, et al.
Veröffentlicht: (2024)
von: Zaphir, Luke, et al.
Veröffentlicht: (2024)
KnowRL: Boosting LLM Reasoning via Reinforcement Learning with Minimal-Sufficient Knowledge Guidance
von: Yu, Linhao, et al.
Veröffentlicht: (2026)
von: Yu, Linhao, et al.
Veröffentlicht: (2026)
Multimodal Crystal Flow: Any-to-Any Modality Generation for Unified Crystal Modeling
von: Seong, Kiyoung, et al.
Veröffentlicht: (2026)
von: Seong, Kiyoung, et al.
Veröffentlicht: (2026)
Transformer-Enhanced Variational Autoencoder for Crystal Structure Prediction
von: Chen, Ziyi, et al.
Veröffentlicht: (2025)
von: Chen, Ziyi, et al.
Veröffentlicht: (2025)
Enhancing Analogical Reasoning in the Abstraction and Reasoning Corpus via Model-Based RL
von: Lee, Jihwan, et al.
Veröffentlicht: (2024)
von: Lee, Jihwan, et al.
Veröffentlicht: (2024)
AceReason-Nemotron 1.1: Advancing Math and Code Reasoning through SFT and RL Synergy
von: Liu, Zihan, et al.
Veröffentlicht: (2025)
von: Liu, Zihan, et al.
Veröffentlicht: (2025)
Reinforced MLLM: A Survey on RL-Based Reasoning in Multimodal Large Language Models
von: Zhou, Guanghao, et al.
Veröffentlicht: (2025)
von: Zhou, Guanghao, et al.
Veröffentlicht: (2025)
MarsRL: Advancing Multi-Agent Reasoning System via Reinforcement Learning with Agentic Pipeline Parallelism
von: Liu, Shulin, et al.
Veröffentlicht: (2025)
von: Liu, Shulin, et al.
Veröffentlicht: (2025)
Metastable Dynamics of Chain-of-Thought Reasoning: Provable Benefits of Search, RL and Distillation
von: Kim, Juno, et al.
Veröffentlicht: (2025)
von: Kim, Juno, et al.
Veröffentlicht: (2025)
SuperRL: Reinforcement Learning with Supervision to Boost Language Model Reasoning
von: Liu, Yihao, et al.
Veröffentlicht: (2025)
von: Liu, Yihao, et al.
Veröffentlicht: (2025)
RL of Thoughts: Navigating LLM Reasoning with Inference-time Reinforcement Learning
von: Hao, Qianyue, et al.
Veröffentlicht: (2025)
von: Hao, Qianyue, et al.
Veröffentlicht: (2025)
RL Squeezes, SFT Expands: A Comparative Study of Reasoning LLMs
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2025)
von: Matsutani, Kohsei, et al.
Veröffentlicht: (2025)
Accelerating RL for LLM Reasoning with Optimal Advantage Regression
von: Brantley, Kianté, et al.
Veröffentlicht: (2025)
von: Brantley, Kianté, et al.
Veröffentlicht: (2025)
Synthetic Data Generation & Multi-Step RL for Reasoning & Tool Use
von: Goldie, Anna, et al.
Veröffentlicht: (2025)
von: Goldie, Anna, et al.
Veröffentlicht: (2025)
From Frege to chatGPT: Compositionality in language, cognition, and deep neural networks
von: Russin, Jacob, et al.
Veröffentlicht: (2024)
von: Russin, Jacob, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
Multiple Realizability and the Rise of Deep Learning
von: McGrath, Sam Whitman, et al.
Veröffentlicht: (2024) -
DialogueReason: Rule-Based RL Sparks Dialogue Reasoning in LLMs
von: Shu, Yubo, et al.
Veröffentlicht: (2025) -
Beyond Structure: Invariant Crystal Property Prediction with Pseudo-Particle Ray Diffraction
von: Cao, Bin, et al.
Veröffentlicht: (2025) -
Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning
von: Bhattarai, Manish, et al.
Veröffentlicht: (2026) -
Continuum-Interaction-Driven Intelligence: Human-Aligned Neural Architecture via Crystallized Reasoning and Fluid Generation
von: Zhou, Pengcheng, et al.
Veröffentlicht: (2025)