BeamPERL: Parameter-Efficient RL with Verifiable Rewards Specializes Compact LLMs for Structured Beam Mechanics Reasoning
Fuente:
arXiv
Saved in:
| Main Authors: | Hage, Tarjei Paule, Buehler, Markus J. |
|---|---|
| Format: | Preprint |
| Published: |
2026
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
GraphAgents: Knowledge Graph-Guided Agentic AI for Cross-Domain Materials Design
by: Stewart, Isabella A., et al.
Published: (2026)
by: Stewart, Isabella A., et al.
Published: (2026)
Agentic Deep Graph Reasoning Yields Self-Organizing Knowledge Networks
by: Buehler, Markus J.
Published: (2025)
by: Buehler, Markus J.
Published: (2025)
Higher-Order Knowledge Representations for Agentic Scientific Reasoning
by: Stewart, Isabella A., et al.
Published: (2026)
by: Stewart, Isabella A., et al.
Published: (2026)
Accelerating Scientific Discovery with Generative Knowledge Extraction, Graph-Based Representation, and Multimodal Intelligent Graph Reasoning
by: Buehler, Markus J.
Published: (2024)
by: Buehler, Markus J.
Published: (2024)
Graph-Aware Isomorphic Attention for Adaptive Dynamics in Transformers
by: Buehler, Markus J.
Published: (2025)
by: Buehler, Markus J.
Published: (2025)
Self-Revising Discovery Systems for Science: A Categorical Framework for Agentic Artificial Intelligence
by: Wang, Fiona Y., et al.
Published: (2026)
by: Wang, Fiona Y., et al.
Published: (2026)
SciAgents: Automating scientific discovery through multi-agent intelligent graph reasoning
by: Ghafarollahi, Alireza, et al.
Published: (2024)
by: Ghafarollahi, Alireza, et al.
Published: (2024)
PRefLexOR: Preference-based Recursive Language Modeling for Exploratory Optimization of Reasoning and Agentic Thinking
by: Buehler, Markus J.
Published: (2024)
by: Buehler, Markus J.
Published: (2024)
In-situ graph reasoning and knowledge expansion using Graph-PReFLexOR
by: Buehler, Markus J.
Published: (2025)
by: Buehler, Markus J.
Published: (2025)
Fine-tuning large language models for domain adaptation: Exploration of training strategies, scaling, model merging and synergistic capabilities
by: Lu, Wei, et al.
Published: (2024)
by: Lu, Wei, et al.
Published: (2024)
Autonomous Inorganic Materials Discovery via Multi-Agent Physics-Aware Scientific Reasoning
by: Ghafarollahi, Alireza, et al.
Published: (2025)
by: Ghafarollahi, Alireza, et al.
Published: (2025)
Sparks: Multi-Agent Artificial Intelligence Model Discovers Protein Design Principles
by: Ghafarollahi, Alireza, et al.
Published: (2025)
by: Ghafarollahi, Alireza, et al.
Published: (2025)
Blockwise Advantage Estimation for Multi-Objective RL with Verifiable Rewards
by: Pavlenko, Kirill, et al.
Published: (2026)
by: Pavlenko, Kirill, et al.
Published: (2026)
Are LLMs Ready for Real-World Materials Discovery?
by: Miret, Santiago, et al.
Published: (2024)
by: Miret, Santiago, et al.
Published: (2024)
Generative Artificial Intelligence Extracts Structure-Function Relationships from Plants for New Materials
by: Luu, Rachel K., et al.
Published: (2025)
by: Luu, Rachel K., et al.
Published: (2025)
RL Tango: Reinforcing Generator and Verifier Together for Language Reasoning
by: Zha, Kaiwen, et al.
Published: (2025)
by: Zha, Kaiwen, et al.
Published: (2025)
Aligning Reasoning LLMs for Materials Discovery with Physics-aware Rejection Sampling
by: Hyun, Lee, et al.
Published: (2025)
by: Hyun, Lee, et al.
Published: (2025)
Evaluating the Performance and Robustness of LLMs in Materials Science Q&A and Property Predictions
by: Wang, Hongchen, et al.
Published: (2024)
by: Wang, Hongchen, et al.
Published: (2024)
FlowRL: Matching Reward Distributions for LLM Reasoning
by: Zhu, Xuekai, et al.
Published: (2025)
by: Zhu, Xuekai, et al.
Published: (2025)
REASONING GYM: Reasoning Environments for Reinforcement Learning with Verifiable Rewards
by: Stojanovski, Zafir, et al.
Published: (2025)
by: Stojanovski, Zafir, et al.
Published: (2025)
Mitigating Lost in Multi-turn Conversation via Curriculum RL with Verifiable Accuracy and Abstention Rewards
by: Li, Ming, et al.
Published: (2025)
by: Li, Ming, et al.
Published: (2025)
On Designing Effective RL Reward at Training Time for LLM Reasoning
by: Gao, Jiaxuan, et al.
Published: (2024)
by: Gao, Jiaxuan, et al.
Published: (2024)
Matter-of-Fact: A Benchmark for Verifying the Feasibility of Literature-Supported Claims in Materials Science
by: Jansen, Peter, et al.
Published: (2025)
by: Jansen, Peter, et al.
Published: (2025)
Decoupling Reasoning and Confidence: Resurrecting Calibration in Reinforcement Learning from Verifiable Rewards
by: Ma, Zhengzhao, et al.
Published: (2026)
by: Ma, Zhengzhao, et al.
Published: (2026)
X-LoRA: Mixture of Low-Rank Adapter Experts, a Flexible Framework for Large Language Models with Applications in Protein Mechanics and Molecular Design
by: Buehler, Eric L., et al.
Published: (2024)
by: Buehler, Eric L., et al.
Published: (2024)
Cephalo: Multi-Modal Vision-Language Models for Bio-Inspired Materials Analysis and Design
by: Buehler, Markus J.
Published: (2024)
by: Buehler, Markus J.
Published: (2024)
Verifier-Free RL for LLMs via Intrinsic Gradient-Norm Reward
by: Wen, Xuexiang, et al.
Published: (2026)
by: Wen, Xuexiang, et al.
Published: (2026)
Rewarding Graph Reasoning Process makes LLMs more Generalized Reasoners
by: Peng, Miao, et al.
Published: (2025)
by: Peng, Miao, et al.
Published: (2025)
LifeGPT: Topology-Agnostic Generative Pretrained Transformer Model for Cellular Automata
by: Berkovich, Jaime A., et al.
Published: (2024)
by: Berkovich, Jaime A., et al.
Published: (2024)
A Relative-Budget Theory for Reinforcement Learning with Verifiable Rewards in Large Language Model Reasoning
by: Wachi, Akifumi, et al.
Published: (2026)
by: Wachi, Akifumi, et al.
Published: (2026)
AgentV-RL: Scaling Reward Modeling with Agentic Verifier
by: Zhang, Jiazheng, et al.
Published: (2026)
by: Zhang, Jiazheng, et al.
Published: (2026)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
by: Liu, Yixin, et al.
Published: (2026)
by: Liu, Yixin, et al.
Published: (2026)
Hierarchical Multi-agent Large Language Model Reasoning for Autonomous Functional Materials Discovery
by: Rothfarb, Samuel, et al.
Published: (2025)
by: Rothfarb, Samuel, et al.
Published: (2025)
Reinforcement Learning with Verifiable Rewards Implicitly Incentivizes Correct Reasoning in Base LLMs
by: Wen, Xumeng, et al.
Published: (2025)
by: Wen, Xumeng, et al.
Published: (2025)
Language-Native Materials Processing Design by Lightly Structured Text Database and Reasoning Large Language Model
by: Liu, Yuze, et al.
Published: (2025)
by: Liu, Yuze, et al.
Published: (2025)
Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
by: Pronesti, Massimiliano, et al.
Published: (2026)
by: Pronesti, Massimiliano, et al.
Published: (2026)
GUI-Libra: Training Native GUI Agents to Reason and Act with Action-aware Supervision and Partially Verifiable RL
by: Yang, Rui, et al.
Published: (2026)
by: Yang, Rui, et al.
Published: (2026)
CoVerRL: Breaking the Consensus Trap in Label-Free Reasoning via Generator-Verifier Co-Evolution
by: Pan, Teng, et al.
Published: (2026)
by: Pan, Teng, et al.
Published: (2026)
Improving Neutral Point-of-View Generation with Data- and Parameter-Efficient RL
by: Hoffmann, Jessica, et al.
Published: (2025)
by: Hoffmann, Jessica, et al.
Published: (2025)
ProofSketch: Efficient Verified Reasoning for Large Language Models
by: Sheshanarayana, Disha, et al.
Published: (2025)
by: Sheshanarayana, Disha, et al.
Published: (2025)
Similar Items
-
GraphAgents: Knowledge Graph-Guided Agentic AI for Cross-Domain Materials Design
by: Stewart, Isabella A., et al.
Published: (2026) -
Agentic Deep Graph Reasoning Yields Self-Organizing Knowledge Networks
by: Buehler, Markus J.
Published: (2025) -
Higher-Order Knowledge Representations for Agentic Scientific Reasoning
by: Stewart, Isabella A., et al.
Published: (2026) -
Accelerating Scientific Discovery with Generative Knowledge Extraction, Graph-Based Representation, and Multimodal Intelligent Graph Reasoning
by: Buehler, Markus J.
Published: (2024) -
Graph-Aware Isomorphic Attention for Adaptive Dynamics in Transformers
by: Buehler, Markus J.
Published: (2025)