Rewarding Structural Conformance of Reasoning using Process Mining
Fuente:
arXiv
Saved in:
| Main Authors: | Lee, Yongjae, Park, Taekhyun, Sim, Sunghyun, Bae, Hyerim |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models
by: Park, Taekhyun, et al.
Published: (2026)
by: Park, Taekhyun, et al.
Published: (2026)
Correlation recurrent units: A novel neural architecture for improving the predictive performance of time-series data
by: Sim, Sunghyun, et al.
Published: (2022)
by: Sim, Sunghyun, et al.
Published: (2022)
JustDense: Just using Dense instead of Sequence Mixer for Time Series analysis
by: Park, TaekHyun, et al.
Published: (2025)
by: Park, TaekHyun, et al.
Published: (2025)
LAST-RAG: Literature-Anchored Stochastic Trajectory Retrieval-Augmented Generation for Knowledge-Conditioned Degradation Model Selection
by: Park, Hanbyeol, et al.
Published: (2026)
by: Park, Hanbyeol, et al.
Published: (2026)
Domain-Adaptive Health Indicator Learning with Degradation-Stage Synchronized Sampling and Cross-Domain Autoencoder
by: Choo, Jungho, et al.
Published: (2026)
by: Choo, Jungho, et al.
Published: (2026)
Artificial Intelligence-based Smart Port Logistics Metaverse for Enhancing Productivity, Environment, and Safety in Port Logistics: A Case Study of Busan Port
by: Sim, Sunghyun, et al.
Published: (2024)
by: Sim, Sunghyun, et al.
Published: (2024)
IConv: Focusing on Local Variation with Channel Independent Convolution for Multivariate Time Series Forecasting
by: Lee, Gawon, et al.
Published: (2025)
by: Lee, Gawon, et al.
Published: (2025)
Return Prediction for Mean-Variance Portfolio Selection: How Decision-Focused Learning Shapes Forecasting Models
by: Lee, Junhyeong, et al.
Published: (2024)
by: Lee, Junhyeong, et al.
Published: (2024)
Generative AI and Machine Learning Collaboration for Container Dwell Time Prediction via Data Standardization
by: Kim, Minseop, et al.
Published: (2026)
by: Kim, Minseop, et al.
Published: (2026)
FEATHer: Fourier-Efficient Adaptive Temporal Hierarchy Forecaster for Time-Series Forecasting
by: Lee, Jaehoon, et al.
Published: (2026)
by: Lee, Jaehoon, et al.
Published: (2026)
Process-Aware Procurement Lead Time Prediction for Shipyard Delay Mitigation
by: Lee, Yongjae, et al.
Published: (2026)
by: Lee, Yongjae, et al.
Published: (2026)
Verifiable Process Rewards for Agentic Reasoning
by: Yuan, Huining, et al.
Published: (2026)
by: Yuan, Huining, et al.
Published: (2026)
Application of Large Language Models for Container Throughput Forecasting: Incorporating Contextual Information in Port Logistics
by: Kim, Minseop, et al.
Published: (2026)
by: Kim, Minseop, et al.
Published: (2026)
Beyond Outcome Verification: Verifiable Process Reward Models for Structured Reasoning
by: Pronesti, Massimiliano, et al.
Published: (2026)
by: Pronesti, Massimiliano, et al.
Published: (2026)
Coarse-to-Fine Process Reward Modeling for Mathematical Reasoning
by: Hu, Yulan, et al.
Published: (2025)
by: Hu, Yulan, et al.
Published: (2025)
Process Reward Agents for Steering Knowledge-Intensive Reasoning
by: Sohn, Jiwoong, et al.
Published: (2026)
by: Sohn, Jiwoong, et al.
Published: (2026)
GuruAgents: Emulating Wise Investors with Prompt-Guided LLM Agents
by: Kim, Yejin, et al.
Published: (2025)
by: Kim, Yejin, et al.
Published: (2025)
Reward Sharpness-Aware Fine-Tuning for Diffusion Models
by: Kim, Kwanyoung, et al.
Published: (2026)
by: Kim, Kwanyoung, et al.
Published: (2026)
Penalizing Infeasible Actions and Reward Scaling in Reinforcement Learning with Offline Data
by: Kim, Jeonghye, et al.
Published: (2025)
by: Kim, Jeonghye, et al.
Published: (2025)
BitAbuse: A Dataset of Visually Perturbed Texts for Defending Phishing Attacks
by: Lee, Hanyong, et al.
Published: (2025)
by: Lee, Hanyong, et al.
Published: (2025)
Advancing Reasoning in Diffusion Language Models with Denoising Process Rewards
by: Xie, Shaoan, et al.
Published: (2025)
by: Xie, Shaoan, et al.
Published: (2025)
RATIONALYST: Mining Implicit Rationales for Process Supervision of Reasoning
by: Jiang, Dongwei, et al.
Published: (2024)
by: Jiang, Dongwei, et al.
Published: (2024)
Estimating Covariance for Global Minimum Variance Portfolio: A Decision-Focused Learning Approach
by: Kim, Juchan, et al.
Published: (2025)
by: Kim, Juchan, et al.
Published: (2025)
Temporal Representation Learning for Stock Similarities and Its Applications in Investment Management
by: Hwang, Yoontae, et al.
Published: (2024)
by: Hwang, Yoontae, et al.
Published: (2024)
Reinforcement Learning from Reflective Feedback (RLRF): Aligning and Improving LLMs via Fine-Grained Self-Reflection
by: Lee, Kyungjae, et al.
Published: (2024)
by: Lee, Kyungjae, et al.
Published: (2024)
When Model Meets New Normals: Test-time Adaptation for Unsupervised Time-series Anomaly Detection
by: Kim, Dongmin, et al.
Published: (2023)
by: Kim, Dongmin, et al.
Published: (2023)
INEXA: Interactive and Explainable Process Model Abstraction Through Object-Centric Process Mining
by: Benzin, Janik-Vasily, et al.
Published: (2024)
by: Benzin, Janik-Vasily, et al.
Published: (2024)
A Temporal Graph Network Framework for Dynamic Recommendation
by: Kim, Yejin, et al.
Published: (2024)
by: Kim, Yejin, et al.
Published: (2024)
Reasoning Structure Matters for Safety Alignment of Reasoning Models
by: In, Yeonjun, et al.
Published: (2026)
by: In, Yeonjun, et al.
Published: (2026)
Beyond the First Error: Process Reward Models for Reflective Mathematical Reasoning
by: Yang, Zhaohui, et al.
Published: (2025)
by: Yang, Zhaohui, et al.
Published: (2025)
YTCommentQA: Video Question Answerability in Instructional Videos
by: Yang, Saelyne, et al.
Published: (2024)
by: Yang, Saelyne, et al.
Published: (2024)
LLM Reasoning with Process Rewards for Outcome-Guided Steps
by: Rezaei, Mohammad, et al.
Published: (2026)
by: Rezaei, Mohammad, et al.
Published: (2026)
Rubric-Grounded RL: Structured Judge Rewards for Generalizable Reasoning
by: Bhattarai, Manish, et al.
Published: (2026)
by: Bhattarai, Manish, et al.
Published: (2026)
Flow Matching with Injected Noise for Offline-to-Online Reinforcement Learning
by: Shin, Yongjae, et al.
Published: (2026)
by: Shin, Yongjae, et al.
Published: (2026)
Structured Debate Improves Corporate Credit Reasoning in Financial AI
by: Lee, Yoonjin, et al.
Published: (2025)
by: Lee, Yoonjin, et al.
Published: (2025)
GeLoc3r: Enhancing Relative Camera Pose Regression with Geometric Consistency Regularization
by: Li, Jingxing, et al.
Published: (2025)
by: Li, Jingxing, et al.
Published: (2025)
LC-ERD: Mining Latent Logic for Self-Evolving Reasoning via Consistency-Regulated Reward Decomposition
by: Chen, Yanyu, et al.
Published: (2026)
by: Chen, Yanyu, et al.
Published: (2026)
StructVRM: Aligning Multimodal Reasoning with Structured and Verifiable Reward Models
by: Zhang, Xiangxiang, et al.
Published: (2025)
by: Zhang, Xiangxiang, et al.
Published: (2025)
Retrieval-Augmented Process Reward Model for Generalizable Mathematical Reasoning
by: Zhu, Jiachen, et al.
Published: (2025)
by: Zhu, Jiachen, et al.
Published: (2025)
Enhancing Analogical Reasoning in the Abstraction and Reasoning Corpus via Model-Based RL
by: Lee, Jihwan, et al.
Published: (2024)
by: Lee, Jihwan, et al.
Published: (2024)
Similar Items
-
LoopUS: Recasting Pretrained LLMs into Looped Latent Refinement Models
by: Park, Taekhyun, et al.
Published: (2026) -
Correlation recurrent units: A novel neural architecture for improving the predictive performance of time-series data
by: Sim, Sunghyun, et al.
Published: (2022) -
JustDense: Just using Dense instead of Sequence Mixer for Time Series analysis
by: Park, TaekHyun, et al.
Published: (2025) -
LAST-RAG: Literature-Anchored Stochastic Trajectory Retrieval-Augmented Generation for Knowledge-Conditioned Degradation Model Selection
by: Park, Hanbyeol, et al.
Published: (2026) -
Domain-Adaptive Health Indicator Learning with Degradation-Stage Synchronized Sampling and Cross-Domain Autoencoder
by: Choo, Jungho, et al.
Published: (2026)