Efficient Post-Training Refinement of Latent Reasoning in Large Language Models
Fuente:
arXiv
Saved in:
| Main Authors: | Wang, Xinyuan, Wang, Dongjie, Ying, Wangyang, Bai, Haoyue, Gong, Nanxu, Dong, Sixun, Liu, Kunpeng, Fu, Yanjie |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Agentic Feature Augmentation: Unifying Selection and Generation with Teaming, Planning, and Memories
by: Gong, Nanxu, et al.
Published: (2025)
by: Gong, Nanxu, et al.
Published: (2025)
Unsupervised Feature Transformation via In-context Generation, Generator-critic LLM Agents, and Duet-play Teaming
by: Gong, Nanxu, et al.
Published: (2025)
by: Gong, Nanxu, et al.
Published: (2025)
Bridging the Domain Gap in Equation Distillation with Reinforcement Feedback
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
LLM-ML Teaming: Integrated Symbolic Decoding and Gradient Search for Valid and Stable Generative Feature Transformation
by: Wang, Xinyuan, et al.
Published: (2025)
by: Wang, Xinyuan, et al.
Published: (2025)
Sculpting Features from Noise: Reward-Guided Hierarchical Diffusion for Task-Optimal Feature Transformation
by: Gong, Nanxu, et al.
Published: (2025)
by: Gong, Nanxu, et al.
Published: (2025)
Data-Efficient Symbolic Regression via Foundation Model Distillation
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
To Think or Not To Think, That is The Question for Large Reasoning Models in Theory of Mind Tasks
by: Gong, Nanxu, et al.
Published: (2026)
by: Gong, Nanxu, et al.
Published: (2026)
Neuro-Symbolic Embedding for Short and Effective Feature Selection via Autoregressive Generation
by: Gong, Nanxu, et al.
Published: (2024)
by: Gong, Nanxu, et al.
Published: (2024)
Mitigating Shortcut Reasoning in Language Models: A Gradient-Aware Training Approach
by: Cao, Hongyu, et al.
Published: (2026)
by: Cao, Hongyu, et al.
Published: (2026)
A Survey on Data-Centric AI: Tabular Learning from Reinforcement Learning and Generative AI Perspective
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
Supply Chain Optimization via Generative Simulation and Iterative Decision Policies
by: Bai, Haoyue, et al.
Published: (2025)
by: Bai, Haoyue, et al.
Published: (2025)
Brownian Bridge Augmented Surrogate Simulation and Injection Planning for Geological CO$_2$ Storage
by: Bai, Haoyue, et al.
Published: (2025)
by: Bai, Haoyue, et al.
Published: (2025)
Towards Data-Centric AI: A Comprehensive Survey of Traditional, Reinforcement, and Generative Approaches for Tabular Data Transformation
by: Wang, Dongjie, et al.
Published: (2025)
by: Wang, Dongjie, et al.
Published: (2025)
Topology-aware Reinforcement Feature Space Reconstruction for Graph Data
by: Ying, Wangyang, et al.
Published: (2024)
by: Ying, Wangyang, et al.
Published: (2024)
Blind Spot Navigation in Large Language Model Reasoning with Thought Space Explorer
by: Zhang, Jinghan, et al.
Published: (2024)
by: Zhang, Jinghan, et al.
Published: (2024)
Evolutionary Large Language Model for Automated Feature Transformation
by: Gong, Nanxu, et al.
Published: (2024)
by: Gong, Nanxu, et al.
Published: (2024)
Distribution Shift Aware Neural Tabular Learning
by: Ying, Wangyang, et al.
Published: (2025)
by: Ying, Wangyang, et al.
Published: (2025)
Evolving Demonstration Optimization for Chain-of-Thought Feature Transformation
by: Wang, Xinyuan, et al.
Published: (2026)
by: Wang, Xinyuan, et al.
Published: (2026)
Knockoff-Guided Feature Selection via A Single Pre-trained Reinforced Agent
by: Wang, Xinyuan, et al.
Published: (2024)
by: Wang, Xinyuan, et al.
Published: (2024)
Feature Selection as Deep Sequential Generative Learning
by: Ying, Wangyang, et al.
Published: (2024)
by: Ying, Wangyang, et al.
Published: (2024)
Reinforcement Feature Transformation for Polymer Property Performance Prediction
by: Hu, Xuanming, et al.
Published: (2024)
by: Hu, Xuanming, et al.
Published: (2024)
Autonomous Data Agents: A New Opportunity for Smart Data
by: Fu, Yanjie, et al.
Published: (2025)
by: Fu, Yanjie, et al.
Published: (2025)
Rethinking Spatio-Temporal Anomaly Detection: A Vision for Causality-Driven Cybersecurity
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
RATT: A Thought Structure for Coherent and Correct LLM Reasoning
by: Zhang, Jinghan, et al.
Published: (2024)
by: Zhang, Jinghan, et al.
Published: (2024)
Incremental Causal Graph Learning for Online Cyberattack Detection in Cyber-Physical Infrastructures
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
Causal Graph Profiling via Structural Divergence for Robust Anomaly Detection in Cyber-Physical Systems
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
by: Malarkkan, Arun Vignesh, et al.
Published: (2025)
SPOT: Span-level Pause-of-Thought for Efficient and Interpretable Latent Reasoning in Large Language Models
by: Chu, Yunlong, et al.
Published: (2026)
by: Chu, Yunlong, et al.
Published: (2026)
Efficient Post-Training Pruning of Large Language Models with Statistical Correction
by: Yu, Peiqi, et al.
Published: (2026)
by: Yu, Peiqi, et al.
Published: (2026)
Training Large Language Models to Reason in a Continuous Latent Space
by: Hao, Shibo, et al.
Published: (2024)
by: Hao, Shibo, et al.
Published: (2024)
Benchmarking Post-Training Quantization of Large Language Models under Microscaling Floating Point Formats
by: Zhang, Manyi, et al.
Published: (2026)
by: Zhang, Manyi, et al.
Published: (2026)
Restoring Exploration after Post-Training: Latent Exploration Decoding for Large Reasoning Models
by: Tan, Wenhui, et al.
Published: (2026)
by: Tan, Wenhui, et al.
Published: (2026)
MixLLM: Dynamic Routing in Mixed Large Language Models
by: Wang, Xinyuan, et al.
Published: (2025)
by: Wang, Xinyuan, et al.
Published: (2025)
SeLaR: Selective Latent Reasoning in Large Language Models
by: Fu, Renyu, et al.
Published: (2026)
by: Fu, Renyu, et al.
Published: (2026)
Revolutionizing Biomarker Discovery: Leveraging Generative AI for Bio-Knowledge-Embedded Continuous Space Exploration
by: Ying, Wangyang, et al.
Published: (2024)
by: Ying, Wangyang, et al.
Published: (2024)
SSR: Socratic Self-Refine for Large Language Model Reasoning
by: Shi, Haizhou, et al.
Published: (2025)
by: Shi, Haizhou, et al.
Published: (2025)
MentraSuite: Post-Training Large Language Models for Mental Health Reasoning and Assessment
by: Xiao, Mengxi, et al.
Published: (2025)
by: Xiao, Mengxi, et al.
Published: (2025)
Dataforge: Agentic Platform for Autonomous Data Engineering
by: Wang, Xinyuan, et al.
Published: (2025)
by: Wang, Xinyuan, et al.
Published: (2025)
Stepwise Perplexity-Guided Refinement for Efficient Chain-of-Thought Reasoning in Large Language Models
by: Cui, Yingqian, et al.
Published: (2025)
by: Cui, Yingqian, et al.
Published: (2025)
Structured In-context Environment Scaling for Large Language Model Reasoning
by: Yu, Peng, et al.
Published: (2025)
by: Yu, Peng, et al.
Published: (2025)
Feature Interaction Aware Automated Data Representation Transformation
by: Azim, Ehtesamul, et al.
Published: (2023)
by: Azim, Ehtesamul, et al.
Published: (2023)
Similar Items
-
Agentic Feature Augmentation: Unifying Selection and Generation with Teaming, Planning, and Memories
by: Gong, Nanxu, et al.
Published: (2025) -
Unsupervised Feature Transformation via In-context Generation, Generator-critic LLM Agents, and Duet-play Teaming
by: Gong, Nanxu, et al.
Published: (2025) -
Bridging the Domain Gap in Equation Distillation with Reinforcement Feedback
by: Ying, Wangyang, et al.
Published: (2025) -
LLM-ML Teaming: Integrated Symbolic Decoding and Gradient Search for Valid and Stable Generative Feature Transformation
by: Wang, Xinyuan, et al.
Published: (2025) -
Sculpting Features from Noise: Reward-Guided Hierarchical Diffusion for Task-Optimal Feature Transformation
by: Gong, Nanxu, et al.
Published: (2025)