PepThink-R1: LLM for Interpretable Cyclic Peptide Optimization with CoT SFT and Reinforcement Learning
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Wang, Ruheng, Zhang, Hang, Nguyen, Trieu, Feng, Shasha, Pang, Hao-Wei, Yu, Xiang, Xiao, Li, Zhang, Peter Zhiping |
|---|---|
| Format: | Preprint |
| Publié: |
2025
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
PepEVOLVE: Position-Aware Dynamic Peptide Optimization via Group-Relative Advantage
par: Nguyen, Trieu, et autres
Publié: (2025)
par: Nguyen, Trieu, et autres
Publié: (2025)
ReasonGen-R1: CoT for Autoregressive Image generation models through SFT and RL
par: Zhang, Yu, et autres
Publié: (2025)
par: Zhang, Yu, et autres
Publié: (2025)
Empowering Lightweight MLLMs with Reasoning via Long CoT SFT
par: Ou, Linyu, et autres
Publié: (2025)
par: Ou, Linyu, et autres
Publié: (2025)
PepBenchmark: A Standardized Benchmark for Peptide Machine Learning
par: Zhang, Jiahui, et autres
Publié: (2026)
par: Zhang, Jiahui, et autres
Publié: (2026)
CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning
par: Gan, Zeyu, et autres
Publié: (2025)
par: Gan, Zeyu, et autres
Publié: (2025)
Target-Specific De Novo Peptide Binder Design with DiffPepBuilder
par: Wang, Fanhao, et autres
Publié: (2024)
par: Wang, Fanhao, et autres
Publié: (2024)
The Synergy Dilemma of Long-CoT SFT and RL: Investigating Post-Training Techniques for Reasoning VLMs
par: Chen, Jierun, et autres
Publié: (2025)
par: Chen, Jierun, et autres
Publié: (2025)
CreoPep: A Universal Deep Learning Framework for Target-Specific Peptide Design and Optimization
par: Ge, Cheng, et autres
Publié: (2025)
par: Ge, Cheng, et autres
Publié: (2025)
To Think or Not to Think: The Hidden Cost of Meta-Training with Excessive CoT Examples
par: Kothapalli, Vignesh, et autres
Publié: (2025)
par: Kothapalli, Vignesh, et autres
Publié: (2025)
LatentChem: From Textual CoT to Latent Thinking in Chemical Reasoning
par: Ye, Xinwu, et autres
Publié: (2026)
par: Ye, Xinwu, et autres
Publié: (2026)
CodePlot-CoT: Mathematical Visual Reasoning by Thinking with Code-Driven Images
par: Duan, Chengqi, et autres
Publié: (2025)
par: Duan, Chengqi, et autres
Publié: (2025)
An Integrated Deep-Learning Framework for Peptide-Protein Interaction Prediction and Target-Conditioned Peptide Generation with ConGA-PepPI and TC-PepGen
par: Tang, Chupei, et autres
Publié: (2026)
par: Tang, Chupei, et autres
Publié: (2026)
T2I-R1: Reinforcing Image Generation with Collaborative Semantic-level and Token-level CoT
par: Jiang, Dongzhi, et autres
Publié: (2025)
par: Jiang, Dongzhi, et autres
Publié: (2025)
CoTDeceptor:Adversarial Code Obfuscation Against CoT-Enhanced LLM Code Agents
par: Li, Haoyang, et autres
Publié: (2025)
par: Li, Haoyang, et autres
Publié: (2025)
MC-CoT: A Modular Collaborative CoT Framework for Zero-shot Medical-VQA with LLM and MLLM Integration
par: Wei, Lai, et autres
Publié: (2024)
par: Wei, Lai, et autres
Publié: (2024)
PepTune: De Novo Generation of Therapeutic Peptides with Multi-Objective-Guided Discrete Diffusion
par: Tang, Sophia, et autres
Publié: (2024)
par: Tang, Sophia, et autres
Publié: (2024)
CoT-based Synthesizer: Enhancing LLM Performance through Answer Synthesis
par: Zhang, Bohan, et autres
Publié: (2025)
par: Zhang, Bohan, et autres
Publié: (2025)
The CoT Encyclopedia: Analyzing, Predicting, and Controlling how a Reasoning Model will Think
par: Lee, Seongyun, et autres
Publié: (2025)
par: Lee, Seongyun, et autres
Publié: (2025)
VietBinoculars: A Zero-Shot Approach for Detecting Vietnamese LLM-Generated Text
par: Nguyen, Trieu Hai, et autres
Publié: (2025)
par: Nguyen, Trieu Hai, et autres
Publié: (2025)
Nash CoT: Multi-Path Inference with Preference Equilibrium
par: Zhang, Ziqi, et autres
Publié: (2024)
par: Zhang, Ziqi, et autres
Publié: (2024)
R-CoT: A Reasoning-Layer Watermark via Redundant Chain-of-Thought in Large Language Models
par: Zhang, Ziming, et autres
Publié: (2026)
par: Zhang, Ziming, et autres
Publié: (2026)
From Explicit CoT to Implicit CoT: Learning to Internalize CoT Step by Step
par: Deng, Yuntian, et autres
Publié: (2024)
par: Deng, Yuntian, et autres
Publié: (2024)
Syzygy of Thoughts: Improving LLM CoT with the Minimal Free Resolution
par: Li, Chenghao, et autres
Publié: (2025)
par: Li, Chenghao, et autres
Publié: (2025)
Performative Thinking? The Brittle Correlation Between CoT Length and Problem Complexity
par: Palod, Vardhan, et autres
Publié: (2025)
par: Palod, Vardhan, et autres
Publié: (2025)
GRASP: Grounded CoT Reasoning with Dual-Stage Optimization for Multimodal Sarcasm Target Identification
par: Wan, Faxian, et autres
Publié: (2026)
par: Wan, Faxian, et autres
Publié: (2026)
Mol-R1: Towards Explicit Long-CoT Reasoning in Molecule Discovery
par: Li, Jiatong, et autres
Publié: (2025)
par: Li, Jiatong, et autres
Publié: (2025)
Evaluating LLM Reasoning Beyond Correctness and CoT
par: Abbasloo, Soheil
Publié: (2025)
par: Abbasloo, Soheil
Publié: (2025)
LLM-CoT Enhanced Graph Neural Recommendation with Harmonized Group Policy Optimization
par: Luo, Hailong, et autres
Publié: (2025)
par: Luo, Hailong, et autres
Publié: (2025)
NS-Pep: De novo Peptide Design with Non-Standard Amino Acids
par: Guo, Tao, et autres
Publié: (2025)
par: Guo, Tao, et autres
Publié: (2025)
Privacy-preserved LLM Cascade via CoT-enhanced Policy Learning
par: Zhang, Kai, et autres
Publié: (2024)
par: Zhang, Kai, et autres
Publié: (2024)
Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation
par: Zhang, Binquan, et autres
Publié: (2025)
par: Zhang, Binquan, et autres
Publié: (2025)
CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning
par: Chen, Shuxu, et autres
Publié: (2026)
par: Chen, Shuxu, et autres
Publié: (2026)
Parrot: A Training Pipeline Enhances Both Program CoT and Natural Language CoT for Reasoning
par: Jin, Senjie, et autres
Publié: (2025)
par: Jin, Senjie, et autres
Publié: (2025)
FutureSightDrive: Thinking Visually with Spatio-Temporal CoT for Autonomous Driving
par: Zeng, Shuang, et autres
Publié: (2025)
par: Zeng, Shuang, et autres
Publié: (2025)
PepDoRA: A Unified Peptide Language Model via Weight-Decomposed Low-Rank Adaptation
par: Wang, Leyao, et autres
Publié: (2024)
par: Wang, Leyao, et autres
Publié: (2024)
Meta-CoT: Enhancing Granularity and Generalization in Image Editing
par: Zhang, Shiyi, et autres
Publié: (2026)
par: Zhang, Shiyi, et autres
Publié: (2026)
Long or short CoT? Investigating Instance-level Switch of Large Reasoning Models
par: Zhang, Ruiqi, et autres
Publié: (2025)
par: Zhang, Ruiqi, et autres
Publié: (2025)
Ada-R1: Hybrid-CoT via Bi-Level Adaptive Reasoning Optimization
par: Luo, Haotian, et autres
Publié: (2025)
par: Luo, Haotian, et autres
Publié: (2025)
PepEDiff: Zero-Shot Peptide Binder Design via Protein Embedding Diffusion
par: Liang, Po-Yu, et autres
Publié: (2026)
par: Liang, Po-Yu, et autres
Publié: (2026)
S 2 ‐ PepAnalyst : A Web Tool for Predicting Plant Small Signalling Peptides
par: Kelly L. Vomo‐Donfack, et autres
Publié: (2026)
par: Kelly L. Vomo‐Donfack, et autres
Publié: (2026)
Documents similaires
-
PepEVOLVE: Position-Aware Dynamic Peptide Optimization via Group-Relative Advantage
par: Nguyen, Trieu, et autres
Publié: (2025) -
ReasonGen-R1: CoT for Autoregressive Image generation models through SFT and RL
par: Zhang, Yu, et autres
Publié: (2025) -
Empowering Lightweight MLLMs with Reasoning via Long CoT SFT
par: Ou, Linyu, et autres
Publié: (2025) -
PepBenchmark: A Standardized Benchmark for Peptide Machine Learning
par: Zhang, Jiahui, et autres
Publié: (2026) -
CoT-Space: A Theoretical Framework for Internal Slow-Thinking via Reinforcement Learning
par: Gan, Zeyu, et autres
Publié: (2025)