Reflection Pretraining Enables Token-Level Self-Correction in Biological Sequence Models
Fuente:
arXiv
Saved in:
| Main Authors: | Zhang, Xiang, Wei, Jiaqi, Yang, Yuejin, Qiu, Zijie, Chen, Yuhan, Gao, Zhiqiang, Abdul-Mageed, Muhammad, Lakshmanan, Laks V. S., Ouyang, Wanli, You, Chenyu, Sun, Siqi |
|---|---|
| Format: | Preprint |
| Published: |
2025
|
| Subjects: | |
| Online Access: | |
| Tags: |
Add Tag
No Tags, Be the first to tag this record!
|
Similar Items
Unifying Tree Search Algorithm and Reward Design for LLM Reasoning: A Survey
by: Wei, Jiaqi, et al.
Published: (2025)
by: Wei, Jiaqi, et al.
Published: (2025)
Autoregressive + Chain of Thought = Recurrent: Recurrence's Role in Language Models' Computability and a Revisit of Recurrent Transformer
by: Zhang, Xiang, et al.
Published: (2024)
by: Zhang, Xiang, et al.
Published: (2024)
DetoxLLM: A Framework for Detoxification with Explanations
by: Khondaker, Md Tawkat Islam, et al.
Published: (2024)
by: Khondaker, Md Tawkat Islam, et al.
Published: (2024)
LLM Performance Predictors are good initializers for Architecture Search
by: Jawahar, Ganesh, et al.
Published: (2023)
by: Jawahar, Ganesh, et al.
Published: (2023)
Curriculum Learning for Biological Sequence Prediction: The Case of De Novo Peptide Sequencing
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
Universal Biological Sequence Reranking for Improved De Novo Peptide Sequencing
by: Qiu, Zijie, et al.
Published: (2025)
by: Qiu, Zijie, et al.
Published: (2025)
Cross-Modal Consistency in Multimodal Large Language Models
by: Zhang, Xiang, et al.
Published: (2024)
by: Zhang, Xiang, et al.
Published: (2024)
Bidirectional Representations Augmented Autoregressive Biological Sequence Generation
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
When to Think, When to Speak: Learning Disclosure Policies for LLM Reasoning
by: Wei, Jiaqi, et al.
Published: (2026)
by: Wei, Jiaqi, et al.
Published: (2026)
KRAFT: A Knowledge Graph-Based Framework for Automated Map Conflation
by: Hashemi, Farnoosh, et al.
Published: (2025)
by: Hashemi, Farnoosh, et al.
Published: (2025)
Accurate de novo sequencing of the modified proteome with OmniNovo
by: Chen, Yuhan, et al.
Published: (2025)
by: Chen, Yuhan, et al.
Published: (2025)
Retrieval is Not Enough: Enhancing RAG Reasoning through Test-Time Critique and Optimization
by: Wei, Jiaqi, et al.
Published: (2025)
by: Wei, Jiaqi, et al.
Published: (2025)
OCCAM: Towards Cost-Efficient and Accuracy-Aware Classification Inference
by: Ding, Dujian, et al.
Published: (2024)
by: Ding, Dujian, et al.
Published: (2024)
EduAdapt: A Question Answer Benchmark Dataset for Evaluating Grade-Level Adaptability in LLMs
by: Naeem, Numaan, et al.
Published: (2025)
by: Naeem, Numaan, et al.
Published: (2025)
On Barriers to Archival Audio Processing
by: Sullivan, Peter, et al.
Published: (2025)
by: Sullivan, Peter, et al.
Published: (2025)
Tokenization Constraints in LLMs: A Study of Symbolic and Arithmetic Reasoning Limits
by: Zhang, Xiang, et al.
Published: (2025)
by: Zhang, Xiang, et al.
Published: (2025)
To Distill or Not to Distill? On the Robustness of Robust Knowledge Distillation
by: Waheed, Abdul, et al.
Published: (2024)
by: Waheed, Abdul, et al.
Published: (2024)
A Multi-Agent Approach for Claim Verification from Tabular Data Documents
by: Saha, Rudra Ranajee, et al.
Published: (2026)
by: Saha, Rudra Ranajee, et al.
Published: (2026)
A Community-Based Approach for Stance Distribution and Argument Organization
by: Saha, Rudra Ranajee, et al.
Published: (2026)
by: Saha, Rudra Ranajee, et al.
Published: (2026)
Counting Ability of Large Language Models and Impact of Tokenization
by: Zhang, Xiang, et al.
Published: (2024)
by: Zhang, Xiang, et al.
Published: (2024)
Effective Self-Mining of In-Context Examples for Unsupervised Machine Translation with LLMs
by: Mekki, Abdellah El, et al.
Published: (2024)
by: Mekki, Abdellah El, et al.
Published: (2024)
From AI for Science to Agentic Science: A Survey on Autonomous Scientific Discovery
by: Wei, Jiaqi, et al.
Published: (2025)
by: Wei, Jiaqi, et al.
Published: (2025)
Model Decides How to Tokenize: Adaptive DNA Sequence Tokenization with MxDNA
by: Qiao, Lifeng, et al.
Published: (2024)
by: Qiao, Lifeng, et al.
Published: (2024)
Towards Zero-Shot Text-To-Speech for Arabic Dialects
by: Doan, Khai Duy, et al.
Published: (2024)
by: Doan, Khai Duy, et al.
Published: (2024)
A Survey of Densest Subgraph Discovery on Large Graphs
by: Luo, Wensheng, et al.
Published: (2023)
by: Luo, Wensheng, et al.
Published: (2023)
Hyperparametric Robust and Dynamic Influence Maximization
by: Saha, Arkaprava, et al.
Published: (2024)
by: Saha, Arkaprava, et al.
Published: (2024)
FinTral: A Family of GPT-4 Level Multimodal Financial Large Language Models
by: Bhatia, Gagan, et al.
Published: (2024)
by: Bhatia, Gagan, et al.
Published: (2024)
Dallah: A Dialect-Aware Multimodal Large Language Model for Arabic
by: Alwajih, Fakhraddin, et al.
Published: (2024)
by: Alwajih, Fakhraddin, et al.
Published: (2024)
uDistil-Whisper: Label-Free Data Filtering for Knowledge Distillation in Low-Data Regimes
by: Waheed, Abdul, et al.
Published: (2024)
by: Waheed, Abdul, et al.
Published: (2024)
Route Experts by Sequence, not by Token
by: Wen, Tiansheng, et al.
Published: (2025)
by: Wen, Tiansheng, et al.
Published: (2025)
Mixture-of-Supernets: Improving Weight-Sharing Supernet Training with Architecture-Routed Mixture-of-Experts
by: Jawahar, Ganesh, et al.
Published: (2023)
by: Jawahar, Ganesh, et al.
Published: (2023)
Predicting Cascading Failures with a Hyperparametric Diffusion Model
by: Xiang, Bin, et al.
Published: (2024)
by: Xiang, Bin, et al.
Published: (2024)
On Efficient Approximate Aggregate Nearest Neighbor Queries over Learned Representations
by: Wang, Carrie, et al.
Published: (2025)
by: Wang, Carrie, et al.
Published: (2025)
Fast Maximum Common Subgraph Search: A Redundancy-Reduced Backtracking Approach
by: Yu, Kaiqiang, et al.
Published: (2025)
by: Yu, Kaiqiang, et al.
Published: (2025)
Length Value Model: Scalable Value Pretraining for Token-Level Length Modeling
by: Zhang, Zhen, et al.
Published: (2026)
by: Zhang, Zhen, et al.
Published: (2026)
Levels of cadmium in seafood products
by: Lakshmanan, P.T.
Published: (1988)
by: Lakshmanan, P.T.
Published: (1988)
Toucan: Many-to-Many Translation for 150 African Language Pairs
by: Elmadany, AbdelRahim, et al.
Published: (2024)
by: Elmadany, AbdelRahim, et al.
Published: (2024)
Interplay of Machine Translation, Diacritics, and Diacritization
by: Chen, Wei-Rui, et al.
Published: (2024)
by: Chen, Wei-Rui, et al.
Published: (2024)
Zero-Shot Context-Aware ASR for Diverse Arabic Varieties
by: Talafha, Bashar, et al.
Published: (2025)
by: Talafha, Bashar, et al.
Published: (2025)
Cheetah: Natural Language Generation for 517 African Languages
by: Adebara, Ife, et al.
Published: (2024)
by: Adebara, Ife, et al.
Published: (2024)
Similar Items
-
Unifying Tree Search Algorithm and Reward Design for LLM Reasoning: A Survey
by: Wei, Jiaqi, et al.
Published: (2025) -
Autoregressive + Chain of Thought = Recurrent: Recurrence's Role in Language Models' Computability and a Revisit of Recurrent Transformer
by: Zhang, Xiang, et al.
Published: (2024) -
DetoxLLM: A Framework for Detoxification with Explanations
by: Khondaker, Md Tawkat Islam, et al.
Published: (2024) -
LLM Performance Predictors are good initializers for Architecture Search
by: Jawahar, Ganesh, et al.
Published: (2023) -
Curriculum Learning for Biological Sequence Prediction: The Case of De Novo Peptide Sequencing
by: Zhang, Xiang, et al.
Published: (2025)