Lost in Tokenization: Context as the Key to Unlocking Biomolecular Understanding in Scientific LLMs
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhuang, Kai, Zhang, Jiawei, Liu, Yumou, Cao, Hanqun, Gu, Chunbin, Liu, Mengdi, Gao, Zhangyang, Wang, Zitong Jerry, Zhou, Xuanhe, Heng, Pheng-Ann, Wu, Lijun, He, Conghui, Tan, Cheng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Unlocking Potential Binders: Multimodal Pretraining DEL-Fusion for Denoising DNA-Encoded Libraries
von: Gu, Chunbin, et al.
Veröffentlicht: (2024)
von: Gu, Chunbin, et al.
Veröffentlicht: (2024)
CONFIDE: Hallucination Assessment for Reliable Biomolecular Structure Prediction and Design
von: Gao, Zijun, et al.
Veröffentlicht: (2025)
von: Gao, Zijun, et al.
Veröffentlicht: (2025)
DEL-Ranking: Ranking-Correction Denoising Framework for Elucidating Molecular Affinities in DNA-Encoded Libraries
von: Cao, Hanqun, et al.
Veröffentlicht: (2024)
von: Cao, Hanqun, et al.
Veröffentlicht: (2024)
CA-DEL: An Open Multi-Target, Multi-Modal Benchmark for Learning from DNA-Encoded Library Screens
von: He, Mutian, et al.
Veröffentlicht: (2026)
von: He, Mutian, et al.
Veröffentlicht: (2026)
Does Engram Do Memory Retrieval in Autoregressive Image Generation?
von: Wang, Jinghao, et al.
Veröffentlicht: (2026)
von: Wang, Jinghao, et al.
Veröffentlicht: (2026)
Lightweight MSA Design Advances Protein Folding From Evolutionary Embeddings
von: Cao, Hanqun, et al.
Veröffentlicht: (2025)
von: Cao, Hanqun, et al.
Veröffentlicht: (2025)
RiboSphere: Learning Unified and Efficient Representations of RNA Structures
von: Zhang, Zhou, et al.
Veröffentlicht: (2026)
von: Zhang, Zhou, et al.
Veröffentlicht: (2026)
Learning the PTM Code through a Coarse-to-Fine, Mechanism-Aware Framework
von: Zhang, Jingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Jingjie, et al.
Veröffentlicht: (2025)
Bi-TEAM: A Unified Cross-Scale Representation Learning Framework for Chemically Modified Biomolecules
von: Gu, Chunbin, et al.
Veröffentlicht: (2026)
von: Gu, Chunbin, et al.
Veröffentlicht: (2026)
SAGEPhos: Sage Bio-Coupled and Augmented Fusion for Phosphorylation Site Detection
von: Zhang, Jingjie, et al.
Veröffentlicht: (2025)
von: Zhang, Jingjie, et al.
Veröffentlicht: (2025)
TD3B: Transition-Directed Discrete Diffusion for Allosteric Binder Generation
von: Cao, Hanqun, et al.
Veröffentlicht: (2026)
von: Cao, Hanqun, et al.
Veröffentlicht: (2026)
Unlocking Positive Transfer in Incrementally Learning Surgical Instruments: A Self-reflection Hierarchical Prompt Framework
von: Zhu, Yu, et al.
Veröffentlicht: (2026)
von: Zhu, Yu, et al.
Veröffentlicht: (2026)
A deep reinforcement learning platform for antibiotic discovery
von: Cao, Hanqun, et al.
Veröffentlicht: (2025)
von: Cao, Hanqun, et al.
Veröffentlicht: (2025)
GeoCycler: Reward-Aligned 3D Diffusion for Constraint-Conditioned Cyclic Peptide Design
von: Zhang, Jingjie, et al.
Veröffentlicht: (2026)
von: Zhang, Jingjie, et al.
Veröffentlicht: (2026)
Creating Virtual Environments with 3D Gaussian Splatting: A Comparative Study
von: Qiu, Shi, et al.
Veröffentlicht: (2025)
von: Qiu, Shi, et al.
Veröffentlicht: (2025)
Advancing Extended Reality with 3D Gaussian Splatting: Innovations and Prospects
von: Qiu, Shi, et al.
Veröffentlicht: (2024)
von: Qiu, Shi, et al.
Veröffentlicht: (2024)
ATLAS: Agentic or Latent Visual Reasoning? One Word is Enough for Both
von: Guo, Ziyu, et al.
Veröffentlicht: (2026)
von: Guo, Ziyu, et al.
Veröffentlicht: (2026)
ODesign: A World Model for Biomolecular Interaction Design
von: Zhang, Odin, et al.
Veröffentlicht: (2025)
von: Zhang, Odin, et al.
Veröffentlicht: (2025)
ChemMiner: A Large Language Model Agent System for Chemical Literature Data Mining
von: Chen, Kexin, et al.
Veröffentlicht: (2024)
von: Chen, Kexin, et al.
Veröffentlicht: (2024)
SciTS: Scientific Time Series Understanding and Generation with LLMs
von: Wu, Wen, et al.
Veröffentlicht: (2025)
von: Wu, Wen, et al.
Veröffentlicht: (2025)
SciVerse: Unveiling the Knowledge Comprehension and Visual Reasoning of LMMs on Multi-modal Scientific Problems
von: Guo, Ziyu, et al.
Veröffentlicht: (2025)
von: Guo, Ziyu, et al.
Veröffentlicht: (2025)
Point Cloud Understanding via Attention-Driven Contrastive Learning
von: Wang, Yi, et al.
Veröffentlicht: (2024)
von: Wang, Yi, et al.
Veröffentlicht: (2024)
MM-Mixing: Multi-Modal Mixing Alignment for 3D Understanding
von: Wang, Jiaze, et al.
Veröffentlicht: (2024)
von: Wang, Jiaze, et al.
Veröffentlicht: (2024)
CURIE: Evaluating LLMs On Multitask Scientific Long Context Understanding and Reasoning
von: Cui, Hao, et al.
Veröffentlicht: (2025)
von: Cui, Hao, et al.
Veröffentlicht: (2025)
PARM: Pipeline-Adapted Reward Model
von: Fan, Xingyu, et al.
Veröffentlicht: (2026)
von: Fan, Xingyu, et al.
Veröffentlicht: (2026)
The Dual-use Dilemma in LLMs: Do Empowering Ethical Capacities Make a Degraded Utility?
von: Zhang, Yiyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yiyi, et al.
Veröffentlicht: (2025)
Testing Properties of Edge Distributions
von: Fei, Yumou
Veröffentlicht: (2026)
von: Fei, Yumou
Veröffentlicht: (2026)
Unbounded-width CSPs are Untestable in a Sublinear Number of Queries
von: Fei, Yumou
Veröffentlicht: (2025)
von: Fei, Yumou
Veröffentlicht: (2025)
LongLLaDA: Unlocking Long Context Capabilities in Diffusion LLMs
von: Liu, Xiaoran, et al.
Veröffentlicht: (2025)
von: Liu, Xiaoran, et al.
Veröffentlicht: (2025)
Evaluating DNA function understanding in genomic language models using evolutionarily implausible sequences
von: Jiang, Shiyu, et al.
Veröffentlicht: (2025)
von: Jiang, Shiyu, et al.
Veröffentlicht: (2025)
Towards an Information Theoretic Framework of Context-Based Offline Meta-Reinforcement Learning
von: Li, Lanqing, et al.
Veröffentlicht: (2024)
von: Li, Lanqing, et al.
Veröffentlicht: (2024)
Two-State Spin Systems with Negative Interactions
von: Fei, Yumou, et al.
Veröffentlicht: (2023)
von: Fei, Yumou, et al.
Veröffentlicht: (2023)
Towards Synchronous Memorizability and Generalizability with Site-Modulated Diffusion Replay for Cross-Site Continual Segmentation
von: Xu, Dunyuan, et al.
Veröffentlicht: (2024)
von: Xu, Dunyuan, et al.
Veröffentlicht: (2024)
Unifying Physically-Informed Weather Priors in A Single Model for Image Restoration Across Multiple Adverse Weather Conditions
von: Xu, Jiaqi, et al.
Veröffentlicht: (2026)
von: Xu, Jiaqi, et al.
Veröffentlicht: (2026)
Cross-modality Guidance-aided Multi-modal Learning with Dual Attention for MRI Brain Tumor Grading
von: Xu, Dunyuan, et al.
Veröffentlicht: (2024)
von: Xu, Dunyuan, et al.
Veröffentlicht: (2024)
Comprehensive Generative Replay for Task-Incremental Segmentation with Concurrent Appearance and Semantic Forgetting
von: Li, Wei, et al.
Veröffentlicht: (2024)
von: Li, Wei, et al.
Veröffentlicht: (2024)
Memory-Efficient Prompt Tuning for Incremental Histopathology Classification
von: Zhu, Yu, et al.
Veröffentlicht: (2024)
von: Zhu, Yu, et al.
Veröffentlicht: (2024)
Lost in Time: Clock and Calendar Understanding Challenges in Multimodal LLMs
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
von: Saxena, Rohit, et al.
Veröffentlicht: (2025)
EchoInk-R1: Exploring Audio-Visual Reasoning in Multimodal LLMs via Reinforcement Learning
von: Xing, Zhenghao, et al.
Veröffentlicht: (2025)
von: Xing, Zhenghao, et al.
Veröffentlicht: (2025)
Less Data Less Tokens: Multilingual Unification Learning for Efficient Test-Time Reasoning in LLMs
von: Chen, Kang, et al.
Veröffentlicht: (2025)
von: Chen, Kang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Unlocking Potential Binders: Multimodal Pretraining DEL-Fusion for Denoising DNA-Encoded Libraries
von: Gu, Chunbin, et al.
Veröffentlicht: (2024) -
CONFIDE: Hallucination Assessment for Reliable Biomolecular Structure Prediction and Design
von: Gao, Zijun, et al.
Veröffentlicht: (2025) -
DEL-Ranking: Ranking-Correction Denoising Framework for Elucidating Molecular Affinities in DNA-Encoded Libraries
von: Cao, Hanqun, et al.
Veröffentlicht: (2024) -
CA-DEL: An Open Multi-Target, Multi-Modal Benchmark for Learning from DNA-Encoded Library Screens
von: He, Mutian, et al.
Veröffentlicht: (2026) -
Does Engram Do Memory Retrieval in Autoregressive Image Generation?
von: Wang, Jinghao, et al.
Veröffentlicht: (2026)