Bio2Token: All-atom tokenization of any biomolecular structure with Mamba
Fuente:
arXiv
Salvato in:
| Autori principali: | Liu, Andrew, Elaldi, Axel, Russell, Nathan, Viessmann, Olivia |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2024
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
Flash Invariant Point Attention
di: Liu, Andrew, et al.
Pubblicazione: (2025)
di: Liu, Andrew, et al.
Pubblicazione: (2025)
Not all tokens are needed(NAT): token efficient reinforcement learning
di: Sang, Hejian, et al.
Pubblicazione: (2026)
di: Sang, Hejian, et al.
Pubblicazione: (2026)
BioMamba: Leveraging Spectro-Temporal Embedding in Bidirectional Mamba for Enhanced Biosignal Classification
di: Qian, Jian, et al.
Pubblicazione: (2025)
di: Qian, Jian, et al.
Pubblicazione: (2025)
Improving Inverse Folding for Peptide Design with Diversity-regularized Direct Preference Optimization
di: Park, Ryan, et al.
Pubblicazione: (2024)
di: Park, Ryan, et al.
Pubblicazione: (2024)
COMPACT: Common-token Optimized Model Pruning Across Channels and Tokens
di: Kwek, Eugene, et al.
Pubblicazione: (2025)
di: Kwek, Eugene, et al.
Pubblicazione: (2025)
Physics in Next-token Prediction
di: An, Hongjun, et al.
Pubblicazione: (2024)
di: An, Hongjun, et al.
Pubblicazione: (2024)
All or None: Identifiable Linear Properties of Next-token Predictors in Language Modeling
di: Marconato, Emanuele, et al.
Pubblicazione: (2024)
di: Marconato, Emanuele, et al.
Pubblicazione: (2024)
Next-token pretraining implies in-context learning
di: Riechers, Paul M., et al.
Pubblicazione: (2025)
di: Riechers, Paul M., et al.
Pubblicazione: (2025)
Adaptive Token-Weighted Differential Privacy for LLMs: Not All Tokens Require Equal Protection
di: Yu, Manjiang, et al.
Pubblicazione: (2025)
di: Yu, Manjiang, et al.
Pubblicazione: (2025)
Scaling FP8 training to trillion-token LLMs
di: Fishman, Maxim, et al.
Pubblicazione: (2024)
di: Fishman, Maxim, et al.
Pubblicazione: (2024)
The pitfalls of next-token prediction
di: Bachmann, Gregor, et al.
Pubblicazione: (2024)
di: Bachmann, Gregor, et al.
Pubblicazione: (2024)
Looking beyond the next token
di: Thankaraj, Abitha, et al.
Pubblicazione: (2025)
di: Thankaraj, Abitha, et al.
Pubblicazione: (2025)
Probing the Embedding Space of Transformers via Minimal Token Perturbations
di: Conti, Eddie, et al.
Pubblicazione: (2025)
di: Conti, Eddie, et al.
Pubblicazione: (2025)
Mamba or Transformer for Time Series Forecasting? Mixture of Universals (MoU) Is All You Need
di: Peng, Sijia, et al.
Pubblicazione: (2024)
di: Peng, Sijia, et al.
Pubblicazione: (2024)
All-atom Diffusion Transformers: Unified generative modelling of molecules and materials
di: Joshi, Chaitanya K., et al.
Pubblicazione: (2025)
di: Joshi, Chaitanya K., et al.
Pubblicazione: (2025)
Mamba Modulation: On the Length Generalization of Mamba
di: Lu, Peng, et al.
Pubblicazione: (2025)
di: Lu, Peng, et al.
Pubblicazione: (2025)
A Survey of Mamba
di: Qu, Haohao, et al.
Pubblicazione: (2024)
di: Qu, Haohao, et al.
Pubblicazione: (2024)
Shaping capabilities with token-level data filtering
di: Rathi, Neil, et al.
Pubblicazione: (2026)
di: Rathi, Neil, et al.
Pubblicazione: (2026)
DeciMamba: Exploring the Length Extrapolation Potential of Mamba
di: Ben-Kish, Assaf, et al.
Pubblicazione: (2024)
di: Ben-Kish, Assaf, et al.
Pubblicazione: (2024)
Don't flatten, tokenize! Unlocking the key to SoftMoE's efficacy in deep RL
di: Sokar, Ghada, et al.
Pubblicazione: (2024)
di: Sokar, Ghada, et al.
Pubblicazione: (2024)
SUN: Shared Use of Next-token Prediction for Efficient Multi-LLM Disaggregated Serving
di: Woo, Sunghyeon, et al.
Pubblicazione: (2026)
di: Woo, Sunghyeon, et al.
Pubblicazione: (2026)
Genomic Next-Token Predictors are In-Context Learners
di: Breslow, Nathan, et al.
Pubblicazione: (2025)
di: Breslow, Nathan, et al.
Pubblicazione: (2025)
DiffuMamba: High-Throughput Diffusion LMs with Mamba Backbone
di: Singh, Vaibhav, et al.
Pubblicazione: (2025)
di: Singh, Vaibhav, et al.
Pubblicazione: (2025)
ms-Mamba: Multi-scale Mamba for Time-Series Forecasting
di: Karadag, Yusuf Meric, et al.
Pubblicazione: (2025)
di: Karadag, Yusuf Meric, et al.
Pubblicazione: (2025)
InfoMamba: An Attention-Free Hybrid Mamba-Transformer Model
di: Wang, Youjin, et al.
Pubblicazione: (2026)
di: Wang, Youjin, et al.
Pubblicazione: (2026)
Gym-Anything: Turn any Software into an Agent Environment
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2026)
di: Aggarwal, Pranjal, et al.
Pubblicazione: (2026)
Nd-BiMamba2: A Unified Bidirectional Architecture for Multi-Dimensional Data Processing
di: Liu, Hao
Pubblicazione: (2024)
di: Liu, Hao
Pubblicazione: (2024)
Scaling Transformer to 1M tokens and beyond with RMT
di: Bulatov, Aydar, et al.
Pubblicazione: (2023)
di: Bulatov, Aydar, et al.
Pubblicazione: (2023)
MambaSL: Exploring Single-Layer Mamba for Time Series Classification
di: Jung, Yoo-Min, et al.
Pubblicazione: (2026)
di: Jung, Yoo-Min, et al.
Pubblicazione: (2026)
eMamba: Efficient Acceleration Framework for Mamba Models in Edge Computing
di: Kim, Jiyong, et al.
Pubblicazione: (2025)
di: Kim, Jiyong, et al.
Pubblicazione: (2025)
any4: Learned 4-bit Numeric Representation for LLMs
di: Elhoushi, Mostafa, et al.
Pubblicazione: (2025)
di: Elhoushi, Mostafa, et al.
Pubblicazione: (2025)
Forward Only Learning for Orthogonal Neural Networks of any Depth
di: Caillon, Paul, et al.
Pubblicazione: (2025)
di: Caillon, Paul, et al.
Pubblicazione: (2025)
Decision Mamba Architectures
di: Correia, André, et al.
Pubblicazione: (2024)
di: Correia, André, et al.
Pubblicazione: (2024)
FoldToken2: Learning compact, invariant and generative protein structure language
di: Gao, Zhangyang, et al.
Pubblicazione: (2024)
di: Gao, Zhangyang, et al.
Pubblicazione: (2024)
Essential-Web v1.0: 24T tokens of organized web data
di: AI, Essential, et al.
Pubblicazione: (2025)
di: AI, Essential, et al.
Pubblicazione: (2025)
Interpretable Next-token Prediction via the Generalized Induction Head
di: Kim, Eunji, et al.
Pubblicazione: (2024)
di: Kim, Eunji, et al.
Pubblicazione: (2024)
Language models are better than humans at next-token prediction
di: Shlegeris, Buck, et al.
Pubblicazione: (2022)
di: Shlegeris, Buck, et al.
Pubblicazione: (2022)
A Mamba Foundation Model for Time Series Forecasting
di: Ma, Haoyu, et al.
Pubblicazione: (2024)
di: Ma, Haoyu, et al.
Pubblicazione: (2024)
You only need 4 extra tokens: Synergistic Test-time Adaptation for LLMs
di: Xu, Yijie, et al.
Pubblicazione: (2025)
di: Xu, Yijie, et al.
Pubblicazione: (2025)
TransferLight: Zero-Shot Traffic Signal Control on any Road-Network
di: Schmidt, Johann, et al.
Pubblicazione: (2024)
di: Schmidt, Johann, et al.
Pubblicazione: (2024)
Documenti analoghi
-
Flash Invariant Point Attention
di: Liu, Andrew, et al.
Pubblicazione: (2025) -
Not all tokens are needed(NAT): token efficient reinforcement learning
di: Sang, Hejian, et al.
Pubblicazione: (2026) -
BioMamba: Leveraging Spectro-Temporal Embedding in Bidirectional Mamba for Enhanced Biosignal Classification
di: Qian, Jian, et al.
Pubblicazione: (2025) -
Improving Inverse Folding for Peptide Design with Diversity-regularized Direct Preference Optimization
di: Park, Ryan, et al.
Pubblicazione: (2024) -
COMPACT: Common-token Optimized Model Pruning Across Channels and Tokens
di: Kwek, Eugene, et al.
Pubblicazione: (2025)