Beyond Language Models: Byte Models are Digital World Simulators
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Wu, Shangda, Tan, Xu, Wang, Zili, Wang, Rui, Li, Xiaobing, Sun, Maosong |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing
von: Wu, Shangda, et al.
Veröffentlicht: (2024)
von: Wu, Shangda, et al.
Veröffentlicht: (2024)
ByteFlow: Language Modeling through Adaptive Byte Compression without a Tokenizer
von: Deng, Chunyuan, et al.
Veröffentlicht: (2026)
von: Deng, Chunyuan, et al.
Veröffentlicht: (2026)
The Efficiency Gap in Byte Modeling
von: Lee, Celine, et al.
Veröffentlicht: (2026)
von: Lee, Celine, et al.
Veröffentlicht: (2026)
PANTHER: Generative Pretraining Beyond Language for Sequential User Behavior Modeling
von: Li, Guilin, et al.
Veröffentlicht: (2025)
von: Li, Guilin, et al.
Veröffentlicht: (2025)
Robust and Scalable Model Editing for Large Language Models
von: Chen, Yingfa, et al.
Veröffentlicht: (2024)
von: Chen, Yingfa, et al.
Veröffentlicht: (2024)
TritonBench: Benchmarking Large Language Model Capabilities for Generating Triton Operators
von: Li, Jianling, et al.
Veröffentlicht: (2025)
von: Li, Jianling, et al.
Veröffentlicht: (2025)
Pruning and Malicious Injection: A Retraining-Free Backdoor Attack on Transformer Models
von: Zhao, Taibiao, et al.
Veröffentlicht: (2025)
von: Zhao, Taibiao, et al.
Veröffentlicht: (2025)
A Closer Look into Mixture-of-Experts in Large Language Models
von: Lo, Ka Man, et al.
Veröffentlicht: (2024)
von: Lo, Ka Man, et al.
Veröffentlicht: (2024)
Exponential Concentration in Stochastic Approximation
von: Law, Kody, et al.
Veröffentlicht: (2022)
von: Law, Kody, et al.
Veröffentlicht: (2022)
Diffusion Language Models are Super Data Learners
von: Ni, Jinjie, et al.
Veröffentlicht: (2025)
von: Ni, Jinjie, et al.
Veröffentlicht: (2025)
World Models with Hints of Large Language Models for Goal Achieving
von: Liu, Zeyuan, et al.
Veröffentlicht: (2024)
von: Liu, Zeyuan, et al.
Veröffentlicht: (2024)
Do Large Language Models Truly Grasp Mathematics? An Empirical Exploration From Cognitive Psychology
von: Xie, Wei, et al.
Veröffentlicht: (2024)
von: Xie, Wei, et al.
Veröffentlicht: (2024)
Beyond Euclidean Proximity: Repairing Latent World Models with Horizon-Matched Trajectory Reachability Metrics
von: Li, Liangyu, et al.
Veröffentlicht: (2026)
von: Li, Liangyu, et al.
Veröffentlicht: (2026)
MambaByte: Token-free Selective State Space Model
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
Towards Practical World Model-based Reinforcement Learning for Vision-Language-Action Models
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
von: Zhang, Zhilong, et al.
Veröffentlicht: (2026)
Byte-token Enhanced Language Models for Temporal Point Processes Analysis
von: Kong, Quyu, et al.
Veröffentlicht: (2025)
von: Kong, Quyu, et al.
Veröffentlicht: (2025)
Latent Geometry Beyond Search: Amortizing Planning in World Models
von: Nguyen, Hoang, et al.
Veröffentlicht: (2026)
von: Nguyen, Hoang, et al.
Veröffentlicht: (2026)
Closing the Loop: Learning to Generate Writing Feedback via Language Model Simulated Student Revisions
von: Nair, Inderjeet, et al.
Veröffentlicht: (2024)
von: Nair, Inderjeet, et al.
Veröffentlicht: (2024)
AcceRL: A Distributed Asynchronous Reinforcement Learning and World Model Framework for Vision-Language-Action Models
von: Lu, Chengxuan, et al.
Veröffentlicht: (2026)
von: Lu, Chengxuan, et al.
Veröffentlicht: (2026)
Multi-view Fake News Detection Model Based on Dynamic Hypergraph
von: Ye, Rongping, et al.
Veröffentlicht: (2024)
von: Ye, Rongping, et al.
Veröffentlicht: (2024)
Digital Player: Evaluating Large Language Models based Human-like Agent in Games
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
von: Wang, Jiawei, et al.
Veröffentlicht: (2025)
CEQuest: Benchmarking Large Language Models for Construction Estimation
von: Wu, Yanzhao, et al.
Veröffentlicht: (2025)
von: Wu, Yanzhao, et al.
Veröffentlicht: (2025)
Exploring Perceptual Limitation of Multimodal Large Language Models
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
von: Zhang, Jiarui, et al.
Veröffentlicht: (2024)
GraphSB: Boosting Imbalanced Node Classification on Graphs through Structural Balance
von: Zhu, Chaofan, et al.
Veröffentlicht: (2025)
von: Zhu, Chaofan, et al.
Veröffentlicht: (2025)
Exact Byte-Level Probabilities from Tokenized Language Models for FIM-Tasks and Model Ensembles
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
Exploring Tokenization Methods for Multitrack Sheet Music Generation
von: Wang, Yashan, et al.
Veröffentlicht: (2024)
von: Wang, Yashan, et al.
Veröffentlicht: (2024)
Language Models over Canonical Byte-Pair Encodings
von: Vieira, Tim, et al.
Veröffentlicht: (2025)
von: Vieira, Tim, et al.
Veröffentlicht: (2025)
Beyond State Consistency: Behavior Consistency in Text-Based World Models
von: Huang, Youling, et al.
Veröffentlicht: (2026)
von: Huang, Youling, et al.
Veröffentlicht: (2026)
World Model-Enabled Causal Digital Twins for Semantic Communications in Physical AI Systems
von: Wang, Lingyi, et al.
Veröffentlicht: (2026)
von: Wang, Lingyi, et al.
Veröffentlicht: (2026)
HoloByte: Continuous Hyperspherical Distillation for Tokenizer-Free Modeling
von: Khasia, Vladimer
Veröffentlicht: (2026)
von: Khasia, Vladimer
Veröffentlicht: (2026)
Beyond Accuracy: On the Effects of Fine-tuning Towards Vision-Language Model's Prediction Rationality
von: Wang, Qitong, et al.
Veröffentlicht: (2024)
von: Wang, Qitong, et al.
Veröffentlicht: (2024)
SpaceByte: Towards Deleting Tokenization from Large Language Modeling
von: Slagle, Kevin
Veröffentlicht: (2024)
von: Slagle, Kevin
Veröffentlicht: (2024)
Understanding the Natural Language of DNA using Encoder-Decoder Foundation Models with Byte-level Precision
von: Malusare, Aditya, et al.
Veröffentlicht: (2023)
von: Malusare, Aditya, et al.
Veröffentlicht: (2023)
FR-Spec: Accelerating Large-Vocabulary Language Models via Frequency-Ranked Speculative Sampling
von: Zhao, Weilin, et al.
Veröffentlicht: (2025)
von: Zhao, Weilin, et al.
Veröffentlicht: (2025)
Scratchpad Patching: Decoupling Compute from Patch Size in Byte-Level Language Models
von: Zheng, Lin, et al.
Veröffentlicht: (2026)
von: Zheng, Lin, et al.
Veröffentlicht: (2026)
Out-of-Distribution Graph Models Merging
von: Wang, Yidi, et al.
Veröffentlicht: (2025)
von: Wang, Yidi, et al.
Veröffentlicht: (2025)
NotaGen: Advancing Musicality in Symbolic Music Generation with Large Language Model Training Paradigms
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
von: Wang, Yashan, et al.
Veröffentlicht: (2025)
Fast Distributed Inference Serving for Large Language Models
von: Wu, Bingyang, et al.
Veröffentlicht: (2023)
von: Wu, Bingyang, et al.
Veröffentlicht: (2023)
TWIN-GPT: Digital Twins for Clinical Trials via Large Language Model
von: Wang, Yue, et al.
Veröffentlicht: (2024)
von: Wang, Yue, et al.
Veröffentlicht: (2024)
Policy and World Modeling Co-Training for Language Agents
von: Lu, Ning, et al.
Veröffentlicht: (2026)
von: Lu, Ning, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
MelodyT5: A Unified Score-to-Score Transformer for Symbolic Music Processing
von: Wu, Shangda, et al.
Veröffentlicht: (2024) -
ByteFlow: Language Modeling through Adaptive Byte Compression without a Tokenizer
von: Deng, Chunyuan, et al.
Veröffentlicht: (2026) -
The Efficiency Gap in Byte Modeling
von: Lee, Celine, et al.
Veröffentlicht: (2026) -
PANTHER: Generative Pretraining Beyond Language for Sequential User Behavior Modeling
von: Li, Guilin, et al.
Veröffentlicht: (2025) -
Robust and Scalable Model Editing for Large Language Models
von: Chen, Yingfa, et al.
Veröffentlicht: (2024)