Byte-token Enhanced Language Models for Temporal Point Processes Analysis
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Kong, Quyu, Zhang, Yixuan, Liu, Yang, Tong, Panrong, Liu, Enqi, Zhou, Feng |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
DanmakuTPPBench: A Multi-modal Benchmark for Temporal Point Process Modeling and Understanding
von: Jiang, Yue, et al.
Veröffentlicht: (2025)
von: Jiang, Yue, et al.
Veröffentlicht: (2025)
Long-range Modeling and Processing of Multimodal Event Sequences
von: Li, Jichu, et al.
Veröffentlicht: (2026)
von: Li, Jichu, et al.
Veröffentlicht: (2026)
Advances in Temporal Point Processes: Bayesian, Neural, and LLM Approaches
von: Zhou, Feng, et al.
Veröffentlicht: (2025)
von: Zhou, Feng, et al.
Veröffentlicht: (2025)
TPP-LLM: Modeling Temporal Point Processes by Efficiently Fine-Tuning Large Language Models
von: Liu, Zefang, et al.
Veröffentlicht: (2024)
von: Liu, Zefang, et al.
Veröffentlicht: (2024)
Negative Binomial Variational Autoencoders for Overdispersed Latent Modeling
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
ByteFlow: Language Modeling through Adaptive Byte Compression without a Tokenizer
von: Deng, Chunyuan, et al.
Veröffentlicht: (2026)
von: Deng, Chunyuan, et al.
Veröffentlicht: (2026)
TPP-SD: Accelerating Transformer Point Process Sampling with Speculative Decoding
von: Gong, Shukai, et al.
Veröffentlicht: (2025)
von: Gong, Shukai, et al.
Veröffentlicht: (2025)
Echo: A Large Language Model with Temporal Episodic Memory
von: Liu, WenTao, et al.
Veröffentlicht: (2025)
von: Liu, WenTao, et al.
Veröffentlicht: (2025)
Sampling from Your Language Model One Byte at a Time
von: Hayase, Jonathan, et al.
Veröffentlicht: (2025)
von: Hayase, Jonathan, et al.
Veröffentlicht: (2025)
Language Models over Canonical Byte-Pair Encodings
von: Vieira, Tim, et al.
Veröffentlicht: (2025)
von: Vieira, Tim, et al.
Veröffentlicht: (2025)
Where is the signal in tokenization space?
von: Geh, Renato Lui, et al.
Veröffentlicht: (2024)
von: Geh, Renato Lui, et al.
Veröffentlicht: (2024)
Proxy Compression for Language Modeling
von: Zheng, Lin, et al.
Veröffentlicht: (2026)
von: Zheng, Lin, et al.
Veröffentlicht: (2026)
Implicit Geometry of Next-token Prediction: From Language Sparsity Patterns to Model Representations
von: Zhao, Yize, et al.
Veröffentlicht: (2024)
von: Zhao, Yize, et al.
Veröffentlicht: (2024)
Fair Bayesian Data Selection via Generalized Discrepancy Measures
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)
PrAg-PO: Prompt Augmented Policy Optimization for Robust and Diverse Mathematical Reasoning
von: Lu, Wenquan, et al.
Veröffentlicht: (2026)
von: Lu, Wenquan, et al.
Veröffentlicht: (2026)
Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
Hierarchical Autoregressive Transformers: Combining Byte- and Word-Level Processing for Robust, Adaptable Language Models
von: Neitemeier, Pit, et al.
Veröffentlicht: (2025)
von: Neitemeier, Pit, et al.
Veröffentlicht: (2025)
Temporal Tokenization Strategies for Event Sequence Modeling with Large Language Models
von: Liu, Zefang, et al.
Veröffentlicht: (2025)
von: Liu, Zefang, et al.
Veröffentlicht: (2025)
GOFA: A Generative One-For-All Model for Joint Graph Language Modeling
von: Kong, Lecheng, et al.
Veröffentlicht: (2024)
von: Kong, Lecheng, et al.
Veröffentlicht: (2024)
Spatial-Temporal Large Language Model for Traffic Prediction
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
von: Liu, Chenxi, et al.
Veröffentlicht: (2024)
Exact Byte-Level Probabilities from Tokenized Language Models for FIM-Tasks and Model Ensembles
von: Phan, Buu, et al.
Veröffentlicht: (2024)
von: Phan, Buu, et al.
Veröffentlicht: (2024)
All or None: Identifiable Linear Properties of Next-token Predictors in Language Modeling
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
von: Marconato, Emanuele, et al.
Veröffentlicht: (2024)
Scratchpad Patching: Decoupling Compute from Patch Size in Byte-Level Language Models
von: Zheng, Lin, et al.
Veröffentlicht: (2026)
von: Zheng, Lin, et al.
Veröffentlicht: (2026)
On multi-token prediction for efficient LLM inference
von: Mehra, Somesh, et al.
Veröffentlicht: (2025)
von: Mehra, Somesh, et al.
Veröffentlicht: (2025)
DeepReview: Improving LLM-based Paper Review with Human-like Deep Thinking Process
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
BRiTE: Bootstrapping Reinforced Thinking Process to Enhance Language Model Reasoning
von: Zhong, Han, et al.
Veröffentlicht: (2025)
von: Zhong, Han, et al.
Veröffentlicht: (2025)
ChunkFT: Byte-Streamed Optimization for Memory-Efficient Full Fine-Tuning
von: Liu, Yongkang, et al.
Veröffentlicht: (2026)
von: Liu, Yongkang, et al.
Veröffentlicht: (2026)
Keyframe-oriented Vision Token Pruning: Enhancing Efficiency of Large Vision Language Models on Long-Form Video Processing
von: Liu, Yudong, et al.
Veröffentlicht: (2025)
von: Liu, Yudong, et al.
Veröffentlicht: (2025)
Language models are better than humans at next-token prediction
von: Shlegeris, Buck, et al.
Veröffentlicht: (2022)
von: Shlegeris, Buck, et al.
Veröffentlicht: (2022)
A Spatio-Temporal Point Process for Fine-Grained Modeling of Reading Behavior
von: Re, Francesco Ignazio, et al.
Veröffentlicht: (2025)
von: Re, Francesco Ignazio, et al.
Veröffentlicht: (2025)
Visualizing token importance for black-box language models
von: Rauba, Paulius, et al.
Veröffentlicht: (2025)
von: Rauba, Paulius, et al.
Veröffentlicht: (2025)
Tokenization counts: the impact of tokenization on arithmetic in frontier LLMs
von: Singh, Aaditya K., et al.
Veröffentlicht: (2024)
von: Singh, Aaditya K., et al.
Veröffentlicht: (2024)
Do language models plan ahead for future tokens?
von: Wu, Wilson, et al.
Veröffentlicht: (2024)
von: Wu, Wilson, et al.
Veröffentlicht: (2024)
Looking beyond the next token
von: Thankaraj, Abitha, et al.
Veröffentlicht: (2025)
von: Thankaraj, Abitha, et al.
Veröffentlicht: (2025)
The pitfalls of next-token prediction
von: Bachmann, Gregor, et al.
Veröffentlicht: (2024)
von: Bachmann, Gregor, et al.
Veröffentlicht: (2024)
Make Some Noise: Unlocking Language Model Parallel Inference Capability through Noisy Training
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
von: Wang, Yixuan, et al.
Veröffentlicht: (2024)
A Systematic Analysis on the Temporal Generalization of Language Models in Social Media
von: Ushio, Asahi, et al.
Veröffentlicht: (2024)
von: Ushio, Asahi, et al.
Veröffentlicht: (2024)
How Can Large Language Models Understand Spatial-Temporal Data?
von: Liu, Lei, et al.
Veröffentlicht: (2024)
von: Liu, Lei, et al.
Veröffentlicht: (2024)
SpaceByte: Towards Deleting Tokenization from Large Language Modeling
von: Slagle, Kevin
Veröffentlicht: (2024)
von: Slagle, Kevin
Veröffentlicht: (2024)
MambaByte: Token-free Selective State Space Model
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
von: Wang, Junxiong, et al.
Veröffentlicht: (2024)
Ähnliche Einträge
-
DanmakuTPPBench: A Multi-modal Benchmark for Temporal Point Process Modeling and Understanding
von: Jiang, Yue, et al.
Veröffentlicht: (2025) -
Long-range Modeling and Processing of Multimodal Event Sequences
von: Li, Jichu, et al.
Veröffentlicht: (2026) -
Advances in Temporal Point Processes: Bayesian, Neural, and LLM Approaches
von: Zhou, Feng, et al.
Veröffentlicht: (2025) -
TPP-LLM: Modeling Temporal Point Processes by Efficiently Fine-Tuning Large Language Models
von: Liu, Zefang, et al.
Veröffentlicht: (2024) -
Negative Binomial Variational Autoencoders for Overdispersed Latent Modeling
von: Zhang, Yixuan, et al.
Veröffentlicht: (2025)