LaTIM: Measuring Latent Token-to-Token Interactions in Mamba Models
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Pitorro, Hugo, Treviso, Marcos |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens
von: Mamidanna, Siddarth, et al.
Veröffentlicht: (2025)
von: Mamidanna, Siddarth, et al.
Veröffentlicht: (2025)
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction
von: Walker, Nicholas
Veröffentlicht: (2024)
von: Walker, Nicholas
Veröffentlicht: (2024)
Steering Language Models in Multi-Token Generation: A Case Study on Tense and Aspect
von: Klerings, Alina, et al.
Veröffentlicht: (2025)
von: Klerings, Alina, et al.
Veröffentlicht: (2025)
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024)
Vocabulary Transfer for Biomedical Texts: Add Tokens if You Can Not Add Data
von: Singh, Priyanka, et al.
Veröffentlicht: (2022)
von: Singh, Priyanka, et al.
Veröffentlicht: (2022)
Just Pass Twice: Efficient Token Classification with LLMs for Zero-Shot NER
von: Ewais, Ahmed, et al.
Veröffentlicht: (2026)
von: Ewais, Ahmed, et al.
Veröffentlicht: (2026)
AraToken: Optimizing Arabic Tokenization with Normalization Pipeline and Language Extension for Qwen3
von: Kashirskiy, Mark, et al.
Veröffentlicht: (2025)
von: Kashirskiy, Mark, et al.
Veröffentlicht: (2025)
RomanLens: The Role Of Latent Romanization In Multilinguality In LLMs
von: Saji, Alan, et al.
Veröffentlicht: (2025)
von: Saji, Alan, et al.
Veröffentlicht: (2025)
Egalitarian Language Representation in Language Models: It All Begins with Tokenizers
von: Velayuthan, Menan, et al.
Veröffentlicht: (2024)
von: Velayuthan, Menan, et al.
Veröffentlicht: (2024)
UIPress: Bringing Optical Token Compression to UI-to-Code Generation
von: Dai, Dasen, et al.
Veröffentlicht: (2026)
von: Dai, Dasen, et al.
Veröffentlicht: (2026)
QuickSilver -- Speeding up LLM Inference through Dynamic Token Halting, KV Skipping, Contextual Token Fusion, and Adaptive Matryoshka Quantization
von: Khanna, Danush, et al.
Veröffentlicht: (2025)
von: Khanna, Danush, et al.
Veröffentlicht: (2025)
Beyond Token Length: Step Pruner for Efficient and Accurate Reasoning in Large Language Models
von: Wu, Canhui, et al.
Veröffentlicht: (2025)
von: Wu, Canhui, et al.
Veröffentlicht: (2025)
Pre-trained Models Perform the Best When Token Distributions Follow Zipf's Law
von: He, Yanjin, et al.
Veröffentlicht: (2025)
von: He, Yanjin, et al.
Veröffentlicht: (2025)
Kronecker Embeddings: Byte-Level Structured Token Representations for Parameter-Efficient Language Models
von: Shravan, Rohan
Veröffentlicht: (2026)
von: Shravan, Rohan
Veröffentlicht: (2026)
Fast Quiet-STaR: Thinking Without Thought Tokens
von: Huang, Wei, et al.
Veröffentlicht: (2025)
von: Huang, Wei, et al.
Veröffentlicht: (2025)
Token Erasure as a Footprint of Implicit Vocabulary Items in LLMs
von: Feucht, Sheridan, et al.
Veröffentlicht: (2024)
von: Feucht, Sheridan, et al.
Veröffentlicht: (2024)
CRISP: Persistent Concept Unlearning via Sparse Autoencoders
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
von: Ashuach, Tomer, et al.
Veröffentlicht: (2025)
Trading Complexity for Expressivity Through Structured Generalized Linear Token Mixing
von: Fagnou, Erwan, et al.
Veröffentlicht: (2026)
von: Fagnou, Erwan, et al.
Veröffentlicht: (2026)
SUBLLM: A Novel Efficient Architecture with Token Sequence Subsampling for LLM
von: Wang, Quandong, et al.
Veröffentlicht: (2024)
von: Wang, Quandong, et al.
Veröffentlicht: (2024)
Tokens with Meaning: A Hybrid Tokenization Approach for Turkish
von: Bayram, M. Ali, et al.
Veröffentlicht: (2025)
von: Bayram, M. Ali, et al.
Veröffentlicht: (2025)
Advancing Polish Language Modeling through Tokenizer Optimization in the Bielik v3 7B and 11B Series
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2026)
von: Ociepa, Krzysztof, et al.
Veröffentlicht: (2026)
TIS-DPO: Token-level Importance Sampling for Direct Preference Optimization With Estimated Weights
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
von: Liu, Aiwei, et al.
Veröffentlicht: (2024)
Why Models Know But Don't Say: Chain-of-Thought Faithfulness Divergence Between Thinking Tokens and Answers in Open-Weight Reasoning Models
von: Young, Richard J.
Veröffentlicht: (2026)
von: Young, Richard J.
Veröffentlicht: (2026)
BrahmicTokenizer-131K: An Indic-Capable Drop-In Replacement for o200k_base
von: Shravan, Rohan
Veröffentlicht: (2026)
von: Shravan, Rohan
Veröffentlicht: (2026)
Tokenization Is More Than Compression
von: Schmidt, Craig W., et al.
Veröffentlicht: (2024)
von: Schmidt, Craig W., et al.
Veröffentlicht: (2024)
Text-Based Approaches to Item Difficulty Modeling in Large-Scale Assessments: A Systematic Review
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
von: Peters, Sydney, et al.
Veröffentlicht: (2025)
Towards Fundamental Language Models: Does Linguistic Competence Scale with Model Size?
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
von: Collado-Montañez, Jaime, et al.
Veröffentlicht: (2025)
Do Latent Tokens Think? A Causal and Adversarial Analysis of Chain-of-Continuous-Thought
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
von: Zhang, Yuyi, et al.
Veröffentlicht: (2025)
Next Token Prediction Is a Dead End for Creativity
von: Olatunji, Ibukun, et al.
Veröffentlicht: (2025)
von: Olatunji, Ibukun, et al.
Veröffentlicht: (2025)
TREX: Tokenizer Regression for Optimal Data Mixture
von: Won, Inho, et al.
Veröffentlicht: (2026)
von: Won, Inho, et al.
Veröffentlicht: (2026)
Textual Entailment is not a Better Bias Metric than Token Probability
von: Felkner, Virginia K., et al.
Veröffentlicht: (2025)
von: Felkner, Virginia K., et al.
Veröffentlicht: (2025)
Detecting Hallucinations in Large Language Model Generation: A Token Probability Approach
von: Quevedo, Ernesto, et al.
Veröffentlicht: (2024)
von: Quevedo, Ernesto, et al.
Veröffentlicht: (2024)
ECG-Byte: A Tokenizer for End-to-End Generative Electrocardiogram Language Modeling
von: Han, William, et al.
Veröffentlicht: (2024)
von: Han, William, et al.
Veröffentlicht: (2024)
Entropy-Based Measurement of Value Drift and Alignment Work in Large Language Models
von: Fadli, Samih
Veröffentlicht: (2025)
von: Fadli, Samih
Veröffentlicht: (2025)
Encoder-Decoder Framework for Interactive Free Verses with Generation with Controllable High-Quality Rhyming
von: Pasini, Tommaso, et al.
Veröffentlicht: (2024)
von: Pasini, Tommaso, et al.
Veröffentlicht: (2024)
SeLeRoSa: Sentence-Level Romanian Satire Detection Dataset
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
von: Smădu, Răzvan-Alexandru, et al.
Veröffentlicht: (2025)
Qtok: A Comprehensive Framework for Evaluating Multilingual Tokenizer Quality in Large Language Models
von: Chelombitko, Iaroslav, et al.
Veröffentlicht: (2024)
von: Chelombitko, Iaroslav, et al.
Veröffentlicht: (2024)
PeLLE: Encoder-based language models for Brazilian Portuguese based on open data
von: de Mello, Guilherme Lamartine, et al.
Veröffentlicht: (2024)
von: de Mello, Guilherme Lamartine, et al.
Veröffentlicht: (2024)
Evaluating the Efficacy of Hybrid Deep Learning Models in Distinguishing AI-Generated Text
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2023)
KinyaColBERT: A Lexically Grounded Retrieval Model for Low-Resource Retrieval-Augmented Generation
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025)
von: Nzeyimana, Antoine, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens
von: Mamidanna, Siddarth, et al.
Veröffentlicht: (2025) -
Future Token Prediction -- Causal Language Modelling with Per-Token Semantic State Vector for Multi-Token Prediction
von: Walker, Nicholas
Veröffentlicht: (2024) -
Steering Language Models in Multi-Token Generation: A Case Study on Tense and Aspect
von: Klerings, Alina, et al.
Veröffentlicht: (2025) -
Tokenization and Morphology in Multilingual Language Models: A Comparative Analysis of mT5 and ByT5
von: Dang, Thao Anh, et al.
Veröffentlicht: (2024) -
Vocabulary Transfer for Biomedical Texts: Add Tokens if You Can Not Add Data
von: Singh, Priyanka, et al.
Veröffentlicht: (2022)