Guardado en:
| Autores principales: | Pham, Quoc Tuan, Jafari, Mehdi, Salim, Flora |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.06266 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Mechanistic Indicators of Steering Effectiveness in Large Language Models
por: Jafari, Mehdi, et al.
Publicado: (2026)
por: Jafari, Mehdi, et al.
Publicado: (2026)
Enhancing Conversational Agents with Theory of Mind: Aligning Beliefs, Desires, and Intentions for Human-Like Interaction
por: Jafari, Mehdi, et al.
Publicado: (2025)
por: Jafari, Mehdi, et al.
Publicado: (2025)
PyRQA -- Conducting Recurrence Quantification Analysis on Very Long Time Series Efficiently
por: Rawald, Tobias, et al.
Publicado: (2024)
por: Rawald, Tobias, et al.
Publicado: (2024)
Is my model perplexed for the right reason? Contrasting LLMs' Benchmark Behavior with Token-Level Perplexity
por: Prins, Zoë, et al.
Publicado: (2026)
por: Prins, Zoë, et al.
Publicado: (2026)
Interpreting token compositionality in LLMs: A robustness analysis
por: Aljaafari, Nura, et al.
Publicado: (2024)
por: Aljaafari, Nura, et al.
Publicado: (2024)
ClozeMath: Improving Mathematical Reasoning in Language Models by Learning to Fill Equations
por: Pham, Quang Hieu, et al.
Publicado: (2025)
por: Pham, Quang Hieu, et al.
Publicado: (2025)
What am I missing here?: Evaluating Large Language Models for Masked Sentence Prediction
por: Wyatt, Charlie, et al.
Publicado: (2025)
por: Wyatt, Charlie, et al.
Publicado: (2025)
LocalRQA: From Generating Data to Locally Training, Testing, and Deploying Retrieval-Augmented QA Systems
por: Yu, Xiao, et al.
Publicado: (2024)
por: Yu, Xiao, et al.
Publicado: (2024)
Harnessing Test-time Adaptation for NLU tasks Involving Dialects of English
por: Nguyen, Duke, et al.
Publicado: (2025)
por: Nguyen, Duke, et al.
Publicado: (2025)
Alternatives To Next Token Prediction In Text Generation -- A Survey
por: Wyatt, Charlie, et al.
Publicado: (2025)
por: Wyatt, Charlie, et al.
Publicado: (2025)
Who's Who: Large Language Models Meet Knowledge Conflicts in Practice
por: Pham, Quang Hieu, et al.
Publicado: (2024)
por: Pham, Quang Hieu, et al.
Publicado: (2024)
Explore Theory of Mind: Program-guided adversarial data generation for theory of mind reasoning
por: Sclar, Melanie, et al.
Publicado: (2024)
por: Sclar, Melanie, et al.
Publicado: (2024)
MAPLE: Mobile App Prediction Leveraging Large Language Model Embeddings
por: Khaokaew, Yonchanok, et al.
Publicado: (2023)
por: Khaokaew, Yonchanok, et al.
Publicado: (2023)
Interpretable Next-token Prediction via the Generalized Induction Head
por: Kim, Eunji, et al.
Publicado: (2024)
por: Kim, Eunji, et al.
Publicado: (2024)
Contextual morphologically-guided tokenization for Latin encoder models
por: Hudspeth, Marisa, et al.
Publicado: (2025)
por: Hudspeth, Marisa, et al.
Publicado: (2025)
Towards Reliable Medical Question Answering: Techniques and Challenges in Mitigating Hallucinations in Language Models
por: Pham, Duy Khoa, et al.
Publicado: (2024)
por: Pham, Duy Khoa, et al.
Publicado: (2024)
PERCORE: A Deep Learning-Based Framework for Persian Spelling Correction with Phonetic Analysis
por: Dashti, Seyed Mohammad Sadegh, et al.
Publicado: (2024)
por: Dashti, Seyed Mohammad Sadegh, et al.
Publicado: (2024)
CP-MoE: Consistency-Preserving Mixture-of-Experts for Continual Learning
por: Liu, Yang, et al.
Publicado: (2026)
por: Liu, Yang, et al.
Publicado: (2026)
Mitigating Data Scarcity in Psychological Defense Classification with Context-Aware Synthetic Augmentation
por: Vu, Hoang-Thuy-Duong, et al.
Publicado: (2026)
por: Vu, Hoang-Thuy-Duong, et al.
Publicado: (2026)
ZARA: Training-Free Motion Time-Series Reasoning via Evidence-Grounded LLM Agents
por: Li, Zechen, et al.
Publicado: (2025)
por: Li, Zechen, et al.
Publicado: (2025)
Do language models plan ahead for future tokens?
por: Wu, Wilson, et al.
Publicado: (2024)
por: Wu, Wilson, et al.
Publicado: (2024)
Visualizing token importance for black-box language models
por: Rauba, Paulius, et al.
Publicado: (2025)
por: Rauba, Paulius, et al.
Publicado: (2025)
Prompt Mining for Language-based Human Mobility Forecasting
por: Xue, Hao, et al.
Publicado: (2024)
por: Xue, Hao, et al.
Publicado: (2024)
Collaborative decoding of critical tokens for boosting factuality of large language models
por: Jin, Lifeng, et al.
Publicado: (2024)
por: Jin, Lifeng, et al.
Publicado: (2024)
SensorLLM: Aligning Large Language Models with Motion Sensors for Human Activity Recognition
por: Li, Zechen, et al.
Publicado: (2024)
por: Li, Zechen, et al.
Publicado: (2024)
AdaCS: Adaptive Normalization for Enhanced Code-Switching ASR
por: Chu, The Chuong, et al.
Publicado: (2025)
por: Chu, The Chuong, et al.
Publicado: (2025)
Distributional reasoning in LLMs: Parallel reasoning processes in multi-hop reasoning
por: Shalev, Yuval, et al.
Publicado: (2024)
por: Shalev, Yuval, et al.
Publicado: (2024)
Leveraging Sentence-oriented Augmentation and Transformer-Based Architecture for Vietnamese-Bahnaric Translation
por: Nguyen, Tan Sang, et al.
Publicado: (2026)
por: Nguyen, Tan Sang, et al.
Publicado: (2026)
On the scaling relationship between cloze probabilities and language model next-token prediction
por: Jacobs, Cassandra L., et al.
Publicado: (2026)
por: Jacobs, Cassandra L., et al.
Publicado: (2026)
Where is the signal in tokenization space?
por: Geh, Renato Lui, et al.
Publicado: (2024)
por: Geh, Renato Lui, et al.
Publicado: (2024)
Automatic Real-word Error Correction in Persian Text
por: Dashti, Seyed Mohammad Sadegh, et al.
Publicado: (2024)
por: Dashti, Seyed Mohammad Sadegh, et al.
Publicado: (2024)
Byte-token Enhanced Language Models for Temporal Point Processes Analysis
por: Kong, Quyu, et al.
Publicado: (2025)
por: Kong, Quyu, et al.
Publicado: (2025)
On the token distance modeling ability of higher RoPE attention dimension
por: Hong, Xiangyu, et al.
Publicado: (2024)
por: Hong, Xiangyu, et al.
Publicado: (2024)
Why do LLMs attend to the first token?
por: Barbero, Federico, et al.
Publicado: (2025)
por: Barbero, Federico, et al.
Publicado: (2025)
Revisiting subword tokenization: A case study on affixal negation in large language models
por: Truong, Thinh Hung, et al.
Publicado: (2024)
por: Truong, Thinh Hung, et al.
Publicado: (2024)
Practical token pruning for foundation models in few-shot conversational virtual assistant systems
por: Qi, Haode, et al.
Publicado: (2024)
por: Qi, Haode, et al.
Publicado: (2024)
Is continuous CoT better suited for multi-lingual reasoning?
por: Bashir, Ali Hamza, et al.
Publicado: (2026)
por: Bashir, Ali Hamza, et al.
Publicado: (2026)
Language models are better than humans at next-token prediction
por: Shlegeris, Buck, et al.
Publicado: (2022)
por: Shlegeris, Buck, et al.
Publicado: (2022)
RELOOP: Recursive Retrieval with Multi-Hop Reasoner and Planners for Heterogeneous QA
por: Yang, Ruiyi, et al.
Publicado: (2025)
por: Yang, Ruiyi, et al.
Publicado: (2025)
Automated stereotactic radiosurgery planning using a human-in-the-loop reasoning large language model agent
por: Nusrat, Humza, et al.
Publicado: (2025)
por: Nusrat, Humza, et al.
Publicado: (2025)
Ejemplares similares
-
Mechanistic Indicators of Steering Effectiveness in Large Language Models
por: Jafari, Mehdi, et al.
Publicado: (2026) -
Enhancing Conversational Agents with Theory of Mind: Aligning Beliefs, Desires, and Intentions for Human-Like Interaction
por: Jafari, Mehdi, et al.
Publicado: (2025) -
PyRQA -- Conducting Recurrence Quantification Analysis on Very Long Time Series Efficiently
por: Rawald, Tobias, et al.
Publicado: (2024) -
Is my model perplexed for the right reason? Contrasting LLMs' Benchmark Behavior with Token-Level Perplexity
por: Prins, Zoë, et al.
Publicado: (2026) -
Interpreting token compositionality in LLMs: A robustness analysis
por: Aljaafari, Nura, et al.
Publicado: (2024)