Guardado en:
| Autores principales: | Saurez, Andres, Sengar, Neha, Har, Dongsoo |
|---|---|
| Formato: | Preprint |
| Publicado: |
2026
|
| Materias: | |
| Acceso en línea: | https://arxiv.org/abs/2602.09784 |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
Why Linear Interpretability Works: Invariant Subspaces as a Result of Architectural Constraints
por: Saurez, Andres, et al.
Publicado: (2026)
por: Saurez, Andres, et al.
Publicado: (2026)
Continuous Adversarial Text Representation Learning for Affective Recognition
por: Son, Seungah, et al.
Publicado: (2025)
por: Son, Seungah, et al.
Publicado: (2025)
Steering Sparse Autoencoder Latents to Control Dynamic Head Pruning in Vision Transformers (Student Abstract)
por: Lee, Yousung, et al.
Publicado: (2026)
por: Lee, Yousung, et al.
Publicado: (2026)
PAST: Phonetic-Acoustic Speech Tokenizer
por: Har-Tuv, Nadav, et al.
Publicado: (2025)
por: Har-Tuv, Nadav, et al.
Publicado: (2025)
Understanding Token Probability Encoding in Output Embeddings
por: Cho, Hakaze, et al.
Publicado: (2024)
por: Cho, Hakaze, et al.
Publicado: (2024)
Optimized Multi-Token Joint Decoding with Auxiliary Model for LLM Inference
por: Qin, Zongyue, et al.
Publicado: (2024)
por: Qin, Zongyue, et al.
Publicado: (2024)
On the Semantic and Syntactic Information Encoded in Proto-Tokens for One-Step Text Reconstruction
por: Bondarenko, Ivan, et al.
Publicado: (2026)
por: Bondarenko, Ivan, et al.
Publicado: (2026)
Do LLMs Encode Functional Importance of Reasoning Tokens?
por: Singh, Janvijay, et al.
Publicado: (2026)
por: Singh, Janvijay, et al.
Publicado: (2026)
Language Model Cascades: Token-level uncertainty and beyond
por: Gupta, Neha, et al.
Publicado: (2024)
por: Gupta, Neha, et al.
Publicado: (2024)
Parity-Aware Byte-Pair Encoding: Improving Cross-lingual Fairness in Tokenization
por: Foroutan, Negar, et al.
Publicado: (2025)
por: Foroutan, Negar, et al.
Publicado: (2025)
Student Answer Forecasting: Transformer-Driven Answer Choice Prediction for Language Learning
por: Gado, Elena Grazia, et al.
Publicado: (2024)
por: Gado, Elena Grazia, et al.
Publicado: (2024)
Discursive Circuits: How Do Language Models Understand Discourse Relations?
por: Miao, Yisong, et al.
Publicado: (2025)
por: Miao, Yisong, et al.
Publicado: (2025)
Antidistillation Fingerprinting
por: Xu, Yixuan Even, et al.
Publicado: (2026)
por: Xu, Yixuan Even, et al.
Publicado: (2026)
DropBP: Accelerating Fine-Tuning of Large Language Models by Dropping Backward Propagation
por: Woo, Sunghyeon, et al.
Publicado: (2024)
por: Woo, Sunghyeon, et al.
Publicado: (2024)
Promote, Suppress, Iterate: How Language Models Answer One-to-Many Factual Queries
por: Yan, Tianyi Lorena, et al.
Publicado: (2025)
por: Yan, Tianyi Lorena, et al.
Publicado: (2025)
Generative Artificial Intelligence: A Systematic Review and Applications
por: Sengar, Sandeep Singh, et al.
Publicado: (2024)
por: Sengar, Sandeep Singh, et al.
Publicado: (2024)
On the Robustness of Answer Formats in Medical Reasoning Models
por: Taveekitworachai, Pittawat, et al.
Publicado: (2025)
por: Taveekitworachai, Pittawat, et al.
Publicado: (2025)
Detecting and Suppressing Reward Hacking with Gradient Fingerprints
por: Wang, Songtao, et al.
Publicado: (2026)
por: Wang, Songtao, et al.
Publicado: (2026)
On Next-Token Prediction in LLMs: How End Goals Determine the Consistency of Decoding Algorithms
por: Trauger, Jacob, et al.
Publicado: (2025)
por: Trauger, Jacob, et al.
Publicado: (2025)
Representation Consistency for Accurate and Coherent LLM Answer Aggregation
por: Jiang, Junqi, et al.
Publicado: (2025)
por: Jiang, Junqi, et al.
Publicado: (2025)
Found in the Middle: How Language Models Use Long Contexts Better via Plug-and-Play Positional Encoding
por: Zhang, Zhenyu, et al.
Publicado: (2024)
por: Zhang, Zhenyu, et al.
Publicado: (2024)
How Do Transformers Learn to Associate Tokens: Gradient Leading Terms Bring Mechanistic Interpretability
por: Im, Shawn, et al.
Publicado: (2026)
por: Im, Shawn, et al.
Publicado: (2026)
The Truncation Blind Spot: How Decoding Strategies Systematically Exclude Human-Like Token Choices
por: Arias, Esteban Garces, et al.
Publicado: (2026)
por: Arias, Esteban Garces, et al.
Publicado: (2026)
Fingerprint Vector: Enabling Scalable and Efficient Model Fingerprint Transfer via Vector Addition
por: Xu, Zhenhua, et al.
Publicado: (2024)
por: Xu, Zhenhua, et al.
Publicado: (2024)
Graph-based Molecular In-context Learning Grounded on Morgan Fingerprints
por: Al-Lawati, Ali, et al.
Publicado: (2025)
por: Al-Lawati, Ali, et al.
Publicado: (2025)
How Important Is Tokenization in French Medical Masked Language Models?
por: Labrak, Yanis, et al.
Publicado: (2024)
por: Labrak, Yanis, et al.
Publicado: (2024)
When Answers Stray from Questions: Hallucination Detection via Question-Answer Orthogonal Decomposition
por: Yao, Siyang, et al.
Publicado: (2026)
por: Yao, Siyang, et al.
Publicado: (2026)
ReasonBENCH: Benchmarking the (In)Stability of LLM Reasoning
por: Potamitis, Nearchos, et al.
Publicado: (2025)
por: Potamitis, Nearchos, et al.
Publicado: (2025)
LLMs Have Rhythm: Fingerprinting Large Language Models Using Inter-Token Times and Network Traffic Analysis
por: Alhazbi, Saeif, et al.
Publicado: (2025)
por: Alhazbi, Saeif, et al.
Publicado: (2025)
DOTResize: Reducing LLM Width via Discrete Optimal Transport-based Neuron Merging
por: Verma, Neha, et al.
Publicado: (2025)
por: Verma, Neha, et al.
Publicado: (2025)
Merging Feed-Forward Sublayers for Compressed Transformers
por: Verma, Neha, et al.
Publicado: (2025)
por: Verma, Neha, et al.
Publicado: (2025)
Judge Circuits
por: Feldhus, Nils, et al.
Publicado: (2026)
por: Feldhus, Nils, et al.
Publicado: (2026)
On the Geometry of Positional Encodings in Transformers
por: Cirrincione, Giansalvo
Publicado: (2026)
por: Cirrincione, Giansalvo
Publicado: (2026)
InterrogateLLM: Zero-Resource Hallucination Detection in LLM-Generated Answers
por: Yehuda, Yakir, et al.
Publicado: (2024)
por: Yehuda, Yakir, et al.
Publicado: (2024)
Uncertainty-Aware Answer Selection for Improved Reasoning in Multi-LLM Systems
por: Agrawal, Aakriti, et al.
Publicado: (2025)
por: Agrawal, Aakriti, et al.
Publicado: (2025)
On the Bias of Next-Token Predictors Toward Systematically Inefficient Reasoning: A Shortest-Path Case Study
por: Alberghi, Riccardo, et al.
Publicado: (2025)
por: Alberghi, Riccardo, et al.
Publicado: (2025)
X-Token: Projection-Guided Cross-Tokenizer Knowledge Distillation
por: Sreenivas, Sharath Turuvekere, et al.
Publicado: (2026)
por: Sreenivas, Sharath Turuvekere, et al.
Publicado: (2026)
TokenShapley: Token Level Context Attribution with Shapley Value
por: Xiao, Yingtai, et al.
Publicado: (2025)
por: Xiao, Yingtai, et al.
Publicado: (2025)
Token Distillation: Attention-aware Input Embeddings For New Tokens
por: Dobler, Konstantin, et al.
Publicado: (2025)
por: Dobler, Konstantin, et al.
Publicado: (2025)
How Alignment Routes: Localizing, Scaling, and Controlling Policy Circuits in Language Models
por: Frank, Gregory N.
Publicado: (2026)
por: Frank, Gregory N.
Publicado: (2026)
Ejemplares similares
-
Why Linear Interpretability Works: Invariant Subspaces as a Result of Architectural Constraints
por: Saurez, Andres, et al.
Publicado: (2026) -
Continuous Adversarial Text Representation Learning for Affective Recognition
por: Son, Seungah, et al.
Publicado: (2025) -
Steering Sparse Autoencoder Latents to Control Dynamic Head Pruning in Vision Transformers (Student Abstract)
por: Lee, Yousung, et al.
Publicado: (2026) -
PAST: Phonetic-Acoustic Speech Tokenizer
por: Har-Tuv, Nadav, et al.
Publicado: (2025) -
Understanding Token Probability Encoding in Output Embeddings
por: Cho, Hakaze, et al.
Publicado: (2024)