Language Models Encode Numbers Using Digit Representations in Base 10
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Levy, Amit Arnold, Geva, Mor |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations
von: Huang, Jing, et al.
Veröffentlicht: (2024)
von: Huang, Jing, et al.
Veröffentlicht: (2024)
Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models
von: Ghandeharioun, Asma, et al.
Veröffentlicht: (2024)
von: Ghandeharioun, Asma, et al.
Veröffentlicht: (2024)
Constructing Interpretable Features from Compositional Neuron Groups
von: Shafran, Or, et al.
Veröffentlicht: (2025)
von: Shafran, Or, et al.
Veröffentlicht: (2025)
Backward Lens: Projecting Language Model Gradients into the Vocabulary Space
von: Katz, Shahar, et al.
Veröffentlicht: (2024)
von: Katz, Shahar, et al.
Veröffentlicht: (2024)
Routers Learn the Geometry of Their Experts: Geometric Coupling in Sparse Mixture-of-Experts
von: Ahrac, Sagi, et al.
Veröffentlicht: (2026)
von: Ahrac, Sagi, et al.
Veröffentlicht: (2026)
Mechanistically Interpretable Neural Encoding Reveals Fine-Grained Functional Selectivity in Human Visual Cortex
von: Grosbard, Idan Daniel, et al.
Veröffentlicht: (2026)
von: Grosbard, Idan Daniel, et al.
Veröffentlicht: (2026)
Towards Interpreting Visual Information Processing in Vision-Language Models
von: Neo, Clement, et al.
Veröffentlicht: (2024)
von: Neo, Clement, et al.
Veröffentlicht: (2024)
On Encoding Matrices using Quantum Circuits
von: Yosef, Liron Mor, et al.
Veröffentlicht: (2025)
von: Yosef, Liron Mor, et al.
Veröffentlicht: (2025)
Don't Blame the Annotator: Bias Already Starts in the Annotation Instructions
von: Parmar, Mihir, et al.
Veröffentlicht: (2022)
von: Parmar, Mihir, et al.
Veröffentlicht: (2022)
Inferring Functionality of Attention Heads from their Parameters
von: Elhelo, Amit, et al.
Veröffentlicht: (2024)
von: Elhelo, Amit, et al.
Veröffentlicht: (2024)
Comparison of Autoencoder Encodings for ECG Representation in Downstream Prediction Tasks
von: Harvey, Christopher J., et al.
Veröffentlicht: (2024)
von: Harvey, Christopher J., et al.
Veröffentlicht: (2024)
Large Language Models Encode Semantics and Alignment in Linearly Separable Representations
von: Saglam, Baturay, et al.
Veröffentlicht: (2025)
von: Saglam, Baturay, et al.
Veröffentlicht: (2025)
Blurred Encoding for Trajectory Representation Learning
von: Zhou, Silin, et al.
Veröffentlicht: (2025)
von: Zhou, Silin, et al.
Veröffentlicht: (2025)
An Analytical Model for Overparameterized Learning Under Class Imbalance
von: Mor, Eliav, et al.
Veröffentlicht: (2025)
von: Mor, Eliav, et al.
Veröffentlicht: (2025)
Estimating Knowledge in Large Language Models Without Generating a Single Token
von: Gottesman, Daniela, et al.
Veröffentlicht: (2024)
von: Gottesman, Daniela, et al.
Veröffentlicht: (2024)
When Can Transformers Count to n?
von: Yehudai, Gilad, et al.
Veröffentlicht: (2024)
von: Yehudai, Gilad, et al.
Veröffentlicht: (2024)
The Condition Number as a Scale-Invariant Proxy for Information Encoding in Neural Units
von: Ludwig, Oswaldo
Veröffentlicht: (2025)
von: Ludwig, Oswaldo
Veröffentlicht: (2025)
On the Power of Randomization in Fair Classification and Representation
von: Agarwal, Sushant, et al.
Veröffentlicht: (2024)
von: Agarwal, Sushant, et al.
Veröffentlicht: (2024)
Physics Encoded Blocks in Residual Neural Network Architectures for Digital Twin Models
von: Zia, Muhammad Saad, et al.
Veröffentlicht: (2024)
von: Zia, Muhammad Saad, et al.
Veröffentlicht: (2024)
Transpose Attack: Stealing Datasets with Bidirectional Training
von: Amit, Guy, et al.
Veröffentlicht: (2023)
von: Amit, Guy, et al.
Veröffentlicht: (2023)
Pre-trained Language Models Learn Remarkably Accurate Representations of Numbers
von: Kadlčík, Marek, et al.
Veröffentlicht: (2025)
von: Kadlčík, Marek, et al.
Veröffentlicht: (2025)
Extracting and Encoding: Leveraging Large Language Models and Medical Knowledge to Enhance Radiological Text Representation
von: Messina, Pablo, et al.
Veröffentlicht: (2024)
von: Messina, Pablo, et al.
Veröffentlicht: (2024)
Convergent Evolution: How Different Language Models Learn Similar Number Representations
von: Fu, Deqing, et al.
Veröffentlicht: (2026)
von: Fu, Deqing, et al.
Veröffentlicht: (2026)
Group Representational Position Encoding
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
von: Zhang, Yifan, et al.
Veröffentlicht: (2025)
Evaluating Large Language Models for Gait Classification Using Text-Encoded Kinematic Waveforms
von: Dindorf, Carlo, et al.
Veröffentlicht: (2026)
von: Dindorf, Carlo, et al.
Veröffentlicht: (2026)
Causality-Induced Positional Encoding for Transformer-Based Representation Learning of Non-Sequential Features
von: Xu, Kaichen, et al.
Veröffentlicht: (2025)
von: Xu, Kaichen, et al.
Veröffentlicht: (2025)
Dataset Distillation Efficiently Encodes Low-Dimensional Representations from Gradient-Based Learning of Non-Linear Tasks
von: Kinoshita, Yuri, et al.
Veröffentlicht: (2026)
von: Kinoshita, Yuri, et al.
Veröffentlicht: (2026)
Semimage: HSV-Based Semantic Image Encoding for Disentangled Text Representation
von: Zare, Mohammad
Veröffentlicht: (2025)
von: Zare, Mohammad
Veröffentlicht: (2025)
LASERS: LAtent Space Encoding for Representations with Sparsity for Generative Modeling
von: Li, Xin, et al.
Veröffentlicht: (2024)
von: Li, Xin, et al.
Veröffentlicht: (2024)
Positional Encoding in Transformer-Based Time Series Models: A Survey
von: Irani, Habib, et al.
Veröffentlicht: (2025)
von: Irani, Habib, et al.
Veröffentlicht: (2025)
Emergent Stack Representations in Modeling Counter Languages Using Transformers
von: Tiwari, Utkarsh, et al.
Veröffentlicht: (2025)
von: Tiwari, Utkarsh, et al.
Veröffentlicht: (2025)
AI Accelerators for Large Language Model Inference: Architecture Analysis and Scaling Strategies
von: Sharma, Amit
Veröffentlicht: (2025)
von: Sharma, Amit
Veröffentlicht: (2025)
Colorful Talks with Graphs: Human-Interpretable Graph Encodings for Large Language Models
von: Zangari, Angelo, et al.
Veröffentlicht: (2026)
von: Zangari, Angelo, et al.
Veröffentlicht: (2026)
Vision-Language Models Encode Clinical Guidelines for Concept-Based Medical Reasoning
von: Harmanani, Mohamed, et al.
Veröffentlicht: (2026)
von: Harmanani, Mohamed, et al.
Veröffentlicht: (2026)
TabText: Language-Based Representations of Tabular Health Data for Predictive Modelling
von: Carballo, Kimberly Villalobos, et al.
Veröffentlicht: (2022)
von: Carballo, Kimberly Villalobos, et al.
Veröffentlicht: (2022)
HyperHELM: Hyperbolic Hierarchy Encoding for mRNA Language Modeling
von: van Spengler, Max, et al.
Veröffentlicht: (2025)
von: van Spengler, Max, et al.
Veröffentlicht: (2025)
Data-driven Circuit Discovery for Interpretability of Language Models
von: Rai, Daking, et al.
Veröffentlicht: (2026)
von: Rai, Daking, et al.
Veröffentlicht: (2026)
Multilingual Language Models Encode Script Over Linguistic Structure
von: Verma, Aastha A K, et al.
Veröffentlicht: (2026)
von: Verma, Aastha A K, et al.
Veröffentlicht: (2026)
Zero-Direction Probing: A Linear-Algebraic Framework for Deep Analysis of Large-Language-Model Drift
von: Pandey, Amit
Veröffentlicht: (2025)
von: Pandey, Amit
Veröffentlicht: (2025)
Molecular Representations for Large Language Models
von: Runcie, Nicholas T., et al.
Veröffentlicht: (2026)
von: Runcie, Nicholas T., et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
RAVEL: Evaluating Interpretability Methods on Disentangling Language Model Representations
von: Huang, Jing, et al.
Veröffentlicht: (2024) -
Patchscopes: A Unifying Framework for Inspecting Hidden Representations of Language Models
von: Ghandeharioun, Asma, et al.
Veröffentlicht: (2024) -
Constructing Interpretable Features from Compositional Neuron Groups
von: Shafran, Or, et al.
Veröffentlicht: (2025) -
Backward Lens: Projecting Language Model Gradients into the Vocabulary Space
von: Katz, Shahar, et al.
Veröffentlicht: (2024) -
Routers Learn the Geometry of Their Experts: Geometric Coupling in Sparse Mixture-of-Experts
von: Ahrac, Sagi, et al.
Veröffentlicht: (2026)