ALTA: Compiler-Based Analysis of Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Shaw, Peter, Cohan, James, Eisenstein, Jacob, Lee, Kenton, Berant, Jonathan, Toutanova, Kristina |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformers
von: Shaw, Peter, et al.
Veröffentlicht: (2025)
von: Shaw, Peter, et al.
Veröffentlicht: (2025)
Effective Reasoning Chains Reduce Intrinsic Dimensionality
von: Prasad, Archiki, et al.
Veröffentlicht: (2026)
von: Prasad, Archiki, et al.
Veröffentlicht: (2026)
BgGPT 1.0: Extending English-centric LLMs to other languages
von: Alexandrov, Anton, et al.
Veröffentlicht: (2024)
von: Alexandrov, Anton, et al.
Veröffentlicht: (2024)
MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games
von: Eisenstein, Jacob, et al.
Veröffentlicht: (2026)
von: Eisenstein, Jacob, et al.
Veröffentlicht: (2026)
Robust Preference Optimization through Reward Model Distillation
von: Fisch, Adam, et al.
Veröffentlicht: (2024)
von: Fisch, Adam, et al.
Veröffentlicht: (2024)
Plantain: Plan-Answer Interleaved Reasoning
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
von: Liang, Anthony, et al.
Veröffentlicht: (2025)
SEMQA: Semi-Extractive Multi-Source Question Answering
von: Schuster, Tal, et al.
Veröffentlicht: (2023)
von: Schuster, Tal, et al.
Veröffentlicht: (2023)
On Evaluating LLM Alignment by Evaluating LLMs as Judges
von: Liu, Yixin, et al.
Veröffentlicht: (2025)
von: Liu, Yixin, et al.
Veröffentlicht: (2025)
Re-evaluating Automatic LLM System Ranking for Alignment with Human Preference
von: Gao, Mingqi, et al.
Veröffentlicht: (2024)
von: Gao, Mingqi, et al.
Veröffentlicht: (2024)
Transforming and Combining Rewards for Aligning Large Language Models
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
von: Wang, Zihao, et al.
Veröffentlicht: (2024)
Algorithmic Capabilities of Random Transformers
von: Zhong, Ziqian, et al.
Veröffentlicht: (2024)
von: Zhong, Ziqian, et al.
Veröffentlicht: (2024)
Observable Propagation: Uncovering Feature Vectors in Transformers
von: Dunefsky, Jacob, et al.
Veröffentlicht: (2023)
von: Dunefsky, Jacob, et al.
Veröffentlicht: (2023)
Calibrating Long-form Generations from Large Language Models
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
von: Huang, Yukun, et al.
Veröffentlicht: (2024)
Optimizing Pre-Training Data Mixtures with Mixtures of Data Expert Models
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
von: Belenki, Lior, et al.
Veröffentlicht: (2025)
Don't lie to your friends: Learning what you know from collaborative self-play
von: Eisenstein, Jacob, et al.
Veröffentlicht: (2025)
von: Eisenstein, Jacob, et al.
Veröffentlicht: (2025)
References Improve LLM Alignment in Non-Verifiable Domains
von: Shi, Kejian, et al.
Veröffentlicht: (2026)
von: Shi, Kejian, et al.
Veröffentlicht: (2026)
Can Large Language Models Understand Intermediate Representations in Compilers?
von: Jiang, Hailong, et al.
Veröffentlicht: (2025)
von: Jiang, Hailong, et al.
Veröffentlicht: (2025)
Latent Context Compilation: Distilling Long Context into Compact Portable Memory
von: Li, Zeju, et al.
Veröffentlicht: (2026)
von: Li, Zeju, et al.
Veröffentlicht: (2026)
The Anxiety of Influence: Bloom Filters in Transformer Attention Heads
von: Balogh, Peter
Veröffentlicht: (2026)
von: Balogh, Peter
Veröffentlicht: (2026)
Contrast-CAT: Contrasting Activations for Enhanced Interpretability in Transformer-based Text Classifiers
von: Han, Sungmin, et al.
Veröffentlicht: (2025)
von: Han, Sungmin, et al.
Veröffentlicht: (2025)
Optimizing Diversity and Quality through Base-Aligned Model Collaboration
von: Wang, Yichen, et al.
Veröffentlicht: (2025)
von: Wang, Yichen, et al.
Veröffentlicht: (2025)
Limits of Transformer Language Models on Learning to Compose Algorithms
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
Rewarding Progress: Scaling Automated Process Verifiers for LLM Reasoning
von: Setlur, Amrith, et al.
Veröffentlicht: (2024)
von: Setlur, Amrith, et al.
Veröffentlicht: (2024)
ReIFE: Re-evaluating Instruction-Following Evaluation
von: Liu, Yixin, et al.
Veröffentlicht: (2024)
von: Liu, Yixin, et al.
Veröffentlicht: (2024)
Survey on Evaluation of LLM-based Agents
von: Yehudai, Asaf, et al.
Veröffentlicht: (2025)
von: Yehudai, Asaf, et al.
Veröffentlicht: (2025)
AIOS Compiler: LLM as Interpreter for Natural Language Programming and Flow Programming of AI Agents
von: Xu, Shuyuan, et al.
Veröffentlicht: (2024)
von: Xu, Shuyuan, et al.
Veröffentlicht: (2024)
Symbolic Prompt Program Search: A Structure-Aware Approach to Efficient Compile-Time Prompt Optimization
von: Schnabel, Tobias, et al.
Veröffentlicht: (2024)
von: Schnabel, Tobias, et al.
Veröffentlicht: (2024)
An Enhanced Dual Transformer Contrastive Network for Multimodal Sentiment Analysis
von: Dao, Phuong Q., et al.
Veröffentlicht: (2025)
von: Dao, Phuong Q., et al.
Veröffentlicht: (2025)
Power Transformer Fault Prediction Based on Knowledge Graphs
von: Wang, Chao, et al.
Veröffentlicht: (2024)
von: Wang, Chao, et al.
Veröffentlicht: (2024)
Compile to Compress: Boosting Formal Theorem Provers by Compiler Outputs
von: Li, Guchan, et al.
Veröffentlicht: (2026)
von: Li, Guchan, et al.
Veröffentlicht: (2026)
Beyond Components: Singular Vector-Based Interpretability of Transformer Circuits
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
von: Ahmad, Areeb, et al.
Veröffentlicht: (2025)
Classification of Hope in Textual Data using Transformer-Based Models
von: Ijezue, Chukwuebuka Fortunate, et al.
Veröffentlicht: (2025)
von: Ijezue, Chukwuebuka Fortunate, et al.
Veröffentlicht: (2025)
Latent Chain-of-Thought? Decoding the Depth-Recurrent Transformer
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
von: Lu, Wenquan, et al.
Veröffentlicht: (2025)
Theoretical guarantees on the best-of-n alignment policy
von: Beirami, Ahmad, et al.
Veröffentlicht: (2024)
von: Beirami, Ahmad, et al.
Veröffentlicht: (2024)
Integrating Locality-Aware Attention with Transformers for General Geometry PDEs
von: Koh, Minsu, et al.
Veröffentlicht: (2025)
von: Koh, Minsu, et al.
Veröffentlicht: (2025)
Examining Reasoning LLMs-as-Judges in Non-Verifiable LLM Post-Training
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
von: Liu, Yixin, et al.
Veröffentlicht: (2026)
CAST: Compositional Analysis via Spectral Tracking for Understanding Transformer Layer Functions
von: Fu, Zihao, et al.
Veröffentlicht: (2025)
von: Fu, Zihao, et al.
Veröffentlicht: (2025)
Interpreting Key Mechanisms of Factual Recall in Transformer-Based Language Models
von: Lv, Ang, et al.
Veröffentlicht: (2024)
von: Lv, Ang, et al.
Veröffentlicht: (2024)
Transformer-Based Multimodal Knowledge Graph Completion with Link-Aware Contexts
von: Ma, Haodi, et al.
Veröffentlicht: (2025)
von: Ma, Haodi, et al.
Veröffentlicht: (2025)
One-shot Optimized Steering Vectors Mediate Safety-relevant Behaviors in LLMs
von: Dunefsky, Jacob, et al.
Veröffentlicht: (2025)
von: Dunefsky, Jacob, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformers
von: Shaw, Peter, et al.
Veröffentlicht: (2025) -
Effective Reasoning Chains Reduce Intrinsic Dimensionality
von: Prasad, Archiki, et al.
Veröffentlicht: (2026) -
BgGPT 1.0: Extending English-centric LLMs to other languages
von: Alexandrov, Anton, et al.
Veröffentlicht: (2024) -
MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games
von: Eisenstein, Jacob, et al.
Veröffentlicht: (2026) -
Robust Preference Optimization through Reward Model Distillation
von: Fisch, Adam, et al.
Veröffentlicht: (2024)