Algorithmic Capabilities of Random Transformers
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Zhong, Ziqian, Andreas, Jacob |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2024
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
The Geometric Anatomy of Capability Acquisition in Transformers
von: Billa, Jayadev
Veröffentlicht: (2026)
von: Billa, Jayadev
Veröffentlicht: (2026)
Base Models Look Human To AI Detectors
von: Xu, Yixuan Even, et al.
Veröffentlicht: (2026)
von: Xu, Yixuan Even, et al.
Veröffentlicht: (2026)
Toward In-Context Teaching: Adapting Examples to Students' Misconceptions
von: Ross, Alexis, et al.
Veröffentlicht: (2024)
von: Ross, Alexis, et al.
Veröffentlicht: (2024)
MMLU-SR: A Benchmark for Stress-Testing Reasoning Capability of Large Language Models
von: Wang, Wentian, et al.
Veröffentlicht: (2024)
von: Wang, Wentian, et al.
Veröffentlicht: (2024)
Learning a Decision Tree Algorithm with Transformers
von: Zhuang, Yufan, et al.
Veröffentlicht: (2024)
von: Zhuang, Yufan, et al.
Veröffentlicht: (2024)
Discovering Interpretable Algorithms by Decompiling Transformers to RASP
von: Huang, Xinting, et al.
Veröffentlicht: (2026)
von: Huang, Xinting, et al.
Veröffentlicht: (2026)
Limits of Transformer Language Models on Learning to Compose Algorithms
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
von: Thomm, Jonathan, et al.
Veröffentlicht: (2024)
A Hitchhiker's Guide to Scaling Law Estimation
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
von: Choshen, Leshem, et al.
Veröffentlicht: (2024)
Lexicon-Level Contrastive Visual-Grounding Improves Language Modeling
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2024)
von: Zhuang, Chengxu, et al.
Veröffentlicht: (2024)
Depth-Width tradeoffs in Algorithmic Reasoning of Graph Tasks with Transformers
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
von: Yehudai, Gilad, et al.
Veröffentlicht: (2025)
ALTA: Compiler-Based Analysis of Transformers
von: Shaw, Peter, et al.
Veröffentlicht: (2024)
von: Shaw, Peter, et al.
Veröffentlicht: (2024)
Explaining Datasets in Words: Statistical Models with Natural Language Parameters
von: Zhong, Ruiqi, et al.
Veröffentlicht: (2024)
von: Zhong, Ruiqi, et al.
Veröffentlicht: (2024)
Training Language Models to Explain Their Own Computations
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
Bridging Kolmogorov Complexity and Deep Learning: Asymptotically Optimal Description Length Objectives for Transformers
von: Shaw, Peter, et al.
Veröffentlicht: (2025)
von: Shaw, Peter, et al.
Veröffentlicht: (2025)
ConsistentEE: A Consistent and Hardness-Guided Early Exiting Method for Accelerating Language Models Inference
von: Zeng, Ziqian, et al.
Veröffentlicht: (2023)
von: Zeng, Ziqian, et al.
Veröffentlicht: (2023)
PMPO: Probabilistic Metric Prompt Optimization for Small and Large Language Models
von: Zhao, Chenzhuo, et al.
Veröffentlicht: (2025)
von: Zhao, Chenzhuo, et al.
Veröffentlicht: (2025)
Policy Learning with a Language Bottleneck
von: Srivastava, Megha, et al.
Veröffentlicht: (2024)
von: Srivastava, Megha, et al.
Veröffentlicht: (2024)
(How) Do Language Models Track State?
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
von: Li, Belinda Z., et al.
Veröffentlicht: (2025)
Out-of-distribution generalization via composition: a lens through induction heads in Transformers
von: Song, Jiajun, et al.
Veröffentlicht: (2024)
von: Song, Jiajun, et al.
Veröffentlicht: (2024)
Neural Algorithmic Reasoning for Hypergraphs with Looped Transformers
von: Huang, Zekai, et al.
Veröffentlicht: (2025)
von: Huang, Zekai, et al.
Veröffentlicht: (2025)
Quantifying the Capabilities of LLMs across Scale and Precision
von: Badshah, Sher, et al.
Veröffentlicht: (2024)
von: Badshah, Sher, et al.
Veröffentlicht: (2024)
Exploring and Benchmarking the Planning Capabilities of Large Language Models
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
von: Bohnet, Bernd, et al.
Veröffentlicht: (2024)
AI Scientists Fail Without Strong Implementation Capability
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
von: Zhu, Minjun, et al.
Veröffentlicht: (2025)
On Calibration of Large Language Models: From Response To Capability
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
von: Yang, Sin-Han, et al.
Veröffentlicht: (2026)
Prescriptive Scaling Reveals the Evolution of Language Model Capabilities
von: Zhang, Hanlin, et al.
Veröffentlicht: (2026)
von: Zhang, Hanlin, et al.
Veröffentlicht: (2026)
Learning How Hard to Think: Input-Adaptive Allocation of LM Computation
von: Damani, Mehul, et al.
Veröffentlicht: (2024)
von: Damani, Mehul, et al.
Veröffentlicht: (2024)
Sample-Efficient Online Learning in LM Agents via Hindsight Trajectory Rewriting
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
von: Hu, Michael Y., et al.
Veröffentlicht: (2025)
ALPINE: Unveiling the Planning Capability of Autoregressive Learning in Language Models
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
von: Wang, Siwei, et al.
Veröffentlicht: (2024)
ForecastBench: A Dynamic Benchmark of AI Forecasting Capabilities
von: Karger, Ezra, et al.
Veröffentlicht: (2024)
von: Karger, Ezra, et al.
Veröffentlicht: (2024)
ModelGPT: Unleashing LLM's Capabilities for Tailored Model Generation
von: Tang, Zihao, et al.
Veröffentlicht: (2024)
von: Tang, Zihao, et al.
Veröffentlicht: (2024)
What is it for a Machine Learning Model to Have a Capability?
von: Harding, Jacqueline, et al.
Veröffentlicht: (2024)
von: Harding, Jacqueline, et al.
Veröffentlicht: (2024)
How Numerical Precision Affects Arithmetical Reasoning Capabilities of LLMs
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
von: Feng, Guhao, et al.
Veröffentlicht: (2024)
Unlocking Reasoning Capabilities in LLMs via Reinforcement Learning Exploration
von: Deng, Wenhao, et al.
Veröffentlicht: (2025)
von: Deng, Wenhao, et al.
Veröffentlicht: (2025)
Automated Capability Discovery via Foundation Model Self-Exploration
von: Lu, Cong, et al.
Veröffentlicht: (2025)
von: Lu, Cong, et al.
Veröffentlicht: (2025)
Enhancing Reasoning Capabilities in SLMs with Reward Guided Dataset Distillation
von: Padarha, Shreyansh
Veröffentlicht: (2025)
von: Padarha, Shreyansh
Veröffentlicht: (2025)
Assessing the Impact of Prompting Methods on ChatGPT's Mathematical Capabilities
von: Chen, Yuhao, et al.
Veröffentlicht: (2023)
von: Chen, Yuhao, et al.
Veröffentlicht: (2023)
Jailbreaking in the Haystack
von: Shah, Rishi Rajesh, et al.
Veröffentlicht: (2025)
von: Shah, Rishi Rajesh, et al.
Veröffentlicht: (2025)
CityBench: Evaluating the Capabilities of Large Language Models for Urban Tasks
von: Feng, Jie, et al.
Veröffentlicht: (2024)
von: Feng, Jie, et al.
Veröffentlicht: (2024)
Disentangling Logic: The Role of Context in Large Language Model Reasoning Capabilities
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
von: Hua, Wenyue, et al.
Veröffentlicht: (2024)
LimiX: Unleashing Structured-Data Modeling Capability for Generalist Intelligence
von: Zhang, Xingxuan, et al.
Veröffentlicht: (2025)
von: Zhang, Xingxuan, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
The Geometric Anatomy of Capability Acquisition in Transformers
von: Billa, Jayadev
Veröffentlicht: (2026) -
Base Models Look Human To AI Detectors
von: Xu, Yixuan Even, et al.
Veröffentlicht: (2026) -
Toward In-Context Teaching: Adapting Examples to Students' Misconceptions
von: Ross, Alexis, et al.
Veröffentlicht: (2024) -
MMLU-SR: A Benchmark for Stress-Testing Reasoning Capability of Large Language Models
von: Wang, Wentian, et al.
Veröffentlicht: (2024) -
Learning a Decision Tree Algorithm with Transformers
von: Zhuang, Yufan, et al.
Veröffentlicht: (2024)