An empirical study on the limitation of Transformers in program trace generation
Fuente:
arXiv
Guardado en:
| Autor principal: | Sun, Simeng |
|---|---|
| Formato: | Preprint |
| Publicado: |
2025
|
| Materias: | |
| Acceso en línea: | |
| Etiquetas: |
Agregar Etiqueta
Sin Etiquetas, Sea el primero en etiquetar este registro!
|
Ejemplares similares
How much do contextualized representations encode long-range context?
por: Sun, Simeng, et al.
Publicado: (2024)
por: Sun, Simeng, et al.
Publicado: (2024)
Suri: Multi-constraint Instruction Following for Long-form Text Generation
por: Pham, Chau Minh, et al.
Publicado: (2024)
por: Pham, Chau Minh, et al.
Publicado: (2024)
SugarTextNet: A Transformer-Based Framework for Detecting Sugar Dating-Related Content on Social Media with Context-Aware Focal Loss
por: Wang, Lionel Z., et al.
Publicado: (2025)
por: Wang, Lionel Z., et al.
Publicado: (2025)
TopicGPT: A Prompt-based Topic Modeling Framework
por: Pham, Chau Minh, et al.
Publicado: (2023)
por: Pham, Chau Minh, et al.
Publicado: (2023)
L0-Reasoning Bench: Evaluating Procedural Correctness in Language Models via Simple Program Execution
por: Sun, Simeng, et al.
Publicado: (2025)
por: Sun, Simeng, et al.
Publicado: (2025)
The optimality of word lengths. Theoretical foundations and an empirical study
por: Petrini, Sonia, et al.
Publicado: (2022)
por: Petrini, Sonia, et al.
Publicado: (2022)
Comparing large language models and human programmers for generating programming code
por: Hou, Wenpin, et al.
Publicado: (2024)
por: Hou, Wenpin, et al.
Publicado: (2024)
Combining psychoanalysis and computer science: an empirical study of the relationship between emotions and the Lacanian discourses
por: Gadalla, Minas, et al.
Publicado: (2024)
por: Gadalla, Minas, et al.
Publicado: (2024)
HYBRIDMIND: Meta Selection of Natural Language and Symbolic Language for Enhanced LLM Reasoning
por: Han, Simeng, et al.
Publicado: (2024)
por: Han, Simeng, et al.
Publicado: (2024)
RULER: What's the Real Context Size of Your Long-Context Language Models?
por: Hsieh, Cheng-Ping, et al.
Publicado: (2024)
por: Hsieh, Cheng-Ping, et al.
Publicado: (2024)
Learning to Reason via Mixture-of-Thought for Logical Reasoning
por: Zheng, Tong, et al.
Publicado: (2025)
por: Zheng, Tong, et al.
Publicado: (2025)
Are most sentences unique? An empirical examination of Chomskyan claims
por: Ring, Hiram
Publicado: (2025)
por: Ring, Hiram
Publicado: (2025)
Assessing Pause Thresholds for empirical Translation Process Research
por: Bandaru, Devi Sri, et al.
Publicado: (2026)
por: Bandaru, Devi Sri, et al.
Publicado: (2026)
Evaluating Legal Reasoning Traces with Legal Issue Tree Rubrics
por: Lee, Jinu, et al.
Publicado: (2025)
por: Lee, Jinu, et al.
Publicado: (2025)
Multilingual Generative Retrieval via Cross-lingual Semantic Compression
por: Huang, Yuxin, et al.
Publicado: (2025)
por: Huang, Yuxin, et al.
Publicado: (2025)
Scheherazade: Evaluating Chain-of-Thought Math Reasoning in LLMs with Chain-of-Problems
por: Miner, Stephen, et al.
Publicado: (2024)
por: Miner, Stephen, et al.
Publicado: (2024)
The effect of source disclosure on evaluation of AI-generated messages: A two-part study
por: Lim, Sue, et al.
Publicado: (2023)
por: Lim, Sue, et al.
Publicado: (2023)
Transformer Layers as Painters
por: Sun, Qi, et al.
Publicado: (2024)
por: Sun, Qi, et al.
Publicado: (2024)
Fact-checking AI-generated news reports: Can LLMs catch their own lies?
por: Yao, Jiayi, et al.
Publicado: (2025)
por: Yao, Jiayi, et al.
Publicado: (2025)
Can reasoning models comprehend mathematical problems in Chinese ancient texts? An empirical study based on data from Suanjing Shishu
por: Liu, Chang, et al.
Publicado: (2025)
por: Liu, Chang, et al.
Publicado: (2025)
Probabilistic energy profiler for statically typed JVM-based programming languages
por: Nyholm, Joel, et al.
Publicado: (2025)
por: Nyholm, Joel, et al.
Publicado: (2025)
Optimizing Language Model's Reasoning Abilities with Weak Supervision
por: Tong, Yongqi, et al.
Publicado: (2024)
por: Tong, Yongqi, et al.
Publicado: (2024)
Statistical investigations into the geometry and homology of random programs
por: Sporring, Jon, et al.
Publicado: (2024)
por: Sporring, Jon, et al.
Publicado: (2024)
SWAN-GPT: An Efficient and Scalable Approach for Long-Context Language Modeling
por: Puvvada, Krishna C., et al.
Publicado: (2025)
por: Puvvada, Krishna C., et al.
Publicado: (2025)
Empirical study of pretrained multilingual language models for zero-shot cross-lingual knowledge transfer in generation
por: Chirkova, Nadezhda, et al.
Publicado: (2023)
por: Chirkova, Nadezhda, et al.
Publicado: (2023)
Controllable Text Generation with Residual Memory Transformer
por: Zhang, Hanqing, et al.
Publicado: (2023)
por: Zhang, Hanqing, et al.
Publicado: (2023)
Deep learning and abstractive summarisation for radiological reports: an empirical study for adapting the PEGASUS models' family with scarce data
por: Benzoni, Claudio, et al.
Publicado: (2025)
por: Benzoni, Claudio, et al.
Publicado: (2025)
ATEB: Evaluating and Improving Advanced NLP Tasks for Text Embedding Models
por: Han, Simeng, et al.
Publicado: (2025)
por: Han, Simeng, et al.
Publicado: (2025)
Minimizing Mismatch Risk: A Prototype-Based Routing Framework for Zero-shot LLM-generated Text Detection
por: Sun, Ke, et al.
Publicado: (2026)
por: Sun, Ke, et al.
Publicado: (2026)
Mixture of Hidden-Dimensions Transformer
por: Chen, Yilong, et al.
Publicado: (2024)
por: Chen, Yilong, et al.
Publicado: (2024)
Metronome: tracing variation in poetic meters via local sequence alignment
por: Nagy, Ben, et al.
Publicado: (2024)
por: Nagy, Ben, et al.
Publicado: (2024)
Grokked Transformers are Implicit Reasoners: A Mechanistic Journey to the Edge of Generalization
por: Wang, Boshi, et al.
Publicado: (2024)
por: Wang, Boshi, et al.
Publicado: (2024)
H$^{2}$MT: Semantic Hierarchy-Aware Hierarchical Memory Transformer
por: Haghifam, Maryam, et al.
Publicado: (2026)
por: Haghifam, Maryam, et al.
Publicado: (2026)
Evaluation of GPT-based large language generative AI models as study aids for the national licensure examination for registered dietitians in Japan
por: Nagamori, Yuta, et al.
Publicado: (2025)
por: Nagamori, Yuta, et al.
Publicado: (2025)
MappedTrace: Tracing Pointer Remotely with Compiler-generated Maps
por: Ma, Zhiyao, et al.
Publicado: (2025)
por: Ma, Zhiyao, et al.
Publicado: (2025)
Automatic generation of DRI Statements
por: Flechtner, Maurice
Publicado: (2025)
por: Flechtner, Maurice
Publicado: (2025)
Towards an empirical understanding of MoE design choices
por: Fan, Dongyang, et al.
Publicado: (2024)
por: Fan, Dongyang, et al.
Publicado: (2024)
Early Transformers: A study on Efficient Training of Transformer Models through Early-Bird Lottery Tickets
por: Cheekati, Shravan
Publicado: (2024)
por: Cheekati, Shravan
Publicado: (2024)
The creative psychometric item generator: a framework for item generation and validation using large language models
por: Laverghetta Jr., Antonio, et al.
Publicado: (2024)
por: Laverghetta Jr., Antonio, et al.
Publicado: (2024)
Linguistic traces of stochastic empathy in language models
por: Kleinberg, Bennett, et al.
Publicado: (2024)
por: Kleinberg, Bennett, et al.
Publicado: (2024)
Ejemplares similares
-
How much do contextualized representations encode long-range context?
por: Sun, Simeng, et al.
Publicado: (2024) -
Suri: Multi-constraint Instruction Following for Long-form Text Generation
por: Pham, Chau Minh, et al.
Publicado: (2024) -
SugarTextNet: A Transformer-Based Framework for Detecting Sugar Dating-Related Content on Social Media with Context-Aware Focal Loss
por: Wang, Lionel Z., et al.
Publicado: (2025) -
TopicGPT: A Prompt-based Topic Modeling Framework
por: Pham, Chau Minh, et al.
Publicado: (2023) -
L0-Reasoning Bench: Evaluating Procedural Correctness in Language Models via Simple Program Execution
por: Sun, Simeng, et al.
Publicado: (2025)