Mechanistic Understanding of Language Models in Syntactic Code Completion
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Miller, Samuel, Rai, Daking, Yao, Ziyu |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Failure by Interference: Language Models Make Balanced Parentheses Errors When Faulty Mechanisms Overshadow Sound Ones
von: Rai, Daking, et al.
Veröffentlicht: (2025)
von: Rai, Daking, et al.
Veröffentlicht: (2025)
A Practical Review of Mechanistic Interpretability for Transformer-Based Language Models
von: Rai, Daking, et al.
Veröffentlicht: (2024)
von: Rai, Daking, et al.
Veröffentlicht: (2024)
An Investigation of Neuron Activation as a Unified Lens to Explain Chain-of-Thought Eliciting Arithmetic Reasoning of LLMs
von: Rai, Daking, et al.
Veröffentlicht: (2024)
von: Rai, Daking, et al.
Veröffentlicht: (2024)
Vibe Code Bench: Evaluating AI Models on End-to-End Web Application Development
von: Tran, Hung, et al.
Veröffentlicht: (2026)
von: Tran, Hung, et al.
Veröffentlicht: (2026)
When Retrieval Hurts Code Completion: A Diagnostic Study of Stale Repository Context
von: Weng, Haojun, et al.
Veröffentlicht: (2026)
von: Weng, Haojun, et al.
Veröffentlicht: (2026)
Data-driven Circuit Discovery for Interpretability of Language Models
von: Rai, Daking, et al.
Veröffentlicht: (2026)
von: Rai, Daking, et al.
Veröffentlicht: (2026)
Automated Web Application Testing: End-to-End Test Case Generation with Large Language Models and Screen Transition Graphs
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Plan with Code: Comparing approaches for robust NL to DSL generation
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
A Comparative Study of DSL Code Generation: Fine-Tuning vs. Optimized Retrieval Augmentation
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
von: Bassamzadeh, Nastaran, et al.
Veröffentlicht: (2024)
All for One: LLMs Solve Mental Math at the Last Token With Information Transferred From Other Tokens
von: Mamidanna, Siddarth, et al.
Veröffentlicht: (2025)
von: Mamidanna, Siddarth, et al.
Veröffentlicht: (2025)
Tool-Schema Compression Enables Agentic RAG Under Constrained Context Budgets
von: Sakizli, Furkan
Veröffentlicht: (2026)
von: Sakizli, Furkan
Veröffentlicht: (2026)
Narrow Transformer: StarCoder-Based Java-LM For Desktop
von: Rathinasamy, Kamalkumar, et al.
Veröffentlicht: (2024)
von: Rathinasamy, Kamalkumar, et al.
Veröffentlicht: (2024)
Simple and Effective Baselines for Code Summarisation Evaluation
von: Robinson, Jade, et al.
Veröffentlicht: (2025)
von: Robinson, Jade, et al.
Veröffentlicht: (2025)
REPOT: Recoverable Program-of-Thought via Checkpoint Repair
von: Mazaheri, Parsa
Veröffentlicht: (2026)
von: Mazaheri, Parsa
Veröffentlicht: (2026)
AcTracer: Active Testing of Large Language Model via Multi-Stage Sampling
von: Huang, Yuheng, et al.
Veröffentlicht: (2024)
von: Huang, Yuheng, et al.
Veröffentlicht: (2024)
Engineering A Large Language Model From Scratch
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
von: Oketunji, Abiodun Finbarrs
Veröffentlicht: (2024)
Learning Software Bug Reports: A Systematic Literature Review
von: Long, Guoming, et al.
Veröffentlicht: (2025)
von: Long, Guoming, et al.
Veröffentlicht: (2025)
Distilling Desired Comments for Enhanced Code Review with Large Language Models
von: Yu, Yongda, et al.
Veröffentlicht: (2024)
von: Yu, Yongda, et al.
Veröffentlicht: (2024)
Improved IR-based Bug Localization with Intelligent Relevance Feedback
von: Samir, Asif Mohammed, et al.
Veröffentlicht: (2025)
von: Samir, Asif Mohammed, et al.
Veröffentlicht: (2025)
Safety Under Scaffolding: How Evaluation Conditions Shape Measured Safety
von: Gringras, David
Veröffentlicht: (2026)
von: Gringras, David
Veröffentlicht: (2026)
CIDR: A Large-Scale Industrial Source Code Dataset for Software Engineering Research
von: Savenkov, Vladislav
Veröffentlicht: (2026)
von: Savenkov, Vladislav
Veröffentlicht: (2026)
LLM4PLC: Harnessing Large Language Models for Verifiable Programming of PLCs in Industrial Control Systems
von: Fakih, Mohamad, et al.
Veröffentlicht: (2024)
von: Fakih, Mohamad, et al.
Veröffentlicht: (2024)
Comprehensive Evaluation and Insights into the Use of Large Language Models in the Automation of Behavior-Driven Development Acceptance Test Formulation
von: Karpurapu, Shanthi, et al.
Veröffentlicht: (2024)
von: Karpurapu, Shanthi, et al.
Veröffentlicht: (2024)
CoTran: An LLM-based Code Translator using Reinforcement Learning with Feedback from Compiler and Symbolic Execution
von: Jana, Prithwish, et al.
Veröffentlicht: (2023)
von: Jana, Prithwish, et al.
Veröffentlicht: (2023)
LLMORPH: Automated Metamorphic Testing of Large Language Models
von: Cho, Steven, et al.
Veröffentlicht: (2026)
von: Cho, Steven, et al.
Veröffentlicht: (2026)
A Framework for Testing and Adapting REST APIs as LLM Tools
von: Bandlamudi, Jayachandu, et al.
Veröffentlicht: (2025)
von: Bandlamudi, Jayachandu, et al.
Veröffentlicht: (2025)
Finetuning LLMs for Automatic Form Interaction on Web-Browser in Selenium Testing Framework
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
von: Le, Nguyen-Khang, et al.
Veröffentlicht: (2025)
Mind the GAP: Text Safety Does Not Transfer to Tool-Call Safety in LLM Agents
von: Cartagena, Arnold, et al.
Veröffentlicht: (2026)
von: Cartagena, Arnold, et al.
Veröffentlicht: (2026)
Improving Existing Optimization Algorithms with LLMs
von: Sartori, Camilo Chacón, et al.
Veröffentlicht: (2025)
von: Sartori, Camilo Chacón, et al.
Veröffentlicht: (2025)
Comparative Analysis of LLM Abliteration Methods: A Cross-Architecture Evaluation
von: Young, Richard J.
Veröffentlicht: (2025)
von: Young, Richard J.
Veröffentlicht: (2025)
The Impact of Large Language Models on Open-source Innovation: Evidence from GitHub Copilot
von: Yeverechyahu, Doron, et al.
Veröffentlicht: (2024)
von: Yeverechyahu, Doron, et al.
Veröffentlicht: (2024)
Leveraging Large Language Models for Use Case Model Generation from Software Requirements
von: Eisenreich, Tobias, et al.
Veröffentlicht: (2025)
von: Eisenreich, Tobias, et al.
Veröffentlicht: (2025)
TSCG: Deterministic Tool-Schema Compilation for Agentic LLM Deployments
von: Sakizli, Furkan
Veröffentlicht: (2026)
von: Sakizli, Furkan
Veröffentlicht: (2026)
Benchmarking Energy Efficiency of Large Language Models Using vLLM
von: Pronk, K., et al.
Veröffentlicht: (2025)
von: Pronk, K., et al.
Veröffentlicht: (2025)
Leveraging Test Driven Development with Large Language Models for Reliable and Verifiable Spreadsheet Code Generation: A Research Framework
von: Thorne, Simon, et al.
Veröffentlicht: (2025)
von: Thorne, Simon, et al.
Veröffentlicht: (2025)
MEC$^3$O: Multi-Expert Consensus for Code Time Complexity Prediction
von: Hahn, Joonghyuk, et al.
Veröffentlicht: (2025)
von: Hahn, Joonghyuk, et al.
Veröffentlicht: (2025)
ContractEval: A Benchmark for Evaluating Contract-Satisfying Assertions in Code Generation
von: Lim, Soohan, et al.
Veröffentlicht: (2025)
von: Lim, Soohan, et al.
Veröffentlicht: (2025)
VeriGuard: Enhancing LLM Agent Safety via Verified Code Generation
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
von: Miculicich, Lesly, et al.
Veröffentlicht: (2025)
Source Code Summarization in the Era of Large Language Models
von: Sun, Weisong, et al.
Veröffentlicht: (2024)
von: Sun, Weisong, et al.
Veröffentlicht: (2024)
A Serverless Architecture for Real-Time Stock Analysis using Large Language Models: An Iterative Development and Debugging Case Study
von: Ashraf, Taniv
Veröffentlicht: (2025)
von: Ashraf, Taniv
Veröffentlicht: (2025)
Ähnliche Einträge
-
Failure by Interference: Language Models Make Balanced Parentheses Errors When Faulty Mechanisms Overshadow Sound Ones
von: Rai, Daking, et al.
Veröffentlicht: (2025) -
A Practical Review of Mechanistic Interpretability for Transformer-Based Language Models
von: Rai, Daking, et al.
Veröffentlicht: (2024) -
An Investigation of Neuron Activation as a Unified Lens to Explain Chain-of-Thought Eliciting Arithmetic Reasoning of LLMs
von: Rai, Daking, et al.
Veröffentlicht: (2024) -
Vibe Code Bench: Evaluating AI Models on End-to-End Web Application Development
von: Tran, Hung, et al.
Veröffentlicht: (2026) -
When Retrieval Hurts Code Completion: A Diagnostic Study of Stale Repository Context
von: Weng, Haojun, et al.
Veröffentlicht: (2026)