Evaluating the Formal Reasoning Capabilities of Large Language Models through Chomsky Hierarchy
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Dong, Yihong, Xiao, Jianha, Jiang, Xue, Guo, Xuyuan, Fan, Zhiyuan, Qian, Jiaru, Zhang, Kechi, Li, Jia, Jin, Zhi, Li, Ge |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
A Survey on Code Generation with LLM-based Agents
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
Focused-DPO: Enhancing Code Generation Through Focused Preference Optimization on Error-Prone Points
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
Saber: An Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
StackTrans: From Large Language Model to Large Pushdown Automata Model
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
A Viable Paradigm of Software Automation: Iterative End-to-End Automated Software Development
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
Computational Thinking Reasoning in Large Language Models
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
Exploring Data-Efficient Adaptation of Large Language Models for Code Generation
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
CodeAgent: Enhancing Code Generation with Tool-Integrated Agent Systems for Real-World Repo-level Coding Challenges
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
Self-collaboration Code Generation via ChatGPT
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
KOCO-BENCH: Can Large Language Models Leverage Domain Knowledge in Software Development?
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
CodeScore: Evaluating Code Generation by Learning Code Execution
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
HiRoPE: Length Extrapolation for Code Models Using Hierarchical Position
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
PACE: Improving Prompt with Actor-Critic Editing for Large Language Model
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
von: Dong, Yihong, et al.
Veröffentlicht: (2023)
LONGCODEU: Benchmarking Long-Context Language Models on Long Code Understanding
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
CodeDPO: Aligning Code Models with Self Generated and Verified Source Code
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
von: Zhang, Kechi, et al.
Veröffentlicht: (2024)
aiXcoder-7B-v2: Training LLMs to Fully Utilize the Long Context in Repository-level Code Completion
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
Uncertainty-Guided Chain-of-Thought for Code Generation with LLMs
von: Zhu, Yuqi, et al.
Veröffentlicht: (2025)
von: Zhu, Yuqi, et al.
Veröffentlicht: (2025)
VulInstruct: Teaching LLMs Root-Cause Reasoning for Vulnerability Detection via Security Specifications
von: Zhu, Hao, et al.
Veröffentlicht: (2025)
von: Zhu, Hao, et al.
Veröffentlicht: (2025)
ROCODE: Integrating Backtracking Mechanism and Program Analysis in Large Language Models for Code Generation
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
von: Jiang, Xue, et al.
Veröffentlicht: (2024)
SEAlign: Alignment Training for Software Engineering Agent
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
von: Zhang, Kechi, et al.
Veröffentlicht: (2025)
Line-level Semantic Structure Learning for Code Vulnerability Detection
von: Wang, Ziliang, et al.
Veröffentlicht: (2024)
von: Wang, Ziliang, et al.
Veröffentlicht: (2024)
From I/O to Code with Discovery Agent
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
von: Dong, Yihong, et al.
Veröffentlicht: (2026)
EvoCodeBench: An Evolving Code Generation Benchmark Aligned with Real-World Code Repositories
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
RealBench: A Repo-Level Code Generation Benchmark Aligned with Real-World Software Development Practices
von: Li, Jia, et al.
Veröffentlicht: (2026)
von: Li, Jia, et al.
Veröffentlicht: (2026)
Self-planning Code Generation with Large Language Models
von: Jiang, Xue, et al.
Veröffentlicht: (2023)
von: Jiang, Xue, et al.
Veröffentlicht: (2023)
GraphCodeAgent: Dual Graph-Guided LLM Agent for Retrieval-Augmented Repo-Level Code Generation
von: Li, Jia, et al.
Veröffentlicht: (2025)
von: Li, Jia, et al.
Veröffentlicht: (2025)
Generalization or Memorization: Data Contamination and Trustworthy Evaluation for Large Language Models
von: Dong, Yihong, et al.
Veröffentlicht: (2024)
von: Dong, Yihong, et al.
Veröffentlicht: (2024)
Rethinking Repetition Problems of LLMs in Code Generation
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
von: Dong, Yihong, et al.
Veröffentlicht: (2025)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
von: Li, Jia, et al.
Veröffentlicht: (2024)
von: Li, Jia, et al.
Veröffentlicht: (2024)
Large Language Model Unlearning for Source Code
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
von: Jiang, Xue, et al.
Veröffentlicht: (2025)
IntentCoding: Amplifying User Intent in Code Generation
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
von: Fang, Zheng, et al.
Veröffentlicht: (2026)
M2CVD: Enhancing Vulnerability Semantic through Multi-Model Collaboration for Code Vulnerability Detection
von: Wang, Ziliang, et al.
Veröffentlicht: (2024)
von: Wang, Ziliang, et al.
Veröffentlicht: (2024)
AdapTrack: Constrained Decoding without Distorting LLM's Output Intent
von: Li, Yongmin, et al.
Veröffentlicht: (2025)
von: Li, Yongmin, et al.
Veröffentlicht: (2025)
aiXcoder-7B: A Lightweight and Effective Large Language Model for Code Processing
von: Jiang, Siyuan, et al.
Veröffentlicht: (2024)
von: Jiang, Siyuan, et al.
Veröffentlicht: (2024)
Enhancing the Capabilities of Large Language Models for API calls through Knowledge Graphs
von: Yang, Ye, et al.
Veröffentlicht: (2025)
von: Yang, Ye, et al.
Veröffentlicht: (2025)
Sifting through the Chaff: On Utilizing Execution Feedback for Ranking the Generated Code Candidates
von: Sun, Zhihong, et al.
Veröffentlicht: (2024)
von: Sun, Zhihong, et al.
Veröffentlicht: (2024)
VulAgent: Hypothesis-Validation based Multi-Agent Vulnerability Detection
von: Wang, Ziliang, et al.
Veröffentlicht: (2025)
von: Wang, Ziliang, et al.
Veröffentlicht: (2025)
Ensembling Large Language Models for Code Vulnerability Detection: An Empirical Evaluation
von: Sun, Zhihong, et al.
Veröffentlicht: (2025)
von: Sun, Zhihong, et al.
Veröffentlicht: (2025)
Think Anywhere in Code Generation
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
von: Jiang, Xue, et al.
Veröffentlicht: (2026)
Evaluating Large Language Models for Time Series Anomaly Detection in Aerospace Software
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
A Survey on Code Generation with LLM-based Agents
von: Dong, Yihong, et al.
Veröffentlicht: (2025) -
Focused-DPO: Enhancing Code Generation Through Focused Preference Optimization on Error-Prone Points
von: Zhang, Kechi, et al.
Veröffentlicht: (2025) -
Saber: An Efficient Sampling with Adaptive Acceleration and Backtracking Enhanced Remasking for Diffusion Language Model
von: Dong, Yihong, et al.
Veröffentlicht: (2025) -
StackTrans: From Large Language Model to Large Pushdown Automata Model
von: Zhang, Kechi, et al.
Veröffentlicht: (2025) -
A Viable Paradigm of Software Automation: Iterative End-to-End Automated Software Development
von: Li, Jia, et al.
Veröffentlicht: (2025)