Using Grammar Masking to Ensure Syntactic Validity in LLM-based Modeling Tasks
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Netz, Lukas, Reimer, Jan, Rumpe, Bernhard |
|---|---|
| Format: | Preprint |
| Publié: |
2024
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks
par: Ganguly, Debargha, et autres
Publié: (2025)
par: Ganguly, Debargha, et autres
Publié: (2025)
Transducer Tuning: Efficient Model Adaptation for Software Tasks Using Code Property Graphs
par: Yusuf, Imam Nur Bani, et autres
Publié: (2024)
par: Yusuf, Imam Nur Bani, et autres
Publié: (2024)
TokDrift: When LLM Speaks in Subwords but Code Speaks in Grammar
par: Li, Yinxi, et autres
Publié: (2025)
par: Li, Yinxi, et autres
Publié: (2025)
CodeJudgeBench: Benchmarking LLM-as-a-Judge for Coding Tasks
par: Jiang, Hongchao, et autres
Publié: (2025)
par: Jiang, Hongchao, et autres
Publié: (2025)
An LLM-as-Judge Metric for Bridging the Gap with Human Evaluation in SE Tasks
par: Zhou, Xin, et autres
Publié: (2025)
par: Zhou, Xin, et autres
Publié: (2025)
Mechanistic Understanding of Language Models in Syntactic Code Completion
par: Miller, Samuel, et autres
Publié: (2025)
par: Miller, Samuel, et autres
Publié: (2025)
Advancing Language Models for Code-related Tasks
par: Tian, Zhao
Publié: (2026)
par: Tian, Zhao
Publié: (2026)
ChainStream: An LLM-based Framework for Unified Synthetic Sensing
par: Liu, Jiacheng, et autres
Publié: (2024)
par: Liu, Jiacheng, et autres
Publié: (2024)
Text2BIM: Generating Building Models Using a Large Language Model-based Multi-Agent Framework
par: Du, Changyu, et autres
Publié: (2024)
par: Du, Changyu, et autres
Publié: (2024)
E2Edev: Benchmarking Large Language Models in End-to-End Software Development Task
par: Liu, Jingyao, et autres
Publié: (2025)
par: Liu, Jingyao, et autres
Publié: (2025)
BenchBrowser: Retrieving Evidence for Evaluating Benchmark Validity
par: Diddee, Harshita, et autres
Publié: (2026)
par: Diddee, Harshita, et autres
Publié: (2026)
WorkflowLLM: Enhancing Workflow Orchestration Capability of Large Language Models
par: Fan, Shengda, et autres
Publié: (2024)
par: Fan, Shengda, et autres
Publié: (2024)
From LLMs to LLM-based Agents for Software Engineering: A Survey of Current, Challenges and Future
par: Jin, Haolin, et autres
Publié: (2024)
par: Jin, Haolin, et autres
Publié: (2024)
SoAy: A Solution-based LLM API-using Methodology for Academic Information Seeking
par: Wang, Yuanchun, et autres
Publié: (2024)
par: Wang, Yuanchun, et autres
Publié: (2024)
SWE-MERA: A Dynamic Benchmark for Agenticly Evaluating Large Language Models on Software Engineering Tasks
par: Adamenko, Pavel, et autres
Publié: (2025)
par: Adamenko, Pavel, et autres
Publié: (2025)
GoNoGo: An Efficient LLM-based Multi-Agent System for Streamlining Automotive Software Release Decision-Making
par: Khoee, Arsham Gholamzadeh, et autres
Publié: (2024)
par: Khoee, Arsham Gholamzadeh, et autres
Publié: (2024)
CodeR: Issue Resolving with Multi-Agent and Task Graphs
par: Chen, Dong, et autres
Publié: (2024)
par: Chen, Dong, et autres
Publié: (2024)
Toward Reusability of AI Models Using Dynamic Updates of AI Documentation
par: Bajcsy, Peter, et autres
Publié: (2026)
par: Bajcsy, Peter, et autres
Publié: (2026)
Verification Limits Code LLM Training
par: Gureja, Srishti, et autres
Publié: (2025)
par: Gureja, Srishti, et autres
Publié: (2025)
SUPER: Evaluating Agents on Setting Up and Executing Tasks from Research Repositories
par: Bogin, Ben, et autres
Publié: (2024)
par: Bogin, Ben, et autres
Publié: (2024)
MERA Code: A Unified Framework for Evaluating Code Generation Across Tasks
par: Chervyakov, Artem, et autres
Publié: (2025)
par: Chervyakov, Artem, et autres
Publié: (2025)
A Reference Architecture for Designing Foundation Model based Systems
par: Lu, Qinghua, et autres
Publié: (2023)
par: Lu, Qinghua, et autres
Publié: (2023)
Demo-Craft: Using In-Context Learning to Improve Code Generation in Large Language Models
par: Kapu, Nirmal Joshua, et autres
Publié: (2024)
par: Kapu, Nirmal Joshua, et autres
Publié: (2024)
Crystal: Illuminating LLM Abilities on Language and Code
par: Tao, Tianhua, et autres
Publié: (2024)
par: Tao, Tianhua, et autres
Publié: (2024)
Pragmatic Reasoning improves LLM Code Generation
par: Cao, Zhuchen, et autres
Publié: (2025)
par: Cao, Zhuchen, et autres
Publié: (2025)
A Taxonomy of Foundation Model based Systems through the Lens of Software Architecture
par: Lu, Qinghua, et autres
Publié: (2023)
par: Lu, Qinghua, et autres
Publié: (2023)
SlopCodeBench: Benchmarking How Coding Agents Degrade Over Long-Horizon Iterative Tasks
par: Orlanski, Gabriel, et autres
Publié: (2026)
par: Orlanski, Gabriel, et autres
Publié: (2026)
EHR-Based Mobile and Web Platform for Chronic Disease Risk Prediction Using Large Language Multimodal Models
par: Liao, Chun-Chieh, et autres
Publié: (2024)
par: Liao, Chun-Chieh, et autres
Publié: (2024)
An evaluation of LLM code generation capabilities through graded exercises
par: Jiménez, Álvaro Barbero
Publié: (2024)
par: Jiménez, Álvaro Barbero
Publié: (2024)
LLM Hallucinations in Practical Code Generation: Phenomena, Mechanism, and Mitigation
par: Zhang, Ziyao, et autres
Publié: (2024)
par: Zhang, Ziyao, et autres
Publié: (2024)
Learning to Ask: When LLM Agents Meet Unclear Instruction
par: Wang, Wenxuan, et autres
Publié: (2024)
par: Wang, Wenxuan, et autres
Publié: (2024)
Planning to Explore: Curiosity-Driven Planning for LLM Test Generation
par: Amayuelas, Alfonso, et autres
Publié: (2026)
par: Amayuelas, Alfonso, et autres
Publié: (2026)
LocAgent: Graph-Guided LLM Agents for Code Localization
par: Chen, Zhaoling, et autres
Publié: (2025)
par: Chen, Zhaoling, et autres
Publié: (2025)
Specifications: The missing link to making the development of LLM systems an engineering discipline
par: Stoica, Ion, et autres
Publié: (2024)
par: Stoica, Ion, et autres
Publié: (2024)
How Toxic Can You Get? Search-based Toxicity Testing for Large Language Models
par: Corbo, Simone, et autres
Publié: (2025)
par: Corbo, Simone, et autres
Publié: (2025)
Automated Business Process Analysis: An LLM-Based Approach to Value Assessment
par: De Michele, William, et autres
Publié: (2025)
par: De Michele, William, et autres
Publié: (2025)
Collaboration is all you need: LLM Assisted Safe Code Translation
par: Karanjai, Rabimba, et autres
Publié: (2025)
par: Karanjai, Rabimba, et autres
Publié: (2025)
Issue Localization via LLM-Driven Iterative Code Graph Searching
par: Jiang, Zhonghao, et autres
Publié: (2025)
par: Jiang, Zhonghao, et autres
Publié: (2025)
RGD: Multi-LLM Based Agent Debugger via Refinement and Generation Guidance
par: Jin, Haolin, et autres
Publié: (2024)
par: Jin, Haolin, et autres
Publié: (2024)
Adaptable and Precise: Enterprise-Scenario LLM Function-Calling Capability Training Pipeline
par: Zeng, Guancheng, et autres
Publié: (2024)
par: Zeng, Guancheng, et autres
Publié: (2024)
Documents similaires
-
Grammars of Formal Uncertainty: When to Trust LLMs in Automated Reasoning Tasks
par: Ganguly, Debargha, et autres
Publié: (2025) -
Transducer Tuning: Efficient Model Adaptation for Software Tasks Using Code Property Graphs
par: Yusuf, Imam Nur Bani, et autres
Publié: (2024) -
TokDrift: When LLM Speaks in Subwords but Code Speaks in Grammar
par: Li, Yinxi, et autres
Publié: (2025) -
CodeJudgeBench: Benchmarking LLM-as-a-Judge for Coding Tasks
par: Jiang, Hongchao, et autres
Publié: (2025) -
An LLM-as-Judge Metric for Bridging the Gap with Human Evaluation in SE Tasks
par: Zhou, Xin, et autres
Publié: (2025)