FASTRIC: Prompt Specification Language for Verifiable LLM Interactions
Fuente:
arXiv
Salvato in:
| Autore principale: | Jin, Wen-Long |
|---|---|
| Natura: | Preprint |
| Pubblicazione: |
2025
|
| Soggetti: | |
| Accesso online: | |
| Tags: |
Aggiungi Tag
Nessun Tag, puoi essere il primo ad aggiungerne!!
|
Documenti analoghi
What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts
di: Yang, Chenyang, et al.
Pubblicazione: (2025)
di: Yang, Chenyang, et al.
Pubblicazione: (2025)
PACE: Improving Prompt with Actor-Critic Editing for Large Language Model
di: Dong, Yihong, et al.
Pubblicazione: (2023)
di: Dong, Yihong, et al.
Pubblicazione: (2023)
Talk Less, Verify More: Improving LLM Assistants with Semantic Checks and Execution Feedback
di: Sun, Yan, et al.
Pubblicazione: (2026)
di: Sun, Yan, et al.
Pubblicazione: (2026)
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast
di: Fan, Wen, et al.
Pubblicazione: (2024)
di: Fan, Wen, et al.
Pubblicazione: (2024)
Show and Tell: Prompt Strategies for Style Control in Multi-Turn LLM Code Generation
di: Bohr, Jeremiah
Pubblicazione: (2025)
di: Bohr, Jeremiah
Pubblicazione: (2025)
(Why) Is My Prompt Getting Worse? Rethinking Regression Testing for Evolving LLM APIs
di: Ma, Wanqin, et al.
Pubblicazione: (2023)
di: Ma, Wanqin, et al.
Pubblicazione: (2023)
Training Software Engineering Agents and Verifiers with SWE-Gym
di: Pan, Jiayi, et al.
Pubblicazione: (2024)
di: Pan, Jiayi, et al.
Pubblicazione: (2024)
Interpretable Online Log Analysis Using Large Language Models with Prompt Strategies
di: Liu, Yilun, et al.
Pubblicazione: (2023)
di: Liu, Yilun, et al.
Pubblicazione: (2023)
ArtifactsBench: Bridging the Visual-Interactive Gap in LLM Code Generation Evaluation
di: Zhang, Chenchen, et al.
Pubblicazione: (2025)
di: Zhang, Chenchen, et al.
Pubblicazione: (2025)
SwiftEval: Developing a Language-Specific Benchmark for LLM-generated Code Evaluation
di: Petrukha, Ivan, et al.
Pubblicazione: (2025)
di: Petrukha, Ivan, et al.
Pubblicazione: (2025)
EvoCodeBench: An Evolving Code Generation Benchmark with Domain-Specific Evaluations
di: Li, Jia, et al.
Pubblicazione: (2024)
di: Li, Jia, et al.
Pubblicazione: (2024)
Prompting Large Language Models to Tackle the Full Software Development Lifecycle: A Case Study
di: Li, Bowen, et al.
Pubblicazione: (2024)
di: Li, Bowen, et al.
Pubblicazione: (2024)
The Prompt Alchemist: Automated LLM-Tailored Prompt Optimization for Test Case Generation
di: Gao, Shuzheng, et al.
Pubblicazione: (2025)
di: Gao, Shuzheng, et al.
Pubblicazione: (2025)
Generating and Evaluating Sustainable Procurement Criteria for the Swiss Public Sector using In-Context Prompting with Large Language Models
di: Gao, Yingqiang, et al.
Pubblicazione: (2026)
di: Gao, Yingqiang, et al.
Pubblicazione: (2026)
Firefly: Illuminating Large-Scale Verified Tool-Call Data Generation from Real APIs
di: Lu, Yuxuan, et al.
Pubblicazione: (2026)
di: Lu, Yuxuan, et al.
Pubblicazione: (2026)
Showing LLM-Generated Code Selectively Based on Confidence of LLMs
di: Li, Jia, et al.
Pubblicazione: (2024)
di: Li, Jia, et al.
Pubblicazione: (2024)
A Taxonomy of Prompt Defects in LLM Systems
di: Tian, Haoye, et al.
Pubblicazione: (2025)
di: Tian, Haoye, et al.
Pubblicazione: (2025)
Sphinx: Benchmarking and Modeling for LLM-Driven Pull Request Review
di: Zhang, Daoan, et al.
Pubblicazione: (2026)
di: Zhang, Daoan, et al.
Pubblicazione: (2026)
Suggesting Code Edits in Interactive Machine Learning Notebooks Using Large Language Models
di: Jin, Bihui, et al.
Pubblicazione: (2025)
di: Jin, Bihui, et al.
Pubblicazione: (2025)
Specifications: The missing link to making the development of LLM systems an engineering discipline
di: Stoica, Ion, et al.
Pubblicazione: (2024)
di: Stoica, Ion, et al.
Pubblicazione: (2024)
AlgoVeri: An Aligned Benchmark for Verified Code Generation on Classical Algorithms
di: Zhao, Haoyu, et al.
Pubblicazione: (2026)
di: Zhao, Haoyu, et al.
Pubblicazione: (2026)
LLM-as-a-Judge for Reference-less Automatic Code Validation and Refinement for Natural Language to Bash in IT Automation
di: Vo, Ngoc Phuoc An, et al.
Pubblicazione: (2025)
di: Vo, Ngoc Phuoc An, et al.
Pubblicazione: (2025)
Evaluate-and-Purify: Fortifying Code Language Models Against Adversarial Attacks Using LLM-as-a-Judge
di: Mu, Wenhan, et al.
Pubblicazione: (2025)
di: Mu, Wenhan, et al.
Pubblicazione: (2025)
Enhancing Project-Specific Code Completion by Inferring Internal API Information
di: Deng, Le, et al.
Pubblicazione: (2025)
di: Deng, Le, et al.
Pubblicazione: (2025)
CodeSpecBench: Benchmarking LLMs for Executable Behavioral Specification Generation
di: Chen, Zaoyu, et al.
Pubblicazione: (2026)
di: Chen, Zaoyu, et al.
Pubblicazione: (2026)
Skill over Scale: The Case for Medium, Domain-Specific Models for SE
di: Mukherjee, Manisha, et al.
Pubblicazione: (2023)
di: Mukherjee, Manisha, et al.
Pubblicazione: (2023)
From Prediction to Application: Language Model-based Code Knowledge Tracing with Domain Adaptive Pre-Training and Automatic Feedback System with Pedagogical Prompting for Comprehensive Programming Education
di: Lee, Unggi, et al.
Pubblicazione: (2024)
di: Lee, Unggi, et al.
Pubblicazione: (2024)
ProbeLLM: Automating Principled Diagnosis of LLM Failures
di: Huang, Yue, et al.
Pubblicazione: (2026)
di: Huang, Yue, et al.
Pubblicazione: (2026)
Evaluating LLM-Based Goal Extraction in Requirements Engineering: Prompting Strategies and Their Limitations
di: Arnaudo, Anna, et al.
Pubblicazione: (2026)
di: Arnaudo, Anna, et al.
Pubblicazione: (2026)
From Solitary Directives to Interactive Encouragement! LLM Secure Code Generation by Natural Language Prompting
di: Liu, Shigang, et al.
Pubblicazione: (2024)
di: Liu, Shigang, et al.
Pubblicazione: (2024)
SERA: Soft-Verified Efficient Repository Agents
di: Shen, Ethan, et al.
Pubblicazione: (2026)
di: Shen, Ethan, et al.
Pubblicazione: (2026)
On the Potential and Limitations of Few-Shot In-Context Learning to Generate Metamorphic Specifications for Tax Preparation Software
di: Srinivas, Dananjay, et al.
Pubblicazione: (2023)
di: Srinivas, Dananjay, et al.
Pubblicazione: (2023)
ConCodeEval: Evaluating Large Language Models for Code Constraints in Domain-Specific Languages
di: Kammakomati, Mehant, et al.
Pubblicazione: (2024)
di: Kammakomati, Mehant, et al.
Pubblicazione: (2024)
Automating Computational Reproducibility in Social Science: Comparing Prompt-Based and Agent-Based Approaches
di: Shah, Syed Mehtab Hussain, et al.
Pubblicazione: (2026)
di: Shah, Syed Mehtab Hussain, et al.
Pubblicazione: (2026)
Debug like a Human: A Large Language Model Debugger via Verifying Runtime Execution Step-by-step
di: Zhong, Li, et al.
Pubblicazione: (2024)
di: Zhong, Li, et al.
Pubblicazione: (2024)
RCAgent: Cloud Root Cause Analysis by Autonomous Agents with Tool-Augmented Large Language Models
di: Wang, Zefan, et al.
Pubblicazione: (2023)
di: Wang, Zefan, et al.
Pubblicazione: (2023)
Comparing Developer and LLM Biases in Code Evaluation
di: Mittal, Aditya, et al.
Pubblicazione: (2026)
di: Mittal, Aditya, et al.
Pubblicazione: (2026)
Evaluating and Achieving Controllable Code Completion in Code LLM
di: Zhang, Jiajun, et al.
Pubblicazione: (2026)
di: Zhang, Jiajun, et al.
Pubblicazione: (2026)
Code Fingerprints: Disentangled Attribution of LLM-Generated Code
di: Guo, Jiaxun, et al.
Pubblicazione: (2026)
di: Guo, Jiaxun, et al.
Pubblicazione: (2026)
Multi-Programming Language Sandbox for LLMs
di: Dou, Shihan, et al.
Pubblicazione: (2024)
di: Dou, Shihan, et al.
Pubblicazione: (2024)
Documenti analoghi
-
What Prompts Don't Say: Understanding and Managing Underspecification in LLM Prompts
di: Yang, Chenyang, et al.
Pubblicazione: (2025) -
PACE: Improving Prompt with Actor-Critic Editing for Large Language Model
di: Dong, Yihong, et al.
Pubblicazione: (2023) -
Talk Less, Verify More: Improving LLM Assistants with Semantic Checks and Execution Feedback
di: Sun, Yan, et al.
Pubblicazione: (2026) -
Evaluating the Ability of Large Language Models to Generate Verifiable Specifications in VeriFast
di: Fan, Wen, et al.
Pubblicazione: (2024) -
Show and Tell: Prompt Strategies for Style Control in Multi-Turn LLM Code Generation
di: Bohr, Jeremiah
Pubblicazione: (2025)