Prompt Stability in Code LLMs: Measuring Sensitivity across Emotion- and Personality-Driven Variations
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Ma, Wei, Yang, Yixiao, Ge, Jingquan, Xie, Xiaofei, Jiang, Lingxiao |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2025
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Beyond Final Code: A Process-Oriented Error Analysis of Software Development Agents in Real-World GitHub Scenarios
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
von: Chen, Zhi, et al.
Veröffentlicht: (2025)
Evaluating Software Development Agents: Patch Patterns, Code Quality, and Issue Complexity in Real-World GitHub Scenarios
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities
von: Ma, Wei, et al.
Veröffentlicht: (2022)
von: Ma, Wei, et al.
Veröffentlicht: (2022)
Promise and Peril of Collaborative Code Generation Models: Balancing Effectiveness and Memorization
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
von: Chen, Zhi, et al.
Veröffentlicht: (2024)
Do Code Semantics Help? A Comprehensive Study on Execution Trace-Based Information for Code Large Language Models
von: Wang, Jian, et al.
Veröffentlicht: (2025)
von: Wang, Jian, et al.
Veröffentlicht: (2025)
Open-Source AI-based SE Tools: Opportunities and Challenges of Collaborative Software Learning
von: Lin, Zhihao, et al.
Veröffentlicht: (2024)
von: Lin, Zhihao, et al.
Veröffentlicht: (2024)
RedCoder: Automated Multi-Turn Red Teaming for Code LLMs
von: Mo, Wenjie Jacky, et al.
Veröffentlicht: (2025)
von: Mo, Wenjie Jacky, et al.
Veröffentlicht: (2025)
Personality-Guided Code Generation Using Large Language Models
von: Guo, Yaoqi, et al.
Veröffentlicht: (2024)
von: Guo, Yaoqi, et al.
Veröffentlicht: (2024)
Transducer Tuning: Efficient Model Adaptation for Software Tasks Using Code Property Graphs
von: Yusuf, Imam Nur Bani, et al.
Veröffentlicht: (2024)
von: Yusuf, Imam Nur Bani, et al.
Veröffentlicht: (2024)
Tests as Prompt: A Test-Driven-Development Benchmark for LLM Code Generation
von: Cui, Yi
Veröffentlicht: (2025)
von: Cui, Yi
Veröffentlicht: (2025)
Same Signal, Different Semantics: A Cross-Framework Behavioral Analysis of Software Engineering Agents
von: Ma, Wei, et al.
Veröffentlicht: (2026)
von: Ma, Wei, et al.
Veröffentlicht: (2026)
Beyond Accuracy: Policy Invariance as a Reliability Test for LLM Safety Judges
von: Weng, Shihao, et al.
Veröffentlicht: (2026)
von: Weng, Shihao, et al.
Veröffentlicht: (2026)
LeDex: Training LLMs to Better Self-Debug and Explain Code
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
von: Jiang, Nan, et al.
Veröffentlicht: (2024)
Effective Code Membership Inference for Code Completion Models via Adversarial Prompts
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
von: Jiang, Yuan, et al.
Veröffentlicht: (2025)
EduBot -- Can LLMs Solve Personalized Learning and Programming Assignments?
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
von: Wang, Yibin, et al.
Veröffentlicht: (2025)
Bias Testing and Mitigation in LLM-based Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
GenCode: A Generic Data Augmentation Framework for Boosting Deep Learning-Based Code Understanding
von: Dong, Zeming, et al.
Veröffentlicht: (2024)
von: Dong, Zeming, et al.
Veröffentlicht: (2024)
Your Instructions Are Not Always Helpful: Assessing the Efficacy of Instruction Fine-tuning for Software Vulnerability Detection
von: Yusuf, Imam Nur Bani, et al.
Veröffentlicht: (2024)
von: Yusuf, Imam Nur Bani, et al.
Veröffentlicht: (2024)
Protocode: Prototype-Driven Interpretability for Code Generation in LLMs
von: Bodla, Krishna Vamshi, et al.
Veröffentlicht: (2025)
von: Bodla, Krishna Vamshi, et al.
Veröffentlicht: (2025)
INTERVENOR: Prompting the Coding Ability of Large Language Models with the Interactive Chain of Repair
von: Wang, Hanbin, et al.
Veröffentlicht: (2023)
von: Wang, Hanbin, et al.
Veröffentlicht: (2023)
Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation
von: Catal, Melih, et al.
Veröffentlicht: (2026)
von: Catal, Melih, et al.
Veröffentlicht: (2026)
Chimera: Harnessing Multi-Agent LLMs for Automatic Insider Threat Simulation
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
von: Yu, Jiongchi, et al.
Veröffentlicht: (2025)
Schedule-and-Calibrate: Utility-Guided Multi-Task Reinforcement Learning for Code LLMs
von: Chen, Yujia, et al.
Veröffentlicht: (2026)
von: Chen, Yujia, et al.
Veröffentlicht: (2026)
Coding in a Bubble? Evaluating LLMs in Resolving Context Adaptation Bugs During Code Adaptation
von: Zhang, Tanghaoran, et al.
Veröffentlicht: (2026)
von: Zhang, Tanghaoran, et al.
Veröffentlicht: (2026)
Prompt Driven Development with Claude Code: Building a Complete TUI Framework for the Ring Programming Language
von: Fayed, Mahmoud Samir, et al.
Veröffentlicht: (2026)
von: Fayed, Mahmoud Samir, et al.
Veröffentlicht: (2026)
ECO: Enhanced Code Optimization via Performance-Aware Prompting for Code-LLMs
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2025)
von: Kim, Su-Hyeon, et al.
Veröffentlicht: (2025)
MORTAR: A Model-based Runtime Action Repair Framework for AI-enabled Cyber-Physical Systems
von: Wang, Renzhi, et al.
Veröffentlicht: (2024)
von: Wang, Renzhi, et al.
Veröffentlicht: (2024)
COAST: Enhancing the Code Debugging Ability of LLMs through Communicative Agent Based Data Synthesis
von: Yang, Weiqing, et al.
Veröffentlicht: (2024)
von: Yang, Weiqing, et al.
Veröffentlicht: (2024)
The Foundation Cracks: A Comprehensive Study on Bugs and Testing Practices in LLM Libraries
von: Jiang, Weipeng, et al.
Veröffentlicht: (2025)
von: Jiang, Weipeng, et al.
Veröffentlicht: (2025)
Rectifier: Code Translation with Corrector via LLMs
von: Yin, Xin, et al.
Veröffentlicht: (2024)
von: Yin, Xin, et al.
Veröffentlicht: (2024)
CodeFort: Robust Training for Code Generation Models
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
von: Zhang, Yuhao, et al.
Veröffentlicht: (2024)
DriveTester: A Unified Platform for Simulation-Based Autonomous Driving Testing
von: Cheng, Mingfei, et al.
Veröffentlicht: (2024)
von: Cheng, Mingfei, et al.
Veröffentlicht: (2024)
From PowerPoint UI Sketches to Web-Based Applications: Pattern-Driven Code Generation for GIS Dashboard Development Using Knowledge-Augmented LLMs, Context-Aware Visual Prompting, and the React Framework
von: Xu, Haowen, et al.
Veröffentlicht: (2025)
von: Xu, Haowen, et al.
Veröffentlicht: (2025)
Empowering AI to Generate Better AI Code: Guided Generation of Deep Learning Projects with LLMs
von: Xie, Chen, et al.
Veröffentlicht: (2025)
von: Xie, Chen, et al.
Veröffentlicht: (2025)
Function-to-Style Guidance of LLMs for Code Translation
von: Zhang, Longhui, et al.
Veröffentlicht: (2025)
von: Zhang, Longhui, et al.
Veröffentlicht: (2025)
Selection of Prompt Engineering Techniques for Code Generation through Predicting Code Complexity
von: Wang, Chung-Yu, et al.
Veröffentlicht: (2024)
von: Wang, Chung-Yu, et al.
Veröffentlicht: (2024)
LLMs as Continuous Learners: Improving the Reproduction of Defective Code in Software Issues
von: Lin, Yalan, et al.
Veröffentlicht: (2024)
von: Lin, Yalan, et al.
Veröffentlicht: (2024)
Automated Customization of LLMs for Enterprise Code Repositories Using Semantic Scopes
von: Finkler, Ulrich, et al.
Veröffentlicht: (2026)
von: Finkler, Ulrich, et al.
Veröffentlicht: (2026)
Benchmarking LLMs for Fine-Grained Code Review with Enriched Context in Practice
von: Hu, Ruida, et al.
Veröffentlicht: (2025)
von: Hu, Ruida, et al.
Veröffentlicht: (2025)
MIMIC-Py: An Extensible Tool for Personality-Driven Automated Game Testing with Large Language Models
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
von: Chen, Yifei, et al.
Veröffentlicht: (2026)
Ähnliche Einträge
-
Beyond Final Code: A Process-Oriented Error Analysis of Software Development Agents in Real-World GitHub Scenarios
von: Chen, Zhi, et al.
Veröffentlicht: (2025) -
Evaluating Software Development Agents: Patch Patterns, Code Quality, and Issue Complexity in Real-World GitHub Scenarios
von: Chen, Zhi, et al.
Veröffentlicht: (2024) -
Unveiling Code Pre-Trained Models: Investigating Syntax and Semantics Capacities
von: Ma, Wei, et al.
Veröffentlicht: (2022) -
Promise and Peril of Collaborative Code Generation Models: Balancing Effectiveness and Memorization
von: Chen, Zhi, et al.
Veröffentlicht: (2024) -
Do Code Semantics Help? A Comprehensive Study on Execution Trace-Based Information for Code Large Language Models
von: Wang, Jian, et al.
Veröffentlicht: (2025)