Structural Anchors and Reasoning Fragility:Understanding CoT Robustness in LLM4Code
Fuente:
arXiv
Gespeichert in:
| Hauptverfasser: | Liu, Yang, Song, Da, Foundjem, Armstrong, Li, Heng, Khomh, Foutse |
|---|---|
| Format: | Preprint |
| Veröffentlicht: |
2026
|
| Schlagworte: | |
| Online-Zugang: | |
| Tags: |
Tag hinzufügen
Keine Tags, Fügen Sie den ersten Tag hinzu!
|
Ähnliche Einträge
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
von: Liu, Yang, et al.
Veröffentlicht: (2025)
von: Liu, Yang, et al.
Veröffentlicht: (2025)
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
von: Liu, Yang, et al.
Veröffentlicht: (2026)
von: Liu, Yang, et al.
Veröffentlicht: (2026)
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
von: Taraghi, Mina, et al.
Veröffentlicht: (2024)
von: Taraghi, Mina, et al.
Veröffentlicht: (2024)
A Taxonomy of Inefficiencies in LLM-Generated Python Code
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2025)
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2025)
Understanding Web Application Workloads and Their Applications: Systematic Literature Review and Characterization
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
An empirical study of testing machine learning in the wild
von: Openja, Moses, et al.
Veröffentlicht: (2023)
von: Openja, Moses, et al.
Veröffentlicht: (2023)
Empirical Characterization of Logging Smells in Machine Learning Code
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2026)
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2026)
Empirical Characterization of Logging Smells in Machine Learning Code
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2026)
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2026)
Protecting Privacy in Software Logs: What Should Be Anonymized?
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)
On the Effectiveness of Log Representation for Log-based Anomaly Detection
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
An Empirical Study of Policy-as-Code Adoption in Open-Source Software Projects
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2026)
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2026)
Trained Without My Consent: Detecting Code Inclusion In Language Models Trained on Code
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
Exploring Security Practices in Infrastructure as Code: An Empirical Study
von: Verdet, Alexandre, et al.
Veröffentlicht: (2023)
von: Verdet, Alexandre, et al.
Veröffentlicht: (2023)
Logging Requirement for Continuous Auditing of Responsible Machine Learning-based Applications
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2025)
von: Foalem, Patrick Loic, et al.
Veröffentlicht: (2025)
Tracing Optimization for Performance Modeling and Regression Detection
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2024)
What Information Contributes to Log-based Anomaly Detection? Insights from a Configurable Transformer-Based Approach
von: Wu, Xingfang, et al.
Veröffentlicht: (2024)
von: Wu, Xingfang, et al.
Veröffentlicht: (2024)
Towards Understanding the Impact of Data Bugs on Deep Learning Models in Software Engineering
von: Shah, Mehil B, et al.
Veröffentlicht: (2024)
von: Shah, Mehil B, et al.
Veröffentlicht: (2024)
ReCatcher: Towards LLMs Regression Testing for Code Generation
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2025)
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2025)
Are They All Good? Evaluating the Quality of CoTs in LLM-based Code Generation
von: Zhang, Binquan, et al.
Veröffentlicht: (2025)
von: Zhang, Binquan, et al.
Veröffentlicht: (2025)
Machine Learning Robustness: A Primer
von: Braiek, Houssem Ben, et al.
Veröffentlicht: (2024)
von: Braiek, Houssem Ben, et al.
Veröffentlicht: (2024)
LLMs and Stack Overflow Discussions: Reliability, Impact, and Challenges
von: Da Silva, Leuson, et al.
Veröffentlicht: (2024)
von: Da Silva, Leuson, et al.
Veröffentlicht: (2024)
RefAgent: A Multi-agent LLM-based Framework for Automatic Software Refactoring
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2025)
von: Oueslati, Khouloud, et al.
Veröffentlicht: (2025)
An Efficient Model Maintenance Approach for MLOps
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
von: Majidi, Forough, et al.
Veröffentlicht: (2024)
SDLog: A Deep Learning Framework for Detecting Sensitive Information in Software Logs
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2025)
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2025)
Mitigating False Positives in Static Memory Safety Analysis of Rust Programs via Reinforcement Learning
von: P, Akilesh, et al.
Veröffentlicht: (2026)
von: P, Akilesh, et al.
Veröffentlicht: (2026)
Performance Smells in ML and Non-ML Python Projects: A Comparative Study
von: Belias, François, et al.
Veröffentlicht: (2025)
von: Belias, François, et al.
Veröffentlicht: (2025)
From Technical Excellence to Practical Adoption: Lessons Learned Building an ML-Enhanced Trace Analysis Tool
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
QMon: Monitoring the Execution of Quantum Circuits with Mid-Circuit Measurement and Reset
von: Ma, Ning, et al.
Veröffentlicht: (2025)
von: Ma, Ning, et al.
Veröffentlicht: (2025)
DeepCodeProbe: Towards Understanding What Models Trained on Code Learn
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2024)
PathOCL: Path-Based Prompt Augmentation for OCL Generation with GPT-4
von: Abukhalaf, Seif, et al.
Veröffentlicht: (2024)
von: Abukhalaf, Seif, et al.
Veröffentlicht: (2024)
Evaluating and Enhancing Segmentation Model Robustness with Metamorphic Testing
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
von: Mzoughi, Seif, et al.
Veröffentlicht: (2025)
An Empirical Study on Method-Level Performance Evolution in Open-Source Java Projects
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
von: Shahedi, Kaveh, et al.
Veröffentlicht: (2025)
On Fixing Insecure AI-Generated Code through Model Fine-Tuning and Prompting Strategies
von: Jahromi, Ali Soltanian Fard, et al.
Veröffentlicht: (2026)
von: Jahromi, Ali Soltanian Fard, et al.
Veröffentlicht: (2026)
Prism: Dynamic and Flexible Benchmarking of LLMs Code Generation with Monte Carlo Tree Search
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
von: Majdinasab, Vahid, et al.
Veröffentlicht: (2025)
Refining GPT-3 Embeddings with a Siamese Structure for Technical Post Duplicate Detection
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
von: Wu, Xingfang, et al.
Veröffentlicht: (2023)
Real Faults in Model Context Protocol (MCP) Software: a Comprehensive Taxonomy
von: Taraghi, Mina, et al.
Veröffentlicht: (2026)
von: Taraghi, Mina, et al.
Veröffentlicht: (2026)
CodeCoT: Tackling Code Syntax Errors in CoT Reasoning for Code Generation
von: Huang, Dong, et al.
Veröffentlicht: (2023)
von: Huang, Dong, et al.
Veröffentlicht: (2023)
BloomAPR: A Bloom's Taxonomy-based Framework for Assessing the Capabilities of LLM-Powered APR Solutions
von: Ma, Yinghang, et al.
Veröffentlicht: (2025)
von: Ma, Yinghang, et al.
Veröffentlicht: (2025)
Continuously Learning Bug Locations
von: Mindom, Paulina Stevia Nouwou, et al.
Veröffentlicht: (2024)
von: Mindom, Paulina Stevia Nouwou, et al.
Veröffentlicht: (2024)
AdaptiveLLM: A Framework for Selecting Optimal Cost-Efficient LLM for Code-Generation Based on CoT Length
von: Cheng, Junhang, et al.
Veröffentlicht: (2025)
von: Cheng, Junhang, et al.
Veröffentlicht: (2025)
Ähnliche Einträge
-
Adversarial Attack Classification and Robustness Testing for Large Language Models for Code
von: Liu, Yang, et al.
Veröffentlicht: (2025) -
Improving the Robustness of Large Language Models for Code Tasks via Fine-tuning with Perturbed Data
von: Liu, Yang, et al.
Veröffentlicht: (2026) -
Deep Learning Model Reuse in the HuggingFace Community: Challenges, Benefit and Trends
von: Taraghi, Mina, et al.
Veröffentlicht: (2024) -
A Taxonomy of Inefficiencies in LLM-Generated Python Code
von: Abbassi, Altaf Allah, et al.
Veröffentlicht: (2025) -
Understanding Web Application Workloads and Their Applications: Systematic Literature Review and Characterization
von: Aghili, Roozbeh, et al.
Veröffentlicht: (2024)