Five Fatal Assumptions: Why T-Shirt Sizing Systematically Fails for AI Projects
Fuente:
arXiv
Enregistré dans:
| Auteurs principaux: | Soundaramourty, Raja, Kilic, Ozkan, Chenchaiah, Ramu |
|---|---|
| Format: | Preprint |
| Publié: |
2026
|
| Sujets: | |
| Accès en ligne: | |
| Tags: |
Ajouter un tag
Pas de tags, Soyez le premier à ajouter un tag!
|
Documents similaires
WebSuite: Systematically Evaluating Why Web Agents Fail
par: Li, Eric, et autres
Publié: (2024)
par: Li, Eric, et autres
Publié: (2024)
Why Attention Fails: A Taxonomy of Faults in Attention-Based Neural Networks
par: Jahan, Sigma, et autres
Publié: (2025)
par: Jahan, Sigma, et autres
Publié: (2025)
Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code
par: Kim, Myeongsoo, et autres
Publié: (2026)
par: Kim, Myeongsoo, et autres
Publié: (2026)
Exploring Autonomous Agents: A Closer Look at Why They Fail When Completing Tasks
par: Lu, Ruofan, et autres
Publié: (2025)
par: Lu, Ruofan, et autres
Publié: (2025)
Where Do AI Coding Agents Fail? An Empirical Study of Failed Agentic Pull Requests in GitHub
par: Ehsani, Ramtin, et autres
Publié: (2026)
par: Ehsani, Ramtin, et autres
Publié: (2026)
TriCEGAR: A Trace-Driven Abstraction Mechanism for Agentic AI
par: Koohestani, Roham, et autres
Publié: (2026)
par: Koohestani, Roham, et autres
Publié: (2026)
LLMs for Qualitative Data Analysis Fail on Security-specificComments in Human Experiments
par: Camporese, Maria, et autres
Publié: (2026)
par: Camporese, Maria, et autres
Publié: (2026)
Empowering AI to Generate Better AI Code: Guided Generation of Deep Learning Projects with LLMs
par: Xie, Chen, et autres
Publié: (2025)
par: Xie, Chen, et autres
Publié: (2025)
Why Are AI Agent Involved Pull Requests (Fix-Related) Remain Unmerged? An Empirical Study
par: Alam, Khairul, et autres
Publié: (2026)
par: Alam, Khairul, et autres
Publié: (2026)
The Machine Learning Canvas: Empirical Findings on Why Strategy Matters More Than AI Code Generation
par: Prause, Martin
Publié: (2026)
par: Prause, Martin
Publié: (2026)
Generative AI for Requirements Engineering: A Systematic Literature Review
par: Cheng, Haowei, et autres
Publié: (2024)
par: Cheng, Haowei, et autres
Publié: (2024)
Describing Agentic AI Systems with C4: Lessons from Industry Projects
par: Rausch, Andreas, et autres
Publié: (2026)
par: Rausch, Andreas, et autres
Publié: (2026)
AI Techniques in the Microservices Life-Cycle: A Systematic Mapping Study
par: Moreschini, Sergio, et autres
Publié: (2023)
par: Moreschini, Sergio, et autres
Publié: (2023)
A3Rank: Augmentation Alignment Analysis for Prioritizing Overconfident Failing Samples for Deep Learning Models
par: Wei, Zhengyuan, et autres
Publié: (2024)
par: Wei, Zhengyuan, et autres
Publié: (2024)
Butterfly Effects in Toolchains: A Comprehensive Analysis of Failed Parameter Filling in LLM Tool-Agent Systems
par: Xiong, Qian, et autres
Publié: (2025)
par: Xiong, Qian, et autres
Publié: (2025)
ProjDevBench: Benchmarking AI Coding Agents on End-to-End Project Development
par: Lu, Pengrui, et autres
Publié: (2026)
par: Lu, Pengrui, et autres
Publié: (2026)
Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects
par: Kashif, Syed Mohammad, et autres
Publié: (2026)
par: Kashif, Syed Mohammad, et autres
Publié: (2026)
An Empirical Framework for Evaluating Semantic Preservation Using Hugging Face
par: Jia, Nan, et autres
Publié: (2025)
par: Jia, Nan, et autres
Publié: (2025)
Generative AI for Software Project Management: Insights from a Review of Software Practitioner Literature
par: Assalaarachchi, Lakshana Iruni, et autres
Publié: (2025)
par: Assalaarachchi, Lakshana Iruni, et autres
Publié: (2025)
Bug Analysis Towards Bug Resolution Time Prediction
par: Ozkan, Hasan Yagiz, et autres
Publié: (2024)
par: Ozkan, Hasan Yagiz, et autres
Publié: (2024)
The Future of Generative AI in Software Engineering: A Vision from Industry and Academia in the European GENIUS Project
par: Gröpler, Robin, et autres
Publié: (2025)
par: Gröpler, Robin, et autres
Publié: (2025)
Impact and Implications of Generative AI for Enterprise Architects in Agile Environments: A Systematic Literature Review
par: Kooy, Stefan Julian, et autres
Publié: (2025)
par: Kooy, Stefan Julian, et autres
Publié: (2025)
Why you shouldn't fully trust ChatGPT: A synthesis of this AI tool's error rates across disciplines and the software engineering lifecycle
par: Garousi, Vahid
Publié: (2025)
par: Garousi, Vahid
Publié: (2025)
How Do LLMs Fail In Agentic Scenarios? A Qualitative Analysis of Success and Failure Scenarios of Various LLMs in Agentic Simulations
par: Roig, JV
Publié: (2025)
par: Roig, JV
Publié: (2025)
How to Trick Your AI TA: A Systematic Study of Academic Jailbreaking in LLM Code Evaluation
par: Sahoo, Devanshu, et autres
Publié: (2025)
par: Sahoo, Devanshu, et autres
Publié: (2025)
Speed at the Cost of Quality: How Cursor AI Increases Short-Term Velocity and Long-Term Complexity in Open-Source Projects
par: He, Hao, et autres
Publié: (2025)
par: He, Hao, et autres
Publié: (2025)
LLMs Integration in Software Engineering Team Projects: Roles, Impact, and a Pedagogical Design Space for AI Tools in Computing Education
par: Kharrufa, Ahmed, et autres
Publié: (2024)
par: Kharrufa, Ahmed, et autres
Publié: (2024)
Neuron-Guided Interpretation of Code LLMs: Where, Why, and How?
par: Yin, Zhe, et autres
Publié: (2025)
par: Yin, Zhe, et autres
Publié: (2025)
How Mature is Requirements Engineering for AI-based Systems? A Systematic Mapping Study on Practices, Challenges, and Future Research Directions
par: Habiba, Umm-e-, et autres
Publié: (2024)
par: Habiba, Umm-e-, et autres
Publié: (2024)
From Expectation to Habit: Why Do Software Practitioners Adopt Fairness Toolkits?
par: Voria, Gianmario, et autres
Publié: (2024)
par: Voria, Gianmario, et autres
Publié: (2024)
When the Code Autopilot Breaks: Why LLMs Falter in Embedded Machine Learning
par: Morabito, Roberto, et autres
Publié: (2025)
par: Morabito, Roberto, et autres
Publié: (2025)
RefusalBench: Why Refusal Rate Misranks Frontier LLMs on Biological Research Prompts
par: Weidener, Lukas, et autres
Publié: (2026)
par: Weidener, Lukas, et autres
Publié: (2026)
A Deep Dive Into Large Language Model Code Generation Mistakes: What and Why?
par: Chen, QiHong, et autres
Publié: (2024)
par: Chen, QiHong, et autres
Publié: (2024)
How does information access affect LLM monitors' ability to detect sabotage?
par: Arike, Rauno, et autres
Publié: (2026)
par: Arike, Rauno, et autres
Publié: (2026)
The Tool-Overuse Illusion: Why Does LLM Prefer External Tools over Internal Knowledge?
par: Zeng, Yirong, et autres
Publié: (2026)
par: Zeng, Yirong, et autres
Publié: (2026)
What's documented in AI? Systematic Analysis of 32K AI Model Cards
par: Liang, Weixin, et autres
Publié: (2024)
par: Liang, Weixin, et autres
Publié: (2024)
Projectional Decoding: Towards Semantic-Aware LLM Generation
par: Chen, Boqi, et autres
Publié: (2026)
par: Chen, Boqi, et autres
Publié: (2026)
Unveiling Project-Specific Bias in Neural Code Models
par: Li, Zhiming, et autres
Publié: (2022)
par: Li, Zhiming, et autres
Publié: (2022)
Demystifying Issues, Causes and Solutions in LLM Open-Source Projects
par: Cai, Yangxiao, et autres
Publié: (2024)
par: Cai, Yangxiao, et autres
Publié: (2024)
Automated Bug Report Prioritization in Large Open-Source Projects
par: Pierson, Riley, et autres
Publié: (2025)
par: Pierson, Riley, et autres
Publié: (2025)
Documents similaires
-
WebSuite: Systematically Evaluating Why Web Agents Fail
par: Li, Eric, et autres
Publié: (2024) -
Why Attention Fails: A Taxonomy of Faults in Attention-Based Neural Networks
par: Jahan, Sigma, et autres
Publié: (2025) -
Coherence Collapse: Diagnosing Why Code Agents Fail After Reaching the Right Code
par: Kim, Myeongsoo, et autres
Publié: (2026) -
Exploring Autonomous Agents: A Closer Look at Why They Fail When Completing Tasks
par: Lu, Ruofan, et autres
Publié: (2025) -
Where Do AI Coding Agents Fail? An Empirical Study of Failed Agentic Pull Requests in GitHub
par: Ehsani, Ramtin, et autres
Publié: (2026)