"Should I Give Up Now?" Investigating LLM Pitfalls in Software Engineering

Jiessie Tie; Bingsheng Yao; Tianshi Li; Hongbo Fang; Syed Ishtiaque Ahmed; Dakuo Wang; Shurui Zhou

arXiv:2411.09916·cs.SE·April 15, 2026

"Should I Give Up Now?" Investigating LLM Pitfalls in Software Engineering

Jiessie Tie, Bingsheng Yao, Tianshi Li, Hongbo Fang, Syed Ishtiaque Ahmed, Dakuo Wang, Shurui Zhou

PDF

TL;DR

This study investigates the challenges and failure modes of using large language models like ChatGPT in software engineering tasks, highlighting user strategies and abandonment factors.

Contribution

It categorizes common LLM failure types in SE workflows and quantifies their impact on user abandonment, informing future AI integration strategies.

Findings

01

Unhelpful responses increase abandonment likelihood by 11 times.

02

Each additional prompt reduces abandonment probability by 17%.

03

Users employ scaffolding, clarification, and debugging to mitigate issues.

Abstract

Software engineers are increasingly incorporating AI assistants into their workflows to enhance productivity and alleviate cognitive load. However, experiences with large language models (LLMs) such as ChatGPT vary widely. While some engineers find them useful, others deem them counterproductive due to inaccuracies in their responses. Researchers have also observed that ChatGPT often provides incorrect information. Given these limitations, it is crucial to determine how to effectively integrate LLMs into software engineering (SE) workflow. Analyzing data from 26 participants in a complex web development task, we identified nine failure types categorized into incorrect or incomplete responses, cognitive overload, and context loss. Users attempted to mitigate these issues through scaffolding, prompt clarification, and debugging. However, 17 participants ultimately chose to abandon ChatGPT…

Peer Reviews

No public reviews on file for this paper yet. If you reviewed it on a platform where reviews are public (OpenReview, ICLR, NeurIPS, ICML), you can paste yours below so the community can read it here.

Videos

No videos yet. Explain this paper in a talk, walkthrough, or lecture? Add one.