Empirical Analysis and Detection of Hallucinations in LLM-Generated Bug Report Summaries
arXiv:2605.24137v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used to generate summaries of software bug reports, including sections such as Steps-to-Reproduce (S2R),