Similar Items: Exploring the capabilities of vision-language models to detect visual bugs in HTML5 applications
- Line-level bug-finding power of static analysis rules: a case study of Teamscale
- Debian Experiments with AI-Assisted Bug Triage as Open-Source Projects Face Growing Report Overload
- CDBench: Benchmarking the mutation testing capabilities of LLMs with code defenders
- Large language models in model-driven engineering: a systematic mapping study
- Empirical benchmarking of large language models for data science coding: a multidimensional evaluation
- Evaluating foundation model integration strategies for detecting PII in java software engineering pipelines