βœ‰οΈ editor@imrjr.com
International Multidisciplinary Research Journal Reviews (IMRJR)
International Multidisciplinary Research Journal Reviews (IMRJR) A monthly Peer-reviewed journal
ISSN Online 3108-026XISSN Print 3139-4833
← Back to VOLUME 3, ISSUE 3, MARCH 2026

Quantified Explainability and Robustness Analysis of Transformer-Based Bug Detection Models

Debargha Ghosh, Eve Thullen Ph.D, Emmanuel Udoh Ph.D

πŸ‘ 8 viewsπŸ“₯ 0 downloads
Share: 𝕏 f in ✈ βœ‰
Abstract: This paper investigates how to build, systematically configure, and rigorously explains a transformer-based bug detection system. The central argument is that trustworthy explainability requires first establishing which model is worth explaining and in which configuration. We evaluate three approaches on the lrhammond/buggy-apps dataset (8,778 balanced samples): a TF-IDF baseline, DistilBERT, and GraphCodeBERT. DistilBERT exhibited mode collapse across all tested learning rates, confirming that general-purpose language model pretraining is insufficient for code defect detection β€” a prerequisite finding that motivates the choice of GraphCodeBERT for explainability analysis. A systematic ablation across three stride configurations (128, 256, 384 tokens) and three aggregation strategies (max, mean, majority vote) yields nine experimental conditions; stride 256 with mean aggregation is the optimal configuration (70.77% accuracy, 69.56% macro F1, p = 5.26 Γ— 10⁻³³ vs TF-IDF, McNemar's test). Explainability analysis via attention rollout and integrated gradients across ten test samples reveals that integrated gradient signal strength is 11.8Γ— stronger for correctly detected bugs than for misclassified samples β€” providing a gradient-based, quantitative explanation of model failure modes. Attention rollout and integrated gradients show near-zero cross-method correlation (mean r = βˆ’0.017), empirically confirming they are non-redundant and complementary methods.

Keywords: Explainable AI, Bug Detection, Transformer Models, GraphCodeBERT, Mode Collapse, Attention Rollout, Integrated Gradients, Sliding Window Ablation

How to Cite:

[1] Debargha Ghosh, Eve Thullen Ph.D, Emmanuel Udoh Ph.D, β€œQuantified Explainability and Robustness Analysis of Transformer-Based Bug Detection Models,” International Multidisciplinary Research Journal Reviews (IMRJR) (IMRJR), DOI: 10.17148/IMRJR.2026.030304

Creative Commons License This work is licensed under a Creative Commons Attribution 4.0 International License.
Google Scholar
Highest Citations
98+
h-index 3  |  i10-index 1
Peer-reviewed
Author Center
IMRJR Standards
πŸ†
Article of the Year
Award
The Future of Automotive Manufacturing: Integrating AI, ML, and Generative AI for Next-Gen Automatic Cars

Chandrakanth Rao Madhavaram, Janardhana Rao Sunkara, Chandrababu Kuraku, Eswar Prasad Galla, Hemanth Kumar Gollangi

Read Article β†’
πŸ“₯Most Downloaded
  1. 1.
  2. 2.
  3. 3.
  4. 4.
  5. 5.
Conference
Conference
International Conference Call for papers