Skip to main navigation Skip to search Skip to main content

Pre-trained Model-based Actionable Warning Identification: A Feasibility Study

  • Xiuting GE
  • , Chunrong FANG*
  • , Quanjun ZHANG
  • , Daoyuan WU
  • , Bowen YU
  • , Qirui ZHENG
  • , An GUO
  • , Shangwei LIN
  • , Zhihong ZHAO
  • , Yang LIU
  • , Zhenyu CHEN
  • *Corresponding author for this work

Research output: Journal PublicationsJournal Article (refereed)peer-review

Abstract

Actionable Warning Identification (AWI) plays a pivotal role in improving the usability of Static Code Analyzers (SCAs). Currently, Machine Learning (ML)-based AWI approaches, which mainly learn an AWI classifier from labeled warnings, are notably common. However, these approaches still face the problem of restricted performance due to the direct reliance on a limited number of labeled warnings to develop a classifier. Very recently, Pre-trained Models (PTMs), which have been trained through billions of text/code tokens and have demonstrated substantial successful applications in various code-related tasks, could potentially address the above problem. Nevertheless, the performance of PTMs on AWI has not been systematically investigated, leaving a gap in understanding their pros and cons. In this article, we are the first to explore the feasibility of applying various PTMs for AWI. By conducting an extensive evaluation on 12K+ warnings involving four commonly used SCAs (i.e., SpotBugs, Infer, CppCheck, and CSA) and three typical programming languages (i.e., Java, C, and C++), we (1) investigate the overall PTM-based AWI performance compared to the state-of-the-art ML-based AWI approach, (2) analyze the impact of three primary aspects (i.e., data preprocessing, model training, and model prediction) in the typical PTM-based AWI workflow, and (3) identify the reasons for the current underperformance of PTMs on AWI, thereby obtaining a series of findings. Based on the above findings, we further provide several potential directions to enhance PTM-based AWI.

Original languageEnglish
Article number238
Number of pages29
JournalACM Transactions on Software Engineering and Methodology
Volume35
Issue number8
Early online date18 Nov 2025
DOIs
Publication statusPublished - 17 Jul 2026

Bibliographical note

Publisher Copyright:
© 2026 Copyright held by the owner/author(s). Publication rights licensed to ACM.

Funding

The work is partly supported by the National Natural Science Foundation of China (62502476 and U24A20337), Fundamental Research Funds for the Central Universities at China University of Geosciences (Wuhan) (CUG250690), “CUG Scholar” Scientific Research Funds at China University of Geosciences (Wuhan) (2025039), and Natural Science Foundation of Jiangsu Province (BK20251458).

Keywords

  • Actionable warning identification
  • feasibility study
  • pre-trained model
  • static analysis

Fingerprint

Dive into the research topics of 'Pre-trained Model-based Actionable Warning Identification: A Feasibility Study'. Together they form a unique fingerprint.

Cite this