Pre-Trained Model-Based NFR Classification: Overcoming Limited Data Challenges

dc.contributor.authorTariq, Muhammad
dc.contributor.authorAlzahran, Abdulrahman
dc.contributor.authorGhani, Anwar
dc.contributor.authorETAL..
dc.date.accessioned2024-05-24T06:13:27Z
dc.date.available2024-05-24T06:13:27Z
dc.date.issued2023-08-03
dc.descriptionNon-functional requirements (NFR) are the characteristics or qualities that a software system or product must possess, beyond its basic functional requirements such as maintainability, operability, performance, security, and usability [1], [2]. NFR classification is the process of categorizing these requirements to better understand and manage them throughout the software development lifecycle [3], [4]. However, manually classifying NFR can be a time-consuming and error-prone task, especially in large software systems.
dc.description.abstractMachine learning techniques have shown promising results in classifying non-functional requirements (NFR). However, the lack of annotated training data in the domain of requirement engineering poses challenges to the accuracy, generalization, and reliability of ML-based methods, including overfitting, poor performance, biased models, and out-of-vocabulary issues. This study presents an approach for the classification of NFR in software requirements specification documents by extracting features from word embedding pre-trained models. The novel algorithms are specifically designed to extract relevant representative features from pre-trained word embedding models. In addition, each pre-trained model is paired with the four tailored neural network architectures for NFR classification including RPCNN, RPBiLSTM, RPLSTM, and RPANN. This combination results in the creation of twelve unique models, each with its unique configuration and characteristics. The results show that the integration of pre-trained GloVe models with RPBiLSTM demonstrates the highest performance, achieving an impressive average Area Under the Curve (AUC) score of 96%, a precision of 85%, and recall of 82%, highlighting its strong ability to accurately classify NFRs. Furthermore, among the integration of pre-trained Word2Vec models, RPLSTM achieved notable results, with an AUC score of 95%, precision of 86%, and recall of 80%. Similarly, integrated fastText-based pre-trained models the RPBiLSTM yield competitive performance, with an AUC score of 95%, precision of 85%, and recall of 80%. This comprehensive and integrated approach provides a practical solution for effectively analyzing and classifying NFR, thereby facilitating improved software development practices. Keywords Feature extraction, Data models, Classification algorithms, Software algorithms, Training, Requirements engineering, Task analysisen
dc.identifier.citationRahman, K., Ghani, A., Alzahrani, A., Tariq, M. U., & Rahman, A. U. (2023). Pre-trained model-based NFR classification: Overcoming limited data challenges. IEEE Access, 11, 81787-81802.
dc.identifier.doihttps://doi.org/10.1109/ACCESS.2023.3301725
dc.identifier.urihttps://dspace.adu.ac.ae/handle/1/5405
dc.language.isoen
dc.publisherIEEE Xplore
dc.titlePre-Trained Model-Based NFR Classification: Overcoming Limited Data Challenges
dc.typeArticle

Files

License bundle

Now showing 1 - 1 of 1
Loading...
Thumbnail Image
Name:
license.txt
Size:
1.71 KB
Format:
Item-specific license agreed to upon submission
Description: