On the relationship between similar requirements and similar software-Reference-Cited by-同舟云学术

On the relationship between similar requirements and similar software

Published:2022-01-18 Issue: Volume: Page:
ISSN:0947-3602
Container-title:Requirements Engineering
language:en
Short-container-title:Requirements Eng

Author:

Abbas Muhammad^ORCID,Ferrari Alessio,Shatnawi Anas,Enoiu Eduard,Saadatmand Mehrdad,Sundmark Daniel

Abstract

AbstractRecommender systems for requirements are typically built on the assumption that similar requirements can be used as proxies to retrieve similar software. When a stakeholder proposes a new requirement, natural language processing (NLP)-based similarity metrics can be exploited to retrieve existing requirements, and in turn, identify previously developed code. Several NLP approaches for similarity computation between requirements are available. However, there is little empirical evidence on their effectiveness for code retrieval. This study compares different NLP approaches, from lexical ones to semantic, deep-learning techniques, and correlates the similarity among requirements with the similarity of their associated software. The evaluation is conducted on real-world requirements from two industrial projects from a railway company. Specifically, the most similar pairs of requirements across two industrial projects are automatically identified using six language models. Then, the trace links between requirements and software are used to identify the software pairs associated with each requirements pair. The software similarity between pairs is then automatically computed with JPLag. Finally, the correlation between requirements similarity and software similarity is evaluated to see which language model shows the highest correlation and is thus more appropriate for code retrieval. In addition, we perform a focus group with members of the company to collect qualitative data. Results show a moderately positive correlation between requirements similarity and software similarity, with the pre-trained deep learning-based BERT language model with preprocessing outperforming the other models. Practitioners confirm that requirements similarity is generally regarded as a proxy for software similarity. However, they also highlight that additional aspect comes into play when deciding software reuse, e.g., domain/project knowledge, information coming from test cases, and trace links. Our work is among the first ones to explore the relationship between requirements and software similarity from a quantitative and qualitative standpoint. This can be useful not only in recommender systems but also in other requirements engineering tasks in which similarity computation is relevant, such as tracing and change impact analysis.

Funder

itea3

Publisher

Springer Science and Business Media LLC

Subject

Information Systems,Software

Link

https://link.springer.com/content/pdf/10.1007/s00766-021-00370-4.pdf

Reference95 articles.

1. Abbas M, Ferrari A, Shatnawi A, Enoiu EP, Saadatmand M (2021) Is requirements similarity a good proxy for software similarity? an empirical investigation in industry. In: The 27th international working conference on requirements engineering: foundation for Software Quality, pp. 3–18. Springer International Publishing

2. Abbas M, Jongeling R, Lindskog C, Enoiu EP, Saadatmand M, Sundmark D (2020) Product line adoption in industry: An experience report from the railway domain. In: Proceedings of the 24th ACM Conference on Systems and Software Product Line: Volume A - Volume A, SPLC ’20. ACM, New York, NY, USA

3. Abbas M, Saadatmand M, Enoiu E, Sundamark D, Lindskog C (2020) Automated reuse recommendation of product line assets based on natural language requirements. In: S. Ben Sassi, S. Ducasse, H. Mili (eds.) Reuse in Emerging Software Engineering Practices, pp. 173–189. Springer International Publishing, Cham

4. Abualhaija S, Arora C, Sabetzadeh M, Briand LC, Traynor M (2020) Automated demarcation of requirements in textual specifications: a machine learning-based approach. Emp Softw Eng 25(6):5454–5497

5. Ali N, Guéhéneuc YG, Antoniol G (2012) Trustrace: mining software repositories to improve the accuracy of requirement traceability links. IEEE Trans Softw Eng 39(5):725–741

Cited by 7 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Enhancing Legal Compliance and Regulation Analysis with Large Language Models;2024 IEEE 32nd International Requirements Engineering Conference (RE);2024-06-24

2. Towards AI-centric Requirements Engineering for Industrial Systems;Proceedings of the 2024 IEEE/ACM 46th International Conference on Software Engineering: Companion Proceedings;2024-04-14

3. Deep Neural Networks in Natural Language Processing for Classifying Requirements by Origin and Functionality: An Application of BERT in System Requirements;Journal of Mechanical Design;2023-11-13

4. SmartDelta project: Automated quality assurance and optimization across product versions and variants;Microprocessors and Microsystems;2023-11

5. DF4RT: Deep Forest for Requirements Traceability Recovery Between Use Cases and Source Code;2023 IEEE International Conference on Systems, Man, and Cybernetics (SMC);2023-10-01