University of Leicester
Browse

How to Make Best Use of Cross-Company Data for Web Effort Estimation?

conference contribution
posted on 2015-11-30, 10:12 authored by Leandro Lei Minku, Federica Sarro, Emilia Mendes, Filomena Ferrucci
[Context]: The numerous challenges that can hinder software companies from gathering their own data have motivated over the past 15 years research on the use of cross-company (CC) datasets for software effort prediction. Part of this research focused on Web effort prediction, given the large increase worldwide in the development of Web applications. Some of these studies indicate that it may be possible to achieve better performance using CC models if some strategy to make the CC data more similar to the within-company (WC) data is adopted. [Goal]: This study investigates the use of a recently proposed approach called Dycom to assess to what extent Web effort predictions obtained using CC datasets are effective in relation to the predictions obtained using WC data when explicitly mapping the CC models to the WC context. [Method]: Data on 125 Web projects from eight different companies part of the Tukutuku database were used to build prediction models. We benchmarked these models against baseline models (mean and median effort) and a WC base learner that does not benefit of the mapping. We also compared Dycom against a competitive CC approach from the literature (NN-filtering). We report a company-by- company analysis. [Results]: Dycom usually managed to achieve similar or better performance than a WC model while using only half of the WC training data. These results are also an improvement over previous studies that investigated the use of different strategies to adapt CC models to the WC data for Web effort estimation. [Conclusions]: We conclude that the use of Dycom for Web effort prediction is quite promising and in general supports previous results when applying Dycom to conventional software datasets.

History

Citation

Proceedings of the 9th ACM/IEEE International Symposium on Empirical Software Engineering and Measurement (ESEM), pp. 172-181

Author affiliation

/Organisation/COLLEGE OF SCIENCE AND ENGINEERING/Department of Computer Science

Source

Beijing

Version

  • AM (Accepted Manuscript)

Published in

Proceedings of the 9th ACM/IEEE International Symposium on Empirical Software Engineering and Measurement (ESEM)

Publisher

ACM, IEEE

isbn

978-1-4673-7899-4

Copyright date

2015

Available date

2015-11-30

Publisher version

http://ieeexplore.ieee.org/xpl/articleDetails.jsp?arnumber=7321199&filter=AND(p_IS_Number:7321177)

Notes

Archived in accordance with the publisher's posting policy, available at http://www.ieee.org/publications_standards/publications/rights/rights_policies.html

Temporal coverage: start date

2015-10-22

Temporal coverage: end date

2015-10-23

Language

en

Usage metrics

    University of Leicester Publications

    Categories

    No categories selected

    Exports

    RefWorks
    BibTeX
    Ref. manager
    Endnote
    DataCite
    NLM
    DC