The Qualitas Corpus: A Curated Collection of Java Code for Empirical Studies

  • Authors:
  • Ewan Tempero;Craig Anslow;Jens Dietrich;Ted Han;Jing Li;Markus Lumpe;Hayden Melton;James Noble

  • Affiliations:
  • -;-;-;-;-;-;-;-

  • Venue:
  • APSEC '10 Proceedings of the 2010 Asia Pacific Software Engineering Conference
  • Year:
  • 2010

Quantified Score

Hi-index 0.00

Visualization

Abstract

In order to increase our ability to use measurement to support software development practise we need to do more analysis of code. However, empirical studies of code are expensive and their results are difficult to compare. We describe the Qualitas Corpus, a large curated collection of open source Java systems. The corpus reduces the cost of performing large empirical studies of code and supports comparison of measurements of the same artifacts. We discuss its design, organisation, and issues associated with its development.