@sjbutler

Extracting concepts from file names; a new file clustering criterion

, und . Proc. Int'l Conf. on Software Engineering., Seite 84--93. IEEE, (April 1998)
DOI: 10.1109/ICSE.1998.671105

Zusammenfassung

Decomposing complex software systems into conceptually independent subsystems is a significant software engineering activity which received considerable research attention. Most of the research in this domain considers the body of the source code; trying to cluster together files which are conceptually related. We discuss techniques for extracting concepts (abbreviations) from a more informal source of information: file names. The task is difficult because nothing indicates where to split the file names into substrings. In general, finding abbreviations would require domain knowledge to identify the concepts that are referred to in a name and intuition to recognize such concepts in abbreviated forms. We show by experiment that the techniques we propose allow about 90% of the abbreviations to be found automatically

Links und Ressourcen

Tags

Community

  • @sjbutler
  • @dblp
@sjbutlers Tags hervorgehoben