GI LogoGI Logo
  • Anmelden
Digitale Bibliothek
    • Gesamter Bestand

      • Bereiche & Sammlungen
      • Titel
      • Autor
      • Erscheinungsdatum
      • Schlagwort
    • Diese Sammlung

      • Titel
      • Autor
      • Erscheinungsdatum
      • Schlagwort
Digital Bibliothek der Gesellschaft für Informatik e.V.
GI-DL
    • English
    • Deutsch
  • Deutsch 
    • English
    • Deutsch
Dokumentanzeige 
  •   Startseite
  • Lecture Notes in Informatics
  • Proceedings
  • Software Engineering
  • P213 - Software Engineering 2013
  • Dokumentanzeige
JavaScript is disabled for your browser. Some features of this site may not work without it.
  •   Startseite
  • Lecture Notes in Informatics
  • Proceedings
  • Software Engineering
  • P213 - Software Engineering 2013
  • Dokumentanzeige

An adaptive filter-framework for the quality improvement of open-source software analysis

Autor(en):
Hannemann, Anna [DBLP] ;
Hackstein, Michael [DBLP] ;
Klamma, Ralf [DBLP] ;
Jarke, Matthias [DBLP]
Zusammenfassung
Knowledge mining in Open-Source Software (OSS) brings a great benefit for software engineering (SE). The researchers discover, investigate, and even simulate the organization of development processes within open-source communities in order to understand the community-oriented organization and to transform its advantages into conventional SE projects. Despite a great number of different studies on OSS data, not much attention has been paid to the data filtering step so far. The noise within uncleaned data can lead to inaccurate conclusions for SE. A special challenge for data cleaning presents the variety of communicational and development infrastructures used by OSS projects. This paper presents an adaptive filter-framework supporting data cleaning and other preprocessing steps. The framework allows to combine filters in arbitrary order, defining which preprocessing steps should be performed. The filter-portfolio can by extended easily. A schema matching in case of cross-project analysis is available. Three filters - spam detection, quotation elimination and coreperiphery distinction - were implemented within the filter-framework. In the analysis of three large-scale OSS projects (BioJava, Biopython, BioPerl), the filtering led to a significant data modification and reduction. The results of text mining (sentiment analysis) and social network analysis on uncleaned and cleaned data differ significantly, confirming the importance of the data preprocessing step within OSS empirical studies.
  • Vollständige Referenz
  • BibTeX
Hannemann, A., Hackstein, M., Klamma, R. & Jarke, M., (2013). An adaptive filter-framework for the quality improvement of open-source software analysis. In: Kowalewski, S. & Rumpe, B. (Hrsg.), Software Engineering 2013. Bonn: Gesellschaft für Informatik e.V.. (S. 143-156).
@inproceedings{mci/Hannemann2013,
author = {Hannemann, Anna AND Hackstein, Michael AND Klamma, Ralf AND Jarke, Matthias},
title = {An adaptive filter-framework for the quality improvement of open-source software analysis},
booktitle = {Software Engineering 2013},
year = {2013},
editor = {Kowalewski, Stefan AND Rumpe, Bernhard} ,
pages = { 143-156 },
publisher = {Gesellschaft für Informatik e.V.},
address = {Bonn}
}
DateienGroesseFormatAnzeige
143.pdf130.9Kb PDF Öffnen

Haben Sie fehlerhafte Angaben entdeckt? Sagen Sie uns Bescheid: Feedback abschicken

Mehr Information

ISBN: 978-3-88579-607-7
ISSN: 1617-5468
Datum: 2013
Sprache: en (en)
Typ: Text/Conference Paper
Sammlungen
  • P213 - Software Engineering 2013 [36]

Zur Langanzeige


Über uns | FAQ | Hilfe | Impressum | Datenschutz

Gesellschaft für Informatik e.V. (GI), Kontakt: Geschäftsstelle der GI
Diese Digital Library basiert auf DSpace.

 

 


Über uns | FAQ | Hilfe | Impressum | Datenschutz

Gesellschaft für Informatik e.V. (GI), Kontakt: Geschäftsstelle der GI
Diese Digital Library basiert auf DSpace.