GI LogoGI Logo
  • Anmelden
Digitale Bibliothek
    • Gesamter Bestand

      • Bereiche & Sammlungen
      • Titel
      • Autor
      • Erscheinungsdatum
      • Schlagwort
    • Diese Sammlung

      • Titel
      • Autor
      • Erscheinungsdatum
      • Schlagwort
Digital Bibliothek der Gesellschaft für Informatik e.V.
GI-DL
    • English
    • Deutsch
  • Deutsch 
    • English
    • Deutsch
Dokumentanzeige 
  •   Startseite
  • Lecture Notes in Informatics
  • Proceedings
  • BTW - Datenbanksysteme für Business, Technologie und Web
  • P216 - BTW2013 - Datenbanksysteme für Business, Technologie und Web – Workshopband
  • Dokumentanzeige
JavaScript is disabled for your browser. Some features of this site may not work without it.
  •   Startseite
  • Lecture Notes in Informatics
  • Proceedings
  • BTW - Datenbanksysteme für Business, Technologie und Web
  • P216 - BTW2013 - Datenbanksysteme für Business, Technologie und Web – Workshopband
  • Dokumentanzeige

Enhancing named entity extraction by effectively incorporating the crowd

Autor(en):
Braunschweig, Katrin [DBLP] ;
Thiele, Maik [DBLP] ;
Eberius, Julian [DBLP] ;
Lehner, Wolfgang [DBLP]
Zusammenfassung
Named entity extraction is an established research area in the field of information extraction. When tailored to a specific domain and with sufficient pre-labeled training data, state-of-the-art extraction algorithms have achieved near human performance. However, when presented with semi-structured data, informal text or unknown domains where training data is not available, extraction results can deteriorate significantly. Recent research has focused on crowdsourcing as an alternative to automatic named entity extraction or as a tool to generate the required training data. While humans easily adapt to semi-structured data and informal style, a crowd-based approach also introduces new issues due to monetary costs or spamming. We address these issues by combining automatic named entity extraction algorithms with crowdsourcing into a hybrid approach. We have conducted a wide range of experiments on real world data to identify a set of subtasks or operators, that can be performed either by the crowd or automatically. Results show that a meaningful combination of these operators into complex processing pipelines can significantly enhance the quality of named entity extraction in challenging scenarios, while at the same time reducing the monetary costs of crowdsourcing and the risk of misuse.
  • Vollständige Referenz
  • BibTeX
Braunschweig, K., Thiele, M., Eberius, J. & Lehner, W., (2013). Enhancing named entity extraction by effectively incorporating the crowd. In: Saake, G., Henrich, A., Lehner, W., Neumann, T. & Köppen, V. (Hrsg.), Datenbanksysteme für Business, Technologie und Web (BTW) 2013 - Workshopband. Bonn: Gesellschaft für Informatik e.V.. (S. 181-196).
@inproceedings{mci/Braunschweig2013,
author = {Braunschweig, Katrin AND Thiele, Maik AND Eberius, Julian AND Lehner, Wolfgang},
title = {Enhancing named entity extraction by effectively incorporating the crowd},
booktitle = {Datenbanksysteme für Business, Technologie und Web (BTW) 2013 - Workshopband},
year = {2013},
editor = {Saake, Gunter AND Henrich, Andreas AND Lehner, Wolfgang AND Neumann, Thomas AND Köppen, Veit} ,
pages = { 181-196 },
publisher = {Gesellschaft für Informatik e.V.},
address = {Bonn}
}
DateienGroesseFormatAnzeige
181.pdf1.976Mb PDF Öffnen

Haben Sie fehlerhafte Angaben entdeckt? Sagen Sie uns Bescheid: Feedback abschicken

Mehr Information

ISBN: 978-3-88579-610-7
ISSN: 1617-5468
Datum: 2013
Sprache: en (en)
Typ: Text/Conference Paper
Sammlungen
  • P216 - BTW2013 - Datenbanksysteme für Business, Technologie und Web – Workshopband [31]

Zur Langanzeige


Über uns | FAQ | Hilfe | Impressum | Datenschutz

Gesellschaft für Informatik e.V. (GI), Kontakt: Geschäftsstelle der GI
Diese Digital Library basiert auf DSpace.

 

 


Über uns | FAQ | Hilfe | Impressum | Datenschutz

Gesellschaft für Informatik e.V. (GI), Kontakt: Geschäftsstelle der GI
Diese Digital Library basiert auf DSpace.