Zur Hauptnavigation wechseln Zur Suche wechseln Zum Hauptinhalt wechseln

JobOlize - Headhunting by Information Extraction in the era of Web 2.0

Publikation: Beitrag in Buch/Bericht/KonferenzbandKonferenzbeitragBegutachtung

Abstract

E-recruitment is one of the most successful ebusiness applications supporting both, headhunters and job seekers. The explosive growth of online job offers makes the usage of information extraction techniques to build up, e.g., job portals in a semiautomatic way a necessity. Existing approaches, however, hardly cope with the heterogeneous and semistructured nature of job offers nor do they consider potentials offered by Web 2.0 technologies. This paper proposes an information extraction system called “JobOlize”1, realized for arbitrarily structured IT job offers. To improve extraction quality, a hybrid approach is employed, combining existing NLPtechniques with a new form of context-driven extraction, incorporating layout, structure and content information. To allow users a proper adaptation of the extraction results while preserving the look and feel of the original Web pages, a rich client interface is provided. The improvements in extraction quality are justified on basis of a case study and the experiences gained are generalized and critically reflected by discussing lessons learned.
OriginalspracheEnglisch
TitelProceedings of the 7th International Workshop on Web-Oriented Software Technologies (IWWOST 2008), Yorktown Heights, New York, USA, July 14, 2008
PublikationsstatusVeröffentlicht - 2008

UN SDGs

Dieser Output leistet einen Beitrag zu folgendem(n) Ziel(en) für nachhaltige Entwicklung

  1. SDG 9 – Industrie, Innovation und Infrastruktur
    SDG 9 – Industrie, Innovation und Infrastruktur
  2. SDG 16 – Frieden, Gerechtigkeit und starke Institutionen
    SDG 16 – Frieden, Gerechtigkeit und starke Institutionen

Wissenschaftszweige

  • 101004 Biomathematik
  • 101027 Dynamische Systeme
  • 101028 Mathematische Modellierung
  • 101029 Mathematische Statistik
  • 101014 Numerische Mathematik
  • 101015 Operations Research
  • 101016 Optimierung
  • 101017 Spieltheorie
  • 101018 Statistik
  • 101019 Stochastik
  • 101024 Wahrscheinlichkeitstheorie
  • 101026 Zeitreihenanalyse
  • 102 Informatik
  • 102001 Artificial Intelligence
  • 102003 Bildverarbeitung
  • 102004 Bioinformatik
  • 102013 Human-Computer Interaction
  • 102018 Künstliche Neuronale Netze
  • 102019 Machine Learning
  • 103029 Statistische Physik
  • 106005 Bioinformatik
  • 106007 Biostatistik
  • 202017 Embedded Systems
  • 202035 Robotik
  • 202036 Sensorik
  • 202037 Signalverarbeitung
  • 305901 Computerunterstützte Diagnose und Therapie
  • 305905 Medizinische Informatik
  • 305907 Medizinische Statistik
  • 102032 Computational Intelligence
  • 102033 Data Mining
  • 101031 Approximationstheorie
  • 102006 Computer Supported Cooperative Work (CSCW)
  • 102010 Datenbanksysteme
  • 102014 Informationsdesign
  • 102015 Informationssysteme
  • 102016 IT-Sicherheit
  • 102028 Knowledge Engineering
  • 102022 Softwareentwicklung
  • 102025 Verteilte Systeme
  • 502007 E-Commerce
  • 505002 Datenschutz
  • 506002 E-Government
  • 509018 Wissensmanagement
  • 202007 Computer Integrated Manufacturing (CIM)
  • 102035 Data Science

Dieses zitieren