Detecting Occupations in German Texts : Challenges and Data

item.page.bibb.id

784672
Loading...
Thumbnail Image

Date

Journal Title

Journal ISSN

Volume Title

Publisher

item.page.bibb.publisherplace

item.page.bibb.participation

BIBB-Mitarbeiter

item.page.bibb.citedinbibb

item.page.bibb.researchfocus

item.page.bibb.reviewof

item.page.bibb.citationdr

Abstract

"This paper is concerned with the detection and classification of occupational titles in German texts, with a focus on the linguistic and structural challenges that are unique to the German language. The study utilizes an extensive dataset on job title variants to assess rule-based and language model-based methodologies across diverse corpora, encompassing historical documents and parliamentary proceedings. The findings indicate that rule-based methods demonstrate robust performance, particularly in structured texts, while large language models exhibit complementary strengths in recognizing complex terms. The work makes a significant contribution to the field by providing valuable annotated data and methodological insights." (Authors' abstract, BIBB-Doku)

Description

Keywords

Citation

item.page.bibb.voevzlink

item.page.bibb.additionallink

Endorsement

Review

Supplemented By

Referenced By

Creative Commons license

Except where otherwise noted, this item's license is described as Namensnennung 4.0 International