2011/02/18 by Heiko Hellweg, Hellweg, Heiko, Jürgen Krause +11
Computer Science · #Advanced Text Analysis Techniques #FOS: Computer and information sciences #Information Retrieval (cs.IR) #Natural Language Processing Techniques #Semantic Web and Ontologies #cs.IR
paper · pdf · doi:10.48550/arxiv.1102.3866
Technical Report (Arbeitsbericht) GESIS - Leibniz Institute for the Social Sciences
arxiv created 2011/02/18 · openalex publication_date 2011/02/18 · arxiv updated 2011/02/21 · openalex created_date 2025/10/10 · openalex updated_date 2026/07/28
The first step to handle semantic heterogeneity should be the attempt to enrich the semantic information about documents, i.e. to fill up the gaps in the documents meta-data automatically. Section 2 describes a set of cascading deductive and heuristic extraction rules, which were developed in the project CARMEN for the domain of Social Sciences. The mapping between different terminologies can be done by using intellectual, statistical and/or neural network transfer modules. Intellectual transfers use cross-concordances between different classification schemes or thesauri. Section 3 describes the creation, storage and handling of such transfers.