Identifying and Translating Subjective Content Descriptions Among Texts-Reference-Cited by-同舟云学术

Identifying and Translating Subjective Content Descriptions Among Texts

Published:2021-12 Issue:04 Volume:15 Page:461-485
ISSN:1793-351X
Container-title:International Journal of Semantic Computing
language:en
Short-container-title:Int. J. Semantic Computing

Author:

Bender Magnus¹,Braun Tanya¹,Gehrke Marcel¹,Kuhr Felix¹,Möller Ralf¹,Schiff Simon¹

Affiliation:

1. Institute of Information Systems, University of Lübeck, Ratzeburger Allee 160, 23562 Lübeck, Germany

Abstract

An agent pursuing a task may work with a corpus of documents as a reference library. Subjective content descriptions (SCDs) provide additional data that add value in the context of the agent’s task. In the pursuit of documents to add to the corpus, an agent may come across new documents where content text and SCDs from another agent are interleaved and no distinction can be made unless the agent knows the content from somewhere else. Therefore, this paper presents a hidden Markov model-based approach to identify SCDs in a new document where SCDs occur inline among content text. Additionally, we present a dictionary selection approach to identify suitable translations for content text and SCDs based on [Formula: see text]-grams. We end with a case study evaluating both approaches based on simulated and real-world data.

Publisher

World Scientific Pub Co Pte Ltd

Subject

Artificial Intelligence,Computer Networks and Communications,Computer Science Applications,Linguistics and Language,Information Systems,Software

Link

https://www.worldscientific.com/doi/pdf/10.1142/S1793351X21400122

Reference20 articles.

1. To Extend or Not to Extend? Context-Specific Corpus Enrichment

2. Corpus-Driven Annotation Enrichment

3. The viterbi algorithm

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Unsupervised Estimation of Subjective Content Descriptions in an Information System;International Journal of Semantic Computing;2024-01-30

2. Unsupervised Estimation of Subjective Content Descriptions;2023 IEEE 17th International Conference on Semantic Computing (ICSC);2023-02

3. TEI-Based Interactive Critical Editions;Document Analysis Systems;2022