|
|
mehr: a persian coreference resolution corpus
|
|
|
|
|
نویسنده
|
haji mohammadi hassan ,talebpour alireza ,mahmoudi aznaveh ahamd ,yazdani samaneh
|
منبع
|
journal of ai and data mining - 2023 - دوره : 11 - شماره : 3 - صفحه:407 -416
|
چکیده
|
Coreference resolution is one of the essential tasks of natural languageprocessing. this task identifies all in-text expressions that refer to thesame entity in the real world. coreference resolution is used in otherfields of natural language processing, such as information extraction,machine translation, and question-answering.this article presents a new coreference resolution corpus in persiannamed mehr corpus. the article’s primary goal is to develop a persiancoreference corpus that resolves some of the previous persian corpus’sshortcomings while maintaining a high inter-annotator agreement. thiscorpus annotates coreference relations for noun phrases, namedentities, pronouns, and nested named entities. two baseline pronounresolution systems are developed, and the results are reported. thecorpus size includes 400 documents and about 170k tokens. corpusannotation is done by webanno preprocessing tool.
|
کلیدواژه
|
natural language processing ,mention ,anaphora resolution ,antecedent
|
آدرس
|
islamic azad university, north tehran branch, department of computer engineering, iran, shahid beheshti university, department of computer engineering, iran, shahid beheshti university, department of computer engineering, iran, islamic azad university, north tehran branch, department of computer engineering, iran
|
پست الکترونیکی
|
samaneh.yazdani@gmail.com
|
|
|
|
|
|
|
|
|
|
|
|
Authors
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|
|