Korean AMR Corpus
收藏资源简介:
为了探索韩语在语义库中的潜力及韩语句子的意义表示方法,本文报告了将抽象意义表示应用于韩语的过程及其输出:韩语AMR语料库。该语料库目前包含1,253个句子,原始文本来自ExoBrain Corpus,这是一个由国家领导的语言AI研发项目。本文还从定性和定量两个角度分析了结果,并提出了进一步发展的讨论。
To explore the potential of the Korean language in semantic repositories and the methods for representing the meaning of Korean sentences, this paper reports on the process of applying Abstract Meaning Representation (AMR) to Korean and its output: the Korean AMR corpus. This corpus currently contains 1,253 sentences, with the original texts sourced from the ExoBrain Corpus, a national-led language AI research and development project. The paper also analyzes the results from both qualitative and quantitative perspectives and proposes discussions for further development.
韩国抽象意义表示语料库(Korean AMR Corpus)
数据集概述
- 名称: 韩国抽象意义表示语料库(Korean AMR Corpus)
- 创建者: Hyonsu Choe, Jiyoon Han, Hyejin Park, Tae Hwan Oh, Hansaem Kim
- 创建年份: 2020
- 出版物: 《构建韩国抽象意义表示语料库》发表于第二国际设计意义表示研讨会
- 出版地点: 巴塞罗那(在线)
- 出版机构: 计算语言学协会
- 论文页码: 21-29
- 数据集大小: 1,253个句子
- 原始文本来源: ExoBrain Corpus
数据集目的
- 探索韩国语义银行化的潜力及韩国句子意义的表示方法
- 应用抽象意义表示(AMR)框架至韩语,并构建相应的语料库
数据集分析
- 通过定性和定量分析,对结果进行评估
- 提出未来发展的讨论
引用信息
@inproceedings{choe-etal-2020-building, title = "Building {K}orean {A}bstract {M}eaning {R}epresentation Corpus", author = "Choe, Hyonsu and Han, Jiyoon and Park, Hyejin and Oh, Tae Hwan and Kim, Hansaem", booktitle = "Proceedings of the Second International Workshop on Designing Meaning Representations", month = dec, year = "2020", address = "Barcelona Spain (online)", publisher = "Association for Computational Linguistics", url = "https://aclanthology.org/2020.dmr-1.3", pages = "21--29", abstract = "To explore the potential sembanking in Korean and ways to represent the meaning of Korean sentences, this paper reports on the process of applying Abstract Meaning Representation to Korean, a semantic representation framework that has been studied in wide range of languages, and its output: the Korean AMR corpus. The corpus which is constructed so far is a size of 1,253 sentences and its raw texts are from ExoBrain Corpus, a state-led R{&}D project on language AI. This paper also analyzes the result in both qualitative and quantitative manners, proposing discussions for further development.", }




