Syntax tagging is the main and most difficult point of corpus tagging, and should be based on syntax theory.
句法标注是语料标注的重点、难点所在,必须以一定的句法理论为基础。
Meanwhile, a novel and convenient training corpus tagging method was proposed, which made Hidden Markov Model practically usable in Musical Entity Recognition.
同时,我们提出了一种新颖、实用的训练语料标注方案,这使得隐马尔科夫模型在音乐实体识别上变得实际可行。
Our experiment for the large scale real corpus tagging proves that transformation-based algorithm and tri-gram statistic method bring out the best in each other.
对大规模真实语料的标注实验表明基于转换的方法与三元统计模型方法相得益彰;
This paper is to discuss the problem and put forward a practical method of tagging language errors in corpus tagging procedure by XML(Extensible Markup Language).
本文从语料库加工流程的角度,探讨了这一问题,并借助XML(可扩展置标语言)提出了错误标注的具体实现方法。
In Chinese corpus tagging, we have analyzed forefathers' rule based work, and have proposed the method based on rule PRI, finished the work of word segmenting and Chinese corpus tagging finally.
在进行词性标注时,作者分析了前人的基于规则的词性标注的工作,并提出了基于规则优先级的词性标注方法,最后实现了分词和标注系统。
The platform can tag word senses in large-scale corpus, and automatic statistic tagging effect, It is convenient for artificial verifying.
平台可对大规模语料库中的词义进行标注,同时自动统计出标注效果,方便人工进行验证。
In the deep processing of large-scale corpus, it has been a chief problem to assure the consistence of part of speech tagging to build the high-quantity corpus.
在对大规模语料库进行深加工时,保证词性标注的一致性已成为建设高质量语料库的首要问题。
The tagging of Mongolian phrases is the further study with Mongolian corpus linguistics.
蒙古语短语标注是蒙古语语料库语言学研究的进一步深化。
In order to set up part of speech tagging corpus efficiently, one practical Chinese syntax compilation and analyse assistant system is designed and implemented.
为高效地建立句法标注语料库,设计研发了一个实用的中文句法编辑与分析辅助系统。
This paper first explores unsupervised part-of-speech tagging for Chinese via monolingual corpus.
本文首先探索了基于单语料库的无监督中文词性标注。
This paper introduces the idea of a corpus of Chinese NP syntactic positions, as well as the tag set used and the principles for tagging. The potential value of this corpus is also pointed out.
本文以非统计的信息处理方法为出发点,介绍一个汉语名物性短语句法位置语料库的设计思想、所使用的句法位置标记集以及标记加工规范,并指出了这样一个语料库的潜在价值。
This paper introduces the idea of a corpus of Chinese NP syntactic positions, as well as the tag set used and the principles for tagging. The potential value of this corpus is also pointed out.
本文以非统计的信息处理方法为出发点,介绍一个汉语名物性短语句法位置语料库的设计思想、所使用的句法位置标记集以及标记加工规范,并指出了这样一个语料库的潜在价值。
应用推荐