英语语料库LOB语料库

时间:2021-10-30 05:52:33
【文件属性】:

文件名称:英语语料库LOB语料库

文件大小:94.94MB

文件格式:RAR

更新时间:2021-10-30 05:52:33

LOB语料库 英语语料库

LOB语料库 创建时间: 1970年代初 创建单位:英国Lancaster大学和挪威Oslo大学以及Bergen大学 规模层级: 100万词次 基本情况:研究当代英国英语,与美国英语对比,使用了TAGIT系统,以统计方式建立换算几率矩阵,提高标注正确率。 The Lancaster-Oslo Bergen Corpus (LOB) was compiled by researchers in Lancaster, Oslo and Bergen. It consists of one million words of British En glish texts from 1961. The texts for the corpus were sampled from 15 different text categories. Each text is just over 2.000 words long (longer texts have b een cut at the first sentence boundary after 2.000 words) and the number of texts in each category varies (see table below). Further information about the t exts can be found in the LOB manual (external link). This corpus is the British counterpart of the Brown Corpus of American English. which contains texts printed in the same year so that comparison bet ween both varieties could be made


【文件预览】:
weibo_users_corpus.rar
weibo_users_corpus
----NLPIR微博博主语料库说明.txt(2KB)
----NLPIR微博博主语料库.txt(389.41MB)
lob
----LU-J-BL(1.1MB)
----LU-F-BL(601KB)
----LU-R-BL(119KB)
----LU-H-BL(420KB)
----LU-K-BL(378KB)
----LU-C-BL(241KB)
----LU-M-BL(82KB)
----LU-B-BL(377KB)
----LU-A-BL(631KB)
----LU-E-BL(513KB)
----LU-D-BL(226KB)
----LU-L-BL(314KB)
----LU-G-BL(1.02MB)
----LU-N-BL(383KB)
----LU-P-BL(379KB)

网友评论

  • 感谢楼主,能下载。但它是没有标注词性的LOB语料库,请问有没有标注了词性的LOB语料库可以下载?