遇见数据集

Japanese Novel Data

收藏
NIAID Data Ecosystem2026-03-14 收录
数据链接:
官方服务:

资源简介:

A dataset on Japanese novels. This dataset contains 1113 novels. This dataset contains the following variables: Title, words, unique words, unique words used once, UWUU%, kanji, kanji used once, kanji readings, difficulty, average sentence length, characters, publisher, pages, ASIN, and the Japanese title. Note: not all observations have complete publisher, page, ASIN, and Japanese title data. The variables title, words, unique words, unique words used once, UWUU%, kanji, kanji used once, kanji readings, difficulty, average sentence length, and characters were obtained from Jpdb.io. Publisher, pages, ASIN, and the Japanese title were obtained from Amazon.co.jp. The dataset was mined using python and BS4 by CDT Budwell. This dataset was created in support of MA206 (Intro to Statistics) at USMA West Point by CDT Jackson Budwell '25.

创建时间:
2022-11-27
二维码
社区交流群
二维码
科研交流群
商业服务